espnet/yodas-granary
Dataset Card for YODAS-Granary Repository: NeMo-speech-data-processor: Granary Paper: Granary: Speech Recognition and Translation Dataset in 25 European Languages Shared by: ESPnet Dataset Description YODAS-Granary is a curated subset of the larger nvidia/Granary dataset, focusing on high-quality pseudo-labeled speech data for Automatic Speech Recognition (ASR) and Automatic Speech Translation (AST) across 23 European languages. Overview… See the full description on the dataset page: https://huggingface.co/datasets/espnet/yodas-granary.
mlforge datasets pull espnet/yodas-granary
Dataset details
About espnet/yodas-granary
Table of Contents - Dataset Description - Overview - Data Distribution - How to Use - Standard Loading - Streaming - Dataset Structure - Data Instance - Data Fields - Data Splits - Reference