fleurs google
FLEURS Fleurs is the speech version of the FLoRes machine translation benchmark. We use 2009 n-way parallel sentences from the FLoRes dev and devtest publicly available sets, in 102 languages. Training sets have around 10 hours of supervision. Speakers of the train sets are different than speakers from the dev/test sets. Multilingual fine-tuning is used and ”unit error rate” (characters, signs) of all languages is averaged. Languages and results are also grouped into seven… See the full description on the dataset page: https://huggingface.co/datasets/google/fleurs.
- Tipo
- dataset
- Licencia
- cc-by-4.0
- Lenguaje
- afr
- Descargas
- 97,349
- Me gusta
- 443
- Acceso
- public
- Archivos
- 0
Etiquetas
- reconocimiento de voz
- audío
- multilingüe
- benchmark
- dataset de audio
- evaluación de modelos
- NLP
- habla
Resumen
FLEURS es la versión de voz del benchmark de traducción automática FLoRes, que comprende oraciones paralelas en 102 idiomas para evaluación del reconocimiento automático del habla. Forma parte del benchmark XTREME-S y sirve para evaluar representaciones del habla en múltiples tareas como reconocimi…
README
--- annotations_creators: - expert-generated - crowdsourced - machine-generated language_creators: - crowdsourced - expert-generated language: - afr - amh - ara - asm - ast - azj - bel - ben - bos - cat - ceb - cmn - ces - cym - dan - deu - ell - eng - spa - est - fas - ful - fin - tgl - fra - gle …