fleurs google

FLEURS Fleurs is the speech version of the FLoRes machine translation benchmark. We use 2009 n-way parallel sentences from the FLoRes dev and devtest publicly available sets, in 102 languages. Training sets have around 10 hours of supervision. Speakers of the train sets are different than speakers from the dev/test sets. Multilingual fine-tuning is used and ”unit error rate” (characters, signs) of all languages is averaged. Languages and results are also grouped into seven… See the full description on the dataset page: https://huggingface.co/datasets/google/fleurs.

Tipo
dataset
Licencia
cc-by-4.0
Lenguaje
afr
Descargas
97,349
Me gusta
443
Acceso
public
Archivos
0

Etiquetas

  • reconocimiento de voz
  • audío
  • multilingüe
  • benchmark
  • dataset de audio
  • evaluación de modelos
  • NLP
  • habla

Resumen

FLEURS es la versión de voz del benchmark de traducción automática FLoRes, que comprende oraciones paralelas en 102 idiomas para evaluación del reconocimiento automático del habla. Forma parte del benchmark XTREME-S y sirve para evaluar representaciones del habla en múltiples tareas como reconocimi…

README

--- annotations_creators: - expert-generated - crowdsourced - machine-generated language_creators: - crowdsourced - expert-generated language: - afr - amh - ara - asm - ast - azj - bel - ben - bos - cat - ceb - cmn - ces - cym - dan - deu - ell - eng - spa - est - fas - ful - fin - tgl - fra - gle …

查看完整页面 · 查看原文