FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech

Alexis Conneau

Min Ma

Simran Khanuja

Yu Zhang

Vera Axelrod

Siddharth Dalmia

Jason Riesa

Clara Rivera

Ankur Bapna

IEEE Spoken Language Technology Workshop (SLT) (2022)

Google Scholar

Abstract

We introduce FLEURS, the Few-shot Learning Evaluation of Universal Representations of Speech benchmark. FLEURS is an n-way parallel speech dataset in 102 languages built on top of the machine translation FLoRes-101 benchmark, with approximately 12 hours of speech supervision per language. FLEURS can be used for a variety of speech tasks, including Automatic Speech Recognition (ASR), Speech Language Identification (Speech LangID), Translation and Retrieval. In this paper, we provide baselines for the tasks based on multilingual pre-trained models like mSLAM. The goal of FLEURS is to enable speech technology in more languages and catalyze research in low-resource speech understanding.

Research Areas

Speech Processing

Defining the technology of today and tomorrow.

Philosophy

People

Teams

AI/ML Foundations  & Capabilities

Algorithms & Optimization

Computing Paradigms

Responsible Human-Centric Technology

Science & Societal Impact

Projects

Publications

Resources

Shaping the future, together.

Student programs

Faculty programs

Conferences & events

FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech

Abstract

Research Areas

Learn more about how we conduct our research

Defining the technology of today and tomorrow.

Philosophy

People

Teams

AI/ML Foundations & Capabilities

Algorithms & Optimization

Computing Paradigms

Responsible Human-Centric Technology

Science & Societal Impact

Projects

Publications

Resources

Shaping the future, together.

Student programs

Faculty programs

Conferences & events

FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech

Abstract

Research Areas

Learn more about how we conduct our research

AI/ML Foundations  & Capabilities