Project collection · Ongoing
SpeeD-TB
Speech Datasets and Models for Tibeto-Burman Languages of India
Dataset created under the Speech Datasets and Models for Tibeto-Burman Languages (SpeeD-TB), sponsored under Mission Bhashini by Ministry of Electronics and Information Technology (MEITY), Govt of India. The project aimed to create 1,200 hours of speech dataset and ASR models for 6 underresourced, tribal Tibeto-Burman languages of India speoken in Eastern and North-Eastern parts of India.
- 1,476.7
- Hours
- 6
- Languages
- 345465
- Entries
- 2066
- Speakers
Project networkLeadership, partners and funders
Chief Investigator and Consortium LeaderProf. Bornini Lahiri
Principle InvestigatorDr. Meiraba Takhellambam
Principle InvestigatorDr. Amalesh Gope
Co-PIDr. Vivek Sheshadri
Principle InvestigatorManu Chopra
Chief Investigator and Consortium Leader (Former)Dr. Ritesh Kumar
Consortium LeaderIndian Institute of Technology Kharagpur
Consortium PartnerTezpur University
Funders / Funding agenciesMinistry of Electronics and Information Technology, Govt of India
Consortium PartnerManipur University
Technology PartnerUnreal Tece LLP
Consortium Partner (Former)Panlingua Language Processing LLP