Indian Flag
Government Of India
A-
A
A+
ORGANISATION

Indic-Conformer model for ASR

Indo-Aryan Indic-Conformer is a multilingual speech model for North-Indian languages. This model is based on Conformer large architecture, with 115M parameters.

About Model

Bhashini - The Indo-Aryan Indic-Conformer is a multilingual automatic speech recognition (ASR) model designed specifically for North-Indian languages. It is based on the Conformer large architecture, which is known for its efficiency and accuracy in processing speech signals. The model contains 115 million parameters, enabling it to effectively transcribe spoken language into text with high precision.

This ASR model has been trained on the Shrutlip dataset, a rich dataset designed to enhance automatic speech recognition capabilities in Indian languages. The model primarily supports the Odia language and has been developed by AI4Bharat, a leading research initiative focused on advancing AI-driven solutions for Indian languages.

With a batch processing setup, this model is optimized for large-scale speech-to-text tasks across general domains. It is a valuable resource for applications in speech transcription, voice-enabled interfaces, digital accessibility, and natural language processing (NLP) research. Given the increasing demand for multilingual ASR systems, this model serves as a foundational tool for improving speech technology in India’s diverse linguistic landscape.

The Indo-Aryan Indic-Conformer is open-source, and its implementation is available on GitHub, making it accessible for researchers, developers, and AI practitioners working in the domain of Indian language speech processing.

For more details about the use of model, refer to github: https://github.com/AI4Bharat/IndicTrans2/tree/main

Indic-Conformer model for ASR

Metadata Metadata

MIT

AI4Bharat

Speech Recognition Model

Other

Open

Sector Agnostic

05/03/25 15:23:44

Admin

64.91 KB

indic-asr-api-backend-master ( 7 files, 1 directories )


Directory
serving

2 files, 7 directories

undefined
.gitignore

1.76 KB

undefined
api.py

4.95 KB

application/json
conformer.json

211 Bytes

undefined
example_ai4b_asr_rest_api.py

1.67 KB

undefined
LICENSE

1.04 KB

text/markdown
README.md

1.26 KB

text/plain
requirements.txt

66 Bytes

Activity Overview Activity Overview

  • Downloads2
  • Downloads 101
  • File Size 64.91 KB
  • Views 3,385

Tags Tags

  • Automatic Speech Recognition
  • Speech Technology
  • Speech Processing
  • Speech Lab
  • Bhashini

License Control License Control

MIT

Version Control Version Control

FolderVersion 1(64.91 KB)
  • admin·1 year(s) ago
    • chevron_rightFolder
      indic-asr-api-backend-master
      • chevron_rightFolder
        serving
      • undefined
        .gitignore
      • undefined
        api.py
      • application/json
        conformer.json
      • undefined
        example_ai4b_asr_rest_api.py
      • undefined
        LICENSE
      • text/markdown
        README.md
      • text/plain
        requirements.txt

More Models from TechCorp More Models from TechCorp

SPRING-INX-DATA2VEC-AQC-GUJARATI
Automatic Speech Recognition (ASR) model for speech recognition, processing audio and transcribing spoken content into text.The inference code, installation requirements, and usage instructions are available in the SPRING Lab, IIT Madras GitHub repository: https://github.com/Speech-Lab-IITM/Fairseq-Inference
ssl
Low-resource languages
SSL_finetunning
Data2vec_aqc
spring_lab
IITM
gujarati
  • See Upvoters0
  • Downloads2
  • File Size3.52 GB
  • Views31
Updated 1 month(s) ago

DIGITAL INDIA BHASHINI DIVISION

SPRING-INX-DATA2VEC-AQC-HINDI
Automatic Speech Recognition (ASR) model for speech recognition, processing audio and transcribing spoken content into text.The inference code, installation requirements, and usage instructions are available in the SPRING Lab, IIT Madras GitHub repository: https://github.com/Speech-Lab-IITM/Fairseq-Inference
Low-resource languages
ssl
IITM
spring_lab
Data2vec_aqc
SSL_finetunning
hindi
  • See Upvoters0
  • Downloads1
  • File Size3.53 GB
  • Views31
Updated 1 month(s) ago

DIGITAL INDIA BHASHINI DIVISION

SPRING-INX-DATA2VEC-AQC-MANIPURI
Automatic Speech Recognition (ASR) model for speech recognition, processing audio and transcribing spoken content into text.The inference code, installation requirements, and usage instructions are available in the SPRING Lab, IIT Madras GitHub repository: https://github.com/Speech-Lab-IITM/Fairseq-Inference
Manipuri
Low Resource Languages
SSL_finetunning
Data2vec_aqc
spring_lab
IITM
ssl
  • See Upvoters0
  • Downloads1
  • File Size3.52 GB
  • Views25
Updated 1 month(s) ago

DIGITAL INDIA BHASHINI DIVISION

SPRING-INX-DATA2VEC-AQC-ASSAMESE
Automatic Speech Recognition (ASR) model for speech recognition, processing audio and transcribing spoken content into text.The inference code, installation requirements, and usage instructions are available in the SPRING Lab, IIT Madras GitHub repository: https://github.com/Speech-Lab-IITM/Fairseq-Inference
SSL_finetunning
Assamese
ssl
IITM
spring_lab
Data2vec_aqc
Low-resource languages
  • See Upvoters0
  • Downloads1
  • File Size3.52 GB
  • Views17
Updated 1 month(s) ago

DIGITAL INDIA BHASHINI DIVISION

IndicXlit
A Transformer-based multilingual transliteration model
NLP
Language Modeling
Multilingual Translation
Machine Translation
Regional Languages
Indian Languages
transliteration
  • See Upvoters0
  • Downloads64
  • File Size3.94 MB
  • Views1,408
Updated 1 month(s) ago

DIGITAL INDIA BHASHINI DIVISION

Indic Trans2
AI4Bharat's Indic-Trans-v2 is a multilingual Transformer (~1.1BM) NMT model trained on Samanantar v2 dataset which is the largest publicly available parallel corpora collection for languages of India at the time of writing (23 March 2023). We currently release two models - Indic to English and English to Indic and support all the 22 scheduled languages of India.
Regional Languages
NLP
Indic-TransV2
Indian Languages
Machine Translation
Multilingual Translation
Bilingual Translation
Language Modeling
Computational Linguistics
Machine Translation
  • See Upvoters1
  • Downloads95
  • File Size214.60 KB
  • Views2,853
Updated 1 month(s) ago

DIGITAL INDIA BHASHINI DIVISION

Bhashini - Fastspeech2 Model using (HS)
Text-to-speech models trained using FastPitch and HiFi-GAN vocoder, separately for each language. Supports both 'female' and 'male' voices.
Transformer
Text to Speech
Language Detection
Multilingual
NLP
Text Processing
  • See Upvoters0
  • Downloads112
  • File Size286.72 MB
  • Views2,186
Updated 1 month(s) ago

DIGITAL INDIA BHASHINI DIVISION

Bhashini - IndicNER
IndicNER is a multilingual Named Entity Recognition model fine-tuned on 11 Indian languages to identify named entities in text
NLP
Foreigners
Multilingual
Transformer
Token Classification
Pytorch
Samanantar
Bert
NER
  • See Upvoters2
  • Downloads193
  • File Size591.28 MB
  • Views2,972
Updated 1 month(s) ago

DIGITAL INDIA BHASHINI DIVISION

Bhashini-AI4Bharat Textual Language Detection v1.0
Detect language from provided text, Currently supports 23 languages (English, Bangla, Manipuri, Bodo, Konkani, Oriya, Nepali, Marathi, Sindhi, Sanskrit, Malayalam, Urdu, Assamese, Telugu, Dogri, Gujarati, Kashmiri, Punjabi, Santali, Maithili, Hindi, Tamil, Kannada)
NLP
Text Processing
Deep Learning
Transformer
Text Language Detection
Multilingual
AI4Bharat
Bhashini
  • See Upvoters5
  • Downloads291
  • File Size3 MB
  • Views5,656
Updated 1 month(s) ago

DIGITAL INDIA BHASHINI DIVISION

SPRING-INX-DATA2VEC-AQC-SANSKRIT
Automatic Speech Recognition (ASR) model for speech recognition, processing audio and transcribing spoken content into text. The inference code, installation requirements, and usage instructions are available in the SPRING Lab, IIT Madras GitHub repository: https://github.com/Speech-Lab-IITM/Fairseq-Inference
low-resource-language
SSL_finetunning
Data2vec_aqc
spring_lab
IITM
ssl
Sanskrit
  • See Upvoters0
  • Downloads5
  • File Size3.52 GB
  • Views224
Updated 1 month(s) ago

DIGITAL INDIA BHASHINI DIVISION