Indian Flag
Government Of India
A-
A
A+

AI4Bharat-IndicConformer-STT-MR-Hybrid-CTC-RNNT-Large (Marathi): Automatic Speech Recognition Model

A large-scale Automatic Speech Recognition Model (ASR) model for Marathi, utilizing a hybrid CTC-RNNT decoder.

About Model

The ai4bharat/indicconformer_stt_mr_hybrid_ctc_rnnt_large model is an Automatic Speech Recognition (ASR) system designed for the Marathi language. It employs a Conformer-Large architecture with 120 million parameters, featuring 17 conformer blocks and a model dimension of 512. This model processes 16 kHz mono-channel audio (wav files) and outputs transcriptions in Marathi. Its hybrid CTC-RNNT decoder enhances recognition performance for spoken Marathi.

AI4Bharat-IndicConformer-STT-MR-Hybrid-CTC-RNNT-Large (Marathi): Automatic Speech Recognition Model

Metadata Metadata

MIT

AI4Bharat

Audio-to-text

N.A.

Open

AI4Bharat

Sector Agnostic

21/02/25 13:21:35

0

Activity Overview Activity Overview

  • Downloads0
  • Redirect 28
  • Views 415
  • File Size 0

Tags Tags

  • Automatic Speech Recognition
  • Marathi
  • Conformer
  • Hybrid-CTC-RNNT
  • ASR
  • Speech-to-Text
  • Deep Learning
  • Neural Networks
  • Speech Recognition
  • Indic Languages

License Control License Control

MIT

More Models from AI4Bharat More Models from AI4Bharat

AI4Bharat- 500 M - RomanSetu Multilingual Native-to-Roman Model
RomanSetu is a multilingual continual pretrained transformer model designed for transliteration across six Indic languages
Instruction-Tuning
LLaMA2
Multilingual
Llama
  • See Upvoters1
  • Downloads38
  • File Size0
  • Views633
Updated 8 month(s) ago

AI4BHARAT

AI4Bharat- 400 M - RomanSetu Multilingual Native-to-Roman Model
RomanSetu is a multilingual continual pretrained transformer model designed for transliteration across six Indic languages
Multilingual
Llama
Instruction-Tuning
LLaMA2
  • See Upvoters1
  • Downloads67
  • File Size0
  • Views779
Updated 8 month(s) ago

AI4BHARAT

AI4Bharat- Maithili - IndicConformer Automatic Speech Recognition (ASR) Model
This model takes in mono-channel audio files at a 16,000 Hz sampling rate (WAV format) and outputs the transcribed text of the speech contained in the audio.
Automatic Speech Recognition
Speech-to-Text
NLP
  • See Upvoters0
  • Downloads21
  • File Size0
  • Views481
Updated 8 month(s) ago

AI4BHARAT

AI4Bharat- Konkani - IndicConformer Automatic Speech Recognition (ASR) Model
Automatic Speech Recognition (ASR) model for Konkani speech recognition, processing 16,000 KHz mono WAV audio and transcribing spoken content into text
Speech-to-Text
NLP
Automatic Speech Recognition
  • See Upvoters0
  • Downloads24
  • File Size0
  • Views512
Updated 8 month(s) ago

AI4BHARAT

AI4Bharat- Kashmiri - IndicConformer Automatic Speech Recognition (ASR) Model
This Automatic Speech Recognition (ASR) model transcribes Kashmiri speech from 16,000 KHz mono WAV audio files into text
Kashmiri
Speech-to-Text
NLP
Automatic Speech Recognition
  • See Upvoters0
  • Downloads17
  • File Size0
  • Views508
Updated 8 month(s) ago

AI4BHARAT

AI4Bharat - Romansetu-200M -Multilingual LLM for Indian langauges using romanization
RomanSetu is Efficiently unlocking multilingual (Indian Languages) capabilities of Large Language Models via Romanization.
Instruction-Tuning
LLaMA2
Llama
Multilingual
  • See Upvoters0
  • Downloads4
  • File Size0
  • Views197
Updated 8 month(s) ago

AI4BHARAT

AI4Bharat - Romansetu-100M - Multilingual LLM for Indian langauges using romanization
RomanSetu is Efficiently unlocking multilingual (Indian Languages) capabilities of Large Language Models via Romanization.
Llama
Multilingual
Instruction-Tuning
LLaMA2
  • See Upvoters0
  • Downloads8
  • File Size0
  • Views314
Updated 8 month(s) ago

AI4BHARAT

AI4Bharat- Kannada - IndicConformer Automatic Speech Recognition (ASR) Model
This Kannada Automatic Speech Recognition (ASR) model transcribes 16kHz mono-channel audio into text. It utilizes a Conformer-Large architecture with 120M parameters and a hybrid CTC-RNNT decoder for high-accuracy speech recognition.
Automatic Speech Recognition
Audio Processing
NLP
  • See Upvoters0
  • Downloads18
  • File Size0
  • Views471
Updated 8 month(s) ago

AI4BHARAT

AI4Bharat – Romanized Path – Base to Supervised Fine-Tuning (SFT)
Romansetu model is built on base pretrained model which is supervised fine tuned on instuction-following tasks using romanized Indian languages.
LLaMA2
Instruction-Tuning
Multilingual
Llama
  • See Upvoters0
  • Downloads3
  • File Size0
  • Views141
Updated 8 month(s) ago

AI4BHARAT

AI4Bharat-IndicTrans2 Large-1B -English-to-Hindi (Devanagari) – : Language Translation Model
A large-scale neural machine translation (NMT) model for translating English to Hindi (Devanagari) language, leveraging 1 billion parameters for high-quality translations.
Machine Translation
Transformer
low-resource-NLP
high-quality-translation
Large Model
cross-lingual
NLP
Multilingual
  • See Upvoters0
  • Downloads25
  • File Size0
  • Views577
Updated 8 month(s) ago

AI4BHARAT