Indian Flag
Government Of India
A-
A
A+

Indic-mobile

Indic-mobile is a 0.5B parameter language model built completely from scratch — no fine-tuning, no adapter on top of an existing checkpoint. Every weight was pretrained from zero, purpose-built for all 22 officially recognized Indian languages and designed for efficient deployment on mobile and edge devices.

About Model

Indic-mobile is a 0.5B parameter language model built completely from scratch — no fine-tuning, no adapter on top of an existing checkpoint. Every weight was pretrained from zero, purpose-built for all 22 officially recognized Indian languages and designed for efficient deployment on mobile and edge devices. Supported Languages Indic-mobile covers all 22 languages recognized under the 8th Schedule of the Indian Constitution: Assamese, Bengali, Bodo, Dogri, Gujarati, Hindi, Kannada, Kashmiri, Konkani, Maithili, Malayalam, Manipuri, Marathi, Nepali, Odia, Punjabi, Sanskrit, Santali, Sindhi, Tamil, Telugu, Urdu Why Indic-mobile? India has 1.4 billion people and 22 officially recognized languages — yet most language models were never built with this diversity in mind. Indic-mobile is designed to change that: Built from scratch — not a fine-tune or adapter on an existing English-centric model Truly multilingual — trained across all 22 Indian languages from the ground up Mobile-first — 0.5B parameters means it runs efficiently on edge devices and smartphones Open source — weights, architecture, and everything else, freely available Model Architecture Architecture: Custom (trained from scratch) Parameters: 0.5B Precision: BF16 Training: Pretrained from scratch (no base model used) Objective: Causal language modeling across 22 Indic languages Intended Uses Direct Use Text generation in any of the 22 official Indian languages Multilingual Indic chatbots and assistants On-device / mobile NLP applications Low-resource language research and experimentation Downstream Use Fine-tuning for specific Indic language tasks (classification, summarization, translation, QA) Integration into larger Indic NLP pipelines RAG (Retrieval-Augmented Generation) systems for Indian language content Out-of-Scope Use High-stakes decision making without human oversight Generation of harmful, misleading, or abusive content in any language Tasks requiring deep factual accuracy without verification Bias, Risks, and Limitations As a small 0.5B model, it may struggle with complex reasoning or long-form generation compared to larger models Training data distribution across all 22 languages may not be perfectly balanced; lower-resource languages may underperform Like all language models, it may reflect biases present in the training data Not intended for use in safety-critical or high-stakes applications without further evaluation and fine-tuning Recommendations Users should evaluate the model on their specific use case and language before deployment, particularly for lower-resource Indic languages.

Indic-mobile

Metadata Metadata

Apache 2.0

Purushottam Kumar

Multimodal Language Model

Transformers

Open

Social

26/07/26 12:40:25

953.22 MB

config.json ( 1.25 KB )


To preview this file, you need to be a registered user. Please complete the registration process to gain access and continue viewing the content.

Activity Overview Activity Overview

  • Downloads0
  • Downloads 1
  • File Size 953.22 MB
  • Views 62

Tags Tags

  • LLMs
  • gpt
  • Multimodal AI
  • Indian_languages

License Control License Control

Apache 2.0

Version Control Version Control

FolderVersion 1(953.22 MB)
  • Purushottam Kumar·19 day(s) ago
    • application/json
      config.json
    • application/json
      generation_config.json
    • undefined
      model.safetensors
    • application/json
      tokenizer_config.json
    • application/json
      tokenizer.json

More Models from AIKosh Opensource Community More Models from AIKosh Opensource Community

Indic-mobile
Indic-mobile is a 0.5B parameter language model built completely from scratch — no fine-tuning, no adapter on top of an existing checkpoint. Every weight was pretrained from zero, purpose-built for all 22 officially recognized Indian languages and designed for efficient deployment on mobile and edge devices.
gpt
Multimodal AI
Indian_languages
LLMs
  • See Upvoters0
  • Downloads1
  • File Size953.22 MB
  • Views63
Updated 2 day(s) ago

AIKOSH OPENSOURCE COMMUNITY

SmolLM2-135M-Constitutional-ChitChat-v1-F16-GGUF
A lightweight SmolLM2 135M conversational language model fine tuned on the Constitutional AI ChitChat dataset for empathetic helpful and natural dialogue. Optimized for lightweight local conversational AI applications.
Conversational AI
chatbot
Dialogue
LLM Fine-tuning
LLM
  • See Upvoters0
  • Downloads2
  • File Size229.67 MB
  • Views58
Updated 1 month(s) ago

AIKOSH OPENSOURCE COMMUNITY

NEET-BioBERT
NEET-BioBERT is a fine-tuned version of DistilBERT (base uncased) specifically trained to classify the correct option for NEET-style multiple-choice biology questions. It selects the best answer among four choices (A, B, C, D).
biology
AI in Education
distilbert
NEET
Transformer
Pytorch
safetensors
  • See Upvoters1
  • Downloads4
  • File Size256.10 MB
  • Views71
Updated 1 month(s) ago

AIKOSH OPENSOURCE COMMUNITY

SKT-ST-X-0-3B
Model Overview :-- SKT 3B-MoE is a compact Small Language Model (SLM) built using ST-X-0 Taken Mixtral For Better MoE Stability , it delivers efficient and intelligent responses while maintaining a small footprint.
moe
slm
SKT AI LABS
3b-model
Mixture of Experts
  • See Upvoters1
  • Downloads3
  • File Size0
  • Views38
Updated 1 month(s) ago

AIKOSH OPENSOURCE COMMUNITY

DistilBERT-AI-Text-Detector
DistilBERT-AI-Text-Detector is a binary text classification model built on top of distilbert-base-uncased. It has been fine-tuned to distinguish between AI-generated and human-written text.
ai-text-detector
distilbert
Transformer
Pytorch
Text Classification
safetensors
  • See Upvoters1
  • Downloads5
  • File Size255.43 MB
  • Views48
Updated 1 month(s) ago

AIKOSH OPENSOURCE COMMUNITY

MiniGPT
A lightweight, 38-million parameter custom AI assistant built from scratch. Pre-trained on TinyStories and fine-tuned on Alpaca-Cleaned for basic conversation.
skill development
  • See Upvoters0
  • Downloads5
  • File Size116.14 MB
  • Views90
Updated 1 month(s) ago

AIKOSH OPENSOURCE COMMUNITY