Indian Flag
Government Of India
A-
A
A+

MiniGPT

A lightweight, 38-million parameter custom AI assistant built from scratch. Pre-trained on TinyStories and fine-tuned on Alpaca-Cleaned for basic conversation.

About Model

MyGPT Assistant (38M) A custom 38-million parameter Large Language Model built entirely from scratch in PyTorch. This model was created as an educational project to understand the complete lifecycle of LLM development—from defining the transformer architecture and dataset processing, all the way to exporting weights into GGUF format for Ollama. Training Details This model was trained in two distinct stages: Pre-training: Trained on a 50MB subset of the TinyStories dataset. This allowed the model to learn basic English grammar, syntax, and structural logic. Supervised Fine-Tuning (SFT): Fine-tuned on a subset of the Alpaca-Cleaned dataset to learn the User: / Assistant: conversational format and basic instruction following. Architecture Details: Parameters: ~38 Million Context Window: 256 tokens Layers: 6 Transformer Blocks Attention Heads: 6 Embedding Dimension: 384 Disclaimer Because this model is extremely small (38M parameters) and trained on children’s stories, it is meant purely for educational purposes and demonstrations of the LLM pipeline. It has very limited real-world knowledge and will heavily hallucinate facts.

MiniGPT

Metadata Metadata

Apache 2.0

Abhishek Singh

Transformers

PyTorch

Open

Science, Technology and Research

30/06/26 06:42:25

Abhishek Singh

116.14 MB

sweetspot_assistant.pt ( 116.14 MB )


To preview this file, you need to be a registered user. Please complete the registration process to gain access and continue viewing the content.

Activity Overview Activity Overview

  • Downloads0
  • Downloads 8
  • File Size 116.14 MB
  • Views 114

Tags Tags

  • skill development

License Control License Control

Apache 2.0

Version Control Version Control

FolderVersion 1(116.14 MB)
  • Abhishek Singh·3 month(s) ago
    • undefined
      sweetspot_assistant.pt

More Models from AIKosh Opensource Community More Models from AIKosh Opensource Community

NCERT-TUTOR-v3
NCERT-aligned tutor LoRA for on-device K-12 tutoring. Runs on 4 GB VRAM (RTX 3050). Stage-3 chained adapter: 15.7k curated questions (Grades 6-12, 13 subjects) + 3 DPO stages (math reasoning, mined factuality, multi-subject). 58% on held-out NCERT math with worked steps.
educational
qlora
low-resource
ncert
intelligent_tutoring_system
k12
  • See Upvoters0
  • Downloads3
  • File Size65.36 MB
  • Views90
Updated 13 day(s) ago

AIKOSH OPENSOURCE COMMUNITY

Kaveri-4B
Kaveri-4B is a multilingual text-generation model for conversational AI, instruction following, reasoning, coding assistance, summarization, translation, and general-purpose NLP tasks.
Conversational
safetensors
Text Generation
  • See Upvoters0
  • Downloads0
  • File Size7.51 GB
  • Views14
Updated 19 day(s) ago

AIKOSH OPENSOURCE COMMUNITY

Kaveri-1-5B
Kaveri-1.5B is a lightweight multilingual text-generation assistant by RiyaVibe for general chat, instruction following, basic question answering, simple coding help, basic math, translation, and local or cloud inference.
nlp
multilingual
Transformer
Text Generation
safetensors
Instruction Following
Conversational AI
  • See Upvoters0
  • Downloads0
  • File Size2.89 GB
  • Views7
Updated 19 day(s) ago

AIKOSH OPENSOURCE COMMUNITY

Kaveri-05b
Kaveri 0.5B is a lightweight multilingual conversational AI assistant designed for local inference, instruction following, chat, and agent workflows.
Instruction Following
Text Generation
Conversational AI
multilingual
  • See Upvoters0
  • Downloads0
  • File Size739.29 MB
  • Views16
Updated 21 day(s) ago

AIKOSH OPENSOURCE COMMUNITY

tusgan-v5
TUS-GAN v5 is a WGAN-GP-based generative model for synthesizing realistic time-use diary data, built to address incomplete records in India's Time Use Survey (ITUS). It generates 24-hour activity-location sequences conditioned on demographic and geographic attributes, trained on 445,268 diary records across 71 districts and 36 states.
behavior-change
tus
gan
wasserstein
wgan-gp
  • See Upvoters0
  • Downloads3
  • File Size41.64 KB
  • Views21
Updated 1 month(s) ago

AIKOSH OPENSOURCE COMMUNITY

Indic-mobile
Indic-mobile is a 0.5B parameter language model built completely from scratch — no fine-tuning, no adapter on top of an existing checkpoint. Every weight was pretrained from zero, purpose-built for all 22 officially recognized Indian languages and designed for efficient deployment on mobile and edge devices.
LLMs
gpt
Multimodal AI
Indian_languages
  • See Upvoters0
  • Downloads7
  • File Size953.22 MB
  • Views163
Updated 1 month(s) ago

AIKOSH OPENSOURCE COMMUNITY

SmolLM2-135M-Constitutional-ChitChat-v1-F16-GGUF
A lightweight SmolLM2 135M conversational language model fine tuned on the Constitutional AI ChitChat dataset for empathetic helpful and natural dialogue. Optimized for lightweight local conversational AI applications.
LLM
LLM Fine-tuning
Dialogue
Conversational AI
chatbot
  • See Upvoters0
  • Downloads4
  • File Size229.67 MB
  • Views115
Updated 2 month(s) ago

AIKOSH OPENSOURCE COMMUNITY

NEET-BioBERT
NEET-BioBERT is a fine-tuned version of DistilBERT (base uncased) specifically trained to classify the correct option for NEET-style multiple-choice biology questions. It selects the best answer among four choices (A, B, C, D).
safetensors
Pytorch
biology
AI in Education
distilbert
NEET
Transformer
  • See Upvoters1
  • Downloads6
  • File Size256.10 MB
  • Views116
Updated 2 month(s) ago

AIKOSH OPENSOURCE COMMUNITY

SKT-ST-X-0-3B
Model Overview :-- SKT 3B-MoE is a compact Small Language Model (SLM) built using ST-X-0 Taken Mixtral For Better MoE Stability , it delivers efficient and intelligent responses while maintaining a small footprint.
Mixture of Experts
moe
SKT AI LABS
slm
3b-model
  • See Upvoters1
  • Downloads3
  • File Size0
  • Views50
Updated 2 month(s) ago

AIKOSH OPENSOURCE COMMUNITY

DistilBERT-AI-Text-Detector
DistilBERT-AI-Text-Detector is a binary text classification model built on top of distilbert-base-uncased. It has been fine-tuned to distinguish between AI-generated and human-written text.
Transformer
Pytorch
safetensors
Text Classification
ai-text-detector
distilbert
  • See Upvoters1
  • Downloads6
  • File Size255.43 MB
  • Views63
Updated 2 month(s) ago

AIKOSH OPENSOURCE COMMUNITY