A 272.7M-parameter language model, pretrained from scratch on 11 Indic languages + English (Sangraha corpus), then fine-tuned on grounded multilingual QA for the Indian government-schemes / financial-banking domain (PM-KISAN, Ayushman Bharat, banking products, insurance, savings instruments, etc.). Updated in place — this repo tracks the current best domain checkpoint, not a fixed snapshot; check back for updates as fine-tuning improves. Built from custom composable primitives, structurally equivalent to Qwen3 (confirmed by direct source comparison during HF conversion) and saved in that format for standard transformers loading: tokens/phase across H100 and V100 GPUs (best validation loss…