From the source
Explorations in Post Training and Inferencing Optimizations for a Hybrid Indic LLM Download the model from Hugging Face , try it on our playground , and build with our APIs .
Towards Building a Sovereign AI Ecosystem in India , we plan to have regular model drops and share our detailed technical findings.
This is the first in this series of technical blogs, wherein we share our findings on post-training.
We look forward to hearing feedback and suggestions for collaborations.
In this blog, we share our explorations on post-training and inference optimization of an open pre-trained model to create a cutting-edge hybrid reasoning model specialized for Indic languages.
The blog spans three broad steps: (i) supervised fine-tuning (SFT), (ii) reinforcement learning with verifiable rewards (RLVR), and (iii) inference optimizations.
…




