[ D-01 ] Reasoning
Reasoning Models
Harder STEM problems where the frontier still fails.
We believe superior reasoning emerges from superior environments. Through more rigorous benchmarks and open training ecosystems, we are setting a new standard for how smarter AI is measured and built.
We split our work into four research domains and four methods. Benchmarks expose gaps. Data, environments, and training close them.
[ D-01 ] Reasoning
Harder STEM problems where the frontier still fails.
[ D-02 ] Embodied
Models that see, reason, and act on physical objects. We focus on unstructured environments.
[ D-03 ] Speech
Speech recognition, synthesis, and language understanding for Indian languages.
[ D-04 ] Software
Models that navigate full repos, use tools, interpret execution feedback, and produce verifiable fixes.
[ M-01 ] Evaluation
Evals that resist contamination and stay hard on purpose. We want to know where models stop working, not how close they are to a ceiling everyone already hit.
[ M-02 ] Data
We generate training data around the specific failure modes our benchmarks find.
[ M-03 ] Environments
Deterministic environments for coding in low-resource languages and domains where the frontier still breaks.
[ M-04 ] Training
RL and post-training on open-weight models, using the data we produce and the environments we build around them.
What we've shipped so far.
[ NeurIPS 2025 MATH-AI Workshop ] 7B reasoning model for JEE Main Math, built on Qwen 2.5 with on-policy curriculum SFT and GRPO
Read project log →GPT-OSS 20B post-trained with GRPO for JEE, NEET, and STEM.
Read project log →Multilingual Indic TTS model tuned for natural-sounding speech. Runs on edge hardware.
Read project log →RAG SLM for answering student doubts in vernacular.
Read project log →We fund researchers who build benchmarks that expose real weaknesses in frontier models. We co-develop the full stack from eval design through training, and provide compute.
Send a one-page proposal with your problem statement, approach, timeline, and compute needs. We review on a rolling basis.
Apply for a grant