flâneur

Chinmay Karkar

chinmaykarkar.com · 240 words · saved by 2 readers

I work on the science and craft of training language models — what makes them learn, what makes them stable over long horizons, and how to tell whether they’re actually getting better.

researcher · pre-training & post-training of language models currently at Microsoft Research India · previously Lossfunk & Athena Agents (RL) about I work on the science and craft of training language models — what makes them learn, what makes them stable over long horizons, and how to tell whether they’re actually getting better. At Microsoft Research India I’m focused on reinforcement learning for LLMs and agentic verifier modules for multi-step reasoning evaluation. Before that, probabilistic forecasting at Lossfunk, model merging & RL post-training at Athena, and a fine-tuning /…

saved by

related reading