Fahim Tajwar
I am a PhD Student at the Machine Learning Department of Carnegie Mellon University. I am fortunate to be co-advised by Prof. Ruslan Salakhutdinov and Prof. Jeff Schneider. Previously, I obtained my BS (with distinction, Mathematics) in 2022 and MS (Computer Science) in 2023 from Stanford University. There I am grateful to have my research supervised by Prof. Chelsea Finn. I am also fortunate to have worked with Prof. Percy Liang, Prof. Stefano Ermon, and Prof. Stephen Luby during my time at Stanford. Feel free to reach out to me in case you have any questions or want to chat about my work! Email / CV / Google Scholar / Github / LinkedIn Website template
Selected Conference Publications Maximum Likelihood Reinforcement Learning Fahim Tajwar*, Guanning Zeng*, Yueer Zhou, Yuda Song, Daman Arora, Yiding Jiang, Jeff Schneider, Ruslan Salakhutdinov, Haiwen Feng, and Andrea Zanette International Conference on Machine Learning (ICML), 2026 (Oral) Workshop on Scaling Post-training for LLMs (SPOT) @ ICLR , 2026 (Best Paper Award) [Paper], [Code], [Project Website] Expanding the Capabilities of Reinforcement Learning via Text Feedback Yuda Song*, Lili Chen*, Fahim Tajwar, Rémi Munos, Deepak Pathak, Drew Bagnell, Aarti Singh, and Andrea Zanette…
saved by
related reading
- Jubayer Ibn Hamidjubayer-ibn-hamid.github.io
- Preston Fuprestonfu.com
- Cornell RL Research Seminarxikronz.github.io
- RLHF | John Lambertjohnwlambert.github.io
- State of RL for reasoning LLMs | A. Weersaweers.de
- RLHF & Post-Training Course by Nathan Lambertrlhfbook.com
- RLHF: Reinforcement Learning from Human Feedbackhuyenchip.com
- GenAI Handbookgenai-handbook.github.io
- Animals vs Ghosts – karpathykarpathy.bearblog.dev
- Chinmay Karkarchinmaykarkar.com
- Kushal Thamankushalthaman.github.io
- LLM Resourcesforrestbicker.com