Selected Publications
The full listing can also be found on my Google Scholar Profile. But here you can find handy links to related material like blog posts and presentation slides.
Selected Publications John Schulman's Homepage Selected Publications The full listing can also be found on my Google Scholar Profile . But here you can find handy links to related material like blog posts and presentation slides. 2023 Let’s verify step by step Hunter Lightman, Vineet Kosaraju, Yura Burda, Harri Edwards, Bowen Baker, Teddy Lee, Jan Leike, John Schulman, Ilya Sutskever, Karl Cobbe Paper (arXiv) Scaling laws for single-agent reinforcement learning Jacob Hilton, Jie Tang, John Schulman Paper (arXiv) 2022 Scaling laws for reward model overoptimization Leo Gao, John Schulman, Jacob
saved by
related reading
- Key Papers in Deep RL - Spinning Up documentationspinningup.openai.com
- David Duvenaudcs.toronto.edu
- Andrew Gordon Wilsoncims.nyu.edu
- Deep RL Bootcamp - Lecturessites.google.com
- Explore | alphaXivalphaxiv.org
- Just Ask for Generalization | Eric Jangevjang.com
- State of RL for reasoning LLMs | A. Weersaweers.de
- LLM Resourcesforrestbicker.com
- Q-learning is not yet scalableseohong.me
- [2412.14135] Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspectivearxiv.org
- [1707.06347] Proximal Policy Optimization Algorithmsarxiv.org
- RL Scaling Laws for LLMs - by Cameron R. Wolfe, Ph.D.cameronrwolfe.substack.com