Selected Publications
The full listing can also be found on my Google Scholar Profile. But here you can find handy links to related material like blog posts and presentation slides.
Selected Publications John Schulman's Homepage Selected Publications The full listing can also be found on my Google Scholar Profile . But here you can find handy links to related material like blog posts and presentation slides. 2023 Let’s verify step by step Hunter Lightman, Vineet Kosaraju, Yura Burda, Harri Edwards, Bowen Baker, Teddy Lee, Jan Leike, John Schulman, Ilya Sutskever, Karl Cobbe Paper (arXiv) Scaling laws for single-agent reinforcement learning Jacob Hilton, Jie Tang, John Schulman Paper (arXiv) 2022 Scaling laws for reward model overoptimization Leo Gao, John Schulman, Jacob
Explore this link on the map →saved by
related reading
- Key Papers in Deep RL - Spinning Up documentationspinningup.openai.com
- Metalearning or Learning to Learn Since 1987people.idsia.ch
- Explore | alphaXivalphaxiv.org
- Just Ask for Generalization | Eric Jangevjang.com
- State of RL for reasoning LLMs | A. Weersaweers.de
- Q-learning is not yet scalableseohong.me
- LLM Resourcesforrestbicker.com
- RL Scaling Laws for LLMs - by Cameron R. Wolfe, Ph.D.cameronrwolfe.substack.com
- Learning Beyond Gradientstrinkle23897.github.io
- [1707.06347] Proximal Policy Optimization Algorithmsarxiv.org
- [1707.06347] Proximal Policy Optimization Algorithmsarxiv.org
- IsoCompute Playbook: Optimally Scaling Sampling Compute for RL Training of LLMscompute-optimal-rl-llm-scaling.github.io