flâneur — a map of the web's best reading

Scaling Laws for Value-Based RL

value-scaling.github.io · 4,597 words · saved by 1 readers

With the right design decisions, value-based RL admits predictable scaling.

Scaling Laws for Value-Based RL Compute-Optimal Scaling for Value-Based Deep RL August 2025 Preston Fu *, Oleh Rybkin *, Zhiyuan Zhou , Michal Nauman , Pieter Abbeel , Sergey Levine , Aviral Kumar NeurIPS , 2025 arXiv Code Thread Poster Value-Based Deep RL Scales Predictably February 2025 Oleh Rybkin , Michal Nauman , Preston Fu , Charlie Snell , Pieter Abbeel , Sergey Levine , Aviral Kumar ICML , 2025 ICLR Robot Learning Workshop , 2025 ( oral ) arXiv Code Thread Poster In the era of large-scale AI, it is important to prototype new training methodologies at small scales before running at larg

Explore this link on the map →

saved by

related reading