Learning Sim-to-Real Humanoid Locomotion in 15 Minutes
younggyo.me · 354 words · saved by 1 readers
Learning Sim-to-Real Humanoid Locomotion in 15 Minutes
Amazon FAR (Frontier AI & Robotics) * Equal contribution We provide a simple recipe with FastSAC and FastTD3 for rapid sim-to-real humanoid iterations Abstract Massively parallel simulation has reduced reinforcement learning (RL) training time for robots from days to minutes. However, achieving fast and reliable sim-to-real RL for humanoid control remains difficult due to the challenges introduced by factors such as high dimensionality and domain randomization. In this work, we introduce a simple and practical recipe based on off-policy RL algorithms, i.e., FastSAC and FastTD3, that…
saved by
related reading
- Learning to Walk in Minutes Using Massively Parallel Deep Reinforcement Learningarxiv.org
- State of Robot Learning, December 2025vedder.io
- Emergent Dexterity via Diverse Resets and Large-Scale Reinforcement Learningarxiv.org
- How Claude Performs on Robotics Tasks \ Anthropicanthropic.com
- Stop Simulating, Start Experiencingpaoloai.substack.com
- Precise Manipulation with Efficient Online RLpi.website
- Building Worlds That Train Robotsworldlabs.ai
- Learning dexterity | OpenAIopenai.com
- GitHub - adam-maj/robotics: A deep dive on the history of robotics and the future of humanoidsgithub.com
- BeyondMimic: From Motion Tracking to Versatile Humanoid Control via Guided Diffusionarxiv.org
- Ch. 21 - Imitation Learningunderactuated.mit.edu
- Helix: A Vision-Language-Action Model for Generalist Humanoid Controlfigure.ai