[D] OpenAI Rubik’s cube hype : r/MachineLearning
Beginners -> /r/mlquestions or /r/learnmachinelearning , AGI -> /r/singularity, career advices -> /r/cscareerquestions, datasets -> r/datasets Following the Rubik’s cube video and paper, seems OpenAI is doing their usual business: Brute force search/learning with $$$ GPU hours, viral a video and PR, call it AGI “progress”. I’ve seen better previous works, in sample efficiency and/or intepretability: DeLaNet, Generalize robots as learnable Lagrangian, able to derive and seperate robot kinematics and external forces, e.g. Friction and Coriolis forces. https://arxiv.org/abs/1907.04489 https://openreview.net/forum?id=BklHpjCqKm TossingBot, Residual dynamics model learning, able to near-online adapt and generalize to arbitrary-weight objects throwing to targets with perturbations, https://arxiv.org/abs/1903.11239 https://ai.googleblog.com/2019/03/unifying-physics-and-deep-learning-with.html One may argue OpenAI did prior-less learning, but their sample efficiency is big problem (adaptive ra
Reddit - Please wait for verification
Explore this link on the map →related reading
- Let's Discuss OpenAI's Rubik's Cube Resultalexirpan.com
- AI 2027ai-2027.com
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- State of Robot Learning, December 2025vedder.io
- AI 2027ai-2027.com
- Juergen Schmidhuber's home page - Universal Artificial Intelligence - AI - Deep Learning - Recurrent Neural Networks - Computer Vision - Object Detection - Image segmentation - GANs - Transformers with linearized self-attention - Goedel Macpeople.idsia.ch
- Just Ask for Generalization | Eric Jangevjang.com
- Explore | alphaXivalphaxiv.org
- DeepMind: Generally capable agents emerge from open-ended play — LessWronglesswrong.com
- GenAI Handbookgenai-handbook.github.io
- how we accidentally solved robotics by watching 1 million hours of YouTube – atharva's blogksagar.bearblog.dev
- Pulkit Agrawalpeople.csail.mit.edu