✳flâneur — a map of the web's best reading
MordatchNIPS15.pdf
roboti.us · 5,367 words · saved by 1 readers
N/A
# link_1nli3k5az74.pdf ## Metadata - PDFFormatVersion=1.5 - IsLinearized=true - IsAcroFormPresent=false - IsXFAPresent=false - IsCollectionPresent=false - IsSignaturesPresent=false - CreationDate=D:20151014014055Z - Creator=LaTeX with hyperref package - ModDate=D:20151219172547-08'00' - Custom.PTEX.Fullbanner=This is pdfTeX, Version 3.14159265-2.6-1.40.15 (TeX Live 2014) kpathsea version 6.2.0 - Producer=pdfTeX-1.40.15 - Trapped=False - dc:format=application/pdf - xmp:createdate=2015-10-14T01:40:55Z - xmp:creatortool=LaTeX with hyperref package - xmp:modifydate=2015-12-19T17:25:47-08:00 - xmp:
Explore this link on the map →saved by
related reading
- [1811.01848] Plan Online, Learn Offline: Efficient Learning and Exploration via Model-Based Controlarxiv.org
- Natural Intelligencegreydanus.github.io
- The 37 Implementation Details of Proximal Policy Optimization · The ICLR Blog Trackiclr-blog-track.github.io
- pdfopenreview.net
- Learning Beyond Gradientstrinkle23897.github.io
- The World Inside Neural Networksgoodfire.ai
- Underactuated Roboticsunderactuated.mit.edu
- State of Robot Learning, December 2025vedder.io
- FACTR: Force-Attending Curriculum Training for Contact-Rich Policy Learningarxiv.org
- Ch. 21 - Imitation Learningunderactuated.mit.edu
- [1707.06347] Proximal Policy Optimization Algorithmsarxiv.org
- Ch. 10 - Trajectory Optimizationunderactuated.csail.mit.edu