✳flâneur — a map of the web's best reading
Muon Outperforms Adam in Tail-End Associative Memory Learning
arxiv.org · saved by 1 readers
N/A
Explore this link on the map →related reading
- Deriving Muonjeremybernste.in
- NL.pdfabehrouz.github.io
- GitHub - jacobhilton/deep_learning_curriculum: Language model alignment-focused deep learning curriculum · GitHubgithub.com
- GitHub - mem0ai/mem0: Universal memory layer for AI Agents · GitHubgithub.com
- GitHub - inverse-scaling/prize: A prize for finding tasks that cause large language models to show inverse scaling · GitHubgithub.com
- [2501.00663] Titans: Learning to Memorize at Test Timearxiv.org
- GitHub - google-research/tuning_playbook: A playbook for systematically maximizing the performance of deep learning models. · GitHubgithub.com
- Does Muon improve regulatory DNA learning? Part 1. — Origin Bioorigin.bio
- Learning to be Bayesian without Supervisionpapers.nips.cc
- perceptron learning algo, max margin classifierspeople.eecs.berkeley.edu
- Jacobian Lens – Qwen3.6-27B | Neuronpedianeuronpedia.org
- GitHub - getmetal/motorhead: 🧠 Motorhead is a memory and information retrieval server for LLMs. · GitHubgithub.com