✳flâneur — a map of the web's best reading
Macaron-V1-Preview: 749B MoL Agent Model post-trained from GLM5.1
macaron.im · 3,080 words · saved by 1 readers
Macaron-V1-Preview
Macaron-V1-Preview: 749B MoL Agent Model post-trained from GLM5.1 We introduce Macaron-V1-Preview , a 749B (744B base + 5 × 1B LoRA) Agent Model post-trained from GLM5.1 with MinT . Macaron-V1-Preview leverages a novel Mixture-of-LoRA (MoL) architecture to provide a scalable, resource-efficient foundation for advanced agentic use cases in general life scenarios. By orthogonalizing the parameter space through the MoL architecture, Macaron-V1-Preview reduces the optimization interference that often appears when one model is trained for diverse agentic capabilities. Each capability is optimized i
Explore this link on the map →saved by
related reading
- Adina Yakup on X: "Macaron-V1-Preview-749B 👀 a Mixture-of-LoRA personal agent model from MindLab ✨ 744B base + 5 specialist LoRAs ✨ Generative UI as a core skill ✨ Personal agent focused ✨ 202K context ✨ MIT license https://t.co/OkjxKEThzZx.com
- Composer2.pdfcursor.com
- LLM Powered Autonomous Agents | Lil'Loglilianweng.github.io
- Inkling: Our Open-Weights Model - Thinking Machines Labthinkingmachines.ai
- Demystifying evals for AI agents \ Anthropicanthropic.com
- Arjun Virkarjunvirk.com
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- Agent Evaluation: A Detailed Guidecameronrwolfe.substack.com
- llm assistant personas seem increasingly incoherent (some subjective observations) — LessWronglesswrong.com
- Building reliable AI agents · parth sareenparthsareen.com
- Explore | alphaXivalphaxiv.org
- PostTrainBenchposttrainbench.com