The Thinking Machines Tinker API is good news for AI control and security — LessWrong
Last week, Thinking Machines announced Tinker. It’s an API for running fine-tuning and inference on open-source LLMs that works in a unique way. I think it has some immediate practical implications for AI safety research: I suspect that it will make RL experiments substantially easier, and increase the number of safety papers that involve RL on big models. But it's more interesting to me for another reason: the design of this API makes it possible to do many types of ML research without direct access to the model you’re working with. APIs like this might allow AI companies to reduce how many of their researchers (either human or AI) have access to sensitive model weights, which is good for reducing the probability of weight exfiltration and other rogue deployments. Learning about the design of this API and observing that Thinking Machines was actually able to implement it, has made me moderately more optimistic about mitigating insider threat from researchers (human or AI agents) doing
x The Thinking Machines Tinker API is good news for AI control and security — LessWrong AI Control Computer Security & Cryptography AI Frontpage 92 The Thinking Machines Tinker API is good news for AI control and security by Buck 9th Oct 2025 AI Alignment Forum 8 min read 10 92 Ω 37 Last week, Thinking Machines announced Tinker . It’s an API for running fine-tuning and inference on open-source LLMs that works in a unique way. I think it has some immediate practical implications for AI safety research: I suspect that it will make RL experiments substantially easier, and increase the number of s
Explore this link on the map →related reading
- Anatomy of a Modern Finetuning APIbenanderson.work
- AI 2027ai-2027.com
- Inkling: Our Open-Weights Model - Thinking Machines Labthinkingmachines.ai
- AI in 2025: gestalt — LessWronglesswrong.com
- A basic systems architecture for AI agents that do autonomous research — LessWronglesswrong.com
- Tinker: Call for Community Projects - Thinking Machines Labthinkingmachines.ai
- [2312.06942] AI Control: Improving Safety Despite Intentional Subversionarxiv.org
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- A Safe Path to Open Weights - Thinking Machines Labthinkingmachines.ai
- What I learned this week - Can distillation be stopped, Mythos and the cybersecurity equilibrium, Pipeline RLdwarkesh.com
- How fast is AI improving? - AI Digesttheaidigest.org