DeepSeek-R1 and exploring DeepSeek-R1-Distill-Llama-8B
DeepSeek are the Chinese AI lab who dropped the best currently available open weights LLM on Christmas day, DeepSeek v3. That model was trained in part using their unreleased R1 …
DeepSeek-R1 and exploring DeepSeek-R1-Distill-Llama-8B Simon Willison’s Weblog Subscribe Sponsored by: Atlassian - Give your agents a plan. Not a prompt. New Jira capabilities unlock full-context for AI-native software development. Assign tasks to Claude, Cursor, or GitHub Copilot, now directly from Jira. Learn more DeepSeek-R1 and exploring DeepSeek-R1-Distill-Llama-8B 20th January 2025 DeepSeek are the Chinese AI lab who dropped the best currently available open weights LLM on Christmas day , DeepSeek v3. That model was trained in part using their unreleased R1 “reasoning” model. Today they’
Explore this link on the map →related reading
- deepseek-r1ollama.com
- The Complete Guide to DeepSeek Models: V3, R1, V4 and Beyondbentoml.com
- DeepSeek-R1arxiv.org
- GitHub - deepseek-ai/DeepSeek-R1 · GitHubgithub.com
- Dario Amodei — On DeepSeek and Export Controlsdarioamodei.com
- DeepSeek R1's recipe to replicate o1 and the future of reasoning LMsinterconnects.ai
- Understanding Reasoning LLMs - by Sebastian Raschka, PhDsebastianraschka.com
- DeepSeek-R1 Uncensored, QwQ-32B Puts Reasoning in Smaller Model, Phi-4-multimodal Takes Spoken Input, Training AI May Not Be Fair Useinfo.deeplearning.ai
- GitHub - Jiayi-Pan/TinyZero: Minimal reproduction of DeepSeek R1-Zero · GitHubgithub.com
- Inkling: Our Open-Weights Model - Thinking Machines Labthinkingmachines.ai
- Paper AI Tigersgleech.org
- Import AI 372: Gibberish jailbreak; DeepSeek's great new model; Google's soccer-playing robotsimportai.substack.com