NVIDIA-NeMo/Automodel: 🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support ·
github.com · 3,177 words · saved by 1 readers
🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support
📣 News and Discussions [09/11/2026]DeepSeek-V4.1-Flash We now support SFT and LoRA for DeepSeek's causal encoder/decoder backbone with CSA2, single-pass mHC, and Engram, many thanks to @Khazic. Check out the HellaSwag EP64 16-node recipe and model coverage page. [08/29/2026]GLM-5.3 We now support full-parameter fine-tuning of Z.ai's GLM-5.3 with cuDNN DSA and HybridEP. Check out the EP64/PP4 recipe and model coverage page. [08/28/2026]Ox Alpha / GLM-5.3-Flash We now support fine-tuning Z.ai's 320B-A18B hybrid-attention MoE VLM with packed CP and EP. Check out the MedPix EP72/CP2 recipe…
saved by
related reading
- Tinkerthinkingmachines.ai
- Together AI | The AI Native Cloudtogether.ai
- GitHub - rasbt/LLMs-from-scratch: Implement a ChatGPT-like LLM in PyTorch from scratch, step by stepgithub.com
- GitHub - huggingface/nanotron: Minimalistic large language model 3D-parallelism training · GitHubgithub.com
- GitHub - karpathy/nanochat: The best ChatGPT that $100 can buy.github.com
- The Llama Hitchiking Guide to Local LLMs – hackerllamaosanseviero.github.io
- Hugging Face – The AI community building the future.huggingface.co
- Llama 2 · Hugging Facehuggingface.co
- PostTrainBenchposttrainbench.com
- Finetune LLMs on your own consumer hardware using tools from PyTorch and Hugging Face ecosystem – PyTorchpytorch.org
- GitHub - linkedin/Liger-Kernel: Efficient Triton Kernels for LLM Traininggithub.com
- Efficient LLM Finetuning with Unsloth | Modal Docsmodal.com