nemotron-3-super-120b-a12b Model by NVIDIA | NVIDIA NIM
build.nvidia.com · 6,357 words · saved by 1 readers
Open, efficient hybrid Mamba-Transformer MoE with 1M context, excelling in agentic reasoning, coding, planning, tool calling, and more
Skip to main content nemotron-3-super-120b-a12b Model by NVIDIA | NVIDIA NIM Overview Overview ModelCard++ ModelCard++ Jump to a topic Model Summary Quick Start Model Overview What is Nemotron? Description License/Terms of Use Benchmarks Deployment Geography: Global Use Case Release Date Reference(s) Model Architecture Model Design Training Methodology Input Output Software Integration Model Version(s) Quick Start Guide API Client OpenCode Training and Evaluation Datasets Public Datasets Crawled and Scraped from Online Sources by NVIDIA Private Non-publicly Accessible Datasets of Third Parties
related reading
- Together AI | The AI Native Cloudtogether.ai
- AI Model & API Providers Analysis | Artificial Analysisartificialanalysis.ai
- Parsed | Custom, interpretable AI systems that continuously learnparsed.com
- Datacurve | The data engine for frontier AIdatacurve.ai
- Neuronpedianeuronpedia.org
- Hugging Face – The AI community building the future.huggingface.co
- GitHub - brexhq/prompt-engineering: Tips and tricks for working with Large Language Models like OpenAI's GPT-4.github.com
- Compare AI Models: Pricing, Context & Benchmarks | OpenRouteropenrouter.ai
- PostTrainBenchposttrainbench.com
- LangChain: the open agent platform to own your intelligencelangchain.com
- Goodfire AIgoodfire.ai
- GitHub - huggingface/nanotron: Minimalistic large language model 3D-parallelism training · GitHubgithub.com