NVIDIA Nemotron

NVIDIA's open model family shipping weights, datasets, and full training recipes for agentic systems.

Open SourceSelf HostedOffline CapableGPU Required
0.0 (0)

About

Nemotron is NVIDIA's umbrella for open models aimed at agentic AI, and it is unusual in shipping not just weights but the datasets and step-by-step training recipes behind them. The Nemotron 3 generation uses a hybrid Mamba-transformer mixture-of-experts backbone and spans Nano (31.6B total, 3.6B active) for edge and PC deployment, Nano Omni (30B/3B) for multimodal input, Super (120.6B/12.7B) for single-GPU throughput, and Ultra at 550B/55B for multi-GPU datacenters, with context windows reaching 1 million tokens; earlier Llama-Nemotron variants remain available. The GitHub repository collects cookbooks, deployment guides, and end-to-end examples under Apache-2.0, while model weights use the NVIDIA Open Model License, which permits commercial use, modification, redistribution, and hosted services without royalties. Pre-training, post-training, and reinforcement learning datasets are published openly on Hugging Face, and official documentation lives at docs.nvidia.com/nemotron. For teams standardizing on NVIDIA hardware the family offers a complete open pipeline from data to deployed agent.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Intermediate (3/5)
License
NVIDIA Open Model License
Added
Jul 29, 2026

Related Tools

Featured

Lightweight open-weight LLM by Google available in 1B to 27B sizes.

Open SourceSelf HostedOfflineGPU 4GB+
Easy
0.0 (0)

Open-source code LLM family by IBM for enterprise code generation.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)

Open-weight LLM by Meta in 8B and 70B sizes with strong general capabilities.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Featured

High-performance open-weight MoE LLM with 671B total parameters.

Open SourceSelf HostedOfflineGPU 24GB+
Advanced
0.0 (0)

Hybrid SSM-Transformer model by AI21 Labs combining Mamba with attention layers.

Open SourceSelf HostedOfflineGPU 12GB+
Intermediate
0.0 (0)

Open-weight code LLM trained on 2 trillion tokens of code and natural language.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Browse all Large Language Models (LLMs) tools