Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. Creators
  3. NVIDIA
Loading...

Model Evolution

One challenge, every NVIDIA generation.

Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
NVIDIA

NVIDIA: every model, side by side

Builds the Nemotron model family and AI hardware and software for inference.

Total Models

5

Text Models

5

Active Period

Sep 2025 to Aug 2026

Compare NVIDIA Models

Nemotron 3.5 Lightning

Aug 2026

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that benefit from domain-specific customization. A free-tier endpoint is also available on OpenRouter.

conversationreasoningcode-generationtool-useagentic-tool-use

Nemotron 3 Ultra

Jun 2026

NVIDIA's open reasoning and orchestration model, 55B active parameters out of 550B on a hybrid Transformer-Mamba mixture of experts. Text in, text out, with a context window up to 1M tokens. Built for long-running agent work: orchestration, coding agents, deep research and multi-step reasoning.

conversationreasoningcode-generationanalysisagentic-tool-usetool-useplanning

Nemotron 3.5 Content Safety

Jun 2026

A 4B multimodal guardrail model from NVIDIA, fine-tuned from Gemma-3-4B. It moderates both the input to an LLM and the response, taking text and images and returning a safe or unsafe classification for each, category labels, and an optional reasoning trace. 12 languages, 128K context.

analysisdata-extraction

NVIDIA Nemotron 3 Super (free)

Mar 2026

NVIDIA's Nemotron 3 Super: a 120B open hybrid mixture-of-experts activating 12B parameters per token. The Mamba-Transformer backbone with multi-token prediction generates over 50% more tokens than comparable open models, and Latent MoE calls 4 experts for the cost of one. 1M context, NVIDIA Open License.

conversationreasoningcode-generationanalysisplanningagentic-tool-use

NVIDIA Nemotron Nano 9B V2

Sep 2025

NVIDIA trained Nemotron Nano 9B v2 from scratch as one model for both reasoning and non-reasoning work. It can expose an internal reasoning trace before the answer, or be told by system prompt to give the answer alone.

conversationreasoningcode-generationanalysis