Lab

NVIDIA

Lab summary is pending review.

Llama 3.1 Nemotron 70B Instruct

unknown

Nemotron model for efficient reasoning, coding, and specialized AI agents

Context
128,000
Output
8,192
Offers
0
View model

Llama 3.1 Nemotron Safety Guard 8B v3

unknown

Safety model for policy screening, moderation, and risk-aware routing workflows

Context
128,000
Output
4,096
Offers
1
View model

Llama 3.1 Nemotron Ultra 253B

unknown

Flagship Nemotron model for high-throughput reasoning and complex agents

Context
128,000
Output
8,192
Offers
1
View model

Llama 3.3 Nemotron Super 49B v1

unknown

Nemotron model for efficient reasoning, coding, and specialized AI agents

Context
131,072
Output
131,072
Offers
1
View model

Llama 3.3 Nemotron Super 49B v1.5

unknown

Nemotron model for efficient reasoning, coding, and specialized AI agents

Context
131,072
Output
131,072
Offers
3
View model

Llama Nemotron Embed VL 1B v2

unknown

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

Context
32,768
Output
2,048
Offers
1
View model

Llama Nemotron Rerank VL 1B v2

unknown

Reranking model for improving retrieval quality in search and recommendation systems

Context
128,000
Output
4,096
Offers
1
View model

Mistral Nemotron

unknown

Mistral model for multilingual chat, reasoning, and tool-assisted workflows

Context
128,000
Output
8,192
Offers
1
View model

Nemotron 3.5 Content Safety

unknown

Safety model for policy screening, moderation, and risk-aware routing workflows

Context
128,000
Output
8,192
Offers
0
View model

Nemotron 3.5 Lightning 30B A3B

unknown

Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads

Context
262,144
Output
262,144
Offers
0
View model

Nemotron 3 Content Safety

unknown

Safety model for policy screening, moderation, and risk-aware routing workflows

Context
128,000
Output
4,096
Offers
1
View model

Nemotron 3 Nano 30B A3B

unknown

Small Nemotron 3 MoE for efficient coding, math, and long-context agents

Context
262,144
Output
262,144
Offers
6
View model

Nemotron 3 Nano Omni 30B A3B Reasoning

unknown

Open Nemotron omni model combining reasoning with text, vision, and audio

Context
256,000
Output
65,536
Offers
3
View model

Nemotron 3 Super 120B A12B

unknown

Nemotron middle tier for collaborative agents and high-volume reasoning workloads

Context
262,144
Output
262,144
Offers
8
View model

Nemotron 3 Ultra 550B A55B

unknown

Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy

Context
1,000,000
Output
128,000
Offers
4
View model

Nemotron Cascade 2 30B A3B

unknown

Nemotron model for efficient reasoning, coding, and specialized AI agents

Context
256,000
Output
32,768
Offers
1
View model

Nemotron Content Safety Reasoning 4B

unknown

Safety model for policy screening, moderation, and risk-aware routing workflows

Context
128,000
Output
4,096
Offers
1
View model

Nemotron Mini 4B Instruct

unknown

Compact Nemotron model for efficient reasoning and deployable AI agents

Context
128,000
Output
8,192
Offers
1
View model

Nemotron Nano 12B v2 VL

unknown

Nemotron multimodal model for visual reasoning and agentic AI workflows

Context
128,000
Output
128,000
Offers
2
View model

Nemotron Nano 9B v2

unknown

Compact Nemotron model for efficient reasoning and deployable AI agents

Context
131,072
Output
131,072
Offers
2
View model

Nemotron VoiceChat

unknown

Nemotron multimodal model for visual reasoning and agentic AI workflows

Context
128,000
Output
8,192
Offers
1
View model