Provider

Nvidia

Provider summary is pending review.

Visible model offerings

ModelProvider IDInput priceOutput priceContextOutput
DeepSeek V4 Flash
DeepSeek
deepseek-ai/deepseek-v4-flash$0.14 / 1M tokens$0.28 / 1M tokens1,048,576393,216
DeepSeek V4 Pro
DeepSeek
deepseek-ai/deepseek-v4-pro$0.435 / 1M tokens$0.87 / 1M tokens1,048,576393,216
Gemma 4 31B IT
Google
google/gemma-4-31b-it$0 / 1M tokens$0 / 1M tokens256,00016,384
Llama-3.3-70B-Instruct
Meta
meta/llama-3.3-70b-instruct$0 / 1M tokens$0 / 1M tokens128,0004,096
MiniMax-M2.7
MiniMax
minimaxai/minimax-m2.7$0 / 1M tokens$0 / 1M tokens204,800131,072
MiniMax-M3
MiniMax
minimaxai/minimax-m3$0 / 1M tokens$0 / 1M tokens1,000,00016,384
Mistral Nemotron
NVIDIA
mistralai/mistral-nemotron$0 / 1M tokens$0 / 1M tokens128,0008,192
Kimi K2.6
Moonshot AI
moonshotai/kimi-k2.6$0 / 1M tokens$0 / 1M tokens262,144262,144
Llama 3.1 Nemotron Safety Guard 8B v3
NVIDIA
nvidia/llama-3_1-nemotron-safety-guard-8b-v3$0 / 1M tokens$0 / 1M tokens128,0004,096
Llama Nemotron Embed VL 1B v2
NVIDIA
nvidia/llama-nemotron-embed-vl-1b-v2$0 / 1M tokens$0 / 1M tokens32,7682,048
Llama Nemotron Rerank VL 1B v2
NVIDIA
nvidia/llama-nemotron-rerank-vl-1b-v2$0 / 1M tokens$0 / 1M tokens128,0004,096
Nemotron 3 Content Safety
NVIDIA
nvidia/nemotron-3-content-safety$0 / 1M tokens$0 / 1M tokens128,0004,096
Nemotron 3 Nano 30B A3B
NVIDIA
nvidia/nemotron-3-nano-30b-a3b$0 / 1M tokens$0 / 1M tokens131,072131,072
Nemotron 3 Nano Omni 30B A3B Reasoning
NVIDIA
nvidia/nemotron-3-nano-omni-30b-a3b-reasoning$0 / 1M tokens$0 / 1M tokens256,00065,536
Nemotron 3 Super 120B A12B
NVIDIA
nvidia/nemotron-3-super-120b-a12b$0.2 / 1M tokens$0.8 / 1M tokens262,144262,144
Nemotron 3 Ultra 550B A55B
NVIDIA
nvidia/nemotron-3-ultra-550b-a55b$0.5 / 1M tokens$2.5 / 1M tokens1,000,00065,536
Nemotron Content Safety Reasoning 4B
NVIDIA
nvidia/nemotron-content-safety-reasoning-4b$0 / 1M tokens$0 / 1M tokens128,0004,096
Nemotron Mini 4B Instruct
NVIDIA
nvidia/nemotron-mini-4b-instruct$0 / 1M tokens$0 / 1M tokens128,0008,192
Nemotron VoiceChat
NVIDIA
nvidia/nemotron-voicechat$0 / 1M tokens$0 / 1M tokens128,0008,192
GPT OSS 120B
OpenAI
openai/gpt-oss-120b$0 / 1M tokens$0 / 1M tokens128,0008,192
GPT OSS 20B
OpenAI
openai/gpt-oss-20b$0 / 1M tokens$0 / 1M tokens131,07232,768
Whisper 3 Large
OpenAI
openai/whisper-large-v3$0 / 1M tokens$0 / 1M tokensUnknown4,096
Qwen3.5 122B-A10B
Alibaba
qwen/qwen3.5-122b-a10b$0 / 1M tokens$0 / 1M tokens262,14465,536
Qwen3.5 397B-A17B
Alibaba
qwen/qwen3.5-397b-a17b$0 / 1M tokens$0 / 1M tokens262,1448,192
Qwen3-Coder 480B-A35B Instruct
Alibaba
qwen/qwen3-coder-480b-a35b-instruct$0 / 1M tokens$0 / 1M tokens262,14466,536
Qwen3-Next 80B-A3B Instruct
Alibaba
qwen/qwen3-next-80b-a3b-instruct$0 / 1M tokens$0 / 1M tokens262,14416,384
Step 3.5 Flash
Stepfun
stepfun-ai/step-3.5-flash$0 / 1M tokens$0 / 1M tokens256,00016,384
Step 3.7 Flash
Stepfun
stepfun-ai/step-3.7-flash$0 / 1M tokens$0 / 1M tokens256,00016,384
GLM-5.2
Zhipu AI
z-ai/glm-5.2$0 / 1M tokens$0 / 1M tokens1,000,000131,072