Provider
Provider summary is pending review.
| Model | Provider ID | Input price | Output price | Context | Output |
|---|---|---|---|---|---|
| DeepSeek V4 Flash DeepSeek | deepseek-ai/deepseek-v4-flash | $0.14 / 1M tokens | $0.28 / 1M tokens | 1,048,576 | 393,216 |
| DeepSeek V4 Pro DeepSeek | deepseek-ai/deepseek-v4-pro | $0.435 / 1M tokens | $0.87 / 1M tokens | 1,048,576 | 393,216 |
| Gemma 4 31B IT Google | google/gemma-4-31b-it | $0 / 1M tokens | $0 / 1M tokens | 256,000 | 16,384 |
| Llama-3.3-70B-Instruct Meta | meta/llama-3.3-70b-instruct | $0 / 1M tokens | $0 / 1M tokens | 128,000 | 4,096 |
| MiniMax-M2.7 MiniMax | minimaxai/minimax-m2.7 | $0 / 1M tokens | $0 / 1M tokens | 204,800 | 131,072 |
| MiniMax-M3 MiniMax | minimaxai/minimax-m3 | $0 / 1M tokens | $0 / 1M tokens | 1,000,000 | 16,384 |
| Mistral Nemotron NVIDIA | mistralai/mistral-nemotron | $0 / 1M tokens | $0 / 1M tokens | 128,000 | 8,192 |
| Kimi K2.6 Moonshot AI | moonshotai/kimi-k2.6 | $0 / 1M tokens | $0 / 1M tokens | 262,144 | 262,144 |
| Llama 3.1 Nemotron Safety Guard 8B v3 NVIDIA | nvidia/llama-3_1-nemotron-safety-guard-8b-v3 | $0 / 1M tokens | $0 / 1M tokens | 128,000 | 4,096 |
| Llama Nemotron Embed VL 1B v2 NVIDIA | nvidia/llama-nemotron-embed-vl-1b-v2 | $0 / 1M tokens | $0 / 1M tokens | 32,768 | 2,048 |
| Llama Nemotron Rerank VL 1B v2 NVIDIA | nvidia/llama-nemotron-rerank-vl-1b-v2 | $0 / 1M tokens | $0 / 1M tokens | 128,000 | 4,096 |
| Nemotron 3 Content Safety NVIDIA | nvidia/nemotron-3-content-safety | $0 / 1M tokens | $0 / 1M tokens | 128,000 | 4,096 |
| Nemotron 3 Nano 30B A3B NVIDIA | nvidia/nemotron-3-nano-30b-a3b | $0 / 1M tokens | $0 / 1M tokens | 131,072 | 131,072 |
| Nemotron 3 Nano Omni 30B A3B Reasoning NVIDIA | nvidia/nemotron-3-nano-omni-30b-a3b-reasoning | $0 / 1M tokens | $0 / 1M tokens | 256,000 | 65,536 |
| Nemotron 3 Super 120B A12B NVIDIA | nvidia/nemotron-3-super-120b-a12b | $0.2 / 1M tokens | $0.8 / 1M tokens | 262,144 | 262,144 |
| Nemotron 3 Ultra 550B A55B NVIDIA | nvidia/nemotron-3-ultra-550b-a55b | $0.5 / 1M tokens | $2.5 / 1M tokens | 1,000,000 | 65,536 |
| Nemotron Content Safety Reasoning 4B NVIDIA | nvidia/nemotron-content-safety-reasoning-4b | $0 / 1M tokens | $0 / 1M tokens | 128,000 | 4,096 |
| Nemotron Mini 4B Instruct NVIDIA | nvidia/nemotron-mini-4b-instruct | $0 / 1M tokens | $0 / 1M tokens | 128,000 | 8,192 |
| Nemotron VoiceChat NVIDIA | nvidia/nemotron-voicechat | $0 / 1M tokens | $0 / 1M tokens | 128,000 | 8,192 |
| GPT OSS 120B OpenAI | openai/gpt-oss-120b | $0 / 1M tokens | $0 / 1M tokens | 128,000 | 8,192 |
| GPT OSS 20B OpenAI | openai/gpt-oss-20b | $0 / 1M tokens | $0 / 1M tokens | 131,072 | 32,768 |
| Whisper 3 Large OpenAI | openai/whisper-large-v3 | $0 / 1M tokens | $0 / 1M tokens | Unknown | 4,096 |
| Qwen3-Coder 480B-A35B Instruct Alibaba | qwen/qwen3-coder-480b-a35b-instruct | $0 / 1M tokens | $0 / 1M tokens | 262,144 | 66,536 |
| Qwen3-Next 80B-A3B Instruct Alibaba | qwen/qwen3-next-80b-a3b-instruct | $0 / 1M tokens | $0 / 1M tokens | 262,144 | 16,384 |
| Qwen3.5 122B-A10B Alibaba | qwen/qwen3.5-122b-a10b | $0 / 1M tokens | $0 / 1M tokens | 262,144 | 65,536 |
| Qwen3.5 397B-A17B Alibaba | qwen/qwen3.5-397b-a17b | $0 / 1M tokens | $0 / 1M tokens | 262,144 | 8,192 |
| Step 3.5 Flash Stepfun | stepfun-ai/step-3.5-flash | $0 / 1M tokens | $0 / 1M tokens | 256,000 | 16,384 |
| Step 3.7 Flash Stepfun | stepfun-ai/step-3.7-flash | $0 / 1M tokens | $0 / 1M tokens | 256,000 | 16,384 |
| GLM-5.2 Zhipu AI | z-ai/glm-5.2 | $0 / 1M tokens | $0 / 1M tokens | 1,000,000 | 131,072 |