AI model directory
Choose from the models that matter.
Compare published models by real constraints—provider, modality, capabilities, limits, and explicit per-token pricing.
32
Labs
139
Providers
76
Model families
Maximum-comprehensiveness agentic researcher for multi-step investigation, synthesis, and cited reports
- Context
- 1,048,576
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
High-accuracy OCR model for extracting text from documents, screenshots, receipts, and natural scenes
- Context
- 8,192
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
March 2025 checkpoint of DeepSeek-V3 with improved reasoning and coding
- Context
- 163,840
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
- Context
- 128,000
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
- Context
- 1,000,000
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes
- Context
- 1,000,000
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
R1 reasoning distilled into Qwen 2.5 32B for efficient open-weight step-by-step problem solving
- Context
- 131,072
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
Open DeepSeek MoE chat model for coding, math, and general reasoning
- Context
- 131,072
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes
- Context
- 131,072
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
Specialized Gemini 2.5 model for browser-control agents that automate UI tasks
- Context
- 128,000
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
High-quality, low-latency Live API model for real-time dialogue and voice-first AI applications
- Context
- 131,072
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown
Low-latency speech generation with steerable prompts and expressive audio tags
- Context
- 8,192
- Providers
- 0
- Input / 1M
- Unknown
- Output / 1M
- Unknown