AI model directory
Choose from the models that matter.
Compare published models by real constraints—provider, modality, capabilities, limits, and explicit per-token pricing.
32
Labs
139
Providers
76
Model families
Fast Gemini model balancing multimodal reasoning, tool use, and cost
- Context
- 1,048,576
- Providers
- 22
- Input / 1M
- $0.1857
- Output / 1M
- $1.1142
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
- Context
- 8,192
- Providers
- 3
- Input / 1M
- $0.15
- Output / 1M
- $0
Fast Gemini model balancing multimodal reasoning, tool use, and cost
- Context
- 1,048,756
- Providers
- 7
- Input / 1M
- $1.5
- Output / 1M
- $9
Low-latency Gemini model for high-volume multimodal and agent workloads
- Context
- 1,048,576
- Providers
- 5
- Input / 1M
- $0.25
- Output / 1M
- $1.5
Video generation and editing model for fast, conversational text- and image-to-video workflows
- Context
- 1,048,576
- Providers
- 2
- Input / 1M
- $1.5
- Output / 1M
- $9
Open Gemma instruction model for efficient chat and self-hosted deployments
- Context
- 262,144
- Providers
- 14
- Input / 1M
- $0.06
- Output / 1M
- $0.33
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
- Context
- 262,144
- Providers
- 25
- Input / 1M
- $0
- Output / 1M
- $0
Open Gemma instruction model for efficient chat and self-hosted deployments
- Context
- 131,072
- Providers
- 1
- Input / 1M
- $0.1
- Output / 1M
- $0.1
Open Gemma instruction model for efficient chat and self-hosted deployments
- Context
- 131,072
- Providers
- 1
- Input / 1M
- $0.2
- Output / 1M
- $0.2
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
- Context
- 131,072
- Providers
- 18
- Input / 1M
- $0
- Output / 1M
- $0
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
- Context
- 131,072
- Providers
- 20
- Input / 1M
- $0
- Output / 1M
- $0
Efficient GLM model for fast reasoning, coding, and agent workflows
- Context
- 200,000
- Providers
- 3
- Input / 1M
- $0
- Output / 1M
- $0