AI model directory

Choose from the models that matter.

Compare published models by real constraints—provider, modality, capabilities, limits, and explicit per-token pricing.

32

Labs

139

Providers

76

Model families

Try “claude”, “deepseek”, or “embedding”

Showing 61–72 of 352

Published models

Fast Gemini model balancing multimodal reasoning, tool use, and cost

Context
1,048,576
Providers
22
Input / 1M
$0.1857
Output / 1M
$1.1142
ReasoningText inputImage inputVideo input
Review providers and fit

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

Context
8,192
Providers
3
Input / 1M
$0.15
Output / 1M
$0
Text input
Review providers and fit

Fast Gemini model balancing multimodal reasoning, tool use, and cost

Context
1,048,756
Providers
7
Input / 1M
$1.5
Output / 1M
$9
ReasoningText inputImage inputVideo input
Review providers and fit

Low-latency Gemini model for high-volume multimodal and agent workloads

Context
1,048,576
Providers
5
Input / 1M
$0.25
Output / 1M
$1.5
ReasoningText inputImage inputVideo input
Review providers and fit

Video generation and editing model for fast, conversational text- and image-to-video workflows

Context
1,048,576
Providers
2
Input / 1M
$1.5
Output / 1M
$9
ReasoningText inputImage inputVideo input
Review providers and fit

Open Gemma instruction model for efficient chat and self-hosted deployments

Context
262,144
Providers
14
Input / 1M
$0.06
Output / 1M
$0.33
Open weightsReasoningText inputImage input
Review providers and fit

Largest Gemma 4 instruction model for open, self-hosted chat and reasoning

Context
262,144
Providers
25
Input / 1M
$0
Output / 1M
$0
Open weightsReasoningText inputImage input
Review providers and fit

Open Gemma instruction model for efficient chat and self-hosted deployments

Context
131,072
Providers
1
Input / 1M
$0.1
Output / 1M
$0.1
Open weightsReasoningText inputImage inputAudio input
Review providers and fit

Open Gemma instruction model for efficient chat and self-hosted deployments

Context
131,072
Providers
1
Input / 1M
$0.2
Output / 1M
$0.2
Open weightsReasoningText inputImage inputAudio input
Review providers and fit

Hybrid-reasoning GLM release that made the 4.5 line broadly useful

Context
131,072
Providers
18
Input / 1M
$0
Output / 1M
$0
Open weightsReasoningText input
Review providers and fit

Lighter GLM-4.5 variant for fast coding assistance and cheaper agents

Context
131,072
Providers
20
Input / 1M
$0
Output / 1M
$0
Open weightsReasoningText input
Review providers and fit

Efficient GLM model for fast reasoning, coding, and agent workflows

Context
200,000
Providers
3
Input / 1M
$0
Output / 1M
$0
ReasoningText input
Review providers and fit