AI model directory
Compare published models by real constraints—provider, modality, capabilities, limits, and explicit per-token pricing.
24
Labs
139
Providers
68
Model families
Agentic model for autonomous multi-step research, synthesis, and cited reports
Multimodal embedding model mapping text, images, video, audio, and PDFs into a unified embedding space
Vision-language model for embodied reasoning: spatial understanding, task planning, and physical-world agentic robotics
Video model for image-to-video generation, editing, and extension workflows
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
Agentic coding model from Poolside in the XS size class for local deployment
Compact open Llama base model for lightweight and on-device use
Small open Llama base model for lightweight text generation and self-hosting
Music generation model for short 30-second clips, loops, and previews from text or image prompts
Music generation model for full-length songs from text or images with vocals and structure
Open Mistral reasoning model for transparent step-by-step problem solving