Realtime speech-to-speech model with configurable reasoning, tool use, and robust voice-agent behavior
- Context
- 128,000
- Providers
- 2
- Input / 1M
- $4
- Output / 1M
- $24
ReasoningText inputAudio inputImage input
Review providers and fitStreaming speech-to-text model for low-latency transcript deltas from live audio
- Context
- Unknown
- Providers
- 1
- Input / 1M
- Unknown
- Output / 1M
- Unknown
Audio input
Review providers and fitGrok model for agentic tool use, reasoning, coding, and live assistance
- Context
- 1,000,000
- Providers
- 1
- Input / 1M
- $1.25
- Output / 1M
- $2.5
Text inputImage inputPdf input
Review providers and fitReasoning Grok for document-heavy analysis and long-horizon tool use
- Context
- 1,000,000
- Providers
- 2
- Input / 1M
- $1.25
- Output / 1M
- $2.5
ReasoningText inputImage inputPdf input
Review providers and fitxAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
- Context
- 1,000,000
- Providers
- 19
- Input / 1M
- $0
- Output / 1M
- $0
ReasoningText inputImage inputPdf input
Review providers and fitxAI's latest Grok for chat, coding, agentic tools, and lower hallucination risk
- Context
- 1,000,000
- Providers
- 12
- Input / 1M
- $2
- Output / 1M
- $6
ReasoningText inputImage input
Review providers and fitFast Grok coding model tuned for agentic engineering and iterative edits
- Context
- 256,000
- Providers
- 13
- Input / 1M
- $0
- Output / 1M
- $0
ReasoningText inputImage inputPdf input
Review providers and fitTencent Hy reasoning model for coding, instruction following, and agent tasks
- Context
- 262,144
- Providers
- 8
- Input / 1M
- $0
- Output / 1M
- $0
Open weightsReasoningText input
Review providers and fitMultimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
- Context
- 1,048,576
- Providers
- 6
- Input / 1M
- $1
- Output / 1M
- $4.05
Open weightsReasoningText inputImage inputAudio input
Review providers and fitEarlier Kimi frontier model for long-context agents, coding, and multimodal work
- Context
- 262,144
- Providers
- 45
- Input / 1M
- $0
- Output / 1M
- $0
Open weightsReasoningText inputImage inputVideo input
Review providers and fitMultimodal Kimi workhorse for agent loops, coding tasks, and visual context
- Context
- 262,144
- Providers
- 54
- Input / 1M
- $0
- Output / 1M
- $0
Open weightsReasoningText inputImage inputVideo input
Review providers and fitCoding-focused Kimi model, stronger on long-horizon repo work with less overthinking
- Context
- 262,144
- Providers
- 33
- Input / 1M
- $0
- Output / 1M
- $0
Open weightsReasoningText inputImage inputVideo input
Review providers and fit