Every AI model that matters one grid.
Text, image, audio, video and voice models compared prices, context and honest verdicts.
GPT-5
OpenAI's flagship with unified reasoning + fast modes.
GPT-5 Mini
The workhorse cheap model of the GPT-5 family.
GPT-5.1 Codex
OpenAI's coding-tuned GPT-5 variant.
o5-mini
Small, cheap deep-reasoning follow-up to o4-mini.
o4-mini
Small, cheap deep-reasoning model.
Sora 2
OpenAI's flagship video generation model.
Claude 4.5 Sonnet
The current state of the art in coding.
Claude 4.5 Haiku
Fast, cheap Claude for high-volume workloads.
Claude 4 Opus
The heavyweight for multi-hour autonomous work.
Claude 5 Sonnet
Anthropic's next-gen Sonnet with 500k context.
Gemini 2.5 Pro
2M-token context leader with native video.
Gemini 2.5 Flash
The cheapest 1M-context multimodal model.
Gemini 3 Pro
Google's flagship multimodal frontier.
Veo 3
Google's cinematic text-to-video with audio.
Imagen 4
Google's flagship image model.
Grok 4
Real-time X-native model with heavy reasoning.
DeepSeek V3.2
The best price-per-quality open model.
DeepSeek R2
OSS reasoning heir to R1.
Llama 4 Maverick
Meta's flagship OSS multimodal model.
Qwen 3 235B
Apache-2.0 OSS with thinking mode built in.
Qwen 3 Coder
Alibaba's coding-specialized OSS model.
Qwen-VL Max
Alibaba's flagship vision-language model.
GLM 4.6
Zhipu's frontier bilingual model.
Kimi K2
Moonshot's 2M-context agentic model.
Mistral Large 3
EU-hosted frontier for regulated workloads.
Codestral 2
Mistral's coding-specialized model.
Nova Pro
Amazon's flagship multimodal foundation model.
ElevenLabs v3
State-of-the-art text-to-speech.
Whisper v3 Turbo
Fast, cheap, open speech-to-text.
Flux 2 Pro
Flux's next-gen image model.