Model Directory

    Every AI model that matters one grid.

    Text, image, audio, video and voice models compared prices, context and honest verdicts.

    Provider
    Modality
    Showing 30 of 30 models

    GPT-5

    OpenAI · GPT
    400k ctx

    OpenAI's flagship with unified reasoning + fast modes.

    Text
    Image
    Audio
    In: $1.25/MOut: $10/M

    GPT-5 Mini

    OpenAI · GPT
    400k ctx

    The workhorse cheap model of the GPT-5 family.

    Text
    Image
    In: $0.25/MOut: $2/M

    GPT-5.1 Codex

    OpenAI · Codex
    400k ctx

    OpenAI's coding-tuned GPT-5 variant.

    Text
    In: $1.50/MOut: $12/M

    o5-mini

    OpenAI · o-series
    256k ctx

    Small, cheap deep-reasoning follow-up to o4-mini.

    Text
    Image
    In: $1.20/MOut: $4.80/M

    o4-mini

    OpenAI · o-series
    200k ctx

    Small, cheap deep-reasoning model.

    Text
    Image
    In: $1.10/MOut: $4.40/M

    Sora 2

    OpenAI · Sora
    n/a ctx

    OpenAI's flagship video generation model.

    Video
    In: usage-basedOut: usage-based

    Claude 4.5 Sonnet

    Anthropic · Claude
    200k ctx

    The current state of the art in coding.

    Text
    Image
    In: $3/MOut: $15/M

    Claude 4.5 Haiku

    Anthropic · Claude
    200k ctx

    Fast, cheap Claude for high-volume workloads.

    Text
    In: $0.80/MOut: $4/M

    Claude 4 Opus

    Anthropic · Claude
    200k ctx

    The heavyweight for multi-hour autonomous work.

    Text
    Image
    In: $15/MOut: $75/M

    Claude 5 Sonnet

    Anthropic · Claude
    500k ctx

    Anthropic's next-gen Sonnet with 500k context.

    Text
    Image
    In: $4/MOut: $18/M

    Gemini 2.5 Pro

    Google · Gemini
    2M ctx

    2M-token context leader with native video.

    Text
    Image
    Audio
    Video
    In: $1.25/MOut: $10/M

    Gemini 2.5 Flash

    Google · Gemini
    1M ctx

    The cheapest 1M-context multimodal model.

    Text
    Image
    Audio
    Video
    In: $0.30/MOut: $2.50/M

    Gemini 3 Pro

    Google · Gemini
    2M ctx

    Google's flagship multimodal frontier.

    Text
    Image
    Audio
    Video
    In: $2/MOut: $15/M

    Veo 3

    Google · Veo
    n/a ctx

    Google's cinematic text-to-video with audio.

    Video
    Audio
    In: usage-basedOut: usage-based

    Imagen 4

    Google · Imagen
    n/a ctx

    Google's flagship image model.

    Image
    In: usage-basedOut: usage-based

    Grok 4

    xAI · Grok
    256k ctx

    Real-time X-native model with heavy reasoning.

    Text
    Image
    In: $3/MOut: $15/M

    DeepSeek V3.2

    DeepSeek · DeepSeek
    128k ctx

    The best price-per-quality open model.

    Text
    In: $0.14/MOut: $0.28/M

    DeepSeek R2

    DeepSeek · DeepSeek
    128k ctx

    OSS reasoning heir to R1.

    Text
    In: $0.20/MOut: $0.80/M

    Llama 4 Maverick

    Meta · Llama
    1M ctx

    Meta's flagship OSS multimodal model.

    Text
    Image
    In: $0.50/MOut: $1.50/M

    Qwen 3 235B

    Alibaba · Qwen
    128k ctx

    Apache-2.0 OSS with thinking mode built in.

    Text
    Image
    In: $0.30/MOut: $1.20/M

    Qwen 3 Coder

    Alibaba · Qwen
    256k ctx

    Alibaba's coding-specialized OSS model.

    Text
    In: $0.40/MOut: $1.60/M

    Qwen-VL Max

    Alibaba · Qwen
    128k ctx

    Alibaba's flagship vision-language model.

    Text
    Image
    Video
    In: $1/MOut: $3/M

    GLM 4.6

    Zhipu · GLM
    200k ctx

    Zhipu's frontier bilingual model.

    Text
    Image
    In: $0.60/MOut: $2.20/M

    Kimi K2

    Moonshot · Kimi
    2M ctx

    Moonshot's 2M-context agentic model.

    Text
    Image
    In: $0.50/MOut: $2/M

    Mistral Large 3

    Mistral · Mistral
    256k ctx

    EU-hosted frontier for regulated workloads.

    Text
    In: $2/MOut: $6/M

    Codestral 2

    Mistral · Codestral
    256k ctx

    Mistral's coding-specialized model.

    Text
    In: $0.30/MOut: $0.90/M

    Nova Pro

    Amazon · Nova
    300k ctx

    Amazon's flagship multimodal foundation model.

    Text
    Image
    Video
    In: $0.80/MOut: $3.20/M

    ElevenLabs v3

    ElevenLabs · Voice
    n/a ctx

    State-of-the-art text-to-speech.

    Voice
    Audio
    In: usage-basedOut: usage-based

    Whisper v3 Turbo

    OpenAI · Whisper
    n/a ctx

    Fast, cheap, open speech-to-text.

    Voice
    Audio
    In: $0.004/minOut: n/a

    Flux 2 Pro

    Black Forest Labs · Flux
    n/a ctx

    Flux's next-gen image model.

    Image
    In: usage-basedOut: usage-based
    Need a hand?

    Talk to a builder, not a bot.

    Questions about vibe coding, prompts, launching, pricing, or a partnership? We reply within one business day usually a few hours.