Cody model library

AI Models, Explained Clearly

Compare leading text, image, video, voice, and world models using current provider facts—not a mystery score. See what each model does, what it costs, where you can access it, and what to verify before production.

21 researched modelsFacts checked September 3, 2026

Compare up to four models

Comparisons stay within one category so the measurements mean the same thing.

Select models from the cards

Showing 1–12 of 21 models

OpenAI · Stable

GPT-5.6 Sol

OpenAI's flagship general model for difficult coding, analysis, and professional work, with a very large context window and a broad native tool set.

Context window
1.05M tokens
Maximum output
128K tokens
View model guide

Anthropic · Stable

Claude Fable 5.1

Anthropic's frontier model for ambitious, long-running agentic and coding work, with strong vision and enterprise marketplace availability.

Context window
1M tokens
Maximum output
128K tokens
View model guide

Google · Stable

Gemini 3.8 Flash

Google's stable, high-efficiency multimodal model for agents, software work, and large mixed-media inputs at an introductory Flash-tier price.

Context window
1,048,576 tokens
Maximum output
65,536 tokens
View model guide

SpaceXAI · Stable

Grok 4.6

SpaceXAI's frontier text-and-image model for coding, agentic tasks, and knowledge work, with direct web, X, and code tools.

Context window
500K tokens
Maximum output
No published text cap
View model guide

OpenAI · Stable

GPT-Image-2

OpenAI's current image model for detailed generation and editing, especially complex prompts, text-heavy visuals, and high-fidelity source images.

Typical price basis
$0.006–$0.211 at 1024² on fal
Maximum output
Up to about 4K / 8.29MP on fal
View model guide

Google · Stable

Nano Banana 2

Google's versatile image workhorse for fast generation, conversational edits, multiple references, grounded imagery, and outputs up to 4K.

Typical price basis
$0.067 / 1K image
Maximum output
0.5K, 1K, 2K, 4K
View model guide

Black Forest Labs · Stable

FLUX.2 Pro

Black Forest Labs' production-focused image model for scalable generation and editing, precise composition, and multi-reference creative workflows.

Typical price basis
From $0.03/MP generation
Maximum output
Up to 4MP, any aspect ratio
View model guide

Ideogram · Stable

Ideogram 4.0

A design-oriented image model with open weights, multilingual typography, structured layout control, editable elements, and simple per-image API tiers.

Typical price basis
$0.03 Turbo · $0.06 Default · $0.10 Quality
Maximum output
1K or 2K
View model guide

OpenAI · Legacy

Sora 2

OpenAI's legacy synced-audio video model for short text- or image-guided clips—still documented, but approaching a permanent API shutdown.

Generation price
$0.10 / generated second
Clip duration
4, 8, or 12 seconds
View model guide

Google · Preview

Veo 3.1

Google's preview video model for text-, image-, and video-guided generation with native audio and output options up to 4K.

Generation price
$0.40/s standard with audio
Clip duration
4, 6, or 8 seconds
View model guide

Runway · Stable

Runway Gen-4.5

Runway's flagship developer video model for text- and image-guided clips, with professional formats and a predictable credit-per-second base rate.

Generation price
12 credits / sec = $0.12 / sec
Clip duration
2–10 seconds
View model guide

Kuaishou · Stable

Kling Video 3.0 Pro

Kuaishou's Pro video model delivered through fal for multi-shot text/image generation, optional native audio, references, and clips up to 15 seconds.

Generation price
$0.112/s no audio · $0.168/s audio
Clip duration
Up to 15 seconds
View model guide

How to use this library

Compare the job, not the hype.

A deliberately small starting set of frontier and category-defining models with practical API information.

Cody does not run a universal quality leaderboard. Provider claims are labeled, pricing excludes taxes and optional tools, and production buyers should verify the live provider terms before choosing a model.

  1. 01

    Use first-party documentation for technical limits, lifecycle, pricing, regional access, and data policies whenever it exists.

  2. 02

    Compare models only inside the same category and retain each provider's measurement basis.

  3. 03

    Show Not published when a provider has not published a comparable fact instead of estimating it.

  4. 04

    Treat prices and availability as dated snapshots and link every page to live provider documentation.

Speed keeps its units

Voice latency, text throughput, video render time, and world-model frame rate are not interchangeable.

Unknown stays unknown

A missing release date, country list, or latency number is labeled as unpublished instead of estimated.

Every fact has a date

Model pages link to the underlying source and show when pricing, access, and policies were last checked.

Put the model to work

Pick the model, then give it a better brief.

Use Cody's expanded prompt playbooks and curated agent skills to turn a model choice into a repeatable workflow.