OpenAI · Stable
GPT-5.6 Sol
OpenAI's flagship general model for difficult coding, analysis, and professional work, with a very large context window and a broad native tool set.
- Context window
- 1.05M tokens
- Maximum output
- 128K tokens
Cody model library
Compare leading text, image, video, voice, and world models using current provider facts—not a mystery score. See what each model does, what it costs, where you can access it, and what to verify before production.
Compare up to four models
Comparisons stay within one category so the measurements mean the same thing.
Select models from the cards
Showing 1–12 of 21 models
OpenAI · Stable
OpenAI's flagship general model for difficult coding, analysis, and professional work, with a very large context window and a broad native tool set.
Anthropic · Stable
Anthropic's frontier model for ambitious, long-running agentic and coding work, with strong vision and enterprise marketplace availability.
Google · Stable
Google's stable, high-efficiency multimodal model for agents, software work, and large mixed-media inputs at an introductory Flash-tier price.
SpaceXAI · Stable
SpaceXAI's frontier text-and-image model for coding, agentic tasks, and knowledge work, with direct web, X, and code tools.
OpenAI · Stable
OpenAI's current image model for detailed generation and editing, especially complex prompts, text-heavy visuals, and high-fidelity source images.
Google · Stable
Google's versatile image workhorse for fast generation, conversational edits, multiple references, grounded imagery, and outputs up to 4K.
Black Forest Labs · Stable
Black Forest Labs' production-focused image model for scalable generation and editing, precise composition, and multi-reference creative workflows.
Ideogram · Stable
A design-oriented image model with open weights, multilingual typography, structured layout control, editable elements, and simple per-image API tiers.
OpenAI · Legacy
OpenAI's legacy synced-audio video model for short text- or image-guided clips—still documented, but approaching a permanent API shutdown.
Google · Preview
Google's preview video model for text-, image-, and video-guided generation with native audio and output options up to 4K.
Runway · Stable
Runway's flagship developer video model for text- and image-guided clips, with professional formats and a predictable credit-per-second base rate.
Kuaishou · Stable
Kuaishou's Pro video model delivered through fal for multi-shot text/image generation, optional native audio, references, and clips up to 15 seconds.
How to use this library
A deliberately small starting set of frontier and category-defining models with practical API information.
Cody does not run a universal quality leaderboard. Provider claims are labeled, pricing excludes taxes and optional tools, and production buyers should verify the live provider terms before choosing a model.
Use first-party documentation for technical limits, lifecycle, pricing, regional access, and data policies whenever it exists.
Compare models only inside the same category and retain each provider's measurement basis.
Show Not published when a provider has not published a comparable fact instead of estimating it.
Treat prices and availability as dated snapshots and link every page to live provider documentation.
Voice latency, text throughput, video render time, and world-model frame rate are not interchangeable.
A missing release date, country list, or latency number is labeled as unpublished instead of estimated.
Model pages link to the underlying source and show when pricing, access, and policies were last checked.
Put the model to work
Use Cody's expanded prompt playbooks and curated agent skills to turn a model choice into a repeatable workflow.