OpenAI's legacy synced-audio video model for short text- or image-guided clips—still documented, but approaching a permanent API shutdown.
- Generation price
- $0.10 / generated second
- Clip duration
- 4, 8, or 12 seconds
Resources
Showing 1–12 of 13 models
OpenAI's legacy synced-audio video model for short text- or image-guided clips—still documented, but approaching a permanent API shutdown.
Google's preview video model for text-, image-, and video-guided generation with native audio and output options up to 4K.
Google · Live
Google's lower-cost Veo 3.1 variant for quickly creating short video clips with native audio and output options up to 4K.
fal · Live
fal's speed-focused post-trained version of MiniMax H3 for short video with native audio, fast 768p iteration, and text, image, or multimodal reference input.
Runway · Live
Runway's flagship developer video model for text- and image-guided clips, with professional formats and a predictable credit-per-second base rate.
Kuaishou · Live
Kuaishou's Pro video model delivered through fal for multi-shot text/image generation, optional native audio, references, and clips up to 15 seconds.
A cost-conscious Kling model on fal for fluid five- or ten-second clips generated from text or a starting image.
SpaceXAI's current video model for turning text, images, or references into short clips with optional generated speech and sound.
ByteDance Seed · Live
ByteDance's current audiovisual model for coherent clips up to 30 seconds, with native sound, long-form storytelling, rich references, and targeted editing.
HappyHorse · Live
Alibaba's unified video family for creating short clips with sound from text, a starting image, or a reference video.
Lightricks · Live
Lightricks' quality-focused video model for making 720p or 1080p clips from text, images, or audio with native sound.
Lightricks · Live
LTX's speed-optimized video model for text, image, or audio input, native sound, automatic duration, and output up to 4K.
How to use this library
A curated cross-category set of frontier, specialist, and category-defining models with practical API information for real product decisions.
Cody does not run a universal quality leaderboard. Provider claims are labeled, pricing excludes taxes and optional tools, and production buyers should verify the live provider terms before choosing a model.
Use first-party documentation for technical limits, lifecycle, pricing, regional access, and data policies whenever it exists.
Compare models only inside the same category and retain each provider's measurement basis.
Show Not published when a provider has not published a comparable fact instead of estimating it.
Treat prices and availability as dated snapshots and link every page to live provider documentation.
Voice latency, text throughput, video render time, and world-model frame rate are not interchangeable.
A missing release date, country list, or latency number is labeled as unpublished instead of estimated.
Model pages link to the underlying source and show when pricing, access, and policies were last checked.
Put the model to work
Use Cody's expanded prompt playbooks and curated agent skills to turn a model choice into a repeatable workflow.