All models
HappyHorse

HappyHorse 1.1

Alibaba's unified video family for creating short clips with sound from text, a starting image, or a reference video.

Plain-English overview

What HappyHorse 1.1 actually is

HappyHorse 1.1 is a set of related QwenCloud video endpoints rather than one all-purpose request. You choose text-to-video, image-to-video, or reference-to-video depending on how much visual direction you already have, then set duration, resolution, and framing.

The family can return clips up to 15 seconds and 1080p with audio. That makes it useful for complete short creative beats, while longer sequences still need shot planning, continuity checks, and editing outside the model.

Good fit for

  • Short concept videos with native sound
  • Animating a supplied first frame
  • Reference-guided character or style experiments

Category comparison

The facts that matter for video models

These are provider-published specifications, not Cody benchmark scores. Follow the linked sources for current limits and endpoint-specific exceptions.

Generation price
Per generated second; use live QwenCloud rateCurrent list price per generated second or the provider's closest billing unit.
Clip duration
3–15 secondsDocumented duration choices or maximum clip length.
Maximum output
1080pHighest provider-documented output resolution, with relevant constraints.
Native audio
YesWhether the endpoint generates synchronized audio with video.
Input modes
Text, first image, or reference videoSupported text, image, video, reference, or motion inputs.
Generation time
Not publishedOnly shown when a provider publishes a range; not a Cody benchmark.

Pricing & comparisons

Estimate your cost

Set your usage. Your estimate updates as you type.

Example: 8-second, 720p clips with audio. Models that do not support these settings will not receive an estimate.

HappyHorse 1.1

HappyHorse

Estimated total (USD)

No reviewed rate

For the usage above · USD · API pricing, not a subscription

How this estimate works

Estimates exclude taxes, tools, cache storage/writes, free allowances and custom discounts. Image estimates cover output only, not prompt or reference-image charges. Quality modes differ by model. Unlisted settings are not treated as free.

API and provider access

Where to get HappyHorse 1.1

Availability

Regions and access stage

Available through QwenCloud as separate text-to-video, image-to-video, and reference-to-video endpoints.

Marketplace and regional access depend on the QwenCloud account and endpoint selected.

Check live availability

Data and training

The route matters.

QwenCloud says API inputs and outputs are not retained for model training and are processed in memory. Confirm regional storage and any enabled conversation-storage setting for the exact account.

This is a concise reading of the cited provider material, not legal advice. A third-party gateway can have different storage, routing, training, and residency terms from the model maker's direct API.

Read the provider policy

Frequently asked questions

HappyHorse 1.1 FAQ

What is HappyHorse 1.1?

Alibaba's unified video family for creating short clips with sound from text, a starting image, or a reference video. HappyHorse 1.1 is a set of related QwenCloud video endpoints rather than one all-purpose request. You choose text-to-video, image-to-video, or reference-to-video depending on how much visual direction you already have, then set duration, resolution, and framing.

When was HappyHorse 1.1 released?

HappyHorse 1.1 was released on June 22, 2026 according to the cited provider materials.

Where can I access HappyHorse 1.1?

Available through QwenCloud as separate text-to-video, image-to-video, and reference-to-video endpoints. The access routes listed in this guide are QwenCloud.

How much does HappyHorse 1.1 cost?

Live QwenCloud marketplace rate. HappyHorse is billed through the QwenCloud marketplace. The public API guide does not expose one durable cross-region dollar rate, so use the live console price for production estimates.

Where is HappyHorse 1.1 available?

Available through QwenCloud as separate text-to-video, image-to-video, and reference-to-video endpoints. Marketplace and regional access depend on the QwenCloud account and endpoint selected.

Is my HappyHorse 1.1 API data used for training?

QwenCloud says API inputs and outputs are not retained for model training and are processed in memory. Confirm regional storage and any enabled conversation-storage setting for the exact account. The policy belongs to the provider route and account terms, so verify it again before production use.

Related comparisons

Head-to-head comparisons

  1. HappyHorse 1.1 vs Sora 2

    OpenAI's legacy synced-audio video model for short text- or image-guided clips—still documented, but approaching a permanent API shutdown.

    Open full comparison
  2. HappyHorse 1.1 vs Veo 3.1

    Google's preview video model for text-, image-, and video-guided generation with native audio and output options up to 4K.

    Open full comparison
  3. HappyHorse 1.1 vs Veo 3.1 Fast

    Google's lower-cost Veo 3.1 variant for quickly creating short video clips with native audio and output options up to 4K.

    Open full comparison
  4. HappyHorse 1.1 vs Runway Gen-4.5

    Runway's flagship developer video model for text- and image-guided clips, with professional formats and a predictable credit-per-second base rate.

    Open full comparison
  5. HappyHorse 1.1 vs Kling Video 3.0 Pro

    Kuaishou's Pro video model delivered through fal for multi-shot text/image generation, optional native audio, references, and clips up to 15 seconds.

    Open full comparison
  6. HappyHorse 1.1 vs Kling Video 2.5 Turbo Pro

    A cost-conscious Kling model on fal for fluid five- or ten-second clips generated from text or a starting image.

    Open full comparison
  7. HappyHorse 1.1 vs Grok Imagine Video 1.5

    SpaceXAI's current video model for turning text, images, or references into short clips with optional generated speech and sound.

    Open full comparison
  8. HappyHorse 1.1 vs Seedance 2.5

    ByteDance's current audiovisual model for coherent clips up to 30 seconds, with native sound, long-form storytelling, rich references, and targeted editing.

    Open full comparison
  9. HappyHorse 1.1 vs LTX-2.5 Pro

    Lightricks' quality-focused video model for making 720p or 1080p clips from text, images, or audio with native sound.

    Open full comparison
  10. HappyHorse 1.1 vs LTX-2.5 Fast

    LTX's speed-optimized video model for text, image, or audio input, native sound, automatic duration, and output up to 4K.

    Open full comparison
  11. HappyHorse 1.1 vs LTX-2.3 Pro

    LTX's mature full-workflow video model for generation plus Retake, Extend, Reframe, native audio, and first-to-last-frame control.

    Open full comparison