Wan 2.6
No media selects text-to-video, exactly one image selects image-to-video, and one to three videos select video editing; images and videos cannot be mixed. The default is 1080p, 5 seconds, shot structure Off, and content filtering On.
At a glance
What is this model like?
No media selects text-to-video, exactly one image selects image-to-video, and one to three videos select video editing; images and videos cannot be mixed. Text and image inputs support 5, 10, or 15 seconds; video input supports only 5 or 10 seconds.
Key facts
Quickly assess whether this model fits your use case.
Also known as
The same model may appear under different names across documentation and community discussions; this list keeps them easy to search and compare.
Common questions
These practical questions focus on the task, input conditions, and result requirements you should confirm before choosing.
Selection guide
Use these decision points when choosing a model.
Choose one input path first: No media selects text-to-video, exactly one image selects image-to-video, and one to three videos select video editing; images and videos cannot be mixed.
Filter duration by workflow: Text and image inputs support 5, 10, or 15 seconds; video input supports only 5 or 10 seconds.
Build three separate baselines: The default is 1080p, 5 seconds, shot structure Off, and content filtering On. Compare models with the same prompt, media, duration, and resolution; test returned results instead of assuming marketing claims.
Popular comparisons
Compare common alternatives on the same task to clarify differences in inputs, controls, and cost.
Model comparison
Compare the current model with alternatives at a glance.
Practical usage insights
Practical guidance based on public sources, current on-site limits, and representative tasks.
Wan 2.6 — Media-driven three-workflow selection
Wan 2.6 — Six-entry credit table
Capabilities
Media-driven three-workflow selection
Six-entry credit table
Media and safety boundary
Use cases
Text-only baseline
One-image motion test
1–3 MP4/MOV — Video duration
Prompt tips
Confirm one media path
Change one output control at a time
Shot structure / Content filter
Why choose it
What to know first
Adjustable parameters
Quickly assess whether this model fits your use case.
Credit usage
720p at 5/10/15 seconds costs 70/140/210 credits, and 1080p costs 105/210/315 credits. Fifteen seconds is unavailable with video input.
Before batch generation, run an A/B test with the same assets across the current model and alternatives to validate quality and avoid wasting credits.
FAQ
Related models
Compare these similar candidates before deciding.
Kling 2.6
Kling 2.6 is a two-workflow video service on Kyeo: text requests use text-to-video, while requests with an image use single-image animation. Both routes offer 5/10 seconds and a `sound` toggle; text generation also provides three aspect ratios. The four Kyeo credit tiers are 55, 110, 110, and 220.
Hailuo 2.3
Kyeo has six valid combinations: Standard 768P at 6/10 seconds and 1080P at 6 seconds cost 25/40/40 credits; the matching Pro prices are 40/80/70. 1080P does not support 10 seconds. Exactly 1 JPG/PNG/WebP file, up to 10MB.
Seedance 1.5 Pro
Seedance 1.5 Pro / Capabilities — Seedance 1.5 Pro on Kyeo AI accepts text alone or up to two images and exposes 480p/720p/1080p, 4–12 seconds, generated audio, six ratios, and a fixed-lens switch. The 7–180-credit cost follows the selected combination; the Pro name and higher credit cost do not prove image quality, stability, or publication readiness.
Sources
Content is based on model documentation, feature references, and the settings available on Kyeo AI.
Reviewed against Alibaba Cloud's official Wan 2.6 series announcement and official video-model overview. The public facts below distinguish text, image, and video inputs, callable durations, local credits, verified media limits, and visible controls; marketing claims are not treated as tested results.