Wan 2.2 A14B Turbo
Wan 2.2 A14B Turbo is not a single Kyeo workflow. No media selects text-to-video, one image selects image-to-video, and one image plus one audio file selects speech-to-video. Each workflow has different resolutions, credit rates, and visible controls, so confirm the media combination first.
At a glance
What is this model like?
Text and image workflows output a fixed five seconds: 40/60/80 credits at 480p/580p/720p. Speech output seconds equal `num_frames ÷ frames_per_second`, multiplied by 12/18/24 credits for the selected resolution; Kyeo rounds the request total up. Speech requires both image and audio, exposes frame, negative-prompt and inference controls, and sends the current `enable_safety_checker` field with a true default. The page does not promise generation speed, synchronization, image quality, or identity consistency.
Key facts
Quickly assess whether this model fits your use case.
Also known as
The same model may appear under different names across documentation and community discussions; this list keeps them easy to search and compare.
Common questions
These practical questions focus on the task, input conditions, and result requirements you should confirm before choosing.
Selection guide
Use these decision points when choosing a model.
Use no media for text, one image without audio for image-to-video, or one image plus one audio file for speech.
Text and image produce a fixed five seconds for 40/60/80 credits at 480p/580p/720p. Speech costs 12/18/24 credits per output second, calculated from `num_frames ÷ frames_per_second`, and Kyeo rounds the total up.
For speech mode, set `num_frames` to a multiple of four from 40 to 120 and frame rate from 4 to 60.
Popular comparisons
Compare common alternatives on the same task to clarify differences in inputs, controls, and cost.
Model comparison
Compare the current model with alternatives at a glance.
Practical usage insights
Practical guidance based on public sources, current on-site limits, and representative tasks.
current public source workflow names are not Alibaba direct model IDs
Marketing claims do not guarantee a specific result
Capabilities
Text and image modes
Speech-driven mode
Safety-field boundary
Use cases
Prompt-only shot checks
Single-image animation
Image-plus-audio motion
Prompt tips
Describe visible action first
Do not assume an aspect ratio in image mode
Separate audio from visual direction
Why choose it
What to know first
Adjustable parameters
Quickly assess whether this model fits your use case.
Credit usage
Text and image workflows output a fixed five seconds: 40/60/80 credits at 480p/580p/720p. Speech output seconds equal `num_frames ÷ frames_per_second`, multiplied by 12/18/24 credits for the selected resolution; Kyeo rounds the request total up.
Before batch generation, run an A/B test with the same assets across the current model and alternatives to validate quality and avoid wasting credits.
FAQ
Related models
Compare these similar candidates before deciding.
Wan 2.6
No media selects text-to-video, exactly one image selects image-to-video, and one to three videos select video editing; images and videos cannot be mixed. The default is 1080p, 5 seconds, shot structure Off, and content filtering On.
Wan 2.5 Video
Review Wan 2.5 Video on Kyeo AI with text-to-video or one-image input, 5/10 seconds, 720p/1080p, prompt constraints, and 60–200-credit pricing. Kyeo charges by resolution and duration: 60 credits for 720p at 5 seconds, 120 for 720p at 10 seconds, 100 for 1080p at 5 seconds, and 200 for 1080p at 10 seconds. The 720p 5-second default costs 60 credits. How are Wan 2.5's 60/100/120/200 credits calculated? — Review Wan 2.5 Video on Kyeo AI with text-to-video or one-image input, 5/10 seconds, 720p/1080p, prompt constraints, and 60–200-credit pricing.
Infinitalk
Infinitalk uses 1 image, 1 audio file no longer than 15 seconds, and a required prompt to generate a speaking-portrait video on Kyeo. It offers 480p or 720p plus optional `seed`; cost depends on audio duration and resolution rather than a fixed request price.
Sources
Content is based on model documentation, feature references, and the settings available on Kyeo AI.
Review Wan 2.2 A14B Turbo text, one-image and speech-driven workflows, fixed five-second tiers, per-second speech rates and mode-specific inputs. Text and image workflows output a fixed five seconds: 40/60/80 credits at 480p/580p/720p. Speech output seconds equal `num_frames ÷ frames_per_second`, multiplied by 12/18/24 credits for the selected resolution; Kyeo rounds the request total up.