APIXO groups Standard image-to-video, Pro image-to-video, Master text-to-video, and Master image-to-video. The route therefore covers three Kling 2.1 quality tiers, but not every tier supports the same input workflow.
Kuaishou’s Kling AI 2.1 improves motion performance, physical simulation, and semantic responsiveness for short-form video generation. APIXO groups four workflows across its Standard, Pro, and Master tiers on one route, covering single-image animation, Pro first-and-last-frame transitions, and Master text-to-video or image-to-video production.
APIXO Modes
Standard / Pro / Master
APIXO Price
From $0.20 / videoStandard image-to-video, 5 seconds
Duration
5 or 10 seconds
Image Inputs
1–2 imagesMode-dependent
Released
May 2025
Loading workspace...
Kling AI 2.1 improves the coherence and quality of subject movement across short generated scenes. This supports more natural character action, object motion, and camera-led compositions.
More realistic physical simulation helps generated subjects and objects interact more plausibly with movement, weight, contact, and the surrounding environment. Complex interactions should still be reviewed before production use.
Improved semantic responsiveness helps the model interpret detailed instructions about subjects, actions, camera direction, lighting, atmosphere, and scene development with greater accuracy.
Pro image-to-video accepts one start frame or a start frame plus one end frame. The model generates the transition between the supplied endpoints, supporting defined entrances, exits, and visual reveals.
Kuaishou describes Kling AI 2.1 as improving motion performance and realistic physical simulation over the preceding generation, supporting more coherent movement and interactions within compact generated scenes.
APIXO supports prompts up to 5,000 characters, negative prompts up to 500 characters, and CFG scale from 0 to 1. These controls balance instruction adherence against creative variation.
Inputs and framing depend on the selected tier and workflow.
4
Exposed Modes
5s / 10s
Clip Lengths
16:9 / 9:16 / 1:1
Master T2V Ratios
JPG / PNG / WebP
Image Formats
10 MB per image
Image Limit
1 MP4 URL
APIXO Output
Turn product photography, campaign artwork, or editorial stills into five- or ten-second motion assets. Standard supports economical variation, while Master image-to-video provides the premium route for higher-priority shots.
Use Pro image-to-video with a starting frame and optional ending frame to prototype transformations, time shifts, environment changes, or visual reveals. The model generates the transition rather than performing deterministic interpolation.
Use Master text-to-video to explore scenes without preparing source artwork. Specify subject action, environment, camera direction, lighting, and timing, then select horizontal, vertical, or square framing for the intended placement.
Prototype character movement, object behavior, camera motion, and scene transitions before filming or animation. Review complex contact, rapid action, anatomy, and continuity carefully before treating a generated clip as production-ready.
Select the mode before preparing images, framing, and delivery.
Select Standard for one-image animation, Pro for optional end-frame control, or Master for text-to-video and premium one-image animation. Pricing changes with both tier and duration.
Write a non-empty prompt and choose five or ten seconds. Add one image for Standard or Master image-to-video, or one to two images for Pro. Master text-to-video accepts no images.
Use APIXO polling or a webhook callback to receive the asynchronous result. Download completed MP4 output promptly because the documented result URL expires after 15 days.
Every mode requires a prompt between 1 and 5,000 characters and a duration of 5 or 10 seconds.
Standard and Master image-to-video require exactly one publicly accessible image URL.
Pro image-to-video accepts one start frame or a start frame plus one end frame.
Aspect-ratio selection applies only to Master text-to-video on the documented APIXO route.
APIXO does not expose audio generation, video extension, video input, lip sync, or motion-control modes here.
Outputs remain subject to safety moderation, and completed result URLs expire after 15 days.
APIXO groups Standard image-to-video, Pro image-to-video, Master text-to-video, and Master image-to-video. The route therefore covers three Kling 2.1 quality tiers, but not every tier supports the same input workflow.
Standard image-to-video costs $0.20 for five seconds or $0.40 for ten. Pro image-to-video costs $0.35 or $0.70. Each Master workflow costs $1.00 for five seconds or $2.00 for ten.
Yes, but only through APIXO’s Pro image-to-video mode. Supply one image for the starting frame and optionally a second for the ending frame. Standard and Master image modes accept exactly one image.
Kuaishou officially positioned Kling AI 2.1 Standard at 720p and its High Quality tier at 1080p. APIXO names Standard, Pro, and Master modes but does not publish an explicit resolution field or complete resolution mapping for this route.
Inspect subject identity, hands, anatomy, embedded text, object contact, rapid motion, and the continuity of complex transitions. Start-and-end-frame guidance defines endpoints but does not guarantee a physically exact path between them.
Discover more AI models for your next creative workflow