kling-motion-control on APIMart uses a reference image plus a reference video to transfer subject motion into generated clips. Add an optional prompt, choose character orientation and mode, and control sound or watermark settings for kling motion workflows.
Video link valid for 72 hours
Transparent pricing with no hidden fees. Pay only for what you use.
* Actual costs are subject to final output.
Access kling-motion-control on APIMart for reference image + reference video generation. Upload the subject image and motion video, then choose orientation and quality mode for kling motion production.
50K+
Active Users
99.9%
Uptime
2x
Faster
70%
Cost Savings
Key features that make kling-motion-control different for kling motion teams
Real-world applications powered by reference image and reference video motion control for kling motion use cases and Kling creators
Generate motion-controlled clips in three steps
Register on APIMart and create your API key to access kling-motion-control.
Upload a reference image and reference video, then configure character orientation, mode, original sound, and watermark options for Kling generation.
Use the native motion-control endpoint with image_url, video_url, character_orientation, mode, and optional prompt or element_list fields for Kling app integration.
What teams are saying about kling-motion-control
“kling-motion-control is exactly what we needed for fast iteration. A reference image locks the subject, while a reference video gives us reliable motion timing.”
Sarah Johnson
Creative Director
“We dropped kling-motion-control into our pipeline and immediately cut integration time. The minimal API surface makes it a joy to scale.”
James Liu
Senior Developer
Common questions about kling-motion-control and kling motion implementation
kling-motion-control uses a reference image plus a reference video to transfer motion into generated clips. The playground exposes the required media fields and common generation settings.
Yes. The native motion-control route requires both image_url and video_url. The prompt is optional, while character_orientation and mode are required for Kling workflows and kling motion tasks.
Default output is 5 seconds. Backend extensions may expose longer durations later.
Pricing is charged per second and follows the selected generation mode. See the pricing block on this page for the latest rates.
Yes. Submit model, image_url, video_url, character_orientation, and mode to /kling/v1/videos/motion-control. The API returns a task ID you can poll for the final clip URL.
You can reach us via the live chat in the bottom-right corner, email us at [email protected], or join our Discord community. Our team will get back to you as soon as possible.
Explore more models in the same category.

HappyHorse
HappyHorse-1.0 is an AI model that generates videos based on text input.

SkyReels V4 Fast
SkyReels-V4-Fast is the first unified framework for joint audiovisual generation, restoration, and editing at cinema-grade quality, efficiently achieving 1080p, 15-second multi-shot video generation with synchronized audio and visuals through a dual-stream MMDiT architecture

Wan 2.7
Wan-2.7 is Alibaba’s next-generation multimodal AI video model that generates high-quality videos from text prompts, images, or reference footage. It supports text-to-video, image-to-video, and instruction-based video editing, producing short clips (up to ~15 seconds) with 720p–1080p resolution, realistic motion, and strong character consistency. 

ViduQ 3
Vidu Q3 is an advanced AI video generation model developed by Shengshu Technology that creates cinematic videos from text prompts or images. It supports both text-to-video and image-to-video workflows, generating clips up to around 16 seconds with synchronized native audio, including dialogue and sound effects.