Kling 2.5 Turbo Image to video API

kuaishou / 2.5 Turbo

Generate a short video from text alone or animate a supplied opening frame with optional end-frame guidance.

Choose text-to-video for prompt-led composition, or image-to-video when a source frame should anchor the subject and framing.

Pricing

Pro pricing

All listed rates are for Pro mode. Standard is currently unavailable.

Pro price
$0.042/s
Input

Prompt, assets, and output parameters form one generation request.

Describe the subject motion, scene action, camera movement, and visual style.
Input media

Upload the media inputs configured for this model.

DurationChoose a 5 or 10-second output.
ModePro is the only mode currently available. Standard is coming soon.

Playground Examples

Try production-style prompt starters.

Winged riders race across towering sunlit clouds as the camera tracks beside them, cinematic fantasy realism.

Model overview

Kling 2.5 Turbo text and image video generation

Kling 2.5 Turbo on sjolt provides separate text-to-video and image-to-video routes for 5 or 10-second Pro clips, with optional final-frame control for image animation.

Kling 2.5 Turbo generation example
01 · Video generation

Prompt-led text-to-video

Describe the subject, action, environment, camera movement, and style, then choose a landscape, portrait, or square composition.

  • Supports Image to video and reference-asset generation workflows.
  • Fits product assets, ad previews, visual direction, and social content tests.
  • Adjust configured inputs, media, and switches before generation.
Reference asset control example
02 · Asset control

Start and end frame guidance

Image-to-video anchors the opening composition to one required image and can use a second image to guide the closing frame.

Start frameEnd frame
Unified model API and result management example
03 · API integration

Pro mode

Generate in Pro mode with 5 or 10-second duration choices. Standard is visible in the controls but remains unavailable.

  • Compare available models from Black Forest Labs, Google, MiniMax, OpenAI, ByteDance, Kuaishou, SJolt AI in one place.
  • Copy the request body directly to the server to reduce frontend/backend parameter drift.
  • Failure states, retries, result preview, and review checkpoints stay in the same workflow.

Model characteristics

Built for concise video creation with direct inputs.

01

Text to video

Generate a complete scene from prompt direction with landscape, portrait, or square framing.

02

Image to video

Animate one required opening image while the prompt directs subject, scene, and camera motion.

03

End-frame guidance

Add an optional closing image when the clip should transition toward a specific final composition.

04

Two durations

Choose a 5-second clip for a compact motion beat or 10 seconds for a longer action and camera progression.

FAQ

These are the first questions to answer when evaluating this model.

Which Kling 2.5 Turbo routes are available?

sjolt exposes separate text-to-video and image-to-video routes. Both currently generate in Pro mode.

Which generation modes are supported?

Pro is currently available. Standard appears in the selector but is disabled.

Which durations are supported?

Choose either 5 or 10 seconds.

Can I control the first and last frame?

Yes. Image-to-video requires image_url for the opening frame and accepts optional end_image_url for the closing frame.

Which image formats can I upload?

The Playground accepts one JPG, JPEG, PNG, or WebP start frame and one optional end frame.