Skip to content

fal.ai · Image to video

Kling 3.0 Pro

Kling 3.0 Pro, image to video, served by fal.ai. The OpenAPI schema for the fal-ai/kling-video/v3/pro/image-to-video queue. Price: $0.112 per second audio off, $0.168 with audio, $0.196 with audio and voice control: audio on is billed at the voice-control rate to never underbill. Asynchronous: returns a request_id; poll fal/status and fetch fal/result with model "fal-ai/kling-video".

From $0.112 / unitfal.aiVideofal/kling-v3-pro-image-to-videoListed

Activation pending; your agent will see it as unavailable until then.

Use this tool with your agent

Give this to your agent
Use my connected Agentik MCP server. I want to use the tool fal/kling-v3-pro-image-to-video (Kling 3.0 Pro, image to video). Call inspect_tool with {"tool_id": "fal/kling-v3-pro-image-to-video"} first and tell me the price before anything runs. Ask me for the inputs it needs; never invent values, IDs or file URLs. When I confirm, call run_tool with tool_id "fal/kling-v3-pro-image-to-video" and my inputs, follow it with get_run if it is still running, and give me the result and the price Agentik reports. If the tool is unavailable, say so and use discover_tools to propose an alternative.
Show the prompt

Paste it into Claude, ChatGPT or Codex: your agent takes it from there, and shows the price before it runs.

Parameters

NameTypeRequiredDescription
promptstringNoText prompt for video generation. Either prompt or multi_prompt must be provided, but not both.
elementsobject[]NoElements (characters/objects) to include in the video. Each example can either be an image set (frontal + reference images) or a video. Reference in prompt as @Element1, @Element2, etc.
start_image_urlstringYesURL of the image to be used for the video
end_image_urlstringNoURL of the image to be used for the end of the video
generate_audiobooleanNoWhether to generate native audio for the video. Supports Chinese and English voice output. Other languages are automatically translated to English. For English speech, use lowercase letters; for acronyms or proper nouns, use uppercase.
cfg_scalenumberNo The CFG (Classifier Free Guidance) scale is a measure of how close you want the model to stick to your prompt.
durationstringNoThe duration of the generated video in seconds Options: 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14.
multi_promptobject[]NoList of prompts for multi-shot video generation. If provided, divides the video into multiple shots.
negative_promptstringNo
shot_typestringNoThe type of multi-shot video generation. 'intelligent' lets the model automatically determine shot structure. Options: customize, intelligent.

How your agent calls it

First inspect_tool for the current schema and price, then run_tool with the published example after confirming your inputs:

{
  "tool_id": "fal/kling-v3-pro-image-to-video",
  "input": {
    "prompt": "a red cube slowly rotating on a white table, soft studio light",
    "start_image_url": "https://storage.googleapis.com/falserverless/example_outputs/nano-banana-2-t2i-output.png",
    "duration": "3",
    "generate_audio": false
  }
}

Price and billing

Price
From $0.112 / unit
Model
Per unit
Detail
The provider’s price, with no markup. Billed for measured usage.
  • The provider’s price, 0% markup.
  • One prepaid balance for every provider; your agent sees the price before it runs.
  • Pay for measured usage; unused reservations are released.