fal.ai · Text to video
Kling 3.0 Pro
Kling 3.0 Pro, text to video, served by fal.ai. The OpenAPI schema for the fal-ai/kling-video/v3/pro/text-to-video queue. Price: $0.112 per second audio off, $0.168 with audio, $0.196 with audio and voice control: audio on is billed at the voice-control rate to never underbill. Asynchronous: returns a request_id; poll fal/status and fetch fal/result with model "fal-ai/kling-video".
From $0.112 / unitfal.aiVideofal/kling-v3-pro-text-to-videoListed
Activation pending; your agent will see it as unavailable until then.
Use this tool with your agent
Use my connected Agentik MCP server. I want to use the tool fal/kling-v3-pro-text-to-video (Kling 3.0 Pro, text to video). Call inspect_tool with {"tool_id": "fal/kling-v3-pro-text-to-video"} first and tell me the price before anything runs. Ask me for the inputs it needs; never invent values, IDs or file URLs. When I confirm, call run_tool with tool_id "fal/kling-v3-pro-text-to-video" and my inputs, follow it with get_run if it is still running, and give me the result and the price Agentik reports. If the tool is unavailable, say so and use discover_tools to propose an alternative.Show the prompt
Paste it into Claude, ChatGPT or Codex: your agent takes it from there, and shows the price before it runs.
Parameters
| Name | Type | Required | Description |
|---|---|---|---|
shot_type | string | No | The type of multi-shot video generation. 'intelligent' lets the model automatically determine shot structure. Options: customize, intelligent. |
prompt | string | No | Text prompt for video generation. Either prompt or multi_prompt must be provided, but not both. |
generate_audio | boolean | No | Whether to generate native audio for the video. Supports Chinese and English voice output. Other languages are automatically translated to English. For English speech, use lowercase letters; for acronyms or proper nouns, use uppercase. |
cfg_scale | number | No | The CFG (Classifier Free Guidance) scale is a measure of how close you want the model to stick to your prompt. |
duration | string | No | The duration of the generated video in seconds Options: 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14. |
multi_prompt | object[] | No | List of prompts for multi-shot video generation. If provided, overrides the single prompt and divides the video into multiple shots with specified prompts and durations. |
negative_prompt | string | No | |
aspect_ratio | string | No | The aspect ratio of the generated video frame Options: 16:9, 9:16, 1:1. |
How your agent calls it
First inspect_tool for the current schema and price, then run_tool with the published example after confirming your inputs:
{
"tool_id": "fal/kling-v3-pro-text-to-video",
"input": {
"prompt": "a red cube slowly rotating on a white table, soft studio light",
"duration": "3",
"aspect_ratio": "1:1",
"generate_audio": false
}
}Price and billing
- Price
- From $0.112 / unit
- Model
- Per unit
- Detail
- The provider’s price, with no markup. Billed for measured usage.
- The provider’s price, 0% markup.
- One prepaid balance for every provider; your agent sees the price before it runs.
- Pay for measured usage; unused reservations are released.