Kling 2.5 Turbo Pro Text to Video API

kwaivgi/kling-v2.5-turbo-pro/text-to-video

Kling 2.5 Turbo Pro Text to Video 将文本提示词转化为 5 秒或 10 秒视频。您可以描述主体、动作、光影和镜头运动,并按需填写宽高比及负向提示词。

输入

775/2,500

输出
已就绪
5 秒 · 42 积分 ($0.210) / 视频
继续使用

示例

One continuous five-second realistic portrait of a single adult female cellist seated on a plain chair in an empty wood-paneled chamber-music hall. She wears a simple dark green concert dress. The cello stands correctly between her knees with its endpin on the floor. Medium frontal three-quarter framing includes her face, both hands, complete bow and most of the cello. From the first frame the bow hair already rests across the strings between fingerboard and bridge. Her right arm draws the same straight bow slowly in ONE direction for the entire shot, maintaining contact with the strings; her left hand remains steadily on the neck. Her focused gaze follows the bow. Subtle slow camera push-in, soft warm side light and a quiet background. Preserve rigid instrument and bow geometry. No audience, no cuts, no text or logos.

A charming hand-drawn 2D cartoon of a European mole catching a potato. Fixed side view inside a cozy underground cellar. The mole sits on the left facing right, with a compact round charcoal velvet body, a smooth round head, tiny eyes, small pink snout and two very broad pink spade-shaped digging paws. Both paws are initially empty and held open at floor level. A small whole brown unpeeled potato starts on the far right, separated from the mole by a visible gap. During the first three seconds it rolls smoothly RIGHT TO LEFT across the floor, closing the gap. The mole cups its broad paws around the arriving potato and brings it to a stop against its body. It holds the same potato still for the final two seconds. Warm flat colors, clean outlines, stable earth walls and wicker baskets, one continuous five-second shot.

Kling 2.5 Turbo Pro Text to Video

Kling 2.5 Turbo Pro Text to Video 支持通过必填 prompt 描述视频内容,去除首尾空白后最多 2,500 个 Unicode 字符。您可以按镜头需要选择 5 秒或 10 秒,并通过可选 negative_prompt 描述希望避免的内容。5 秒为 42 积分($0.210),10 秒为 84 积分($0.420),可用于广告创意、动态分镜与概念视频的尝试。

核心优势

  • 文本驱动视频生成通过提示词描述主体、环境、动作与镜头运动,生成视频任务。

  • 可选负向提示词使用 negative_prompt 描述希望避免的内容;接口文档未规定其长度上限。

  • 2,500 字符提示词prompt 为必填非空字符串,去除首尾空白后最多 2,500 个 Unicode 字符。

  • 5 秒与 10 秒时长duration 支持 5 或 10,省略时默认 5 秒。

  • 明确的按次计费5 秒为 42 积分($0.210),10 秒为 84 积分($0.420),任务失败时退还所扣积分。

参数列表

参数必填说明
prompt

必填文本提示词,去除首尾空白后最多支持 2,500 个 Unicode 字符。

默认值-
duration

时长仅支持整数值 5 或 10,省略时默认 5 秒。5.0 和 10.0 按整数值处理;字符串、布尔值、非整数小数和 null 会被拒绝。

默认值5
aspect_ratio

aspect_ratio 为可选字符串,接口文档未规定具体枚举或默认值;不使用时省略。

默认值-
negative_prompt

negative_prompt 为可选字符串,用于描述希望避免的内容。接口文档未规定其长度上限,2,500 字符限制仅适用于 prompt。

默认值-

使用步骤

  1. 撰写详细场景提示词使用中英文自然语言描述主体外貌、动作发展、光影氛围及摄影机运动,支持长达 2,500 字符。

  2. 选择生成时长与画幅比例选择 5 秒或 10 秒时长。aspect_ratio 为可选字符串,接口文档未规定具体枚举或默认值;不使用时省略。

  3. 配置负向提示词过滤瑕疵在 negative_prompt 中添加如模糊、肢体畸变、低分辨率、水印等关键词,提升成片纯净度。

  4. 在体验区运行或通过 REST API 提交在网页工作台点击立即运行快速预览,或向 /api/generate/submit 端点发起 POST 请求发起批量任务。

  5. 轮询任务状态并下载成片使用 task_id 轮询 GET /api/generate/status/{task_id}。finished 时读取并下载视频文件;failed 时停止轮询并查看错误信息。

计费标准

按生成的视频数量计费。1 积分 = $0.005。

计费项价格单价明细
5 秒时长42 积分/次$0.210/次
10 秒时长84 积分/次$0.420/次

应用场景

  • 社交媒体爆款与广告视频创意快速批量生成符合 TikTok、抖音、Instagram Reels 等平台规格的商业推广视效与动态口播背景。

  • 影视导演分镜与动画动态预演将剧本文案快速具象化为动态镜头序列,用于影视团队评估分镜节奏、景别调度与光影氛围。

  • 电商商品动态氛围与品牌视觉无需搭建实景影棚,即可通过文本生成充满高级质感的商品生活化应用场景与艺术动态背景。

  • 创意设计工作室快速概念推演为艺术总监和视觉创作者提供高频迭代手段,在提案阶段直接向客户展示生动的动态概念视频。

实战技巧

  • 明确运镜方式:在提示词中加入专业摄影机运动术语(如“缓慢推镜头”、“低角度平移追踪”、“柔和俯仰环绕”),能有效激活镜头动感。
  • 强化环境与光影质感:使用“黄金时刻丁达尔光”、“电影级明暗对照”、“柔和轮廓边缘光”等词汇,大幅提升画面纵深感。
  • 采用时序递进结构:使用清晰的句式按时间顺序描述动作演进(如“角色转身,随后抬头望向远方天空”),引导模型自然生成连续动作。
  • 合理匹配镜头时长:对于单一步伐或突发性动作选用 5 秒,而对于叙事延展、环境漫游等较复杂镜头选用 10 秒以充分展开动作。
  • 巧用负向提示词净化画面:通过 negative_prompt 排除“画面重影、肢体变形、文字噪点、过曝”等瑕疵,保证专业质感。

注意事项

  • 时长校验:时长仅支持整数值 5 或 10,省略时默认 5 秒。5.0 和 10.0 按整数值处理;字符串、布尔值、非整数小数和 null 会被拒绝。
  • 提示词字符限制:prompt 去除首尾空白后最多 2,500 个 Unicode 字符。negative_prompt 为可选字符串,用于描述希望避免的内容。接口文档未规定其长度上限,2,500 字符限制仅适用于 prompt。
  • 按次计费与失败自动退款:任务提交时扣减对应积分(5 秒 42 积分,10 秒 84 积分),若因系统异常导致生成失败,已扣积分将全额自动返还。

Kling 2.5 Turbo Pro Text to Video API 常见问题

Kling 2.5 Turbo Pro Text to Video API 是什么?

Kling 2.5 Turbo Pro Text to Video 将文本提示词转化为 5 秒或 10 秒视频。您可以描述主体、动作、光影和镜头运动,并按需填写宽高比及负向提示词。您可以在上方体验区试用,或通过 REST API 提交任务。

Kling 2.5 Turbo Pro Text to Video 一次能生成多长时间?

Kling 2.5 Turbo Pro Text to Video 支持 5 秒与 10 秒两种成片时长。您可以在请求参数中使用 duration 字段指定整数值 5 或 10;若未传入该字段,系统默认生成 5 秒视频。

Kling 2.5 Turbo Pro Text to Video 支持哪些画面比例?

aspect_ratio 为可选字符串,接口文档未规定具体枚举或默认值;不使用时省略。

Kling 2.5 Turbo Pro Text to Video 提示词长度限制是多少?

prompt 字段最多支持 2,500 个 Unicode 字符(去除首尾空白字符后计算)。充裕的字数空间能够充分容纳详细的故事情节、镜头调度说明、环境光影及特定风格修饰。

Kling 2.5 Turbo Pro Text to Video 如何使用负向提示词?

negative_prompt 为可选字符串,用于描述希望避免的内容。接口文档未规定其长度上限,2,500 字符限制仅适用于 prompt。

Kling 2.5 Turbo Pro Text to Video 的生成费用是多少?

模型按生成成片数量计费:5 秒视频每次消耗 42 积分($0.210),10 秒视频每次消耗 84 积分($0.420)。费用在提交任务时扣除,若任务因内部异常未能生成有效视频,系统会自动将扣减积分全额退回账户。

Kling 2.5 Turbo Pro Text to Video 如何查询生成结果?

提交后保存 task_id,通过 GET /api/generate/status/{task_id} 查询任务。状态为 finished 时读取 data.files 中的视频链接;状态为 failed 时停止轮询并查看错误信息。

Kling 2.5 Turbo Pro Text to Video 是否支持上传图片素材?

不支持。该端点专为纯文本提示词生成视频设计。如果您需要上传静态图片或使用首尾关键帧引导生成动态视频,请使用专门的 Kling 2.5 Turbo Pro Image to Video 端点。