Kling O3 Image Edit API

kwaivgi/kling-image-o3/edit
1K / 2K / 4K · 1–9 张

Kling O3 Image Edit 结合多达 10 张参考图像与自然语言指令完成精准图像编辑,支持多参考图高维特征融合、免遮罩局部精准重绘,以及原图画幅比例自适应延伸。它能够在执行复杂的角色换装、背景替换或多素材合成时,深度保留输入图像的核心主体特征、五官轮廓与物理光影质感,呈现浑然一体的高保真编辑效果。

获取 API Key
输入
1036/2000
2/10
Reference image 1: Reference 1
Reference image 2: Reference 2
Elements

按顺序使用 @Element1、@Element2 引用主体。可提供主视图与补充视角;仅上传文件时需登录。

输出已就绪
1 × 3.5 = 3.5 积分 · $0.018

继续使用

示例

edit-01-output.png

Use image 1 as the exact reference for the woman and her skateboard, and image 2 as the environment and composition to preserve. Place the woman naturally on the flat central foreground apron of the skatepark in image 2, slightly left of center, full body visible. Keep her recognizable face, short curls, olive helmet, burgundy T-shirt over white sleeves, beige trousers, black shoes and the turquoise skateboard with cream wheels. She stands casually with both feet firmly on the ground, holding her skateboard vertically on the same side and with the same hand as in image 1; do not invent an action trick. Match the camera height, adult scale and late-afternoon light entering from the left in image 2; add realistic contact shadows beneath both shoes and the board. Preserve the bridge columns, right-hand bowl, left bank, distant trees and the overall park layout. Documentary sports photograph, natural skin and fabric. Exactly one person and one skateboard. No additional people, animals, text, brands, advertising or watermark.

edit-02-output.png

Edit this attic photograph with precisely localized changes. Remove every cardboard moving box on the left side of the window and restore the uninterrupted oak floor beneath them. In that cleared area place one comfortable moss-green fabric armchair angled slightly toward the center of the room, and one slender dark-bronze floor lamp with a small ivory shade immediately behind the chair. Fit both naturally under the sloping ceiling with believable proportions and floor contact shadows. Keep the exact original camera position and framing, roof slope, all exposed beams, square window and wooden frame, floorboard directions, foreground woven rug, right-hand built-in shelf and existing books. Preserve the original daylight direction and realistic photographic appearance; the new lamp is switched off. Do not renovate other surfaces or add decorations. No people, animals, boxes, typography, brands, advertising or watermark.

edit-03-output.png

Transform this mountain village photograph into a meticulous handmade layered-paper artwork while preserving the original scene layout and camera framing. Keep the same cluster of tiled-roof houses in the upper left, the S-shaped path from the bottom center to the village, the sweeping terraces on the right, and the distant hills along the top. Construct every terrace level from a distinct stacked sheet with crisp cut edges, visible paper thickness and soft cast shadows between layers. Render the houses as tiny folded-paper buildings and trees as simple carefully cut silhouettes. Use warm ivory, muted forest green, ochre and dusty blue cardstock with subtle fiber texture. Directional studio light from upper left reveals the relief without flattening the recognizable topography. No new buildings or roads, no people, animals, writing, labels, border, frame, logos or watermark.

Kling O3 Image Edit 概览

Kling O3 Image Edit 是快手(Kling AI)专为复杂图像重构与多源素材融合打造的旗舰级多模态图像编辑模型。它全面支持 1 至 10 张参考图片输入,具备卓越的多维特征解耦合成、免遮罩自然语言局部重绘、auto 原图比例自适应继承与最高 4K 超高清原生重绘能力。无论是电商多图换景换装、创意合成,还是系列故事分镜调整,均能在保持核心主体一致性的同时实现精准创作。

为什么选择 Kling O3 Image Edit?

  • 多达 10 张参考图高维融合支持传入 1 至 10 张图像素材,模型能够智能提取不同图片中的主体身份、构图线索、艺术风格与特定材质,完成自然有机的高维特征综合重构。

  • 免遮罩自然语言局部精修无须手动绘制涂抹蒙版,仅需通过文字清晰阐述修改意图,即可实现发丝级精细换装、道具增删与局部光照重绘,周边像素丝毫不乱。

  • auto 原生比例自适应专属支持 auto 比例模式,可完美继承源图宽高比,也可自由选择 8 种常规画幅并智能延展周边透视,绝不发生画面拉伸变形。

  • 深度主体与光照特征延续在执行换背景、跨场景迁移或多图要素融合时,深度锚定主体的面部五官、固有材质与环境光线相互作用,确保前后视觉资产的一致性。

  • 透明计费且零输入图附加费与文生图保持完全相同的阶梯单价,不按传入的参考图片张数额外加收附加费,任务失败自动全额返还扣除积分。

参数

参数要求说明
prompt必填

字符串,去除首尾空白后 1–2,000 字符。

image_urls必填

1–10 个公开 HTTP(S) 图片 URL,按顺序传递。

resolution可选

输出分辨率,按所选档位计费。

默认1K2K4K
size可选

画面比例;auto 仅编辑端点支持。 auto, 16:9, 9:16, 1:1, 4:3, 3:4, 3:2, 2:3, 21:9

默认16:9auto9:161:14:33:43:22:321:9
output_format可选

输出图片格式。

默认pngjpegwebp
n可选

输出数量,整数 1–9;总积分为单价乘以 n。

默认1
elements可选

主体参考对象数组。在提示词中使用 @Element1、@Element2 等引用;无公开数量上限,保留扩展字段。

elements[].frontal_image_url可选

主体主视图的公开 HTTP(S) 图片 URL。

elements[].reference_image_urls可选

同一主体的补充视角 HTTP(S) 图片 URL 数组。

使用教程

  1. 准备参考素材图片整理 1 至 10 张清晰的公开 HTTP(S) 图片 URL(支持人像、商品图、背景环境或风格样片)。

  2. 编写明确的编辑指令用自然语言清楚指明需要改变的内容以及需要保留的边界(例如“保留图 1 中人物的容貌与姿态,换上图 2 中的复古西装外套,背景保持不变”)。

  3. 设定输出分辨率与画幅选择目标分辨率(1K、2K 或 4K),画幅可选择 auto 继承原图比例,或指定标准比例(如 16:9、1:1 等)。

  4. 指定文件格式与输出张数配置交付图片格式(PNG、JPEG 或 WebP)与单次生成张数 n(1–9 张)。

  5. 提交任务并获取修改成品调用 API 发起异步编辑任务,通过轮询或 Webhook 回调获取最终高保真输出图像。

价格

1K/2K:3.5 积分/张(约 $0.018);4K:7 积分/张($0.035)。总积分 = 分辨率单价 × n,无输入图片或主体参考附加费。1 积分 = $0.005,3.5 积分精确等于 $0.0175;美元显示保留三位小数,总额计算后再舍入。

用量费率详情
1K / 2K3.5 积分/张 · 约 $0.018乘以输出数量 n。
4K7 积分/张 · $0.035乘以输出数量 n。

最佳应用场景

  • 电商商品模特换装与多素材融合上传模特图、服装平铺图与棚拍背景图,一键完成三者自然融合并重绘光影阴影,大幅降低商业拍摄成本。

  • 多视角与多参考角色定型结合多张不同角度或不同表情的参考照片,生成在全新动作与场景下高度统一的人物写真。

  • 画面画幅智能自适应重构借助 auto 保持原始比例,或将横版视觉资产无损重构为 9:16 移动端竖屏海报,智能补全周边环境透视。

  • 局部细节精修与无痕移除替换精准增减画面中的道具配饰、修改特定发色发型或替换背景为洁净工作室布景,边缘自然过度。

专家技巧

  • 传入多张参考图时,建议在提示词中通过明确的文字描述区分各图角色(如“按照第一张图的人物面部,穿着第二张图展示的驼色大衣”)。
  • 进行局部修改时,明确界定“改变”与“保留”的范围,例如“仅将手提包替换为纸质咖啡杯,保持人物站姿、面部表情以及背景虚化完全不变”。
  • 当需要完整沿用原始图像的构图与尺寸时,将 size 参数设置为 auto,模型将自动对齐源图比例生成结果。
  • 准备参考图时建议优先提供主体清晰、光线明亮的图片,有助于多模态特征解耦算法更精确地捕捉主体纹理。

注意事项

  • image_urls 为必填数组,必须包含 1 至 10 个有效的公开 HTTP(S) 图像 URL。
  • prompt 为必填参数,去除首尾空格后字符数必须在 1 至 2,000 字符之间。
  • 单次输出图片张数 n 范围为 1 至 9,生成费用按输出张数结算,不收取输入图附加费。
  • 任务执行采用异步处理机制,可根据 task_id 轮询状态或设置 callback_url 接收推送通知。

相关模型推荐

Kling O3 Image Edit API 常见问题

Kling O3 Image Edit API 是什么?

Kling O3 Image Edit 是快手(Kling AI)研发的多模态图像精准编辑与多图合成模型。它结合 1 至 10 张参考图片与自然语言指令,支持多参考特征融合、免遮罩局部重绘与 auto 画幅自适应扩展。基于先进的多模态特征解耦与扩散重构架构,它能够在完成角色换装、换景或元素增删的同时,深度保留主体的五官结构、固有材质与环境光照一致性。你可以通过 API 进行程序化调用,也可以在上方体验区直接在线试用。

Kling O3 Image Edit 一次最多支持传入多少张参考图片?

支持传入 1 至 10 张公开 HTTP(S) 图像 URL。模型能够同时解析多张图片中的主体形象、色彩风格、构图透视与细节道具,并在单次编辑任务中将它们有机融合成一张全新画面。

Kling O3 Image Edit 进行局部修改时如何保留原图人物特征与光影?

模型具备前沿的语义驱动区域定位能力。只需在提示词中明确说明修改目标与保留范围,算法便会自动在特征空间实施针对性重绘,原图的主体身份轮廓、皮肤纹理以及环境反光均会获得完整保留。

Kling O3 Image Edit 支持保持原图画幅比例吗?

支持。通过将 size 参数设置为 auto,模型会自动检测并遵循输入首张图像的原始宽高比输出成品;此外也支持选择 16:9、1:1 等 8 种常规画幅并对周边场景进行无缝延伸。

Kling O3 Image Edit 上传参考图片需要额外收费吗?

不需要。Kling O3 Image Edit 与文生图保持相同的费率标准,不设任何输入参考图片附加费。费用仅按最终输出的图片张数和选定的分辨率档位核算,若任务失败全额自动返还积分。

Kling O3 Image Edit 支持哪些分辨率与输出图片格式?

支持 1K、2K 和 4K 原生输出分辨率,文件格式支持 PNG、JPEG 和 WebP(默认 PNG)。即使在复杂的局部重绘任务中,也能输出细节锐利的原生高保真成品。

如何编写 Kling O3 Image Edit 的多图融合提示词以获得最佳效果?

建议使用“指定对象 + 关联参考来源 + 执行动作 + 约束保留范围”的句式结构,例如“以图 1 中的人物为主体,穿上图 2 所示的黑色皮夹克,置于图 3 的都市街道夜景中,保持人物面部五官与原有视线完全不变”。