Seedance 1.5 Pro Text to Video API

bytedance/seedance-v1.5-pro/text-to-video

Seedance 1.5 Pro Text to Video 将文本提示词转化为视频,支持同步音频、电影感运镜与人物表现。通过场景、对白和动作描述引导画面发展,让人物表情、动作与声音构成连贯叙事短片。

输入
889/2500
输出已就绪
720p · 8 秒 · 有声 · 64 积分 = $0.320

示例

Four-second realistic wildlife macro video. Vertical composition, locked camera, tropical rainforest daylight. TWO SEPARATE broad wet leaves are fully visible, one at upper left and one at lower right, separated by a small clear AIR GAP. Exactly ONE red-eyed green tree frog with orange toes crouches at the very tip of the LEFT leaf, facing the RIGHT leaf. The main event is a SINGLE HOP ACROSS THE GAP: within the first second it pushes off with both hind legs; its entire body and all four feet visibly leave the left leaf and travel through the air; it lands once on the separate right leaf by the second second. The right leaf briefly bends under the landing, releasing a few droplets. It then sits still for the remaining time. Show the complete takeoff, airborne body and landing in the same shot. No crawling along a leaf, no extra frog, no transformation. Natural quiet rain with one soft wet-leaf tap at landing; no music, speech, text or logos.

A single continuous four-second cinematic widescreen science-fiction shot inside an unoccupied spacecraft cargo bay in zero gravity. Exactly one ordinary silver open-ended wrench floats freely at the center, slowly translating a few centimeters and rotating smoothly about its long axis. Its rigid metal shape and both ends remain unchanged. A slow short lateral camera slide produces clear parallax between strapped cargo rails in the foreground and a large round porthole behind. The blue curved limb of Earth is visible through the porthole, soft reflected blue light across brushed metal and warm practical lights. All cargo is securely strapped down. Only a quiet steady interior ventilation hum, no wrench impact sound, no explosive effects, no people, speech, music, writing, logos or subtitles.

Seedance 1.5 Pro Text to Video

Seedance 1.5 Pro Text to Video 是字节跳动 Seed 团队研发的文本生成视频模型,采用音视频联合生成机制,将场景、对白与动作描述转化为画面和声音相互配合的短片。通过提示词组织人物表情、动作顺序和环境氛围,并结合固定镜头或运动镜头,适合短剧分镜、广告创意与社交内容制作。

核心能力

  • 从文字构建场景描述场景中的人物、动作和镜头路径,组织一段完整的画面。

  • 声音与动作共同生成开启 Generate Audio,让对白、环境声与动作音效随视频一起生成。

  • 镜头运动表达通过跟拍、缓慢推进或特写,将主体动作与叙事节奏连接起来。

  • 控制镜头与画幅选择画幅、分辨率和时长;需要静止机位时开启 Fixed Lens。

参数说明

参数要求说明
prompt必填

用 3–2500 个 Unicode 字符描述场景、动作、镜头和声音;计数前会去除首尾空白。

aspect_ratio必填

选择输出画幅:1:1、21:9、4:3、3:4、16:9 或 9:16。

1:121:94:33:416:99:16
resolution可选

选择 480p、720p 或 1080p;省略时 Vidgo 默认使用 720p。

默认值720p480p1080p
duration必填

视频时长使用整数 4、8 或 12,单位为秒。

4812
fixed_lens可选

设为 true 时固定镜头,设为 false 时允许镜头运动。

truefalse
generate_audio可选

设为 true 生成同步音频,设为 false 生成静音视频;省略时 Vidgo 默认使用 true。

默认值truefalse

使用步骤

  1. 描述场景在 Prompt 中写清主体、环境与动作顺序。

  2. 描述运动与声音描述主体动作、镜头运动,以及希望出现的对白或环境声音。

  3. 选择输出配置设置 Aspect Ratio、Resolution 和 Duration,再选择 Generate Audio 与 Fixed Lens。

  4. 生成与查看确认显示的费用后点击 Run,完成后预览或下载视频。

用量与计费

每条视频根据分辨率、时长和音频设置计费,1 点数等于 $0.005,两个端点采用相同价格。

计费项费率说明
480p · 4 秒 · 静音9 点数 / $0.045每条视频
720p · 4 秒 · 静音16 点数 / $0.080每条视频
480p · 8 秒 · 静音;480p · 4 秒 · 有声18 点数 / $0.090每条视频
480p · 12 秒 · 静音21 点数 / $0.105每条视频
720p · 8 秒 · 静音;720p · 4 秒 · 有声32 点数 / $0.160每条视频
480p · 8 秒 · 有声36 点数 / $0.180每条视频
1080p · 4 秒 · 有声或静音40 点数 / $0.200每条视频
480p · 12 秒 · 有声;720p · 12 秒 · 静音42 点数 / $0.210每条视频
720p · 8 秒 · 有声64 点数 / $0.320每条视频
1080p · 8 秒 · 有声或静音70 点数 / $0.350每条视频
720p · 12 秒 · 有声84 点数 / $0.420每条视频
1080p · 12 秒 · 有声或静音100 点数 / $0.500每条视频

应用场景

  • 短剧场景在一段短片中安排对话、场景与人物反应。

  • 广告创意将产品场景文案转化为带有同步声音的动态演示。

  • 社交叙事选择竖屏或方形画幅,围绕一个动作构建清晰的视觉片段。

提示词建议

  • 按时间顺序描述动作,并让整段视频的镜头指令保持一致。
  • 为对白加上引号,并写清每句台词的说话者。
  • 生成音频时,在画面动作之外补充环境声音的描述。

相关模型

Seedance 1.5 Pro Text to Video API 常见问题

Seedance 1.5 Pro Text to Video API 是什么?

Seedance 1.5 Pro Text to Video API 是字节跳动用于将文本描述转化为视频的模型接口。它将生动的动作、同步声音与镜头表达结合在短片中。音视频联合生成机制根据场景、对白和动作指令组织叙事,并提供固定镜头控制。你可以通过 API 进行程序化调用,也可以在体验标签中直接在线试用。

Seedance 1.5 Pro Text to Video 如何生成同步音频?

将 Generate Audio 设为 true,并在 Prompt 中描述对白、环境声或动作音效。声音与画面共同生成,让音效随场景动作展开。

Seedance 1.5 Pro Text to Video 的对白提示词怎么写?

先写清说话者,再用引号标出每句台词,并描述语气与语速。模型能够生成多种语言和方言的声音,并让嘴部动作与语音配合。

Seedance 1.5 Pro Text to Video 怎样遵循动作顺序?

先建立人物与环境,再按时间顺序描述动作。让每一步动作围绕同一主体展开,并使用与场景发展一致的镜头指令。

Seedance 1.5 Pro Text to Video 可以生成 12 秒场景吗?

可以,将 Duration 设为 12,即可为动作或对白安排更长的片段;也可以选择 4 或 8 秒,并让提示词中的动作数量与时长相匹配。

Seedance 1.5 Pro Text to Video 怎样固定镜头?

将 Fixed Lens 设为 true,使用固定机位构图,并通过主体动作表现运动;需要镜头跟随提示词运动时设为 false。

Seedance 1.5 Pro Text to Video 可以生成 1080p 视频吗?

可以,在 Resolution 中选择 1080p,并设置 4、8 或 12 秒时长。480p 与 720p 输出也使用这三档时长。