Seedance 2.0 Text-to-Video

seedance-2.0/text-to-video

Seedance 2.0 Text-to-Video turns a written scene into a 4–15 second video without source media. Direct the subject, action, camera, light, atmosphere, and optional audio across four resolutions through 4K.

Input

Prompt
134/20000
Resolution
Duration
Aspect ratio
Advanced

Generate audio

Ask the model to generate an audio track with the video.

Return last frame

Keep the video result and request the final frame as an additional file.

Web search

Send the optional web-assisted generation field.

Output

Idle

Your generated files will appear here

Set the inputs, choose a duration, then run the asynchronous video task.

5 sec × $0.200/sec = $1.000

Continue with

Examples

One continuous four-second shot in an empty black-box theater. A fictional adult performer wearing a plain silver face mask and a fitted indigo rehearsal suit holds one long unprinted ivory ribbon. Beat one: the performer steps forward and draws one low horizontal ribbon arc. Beat two: they pivot once and lift the same ribbon into a vertical spiral. Beat three: they stop in a balanced stance while the ribbon settles behind the right shoulder. The camera makes one smooth waist-height semicircle from front-left to front-right. Preserve the same mask, body proportions, suit, ribbon length, hand count, stage geometry, movement direction, and lighting throughout. Use natural weight transfer and physically believable cloth motion. Synchronized audio: three soft footfalls, ribbon movement through air, and restrained room reflections. No cuts, no extra people, no dialogue, no music, no readable text, no logos, no brands, no products, no advertising, no watermark.

Seedance 2.0 Text-to-Video

Seedance 2.0 Text-to-Video turns a written scene into a 4–15 second video without source media. Direct the subject, action, camera, light, atmosphere, and optional audio across four resolutions through 4K.

Why Choose This?

  • Start from text alone.Describe a complete shot without sending image, video, or audio media.

  • Full resolution range.Choose 480p, 720p, 1080p, or 4K according to the delivery target.

  • Continuous duration control.Select any whole-second duration from 4 through 15 instead of a short preset list.

  • Explicit advanced controls.Request audio, a returned last frame, web assistance, and an optional integer seed.

Parameters

ParameterRequirementDescription
promptRequired

Trimmed length 1–20,000; the Playground supplies a mode-specific example.

durationRequired

Any integer from 4 through 15 inclusive; Playground default 5.

resolutionRequired

480p, 720p, 1080p, 4k; Playground default 720p.

Default720p480p1080p4k
aspect_ratioOptional

auto, 21:9, 16:9, 4:3, 1:1, 3:4, or 9:16; Playground default 16:9.

Default16:99:161:121:94:33:4auto
generate_audioOptional

Boolean. The Playground explicitly sends true by default.

Defaulttrue
return_last_frameOptional

Boolean. The Playground explicitly sends false by default.

Defaultfalse
web_searchOptional

Boolean. The Playground explicitly sends false by default.

Defaultfalse
seedOptional

Integer with no published range; omitted when empty.

How to Use

  1. Define subject and settingState who or what appears, the setting, and details that must remain visible.

  2. Direct the movementUse concrete verbs and separate subject motion from camera motion.

  3. Set the outputChoose 480p, 720p, 1080p, 4k, a 4–15 second duration, and the applicable composition.

  4. Review advanced controlsConfirm audio, last-frame, web-assistance, and seed choices.

  5. Submit and trackRun the request, then poll the task ID or process its terminal callback.

Pricing

One credit equals $0.005. Credits equal output duration times the selected resolution rate. Advanced options, image count, and audio count do not change the current formula.

UsageRateDetails
480p output20 credits / output second$0.100/s; 5 seconds uses 100 credits ($0.500).
720p output40 credits / output second$0.200/s; 5 seconds uses 200 credits ($1.000).
1080p output90 credits / output second$0.450/s; 5 seconds uses 450 credits ($2.250).
4k output200 credits / output second$1.000/s; 5 seconds uses 1000 credits ($5.000).

Best Use Cases

  • Concept clipsTurn a written scene into a reviewable motion concept.

  • Social variantsCompose the same short idea for landscape, square, or vertical delivery.

  • Advertising storyboardsTest one advertising shot's action, light, and timing.

  • Film or game previsValidate staging, camera movement, and atmosphere before production.

Pro Tips

  • Order the prompt as subject and setting, action, camera, light and mood, then sound intent.
  • Give a short text-only clip one primary action to preserve readability.
  • Use concrete motion verbs and separate subject movement from camera movement.
  • Choose composition before describing where the subject sits in frame.
  • When timing matters, describe an opening, development, and final beat.

Notes

  • This endpoint rejects every media field.
  • resolution is required by the public OpenAPI, so examples and the Playground always send it explicitly.
  • Playground upload limits are 30 MB images, 50 MB videos, and 15 MB audio; they are not the complete upstream file contract.
  • Standard accepts 480p, 720p, 1080p, and 4K.
  • Tasks are asynchronous; poll only while status is not_started or running.

Frequently Asked Questions about Seedance 2.0 Text-to-Video API

What is the Seedance 2.0 Text-to-Video API?

Seedance 2.0 is a ByteDance Seedance model. This Text-to-Video endpoint accepts a text prompt without source media and returns an asynchronous video-generation task through the documented REST request contract.

How do I call the Seedance 2.0 Text-to-Video API?

Send POST /api/generate/submit with Authorization: Bearer VIDGO_API_KEY. A successful response includes task_id for the unified status endpoint.

How much does the Seedance 2.0 Text-to-Video API cost?

Multiply output seconds by the selected resolution rate. 480p: 20 credits ($0.100)/s; 720p: 40 credits ($0.200)/s; 1080p: 90 credits ($0.450)/s; 4k: 200 credits ($1.000)/s.

What inputs does the Seedance 2.0 Text-to-Video API accept?

It accepts a 1–20,000 character prompt, 4–15 whole seconds, 480p, 720p, 1080p, 4k, seven ratios, and optional advanced fields; media is rejected.

How do I get the generated video?

Poll GET /api/generate/status/{task_id}, or submit callback_url. Continue only for not_started or running; read every data.files[].file_url after finished, and stop with the error after failed.

Which Seedance 2.0 endpoint should I choose?

Choose Text with no source asset, Image for a start or start/end frame, and Reference for image, video, or audio guidance. Standard offers resolutions through 4K; Fast is limited to 480p/720p at lower current rates.