Seedance 1.0 Pro Text to Video API

bytedance/seedance/v1/pro/text-to-video

Seedance 1.0 Pro Text to Video transforms natural language prompts into high-fidelity cinematic video, with 720p and 1080p resolutions, 5-second and 10-second durations, and realistic physical motion simulation. It faithfully adheres to detailed scene prompts up to 10,000 Unicode characters while preserving consistent subject appearance, lighting, and environmental continuity across smooth camera movements.

Input
670/10000
OutputReady
720p · 5 seconds · 21 credits = $0.105

Examples

One continuous natural-history five-second shot in a dense Amazon rainforest canopy in humid late-morning light. A single adult scarlet macaw with accurate red, yellow, and blue plumage crouches on a dead branch, then pushes off, opens both wings, and flies through a wide gap between two trunks. The camera pans gently to follow as the bird becomes smaller among layered leaves. Maintain believable wingbeat rhythm, body weight, and claw release. No other birds, no people, no cuts, no lettering, no logos, no brands, no advertising, no watermark.

One continuous locked wide five-second shot at dusk on the Atacama plateau. An unmarked white observatory dome sits on dark volcanic gravel under a sky shifting from amber to indigo. The dome slit rotates open slowly on visible hinges, revealing a dark interior. No people, vehicles, antennas, or lettering. Mechanically plausible rotation, consistent panel count, natural twilight, no cuts, no logos, no brands, no products, no advertising, no watermark.

Seedance 1.0 Pro Text to Video

Seedance 1.0 Pro Text to Video is ByteDance's high-fidelity text-to-video generation model designed for commercial storyboarding, cinematic conceptualization, and creative video storytelling. Creators and developers can produce vivid, physically grounded scenes directly from natural language prompts, with flexible 720p and 1080p resolutions and 5- or 10-second clips.

Why Choose This?

  • Pure Text-Driven CreativityConstruct dynamic scenes, lifelike character actions, and cinematic environments directly from natural language prompts without reference assets.

  • Physically Grounded Motion SimulationFaithfully simulate real-world physics, fluid inertia, and natural lighting shifts to ensure coherent movements without morphing or temporal stutter.

  • Flexible Resolution and Duration TiersProduce 5-second quick cuts or 10-second narrative scenes in standard 720p or high-definition 1080p master quality.

  • 10,000-Character Prompt CapacityTake advantage of an extensive 10,000 Unicode character limit to detail multi-beat direction, environmental depth, and camera motion.

  • Predictable Per-Generation PricingSimple per-generation credit rates starting at 21 credits ($0.105) for 720p at 5 seconds, with automatic credit refunds if generation fails.

Parameters

ParameterRequirementDescription
promptRequired

Required string containing 1–10,000 Unicode characters after trimming surrounding whitespace. Non-string and blank prompts are rejected.

resolutionOptional

Only the strings 720p and 1080p are supported. Defaults to 720p only when omitted; aliases, different casing, surrounding whitespace, null and empty strings are rejected.

Default720p1080p
durationOptional

Only numeric integers 5 and 10 seconds are supported. Defaults to 5 only when omitted; strings, booleans, null and fractional values are rejected. Numeric 5.0 is equivalent to 5.

Default510

How to Use

  1. Define Subject and Scene AtmosphereEstablish focal character appearance, setting details, and mood in the opening sentences to anchor the visual composition.

  2. Outline Camera Motion and Action FlowDescribe subject actions chronologically and specify explicit camera directions such as smooth dolly pushes, tracking pans, or crane shots.

  3. Select Resolution and Output DurationChoose between 5 seconds for rapid clips or 10 seconds for narrative sequences, alongside 720p standard or 1080p HD quality.

  4. Review Credits and Submit TaskTrigger generation from the playground or dispatch an asynchronous POST request via the REST API to receive a unique task_id.

  5. Poll Progress and Download VideoQuery task status with your task_id until finished, then retrieve the secure MP4 download link from the response files array.

Pricing

Per-generation pricing by resolution and duration, identical for text-to-video and image-to-video. 1 credit = $0.005.

UsageRateDetails
720p · 5 seconds21 credits / $0.105Per video generation
720p · 10 seconds42 credits / $0.210Per video generation
1080p · 5 seconds43 credits / $0.215Per video generation
1080p · 10 seconds86 credits / $0.430Per video generation

Best Use Cases

  • Commercial Storyboard PrototypingTranslate script lines into dynamic motion previews to validate pacing, camera angles, and visual impact before physical production.

  • Social Media and Digital ContentProduce eye-catching, high-resolution short-form video content rapidly for digital campaigns and social channels.

  • Film and Narrative Concept VisualizationTurn screenplay paragraphs into 5-to-10-second cinematic scenes to assist directors and animators with blocking and tone exploration.

  • Worldbuilding and Sci-Fi Concept ArtBreathe life into imaginative landscapes, futuristic vehicles, and fantastical phenomena in high-definition video.

Pro Tips

  • Separate Subject Motion from Camera Movement: Describing character actions and camera trajectories in distinct clauses helps the model execute cinematic movements accurately.
  • Leverage the 10,000-Character Capacity: Utilize the generous input ceiling to detail ambient lighting, weather conditions, textural nuances, and camera focal depth.
  • Prototype in 720p Before Rendering 1080p: Validate movement dynamics and framing economically at 720p 5s (21 credits) before producing your final 1080p version.
  • Use Specific Physical Action Verbs: Favor concrete dynamic descriptions such as 'strides steadily forward' or 'water ripples outward' over generic superlatives.
  • Establish Consistent Lighting and Setting: Clearly specify atmospheric details like golden hour, overcast daylight, or neon nightscapes to enhance cinematic coherence.

Notes

  • Text-Only Input Specification: This endpoint is strictly for text-to-video generation, accepting a prompt string up to 10,000 Unicode characters after trimming surrounding whitespace.
  • Valid Resolution and Duration Parameters: Supported resolutions are 720p (default) and 1080p; supported durations are 5 seconds (default) and 10 seconds as integers.
  • Asynchronous Task Lifecycle and Refund Guarantee: Each submission returns a task_id for polling or webhook callback delivery; failed tasks are automatically refunded in full.

Related Models

Seedance 1.0 Pro Text to Video API frequently asked questions

What is the Seedance 1.0 Pro Text to Video API?

Seedance 1.0 Pro Text to Video is a ByteDance model for text-to-video generation. It produces high-fidelity videos in 720p and 1080p resolutions directly from natural language prompts, supporting 5-second and 10-second durations with realistic physical dynamics and camera trajectory control. Built on ByteDance's advanced video generation architecture, it preserves lighting consistency and temporal coherence while strictly adhering to sequential action and spatial descriptions. You can call it programmatically or try it from the playground above.

Does Seedance 1.0 Pro Text to Video support 10-second videos?

Yes. The model provides both 5-second and 10-second duration settings (defaulting to 5 seconds). Choosing 10 seconds allows for richer narrative progression, multi-stage character movement, and smoother camera transitions within a single generation.

What resolutions does Seedance 1.0 Pro Text to Video support?

This endpoint supports 720p and 1080p output resolutions. Omitted resolution parameters default to 720p; select 1080p when generating high-definition assets for commercial displays or cinematic productions.

How long can text prompts be for Seedance 1.0 Pro Text to Video?

After trimming surrounding whitespace, prompts can range from 1 to 10,000 Unicode characters. This generous ceiling lets you provide comprehensive scene descriptions, lighting instructions, character actions, and multi-beat camera framing.

How is Seedance 1.0 Pro Text to Video priced?

Pricing is billed per generation based on resolution and duration (1 credit = $0.005). At 720p, 5 seconds costs 21 credits ($0.105) and 10 seconds costs 42 credits ($0.210). At 1080p, 5 seconds costs 43 credits ($0.215) and 10 seconds costs 86 credits ($0.430). Unsuccessful tasks are automatically refunded.

Can you control camera motion with prompts in Seedance 1.0 Pro Text to Video?

Yes. You can include standard cinematic camera terminology in your prompt (such as 'slow camera push in', 'tracking shot following the character', or 'aerial panoramic pan'), and the model will synchronize camera movement with the subject's action.

How do you maintain scene continuity during character actions in Seedance 1.0 Pro Text to Video?

Anchor the character appearance and surrounding setting clearly at the beginning of your prompt, then describe actions chronologically. Avoid conflicting camera directions or abrupt cuts within the same clip to maintain seamless temporal continuity.