Seedance 1.0 Pro Image to Video API

bytedance/seedance/v1/pro/image-to-video

Seedance 1.0 Pro Image to Video animates a single reference image into high-fidelity cinematic video, with 720p and 1080p resolutions, 5-second and 10-second durations, and physically accurate motion simulation. It faithfully preserves the source image's facial likeness, wardrobe textures, and spatial composition while introducing fluid subject actions and camera movements guided by your text prompt.

Input
498/10000
1/1
input.png

Image-to-video requires one HTTP(S) image URL. Playground uploads accept JPG, PNG or WebP up to 10 MiB; this is an upload policy. JSON/API does not accept Base64.

OutputReady
720p · 5 seconds · 21 credits = $0.105

Examples

Animate this exact locked-camera landscape in one continuous five-second shot. Dry steppe grass leans in a steady crosswind and a few high clouds drift. Keep the mosaic shelter, tiles, concrete plinth, distant horizon, and lighting direction unchanged. No people, animals, vehicles, or new objects. Natural wind physics only. No cuts, no lettering, no logos, no brands, no advertising, no watermark.

Animate this exact locked macro composition in one continuous five-second shot. The hermit crab walks a few steps out of the shell across barnacles, and one small concentric ripple travels over the pool. Keep the anemone, rock, shell, waterline, and camera fixed. Natural wet surfaces and tiny fluid motion. No people, no cuts, no lettering, no logos, no brands, no advertising, no watermark.

Seedance 1.0 Pro Image to Video

Seedance 1.0 Pro Image to Video is ByteDance's high-fidelity image-to-video generation model. It accurately captures and preserves subject identity, color grading, and framing from a single input image, generating physically coherent movement and camera trajectories guided by natural language prompts across 720p and 1080p specifications.

Why Choose This?

  • Exceptional Identity and Framing RetentionAnchors character facial contours, hair textures, clothing, and original compositional balance from your reference image, preventing distortion as motion begins.

  • Physically Accurate Dynamic SimulationSimulates subtle fabric folds, hair physics, and environmental light reflections naturally to transform static pictures into fluid, cinematic clips.

  • Flexible Resolution and Duration TiersSelect 5-second quick-motion cuts or 10-second extended shots across 720p standard and 1080p high-definition master resolutions.

  • Prompt-Directed Motion and Camera ControlSteer subject actions, facial expressions, and camera moves including dolly pushes, pans, and tracking orbits cleanly using natural language.

  • Predictable Per-Generation BillingEnjoy transparent per-generation rates starting at 21 credits ($0.105) for 720p at 5 seconds, backed by automatic refunds whenever generation encounters an issue.

Parameters

ParameterRequirementDescription
promptRequired

Required string containing 1–10,000 Unicode characters after trimming surrounding whitespace. Non-string and blank prompts are rejected.

image_urlsRequired

An array of exactly one HTTP(S) image URL with a hostname and no embedded credentials or whitespace. Base64 and bare strings are rejected.

resolutionOptional

Only the strings 720p and 1080p are supported. Defaults to 720p only when omitted; aliases, different casing, surrounding whitespace, null and empty strings are rejected.

Default720p1080p
durationOptional

Only numeric integers 5 and 10 seconds are supported. Defaults to 5 only when omitted; strings, booleans, null and fractional values are rejected. Numeric 5.0 is equivalent to 5.

Default510

How to Use

  1. Prepare a Crisp Reference ImageUpload a clean JPG, PNG, or WebP image under 10 MiB with well-defined subjects and balanced lighting to establish the visual foundation.

  2. Write Motion and Camera DirectionsFocus your text prompt on physical transformations and camera trajectories (e.g., 'subject smiles and glances left as the camera dollies forward slowly').

  3. Select Output Resolution and DurationChoose between 5 seconds or 10 seconds, and pick 720p for rapid iteration or 1080p for crisp high-definition production.

  4. Review Credits and Dispatch GenerationHit Run in the playground or send an asynchronous POST request with image_urls and parameters via the REST API to receive your task_id.

  5. Poll Status and Retrieve Finished VideoQuery task status with task_id until finished, then access and download your high-fidelity MP4 video directly from the response URL.

Pricing

Per-generation pricing by resolution and duration, identical for text-to-video and image-to-video. 1 credit = $0.005.

UsageRateDetails
720p · 5 seconds21 credits / $0.105Per video generation
720p · 10 seconds42 credits / $0.210Per video generation
1080p · 5 seconds43 credits / $0.215Per video generation
1080p · 10 seconds86 credits / $0.430Per video generation

Best Use Cases

  • Portrait and Character AnimationAnimate still character portraits, digital avatars, and photographic model shots with natural eye movement, smiles, and head turns.

  • E-Commerce and Commercial DisplayConvert static product photography and marketing posters into eye-catching promotional clips featuring dynamic lighting and smooth camera sweeps.

  • Photography and Landscape MotionBring still architectural captures, travel photography, and nature shots to life with rippling water, moving clouds, and cinematic fly-throughs.

  • Concept Art and Illustration Motion ExplorationTurn 2D digital paintings, storyboards, or concept sketches into dynamic video previews to test cinematic pacing and mood.

Pro Tips

  • Focus Prompts on Motion Rather Than Appearance: The model automatically retains visual details from your starting image; concentrate your prompt on movement direction, tempo, and camera framing.
  • Use Crisp, High-Contrast Starting Images: High-resolution images with clean background separation and clear facial lighting yield significantly better temporal consistency.
  • Validate Dynamics in 720p Before Rendering 1080p: Test motion trajectories economically at 720p 5s (21 credits) before producing your final 1080p high-definition master.
  • Specify Clear Directional Camera Moves: Employ descriptive cinematography cues like 'slow dolly zoom', 'subtle pan left', or 'low-angle tracking shot' for precise viewpoint motion.
  • Choose 10 Seconds for Multi-Phase Actions: For compound actions such as turning around followed by a gesture, select 10 seconds to allow the action to unfold naturally.

Notes

  • Single-Image Array Specification: Exactly one accessible HTTP(S) image URL must be provided inside the image_urls array; playground uploads accept up to 10 MiB.
  • Parameter Values and Defaults: Output resolution is restricted to 720p (default) or 1080p; duration is restricted to 5 seconds (default) or 10 seconds as integers.
  • Asynchronous Processing and Refund Policy: Submissions return an immediate task_id for polling or webhook callback; failed generation tasks are automatically refunded in full.

Related Models

Seedance 1.0 Pro Image to Video API frequently asked questions

What is the Seedance 1.0 Pro Image to Video API?

Seedance 1.0 Pro Image to Video is a ByteDance model for image-to-video generation. It produces high-fidelity videos in 720p and 1080p resolutions from a single reference image and natural language prompt, supporting 5-second and 10-second durations with realistic physical simulation and camera motion control. Built on ByteDance's advanced video generation architecture, it preserves the source subject's identity, wardrobe details, and environmental framing while adding natural, physically coherent motion. You can call it programmatically or try it from the playground above.

How many reference images can I upload to Seedance 1.0 Pro Image to Video?

The image_urls array must contain exactly one starting reference image. The model anchors this image as the visual origin and subject reference, expanding it temporally according to your prompt.

How does Seedance 1.0 Pro Image to Video maintain subject consistency?

The model extracts facial contours, clothing textures, and lighting characteristics from the reference image and tracks them continuously throughout the generation. Focus your prompt on actions and camera motion rather than introducing conflicting physical descriptions.

Can Seedance 1.0 Pro Image to Video generate 10-second videos?

Yes. The model provides both 5-second and 10-second duration options (defaulting to 5 seconds). A 10-second generation allows more continuous motion progression and smoother camera transitions while maintaining subject fidelity.

What output resolutions are supported by Seedance 1.0 Pro Image to Video?

This endpoint supports 720p and 1080p output resolutions. If resolution is omitted, the model defaults to 720p; select 1080p when you need high-definition footage for commercial publishing or post-production.

How is Seedance 1.0 Pro Image to Video priced?

Pricing is billed per generation according to resolution and duration (1 credit = $0.005). At 720p, 5 seconds is 21 credits ($0.105) and 10 seconds is 42 credits ($0.210). At 1080p, 5 seconds is 43 credits ($0.215) and 10 seconds is 86 credits ($0.430). Failed tasks are automatically refunded.

Does Seedance 1.0 Pro Image to Video support camera motion control via text prompts?

Yes. You can include explicit camera movement instructions in your text prompt (such as 'slow dolly forward', 'subtle tracking pan', or 'low-angle orbit around subject'), and the model will execute the camera motion while preserving subject stability.