Seedream 4.5 Text to Image API

bytedance/seedream/v4.5/text-to-image
2K/4K · 8 ratios · Custom size

Seedream 4.5 Text to Image transforms text prompts into 2K and 4K high-fidelity visuals, supporting crisp bilingual typography, refined portrait lighting, and batch exploration of 1–15 images. It follows spatial semantics and prompt instructions closely while maintaining balanced compositions across diverse aspect ratios for ready-to-use production assets.

Get API Key
Input
Required0/3000
Optional
Optional
Optional
OutputIdle

Generated images will appear here

Estimated cost: —
Continue using

Examples

text-to-image-01-output.jpg

A detailed isometric cutaway illustration of a small polar research station standing on steel stilts on an Antarctic ice shelf, shown like a page from a science museum guidebook. The roof and front wall are removed so six rooms are visible, each with a small clean sans-serif label plate printed exactly: "LAB" (a scientist examining an ice core under a lamp), "GALLEY" (two people cooking soup at a steaming stove), "BUNKS" (stacked beds with colorful quilts), "RADIO" (an operator with headphones in front of dials), "GREENHOUSE" (lettuce and tomatoes under pink grow lights) and "GARAGE" (an orange snowcat being repaired). A banner across the top reads exactly "AURORA RIDGE STATION", and a small thermometer sign beside the entrance reads "-38°C". Tiny figures, pipes, ladders and cables connect the rooms logically. Outside: blowing snow, a weather mast, a pale green aurora in the dark sky. Crisp linework, soft flat colors, consistent isometric perspective, every label spelled correctly and legible, no extra text, no watermark.

text-to-image-02-output.jpg

A vertical documentary portrait photograph of an elderly South Indian fisherwoman mending a fishing net under the palm-thatched roof of a wooden pier in Kerala during a monsoon evening. She sits cross-legged, wearing a faded turquoise cotton sari with a thin gold border and a small silver nose stud; her deeply lined face, silver hair tied back and weathered hands pulling green nylon mesh are in sharp focus. Blue-hour light from the backwaters mixes with the warm glow of a single hanging kerosene lantern beside her. Heavy rain falls in streaks beyond the roof edge, drops glint on the net, wet planks reflect the lantern, and a few moored wooden boats fade into the misty background. Natural skin texture with pores and fine wrinkles, calm dignified expression looking at her work, 85mm lens, shallow depth of field, believable hands with five fingers, no text, no watermark.

text-to-image-03-output.jpg

An ultra-wide cinematic panorama from inside a gigantic rotating cylindrical space habitat. The camera stands on a gravel path in the foreground among ripening wheat and small orchards; the farmland, rivers, winding roads and small white villages continuously curve upward on both sides and arc overhead, so the opposite side of the cylinder is visible high in the sky, upside down, with tiny fields and lakes. A long glowing sun-line runs along the central axis, casting soft morning light and thin wispy clouds drifting in the middle of the cylinder. At the far end, a huge circular end-cap window shows black space, stars and a slice of a blue planet. A cyclist rides along the path for scale. Coherent curved perspective, consistent lighting, atmospheric haze with distance, grounded realistic sci-fi concept art with photographic detail, no text, no logos, no watermark.

Seedream 4.5 Text to Image

Seedream 4.5 Text to Image is a high-fidelity text-to-image generation model developed by the ByteDance Seed team. Sharing a unified architecture with its companion image editing endpoint and scaled up from Seedream 4.0, it delivers substantial upgrades across spatial reasoning, 3D depth, and complex prompt execution. The model natively generates 2K and 4K images across eight aspect ratios and custom dimensions with zero high-resolution surcharge, excelling at dense bilingual typography and realistic facial details even on small-scale subjects. Generating 30%–40% faster and supporting 1 to 15 batch variations, it powers production pipelines in commercial advertising, e-commerce, and creative design.

Why Choose Seedream 4.5 Text to Image API?

  • Breakthrough Bilingual and Dense Small-Text TypographyRenders legible, sharp Chinese and English text, dense product packaging copy, and brand logos directly within visuals, overcoming typical AI character distortion for production-ready designs.

  • Native 4K Output with Zero Surcharge and Small-Face ClarityProduces native 4K ultra-high-definition visuals without resolution markups, capturing realistic skin micro-textures and cinematic lighting while preserving natural clarity even on small-scale faces.

  • 3D Spatial Reasoning and Compositional PerspectiveLeverages upgraded spatial comprehension to maintain accurate physical proportions, depth of field, and object occlusion, ensuring balanced, coherent perspective across multi-subject scenes.

  • 1–15 Batch Output with 30%–40% Generation SpeedupGenerates up to 15 distinct framing and angle variations in a single run with 30%–40% faster generation speeds than version 4.0, drastically reducing creative turnaround time at 5 credits per image.

Parameters

ParameterRequirementDescription
promptRequired

Describe the image, art style, and composition in 1–3,000 characters.

sizeOptional

Output size: 2K, 4K, or an aspect ratio. For a custom size, enter WIDTHxHEIGHT (for example 1920x4096) or {"width":2304,"height":3072} with positive integers. Defaults to 1:1.

Default1:12K4K4:33:416:99:163:22:321:9
nOptional

Number of images to generate: integer 1–15. Defaults to 1.

Default123456789101112131415
enable_safety_checkerOptional

Set true to enable safety checking or false to turn it off. Defaults to true.

Defaulttrue

How to Use

  1. Craft Your Prompt and Text ContentDescribe the scene, subject, aesthetic style, lighting, and composition in detail. Enclose any required text in quotation marks to guide placement and spelling.

  2. Select Dimensions and Output CountChoose a preset aspect ratio such as 1:1, 16:9, or 9:16, specify 2K or 4K resolution, or provide custom dimensions, and set the image count n between 1 and 15.

  3. Submit Task and Retrieve ResultsSend your request to the unified endpoint to receive a task_id, then poll status or configure a callback_url to access the generated image URLs.

Pricing

Each generation costs 5 credits ($0.025). n is the number of generated images; the total is 5 × n credits.

UsageRateDetails
Standard generation5 credits / generation · $0.025 / generation5 × n credits · $0.025 × n

Best Use Cases

  • Commercial Ads and Marketing PostersGenerate complete promotional creatives featuring legible bilingual headlines, promotional tags, and polished product arrangements.

  • E-Commerce Product Scenes and Lifestyle VisualsProduce photorealistic contextual backgrounds and lifestyle mockups tailored to catalog items with physically accurate ambient lighting.

  • Concept Art and Editorial MediaExplore diverse creative variations across 1–15 images for film pre-production, book covers, social media graphics, and digital storytelling.

Pro Tips

  • Typography prompting: Enclose exact slogans or titles in quotation marks (such as "NEW ARRIVAL" at the top) and explicitly state font weight, placement, color, and language.
  • Activate 3D spatial depth: For complex multi-subject scenes, specify foreground and background relationships (e.g. "foreground right", "deep center") to direct 3D depth and proportions.
  • Staged creation workflow: Explore creative directions at 2K or with n set to 4 for rapid iteration, then produce final deliverables at native 4K with no resolution surcharge.

Notes

  • Text-only input: This endpoint accepts text prompts only. For reference-guided editing or multi-image composition, use the Seedream 4.5 Image Edit endpoint.
  • Prompt length limit: Prompts must contain between 1 and 3,000 characters after trimming whitespace.
  • Billing and refunds: Tasks reserve 5 credits ($0.025) per requested image. Unused credits are automatically refunded for failed tasks or when fewer than n images are returned.

Related Models

Seedream 4.5 Text to Image API frequently asked questions

What is the Seedream 4.5 Text to Image API?

Seedream 4.5 Text to Image is a ByteDance Seed model for high-quality text-to-image generation. It generates 2K and 4K ultra-high-definition images from text prompts with crisp bilingual typography, refined portrait lighting, and batch creation of 1–15 images. Built on an advanced global scaling multimodal foundation, it preserves spatial semantics and prompt adherence while maintaining balanced compositions across diverse aspect ratios. You can call it programmatically or try it from the playground above.

Can Seedream 4.5 Text to Image generate multiple images at once?

Yes, you can generate 1 to 15 independent images in a single call. By configuring the n parameter from 1 to 15, you receive multiple composition and framing variations for the same prompt, billed only for the images successfully produced.

Does Seedream 4.5 Text to Image support rendering Chinese and English text?

Yes. The model features a built-in typography engine that accurately renders legible Chinese and English characters, headlines, and slogans directly within generated graphics when enclosed in quotation marks.

What resolutions and aspect ratios does Seedream 4.5 Text to Image support?

It supports 2K and 4K resolution tiers along with eight standard aspect ratio presets including 1:1, 4:3, 3:4, 16:9, 9:16, 3:2, 2:3, and 21:9, as well as custom positive integer dimensions such as 1920x4096 or an object.

How does Seedream 4.5 Text to Image handle content safety?

An automated safety checker is enabled by default through enable_safety_checker set to true. It screens requests to ensure safe visual content, and credits are automatically refunded if a task produces no output due to moderation.

What is the typical generation latency for Seedream 4.5 Text to Image?

Benefiting from architectural scaling, Seedream 4.5 generates 30%–40% faster than version 4.0. Under standard cluster loads, end-to-end generation latency for a 2K or 4K image has a median of 5 to 15 seconds, and production applications should poll every 2 to 5 seconds or configure a callback_url.

How are credits handled if a Seedream 4.5 Text to Image task fails?

Vidgo uses an automated refund guarantee. Tasks pre-charge 5 credits per image upon submission; if a task fails due to validation errors or timeouts, all credits are returned, and partial output is refunded proportionally.