Seedream 4.5 Text to Image API

bytedance/seedream/v4.5/text-to-image
2K/4K · 8 ratios · Custom size

Seedream 4.5 Text to Image 将自然语言提示词转化为 2K 与 4K 超高清图像,支持专业级中英双语文字排版、细腻人像光影质感与 1–15 张多图批量发散。它在严格遵循提示词空间逻辑与物理规律的同时,保持多画幅构图严谨自然,为海报设计与商业物料提供出图即用的高质量视觉成品。

获取 API 密钥
Input
必填0/3000
可选
可选
可选
Output待生成

生成的图片将在此处显示

预估费用: —
继续使用

示例

text-to-image-01-output.jpg

A detailed isometric cutaway illustration of a small polar research station standing on steel stilts on an Antarctic ice shelf, shown like a page from a science museum guidebook. The roof and front wall are removed so six rooms are visible, each with a small clean sans-serif label plate printed exactly: "LAB" (a scientist examining an ice core under a lamp), "GALLEY" (two people cooking soup at a steaming stove), "BUNKS" (stacked beds with colorful quilts), "RADIO" (an operator with headphones in front of dials), "GREENHOUSE" (lettuce and tomatoes under pink grow lights) and "GARAGE" (an orange snowcat being repaired). A banner across the top reads exactly "AURORA RIDGE STATION", and a small thermometer sign beside the entrance reads "-38°C". Tiny figures, pipes, ladders and cables connect the rooms logically. Outside: blowing snow, a weather mast, a pale green aurora in the dark sky. Crisp linework, soft flat colors, consistent isometric perspective, every label spelled correctly and legible, no extra text, no watermark.

text-to-image-02-output.jpg

A vertical documentary portrait photograph of an elderly South Indian fisherwoman mending a fishing net under the palm-thatched roof of a wooden pier in Kerala during a monsoon evening. She sits cross-legged, wearing a faded turquoise cotton sari with a thin gold border and a small silver nose stud; her deeply lined face, silver hair tied back and weathered hands pulling green nylon mesh are in sharp focus. Blue-hour light from the backwaters mixes with the warm glow of a single hanging kerosene lantern beside her. Heavy rain falls in streaks beyond the roof edge, drops glint on the net, wet planks reflect the lantern, and a few moored wooden boats fade into the misty background. Natural skin texture with pores and fine wrinkles, calm dignified expression looking at her work, 85mm lens, shallow depth of field, believable hands with five fingers, no text, no watermark.

text-to-image-03-output.jpg

An ultra-wide cinematic panorama from inside a gigantic rotating cylindrical space habitat. The camera stands on a gravel path in the foreground among ripening wheat and small orchards; the farmland, rivers, winding roads and small white villages continuously curve upward on both sides and arc overhead, so the opposite side of the cylinder is visible high in the sky, upside down, with tiny fields and lakes. A long glowing sun-line runs along the central axis, casting soft morning light and thin wispy clouds drifting in the middle of the cylinder. At the far end, a huge circular end-cap window shows black space, stars and a slice of a blue planet. A cyclist rides along the path for scale. Coherent curved perspective, consistent lighting, atmospheric haze with distance, grounded realistic sci-fi concept art with photographic detail, no text, no logos, no watermark.

Seedream 4.5 Text to Image

Seedream 4.5 Text to Image 是由字节跳动 Seed 团队研发的高保真文本生成图像大模型。模型与图像编辑端点统一底层架构,并在 4.0 基础上实现全面精准 Scaling 升级,不仅原生支持 2K 与 4K 分辨率直出及 8 种主流画幅,更具备突破性的中英双语密集小字排版能力、三维透视空间推理与远景小比例面容保真表现。生成速度较 4.0 提升约 30%–40%,且单次请求支持批量生成 1 至 15 张多样化创意成品,广泛适用于商业海报制作、电商视觉营销、数字艺术创作及概念设计等高保真工业级生产力场景。

为什么选择 Seedream 4.5 Text to Image API?

  • 突破性中英双语与密集小字排版攻克传统 AI 绘图文字扭曲与字形畸变难题,原生支持在画面中清晰呈现排版规整的多行中英文字符、商品小字标语与品牌包装标识,满足商业视觉出图即用的高标准交付需求。

  • 原生 4K 零加价直出与远景小脸保真支持直接生成最高 4K 原生超高清分辨率且无需额外付费,精准呈现细腻的皮肤毛孔质感、织物纹理与电影级光影层次,即使画面中主体人物占据较小比例,远景面容依然清晰自然。

  • 三维景深与复杂空间几何推理全面升级空间理解底座,准确把握多物体比例、3D 景深层次与前后遮挡关系,在多主体互动或宏大场景叙事中严格遵循透视法则,呈现构图严密协调的真实纵深感。

  • 单次 1–15 张批量发散与 30%–40% 提速单次请求即可并发生成最多 15 张多样化景别与构图的独立画面变体,推理耗时较上一代 4.0 显著缩短 30%–40%,配合单张 5 积分平价计费,全面加速创意选型与方案定稿。

参数说明

参数要求说明
prompt必填

用 1–3,000 个字符描述画面主体、艺术风格与细节要求。

size可选

输出尺寸,可选 2K、4K 或画幅比例;也可填写 WIDTHxHEIGHT(如 1920x4096)或 {"width":2304,"height":3072} 自定义宽高,宽高为正整数。默认 1:1。

默认值1:12K4K4:33:416:99:163:22:321:9
n可选

生成图片数量,取值为整数 1–15,默认 1。

默认值123456789101112131415
enable_safety_checker可选

使用 true 开启安全检查,使用 false 关闭,默认 true。

默认值true

使用方法

  1. 构造提示词与文字内容使用详尽的中英文描述画面主体、艺术风格、光影质感与空间布局;如需画面文字,请使用双引号明确标出文案内容及排版位置。

  2. 配置画幅与输出张数根据展示媒介选择 2K、4K、预设比例(如 16:9、9:16、1:1 等)或自定义宽高,并指定 1 至 15 张输出数量 n。

  3. 提交任务并异步获取成品调用统一接口提交生成任务并获取 task_id,通过轮询或设置 callback_url 接收终态,完成后直接获取高清图片下载链接。

计费

每次生成 5 积分($0.025),n 为生成图片数量,总价 = 5 × n 积分。

计费项目费率详情
标准生成5 积分 / 次 · $0.025 / 次5 × n 积分 · $0.025 × n

适用场景

  • 商业广告与营销海报一键生成带有清晰中英文标题、副标题与促销标语的活动视觉海报,降低设计排版门槛。

  • 电商产品场景与模特图根据商品特性规划逼真光影与空间透视,为电商店铺快速生成质感出众的陈列背景与营销氛围图。

  • 概念设计与自媒体视觉借助 1–15 张批量生成能力,为影视前期概念、小说插画、社交媒体封面提供丰富多元的视觉灵感方案。

实用技巧

  • 精准渲染文字技巧:在提示词中使用半角双引号精确框定文字内容(例如:海报上方写着 "NEW ARRIVAL"),并明确指出字体的排版位置、粗细、色调与语言风格。
  • 激活空间透视构图:面对多主体交互场景时,详细阐述主体在画面中的前后景层级与相对方位(如“前景右侧”、“正中偏远”),引导模型精准排布 3D 深度与空间透视。
  • 阶梯式创作与批量发散:探索新概念时可先以 2K 分辨率或设置 n 为 4 进行快速发散选型,锁定满意构图后再以 4K 原生分辨率生成高精度商业交付级成品。

注意事项

  • 纯文本生成限定:本端点专用于文本生成图像,不接收图片素材;如需基于参考图进行修改或多图融合,请调用 Seedream 4.5 Image Edit 端点。
  • 提示词字符范围:去除首尾空白后的提示词长度需在 1 至 3,000 字符之间。
  • 批量计费规则:生成费用按实际成功张数结算,单张固定为 5 积分($0.025),提交时预扣 5 × n 积分,任务失败或少出图差额均会自动退还。

相关模型

Seedream 4.5 Text to Image API 常见问题

Seedream 4.5 Text to Image API 是什么?

Seedream 4.5 Text to Image 是字节跳动 Seed 团队用于高质量文本生成图像的多模态大模型。它根据文本提示词生成 2K 与 4K 超高清图像,具备行业领先的中英双语密集小字排版渲染能力、细腻真实的人像光影质感,并支持单次 1–15 张批量生成。基于先进的全局 scaling 多模态统一生成底座,它在严格遵循提示词空间几何与物理规律的同时,保持多画幅构图严谨自然。你可以通过 API 进行程序化调用,也可以在上方体验区直接在线试用。

Seedream 4.5 Text to Image 一次能生成多张图片吗?

支持单次生成 1 至 15 张独立图片。通过在请求中配置 n 参数(取值 1–15),即可批量获得针对同一提示词的多样化构图与视觉变体,计费按实际成功输出张数结算。

Seedream 4.5 Text to Image 支持在画面中渲染中文文字吗?

支持。模型内置强大的文字渲染引擎,原生支持在生成图像中精准呈现清晰的中英文字符、标语与品牌文案,在提示词中用半角双引号标明文本即可生成端正锐利的字迹。

Seedream 4.5 Text to Image 支持哪些原生分辨率与画幅?

支持 2K 与 4K 清晰度档位,并提供 1:1、4:3、3:4、16:9、9:16、3:2、2:3、21:9 等 8 种主流比例预设,同时支持传入自定义正整数像素宽高(如 1920x4096 或对象格式)。

Seedream 4.5 Text to Image 如何处理敏感内容与安全审查?

内置可选的安全检查机制。默认参数 enable_safety_checker 为 true,会自动拦截不符合合规要求的生成内容;若因合规拦截导致任务未产出图像,系统会自动退还预扣积分。

Seedream 4.5 Text to Image 生成耗时通常在什么范围?

得益于架构升级,Seedream 4.5 的生成速度较 4.0 提升了约 30%–40%。在常规网络与标准并发负载下,生成单张 2K 或 4K 图像的端到端耗时中位数约为 5 至 15 秒;批量生成多张图片时耗时会略有增加,生产环境建议配置 2 至 5 秒轮询间隔或使用 callback_url 异步接收终态。

Seedream 4.5 Text to Image 任务生成失败时如何结算积分?

系统具备自动退款保障机制。任务提交时会按 5 × n 积分进行预扣除;若任务因校验未通过、网络异常或超时导致生成失败,预扣积分将全额退还;若实际成功生成图片少于 n 张,系统会按差额自动退还未生成张数的积分。