Seedream 5.0 Lite Text to Image API

bytedance/seedream/v5/lite/text-to-image
2K/3K · 8 aspect ratios · n 1–15

Seedream 5.0 Lite Text to Image 将自然语言提示词转化为 2K 与 3K 高保真图像,支持多步视觉逻辑推理、实时网络检索增强与 1–15 张多图批量发散。它在严格遵循复杂空间几何与物理光影规律的同时,精准呈现时效性热点细节与结构化知识图表,为商业营销与创意设计提供出图即用的高质量视觉资产。

获取 API 密钥
Input
必填0/3000
可选
可选
可选
Output待生成

生成的图片将在此处显示

预估费用: —
继续使用

示例

text-to-image-01-output.png

A scientifically accurate educational infographic explaining the honeybee waggle dance, drawn in a warm naturalist illustration style on cream paper. A title across the top reads exactly "THE WAGGLE DANCE". Below it, four numbered panels in a 2x2 grid, each with one short caption in clean capital letters. Panel 1 caption "1. A FORAGER FINDS FLOWERS": a worker bee flies from a wooden hive to a field of purple lavender; a dotted flight line leads to the flowers, a small sun sits in the sky, and a thin arc marks a 40 degree angle between the sun direction and the flight line. Panel 2 caption "2. SHE DANCES ON THE COMB": a vertical honeycomb of hexagonal wax cells; the returning bee traces a figure-eight path whose straight waggle run is tilted 40 degrees to the right of vertical, with a dashed vertical reference line, a thin arc marking the 40 degree angle between that line and the waggle run, and a small upward arrow whose tip carries the single word "SUN" to show that straight up on the comb stands for the direction of the sun. Panel 3 caption "3. LONGER WAGGLE, FARTHER FOOD": three waggle runs of increasing length drawn side by side, each connected to a flower patch placed at an increasing distance from the hive. Panel 4 caption "4. SISTERS FOLLOW THE MAP": several bees that watched the dance fly out from the hive along the same 40 degree angle to the sun and reach the lavender field. The 40 degree angle must stay consistent in panels 1, 2 and 4. Anatomically correct bees with six legs and four wings, readable fine linework, every word spelled correctly, no other text, no watermark.

text-to-image-02-output.png

A cinematic documentary photograph inside a centuries-old glassblowing furnace room on the island of Murano, Venice, late at night. In the foreground, a master glassblower in a sweat-darkened grey linen shirt and a scorched leather apron turns a long blowpipe with a glowing orange gather of molten glass at its tip; his face is lit from below by the glow. In the middle ground, a young female assistant shields her face with a wooden paddle as she opens the round furnace door, revealing a white-hot interior that is the brightest light source in the room. In the background, partly hidden behind the assistant's shoulder, an older artisan sits at a steel bench shaping a translucent cobalt-blue vase with folded wet newspaper, a wisp of steam rising. Shelves along the soot-darkened brick wall hold finished clear, amber and green glass pieces that refract the orange light and cast colored caustics onto the bench. Correct occlusion and depth between the three people, every shadow falling away from the furnace, warm firelight contrasted with cool blue moonlight from one small window, visible heat shimmer above the furnace mouth, 35mm lens, natural skin texture, believable hands with five fingers, no text, no logos, no watermark.

text-to-image-03-output.png

An ultra-wide cinematic landscape photograph of the Altai Mountains in western Mongolia in deep winter at golden hour. In the left third of the foreground, a Kazakh eagle hunter sits on a sturdy brown Mongolian horse, wearing a fox-fur hat and a long embroidered wool coat; he raises his thick leather-gloved forearm as a golden eagle with fully spread wings lands on it, powder snow kicked up around the horse's hooves. In the middle distance on the right, a second rider leads a small caravan of four Bactrian camels across a frozen river, their long shadows stretching over the snow. Far behind, jagged snow-covered peaks glow pink and gold under a clear pale sky, each successive ridge lighter and hazier with distance. Correct relative scale between the eagle, the riders, the camels and the mountains, one consistent low sun from the right, visible breath vapor from the horse, fine detail in feathers and fur, no text, no watermark.

Seedream 5.0 Lite Text to Image

Seedream 5.0 Lite Text to Image 是由字节跳动 Seed 团队研发的新一代轻量高效多模态文本生成图像大模型。模型在 Seedream 系列演进基础上实现了理解、推理与生成的全方位升级,不仅原生支持 2K 与 3K 分辨率直出及 8 种主流画幅预设,更深度融合了实时网络检索能力与思维链风格的多步视觉推理机制。无论是需要融入即时动态资讯的时效性海报、包含复杂前后景深与透视规律的多主体交互画面,还是抽象概念结构化的知识信息图表,模型均能准确领会深层意图并输出细腻逼真的视觉细节。单次请求支持并发批量生成 1 至 15 张多样化创意成品,单张仅需 5 积分,全面赋能内容创作、电商视觉、广告营销与教育科普等工业级高并发场景。

为什么选择 Seedream 5.0 Lite Text to Image API?

  • 多步视觉推理与空间几何物理逻辑引入思维链式深度推理机制,深度解析复杂分层指令,精准把控多主体相对比例、3D 景深层次与遮挡关系,确保真实自然的光影投射与严格符合物理规律的空间透视。

  • 动态实时网络检索突破时限壁垒生成过程中无缝接入实时网络检索模块,动态获取最新互联网事件、消费品趋势与前沿概念资讯,避免传统预训练模型的知识截止限制,让创意视觉始终与现实动态保持同步。

  • 结构化知识呈现与信息可视化图表内置广博的世界知识库与排版美学底座,能够将枯燥复杂的数据、流程架构与科普概念转化为布局规整、图文并茂的专业信息图解、流程图与多栏演示幻灯片。

  • 2K/3K 原生直出与 1–15 张批量发散支持最高 3K 原生高保真画质直出,单次请求即可并发生成最多 15 张不同构图与视角的独立画面变体,配合单张 5 积分平价计费与异步任务流,极速提升创意选型与资产交付效率。

参数说明

参数要求说明
prompt必填

用 3–3,000 个字符描述要生成的图片,首尾空白不计入长度。

size可选

输出尺寸,可选 2K、3K 或画幅比例;也可填写 WIDTHxHEIGHT(如 2304x1728)或 {"width":2304,"height":1728} 自定义宽高,宽高为正整数。默认 1:1。

默认值1:12K3K4:33:416:99:163:22:321:9
n可选

生成图片数量,取值为整数 1–15,默认 1。

默认值123456789101112131415
enable_safety_checker可选

使用 true 开启安全检查,使用 false 关闭,默认 true。

默认值true

使用教程

  1. 撰写提示词与视觉构思使用详尽的中英文描述画面主体、艺术风格、光影氛围与空间布局;若涉及实时热点或专业知识,直接写入具体实体名词即可触发深度理解与知识检索。

  2. 配置画幅比例与生成张数根据终端展示需求选择 2K、3K 预设画质,或指定 16:9、9:16、1:1 等 8 种常用比例与自定义像素宽高,并通过 n 参数指定 1 至 15 张期望生成的图片数量。

  3. 异步提交任务并获取高清成品向统一端点发送请求并获取唯一 task_id,通过轮询查询接口或在顶层配置 callback_url 接收完成通知,任务终态后直接读取并下载高清图片 URL。

价格

每张图片 5 credits($0.025),总价按 n 计算:5 × n credits。

用量价格说明
标准生成5 credits/image · $0.025/image5 × n credits · $0.025 × n
n = 1 / 5 / 10 / 155 / 25 / 50 / 75 credits$0.025 / $0.125 / $0.250 / $0.375

适用场景

  • 时效性营销与热点广告海报结合实时网络检索能力,快速生成契合最新节日事件、流行趋势与商业新品发布的高品质宣传海报与社交媒体物料。

  • 教育科普与信息图表设计将抽象知识、技术原理与业务流程转化为条理清晰、层次分明的知识图解与教学演示插图,大幅降低制图排版成本。

  • 电商商品陈列与品牌视觉包装依托高精度材质渲染与多步光影推理,为商品快速生成质感逼真、透视协调的陈列场景、包装视觉与概念氛围大片。

  • 影视前期概念与分镜批量发散利用 1–15 张多图批量并发生成优势,在统一设定下快速探索多种镜头机位、场景景别与构图方案,加速前期创意定稿。

使用技巧

  • 精准文字与标语排版:若需在画面中渲染特定文字,请使用半角双引号明确标注文案(如:海报正中写着 "SUMMER SALE"),并指明文字的排版层级、字体风格与色彩。
  • 分层描述引导多步推理:按照“核心主体 + 空间相对位置 + 材质表面细节 + 动态光源环境 + 全局色彩基调”的顺序组织提示词,能够更好地激活模型的空间几何与物理逻辑。
  • 善用实时实体关键词:对于近期发布的科技产品、社会热点或流行风尚,直接在提示词中包含准确的专有名词与背景描述,引导模型检索最新上下文生成精准写实的画面。
  • 阶梯式批量发散与定稿:探索新方案时可先以默认尺寸设置 n 为 4 进行快速发散选型,确定理想构图与光影风格后,再针对性优化提示词生成高精度成品。

注意事项

  • 纯文本生成专有端点:本端点专用于纯文本生成图像,不接收图片素材;如需结合参考图进行编辑、风格迁移或多图融合,请调用 Seedream 5.0 Lite Edit 端点。
  • 提示词字符长度规范:提示词去除首尾空白字符后的长度需在 3 至 3,000 字符之间,包含至少一个非空白字符。
  • 批量计费与退款保障:任务提交时按 5 × n 积分预扣除;若任务因校验未通过、网络异常或超时导致生成失败,预扣积分将全额自动退还;若实际成功出图少于 n 张,系统按未出图差额自动退费。

相关模型

Seedream 5.0 Lite Text to Image API 常见问题

Seedream 5.0 Lite Text to Image API 是什么?

Seedream 5.0 Lite Text to Image 是字节跳动 Seed 团队用于高质量文本生成图像的多模态大模型。它根据自然语言提示词生成 2K 与 3K 高保真图像,具备行业领先的多步视觉逻辑推理、实时网络检索增强与结构化信息可视化能力,并支持单次 1–15 张批量生成。基于统一多模态生成底座与深度思考机制,它在严格遵循物理光影与空间几何规律的同时,准确呈现时效性实体细节与复杂构图。你可以通过 API 进行程序化调用,也可以在上方体验区直接在线试用。

Seedream 5.0 Lite Text to Image 如何利用实时网络检索生成图像?

模型在生成前可动态检索互联网公开最新信息与知识,自动补充现实中的产品细节、突发事件背景或潮流元素。当提示词涉及最新时事、新发布设备或实时文化话题时,模型能打破传统预训练数据的时间局限,生成贴合当下事实的准确图像。

Seedream 5.0 Lite Text to Image 的多步视觉推理能改善哪些画面细节?

多步视觉推理支持类似思维链的空间逻辑演化,显著改善多物体空间相对位置、3D 景深遮挡、透视比例以及复杂光源在不同材质表面的物理反射效果,避免传统生成模型常出现的主体漂浮、比例失调或阴影错位问题。

Seedream 5.0 Lite Text to Image 支持生成信息图表和知识图解吗?

支持。模型具备强大的信息可视化与世界知识整合能力,能够理解复杂概念与数据流转关系,自动规划清晰的视觉信息层级,生成排版工整、图文结合的流程图、对比图解、科普插图与演示文稿背景。

Seedream 5.0 Lite Text to Image 一次能批量生成多张图片吗?

支持单次请求生成 1 至 15 张独立图片。通过在请求体 input 中配置 n 参数(取值整数 1–15),即可针对同一提示词并发获得多种构图与视角的发散变体,生成费用按实际成功输出张数结算。

Seedream 5.0 Lite Text to Image 支持哪些原生分辨率与画幅比例?

支持 2K 与 3K 两种清晰度预设档位,并提供 1:1、4:3、3:4、16:9、9:16、3:2、2:3、21:9 等 8 种主流比例;同时支持通过 WIDTHxHEIGHT 字符串或包含正整数 width 与 height 的对象传入自定义像素宽高。

Seedream 5.0 Lite Text to Image 任务生成失败时如何结算积分?

系统具备全自动退款保障机制。任务提交时会按 5 × n 积分进行预扣除;若因输入校验未通过、网络异常或系统超时导致生成失败,预扣积分将即时全额退还;若实际成功生成图片少于 n 张,系统会自动按未出图差额退还多扣除的积分。