Flux Kontext Max Text to Image API

blackforestlabs/flux-kontext-max/text-to-image
size · output_format

Flux Kontext Max Text to Image 将复杂提示词转化为影棚级商业视觉,具备极致的语义遵循与多行文字排版。它在精准解析多主体空间透视的同时,保持电影级光影平衡。

获取 API 密钥
Input
必填0/2000
可选
可选
Output待生成

生成的图片将在此处显示

预估费用: 16 credits · $0.080

每次生成 16 credits($0.080)。

继续使用

示例

max-text-to-image-03-output.png

A photograph taken from inside a small deep-sea research submersible, looking through a thick round acrylic porthole at a depth of 1,200 meters. Outside in the black water, a female deep-sea anglerfish hangs close to the glass, seen from the side: a round lumpy dark-brown body, a huge wide mouth full of long translucent needle teeth, and tiny eyes. Growing out of the top of the fish's own head is a thin, flexible fleshy fishing rod, part of its body, that arches forward over its mouth and ends in a small glowing blue-green bioluminescent bulb, the only strong light source outside. The glow lights the fish's teeth and skin from above. Drifting marine snow particles catch the light. On the inside surface of the porthole, faint reflections of the cabin appear: amber-lit analog gauges, a red emergency switch and the soft silhouette of a scientist's face lit by a tablet screen. Tiny condensation droplets bead on the lower edge of the acrylic. A thick steel ring with heavy hex bolts frames the porthole. Extremely low light, high dynamic range, realistic photographic noise. No text, no logos, no watermark.

max-text-to-image-01-output.png

A single hand-colored botanical plate from a nineteenth-century natural history book, photographed flat and straight on in soft daylight. On warm cream paper with faint foxing, a detailed watercolor and engraving of the carnivorous pitcher plant Nepenthes rajah: one large upright urn-shaped pitcher in deep wine red with yellow-green speckles, a ribbed glossy crimson peristome rim around its mouth, a round lid raised above the opening, and a curling tendril joining the base of the pitcher to the tip of a long leathery green leaf. Next to it, a small cutaway of the pitcher shows digestive fluid and a trapped beetle. Delicate engraved line hatching, soft watercolor washes and blooms, fine paper fibers and a slightly uneven plate-mark border. At the top center, one single line of elegant engraved serif capitals reads exactly "NEPENTHES RAJAH". That title is the only text on the page: there are no numbers, labels, captions, signatures, stamps or handwriting anywhere else.

max-text-to-image-02-output.jpg

An ultra-wide cinematic panorama of the Danakil Depression in Ethiopia at dawn. A long camel caravan of about twenty camels, each loaded with stacked rectangular slabs of white salt tied with rope, walks in single file from left to right across a vast, perfectly flat salt pan. Afar herders in white and ochre wraps walk beside the camels holding long sticks. The low sun sits just above the horizon on the right, casting very long blue shadows of the camels across the cracked hexagonal salt crust and turning the ground into a gradient of pink, apricot and pale lilac. A thin layer of standing water in the foreground mirrors the caravan and the sky. Distant dark volcanic hills line the horizon under a clear sky fading from peach to deep blue with a few high streaks of cloud. Shot on an anamorphic lens, low camera height, gentle heat haze, fine detail in the salt texture and camel fur, realistic scale. No text, no logos, no watermark.

Flux Kontext Max Text to Image

Flux Kontext Max Text to Image 是 Black Forest Labs 研发的旗舰级文生图大模型,代表 Kontext 系列的顶峰画质。基于扩展的流匹配架构,它能严密推理多主体空间透视与复杂微观材质,对长篇叙事提示词及密集多行排版具备极高的还原精度,为高端平面广告与影视视觉提供卓越交付标准。

为什么选择 Flux Kontext Max Text to Image API?

  • 极致的复杂提示词语义遵循攻克多物体排布与空间遮挡关系理解瓶颈,严格还原长提示词中对服装纹理、肢体姿态、镜头距离与背景细节的精细设定。

  • 行业标杆级多行文字排版渲染针对多行英文单词、书刊排版、复古标牌与商标字体展现顶尖的绘制能力,字符结构锐利端正,极大降低后期修字成本。

  • 影棚级真实光影与物理材质表现精准模拟复杂环境下的多光源交互,包括玻璃折射、金属高光拉丝、丝绸次表面散射与人像皮肤毛孔质感,呈现出色的商业级质感。

  • 全比例高精度一致输出原生支持 1:1、16:9、9:16、21:9 等 7 种构图画幅,无论横屏宽幅展布还是竖屏移动端视觉,输出像素均稳定在约 100 万高保真标准。

参数说明

参数要求说明
prompt必填

描述要生成的图片内容,长度为 1–2,000 个字符,包含至少一个非空白字符。

size可选

输出图片的宽高比,默认 1:1。

默认值1:14:33:416:99:1621:99:21
output_format可选

输出图片的文件格式,可选 png 或 jpg。

pngjpg

使用方法

  1. 撰写高精度提示词详尽定义画面构图、核心主体、微观材质与色彩层级;如需排布标题与文案,请用半角双引号精确括出文本并指明其在画面中的排版位置。

  2. 匹配专业画幅与格式依据商业交付媒介选择 1:1 方图、16:9 横版宽幅或 9:16 竖版海报,并指定 png 或 jpg 格式。

  3. 提交请求并获取生成结果调用统一 API 端点提交任务并获取 task_id,通过轮询或 webhook 回调获取生成完成的高保真图像下载链接。

计费说明

每次生成 16 credits($0.080),输出 1 张旗舰级高保真图片。

用量价格说明
标准生成16 credits/次 · $0.080/次适用于所有 size 比例与 output_format 格式选项

适用场景

  • 高端奢侈品与工业产品渲染为腕表、香氛、珠宝及汽车概念生成具备复杂反射、精准折射与无瑕环境布光的商业交付级视觉大片。

  • 大型商业活动主视觉与户外看板生成带有精细多行文字排版、宏大透视空间与丰富细节的广告大图,满足高规格印刷与展布需求。

  • 影视概念艺术与电影级分镜借助模型深邃的氛围渲染与角色神态刻画能力,将导演级构想快速具象化为充满戏剧张力的概念画作。

实用技巧

  • 排版文字多层标注法:若画面包含主标题与副标,可在提示词中清晰分列,例如 A poster with the main title "SOLAR" at the top, and subtitle "FUTURE ENERGY" in smaller sans-serif font below.
  • 细化材质与表面工艺:利用微观质感词汇(如 brushed titanium, velvet drape, subsurface scattering, ambient occlusion)充分释放 Max 版本的物理光影潜力。
  • 掌控电影级景深焦段:指定具体摄影器材参数(如 shot on 70mm IMAX camera, anamorphic lens flare, shallow f/1.2 focus),引导模型渲染独特的胶片色调与光学虚化。

注意事项

  • 纯文本端点边界:本端点严格面向纯文本生成图像,不接收源图片 URL;如需基于参考图进行修改或精修,请调用 Flux Kontext Max Edit 端点。
  • 提示词字数范围:去除首尾空白字符后,文本字符数需在 1 至 2,000 字符之间。
  • 透明计费结算:每次成功调用扣除 16 credits($0.080),若因服务器网络故障未成功返回图片,积分将全额自动返还至账户。

相关模型

Flux Kontext Max Text to Image API 常见问题

Flux Kontext Max Text to Image API 是什么?

Flux Kontext Max Text to Image 是 Black Forest Labs 用于最高品质文本生成图像的旗舰级流匹配大模型。它根据复杂的自然语言提示词生成约 100 万像素高保真图像,具备行业顶尖的复杂文字排版、极其严密的提示词语义遵循与影棚级物理光影渲染能力。基于深层流匹配与高阶多模态 Transformer 架构,它在解析多主体空间布局与精细纹理的同时,呈现出色的色调平衡与空间景深。你可以通过 API 进行程序化调用,也可以在上方体验区直接在线试用。

Flux Kontext Max Text to Image 处理复杂提示词有哪些优势?

Max 版本拥有 Kontext 系列中最顶级的语义解析能力。对于长达数段、涵盖多个人物互动、前后景遮挡关系以及精细微观材质的复杂指令,模型能够逐一解析并协调呈现,极大地避免了传统模型常见的元素遗漏或属性混淆。

Flux Kontext Max Text to Image 支持密集多行文字排版吗?

支持。Max 版本是文字排版渲染的标杆,不仅能渲染单个单词,还能在画面中的报刊、霓虹灯、包装或书本上清晰呈现整段多行英文文本,字形边缘锐利且完美贴合表面透视。

Flux Kontext Max Text to Image 的端到端推理耗时是多少?

在标准并发负载与网络环境下,单张高保真图像的端到端生成耗时中位数约为 7 秒。虽然比 Pro 版本的 5–6 秒略长,但换来的是更为极致的构图逻辑与微观细节保真度。

Flux Kontext Max Text to Image 支持自定义画幅比例吗?

端点提供 1:1、4:3、3:4、16:9、9:16、21:9 与 9:21 等 7 种工业标准画幅比例。无论选择横屏还是竖屏,输出总像素均精准保持在约 100 万像素,确保画面细节均匀一致。

Flux Kontext Max Text to Image 单次生成的计费标准是多少?

单次成功生成固定扣除 16 credits(折合 $0.080)。该价格包含所有画幅预设与输出格式,费用透明且在任务提交时校验,如因系统异常未成功生成结果将自动原路退还。

商业广告级海报设计该如何选择 Pro 与 Max?

如果您的海报概念偏向标准构图、单主体展示且追求快速迭代,选用 Pro 即可获得优秀的品质与高效体验;如果您需要制作包含密集文字标语、多层复杂透视、苛刻奢侈材质反射的最终交付级商业主视觉,推荐使用 Max 以获取最高画质保障。