Grok Imagine Image Text to Image API

xai/grok-imagine-image/text-to-image
5 种画面比例

Grok Imagine Image Text to Image 将文字转化为写实或风格化图像,结合主体细节、光线与构图。文字描述引导场景外观,五种画幅适配人物肖像、产品概念与社交配图。

获取 API Key
输入
629/5000

示例

The Record Conservator

An intimate documentary photograph inside a small old vinyl record restoration room. A young adult male technician with a silver buzz cut and round face gently holds a carbon-fiber record brush against a black vinyl record lying flat on a workbench. His other hand steadies the outer edge. Frame his face, shoulders and both hands together, believable fingers and a single brush. Slanting side-window daylight reveals the fine grooves and a few dust particles, with worn wooden shelves softly out of focus. Quiet concentration, natural skin texture, warm understated color, candid rather than posed. No readable labels, branding, promotional graphics, or watermark.

Lynx Behind the Fallen Trunk

An eye-level wildlife photograph of one Eurasian lynx peeking around the right side of a snow-covered fallen tree trunk in a boreal conifer forest. Show its head and forequarters with accurate feline anatomy, prominent black ear tufts, amber eyes, white cheek ruff and naturally damp whiskers. Its paws stand on undisturbed snow. Soft cool morning light, fine individual fur strands, a small amount of distant mist between spruce trunks, softly blurred background, realistic wildlife observation. No other animals, people, text, logos, or watermark.

Grok Imagine Image Text to Image

Grok Imagine Image Text to Image 是 xAI 用于将文字描述转化为图像的模型,创作方向涵盖摄影感场景与插画风格。它结合提示词中的主体、环境、光线和构图,将创意简报中的人物形象、产品构想或故事场景呈现为可供讨论的画面。五种输出画幅适配方形配图、竖幅肖像与横向场景设计,让同一创作方向对应不同版面。

为什么选择 Grok Imagine Image Text to Image?

  • 探索写实与插画方向通过文字选择摄影感场景或插画表现,在项目构思阶段比较同一题材的不同视觉方向。

  • 从创意简报建立画面提示词将主体、环境、光线和取景组织在一起,让文字简报中的创意获得可见的形态。

  • 画幅对应成品用途方形、竖幅与横幅输出分别适配社交配图、人物肖像和场景概念,使画面构图与使用位置相衔接。

参数

参数要求说明
prompt必填

描述画面与视觉风格,长度为 1–5,000 个字符。

size可选

可选值为 1:1、2:3、3:2、16:9、9:16。

1:12:33:216:99:16

使用步骤

  1. 描述画面填写主体、环境、光线和视觉风格。

  2. 选择成品画幅按人物肖像、社交配图或横向场景选择 Size。

  3. 运行并查看结果核对费用后点击 Run,生成完成后预览并下载。

价格

1 credit 等于 $0.005。

用量费率详情
图片生成6 credits / image$0.030 / image

应用场景

  • 产品概念提案描述设想中的物件、材质与摆放环境,生成用于讨论产品外观的概念画面。

  • 人物视觉设定结合人物特征、姿态、光线和摄影或插画风格,探索肖像的视觉方向。

  • 社交内容插画围绕帖文主题构建场景,选择方形或竖幅构图,制作与内容配套的图像。

提示词建议

  • 先说明主要主体,再补充环境、主体位置和在画面中的大小。
  • 摄影感画面写明材质、光线方向与视角;插画写明媒介和配色。
  • 按使用版面选择 size,并在提示词中说明主体周围需要的留白。
  • 比较创作方向时保留场景主干,每次调整一项视觉选择,例如光线或风格。

使用说明

  • 保存 task_id 以查询结果,或提供 callback_url 接收任务完成通知。

相关模型

Grok Imagine Image Text to Image API 常见问题

Grok Imagine Image Text to Image API 是什么?

Grok Imagine Image Text to Image 是 xAI 用于将文字描述转化为图像的模型。它生成写实与风格化场景,通过主体、光线和构图描述引导画面。文字构成视觉简报,所选比例确定输出画幅,用于人物肖像、产品概念或插画创作。你可以通过 API 调用,也可以在“体验”标签中在线试用。

Grok Imagine Image Text to Image 能生成摄影感场景吗?

可以。在提示词中描述主体、材质、环境与光线,例如柔和窗光下放在木桌上的陶瓷杯,模型根据这段场景描述生成画面。

Grok Imagine Image Text to Image 如何呈现插画风格?

在 prompt 中指定插画媒介与表现方式,例如低饱和配色的水彩或平面编辑插画。保留主体与场景描述,可以探索同一题材的不同风格。

Grok Imagine Image Text to Image 提供哪些画幅?

size 可选 1:1、2:3、3:2、16:9 和 9:16。先按使用位置选择画幅,再用提示词说明主体位置与周围留白。

何时应选择 Grok Imagine Image Text to Image 而非 Grok Imagine Image Edit?

以文字场景或创意简报作为起点时,选择 Grok Imagine Image Text to Image;已有图片并希望描述修改内容时,选择 Grok Imagine Image Edit。

Grok Imagine Image Text to Image 的结果可以制作动画吗?

可以将生成图片作为 Grok Imagine Video Image to Video 的参考图,再描述希望增加的动作。这是一次独立的视频生成,可制作 6 秒或 10 秒片段。

Grok Imagine Image Text to Image 生成一张图片如何计费?

每张图片为 6 credits,即 $0.030。五种 size 取值采用相同的单张价格。