Qwen Image 3.0 Text-to-Image API

alibaba/qwen-image-3.0/text-to-image
Standard/Pro · 1K/2K · 8 ratios

Qwen Image 3.0 Standard Text-to-Image API turns a 1–5,000-character prompt into an asynchronous PNG or JPEG image task without source images. Choose one of eight aspect ratios and 1K or 2K resolution.

Get API Key

Input

779/5000

Output

Ready
Qwen Image 3.0 generated result 1
4.8 credits × $0.005 = $0.024

Continue with

Examples

standard-text-01-output.jpg

Create a full-frame vertical ink-and-watercolor natural-history illustration of one scientifically plausible grey heron standing among dew-covered reeds at sunrise. The heron is the clear focal point, with anatomically accurate legs, beak, eye, layered feathers, and a calm alert pose. Surround it with realistic wetland plants, shallow reflective water, distant mist, and warm ivory dawn light. Use an elegant museum-quality observational illustration style with fine linework, subtle paper texture, generous natural breathing room, and no graphic-design elements. The image must contain no writing or typography of any kind: no title, letters, numbers, symbols, panels, article columns, captions, annotations, labels, legends, logos, products, promotional marks, or watermarks.

standard-text-02-output.jpg

Create a landscape museum observation sheet titled exactly "OBJECT NOTES / 器物观察". Center one ancient bronze ritual vessel on a neutral plinth and surround it with exactly four precise study labels: "SILHOUETTE / 轮廓", "PATINA / 铜锈", "CAST MARK / 铸痕", and "HANDLE / 器耳". Include a small scale ruler and two thumbnail contour drawings. Quiet archival photography blended with pencil annotation, charcoal and oxidized green palette, realistic metal surface, balanced editorial grid, readable bilingual typography. Cultural education only; no institution logo, no brand, no sale context, no advertisement, no watermark, no extra text.

standard-text-03-output.jpg

Create a wide three-panel public-education diagram titled exactly "A RAINDROP'S JOURNEY / 一滴雨的旅程". The three equal panels must be labeled exactly "CLOUD / 云", "STREAM / 溪流", and "WETLAND / 湿地". Use one continuous blue path to show a raindrop moving from a mountain cloud through a clear stream into a reed wetland. Add simple arrows, one elevation line, and a tiny legend with only the words "FLOW / 流动" and "REST / 停留". Friendly paper-cut illustration, clean geometry, readable bilingual text, calm blue and moss palette, clear sequence. Educational content only; no campaign logo, no sponsor, no product, no promotional callout, no watermark, no extra text.

Qwen Image 3.0 Standard Text-to-Image

Qwen Image 3.0 Standard is the everyday generation tier in Alibaba's Qwen image family. It suits repeatable creative work such as product scenes, campaign concepts, posters, and typography-led assets when a text prompt is the complete source.

Why Choose Qwen Image 3.0 Standard Text-to-Image API?

  • Start with a written brief.Describe the subject, setting, visual hierarchy, and visible text in one instruction.

  • Design for the destination.Create square, portrait, landscape, or ultrawide assets for the intended placement.

  • Repeat requests with a recorded seed.Leave seed empty while testing prompts, then submit an integer from 0 to 2,147,483,647 when you need to reuse the same request value.

  • Choose PNG or JPEG output.Use PNG or JPEG for the returned image; PNG is the API default when output_format is omitted.

Parameters

ParameterRequirementDescription
promptRequired

Text instruction containing 1–5,000 characters after trimming.

sizeOptional

Output aspect ratio.

Default1:13:22:34:33:416:99:1621:9
resolutionOptional

Output resolution.

Default1K2K
output_formatOptional

Output image encoding.

Defaultpngjpeg
prompt_extendOptional

Boolean that enables automatic prompt expansion.

negative_promptOptional

Content to avoid, up to 5,000 characters.

seedOptional

Integer from 0 through 2,147,483,647, inclusive.

enable_safety_checkerOptional

Boolean that controls content safety checking.

How to Use

  1. Set the visual hierarchy.Name the subject, focal point, supporting elements, and any text that must appear.

  2. Describe the art direction.Add the setting, lighting, materials, color relationships, and camera viewpoint.

  3. Choose the final canvas.Match the aspect ratio and resolution to the placement where the image will be used.

  4. Generate and refine.Review composition and text accuracy, then revise concrete parts of the prompt when needed.

Pricing

Standard Text-to-Image has one base rate for both supported resolutions. Credits convert at $0.005 each.

UsageRateDetails
1K or 2K generation4.8 credits$0.024 per generation at the current credit conversion.

Best Use Cases

  • Product postersCombine a product concept, headline, supporting copy, and layout direction in one prompt.

  • Campaign concept framesExplore key visual directions at the aspect ratio required by the campaign.

  • Marketplace scenesPlace a described product in a controlled environment with purposeful lighting and props.

  • Typography-led artworkBuild an image around a specified title, caption, or concise message.

Pro Tips

  • Put exact visible text in quotation marks and state its position in the hierarchy.
  • Describe foreground, subject, background, and lighting separately when the scene is dense.
  • Choose the delivery ratio before prompting so the composition is built for the final crop.
  • Use the negative prompt for concrete unwanted objects or visual traits.
  • Leave seed empty while exploring; set an integer when you need repeatable input.

Notes

  • Qwen Image 3.0 Standard Text-to-Image API does not accept source-image URLs.
  • The Playground enables prompt extension and safety checking and sends both values explicitly.
  • Generation is asynchronous; continue polling only while the task is not_started or running.
  • Read generated URLs from data.files after finished, or from a configured terminal callback.

Frequently Asked Questions about Qwen Image 3.0 Text-to-Image

What is the Qwen Image 3.0 Text-to-Image API?

Qwen Image 3.0 Text-to-Image is Alibaba's Standard model for creating new images from text. You describe the subject, composition, style, and visible copy, and it returns a PNG or JPEG image suited to everyday posters, product visuals, campaign concepts, and typography-led assets.

How do I call Qwen Image 3.0 Text-to-Image?

Submit the model ID and input object with a Bearer API key, then use the returned task ID to check status or configure a callback.

How much does Standard Text-to-Image cost?

A generation costs 4.8 credits at either 1K or 2K. The pricing section shows the current dollar conversion.

What inputs can I control?

Provide a prompt and optionally set aspect ratio, resolution, format, prompt expansion, negative prompt, seed, and safety checking. Exact values are listed above.

Where is the generated image returned?

A finished status response places each image URL in data.files[].file_url. A failed task returns error_message instead.

Which Qwen Image 3.0 API mode should I choose?

Use Text-to-Image when the request has no source images and Edit when it includes one to three ordered source-image URLs. Standard generation costs 4.8 credits at 1K or 2K; Pro generation costs 6.4 credits at 1K and 12 credits at 2K.