Qwen/qwen-image/Released Aug 2025

Qwen Image

Text-to-image with strong text rendering

Fine-tunableCommercial use
Image

About Qwen Image

Qwen/Qwen-Image is an image generation model that excels in complex text rendering—including for both alphabetic and logographic languages such as English and Chinese—and precise image editing.

It is particularly strong in producing images with high-fidelity embedded text, making it well-suited for tasks where maintaining the integrity and clarity of text within generated images is critical. The model also provides consistent and realistic editing capabilities, such as style transfer, object addition/removal, and detailed attribute editing, with improved preservation of identity for people and products.

Some other noteworthy features of Qwen/Qwen-Image include robust support for image understanding tasks—such as object detection, semantic segmentation, depth and edge estimation, novel view synthesis, and super-resolution—and native integration with ControlNet for enhanced conditioning and control of outputs.

MetricValue
Parameter Count20 billion
Mixture of ExpertsNo
MultilingualYes
Quantized*No

*Quantization is specific to the inference provider and the model may be offered with different quantization levels by other providers.

Ready to build with Qwen Image?

Try Qwen Image in the Workbench to prompt it, compare outputs, and iterate on prompts without writing any code. When you're ready to ship, call the same model from our API and build your own apps on top of it.

curl -sSf -X POST https://hub.oxen.ai/api/ai/images/generate \
    -H "Content-Type: application/json" \
    -H "Authorization: Bearer $OXEN_API_KEY" \
    -d '{
  "model": "qwen-image",
  "prompt": "A beautiful landscape painting of a serene lake with mountains in the background",
  "negative_prompt": " ",
  "aspect_ratio": "16:9",
  "image_size": "optimize_for_quality",
  "num_inference_steps": 30,
  "guidance": 3,
  "output_format": "jpg",
  "output_quality": 80,
  "disable_safety_checker": false
}'

Request parameters

Every field Qwen Image accepts in the request body. See the API reference for response formats.

POSThttps://hub.oxen.ai/api/ai/images/generate
FieldTypeRequiredDefaultDescription
promptstringYes
Prompt for generated image
negative_promptstringNo
Negative prompt for generated image
aspect_ratiostringNo"16:9"
1:116:99:164:33:43:22:3
image_sizestringNo"optimize_for_quality"Image size for the generated image
optimize_for_qualityoptimize_for_speed
num_inference_stepsintegerNo30Number of denoising steps. Recommended range is 28-50, and lower number of steps produce lower quality outputs, faster.
Range: 150
guidancenumberNo3Guidance for generated image. Lower values can give more realistic images. Good values to try are 2, 2.5, 3 and 3.5
Range: 010
seedinteger
nullable
No
Random seed. Set for reproducible generation
output_formatstringNo"jpg"Format of the output images
webpjpgpng
output_qualityintegerNo80Quality when saving the output images, from 0 to 100. 100 is best quality, 0 is lowest quality. Not relevant for .png outputs
Range: 0100
disable_safety_checkerbooleanNofalseDisable safety checker for generated images.