About Qwen Image
Qwen/Qwen-Image is an image generation model that excels in complex text rendering—including for both alphabetic and logographic languages such as English and Chinese—and precise image editing.
It is particularly strong in producing images with high-fidelity embedded text, making it well-suited for tasks where maintaining the integrity and clarity of text within generated images is critical. The model also provides consistent and realistic editing capabilities, such as style transfer, object addition/removal, and detailed attribute editing, with improved preservation of identity for people and products.
Some other noteworthy features of Qwen/Qwen-Image include robust support for image understanding tasks—such as object detection, semantic segmentation, depth and edge estimation, novel view synthesis, and super-resolution—and native integration with ControlNet for enhanced conditioning and control of outputs.
| Metric | Value |
|---|---|
| Parameter Count | 20 billion |
| Mixture of Experts | No |
| Multilingual | Yes |
| Quantized* | No |
*Quantization is specific to the inference provider and the model may be offered with different quantization levels by other providers.
Ready to build with Qwen Image?
Try Qwen Image in the Workbench to prompt it, compare outputs, and iterate on prompts without writing any code. When you're ready to ship, call the same model from our API and build your own apps on top of it.
curl -sSf -X POST https://hub.oxen.ai/api/ai/images/generate \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $OXEN_API_KEY" \
-d '{
"model": "qwen-image",
"prompt": "A beautiful landscape painting of a serene lake with mountains in the background",
"negative_prompt": " ",
"aspect_ratio": "16:9",
"image_size": "optimize_for_quality",
"num_inference_steps": 30,
"guidance": 3,
"output_format": "jpg",
"output_quality": 80,
"disable_safety_checker": false
}'Request parameters
Every field Qwen Image accepts in the request body. See the API reference for response formats.
https://hub.oxen.ai/api/ai/images/generate| Field | Type | Required | Default | Description |
|---|---|---|---|---|
prompt | string | Yes | — | Prompt for generated image |
negative_prompt | string | No | — | Negative prompt for generated image |
aspect_ratio | string | No | "16:9" | 1:116:99:164:33:43:22:3 |
image_size | string | No | "optimize_for_quality" | Image size for the generated imageoptimize_for_qualityoptimize_for_speed |
num_inference_steps | integer | No | 30 | Number of denoising steps. Recommended range is 28-50, and lower number of steps produce lower quality outputs, faster. Range: 1 – 50 |
guidance | number | No | 3 | Guidance for generated image. Lower values can give more realistic images. Good values to try are 2, 2.5, 3 and 3.5 Range: 0 – 10 |
seed | integernullable | No | — | Random seed. Set for reproducible generation |
output_format | string | No | "jpg" | Format of the output imageswebpjpgpng |
output_quality | integer | No | 80 | Quality when saving the output images, from 0 to 100. 100 is best quality, 0 is lowest quality. Not relevant for .png outputs Range: 0 – 100 |
disable_safety_checker | boolean | No | false | Disable safety checker for generated images. |