Qwen/qwen-image-2512/Released Dec 2025

Qwen Image - 2512

Photorealistic portraits and natural scenes

Fine-tunableCommercial use
Image

About Qwen Image - 2512

Qwen Image 2512 is a Large Vision Model. It excels in text-to-image generation with improved realism in human portraits, finer natural textures, and stronger text rendering, particularly for Chinese characters.

Some other noteworthy use cases of Qwen Image 2512 include instruction-based image editing and generating structured visuals like posters or UI mockups.

MetricValue
Parameter Count20 billion
MultilingualYes

*Quantization is specific to the inference provider and the model may be offered with different quantization levels by other providers.

Ready to build with Qwen Image - 2512?

Try Qwen Image - 2512 in the Workbench to prompt it, compare outputs, and iterate on prompts without writing any code. When you're ready to ship, call the same model from our API and build your own apps on top of it.

curl -sSf -X POST https://hub.oxen.ai/api/ai/images/generate \
    -H "Content-Type: application/json" \
    -H "Authorization: Bearer $OXEN_API_KEY" \
    -d '{
  "model": "qwen-image-2512",
  "prompt": "A bald eagle sitting on a vast frozen lake, centered in the shot facing the camera. The eagle is in it'\''s natural habitat and being photographed from a medium distance. The background is a vast lake surrounded by a forest going up a hill. Photorealistic - it is high enough quality that it could be used for a National Geographic cover, but is just a stand alone photo without any graphics.",
  "negative_prompt": " ",
  "aspect_ratio": "16:9",
  "image_size": "optimize_for_quality",
  "num_inference_steps": 30,
  "guidance": 3,
  "output_format": "webp",
  "output_quality": 80,
  "disable_safety_checker": false
}'

Request parameters

Every field Qwen Image - 2512 accepts in the request body. See the API reference for response formats.

POSThttps://hub.oxen.ai/api/ai/images/generate
FieldTypeRequiredDefaultDescription
promptstringYes
Prompt for generated image
negative_promptstringNo
Negative prompt for generated image
aspect_ratiostringNo"16:9"
1:116:99:164:33:43:22:3
image_sizestringNo"optimize_for_quality"Image size for the generated image
optimize_for_qualityoptimize_for_speed
num_inference_stepsintegerNo30Number of denoising steps. Recommended range is 28-50, and lower number of steps produce lower quality outputs, faster.
Range: 150
guidancenumberNo3Guidance for generated image. Lower values can give more realistic images. Good values to try are 2, 2.5, 3 and 3.5
Range: 010
seedinteger
nullable
No
Random seed. Set for reproducible generation
output_formatstringNo"webp"Format of the output images
webpjpgpng
output_qualityintegerNo80Quality when saving the output images, from 0 to 100. 100 is best quality, 0 is lowest quality. Not relevant for .png outputs
Range: 0100
disable_safety_checkerbooleanNofalseDisable safety checker for generated images.