Alibaba: qwen-image-plus-2026-01-09 API

Qwen-Image-Edit-Max is the most powerful flagship image editing model in the Tongyi Wanxiang (Qwen-Image) series. Built on a 20B-parameter multimodal diffusion architecture (MMDiT), it utilizes an innovative framework that feeds input images into both Qwen2.5-VL (for visual semantic control) and a VAE Encoder (for visual appearance control). Compared to the Plus version, the Max model significantly enhances the understanding of physical laws (e.g., auto-generating matching shadows and reflections), geometric reasoning for industrial design, and character consistency. It supports multi-image fusion (up to 3 images), enabling precise bilingual text modification, object addition/removal, pose changes, and style transfer, while greatly mitigating image drift issues.

Frequently Asked Questions

What is the context window of qwen-image-plus-2026-01-09?

qwen-image-plus-2026-01-09 supports a context window of up to 2,048 tokens.

What is the knowledge cutoff of qwen-image-plus-2026-01-09?

The knowledge cutoff of qwen-image-plus-2026-01-09 is 2025-11-30.