AlibabaAlibaba·🖌️ Inpainting·New

Qwen Image 2.1 Turbo

anonymized
Try on Venice.ai ↗
Quick reference
Qwen Image 2.1 Turbo — TLDR
  • 🏢 Alibaba Qwen's editing variant of the Qwen-Image-2.1 family
  • ⚡ Accelerated low-step checkpoint for faster image editing
  • 🧠 Compact visual generation component in a unified architecture
  • 🔧 Instruction-based editing and inpainting workflows
  • 👁️ Composes from up to 10 reference images in one pass
  • 🎯 Qwen reports improved portrait identity and product consistency
  • 📚 Carries the family's signature in-image text rendering
  • 🆕 Mixed-granularity attention plus prefix KV cache reuse cuts memory
💰 Pricing
$0.020
per edit
📅 On Venice since
Oct 9, 2026
1 day ago
Provider

Alibaba Group is a Chinese multinational technology company founded in 1999 and headquartered in Hangzhou, Zhejiang. Originally built around e-commerce and cloud computing, Alibaba has become one of the most prolific contributors to open-weight AI research,…

Read full profile →
79 models on Venice
35 video · 23 text · 9 image · 8 inpaint · 2 embedding · 2 tts
Since Jan 11, 2025

About this model

Qwen Image 2.1 Turbo (edit) is the inpainting and instruction-editing face of Alibaba's Qwen-Image-2.1 release, paired with the text-to-image sibling Qwen Image 2.1 Turbo. Rather than shipping separate checkpoints for generation, local editing, and transparency, Qwen unified these workflows into one compact model. Edits are driven by natural-language instructions applied to an input image, the same interaction pattern documented for Qwen's image-editing API.

The generational story is largely about size and speed. The original Qwen-Image was a 20B multimodal diffusion transformer, and its editing version built directly on that backbone. The 2.1 line slims this down and adds mixed-granularity attention — token-level causal masking for text, chunk-level for image generation — plus prefix KV cache reuse, so reference images and instructions are encoded once and reused across denoising steps, improving inference efficiency and lowering memory use. The Turbo checkpoint is the accelerated, low-step variant of that line.

On capability, Qwen describes four editing upgrades in 2.1: support for multiple reference images (up to 10), more flexible local editing, better fidelity preservation, and broader task coverage — specifically better retention of facial identity across edits and of a product's text, texture, and shape. Compared with earlier catalog entries such as Qwen Image 2 and Qwen Image 3, it targets the same unified editing tasks with fewer sampling steps. Legible in-image typography, long a hallmark of the family, is retained.

This About section is AI-generated from public sources (Claude Opus 5), with no human editing. It may contain inaccuracies — verify critical details against the sources listed above.

Data sources: Venice API · HuggingFace · Wikipedia — enrichment updated 16h ago