AlibabaAlibaba·💬 Text Generation·New

Qwen 3.8 Max

ReasoningVisionCodeFunction CallingWeb Searchanonymized
🧠 Try in Intelligence →Try on Venice.ai ↗
Quick reference
Qwen 3.8 Max — TLDR
  • 🧠 2.4-trillion-parameter mixture-of-experts model, Alibaba's Max tier
  • 📏 Massive 1M-token context window
  • 👁️ Accepts text, images and video as input
  • 🔧 Function calling, web search, code-optimized
  • 🎯 Thinking mode only — built for long-horizon reasoning
  • 🧩 Strong on software engineering and multi-agent workflows
💰 Pricing
$2.50 / $7.50
per 1M · input / output
📏 Context
1M tokens
📅 On Venice since
Jul 22, 2026
12 days ago
Provider

Alibaba Group is a Chinese multinational technology company founded in 1999 and headquartered in Hangzhou, Zhejiang. Originally built around e-commerce and cloud computing, Alibaba has become one of the most prolific contributors to open-weight AI research,…

Read full profile →
53 models on Venice
20 text · 20 video · 5 image · 4 inpaint · 2 embedding · 2 tts
Since Jan 11, 2025

About this model

Qwen 3.8 Max sits at the top of Alibaba's proprietary Max tier, a 2.4-trillion-parameter mixture-of-experts system released in July 2026 that runs exclusively in thinking mode. Rather than offering a toggle between fast and deliberate responses, every request is reasoned through, which shapes its strengths: multi-step software engineering, office-productivity workflows, and long-horizon agentic tasks where a model must plan, call tools, and recover across many turns.

Within Alibaba's lineup, the Max models are the closed, hosted flagships that sit above the open-weight Qwen releases like Qwen 3.5 397B and the compact Qwen 3.6 35B A3B, and above the mid-weight Qwen 3.7 Plus service tier. It builds directly on Qwen 3.7 Max, with the headline gains concentrated in coding and structured work tasks. Native vision-language input covers both images and video, and the model supports function calling and web search.

The one-million-token context makes it practical to load entire repositories, long document sets, or extended agent traces in a single pass. Best suited to complex engineering work, document-heavy analysis, and multi-agent orchestration where reasoning depth matters more than raw latency.

This About section is AI-generated from public sources (Claude Opus 5), with no human editing. It may contain inaccuracies — verify critical details against the sources listed above.

Data sources: Venice API · HuggingFace · Wikipedia — enrichment updated 9h ago