Qwen icon

Qwen Qwen3.8-Max

Qwen3.8-Max

Scale Qwen's Max line to 2.4 trillion parameters with 95 billion active parameters for stronger coding, cowork, research, and long-horizon task execution Build dependable deliverables from open-ended goals with improved autonomous coding, real-world work, multimodal agent, and closed-loop optimization capabilities Prepare for the first Qwen-Max-class open-weight release, with official weights planned for the week after the August 3, 2026 launch

Reviewed by ToolWorthy Editors·updated today·Qwen3.8-Max released today

Pricing:Free + from $0/per 1M input tokens
Categories:
Jump to section
Qwen3.8-Max official release cover

More tools to compare

MakersClaw icon

MakersClaw

TypingMind icon

TypingMind

Doubao icon

Doubao

Z.ai icon

Z.ai

Odysseus icon

Odysseus

ReleaseDock icon

ReleaseDock

Pros & Cons

Pros

  • 2.4T parameter Max-scale model with 95B active parameters.
  • Strong focus on autonomous coding, research workflows, professional cowork tasks, and long-horizon agents.
  • Official examples cover OpenAI-compatible and Anthropic-compatible API usage.
  • 1M context and 65,536 max output token configuration appears in official agent examples.
  • Planned open weights make Qwen3.8-Max unusually important for teams watching open frontier models.

Cons

  • Open weights were announced but not available on launch day.
  • Public pricing for Qwen3.8-Max was not included in the launch article.
  • Vendor benchmarks are promising but should be cross-checked with independent evaluations.
  • Hardware requirements, license details, and exact self-hosting stack guidance depend on the upcoming open-weight release.
  • Video-specific agent limits should be verified in current QwenCloud docs before production use.

Overview

Qwen3.8-Max is Alibaba Qwen's new flagship model, officially released on August 3, 2026. The release moves the Max line to a 2.4 trillion parameter model with 95 billion active parameters, while keeping Qwen focused on practical agent work rather than only short benchmark prompts. Qwen describes it as the most capable model in the Qwen family to date, with stronger coding, cowork, research, long-horizon task execution, and multimodal agent capability.

The launch is also important for model access strategy. Qwen says this is the first Qwen-Max-class model that will receive open weights, but those weights were not downloadable at launch. The official plan is to release them during the week after August 3, 2026, so teams should treat hosted QwenCloud access as the immediate production route and open-weight deployment as a near-term follow-up.

What's New

2.4T Max-Scale Model

Qwen3.8-Max succeeds the Qwen3.7-Max flagship line with 2.4 trillion total parameters and 95 billion active parameters. The model builds on the architectural foundation of Qwen3.5, but targets harder end-to-end workflows: codebases, research reproduction, competition pipelines, professional document work, and agent systems that need to keep state across many tool calls.

For ToolWorthy readers comparing AI coding tools, the practical signal is that Qwen is positioning Qwen3.8-Max as a coding agent model, not just a chat model with coding answers.

Autonomous Coding and Research Workflows

The official release examples focus on long-running coding tasks. Qwen reports a 10+ day autonomous coding run for a self-evolving CLI project, a research-paper reproduction and improvement task, and an online multimodal-dialogue competition workflow. These examples emphasize planning, code generation, test execution, feedback loops, and leaderboard-style iteration.

The more relevant takeaway for builders is not that every task should run unattended for days. It is that Qwen3.8-Max is designed for agent harnesses where the model has to plan, edit, run tools, inspect failures, and continue without constant human prompting.

Stronger Cowork and Professional Task Coverage

Qwen's launch article also tests the model against professional workflows across law, finance, design, manufacturing, healthcare-style demos, sports analytics, menu planning, and structural modeling. The common pattern is broad work execution: ingesting messy documents or briefs, producing a structured deliverable, and maintaining enough task context to complete the workflow in one session.

That makes Qwen3.8-Max relevant for teams building AI agent systems that need dependable outputs, not only conversational assistance.

Long-Horizon and Closed-Loop Optimization

Qwen highlights long-horizon tasks such as autonomous chip-design optimization and continuous learning in operational settings. These examples show the model using feedback loops instead of a fixed single-shot plan: propose an action, run a tool or simulation, inspect the result, and update the next step.

This is a useful distinction from ordinary long-context summarization. Qwen3.8-Max is meant to use context, tools, and iterative evaluation together.

Multimodal Agent Inputs

The official assistant and agent configuration examples list text and image input support. Qwen also describes the model as improving multimodal agents, including workflows involving screenshots, images, manuscripts, and video-like work inputs. For production planning, treat text and image support as the directly documented model configuration, and verify any video-specific workflow limits in current QwenCloud documentation before building around them.

Benchmark Notes

Qwen published a full benchmark table comparing Qwen3.8-Max with Opus 4.8, Fable 5, GPT-5.6 Sol, and Qwen3.7-Max. Selected Qwen3.8-Max results include:

Area Benchmark Qwen3.8-Max
Coding agent Terminal Bench 2.1 86.6
Coding agent SWE-bench Pro 67.7
Coding agent PaperBench 93.0
Coding agent QwenSWEBench 80.7
General agent CoWorkBench 74.8
General agent WorkSpaceBench 67.7
General agent WideSearch 81.9
General capability IFBench 82.8
General capability GPQA Diamond 92.6
Long context MRCR v2 256K, 8-needle 92.9

These are vendor-published numbers, so they are best read as launch-positioning data until independent evaluations appear. Still, the table is directionally useful: Qwen3.8-Max is strongest in the release narrative where agentic coding, professional work, and long-context execution overlap.

Availability & Access

Qwen3.8-Max is available through QwenCloud at launch. The official article provides API examples for QwenCloud's OpenAI-compatible endpoint and Anthropic-compatible endpoint, plus integration examples for Claude Code, Codex, Qoder CLI, Qwen Code, and OpenClaw.

The model identifier shown in Qwen's examples is qwen3.8-max. Agent configuration examples list a 1,000,000 token context window and 65,536 maximum output tokens. The same examples list reasoning support and text-plus-image inputs.

Open-weight access is not live on August 3, 2026. Qwen states that the Qwen-Max-class open weights are planned for release the following week. Teams that need self-hosting should wait for the official weights, license, model card, and hardware guidance before committing deployment architecture.

Pricing & Plans

Qwen's launch article does not publish a standalone token-pricing table for Qwen3.8-Max. Immediate hosted access is through QwenCloud / Alibaba Cloud Model Studio, so production costs should be verified on the current Model Studio pricing page or inside the QwenCloud console before launch.

For now, the safest planning assumption is:

  • Qwen Chat / Qwen Studio: useful for hands-on evaluation where access is available.
  • QwenCloud API: the production route at launch, billed according to Model Studio account terms and current regional pricing.
  • Open weights: planned for the week after August 3, 2026, but not yet a usable self-hosting option at launch.

Best For

  • Engineering teams building long-running coding agents that edit, test, and recover from failed runs.
  • Research teams reproducing papers, running experiments, and iterating on results.
  • Product teams building AI chatbot or cowork assistants that need structured deliverables rather than short answers.
  • Enterprises testing agent workflows across legal review, financial research, manufacturing analysis, or document-heavy operations.
  • Open-model teams waiting for a Max-class Qwen release they can evaluate for self-hosting after the weights ship.

FAQ

Is Qwen3.8-Max officially released?

Yes. Qwen officially released Qwen3.8-Max on August 3, 2026, through the Qwen blog and made hosted access available via QwenCloud.

Is Qwen3.8-Max open source?

Not on launch day. Qwen says Qwen3.8-Max will be the first Qwen-Max-class model with open weights, and that the weights are planned for release the week after August 3, 2026. Until those files and license details are live, treat it as hosted-access first.

What is the Qwen3.8-Max model ID?

The official API and agent examples use qwen3.8-max.

Does Qwen3.8-Max support a 1M context window?

Yes. Qwen's official agent configuration examples list contextWindow as 1,000,000 and maxTokens as 65,536.

Does Qwen3.8-Max support images?

Yes. Official configuration examples list text and image input support. The launch article also describes Qwen3.8-Max as stronger for multimodal agents, but teams should verify any video-specific input limits in current QwenCloud documentation.

How much does Qwen3.8-Max cost?

Qwen did not publish a dedicated Qwen3.8-Max pricing table in the launch article. For production API use, check the current Alibaba Cloud Model Studio or QwenCloud pricing page for your account region and token tier.

Version History

Qwen3.8-Max

Current Version

Released on August 3, 2026

+What's new
3 updates
  • Scale Qwen's Max line to 2.4 trillion parameters with 95 billion active parameters for stronger coding, cowork, research, and long-horizon task execution
  • Build dependable deliverables from open-ended goals with improved autonomous coding, real-world work, multimodal agent, and closed-loop optimization capabilities
  • Prepare for the first Qwen-Max-class open-weight release, with official weights planned for the week after the August 3, 2026 launch

Qwen3.7-Max

Released on May 21, 2026

+What's new
3 updates
  • Use the next-generation Qwen Max flagship for coding, office productivity, and long-horizon autonomous execution through Alibaba Cloud Model Studio
  • Run text-only agent workflows with thinking mode enabled by default, explicit cache support, built-in tools, and a 1M-token context window
  • Upgrade from earlier Qwen Max snapshots with stronger planning, tool use, and sustained execution for complex software and professional work

Qwen3.6-Plus

Released on April 1, 2026

View Update
+What's new
3 updates
  • Process up to 1 million tokens of context with 65,536 output tokens per response, enabling analysis of entire codebases and multi-thousand-page documents in a single request
  • Reach 80.9 on SWE-bench Verified and 77.5 on SWE-bench Multilingual, strengthening agentic coding reliability and multi-step software workflows over Qwen3.5
  • Build real-world coding workflows more reliably with significantly improved agentic coding, stronger frontend development, and sharper multimodal reasoning for complex tasks

Qwen3.5 Small Series

Released on March 9, 2026

View Update
+What's new
3 updates
  • Run the 0.8B, 2B, 4B, and 9B Qwen3.5 small models locally through standard runtimes, giving developers lightweight multimodal options for edge and consumer hardware
  • Scale reinforcement learning across the small-model line to improve real-world adaptability, with the 9B model card emphasizing stronger generalization under progressively harder tasks
  • Use native text-plus-vision models across the small line, with 4B and 9B variants exposing 248K-vocabulary multimodal architectures designed for OCR, spatial understanding, and agents

Qwen3.5

Released on February 15, 2026

View Update
+What's new
3 updates
  • Process text, images, and video natively in a unified model with early-fusion architecture—no separate VL variant needed—enabling seamless cross-modal reasoning at frontier quality
  • Decode 8.6× faster at 32K context and 19× faster at 256K context than Qwen3-Max, with a hybrid Gated DeltaNet + sparse MoE activating only 17B of 397B parameters per token
  • Post strong agentic results including AndroidWorld (66.8), BrowseComp (69.0), and NOVA-63 (59.1), while expanding support to 201 languages and a roughly 250K-token vocabulary

Qwen3-VL-Embedding

Released on January 7, 2026

+What's new
2 updates
  • Improve multimodal retrieval accuracy for text, image, and mixed content search across large document collections and knowledge bases
  • Represent text, images, visual documents, and video in one shared embedding space, supporting modern multimodal retrieval and reranking workflows from a single model family

Qwen3-TTS Voice Cloning & Voice Design

Released on December 22, 2025

+What's new
2 updates
  • Design custom voices with Qwen3-TTS-VD-Flash voice design model for personalized audio experiences in podcasts and audiobook production
  • Clone voices naturally with Qwen3-TTS-VC-Flash for high-fidelity speech synthesis that preserves speaker characteristics and emotional tones

Qwen3-Max

Released on September 23, 2025

+What's new
3 updates
  • Process up to 256K tokens with Alibaba's largest model featuring over 1 trillion parameters trained on 36 trillion tokens for handling extensive codebases
  • Solve real-world coding challenges with strong performance on industry benchmarks including SWE-Bench Verified and similar software engineering evaluation suites
  • Execute complex agent workflows effectively with advanced tool-calling capabilities demonstrated across multiple multi-step reasoning and automation benchmarks

Qwen VLo

Released on June 26, 2025

+What's new
3 updates
  • Edit images using natural language instructions with unified multimodal understanding and generation capabilities for design refinement workflows
  • Generate high-quality images from text descriptions while maintaining semantic consistency and artistic coherence across multiple generation iterations
  • Process multilingual instructions for global creative workflows enabling teams worldwide to collaborate on visual content creation seamlessly

Qwen3

Released on April 29, 2025

+What's new
3 updates
  • Think deeper with hybrid thinking modes that combine fast and slow reasoning for complex problem-solving across diverse scenarios
  • Choose between dense and Mixture-of-Expert (MoE) architectures to optimize for your specific performance and efficiency requirements
  • Communicate naturally in multiple languages with significantly enhanced multilingual understanding and generation capabilities

Qwen2.5-VL

Released on January 26, 2025

+What's new
3 updates
  • Process and understand videos up to 1+ hour in length with advanced long video comprehension capabilities for analyzing presentations and tutorials
  • Build visual agents that can interact with UI elements and analyze screenshots for automation workflows and quality assurance testing scenarios
  • Generate structured JSON outputs from images and multimodal inputs for seamless integration with business systems and data processing pipelines

Qwen2.5-Coder Family

Released on November 12, 2024

+What's new
2 updates
  • Access six model sizes ranging from 0.5B to 32B parameters optimized for different coding scenarios from edge devices to complex enterprise systems
  • Generate more accurate code completions with expanded training on diverse programming languages and frameworks including Python, JavaScript, Java, and modern web stacks

Qwen2.5

Released on September 19, 2024

+What's new
3 updates
  • Benefit from training on up to 18 trillion tokens delivering significantly improved knowledge base and reasoning capabilities across diverse domains
  • Achieve strong performance across benchmarks including 85+ on MMLU knowledge tests, 85+ on HumanEval coding challenges, and 80+ on MATH problem-solving tasks
  • Choose from seven model sizes ranging from 0.5B to 72B parameters plus specialized variants including Qwen2.5-Coder and Qwen2.5-Math for domain-specific tasks

Qwen2

Released on June 7, 2024

+What's new
3 updates
  • Communicate in 29 languages with comprehensive multilingual support extending far beyond English and Chinese for global application deployment
  • Process up to 128K tokens in a single context window for handling extensive documents, long conversations, and multi-document analysis workflows
  • Deploy across five model sizes ranging from 0.5B to 72B parameters to match different computational requirements and infrastructure constraints

Qwen1.5-110B

Released on April 25, 2024

+What's new
2 updates
  • Scale to 110 billion parameters as the largest model in the Qwen1.5 series designed specifically for handling complex reasoning tasks and advanced applications
  • Achieve superior performance on advanced benchmarks compared to smaller Qwen1.5 variants with enhanced capabilities in mathematics, coding, and logical inference

Qwen1.5

Released on February 4, 2024

+What's new
3 updates
  • Choose from eight model sizes including Mixture-of-Experts (MoE) architecture for flexible deployment options matching your performance and cost requirements
  • Process 32,768 tokens uniformly across all model variants with consistent context length support for reliable multi-document processing workflows
  • Communicate more naturally with enhanced multilingual capabilities and improved human alignment delivering better instruction-following and conversational quality

Qwen-VL-Plus

Released on January 25, 2024

+What's new
2 updates
  • Extract text accurately from ultra-high-resolution images with millions of pixels enabling professional document processing and OCR workflows at scale
  • Recognize complex visual patterns with significantly improved image understanding capabilities for detailed scene analysis and object detection applications

Qwen-VL

Released on August 22, 2023

+What's new
2 updates
  • Understand and analyze images with Qwen's first vision-language model enabling visual question answering and image captioning capabilities
  • Answer questions about visual content for multimodal AI applications including document analysis, scene understanding, and visual reasoning tasks

Qwen1

Released on August 3, 2023

+What's new
3 updates
  • Access Qwen's first open-source 7B-parameter model trained on over 2 trillion multilingual tokens covering Chinese, English, code, and mathematics
  • Process up to 8,000 tokens in a single context window for handling medium-length documents and multi-turn conversations efficiently
  • Deploy for commercial use following official licensing terms which may require registration or approval for certain business applications

Top alternatives

Related categories

From the blog

View all →

Track Qwen in ToolWorthy Weekly

Important tool updates, better alternatives, and selected AI signals in one weekly brief.

Weekly only. Unsubscribe anytime.