
Qwen 3.8 Max costs $2/M input and $6/M output. Compare cache fees, tool costs, real examples, limits, and Qwen 3.7 migration advice.

On Lunar New Year’s Eve (Feb 16–17, 2026), Alibaba Group released its next-generation model, Qwen 3.5 — a multimodal, agent-capable model positioned for what the company calls an “agentic AI” era. Industry coverage highlighted claims of large gains in efficiency and cost, and rapid support from hardware and cloud vendors. CometAPI is options for developers who want hosted API access or an OpenAI-compatible integration, while AMD announced Day-0 GPU support for the model on its Instinct line. ByteDance is one of the principal domestic competitors that released upgrades around the same holiday window. OpenAI remains a reference point for comparison in benchmarks and integration style.

Alibaba’s new Qwen3.5 is a major step forward — it closes the gap with, and in some agentic / multimodal workloads claims parity or advantage over, certain frontier closed-source models on a number of public benchmarks and internal tests. However, “outperform” depends on the workload: on agentic tool-use, multimodal document/video understanding, and cost-per-inference Qwen3.5 is reported to be extremely competitive (and in some vendor charts ahead). The practical takeaway: Qwen3.5 appears to be a genuine frontier contender in early 2026 — for many enterprise agentic and multimodal use cases it is now viable as a primary option.
Alibaba’s Qwen3-Max-Thinking — the “thinking” variant of the massive Qwen3 family — has become one of the headline stories in AI this year: a trillion-plus parameter flagship tuned for deep reasoning, long-context understanding and agentic workflows. In short, it’s the vendor’s move to give applications a slower, more traceable “System-2” mode of thought: the model doesn’t just answer, it can show (and use) steps, tools, and intermediate checks in a controlled way.

Alibaba’s Qwen team has released Qwen3-Max-Preview (Instruct) — the company’s largest model to date, with more than 1 trillion parameters — and made it

Qwen3-Max-Preview is Alibaba’s latest flagship preview model in the Qwen3 family — a trillion+-parameter, Mixture-of-Experts (MoE) style model with an ultra-long 262k token context window, released in preview for enterprise/cloud use. It targets *deep reasoning, long-document understanding, coding, and agentic workflows.

Alibaba’s latest advance in artificial intelligence, Qwen3-Coder, marks a significant milestone in the rapidly evolving landscape of AI-driven software

The Qwen3-Coder API is a code generation and completion API based on Alibaba’s Qwen3 language model family, optimized for software development tasks such as writing, understanding, and debugging code across multiple programming languages.