GPT-5.6 Luna price down 80%, Terra down 20% →
Qwen3.8-Max
Aliyun
Q
Text Generation

Qwen3.8-Max

qwen3.8-max
Text modeltext-to-textimage-to-textpdf-to-textvideo-to-text

Qwen3.8-Max is Alibaba Qwen’s flagship large language model designed for advanced reasoning, agentic workflows, multimodal understanding, and enterprise-scale AI applications.. It has 2.4T parameters, adopts the MoE architecture, supports switching between thinking and fast inference modes, can handle various content formats, and performs excellently in scenarios such as code engineering, professional office work, and complex logical reasoning, second only to Anthropic's Fable 5.

From
$1.6/1M tokens
View model
Happy Horse 1.1
Aliyun
Popular
Q
Video Generation

Happy Horse 1.1

happyhorse-1.1
Video modelPopulartext-to-videoimage-to-video

HappyHorse 1.1 is a multimodal video-generation model designed for professional content creation, advertising, short films, social media production, and storytelling. It extends the capabilities of HappyHorse 1.0—which gained significant attention after ranking highly in independent video-generation evaluations—with stronger scene coherence and improved visual fidelity.

From
$0.112/s
View model
Happy Horse 1.0
Aliyun
Q
Video Generation

Happy Horse 1.0

happyhorse-1.0
Video modeltext-to-videoimage-to-video

Happy Horse 1.0 — A high-quality audio-video generation model that supports text-to-video and image-to-video creation. It can generate synchronized visuals, audio, and lip movements, making it suitable for short films, advertising creatives, and product showcases.

From
$0.112/s
View model
Qwen3.7 Plus
Aliyun
Q
Text Generation

Qwen3.7 Plus

qwen3.7-plus
Text modeltext-to-text

Qwen3.7 Plus is a high-performance large language model developed by Alibaba Cloud. It supports long-context understanding up to 128K tokens, function calling, and multilingual tasks. Designed for complex reasoning, coding, and instruction-following scenarios.

From
$0.32/1M tokens
View model
Qwen3.7-Max
Aliyun
Q
Text Generation

Qwen3.7-Max

qwen3.7-max
Text modeltext-to-text

Qwen3.7-Max's core strength lies in the breadth and depth of its agentic capabilities. In coding, it handles everything from front-end prototyping to complex multi-file engineering projects. For office and productivity work, it enables workflow automation through MCP integration and multi-agent collaboration. In long-horizon autonomous execution, it maintained coherent reasoning throughout a 35-hour, fully autonomous kernel optimization experiment involving over 1,000 tool calls — convincingly demonstrating its sustained, stable execution. Furthermore, it delivers consistently strong cross-framework generalization, performing reliably whether deployed in Claude Code, OpenClaw, Qwen Code, or other frameworks.

From
$2/1M tokens
View model
Q

Wan2.7

Aliyun
Q
Video Generation

Wan2.7

wan2.7
Video modeltext-to-videoimage-to-video

Wan2.7 is a video generation model designed for high-quality visual synthesis and improved motion consistency. It is suitable for cinematic content creation and professional video production workflows.

From
$0.08/s
View model
Q

Wan2.6

Aliyun
Q
Video Generation

Wan2.6

wan2.6
Video modelimage-to-videotext-to-video

Wan2.6 is a video generation model designed for stable and efficient video synthesis. It provides reliable visual quality and smooth motion generation for general video creation tasks.

From
$0.08/s
View model
Qwen3.6-Plus
Aliyun
Q
Text Generation

Qwen3.6-Plus

qwen3.6-plus
Text modeltext-to-text

Qwen 3.6-Plus is now available, featuring enhanced code development capabilities and improved efficiency in multimodal recognition and inference, making the Vibe Coding experience even better.

From
$0.4/1M tokens
View model
Q

Qwen 3.5 Flash

Aliyun
Q
Text Generation

Qwen 3.5 Flash

qwen3.5-flash
Text modeltext-to-text

The Qwen-3.5 Flash Series is a production-oriented family of large language models (LLMs) developed by the Alibaba Group under its Qwen initiative. It represents the deployment (hosted/API) layer of the broader Qwen-3.5 model family, optimized for high speed, long-context processing, and agent-based applications. In simple terms: Qwen-3.5 Flash = fast, scalable, long-context, tool-using versions of Qwen-3.5 models designed for real-world production use.

From
$0.16/1M tokens
View model
qwen3.5-plus
Aliyun
Q
Text Generation

qwen3.5-plus

qwen3.5-plus-2026-02-15
Text modeltext-to-text

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency.

From
$0.32/1M tokens
View model
qwen3.5-397b-a17b
Aliyun
Q
Text Generation

qwen3.5-397b-a17b

qwen3.5-397b-a17b
Text modeltext-to-text

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency.

From
$0.48/1M tokens
View model
qwen3 max
Aliyun
Q
Text Generation

qwen3 max

qwen3-max-2026-01-23
Text modeltext-to-text

- qwen3-max: Alibaba Tongyi Qianwen team's latest Qwen3-Max model, positioned as the series' performance peak. - 🧠 Powerful Multimodal and Inference: Supports ultra-long context (up to 128k tokens) and Multimodal input, excels at complex Inference, code generation, translation, and creative content. - ⚡️ Breakthrough Improvement: Significantly optimized across multiple technical indicators, faster response speed, knowledge cutoff up to 2025, suitable for enterprise-level high-precision AI applications.

From
$0.96/1M tokens
View model
Q

Qwen Image

Aliyun
Q
Image Generation

Qwen Image

qwen-image
Image modeltext-to-image

Qwen-Image is a revolutionary image generation foundational model released by Alibaba's Tongyi Qianwen team in 2025. With a parameter scale of 20 billion, it is based on the MMDiT (Multimodal Diffusion Transformer) architecture. The model has achieved significant breakthroughs in complex text rendering and precise image editing, demonstrating exceptional performance particularly in Chinese text rendering. Translated with DeepL.com (free version)

From
$0.028/request
View model
Q

qwen-image-2

Aliyun
Coming soon
Q
Image Generation

qwen-image-2

qwen-image-2
Image modelComing soontext-to-image

qwen-image-2 coming soon

From
Coming soon
View model
Q

qwen3-vl-30b-a3b

Aliyun
Q
Text Generation

qwen3-vl-30b-a3b

qwen3-vl-30b-a3b
Text modeltext-to-text2M context

Qwen3-VL-30B-A3B is a state-of-the-art multimodal AI model in the Qwen3 AI family, developed by Alibaba’s Qwen team. It’s designed to unify language understanding and visual comprehension — including text, images, and video — in a single foundation model.

From
$0.12/1M tokens
View model
Q

qwen3-vl-32b

Aliyun
Q
Text Generation

qwen3-vl-32b

qwen3-vl-32b
Text modeltext-to-text

Qwen3-VL-32B is the 32-billion-parameter dense variant in Alibaba’s Qwen3 vision-language model family. It is a multimodal (vision + language + video) transformer designed for unified perception, long-context reasoning, robust OCR and visual grounding, and agentic/toolified workflows.

From
$0.24/1M tokens
View model
Q

qwen3-vl-235b-a22b

Aliyun
Q
Text Generation

qwen3-vl-235b-a22b

qwen3-vl-235b-a22b
Text modeltext-to-text2M context

qwen3-vl-235b-a22b is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception of real-world/synthetic categories, 2D/3D spatial grounding, and long-form visual comprehension, achieving competitive multimodal benchmark results.

From
$0.24/1M tokens
View model
Q

qwen3-30b-a3b

Aliyun
Popular
Q
AI Model

qwen3-30b-a3b

qwen3-30b-a3b
AI modelPopular

Has 3 billion parameters, balancing performance and resource requirements, suitable for enterprise-level applications. - This model may employ MoE or other optimized architectures, suitable for scenarios requiring efficient processing of complex tasks, such as intelligent customer service and content generation.

From
$0.12/1M tokens
View model
Q

qwen3-coder-plus

Aliyun
Popular
Q
AI Model

qwen3-coder-plus

qwen3-coder-plus
AI modelPopular

Explore the qwen3-coder-plus API.

From
$0.52/1M tokens
View model
Q

qwen3-coder-480b-a35b-instruct

Aliyun
Popular
Q
AI Model

qwen3-coder-480b-a35b-instruct

qwen3-coder-480b-a35b-instruct
AI modelPopular

Explore the qwen3-coder-480b-a35b-instruct API.

From
$0.24/1M tokens
View model
Q

qwen3-coder

Aliyun
Q
AI Model

qwen3-coder

qwen3-coder
AI model

CometAPI’s qwen3-coder is an affordable, OpenAI-compatible coding model API for Qwen3 Coder, optimized for code generation, debugging, and repository-level engineering workflows with ~20% lower pricing.

From
$0.24/1M tokens
View model
Q

qwen3-235b-a22b

Aliyun
Popular
Q
AI Model

qwen3-235b-a22b

qwen3-235b-a22b
AI modelPopularInference,Tool,131.1K

Qwen3-235B-A22B is the flagship model of the Qwen3 series, with 23.5 billion parameters, using a Mixture of Experts (MoE) architecture. - Particularly suitable for complex tasks requiring high-performance Inference, such as coding, mathematics, and Multimodal applications.

From
$0.336/1M tokens
View model
Q

Wan3.0

Aliyun
Coming soon
Q
Video Generation

Wan3.0

wan3.0
Video modelComing soontext-to-videoimage-to-video

coming soon

From
Coming soon
View model
Q

Qwen3.6-Max-Preview

Aliyun
Coming soon
Q
Text Generation

Qwen3.6-Max-Preview

qwen3.6-max-preview
Text modelComing soontext-to-text

Qwen3.6-Max-Preview Compared with Qwen3.6-Plus, this preview version brings stronger world knowledge and instruction compliance capabilities, as well as significantly improved agent programming performance on multiple benchmarks

From
Coming soon
View model