Claude Opus 5 is now live on CometAPI โ†’

Models

Claude Opus 5
A

Claude Opus 5

Input:$4/M
Output:$20/M

Claude Opus 5 is available today. Itโ€™s a thoughtful and proactive model that comes close to the frontier intelligence of Claude Fable 5 at half the price.

Flux 3
F

Flux 3

Coming soon
Input:$60/M
Output:$240/M

coming soon

GPT 5.6
O

GPT 5.6

Input:$4/M
Output:$24/M

Start with GPT-5.6 Sol for complex reasoning and coding, choose GPT-5.6 Terra to balance intelligence and cost, or use GPT-5.6 Luna for cost-sensitive, high-volume workloads.

Gemini 3.6 Flash
G

Gemini 3.6 Flash

Input:$1.2/M
Output:$6/M

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world tasks at a higher speed and lower cost. Designed for the agentic era, it excels at code generation, agentic execution, and spatial reasoning. This model is particularly effective for rapid agentic loops involving complex coding cycles and iterations.

Nano Banana 2 lite
G

Nano Banana 2 lite

Popular
Input:$0.2/M
Output:$1.2/M

Gemini 3.1 Flash Lite Image model is an efficiency expert in the image generation family, designed for ultra-low latency and cost-effective image generation and modification.

Claude Sonnet 5
C

Claude Sonnet 5

Popular
Context:1M
Input:$1.6/M
Output:$8/M

Claude Sonnet 5 API is live on CometAPI at $1.6 per million input tokens and $8 per million output tokens, 20 percent below Anthropic list price. One key gives you Claude Sonnet 5 plus 500+ models from OpenAI, Google, and ByteDance under all in one pricing and a single invoice. No seat fees. No monthly minimum. Pay only for what you use. Start free and make your first call in under five minutes.

Seedance-2-5
B

Seedance-2-5

Coming soon
Input:$60/M
Output:$240/M

coming soon

Happy Horse 1.1
Q

Happy Horse 1.1

Popular
Per Second:$0.112

HappyHorse 1.1 is a multimodal video-generation model designed for professional content creation, advertising, short films, social media production, and storytelling. It extends the capabilities of HappyHorse 1.0โ€”which gained significant attention after ranking highly in independent video-generation evaluationsโ€”with stronger scene coherence and improved visual fidelity.

Claude Fable 5
C

Claude Fable 5

Popular
Input:$8/M
Output:$40/M

Access to Claude Fable 5 has been restored. It brings 5th-generation intelligence to your most ambitious coding and professional work.

GPT Image 2
O

GPT Image 2

Popular
Input:$4/M
Output:$24/M

GPT Image 2 is openai state-of-the-art image generation model for fast, high-quality image generation and editing. It supports flexible image sizes and high-fidelity image inputs.

Seedance 2-0
B

Seedance 2-0

Popular
Per Second:$0.056

Seedance 2.0 is ByteDanceโ€™s next-generation multimodal video foundation model focused on cinematic, multi-shot narrative video generation. Unlike single-shot text-to-video demos, Seedance 2.0 emphasizes reference-based control (images, short clips, audio), coherent character/style consistency across shots, and native audio/video synchronization โ€” aiming to make AI video useful for professional creative and previsualization workflows.

Claude Opus 4.8
C

Claude Opus 4.8

Popular
Context:200K tokens
Input:$4/M
Output:$20/M

Claude Opus 4.8 is a premium AI model designed for advanced reasoning, deep analysis, and high-quality content generation. It excels at handling complex instructions, long-context understanding, and sophisticated problem-solving across professional and technical domains.

Gemini 3.5 Flash
G

Gemini 3.5 Flash

Popular
Input:$1.2/M
Output:$7.2/M

Gemini 3.5 Flash is a high-speed AI model designed for fast response and efficient coding performance. It delivers significantly improved generation speed while maintaining strong reasoning ability, making it suitable for real-time applications and developer workflows.

Gemini 3.1 Pro
G

Gemini 3.1 Pro

Input:$1.6/M
Output:$9.6/M

Gemini 3.1 Pro is the next generation in the Gemini series of models, a suite of highly-capable, natively multimodal, reasoning models. Gemini 3 Pro is now Googleโ€™s most advanced model for complex tasks, and can comprehend vast datasets, challenging problems from different information sources, including text, audio, images, video, and entire code repositories

Kimi K3
M

Kimi K3

Popular
Context:1,000k tokens
Input:$2.4/M
Output:$12/M

Kimi K3 is Kimi's flagship model, designed for long-range programming and end-to-end knowledge work, featuring 1M token context and leading-edge comprehensive intelligence.

Kimi K2.7 Code
M

Kimi K2.7 Code

Input:$0.76/M
Output:$3.19998/M

Kimi K2.7 Code is Kimi's most intelligent coding model to date, reliably following instructions in long contexts and completing programming tasks with a higher success rate. It supports text, image, and video input, and only supports thought mode, dialogue, and agent tasks.

Happy Horse 1.0
Q

Happy Horse 1.0

Per Second:$0.112

Happy Horse 1.0 โ€” A high-quality audio-video generation model that supports text-to-video and image-to-video creation. It can generate synchronized visuals, audio, and lip movements, making it suitable for short films, advertising creatives, and product showcases.

Claude Mythos 5
C

Claude Mythos 5

Coming soon
Input:$8/M
Output:$40/M

Anthropic's most capable, widely released model, for the most demanding reasoning and long-horizon agentic work

Claude Opus 4.7
C

Claude Opus 4.7

Popular
Input:$4/M
Output:$20/M

Claude Opus 4.7 is a hybrid reasoning model designed specifically for frontier-level coding, AI agents, and complex multi-step professional work. Unlike lighter models (e.g., Sonnet or Haiku variants), Opus 4.7 prioritizes depth, consistency, and autonomy on the hardest tasks.

MiniMax-M3
M

MiniMax-M3

Input:$0.48/M
Output:$1.92/M

Minimax-m3 is a multimodal AI model designed for strong reasoning, natural conversation, and creative content generation. It provides balanced performance across text and visual understanding tasks, making it suitable for general-purpose AI applications.

GPT 5.5 Pro
O

GPT 5.5 Pro

Input:$24/M
Output:$144/M

GPT-5.5 Pro combines state-of-the-art intelligence, precision, and efficiency to tackle sophisticated challenges. From software development and data analysis to research and decision support, it delivers expert-level assistance with speed and consistency.

DeepSeek V4 Pro
D

DeepSeek V4 Pro

Popular
Input:$0.416/M
Output:$0.832/M

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding, and long-horizon agent workflows, with strong performance across knowledge, math, and software engineering benchmarks.

GPT 5.5
O

GPT 5.5

Popular
Input:$4/M
Output:$24/M

Model 5.5 is a next-generation AI model designed for stronger reasoning, faster responses, and improved accuracy across a wide range of tasks. It excels at understanding complex instructions, generating high-quality content, and assisting with coding, analysis, and problem-solving.

MiniMax-M2.7
M

MiniMax-M2.7

Input:$0.24/M
Output:$0.96/M

MiniMax-M2.7 offers the same top-tier intelligence as the standard versionโ€”including recursive self-evolution and expert-level office productivityโ€”but is designed for applications requiring sub-second latency and high-speed token generation. Leveraging an enhanced inference backbone architecture, its output speed is 66% faster than the standard model (reaching 100 tps). It is the preferred choice for interactive programming assistants, real-time agent loop execution, and high-throughput enterprise pipelines with stringent completion time requirements.

DeepSeek V4 Flash
D

DeepSeek V4 Flash

Popular
Input:$0.12/M
Output:$0.24/M

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and high-throughput workloads, while maintaining strong reasoning and coding performance.

GPT-5.4 nano
O

GPT-5.4 nano

Context:400,000
Input:$0.16/M
Output:$1/M

GPT-5.4 Nano is an ultra-lightweight AI model built for maximum speed and efficiency. It is optimized for simple tasks, real-time interactions, and large-scale deployments where low latency and minimal resource consumption are essential.

GPT-5.4 mini
O

GPT-5.4 mini

Context:400,000
Input:$0.6/M
Output:$3.6/M

GPT-5.4 Mini is a lightweight and efficient AI model optimized for speed and everyday productivity. It provides reliable conversational capabilities, content generation, and task assistance while maintaining low latency and resource usage.

GPT-5.4 pro
O

GPT-5.4 pro

Context:1,050,000
Input:$24/M
Output:$144/M

GPT-5.4 Pro is a high-performance AI model designed for professional and business applications. It offers strong reasoning, reliable accuracy, and efficient execution across tasks such as content creation, coding, research, and data analysis.

Nano Banana 2
G

Nano Banana 2

Popular
Input:$0.4/M
Output:$2.4/M

Core Capabilities Overview: Resolution: Up to 4K (4096ร—4096), on par with Pro. Reference Image Consistency: Up to 14 reference images (10 objects + 4 characters), maintaining style/character consistency. Extreme Aspect Ratios: New 1:4, 4:1, 1:8, 8:1 ratios added, suitable for long images, posters, and banners. Text Rendering: Advanced text generation, suitable for infographics and marketing poster layouts. Search Enhancement: Integrated Google Search + Image Search. Grounding: Built-in thinking process; complex prompts are reasoned before generation.

Claude Sonnet 4.6
C

Claude Sonnet 4.6

Popular
Input:$2.4/M
Output:$12/M

Claude Sonnet 4.6 is our most capable Sonnet model yet. Itโ€™s a full upgrade of the modelโ€™s skills across coding, computer use, long-context reasoning, agent planning, knowledge work, and design. Sonnet 4.6 also features a 1M token context window in beta.