Claude Opus 5 is now live on CometAPI →
Gemini 3.6 Flash
G

Gemini 3.6 Flash

Input:$1.2/M
Output:$6/M

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world tasks at a higher speed and lower cost. Designed for the agentic era, it excels at code generation, agentic execution, and spatial reasoning. This model is particularly effective for rapid agentic loops involving complex coding cycles and iterations.

Nano Banana 2 lite
G

Nano Banana 2 lite

Popular
Input:$0.2/M
Output:$1.2/M

Gemini 3.1 Flash Lite Image model is an efficiency expert in the image generation family, designed for ultra-low latency and cost-effective image generation and modification.

Gemini 3.5 Flash
G

Gemini 3.5 Flash

Popular
Input:$1.2/M
Output:$7.2/M

Gemini 3.5 Flash is a high-speed AI model designed for fast response and efficient coding performance. It delivers significantly improved generation speed while maintaining strong reasoning ability, making it suitable for real-time applications and developer workflows.

Gemini 3.1 Pro
G

Gemini 3.1 Pro

Input:$1.6/M
Output:$9.6/M

Gemini 3.1 Pro is the next generation in the Gemini series of models, a suite of highly-capable, natively multimodal, reasoning models. Gemini 3 Pro is now Google’s most advanced model for complex tasks, and can comprehend vast datasets, challenging problems from different information sources, including text, audio, images, video, and entire code repositories

Nano Banana 2
G

Nano Banana 2

Popular
Input:$0.4/M
Output:$2.4/M

Core Capabilities Overview: Resolution: Up to 4K (4096×4096), on par with Pro. Reference Image Consistency: Up to 14 reference images (10 objects + 4 characters), maintaining style/character consistency. Extreme Aspect Ratios: New 1:4, 4:1, 1:8, 8:1 ratios added, suitable for long images, posters, and banners. Text Rendering: Advanced text generation, suitable for infographics and marketing poster layouts. Search Enhancement: Integrated Google Search + Image Search. Grounding: Built-in thinking process; complex prompts are reasoned before generation.

Gemini omni fast
G

Gemini omni fast

Per Request:$0.4

Gemini Omni Fast is a lightweight multimodal video generation model designed for fast and flexible content creation. It enables efficient video generation with support for multiple input types, making it suitable for interactive and iterative workflows.

G

Veo 3.1 Fast

G

Veo 3.1 Fast

Popular
Per Second:$0.08

Veo 3.1 Fast is an optimized version of Veo 3.1 designed for rapid video generation and high-throughput production. It maintains strong visual quality while significantly improving speed, making it suitable for batch content creation and scalable workflows.

Veo 3.1
G

Veo 3.1

Popular
Per Second:$0.32

Veo 3.1 is a flagship AI video generation model designed for high-quality cinematic output with native audio synchronization. It delivers realistic motion, strong visual fidelity, and tightly aligned audio-visual generation for professional video production.

Gemini 3.1 Flash-Lite
G

Gemini 3.1 Flash-Lite

Input:$0.2/M
Output:$1.2/M

Gemini 3.1 Flash-Lite is a highly cost-efficient and low-latency Tier-3 model in Google’s Gemini 3 series, designed for high-volume production AI workflow where throughput and speed matter more than maximal reasoning depth. It combines a large multimodal context window with efficient inference performance at a lower cost than most flagship counterparts.

Nano Banana Pro
G

Nano Banana Pro

Popular
Input:$1.5616/M
Output:$9.3696/M

Nano Banana Pro is an AI model for general-purpose assistance in text-centric workflows. It is suitable for instruction-style prompting to generate, transform, and analyze content with controllable structure. Typical uses include chat assistants, document summarization, knowledge QA, and workflow automation. Public technical details are limited; integration aligns with common AI assistant patterns such as structured outputs, retrieval-augmented prompts, and tool or function calling.

G

Veo 3 Fast

G

Veo 3 Fast

Per Second:$0.08

Veo 3 Fast is Google’s speed-optimized variant of the Veo family of generative video models (Veo 3 / Veo 3.1 etc.). It is engineered to produce short, high-quality video clips with natively generated audio while prioritizing throughput and cost per second—trading some top-end visual fidelity and/or longer single-shot duration for much faster generation and lower price. What is Veo 3 Fast — concise introduction

G

Veo 3

G

Veo 3

Per Second:$0.32

Google DeepMind’s Veo 3 represents the cutting edge of text-to-video generation, marking the first time a large-scale generative AI model seamlessly synchronizes high-fidelity video with accompanying audio—including dialogue, sound effects, and ambient soundscapes.

Gemini 3 Flash
G

Gemini 3 Flash

Context:1,048,576
Input:$0.4/M
Output:$2.4/M

Gemini 3 Flash is a lightweight, efficient multimodal large-scale model from Google tailored for real-world scenarios that require fast responses and low latency.

G

gemini-3.1-flash

G

gemini-3.1-flash

Coming soon
Input:$0.4/M
Output:$2.4/M

gemini-3.1-flash coming soon

G

Gemini 3.5 Pro

G

Gemini 3.5 Pro

Coming soon
Input:$60/M
Output:$240/M

coming soon

G

Gemini 3.5 Flash Lite

G

Gemini 3.5 Flash Lite

Input:$0.24/M
Output:$2.016/M

Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for high-throughput, low-cost execution for subagent tasks and document parsing. The model supports text, image, video, audio, and PDF inputs, and is designed for high-volume agentic workflows, simple data extraction, and applications where latency and API cost are the primary constraints

G

Veo 4

G

Veo 4

Coming soon
Input:$60/M
Output:$240/M

coming soon