Claude Opus 5 is now live on CometAPI →
Gemini 3.6 Flash
Google
G
Text Generation

Gemini 3.6 Flash

gemini-3.6-flash
Text model20% offPrice droptext-to-text

Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world tasks at a higher speed and lower cost. Designed for the agentic era, it excels at code generation, agentic execution, and spatial reasoning. This model is particularly effective for rapid agentic loops involving complex coding cycles and iterations.

From
$1.88$1.5/1M tokens
View model
Nano Banana 2 lite
Google
Popular
G
Image Generation

Nano Banana 2 lite

gemini-3.1-flash-lite-image
Image modelPopular20% offPrice droptext-to-image

Gemini 3.1 Flash Lite Image model is an efficiency expert in the image generation family, designed for ultra-low latency and cost-effective image generation and modification.

From
$0.31$0.25/1M tokens
View model
Gemini 3.5 Flash
Google
Popular
G
Text Generation

Gemini 3.5 Flash

gemini-3.5-flash
Text modelPopular40% offPrice droptext-to-textimage-to-textvideo-to-text

Gemini 3.5 Flash is a high-speed AI model designed for fast response and efficient coding performance. It delivers significantly improved generation speed while maintaining strong reasoning ability, making it suitable for real-time applications and developer workflows.

From
$2.50$1.5/1M tokens
View model
Gemini 3.1 Pro
Google
G
Text Generation

Gemini 3.1 Pro

gemini-3.1-pro-preview
Text model25% offPrice droptext-to-textimage-to-textvideo-to-textspeech-to-text

Gemini 3.1 Pro is the next generation in the Gemini series of models, a suite of highly-capable, natively multimodal, reasoning models. Gemini 3 Pro is now Google’s most advanced model for complex tasks, and can comprehend vast datasets, challenging problems from different information sources, including text, audio, images, video, and entire code repositories

From
$2.67$2/1M tokens
View model
Nano Banana 2
Google
Popular
G
Image Generation

Nano Banana 2

gemini-3.1-flash-image-preview
Image modelPopular20% offPrice droptext-to-imageimage-editing

Core Capabilities Overview: Resolution: Up to 4K (4096×4096), on par with Pro. Reference Image Consistency: Up to 14 reference images (10 objects + 4 characters), maintaining style/character consistency. Extreme Aspect Ratios: New 1:4, 4:1, 1:8, 8:1 ratios added, suitable for long images, posters, and banners. Text Rendering: Advanced text generation, suitable for infographics and marketing poster layouts. Search Enhancement: Integrated Google Search + Image Search. Grounding: Built-in thinking process; complex prompts are reasoned before generation.

From
$0.63$0.5/1M tokens
View model
Gemini omni fast
Google
G
Video Generation

Gemini omni fast

omni-fast
Video model20% offPrice droptext-to-videovideo-editing

Gemini Omni Fast is a lightweight multimodal video generation model designed for fast and flexible content creation. It enables efficient video generation with support for multiple input types, making it suitable for interactive and iterative workflows.

From
$0.63$0.5/request
View model
G

Veo 3.1 Fast

Google
Popular
G
Video Generation

Veo 3.1 Fast

veo3.1-fast
Video modelPopular25% offPrice droptext-to-videoimage-to-video

Veo 3.1 Fast is an optimized version of Veo 3.1 designed for rapid video generation and high-throughput production. It maintains strong visual quality while significantly improving speed, making it suitable for batch content creation and scalable workflows.

From
$0.13$0.1/s
View model
Veo 3.1
Google
Popular
G
Video Generation

Veo 3.1

veo3.1
Video modelPopular25% offPrice dropimage-to-videotext-to-video

Veo 3.1 is a flagship AI video generation model designed for high-quality cinematic output with native audio synchronization. It delivers realistic motion, strong visual fidelity, and tightly aligned audio-visual generation for professional video production.

From
$0.53$0.4/s
View model
Gemini 3.1 Flash-Lite
Google
G
Text Generation

Gemini 3.1 Flash-Lite

gemini-3.1-flash-lite-preview
Text model20% offPrice droptext-to-textimage-to-textspeech-to-textvideo-to-text

Gemini 3.1 Flash-Lite is a highly cost-efficient and low-latency Tier-3 model in Google’s Gemini 3 series, designed for high-volume production AI workflow where throughput and speed matter more than maximal reasoning depth. It combines a large multimodal context window with efficient inference performance at a lower cost than most flagship counterparts.

From
$0.31$0.25/1M tokens
View model
Nano Banana Pro
Google
Popular
G
Image Generation

Nano Banana Pro

gemini-3-pro-image
Image modelPopular25% offPrice droptext-to-imageimage-editing

Nano Banana Pro is an AI model for general-purpose assistance in text-centric workflows. It is suitable for instruction-style prompting to generate, transform, and analyze content with controllable structure. Typical uses include chat assistants, document summarization, knowledge QA, and workflow automation. Public technical details are limited; integration aligns with common AI assistant patterns such as structured outputs, retrieval-augmented prompts, and tool or function calling.

From
$2.60$1.952/1M tokens
View model
G

Veo 3 Fast

Google
G
Video Generation

Veo 3 Fast

veo3-fast
Video model25% offPrice droptext-to-videoimage-to-video

Veo 3 Fast is Google’s speed-optimized variant of the Veo family of generative video models (Veo 3 / Veo 3.1 etc.). It is engineered to produce short, high-quality video clips with natively generated audio while prioritizing throughput and cost per second—trading some top-end visual fidelity and/or longer single-shot duration for much faster generation and lower price. What is Veo 3 Fast — concise introduction

From
$0.13$0.1/s
View model
G

Veo 3

Google
G
Video Generation

Veo 3

veo3
Video model25% offPrice droptext-to-videoimage-to-video

Google DeepMind’s Veo 3 represents the cutting edge of text-to-video generation, marking the first time a large-scale generative AI model seamlessly synchronizes high-fidelity video with accompanying audio—including dialogue, sound effects, and ambient soundscapes.

From
$0.53$0.4/s
View model
Gemini 3 Flash
Google
G
Text Generation

Gemini 3 Flash

gemini-3-flash
Text model20% offPrice droptext-to-textimage-to-textspeech-to-textvideo-to-text

Gemini 3 Flash is a lightweight, efficient multimodal large-scale model from Google tailored for real-world scenarios that require fast responses and low latency.

From
$0.63$0.5/1M tokens
View model
G

gemini-3.1-flash

Google
Coming soon
G
Text Generation

gemini-3.1-flash

gemini-3.1-flash
Text modelComing soontext-to-text

gemini-3.1-flash coming soon

From
Coming soon
View model
G

Gemini 3.5 Pro

Google
Coming soon
G
Text Generation

Gemini 3.5 Pro

gemini-3.5-pro
Text modelComing soontext-to-text

coming soon

From
Coming soon
View model
G

Gemini 3.5 Flash Lite

Google
G
Text Generation

Gemini 3.5 Flash Lite

gemini-3.5-flash-lite
Text model40% offPrice droptext-to-text

Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for high-throughput, low-cost execution for subagent tasks and document parsing. The model supports text, image, video, audio, and PDF inputs, and is designed for high-volume agentic workflows, simple data extraction, and applications where latency and API cost are the primary constraints

From
$0.50$0.3/1M tokens
View model
G

Veo 4

Google
Coming soon
G
Video Generation

Veo 4

veo4
Video modelComing soontext-to-video

coming soon

From
Coming soon
View model