
GPT-6.1 Sol
Explore the GPT-6.1 Sol API.
Choose your path
Choose your path
Search and compare text, image, video and audio models from leading providers. Review capabilities and pricing, then integrate with one API key.

Explore the GPT-6.1 Sol API.

GPT-6 Sol (gpt-6-sol) is an OpenAI GPT-6 reasoning model designed for complex coding, agentic workflows, and demanding knowledge work. OpenAI describes it as a model for complex coding and agentic workflows, with adaptive reasoning, long-context processing, image input, and access to a broad set of tools through the Responses API.

GPT-6 Luna (gpt-6-luna) is OpenAI's efficient GPT-6 model for focused, high-volume workloads. OpenAI positions Luna below GPT-6 Sol in the GPT-6 family, emphasizing efficiency, repeatability, and cost-sensitive production use.

GPT-6 Astra, the flagship model for complex reasoning and coding. Choose GPT-5.6 Terra to balance intelligence and cost,

GPT Image 2 is openai state-of-the-art image generation model for fast, high-quality image generation and editing. It supports flexible image sizes and high-fidelity image inputs.

Explore the gpt-oss-20b(free) API.

GPT-5.4 Mini is a lightweight and efficient AI model optimized for speed and everyday productivity. It provides reliable conversational capabilities, content generation, and task assistance while maintaining low latency and resource usage.

Explore the GPT Image 2.5(sunburst) API.
GPT-5 Nano is an artificial intelligence model provided by OpenAI.
GPT-5 mini is OpenAI’s cost- and latency-optimized member of the GPT-5 family, intended to deliver much of GPT-5’s multimodal and instruction-following strengths at substantially lower cost for large-scale production use. It targets environments where throughput, predictable per-token pricing, and fast responses are the primary constraints while still providing strong general-purpose capabilities.
GPT-Realtime-2 is our most capable realtime voice model. It supports speech-to-speech interactions with configurable reasoning effort, stronger instruction following, and more reliable tool use for complex voice-agent workflows
coming soon

GPT-5.4 Nano is an ultra-lightweight AI model built for maximum speed and efficiency. It is optimized for simple tasks, real-time interactions, and large-scale deployments where low latency and minimal resource consumption are essential.
gpt-4o-image generate images as output, optionally using images as input

Sora 2 Pro is our most advanced and powerful media generation model, capable of generating videos with synchronized Audio. It can create detailed, dynamic video clips from natural language or images.

GPT-5.4 Pro is a high-performance AI model designed for professional and business applications. It offers strong reasoning, reliable accuracy, and efficient execution across tasks such as content creation, coding, research, and data analysis.

GPT-5.2 is a multi-flavored model suite (Instant, Thinking, Pro) engineered for better long-context understanding, stronger coding and tool use, and materially higher performance on professional “knowledge-work” benchmarks.

Explore the GPT Image 2.5 Flare API.
OpenAI o3‑pro is a “pro” variant of the o3 reasoning model engineered to think longer and deliver the most dependable responses by employing private chain‑of‑thought reinforcement learning and setting new state‑of‑the‑art benchmarks across domains like science, programming, and business—while autonomously integrating tools such as web search, file analysis, Python execution, and visual reasoning within API.

The best voice model for audio in, audio out.

OpenAI Text-to-Speech

Super powerful video generation model, with sound effects, supports chat format.

GPT-5.5 Pro combines state-of-the-art intelligence, precision, and efficiency to tackle sophisticated challenges. From software development and data analysis to research and decision support, it delivers expert-level assistance with speed and consistency.

Speech to text, creating translations
.jpeg&w=3840&q=75)
Model 5.5 is a next-generation AI model designed for stronger reasoning, faster responses, and improved accuracy across a wide range of tasks. It excels at understanding complex instructions, generating high-quality content, and assisting with coding, analysis, and problem-solving.

GPT-5.3-Codex is optimized for agentic coding tasks in Codex or similar environments. GPT-5.3-Codex supports low, medium, high, and xhigh reasoning effort settings.
Explore the tts-1 API.
GPT-4o mini TTS is a neural text-to-speech model designed for natural, low-latency voice generation in user-facing applications. It converts text to natural-sounding speech with selectable voices, multi-format output, and streaming synthesis for responsive experiences. Typical uses include voice assistants, IVR and contact flows, product read-aloud, and media narration. Technical highlights include API-based streaming and export to common audio formats such as MP3 and WAV.

GPT Image 2 ALL is a comprehensive image generation model designed to handle a wide range of creative and professional visual tasks. It combines high-quality image creation, advanced prompt understanding, and versatile style support to deliver exceptional results across diverse use cases.
Explore the o1-pro-2025-03-19 API.

GPT-Image-1.5 is OpenAI’s image model in the GPT Image family . It is a natively multimodal GPT model designed to generate images from text prompts and to perform high-fidelity edits of input images while following user instructions closely.
GPT-4o Transcribe is an audio-to-text model for multilingual, low-latency speech recognition. It supports real-time streaming and batch transcription from common audio formats with punctuation and sentence segmentation. Typical uses include live captions, voice assistant input, meeting notes, and media or call recording transcription. Technical highlights include audio modality support, long-form processing, and APIs suited for interactive and server-side workflows.