Models
| Provider | Model | Input | Output | Context |
|---|---|---|---|---|
GPT-5.6 Luna Pro uses the same underlying model as GPT-5.6 Luna, but runs with reasoning.mode set to pro to deliver higher-quality responses on complex tasks. It is optimized for deeper reasoning, advanced coding, and multi-step agentic workflows, offering improved accuracy and solution quality while retaining the efficiency and scalability of the Luna tier. ChatJul 8, 2026 | Input$1/1M tokens | Output$6/1M tokens | Context1M | |
GPT-5.6 Luna is the fast, cost-efficient model in OpenAI's GPT-5.6 series, optimized for high-volume, latency-sensitive workloads. It delivers capable reasoning at an affordable price point, making it ideal for chat applications, classification, and lightweight agentic workflows. Designed for scalable production deployments, GPT-5.6 Luna balances speed, cost, and reliability, providing efficient performance for real-time applications and large-scale automation tasks. ChatJul 8, 2026 | Input$1/1M tokens | Output$6/1M tokens | Context1M | |
GPT-5.6 Terra Pro uses the same underlying model as GPT-5.6 Terra, but runs with reasoning.mode set to pro to deliver higher-quality responses on complex tasks. Optimized for deeper reasoning and greater reliability, it is well suited for advanced coding, multi-step reasoning, and agentic workflows where improved accuracy and solution quality are more important than maximizing speed or minimizing cost. ChatJul 8, 2026 | Input$2.50/1M tokens | Output$15/1M tokens | Context1M | |
GPT-5.6 Terra is the balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is designed for everyday coding, reasoning, and agentic workflows, delivering strong performance while balancing capability and cost. Offering near-flagship quality at approximately half the cost of Sol, GPT-5.6 Terra is well suited for production applications that require reliable reasoning, software development, and scalable agent execution. ChatJul 8, 2026 | Input$2.50/1M tokens | Output$15/1M tokens | Context1M | |
GPT-5.6 Sol Pro uses the same underlying model as GPT-5.6 Sol, but runs with reasoning.mode set to pro for higher-quality responses on complex tasks. Optimized for deeper reasoning and more reliable execution, it is particularly well suited for advanced coding, long-horizon problem solving, and agentic workflows where accuracy and solution quality take priority over speed and cost. ChatJul 8, 2026 | Input$5/1M tokens | Output$30/1M tokens | Context1M | |
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series, designed for complex reasoning, coding, and agentic workflows. It delivers strong performance on multi-step software engineering tasks, command-line workflows, and long-horizon problem solving, making it well suited for advanced development and autonomous execution. Optimized for high-reliability reasoning and end-to-end task completion, GPT-5.6 Sol excels in coding, tool-driven automation, and large-scale engineering workflows that require sustained context and precise execution. ChatJul 8, 2026 | Input$5/1M tokens | Output$30/1M tokens | Context1M | |
GPT-4o Mini TTS is OpenAI's cost-efficient text-to-speech model, designed to convert text into natural-sounding audio output. It supports a variety of voices and tones, enabling flexible and expressive speech generation. Optimized for scalability and low cost, it is well suited for real-time voice applications, content narration, and high-volume audio generation workflows. VoiceApr 30, 2026 | Input$0.3/1M tokens | Output$0/1M tokens | Context4K | |
GPT-4o Mini Transcribe is a smaller, cost-efficient speech-to-text model built on GPT-4o Mini's audio capabilities. It is designed for high-volume transcription workloads, delivering reliable performance with lower cost and latency. Priced per token (input and output), it provides transparent, fine-grained billing, making it well suited for scalable transcription pipelines, real-time applications, and cost-sensitive deployments. VoiceApr 30, 2026 | Input$0.625/1M tokens | Output$0.625/1M tokens | Context128K | |
GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o's audio capabilities. It delivers accurate transcription with strong language understanding, making it suitable for a wide range of audio processing tasks. Priced per token (input and output), it offers transparent, fine-grained billing, making it well suited for workflows that require scalable transcription, integration with LLM pipelines, and cost-aware processing. VoiceApr 30, 2026 | Input$1.25/1M tokens | Output$0/1M tokens | Context128K | |
Whisper Large V3 Turbo is an optimized version of OpenAI's Whisper Large V3 speech recognition model, designed for high-speed and cost-efficient transcription. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm, flac, and ogg. With a ~12% word error rate and real-time speed factors up to 216×, it delivers fast, scalable performance for latency-sensitive and high-throughput transcription workloads, making it ideal for real-time and large-scale speech processing applications. VoiceApr 30, 2026 | Input$3.33/1M tokens | Output$0/1M tokens | - | |
Whisper Large V3 is OpenAI's advanced open-source automatic speech recognition (ASR) model, supporting both audio transcription and translation across 99+ languages. It accepts common audio formats including mp3, mp4, wav, webm, flac, and ogg, and delivers strong performance in noisy, real-world conditions. With 1.55B parameters and a low 10.3% word error rate, it provides accurate, multilingual transcription with support for word- and segment-level timestamps, making it well suited for high-quality, noise-robust speech processing applications. VoiceApr 30, 2026 | Input$9.25/1M tokens | Output$0/1M tokens | - | |
Whisper (whisper-1) is OpenAI's open-source automatic speech recognition (ASR) model, designed for audio transcription and translation. It supports 50+ languages and processes audio files up to 25 MB, accepting formats such as mp3, mp4, wav, and webm. Optimized for reliable speech-to-text conversion across diverse audio inputs, Whisper is priced per minute of audio, billed to the nearest second, making it well suited for transcription, localization, and voice-driven applications. VoiceApr 30, 2026 | Input$75/1M tokens | Output$75/1M tokens | - | |
GPT-5.5 is OpenAI's frontier model for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on challenging tasks. It supports text and image inputs and features a 1M+ token context window (≈922K input, 128K output) for large-scale, high-context workflows. Designed for advanced applications, GPT-5.5 excels in reasoning, coding, and multimodal workflows, enabling efficient execution of complex, multi-step tasks within a single system. ChatApr 23, 2026 | Input$4/1M tokens | Output$24/1M tokens | Context1.1M | |
GPT-5.5 Pro is OpenAI's high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It supports text and image inputs and features a 1M+ token context window (≈922K input, 128K output) for handling large-scale, long-context tasks. Designed for long-horizon problem solving, agentic coding, and precise multi-step execution, GPT-5.5 Pro delivers strong reliability and performance across advanced engineering, research, and complex workflow scenarios. ChatApr 23, 2026 | Input$30/1M tokens | Output$180/1M tokens | Context1.1M | |
GPT Image 2 combines OpenAI's GPT-5.4 with advanced image generation capabilities from GPT Image 2, enabling fully integrated multimodal workflows. It allows users to seamlessly transition between reasoning, coding, and visual generation within a single interaction, making it well suited for creative, development, and agent-driven applications that require both intelligence and visual output. ImageApr 20, 2026 | Input$0/1M tokens | Output$0/1M tokens | Context272K | |
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume workloads. It supports text and image inputs and is designed for low-latency tasks such as classification, data extraction, ranking, and sub-agent execution. Prioritizing responsiveness and efficiency over deep reasoning, GPT-5.4 nano is ideal for real-time systems, background processing, and distributed agent pipelines where minimizing cost and latency is essential. ChatMar 16, 2026 | Input$0.2/1M tokens | Output$1.25/1M tokens | Context400K | |
GPT-5.4 mini brings the core capabilities of GPT-5.4 into a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs and delivers strong performance across reasoning, coding, and tool use, while reducing latency and cost for large-scale deployments. Designed for production environments, GPT-5.4 mini balances capability and efficiency, making it well suited for chat applications, coding assistants, and scalable agent workflows. It provides reliable instruction following, solid multi-step reasoning, and consistent performance across diverse tasks with improved cost efficiency. ChatMar 16, 2026 | Input$0.75/1M tokens | Output$4.50/1M tokens | Context400K | |
GPT-5.4 Pro is OpenAI's most advanced model, built on the unified GPT-5.4 architecture with enhanced reasoning capabilities for complex and high-stakes tasks. It supports text and image inputs and features a 1M+ token context window (≈922K input, 128K output) for handling large-scale workflows and long-context analysis. Optimized for step-by-step reasoning, instruction following, and accuracy, GPT-5.4 Pro excels in agentic coding, long-context problem solving, and complex multi-step workflows, making it well suited for advanced engineering, research, and high-reliability applications. ChatMar 4, 2026 | Input$24/1M tokens | Output$144/1M tokens | Context1.1M | |
GPT-5.4 is OpenAI's latest frontier model, unifying the GPT and Codex lines into a single system designed for both general intelligence and advanced software engineering workflows. It supports text and image inputs and features a 1M+ token context window (≈922K input, 128K output), enabling high-context reasoning, coding, and multimodal analysis within a single workflow. The model delivers improved performance in coding, document understanding, tool use, and instruction following, and is designed as a strong default for complex tasks. It can generate production-quality code, synthesize information across large datasets, and execute multi-step workflows with fewer iterations and greater token efficiency. ChatMar 4, 2026 | Input$2/1M tokens | Output$12/1M tokens | Context1.1M | |
GPT-5.3 Chat is an updated version of ChatGPT's most widely used conversational model, designed to make everyday interactions smoother, more accurate, and more helpful. It improves contextual understanding and response quality while reducing unnecessary refusals, excessive caveats, and overly cautious phrasing that can disrupt conversational flow. Optimized for general-purpose dialogue, GPT-5.3 Chat delivers more natural, reliable responses across a wide range of everyday tasks and discussions. ChatMar 2, 2026 | Input$0.875/1M tokens | Output$7/1M tokens | Context128K | |
GPT-Codex-5.3 is OpenAI's most advanced agentic coding model, designed for software engineering workflows that extend beyond single prompts into long-running, tool-driven execution. It combines the frontier coding performance of earlier Codex models with stronger reasoning and professional knowledge capabilities, enabling reliable handling of complex refactors, multi-step debugging, research-driven development, and autonomous task execution. Optimized for developer productivity, GPT-Codex-5.3 supports interactive collaboration during execution, allowing users to steer tasks in real time without losing context. With improved agentic reliability, faster inference, and stronger performance on long-horizon engineering tasks, it is well suited for coding agents, IDE and CLI workflows, and end-to-end software development pipelines where persistence, tool use, and execution continuity are critical. ChatFeb 5, 2026 | Input$1.75/1M tokens | Output$14/1M tokens | Context400K | |
GPT-5.2-Codex is OpenAI's most advanced agentic coding model yet, built for complex, real-world software engineering and defensive cybersecurity. It’s a version of GPT-5.2 further optimized for Codex, with improvements in long-horizon coding tasks (like refactors and migrations), better handling of long contexts, stronger performance on large code changes, enhanced Windows support, and significantly stronger cybersecurity capabilities. ChatDec 17, 2025 | Input$0.5/1M tokens | Output$4/1M tokens | Context196K | |
gpt-image-1.5 is an upgraded OpenAI image model that produces more detailed, realistic visuals with better composition and prompt fidelity. It improves editing, variation, and style control over earlier versions, making it well suited for creative design, illustration, marketing assets, and visual prototyping. ImageDec 15, 2025 | Input$0/1M tokens | Output$0/1M tokens | Context32K | |
GPT-5.2 Chat (Instant) is the fast, lightweight version of the 5.2 family, tuned for low-latency conversation while keeping strong general intelligence. It uses adaptive reasoning to think more on tough questions, boosting accuracy in math, coding, and multi-step tasks without slowing normal chats. It’s friendlier by default, follows instructions well, and is ideal for high-throughput interactive use where speed and consistency matter more than deep deliberation. ChatDec 9, 2025 | Input$1/1M tokens | Output$8/1M tokens | Context128K | |
GPT-5.2 Pro is OpenAI's most advanced model, with big gains over GPT-5 Pro in long-context reasoning and agentic coding. It’s built for complex, high-stakes tasks that need careful step-by-step thinking and precise instruction following. It supports test-time routing and intent cues like “think hard,” while reducing hallucinations and sycophancy and improving results in coding, writing, and health-related workflows. ChatDec 9, 2025 | Input$16/1M tokens | Output$128/1M tokens | Context400K | |
GPT-5.2 is the newest frontier model in the GPT-5 lineup, with stronger agent abilities and long-context performance than GPT-5.1. It uses adaptive reasoning to stay fast on simple prompts while thinking more deeply on hard ones, and delivers steady gains across math, coding, science, and tool use — with more coherent long-form output and more reliable tooling. ChatDec 9, 2025 | Input$1/1M tokens | Output$8/1M tokens | Context400K | |
GPT-5.1 Codex Max is OpenAI's advanced agentic coding model, built for long-running, high-context development work. Using an upgraded 5.1 reasoning stack and training on real engineering workflows, it delivers faster performance, stronger reasoning, and better token efficiency across the full software lifecycle. ChatDec 3, 2025 | Input$0.375/1M tokens | Output$3/1M tokens | Context400K | |
GPT-5.1 Codex is a coding-focused version of GPT-5.1 designed for both interactive development and long autonomous engineering tasks. It can build projects, add features, debug, refactor, and review code with higher steerability and cleaner outputs than GPT-5.1. It integrates with developer tools (CLI, IDEs, GitHub, cloud), supports adjustable reasoning effort, handles images/screenshots for UI work, and uses tools for search and environment setup — making it purpose-built for agentic coding workflows. ChatNov 12, 2025 | Input$0.375/1M tokens | Output$3/1M tokens | Context400K | |
GPT-5.1 Codex is a coding-focused version of GPT-5.1 designed for both interactive development and long autonomous engineering tasks. It can build projects, add features, debug, refactor, and review code with higher steerability and cleaner outputs than GPT-5.1. It integrates with developer tools (CLI, IDEs, GitHub, cloud), supports adjustable reasoning effort, handles images/screenshots for UI work, and uses tools for search and environment setup — making it purpose-built for agentic coding workflows. ChatNov 12, 2025 | Input$0.125/1M tokens | Output$1/1M tokens | Context400K | |
GPT-5.1 Codex is a coding-focused version of GPT-5.1 designed for both interactive development and long autonomous engineering tasks. It can build projects, add features, debug, refactor, and review code with higher steerability and cleaner outputs than GPT-5.1. It integrates with developer tools (CLI, IDEs, GitHub, cloud), supports adjustable reasoning effort, handles images/screenshots for UI work, and uses tools for search and environment setup — making it purpose-built for agentic coding workflows. ChatNov 12, 2025 | Input$0.375/1M tokens | Output$3/1M tokens | Context400K | |
GPT-5.1 is the full-capability successor to GPT-5, offering stronger general reasoning, better instruction following, and a more natural conversational style. It uses adaptive reasoning to stay fast on simple questions while thinking more deeply on complex tasks, producing clearer, more grounded explanations. It shows steady improvements across math, coding, and structured analysis, with more coherent long-form output and more reliable tool use. ChatNov 12, 2025 | Input$0.875/1M tokens | Output$7/1M tokens | Context400K | |
GPT-5.1 is the full-capability successor to GPT-5, offering stronger general reasoning, better instruction following, and a more natural conversational style. It uses adaptive reasoning to stay fast on simple questions while thinking more deeply on complex tasks, producing clearer, more grounded explanations. It shows steady improvements across math, coding, and structured analysis, with more coherent long-form output and more reliable tool use. ChatNov 12, 2025 | Input$0.375/1M tokens | Output$3/1M tokens | Context400K | |
GPT-5.1 Codex is a coding-focused version of GPT-5.1 designed for both interactive development and long autonomous engineering tasks. It can build projects, add features, debug, refactor, and review code with higher steerability and cleaner outputs than GPT-5.1. It integrates with developer tools (CLI, IDEs, GitHub, cloud), supports adjustable reasoning effort, handles images/screenshots for UI work, and uses tools for search and environment setup — making it purpose-built for agentic coding workflows. ChatNov 12, 2025 | Input$0.375/1M tokens | Output$3/1M tokens | Context400K | |
GPT-5.1 is the full-capability successor to GPT-5, offering stronger general reasoning, better instruction following, and a more natural conversational style. It uses adaptive reasoning to stay fast on simple questions while thinking more deeply on complex tasks, producing clearer, more grounded explanations. It shows steady improvements across math, coding, and structured analysis, with more coherent long-form output and more reliable tool use. ChatNov 12, 2025 | Input$0.375/1M tokens | Output$3/1M tokens | Context400K | |
GPT-5.1 is the full-capability successor to GPT-5, offering stronger general reasoning, better instruction following, and a more natural conversational style. It uses adaptive reasoning to stay fast on simple questions while thinking more deeply on complex tasks, producing clearer, more grounded explanations. It shows steady improvements across math, coding, and structured analysis, with more coherent long-form output and more reliable tool use. ChatNov 12, 2025 | Input$0.625/1M tokens | Output$5/1M tokens | Context128K | |
GPT-5.1 is the full-capability successor to GPT-5, offering stronger general reasoning, better instruction following, and a more natural conversational style. It uses adaptive reasoning to stay fast on simple questions while thinking more deeply on complex tasks, producing clearer, more grounded explanations. It shows steady improvements across math, coding, and structured analysis, with more coherent long-form output and more reliable tool use. ChatNov 12, 2025 | Input$0.375/1M tokens | Output$3/1M tokens | Context400K | |
GPT-5.1 is the full-capability successor to GPT-5, offering stronger general reasoning, better instruction following, and a more natural conversational style. It uses adaptive reasoning to stay fast on simple questions while thinking more deeply on complex tasks, producing clearer, more grounded explanations. It shows steady improvements across math, coding, and structured analysis, with more coherent long-form output and more reliable tool use. ChatNov 12, 2025 | Input$0.875/1M tokens | Output$7/1M tokens | Context400K | |
gpt-oss-safeguard-20b is an OpenAI safety-focused model built on gpt-oss-20b. It's an open-weight 21B MoE system optimized for low-latency safety work such as content classification, LLM moderation, and trust-and-safety labeling, with guidance available in its official user guide. ChatOct 28, 2025 | Input$0.09/1M tokens | Output$0.36/1M tokens | Context131K | |
GPT-5 is OpenAI's most advanced model, built for complex, high-stakes tasks that require careful step-by-step reasoning and precise instruction following. It improves code quality, writing, and reliability, supports test-time routing and intent cues like “think hard,” and reduces hallucinations and sycophancy across demanding workloads. ChatOct 13, 2025 | Input$2.50/1M tokens | Output$10/1M tokens | Context400K | |
GPT-5 is OpenAI's most advanced model, built for complex, high-stakes tasks that require careful step-by-step reasoning and precise instruction following. It improves code quality, writing, and reliability, supports test-time routing and intent cues like “think hard,” and reduces hallucinations and sycophancy across demanding workloads. ChatOct 13, 2025 | Input$2.50/1M tokens | Output$10/1M tokens | Context400K | |
o4-mini-deep-research is a faster, lower-cost version of OpenAI's deep-research model, designed for complex, multi-step investigations. It automatically relies on web_search for information gathering, which always adds extra usage cost. ChatOct 9, 2025 | Input$2/1M tokens | Output$8/1M tokens | Context200K | |
o3-deep-research is OpenAI's advanced research model, built for complex, multi-step investigation and analysis. It automatically performs web searches to gather and synthesize information — but this always incurs additional cost since web_search is used by default. ChatOct 9, 2025 | Input$10/1M tokens | Output$40/1M tokens | Context200K | |
GPT-5 Pro is OpenAI's top model, optimized for complex, high-stakes tasks that require careful step-by-step reasoning and precise instruction following. It delivers stronger code quality, clearer writing, and better factual reliability, with support for test-time routing and intent cues like “think hard about this.” It also reduces hallucinations and sycophancy while improving performance across coding, writing, and health-related workloads. ChatOct 5, 2025 | Input$7.50/1M tokens | Output$60/1M tokens | Context400K | |
Sora 2 is the higher-quality version of OpenAI's Sora 2 text-to-video and audio generation model, designed to produce more realistic, controllable, and detailed AI-generated videos with synchronized audio and advanced world simulation capabilities. It builds on Sora 2's breakthrough in video realism and physics-aware generation, offering enhanced visual fidelity and extended features for creative and professional use. Early access has been available to ChatGPT Pro subscribers via sora.com, with wider availability expected after the invite/beta rollout. VideoSep 29, 2025 | Input$0.1125/request | Output-/request | - | |
Sora 2 is the higher-quality version of OpenAI's Sora 2 text-to-video and audio generation model, designed to produce more realistic, controllable, and detailed AI-generated videos with synchronized audio and advanced world simulation capabilities. It builds on Sora 2's breakthrough in video realism and physics-aware generation, offering enhanced visual fidelity and extended features for creative and professional use. Early access has been available to ChatGPT Pro subscribers via sora.com, with wider availability expected after the invite/beta rollout. VideoSep 29, 2025 | Input$0.1125/request | Output-/request | - | |
Sora 2 Pro is the higher-quality version of OpenAI's Sora 2 text-to-video and audio generation model, designed to produce more realistic, controllable, and detailed AI-generated videos with synchronized audio and advanced world simulation capabilities. It builds on Sora 2's breakthrough in video realism and physics-aware generation, offering enhanced visual fidelity and extended features for creative and professional use. Early access has been available to ChatGPT Pro subscribers via sora.com, with wider availability expected after the invite/beta rollout. VideoSep 29, 2025 | Input$1.26/request | Output-/request | - | |
Sora 2 is the higher-quality version of OpenAI's Sora 2 text-to-video and audio generation model, designed to produce more realistic, controllable, and detailed AI-generated videos with synchronized audio and advanced world simulation capabilities. It builds on Sora 2's breakthrough in video realism and physics-aware generation, offering enhanced visual fidelity and extended features for creative and professional use. Early access has been available to ChatGPT Pro subscribers via sora.com, with wider availability expected after the invite/beta rollout. VideoSep 29, 2025 | Input$0.105/request | Output-/request | - | |
Sora 2 is the higher-quality version of OpenAI's Sora 2 text-to-video and audio generation model, designed to produce more realistic, controllable, and detailed AI-generated videos with synchronized audio and advanced world simulation capabilities. It builds on Sora 2's breakthrough in video realism and physics-aware generation, offering enhanced visual fidelity and extended features for creative and professional use. Early access has been available to ChatGPT Pro subscribers via sora.com, with wider availability expected after the invite/beta rollout. VideoSep 29, 2025 | Input$0.14/request | Output-/request | - | |
GPT-5 Codex (Medium) is a coding-focused version of GPT-5 built for both interactive development and long autonomous engineering tasks. It can create projects, add features, debug, refactor, and review code, producing cleaner and more controllable outputs than GPT-5. It integrates with developer tools (CLI, IDEs, GitHub, cloud), supports adjustable reasoning effort, handles multimodal inputs, and uses tools for search and environment setup — making it purpose-built for agentic coding workflows. ChatSep 22, 2025 | Input$0.625/1M tokens | Output$5/1M tokens | Context400K | |
GPT-5 Codex (High) is a coding-focused version of GPT-5 built for both interactive development and long autonomous engineering tasks. It can create projects, add features, debug, refactor, and review code, producing cleaner and more controllable outputs than GPT-5. It integrates with developer tools (CLI, IDEs, GitHub, cloud), supports adjustable reasoning effort, handles multimodal inputs, and uses tools for search and environment setup — making it purpose-built for agentic coding workflows. ChatSep 22, 2025 | Input$0.625/1M tokens | Output$5/1M tokens | Context400K |