GPT-5.1 (Medium)
gpt-5.1-mediumGPT-5.1 is the full-capability successor to GPT-5, offering stronger general reasoning, better instruction following, and a more natural conversational style. It uses adaptive reasoning to stay fast on simple questions while thinking more deeply on complex tasks, producing clearer, more grounded explanations. It shows steady improvements across math, coding, and structured analysis, with more coherent long-form output and more reliable tool use.
- Context
- 400K tokens
- Endpoint
Pricing
Quick Start
Select an endpoint and copy a working example for this model.
from openai import OpenAI client = OpenAI( api_key="YOUR_API_KEY", base_url="https://api.apertis.ai/v1") response = client.chat.completions.create( model="gpt-5.1-medium", messages=[ {"role": "user", "content": "Hello!"} ], max_tokens=1024, temperature=0.7) print(response.choices[0].message.content) # Optional: Enable context compression to reduce token usage# response = client.chat.completions.create(# model="gpt-5.1-medium",# messages=[{"role": "user", "content": "Hello!"}],# extra_body={"compression": {"enabled": True, "model": "gpt-4.1-mini"}}# )Supported Parameters
API docsmodelmessagesmax_tokenstemperaturetop_pstreamtoolsreasoning_effortstream_optionsthinkingextra_bodyCursor IDE Model IDs
Use these namespaced identifiers in Cursor IDE to avoid conflicts with built-in models.
Compare with Other Models
See how this model compares to others from the same provider.
GPT-6 Astra
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end professional work, designed for advanced analysis, software engineering, deep research, scientific tasks, and document creation. It is particularly strong in long-horizon agentic workflows, including tasks that require sustained reasoning, tool orchestration, and computer and browser use, making it well suited for complex autonomous workflows and production-grade knowledge work.
- Context
- 1M
- Input
- $10.00/M
- Output
- $50.00/M
GPT-6 Astra Pro
GPT-6 Astra Pro uses the same underlying model as GPT-6 Astra, but runs with reasoning.mode set to pro for higher-quality responses on complex tasks. Optimized for deeper reasoning, greater accuracy, and more reliable multi-step execution, it is well suited for demanding coding, analysis, and agentic workflows where solution quality takes priority over speed and cost.
- Context
- 1M
- Input
- $10.00/M
- Output
- $50.00/M
GPT-4o Mini TTS
GPT-4o Mini TTS is OpenAI's cost-efficient text-to-speech model, designed to convert text into natural-sounding audio output. It supports a variety of voices and tones, enabling flexible and expressive speech generation. Optimized for scalability and low cost, it is well suited for real-time voice applications, content narration, and high-volume audio generation workflows.
- Context
- 4.1K
- Input
- $0.30/M
- Output
- $0/M
Whisper Large V3 Turbo
Whisper Large V3 Turbo is an optimized version of OpenAI's Whisper Large V3 speech recognition model, designed for high-speed and cost-efficient transcription. It supports 99+ languages and accepts common audio formats including mp3, mp4, wav, webm, flac, and ogg. With a ~12% word error rate and real-time speed factors up to 216×, it delivers fast, scalable performance for latency-sensitive and high-throughput transcription workloads, making it ideal for real-time and large-scale speech processing applications.
- Context
- N/A
- Input
- $3.33/M
- Output
- $0/M