Skip to content

Claude Sonnet 5

claude-sonnet-5

Sonnet 5 is Anthropic's most capable Sonnet-class model, delivering frontier-level performance across coding, agentic workflows, and professional knowledge tasks. It supports text, image, and file inputs, features a 1M-token context window, and offers adaptive thinking with configurable reasoning levels (low, medium, high, max, and x-high) to balance speed, cost, and reasoning depth. Optimized for complex coding, long-horizon agent execution, and professional workflows, Sonnet 5 combines strong reasoning, robust instruction following, and enhanced safety features, including an updated tokenizer and real-time cyber safeguards for high-risk dual-use scenarios.

Context
1M tokens
Endpoint

Get API KeyCompare

Pricing

Input$2.00 / 1M
Output$10.00 / 1M
Cache Write (5m)$2.50 / 1M
Cache Write (1h)$4.00 / 1M
Cache Read$0.20 / 1M
Web Search$0 / 1M

Quick Start

Select an endpoint and copy a working example for this model.

Endpoint
python
from openai import OpenAI client = OpenAI(    api_key="YOUR_API_KEY",    base_url="https://api.apertis.ai/v1") response = client.chat.completions.create(    model="claude-sonnet-5",    messages=[        {"role": "user", "content": "Hello!"}    ],    max_tokens=1024,    temperature=0.7) print(response.choices[0].message.content) # Optional: Enable context compression to reduce token usage# response = client.chat.completions.create(#     model="claude-sonnet-5",#     messages=[{"role": "user", "content": "Hello!"}],#     extra_body={"compression": {"enabled": True, "model": "gpt-4.1-mini"}}# )

Supported Parameters

API docs
Common7 params
modelmessagesmax_tokenstemperaturetop_pstreamtools
Extended4 params
reasoning_effortstream_optionsthinkingextra_body

Cursor IDE Model IDs

Use these namespaced identifiers in Cursor IDE to avoid conflicts with built-in models.

claude-sonnet-5

Compare with Other Models

See how this model compares to others from the same provider.

Claude Sonnet 5.5

Claude Sonnet 5.5 is Anthropic's Sonnet-class model for well-scoped everyday work, serving as a direct upgrade to Sonnet 5. It excels at feature development, bug fixing, and creating polished documents, presentations, and spreadsheets, while offering clearer writing and communication than its predecessor.

Context
1M
Input
$2.00/M
Output
$10.00/M

Claude Opus 5.5

Claude Opus 5.5 is Anthropic's flagship model for advanced reasoning, coding, and long-horizon agentic workflows, succeeding Opus 5. It excels at multi-step changes across large codebases, code review and bug detection, financial and scientific analysis, and understanding dense charts, diagrams, and screenshots, with stronger grounding when reporting figures and citing sources. Compared with Opus 5, it completes comparable tasks with fewer steps and lower token usage while providing clearer, more concise progress reporting. With adaptive thinking and configurable effort levels, Opus 5.5 can balance reasoning depth, latency, and cost, making it well suited for both demanding autonomous workflows and latency-sensitive professional tasks.

Context
1M
Input
$4.00/M
Output
$20.00/M

Claude Opus 4.6 (Thinking)

Opus 4.6 is Anthropic's most capable model for coding and long-running professional workflows, designed for agents that operate across entire workflows rather than single prompts. It demonstrates strong performance on large codebases, complex refactoring, and multi-step debugging, with improved contextual understanding, deeper problem decomposition, and higher reliability on challenging engineering tasks compared to earlier generations. Beyond software development, Opus 4.6 excels at sustained knowledge work, producing near production-ready documents, technical plans, and analyses in a single pass while maintaining coherence across long outputs and extended sessions. Its strength in persistence, judgment, and structured execution makes it well suited for technical design, migration planning, and end-to-end project execution.

Context
1M
Input
$4.00/M
Output
$20.00/M

Claude Opus 4.5

Claude Opus 4.5 is Anthropic's frontier reasoning model, built for complex engineering, agent workflows, and long computer-use tasks. It offers strong multimodal skills, better security against prompt injection, and flexible effort controls — including a Verbosity setting to trade speed vs. depth and token use. With advanced tool use, long-context handling, and support for coordinated multi-agent setups, it excels at research, debugging, multi-step planning, and UI/spreadsheet automation while improving reliability, alignment, and efficiency over earlier Opus versions.

Context
200K
Input
$4.00/M
Output
$20.00/M