AnthropicProprietaryCodingReasoningVision
Claude Sonnet
Anthropic’s balanced Claude tier — strong quality at lower latency and cost than Opus, widely used for production agents and coding.
Vision · Tool calling · Thinking · MCP · Coding
Last reviewed: 24 July 2026
Overview
Anthropic’s balanced Claude tier — strong quality at lower latency and cost than Opus, widely used for production agents and coding.
Capabilities
- Vision: Yes
- Audio: No
- Tool calling: Yes
- Thinking: Yes
- MCP: Yes
- Coding: Yes
- Structured output: Yes
Technical specifications
- Provider
- Anthropic
- License
- Proprietary
- Context window
- 200K
- Parameters
- Undisclosed
- Release
- 2025
- Modalities
- Text, Image
- Vision
- Yes
- Audio
- No
- Tool calling
- Yes
- Thinking
- Yes
- MCP
- Yes
- Open weights
- No
- API
- Yes
- Pricing (input)
- API tiered (see Anthropic pricing)
- Pricing (output)
- API tiered (see Anthropic pricing)
Supported modalities
Text · Image
Context window
200K (200,000 tokens)
Pricing
Input: API tiered (see Anthropic pricing)
Output: API tiered (see Anthropic pricing)
Availability
API: Yes
Chat UI: Yes
Open weights: No
Use cases
- Production chat and support agents
- IDE coding assistants
- RAG answer generation
- Structured extraction pipelines
Strengths
- Strong quality / cost balance
- Reliable tool use for agents
- Widely available across clouds
Limitations
- Not the absolute peak of Claude capability (vs Opus)
- Closed weights
Related guides
Related benchmarks
Related research
Related GitHub
Related tools
Related rankings
Related comparisons
Companies
Explore more models
- Claude OpusAnthropicAnthropic’s highest-capability Claude tier for deep reasoning, long-context analysis, coding, and careful instruction following.
- Gemini 2.5 ProGoogleGoogle’s Pro-class Gemini model for advanced reasoning and native multimodal workloads across text, images, audio, and long context.
- GPT-5OpenAIOpenAI’s flagship general-purpose model for complex reasoning, coding, multimodal understanding, and agentic tool use.
- Qwen3AlibabaAlibaba’s Qwen3 generation — strong multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.
- Llama 4MetaMeta’s Llama 4 family — open-weight multimodal models designed for research and commercial use under Meta’s community license.
- GrokxAIxAI’s Grok model family — real-time oriented assistants with strong coding and reasoning, available via xAI API and consumer surfaces.