Claude 3.5 Haiku
by Anthropic
Fastest model that surpasses Claude 3 Opus intelligence on many benchmarks while maintaining speed and cost efficiency. Achieves 40.6% on SWE-bench Verified despite being optimized for rapid responses and high-volume applications.
- Superior intelligence outperforming Claude 3 Opus on numerous benchmarks while maintaining speed
- Exceptional coding capabilities for a compact model with 40.6% SWE-bench Verified performance
- Near-instant responses optimized for real-time applications and user-facing products
- Enhanced tool use and instruction following compared to predecessors
- Cost-effective solution ideal for high-volume data processing and automated workflows
JSON Mode
Output responses in valid JSON format.
Web Search
Search the web for current info.
Model Information
Supported Formats
Anthropic's most capable widely released model, for the most demanding reasoning and long-horizon agentic work. Thinking is always on with a 1M token context window (also the maximum) and 128K max output. Priced above the Opus tier for the highest capability ceiling in the Claude lineup.
Anthropic's successor to Opus 4.8 for complex agentic coding and enterprise work. Strongest on deep reasoning, agentic and long-horizon work, and test-time compute scaling, at half the cost of Claude Fable 5. Thinking is on by default with a 1M token context window and 128K max output.
Best combination of speed and intelligence in the Sonnet tier, delivering near-Opus quality on coding and agentic work. Successor to Sonnet 4.6 with adaptive thinking on by default, a 1M token context window, and 128K max output.