Anthropic: Claude 3.5 Haiku
🇺🇸 Anthropic · Claude 3.5
Overview
Anthropic: Claude 3.5 Haiku is a large language model API from Anthropic, part of its Claude 3.5 model family. Priced at $0.800 per million input tokens and $4.00 per million output tokens, it occupies the mid-range, balancing capability against running cost. Output tokens cost about 5× as much as input, so prompt-heavy workloads run noticeably cheaper than generation-heavy ones. A large 200K-token context window (≈300 pages of text) lets it take in whole books, large codebases, or lengthy transcripts in a single call. Beyond plain text it also accepts Image input, so it can be applied to multimodal tasks rather than text alone. On Artificial Analysis's Intelligence Index it scores 9 (F grade), a useful proxy for its general reasoning strength relative to the other models tracked here. All prices on this page reflect OpenRouter's routed rates and are re-synced automatically every day; confirm against the provider's official pricing before committing to production.
| Dimension | Unit | Price (USD) |
|---|---|---|
| Input | per 1M tokens | $0.800 |
| Output | per 1M tokens | $4.00 |
| Cached Input | per 1M tokens | $0.080 |
| Cache Write | per 1M tokens | $1.00 |
| Web Search | per search | $0.010 |
- Provider
- Anthropic
- Model Family
- Claude 3.5
- Version String
- anthropic/claude-3.5-haiku
- Status
- Active
- Modality
- Text, Image
- Context Window
- 200,000 tokens
- Output Limit
- 8,192 tokens
Index Metrics
Cross-domain capability indexes evaluated by Artificial Analysis — Artificial Analysis
Benchmark Scores
Data source: Artificial Analysis
Use Case Analysis
Model strengths, weaknesses, and ideal use cases based on benchmarks and official documentation
Strength
- Fastest Response Times Fastest comparative latency among all Claude models. TTFT 0.36s. Highest throughput 52.54 tok/s. Official
- Exceptional Cost Efficiency Priced at $1/$5 per MTok, 5x cheaper than Sonnet 4.5. Official
Best For
- Real-time User-facing Apps Fastest response times make it ideal for chatbots, autocomplete, and sub-second latency apps. Official
- High-Volume Low-Cost Tasks At $1/$5 pricing, ideal for content moderation, log classification, simple extraction at scale. Official
Weakness
- Limited Reasoning for Complex Tasks Struggles with complex reasoning, multi-step logical problems, and tasks exceeding 150 lines of code. Official
Not Recommended
- Complex Software Engineering For complex coding tasks, multi-file refactoring, or code exceeding 150 lines, use Sonnet or Opus. Official
Key Insights
Key data points from this page for quick reference and citation.
- Anthropic: Claude 3.5 Haiku Input price: $0.8/M tokens
- Anthropic: Claude 3.5 Haiku Output price: $4/M tokens
- Context window: 200,000 tokens
- Provider: Anthropic
- Model family: Claude 3.5
- Modalities: Text, Image
- Data source: OpenRouter, updated daily