Claude 3 Haiku
Claude 3 Haiku is a text generation model from Anthropic with a 200,000 token context window, available via Amazon Bedrock.
Compact text generation with large context
Claude 3 Haiku is a large language model developed by Anthropic and released in March 2024. It is the smallest model in the Claude 3 model family and is designed for tasks where response speed and efficiency matter, such as customer-facing applications, content moderation, and lightweight summarization. It supports a 200,000 token context window, allowing it to process long documents in a single request, and generates responses up to 4,096 tokens.
Claude 3 Haiku is available through Amazon Bedrock under the model identifier anthropic.claude-3-haiku-20240307-v1:0. Its status is listed as deprecated, meaning Anthropic and Amazon Bedrock have moved toward newer model versions for active use. Developers who built workflows around this model should evaluate migration to a current Claude release, as deprecated models may have limited long-term availability on the platform.
What Claude 3 Haiku supports
Long Context Processing
Handles up to 200,000 tokens of input in a single request, enabling analysis of lengthy documents, codebases, or conversation histories without chunking.
Text Generation
Generates coherent prose, summaries, and structured text responses up to 4,096 tokens per reply.
Instruction Following
Responds to natural language instructions for tasks such as classification, extraction, rewriting, and question answering.
Low-Latency Responses
Optimized for speed within the Claude 3 family, making it suitable for real-time or high-throughput applications like chat interfaces and content moderation pipelines.
Amazon Bedrock Integration
Deployed via Amazon Bedrock, allowing access through AWS SDKs and APIs without managing model infrastructure directly.
Ready to build with Claude 3 Haiku?
Get Started FreeBenchmark scores
Scores represent accuracy — the percentage of questions answered correctly on each test.
| Benchmark | What it tests | Score |
|---|---|---|
| GPQA Diamond | PhD-level science questions (biology, physics, chemistry) | 37.4% |
| MATH-500 | Undergraduate and competition-level math problems | 39.4% |
| AIME 2024 | American math olympiad problems | 1.0% |
| LiveCodeBench | Real-world coding tasks from recent competitions | 15.4% |
| HLE | Questions that challenge frontier models across many domains | 3.9% |
| SciCode | Scientific research coding and numerical methods | 18.6% |
Common questions about Claude 3 Haiku
What is the context window for Claude 3 Haiku?
Claude 3 Haiku supports a context window of 200,000 tokens, meaning it can process up to 200,000 tokens of combined input and conversation history in a single request.
What is the maximum response length?
The model can generate up to 4,096 tokens per response.
Is Claude 3 Haiku still available on Amazon Bedrock?
The model's status is listed as deprecated. While it may still be accessible on Amazon Bedrock under the identifier anthropic.claude-3-haiku-20240307-v1:0, Anthropic recommends migrating to a current Claude model for ongoing projects.
What is the pricing for Claude 3 Haiku on Amazon Bedrock?
Pricing information is not included in the available metadata. Refer to the Amazon Bedrock pricing page for current rates, as deprecated model pricing may differ from active models.
What is the knowledge cutoff date for Claude 3 Haiku?
The model was released in March 2024. Anthropic has not published a specific training data cutoff date in the available metadata, but the model's knowledge reflects information available prior to its March 2024 release.
What people think about Claude 3 Haiku
The available Reddit thread focuses on Claude Opus 4.5 rather than Claude 3 Haiku specifically, so direct community sentiment about Haiku is limited in this dataset. General community discussion around the Claude 3 family tends to highlight Haiku's speed and cost efficiency as practical advantages for high-volume tasks.
Developers commonly reference Haiku for use cases where response latency and API cost are primary constraints, such as chatbots and automated pipelines. Concerns in broader Claude 3 discussions often center on capability trade-offs relative to larger models in the same family.
New benchmark: Claude Opus 4.5 broke the efficiency wall.+21% intelligence while getting 66% cheaper
Documentation & links
Parameters & options
Explore similar models
Start building with Claude 3 Haiku
No API keys required. Create AI-powered workflows with Claude 3 Haiku in minutes — free.