Skip to main content
MindStudio
Pricing
BlogAbout
My Workspace
Text Generation ModelDeprecated

Claude 3 Haiku

Claude 3 Haiku is a text generation model from Anthropic with a 200,000 token context window, available via Amazon Bedrock.

PublisherAnthropic
TypeText
Context Window200,000 tokens
ReleasedMarch 2024
Replaced byClaude 4.5 Haiku
ProviderAmazon Bedrock

Compact text generation with large context

Claude 3 Haiku is a large language model developed by Anthropic and released in March 2024. It is the smallest model in the Claude 3 model family and is designed for tasks where response speed and efficiency matter, such as customer-facing applications, content moderation, and lightweight summarization. It supports a 200,000 token context window, allowing it to process long documents in a single request, and generates responses up to 4,096 tokens.

Claude 3 Haiku is available through Amazon Bedrock under the model identifier anthropic.claude-3-haiku-20240307-v1:0. Its status is listed as deprecated, meaning Anthropic and Amazon Bedrock have moved toward newer model versions for active use. Developers who built workflows around this model should evaluate migration to a current Claude release, as deprecated models may have limited long-term availability on the platform.

What Claude 3 Haiku supports

Long Context Processing

Handles up to 200,000 tokens of input in a single request, enabling analysis of lengthy documents, codebases, or conversation histories without chunking.

Text Generation

Generates coherent prose, summaries, and structured text responses up to 4,096 tokens per reply.

Instruction Following

Responds to natural language instructions for tasks such as classification, extraction, rewriting, and question answering.

Low-Latency Responses

Optimized for speed within the Claude 3 family, making it suitable for real-time or high-throughput applications like chat interfaces and content moderation pipelines.

Amazon Bedrock Integration

Deployed via Amazon Bedrock, allowing access through AWS SDKs and APIs without managing model infrastructure directly.

Ready to build with Claude 3 Haiku?

Get Started Free

Benchmark scores

Scores represent accuracy — the percentage of questions answered correctly on each test.

BenchmarkWhat it testsScore
GPQA DiamondPhD-level science questions (biology, physics, chemistry)37.4%
MATH-500Undergraduate and competition-level math problems39.4%
AIME 2024American math olympiad problems1.0%
LiveCodeBenchReal-world coding tasks from recent competitions15.4%
HLEQuestions that challenge frontier models across many domains3.9%
SciCodeScientific research coding and numerical methods18.6%

Common questions about Claude 3 Haiku

What is the context window for Claude 3 Haiku?

Claude 3 Haiku supports a context window of 200,000 tokens, meaning it can process up to 200,000 tokens of combined input and conversation history in a single request.

What is the maximum response length?

The model can generate up to 4,096 tokens per response.

Is Claude 3 Haiku still available on Amazon Bedrock?

The model's status is listed as deprecated. While it may still be accessible on Amazon Bedrock under the identifier anthropic.claude-3-haiku-20240307-v1:0, Anthropic recommends migrating to a current Claude model for ongoing projects.

What is the pricing for Claude 3 Haiku on Amazon Bedrock?

Pricing information is not included in the available metadata. Refer to the Amazon Bedrock pricing page for current rates, as deprecated model pricing may differ from active models.

What is the knowledge cutoff date for Claude 3 Haiku?

The model was released in March 2024. Anthropic has not published a specific training data cutoff date in the available metadata, but the model's knowledge reflects information available prior to its March 2024 release.

What people think about Claude 3 Haiku

The available Reddit thread focuses on Claude Opus 4.5 rather than Claude 3 Haiku specifically, so direct community sentiment about Haiku is limited in this dataset. General community discussion around the Claude 3 family tends to highlight Haiku's speed and cost efficiency as practical advantages for high-volume tasks.

Developers commonly reference Haiku for use cases where response latency and API cost are primary constraints, such as chatbots and automated pipelines. Concerns in broader Claude 3 discussions often center on capability trade-offs relative to larger models in the same family.

View more discussions →

Parameters & options

Max Temperature1
Max Response Size4,096 tokens

Start building with Claude 3 Haiku

No API keys required. Create AI-powered workflows with Claude 3 Haiku in minutes — free.