Skip to main content
MindStudio
Pricing
BlogAbout
My Workspace
Text Generation ModelDeprecated

Llama-2 13B Chat

Llama-2 13B Chat is a conversational text generation model from Meta with a 4096 token context window.

PublisherMeta
TypeText
Context Window4,096 tokens
Replaced byClaude 3 Haiku
ProviderDeepInfra

13B parameter chat model from Meta

Llama-2 13B Chat is a 13-billion parameter language model developed by Meta as part of the Llama 2 family of open-weight models. It is a fine-tuned variant optimized for dialogue and conversational use cases, trained using reinforcement learning from human feedback (RLHF). The model was released in July 2023 alongside other sizes in the Llama 2 lineup, including 7B and 70B variants, and is hosted on Hugging Face under the identifier meta-llama/Llama-2-13b-chat-hf.

The 13B parameter size positions this model between the lighter 7B and the larger 70B variants, offering a middle ground in terms of computational requirements and language understanding depth. It supports a context window of 4096 tokens and a maximum response size of 2500 tokens. The model is suited for tasks involving multi-turn conversation, instruction following, and general-purpose text generation where a fully open-weight model is preferred. Note that this model has been marked as deprecated on MindStudio.

What Llama-2 13B Chat supports

Conversational Chat

Fine-tuned with RLHF specifically for multi-turn dialogue, making it suitable for chatbot and assistant-style interactions.

Instruction Following

Trained to respond to natural language instructions across a range of general-purpose tasks, including summarization and Q&A.

Text Generation

Generates coherent long-form text up to a maximum response size of 2500 tokens within a 4096 token context window.

Open Weights Access

Released as an open-weight model, available for download and self-hosting via Hugging Face under Meta's Llama 2 community license.

Context Retention

Supports up to 4096 tokens of context, allowing the model to reference earlier parts of a conversation within a single session.

Ready to build with Llama-2 13B Chat?

Get Started Free

Common questions about Llama-2 13B Chat

What is the context window size for Llama-2 13B Chat?

Llama-2 13B Chat supports a context window of 4096 tokens, which includes both the input prompt and the generated response.

What is the maximum response length this model can produce?

The model has a maximum response size of 2500 tokens per generation.

Is Llama-2 13B Chat still available on MindStudio?

The model is currently marked as deprecated on MindStudio, which means it may no longer be actively supported or recommended for new projects.

Does Llama-2 13B Chat support image or video inputs?

No. Llama-2 13B Chat is a text-only model and does not support image or video analysis.

Who published Llama-2 13B Chat and under what license?

The model was published by Meta and is available under the Llama 2 community license, which permits use for research and commercial applications subject to Meta's terms.

What is the knowledge cutoff date for Llama-2 13B Chat?

Meta has not published a precise knowledge cutoff date in the available metadata. Based on public information, Llama 2 models were trained on data with a cutoff of approximately September 2022.

Parameters & options

Max Temperature1
Max Response Size2,500 tokens

Start building with Llama-2 13B Chat

No API keys required. Create AI-powered workflows with Llama-2 13B Chat in minutes — free.