Llama 2 7B Chat HF LoRA is Meta's language model with a 8K context window and up to 8K output tokens. Llama 2 7B chat-tuned base model dedicated for inference with LoRA adapters, enabling parameter-efficient fine-tuning for conversational applications.
Specifications
Canonical IDmeta-llama-2-7b-chat-hf-lora
TypeLanguage
StatusActive
CreatorMetaMeta
Context Window8K tokens
Max Output8K tokens
Input ModalitiesText
Output ModalitiesText

Capabilities

Input1/5
Text
Image·
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities0/13
Reasoning·
Adaptive Reasoning·
Function Calling·
Parallel Function Calling·
Structured Outputs·
Native JSON Schema·
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
Llama 3.3 70B Instruct131K$0.120$0.200Deprecated
Llama 3.2 3B Instruct131K$0.015$0.020Deprecated
Llama 3.2 1B Instruct128K$0.027$0.080Deprecated
Llama 3.2 11B128K$0.160$0.160Available
Llama 3.1 405B Instruct131K$0.120$0.300Deprecated
Llama 3.1 8B Instruct200K$0.020$0.030Deprecated
Llama 3.1 70B Instruct131K$0.120$0.300Available
Llama 3.1 70B128K$0.360$0.360Available
Llama 3.1 8B131K$0.030$0.050Available
Llama 3 70B Instruct131K$0.120$0.300Deprecated
Llama 2 7B Chat HF LoRA8KCurrent

Model IDs

@cf/meta-llama/llama-2-7b-chat-hf-lora
cloudflare/@cf/meta-llama/llama-2-7b-chat-hf-lora
meta-llama-2-7b-chat-hf-lora