Skip to content
01

Providers + pricing

CapabilitiesProvider details
● Credits · billed by AnyRouter
cloudflare
cloudflare
$0.11
$0.11
● Free pool · donated keys · not available for this model — donate a key
02

Try it

POST /api/v1/chat/completions
import { createAnyRouter } from "@anyr/ai-sdk-provider"import { generateText } from "ai" const anyrouter = createAnyRouter() const { text } = await generateText({  model: anyrouter("meta-llama/llama-2-7b-chat-hf-lora"),  prompt: "Say hi in 3 words.",})console.log(text)
Set ANYROUTER_API_KEY · edits on the demo update this code
03

Uptime + latency

—recent checks · all routes

No health checks recorded for this model yet.

24h
No traffic in the last 24h.
More performance detail
Usage analytics

Loading usage…

Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Llama 2 7B Chat HF LoRADisabled since Aug 14, 2026

Workers AI LoRA host — sibling Gemma LoRA SKUs already 404 on this account. Disabled so smoke does not pick it as the next Cloudflare cheapest model.

  • cloudflare
Texttext → text8Ktoken context
Context
8K
Input
$0.11
Output
$0.11
TTFT
—
Uptime
—
Routes
1
04

Your access

Credits
Free poolnot in pool yetDonate key →
Create API key for this model
Text generation
Context length8,192 tokens
Max output6,553 tokens
ArchitectureTransformer
Categorytext
ReleasedApr 2, 2024
Modalities
Capabilities
05

About

This is a Llama2 base model that Cloudflare dedicated for inference with LoRA adapters. Llama 2 is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion to 70 billion parameters. This is the repository for the 7B fine-tuned model, optimized for dialogue use cases and converted for the Hugging Face Transformers format.

Released 2024-04-02

API & code
Share cards
Llama 2 7B Chat HF LoRA share card
Llama 2 7B Chat HF LoRA
cloudflare upstream share card
cloudflare upstream
Back to models