Skip to content
01

Providers + pricing

CapabilitiesProvider details
● Your own key · billed by your provider · AnyRouter fee $0
Hugging Face
huggingface-byok
$0.15list$0.29list
Unavailable
OpenRouter
openrouter-byok
$0.05list$0.08list
Unavailable
● Free pool · donated keys · not available for this model — donate a key
02

Try it

POST /api/v1/chat/completions
import { createAnyRouter } from "@anyr/ai-sdk-provider"import { generateText } from "ai" const anyrouter = createAnyRouter() const { text } = await generateText({  model: anyrouter("meta/llama-3.1-8b-instruct"),  prompt: "Say hi in 3 words.",})console.log(text)
Set ANYROUTER_API_KEY · edits on the demo update this code
Live demo

Type a prompt and run it on this model.

03

Uptime + latency

—recent checks · all routes

No health checks recorded for this model yet.

24h
No traffic in the last 24h.
More performance detail
Usage analytics

Loading usage…

Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Llama 3.1 8B Instruct

Also accepted:meta-llama/llama-3.1-8b-instructmeta/llama-3.1-8b-instruct-fp8
  • Hugging Face
  • OpenRouter
Texttext → text131Ktoken context
Context
131K
Input
$0.05
Output
$0.08
TTFT
200ms
Uptime
—
Routes
2
04

Your access

Credits
Your own key · AnyRouter fee $0

Run Llama 3.1 8B Instruct on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Free poolnot in pool yetDonate key →
Create API key for this model
Text generation
Context length131,072 tokens
Max output104,857 tokens
ArchitectureTransformer
Categorytext
ReleasedJul 23, 2024
Modalities
Capabilities
05

About

Meta's compact Llama 3.1 8B instruction-tuned model optimized for fast inference and edge deployments.

Released 2024-07-23 · params: max_tokens · temperature · top_p · stop · tools · tool_choice · response_format

API & code
Share cards
Llama 3.1 8B Instruct share card
Llama 3.1 8B Instruct
Hugging Face upstream share card
Hugging Face upstream
OpenRouter upstream share card
OpenRouter upstream
Back to models