Meta (Llama)
Llama 3.1 8B (Cloudflare)
model
cf-llama-3.1-8bBudgetStableStreamingGlobal
Context
8K
Input / 1M tokens
—
Output / 1M tokens
—
Weighted tokens formula
1× provider × 1× tier
About this model
Llama 3.1 8B (Cloudflare)
Llama 3.1 8B (Cloudflare) is a budget-tier model available through the Anoman gateway. Detailed performance benchmarks and use-case guidance for this model are coming soon. In the meantime, see the code example below to integrate it into your workflow, and the tier + region information in the hero for pricing and residency.
Code example
Chat completion with Llama 3.1 8B (Cloudflare)
from openai import OpenAI
client = OpenAI(
base_url="https://api.anoman.io/v1",
api_key="anm-sk-..."
)
response = client.chat.completions.create(
model="cf-llama-3.1-8b",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)Weighted tokens
weighted_tokens = raw_tokens × 1 (provider) × 1 (tier)
Pro plan: 20M weighted tokens/month. Combined multiplier 1×: 20,000,000 raw tokens available.