Meta (Llama)
meta-llama/llama-4-scout-17b-16e-instruct
model
meta-llama/llama-4-scout-17b-16e-instructBudgetStableVisionStreamingVisionToolsBatchLong contextUnited States
Context
131K
Input / 1M tokens
$0.1100
Output / 1M tokens
$0.3400
Weighted tokens formula
1× provider × 1× tier
About this model
meta-llama/llama-4-scout-17b-16e-instruct
meta-llama/llama-4-scout-17b-16e-instruct is a budget-tier model available through the Anoman gateway. Detailed performance benchmarks and use-case guidance for this model are coming soon. In the meantime, see the code example below to integrate it into your workflow, and the tier + region information in the hero for pricing and residency.
Code example
Chat completion with meta-llama/llama-4-scout-17b-16e-instruct
from openai import OpenAI
client = OpenAI(
base_url="https://api.anoman.io/v1",
api_key="anm-sk-..."
)
response = client.chat.completions.create(
model="meta-llama/llama-4-scout-17b-16e-instruct",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)Weighted tokens
weighted_tokens = raw_tokens × 1 (provider) × 1 (tier)
Pro plan: 20M weighted tokens/month. Combined multiplier 1×: 20,000,000 raw tokens available.