anoman
Pricing

Simple, transparent pricing.

0% markup — you pay exactly what providers charge. Revenue comes from the platform fee, not a tax on your tokens.

Your bill, itemizedper 1M tokens
Provider costpass-through
Platform feeflat
Token markup0%
You pay providers' rates — we never tax your tokens.

Free

$0

Free

20 RPM40K TPM
  • Up to 2M tokens/day, best-effort
  • Full guardrails on every call
  • $3 credit + premium-model trial
  • Local AI model (Sahabat-AI) + open-source
  • No credit card, free forever
Start free

Budget Pass

Rp 5rb/hari

or Rp 30k/week

Coming soon
30 RPM60K TPM
  • 2M tokens/day, guaranteed
  • Budget-tier models
  • Full guardrails
  • Pay per day or week — no subscription

Plus Pass

Rp 15rb/hari

or Rp 90k/week

Coming soon
60 RPM120K TPM
  • 3M tokens/day, guaranteed
  • Budget + Mid-tier models
  • Full guardrails
  • Pay per day or week — no subscription
MOST POPULAR

Pro

$27/mo

Rp 499rb

≈ Rp 17rb/hari ☕

120 RPM500K TPM
  • Budget + Mid + Premium
  • Batch routing (50% off)
  • Semantic cache
  • 5 API keys
Get Started

Pay-As-You-Go

$9/mo + raw cost

Rp 159rb

60 RPM200K TPM
  • All 400+ LLMs
  • 0% base markup
  • Full guardrails
  • PAYG billing
Start Free

Enterprise

$499+/mo

Rp 8,5jt+

30,000 RPM100M TPM
  • All tiers incl. Ultra
  • SSO/SAML
  • Audit log export
  • SLA guarantee
  • Net-30 invoicing
Contact Sales

Free tier

Guardrails on every call — even the free ones

Every signup gets up to 2M tokens/day of guarded calls on low-cost models — best-effort by design, free forever. Plus a $3 prepaid credit and a premium-model trial to start. No credit card, billed in Rupiah.

Guarded, not raw

Every free call runs through the full guardrail stack — prompt-injection detection, PII handling and content moderation — the same defense your paid traffic gets.

Best-effort, honest

The free bucket is best-effort and resets daily. When you need reliable throughput that is never throttled, that is what the paid plans are for.

Pay in Rupiah when you grow

Upgrade with local payment — QRIS, GoPay and Rupiah, no international credit card required.

See the guardrails catch an attack in real time: Try the injection comparison →

Weighted token formula

Not all tokens are equal

weighted_tokens = raw_tokens × provider_multiplier × tier_multiplier × quota_discount

Provider Multiplier

Cloud Direct (OpenAI, Anthropic, Google)
Bedrock / self-hosted0.5×
Local ID (Jakarta edge)0.3×

Model Tier Multiplier

Budget (Gemini Flash, Llama, Haiku)
Mid (Claude Sonnet)
Premium (Claude Opus)17×
Ultra (largest models)42×

Quota Discounts

Cache hit (semantic)0% of quota used
Provider cache (prefix)10% discount
Batch routing50% discount
Real-time100% (full cost)

Token consumption

How far do your tokens go?

Estimated at 1,000 tokens per request (typical chat turn)

ModelTierWeightStarter (500K wt)Pro (20M wt)
Gemini 2.5 FlashBudget500K calls20M calls
Llama 3.3 70BBudget500K calls20M calls
Claude Haiku 4.5Budget500K calls20M calls
Claude Sonnet 4.6Mid125K calls5M calls
Claude Opus 4.7Premium17×~29K calls~1.2M calls

Pay-As-You-Go markup

Volume-based progressive markup

PAYG customers pay a $9/mo platform fee + raw provider cost, with a small progressive markup based on monthly call volume.

Monthly callsMarkupNote
0–1,0000%Pure pass-through
1,001–5,0001%Near pass-through
5,001–20,0002%Below OpenRouter (10–25%)
20,001–50,0003%3–5× cheaper than OpenRouter
50,001+5% capStill 2–5× cheaper than OpenRouter

Models by region

Pick where your data is processed

Every model declares the region where the request is processed. Filter the catalog by region to align your model choice with your compliance posture. Cross-border data movement is opt-in.

Indonesia

Framework: UU PDP

Models: Frontier + curated chat

Best for: All Indonesian-resident workloads

Singapore

Framework: PDPA

Models: Frontier + curated chat

Best for: PDPA workloads (Q3 2026)

China

Framework: PIPL

Models: 9 OSS text + 3 OSS vision

Best for: Cost-optimized OSS, ~50–75% cheaper

United States

Framework:

Models: Most frontier flagship

Best for: Latency-insensitive global

Europe

Framework: GDPR

Models: EU-resident frontier

Best for: GDPR workloads

Global

Framework: Multi-region

Models: Routed for best latency

Best for: Default residency-agnostic

Pricing across regions is pass-through — Anoman applies 0% markup. The platform fee covers the gateway, guardrails, and observability.

Compare plans

Everything you get

Full feature breakdown across all tiers.

FeatureFreeStarterProPay-As-You-GoEnterprise
Weighted tokens/mo2M/day, best-effort500K20MUnlimitedCustom
OverageHard capHard cap$1.20/1M wt0–5% markupNegotiated
Token cap (budget, mo)2M/day25M120MNo capNo cap
Token cap (mid, mo)24MNo capNo cap
Token cap (premium, mo)Trial4MNo capNo cap
Ultra-class token capNo capNo cap
Requests / min (rate)20301206030,000
Tokens / min (rate)40K60K500K200K100M
Max concurrent requests2560150125500
Budget models
Mid models
Premium modelsTrial
Ultra models
Community models (opt-in)
API keys1155Unlimited
Prompt injection detection
PII masking (4 modes)
Content moderation
Conversational guardrails
Agent policy control
MCP governance
Batch routing (50% savings)
Token packs
Semantic cache
Team management + RBAC
Organization workspaces
Member groups + entitlements
SSO/SAML
Audit log export
Custom rate limits
SLA guarantee
Net-30 invoicing
IDR billing
SupportCommunityCommunityEmailEmailDedicated

Ready to secure every AI call?

Start for free. No credit card required. Upgrade when you need more.

Curious how we stack up against other providers? Compare Anoman vs OpenRouter, OpenAI & more →

Frequently Asked Questions

Everything you need to know about Anoman AI and LLM gateway security.

An LLM gateway is a proxy between your application and the upstream model providers. It routes requests, enforces security policies, meters usage, and gives you observability — through one OpenAI-compatible endpoint. Anoman is the guarded LLM gateway: prompt-injection defense, PII masking, and policy control run on every call.