Gemini 3.5 Flash Lite
model
gemini-3.5-flash-liteBudgetStableVisionStreamingVisionToolsBatchLong contextUnited States
Context
1.0M
Input / 1M tokens
$0.3000
Output / 1M tokens
$2.50
Cached input / 1M tokens
$0.0300
Cache write / 1M tokens
β (bills at input price)
About this model
Gemini 3.5 Flash Lite
Gemini 3.5 Flash Lite is a model by Google, available through the Anoman AI OpenAI-compatible API with a 1.0M-token context window. It costs $0.3000 per 1M input tokens and $2.50 per 1M output tokens, and is available from the Pro plan.
Key facts
- Maker
- Context window
- 1.0M tokens
- Max output
- 66K tokens
- Lane
- Guarded: the provider has confirmed it does not train on your data
- Modality
- Text and image input
- Input price
- $0.3000 per 1M tokens
- Output price
- $2.50 per 1M tokens
- Processing region(s)
- us
- Route
- Via Google's API
- Cheapest plan
- Available from the Pro plan
- Price comparison
- Priced below 121 of 241 pro-band guarded models on Anoman (blended price: 1 part input to 3 parts output).
Ways to call
Ways to call Gemini 3.5 Flash Lite
Each model ID below is callable through the Anoman AI API. Price and processing region depend on the route.
| Model ID | Route | Input / output per 1M tokens | Region(s) | Lane | Context |
|---|---|---|---|---|---|
| gemini-3.5-flash-lite | Via Google's API | $0.3000 / $2.50 | us | Guarded | 1.0M |
| openrouter/google/gemini-3.5-flash-lite | Via OpenRouter | $0.3000 / $2.50 | global | Guarded | 1.0M |
Code example
Chat completion with Gemini 3.5 Flash Lite
from openai import OpenAI
client = OpenAI(
base_url="https://api.anoman.io/v1",
api_key="anm-sk-..."
)
response = client.chat.completions.create(
model="gemini-3.5-flash-lite",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)Alternatives
Similar models in our catalog
Same maker first, then the closest price to Gemini 3.5 Flash Lite in the Anoman catalog.