Open-source LLM (via Fireworks AI)
Nemotron 3 Ultra NVFP4
model
nemotron-3-ultra-nvfp4MidStableStreamingToolsBatchLong contextUnited States
Context
262K
Input / 1M tokens
$0.6000
Output / 1M tokens
$2.40
Cached input / 1M tokens
$0.1200
Cache write / 1M tokens
— (bills at input price)
About this model
Nemotron 3 Ultra NVFP4
Nemotron 3 Ultra NVFP4 is available through the Anoman AI OpenAI-compatible API with a 262K-token context window. It costs $0.6000 per 1M input tokens and $2.40 per 1M output tokens, and is available from the Pro plan.
Key facts
- Context window
- 262K tokens
- Max output
- 4K tokens
- Lane
- Guarded: the provider has confirmed it does not train on your data
- Modality
- Text
- Input price
- $0.6000 per 1M tokens
- Output price
- $2.40 per 1M tokens
- Processing region(s)
- us
- Route
- Via Fireworks AI
- Cheapest plan
- Available from the Pro plan
- Price comparison
- Priced below 125 of 241 pro-band guarded models on Anoman (blended price: 1 part input to 3 parts output).
Code example
Chat completion with Nemotron 3 Ultra NVFP4
from openai import OpenAI
client = OpenAI(
base_url="https://api.anoman.io/v1",
api_key="anm-sk-..."
)
response = client.chat.completions.create(
model="nemotron-3-ultra-nvfp4",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)Alternatives
Similar models in our catalog
Same maker first, then the closest price to Nemotron 3 Ultra NVFP4 in the Anoman catalog.