anoman
Sign InGet API Key

Open-source LLM (via Fireworks AI)

Nemotron 3 Ultra NVFP4

modelnemotron-3-ultra-nvfp4
MidStableStreamingToolsBatchLong contextUnited States

Context

262K

Input / 1M tokens

$0.6000

Output / 1M tokens

$2.40

Cached input / 1M tokens

$0.1200

Cache write / 1M tokens

— (bills at input price)

About this model

Nemotron 3 Ultra NVFP4

Nemotron 3 Ultra NVFP4 is available through the Anoman AI OpenAI-compatible API with a 262K-token context window. It costs $0.6000 per 1M input tokens and $2.40 per 1M output tokens, and is available from the Pro plan.

Key facts

Context window
262K tokens
Max output
4K tokens
Lane
Guarded: the provider has confirmed it does not train on your data
Modality
Text
Input price
$0.6000 per 1M tokens
Output price
$2.40 per 1M tokens
Processing region(s)
us
Route
Via Fireworks AI
Cheapest plan
Available from the Pro plan
Price comparison
Priced below 125 of 241 pro-band guarded models on Anoman (blended price: 1 part input to 3 parts output).

Code example

Chat completion with Nemotron 3 Ultra NVFP4

from openai import OpenAI

client = OpenAI(
    base_url="https://api.anoman.io/v1",
    api_key="anm-sk-..."
)

response = client.chat.completions.create(
    model="nemotron-3-ultra-nvfp4",
    messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)

Alternatives

Similar models in our catalog

Same maker first, then the closest price to Nemotron 3 Ultra NVFP4 in the Anoman catalog.

Use Nemotron 3 Ultra NVFP4 through Anoman.