DzRouter

Nemotron 3 Ultra NVFP4

ReasoningTool use

NVIDIA's frontier-scale hybrid Latent-MoE model (550B total, 55B active) for the most demanding agentic, reasoning, and long-context workloads across code, math, and science. Trained with an NVFP4 recipe.

Model id — paste this in your code
nvidia/nemotron-3-ultra-nvfp4

Pricing · per 1M tokens

Show prices in
Input150 DA$0.6
Output660 DA$2.64
Cached input30 DA$0.12
Context window262KMax output: 66K

Capabilities

  • Reasoning — Thinks step by step before answering
  • Tool use — Can call your functions / tools

Estimate your cost

Estimated cost

Use this model

Create an account, top up in dinars and call it with your key:

from openai import OpenAI

client = OpenAI(
    base_url="https://dzrouter.com/v1",
    api_key="sk-dz-…",
)

reply = client.chat.completions.create(
    model="nvidia/nemotron-3-ultra-nvfp4",
    messages=[{"role": "user", "content": "Salam!"}],
)
print(reply.choices[0].message.content)
Create a free account

Start building with AI today

Create your account, top up from 500 DA and send your first request in minutes.

Create a free account