Nvidia

NVIDIA: Nemotron 3 Ultra

nvidia/nemotron-3-ultra-550b-a55b-20260604:freeHugging Face ↗

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

toolsstructuredreasoning
Data source

NVIDIA: Nemotron 3 Ultra usage over time

OpenRouter share of tokens, by day.

View
Tidelines — LLM usage and market share

NVIDIA: Nemotron 3 Ultra holds 3.04% today, up from 0.43% at the start of the visible window.

Data: OpenRouter as of 2026-09-01 · Source: tidelines.ai/models/nvidia-nemotron-3-ultra-550b-a55b-20260604-free

Rank history

Tidelines — LLM usage and market share
Data: OpenRouter as of 2026-09-01 · Source: tidelines.ai/models/nvidia-nemotron-3-ultra-550b-a55b-20260604-free

Adoption curve vs the pantheon

NVIDIA: Nemotron 3 Ultra's share plotted against days since first tracked, overlaid on other current leaders on the same axis. The question every new-model page asks: is this one going to make it?

Tidelines — LLM usage and market share
Data: OpenRouter as of 2026-09-01 · Source: tidelines.ai/models/nvidia-nemotron-3-ultra-550b-a55b-20260604-free

Citable stat

As of 2026-09-01, NVIDIA: Nemotron 3 Ultra holds 3.0% of OpenRouter token share, down 1.56pp over 30 days. Source: Tidelines (OpenRouter).

Cost per session

Median dollars spent on one NVIDIA: Nemotron 3 Ultra session, by coding agent and session length.

Tidelines — LLM usage and market share

Token rate sets the price; session length and reasoning set the bill.

Agent1 turn2–9 turns10–49 turns
Hermes Agent$0.012$0.035$0.307

Compare every model's session cost →

Data: OpenRouter session-cost dataset as of 2026-08-30 · Source: tidelines.ai/models/nvidia-nemotron-3-ultra-550b-a55b-20260604-free

Where to run it

NVIDIA: Nemotron 3 Ultra pricing by provider

Tidelines — LLM usage and market share

Per-provider pricing, context, and uptime via OpenRouter. Sorted by output price.

NVIDIA: Nemotron 3 Ultra is served by 3 tracked endpoints on OpenRouter. The cheapest output price is $2.20 per million tokens from DeepInfra, and the most expensive is $3.13 from Venice. The largest context window offered is 262k tokens. Prices are list rates and refresh daily.

ProviderInput $/MOutput $/MContextUptime
DeepInfra$0.50$2.20262k
BaseTen$0.60$2.40203k99.33%
Venice$0.63$3.13256k100.00%