← All Models

NVIDIA: Nemotron 3 Super

🇺🇸 NVIDIA · Nemotron 3

Input Price $0.080 per million tokens NT$2.6
Output Price $0.450 per million tokens NT$14.4
Context Window 262K tokens Output limit: 236K
OpenRouter Route Price Please verify with official pricing pages
Use this model via OpenRouter →

Overview

NVIDIA: Nemotron 3 Super is a large language model API from NVIDIA, part of its Nemotron 3 model family. At $0.080 per million input tokens and $0.450 per million output tokens, it sits in the budget tier — among the cheaper options for high-throughput or cost-sensitive workloads. Output tokens cost about 6× as much as input, so prompt-heavy workloads run noticeably cheaper than generation-heavy ones. A large 262K-token context window (≈393 pages of text) lets it take in whole books, large codebases, or lengthy transcripts in a single call. On Artificial Analysis's Intelligence Index it scores 14 (F grade), a useful proxy for its general reasoning strength relative to the other models tracked here. All prices on this page reflect OpenRouter's routed rates and are re-synced automatically every day; confirm against the provider's official pricing before committing to production.

Dimension Unit Price (USD) Price (TWD) Effective From
Input per 1M tokens $0.080 NT$2.6
Input per 1M tokens $0.080 NT$2.6
Output per 1M tokens $0.450 NT$14.4
Output per 1M tokens $0.450 NT$14.4
Cached Input per 1M tokens $0.060 NT$1.9
Cached Input per 1M tokens $0.060 NT$1.9

Provider
NVIDIA
Model Family
Nemotron 3
Version String
nvidia/nemotron-3-super-120b-a12b
Status
Active
Modality
Text
Context Window
262,144 tokens
Output Limit
235,929 tokens

Index Metrics

Cross-domain capability indexes evaluated by Artificial Analysis — Artificial Analysis

Agentic Index 4 F Measured: 2026-09-08
Coding Index 38 B Measured: 2026-09-08
Intelligence Index 14 F Measured: 2026-09-08

Benchmark Scores

Data source: Artificial Analysis

AA-LCR 65.7% B Measured: 2026-09-08
GPQA Diamond 80.0% A Measured: 2026-09-08
HLE 20.8% A Measured: 2026-09-08
IFBench 71.5% B Measured: 2026-09-08
Non-Hallucination 13.0% Measured: 2026-09-08
Omniscience Accuracy 24.3% Measured: 2026-09-08
SciCode 36.2% B Measured: 2026-09-08
Tau2 67.8% Measured: 2026-09-08
TerminalBench 28.8% Measured: 2026-09-08

Performance Metrics

Real-world benchmarks, updated every 72 hours by Artificial Analysis — Artificial Analysis

First Token Latency 3.8s Measured: 2026-09-08
Output Speed 99 t/s Measured: 2026-09-08
Response Time 29.0s Measured: 2026-09-08

90-Day Price Trend

Input / Output price (USD per 1M tokens)

Past 90 days of records; every price change is shown here

Date Dimension Price (USD) Source
Output $0.450 OpenRouter
Input $0.080 OpenRouter
Output $0.450 OpenRouter
Input $0.080 OpenRouter
Output $0.450 OpenRouter
Input $0.080 OpenRouter
Output $0.450 OpenRouter
Input $0.080 OpenRouter
Output $0.450 OpenRouter
Input $0.080 OpenRouter
Output $0.450 OpenRouter
Input $0.080 OpenRouter
Output $0.450 OpenRouter
Input $0.080 OpenRouter
Output $0.450 OpenRouter
Input $0.080 OpenRouter
Output $0.450 OpenRouter
Input $0.080 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter
Output $0.400 OpenRouter
Input $0.085 OpenRouter

Key Insights

Key data points from this page for quick reference and citation.

  • NVIDIA: Nemotron 3 Super Input price: $0.08/M tokens
  • NVIDIA: Nemotron 3 Super Output price: $0.45/M tokens
  • Context window: 262,144 tokens
  • Provider: NVIDIA
  • Model family: Nemotron 3
  • Modalities: Text
  • Data source: OpenRouter, updated daily