Post-Cutoff.com
  1. Home
  2. Models
  3. NVIDIA Nemotron 3 Super (120B-A12B)

NVIDIA Nemotron 3 Super (120B-A12B)

NVIDIAcurrentreasoning-llmNemotron 3open weights

Knowledge cutoff = pre-training (Jun 2025); post-training to Feb 2026. Also FP8/NVFP4 repos. Free tier on OpenRouter (:free).

Context window
1,000,000 tokens
Knowledge cutoff
2025-06
Input
text
Output
text
License
nvidia-open-model-license
Pricing
input: $0.08 · output: $0.45 (per 1M tokens (USD) on OpenRouter; NVIDIA hosted pricing not verified) source
Verified
2026-09-29

How to call it

ProviderModel idEndpoint / URLDocs
NVIDIA API (build.nvidia.com)nvidia/nemotron-3-super-120b-a12bhttps://integrate.api.nvidia.com/v1/chat/completionsdocs
AWS Bedrocknvidia.nemotron-super-3-120b—docs
OpenRouternvidia/nemotron-3-super-120b-a12bopenrouter.ai/nvidia/nemotron-3-super-120b-a12b—
Hugging Face—huggingface.co/nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16—

Notable capabilities (2)

Mid-size open reasoning model, widely hosted (Bedrock, NIM, OpenRouter).

curl https://integrate.api.nvidia.com/v1/chat/completions -H "Authorization: Bearer $NVIDIA_API_KEY" -H "Content-Type: application/json" \
 -d '{"model":"nvidia/nemotron-3-super-120b-a12b","messages":[{"role":"user","content":"Hello"}]}'

Sources: https://huggingface.co/nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16 , https://docs.aws.amazon.com/bedrock/latest/userguide/model-card-nvidia-nemotron-super-3-120b.html

Other NVIDIA models

NVIDIA Nemotron 3.5 Lightning (30B-A3B) · NVIDIA NemotronLabs VoiceChat 11B (and PersonaPlex-7B) · NVIDIA Nemotron 3 Ultra (550B-A55B) · Cosmos 3 (Nano / Super) · NVIDIA Nemotron 3 Nano Omni (30B-A3B Reasoning) · Isaac GR00T N1.7 · Cosmos Reason 2 · NVIDIA MagpieTTS Multilingual 357M · NVIDIA Parakeet / Canary / Nemotron Speech ASR (open) · Isaac GR00T N2 · Cosmos Predict 2.5 / Transfer 2.5 · Isaac GR00T N1 / N1.5 / N1.6