reasoningNVIDIA
Nemotron 3 Super 120B
NVIDIA's mixture-of-experts reasoning model — 120B total parameters with 12B active, so frontier-class reasoning arrives at small-model latency.
nvidia/nemotron-3-super-120b-a12bRunning· verified 11 min ago · 725ms
What is Nemotron 3 Super 120B?
Nemotron 3 Super 120B is NVIDIA's reasoning model. NVIDIA's mixture-of-experts reasoning model — 120B total parameters with 12B active, so frontier-class reasoning arrives at small-model latency. It runs on Speka's OpenAI-compatible /v1/chat/completions endpoint at $0.40 / 1M input and $0.60 / 1M output tokens, using the model ID nvidia/nemotron-3-super-120b-a12b.
Test it live
Live playground
Nemotron 3 Super 120Bin-browser demo
Test Nemotron 3 Super 120B live
Send a message and watch the model respond. Multi-turn — it remembers the conversation.