By nvidia ยท Updated 2026-07-24 ยท long-context / RAGcheap high-volumestep-by-step reasoning
NVIDIA: Nemotron 3 Super is best suited for long-context / RAG, cheap high-volume, step-by-step reasoning. Access it through one Appmo API key alongside 390 other models โ Appmo's optimizer can route to it automatically when it is the best value for a request.
curl https://appmo.com/v1/chat/completions \
-H "Authorization: Bearer $APPMO_KEY" \
-d '{"model":"nvidia/nemotron-3-super-120b-a12b","messages":[{"role":"user","content":"Hello"}]}'How much does NVIDIA: Nemotron 3 Super cost?
NVIDIA: Nemotron 3 Super costs $0.085 per 1M input tokens and $0.4 per 1M output tokens ($0.164 blended), as of 2026-07-24. Prices are expected to fall as competition increases.
What is NVIDIA: Nemotron 3 Super best for?
NVIDIA: Nemotron 3 Super is best suited for long-context / RAG, cheap high-volume, step-by-step reasoning.
How do I use NVIDIA: Nemotron 3 Super?
Send requests to Appmo's OpenAI-compatible endpoint with model id "nvidia/nemotron-3-super-120b-a12b" and one API key โ no separate provider account needed.