NVIDIA / nvidia/nemotron-3-super-120b-a12b
Nemotron 3 Super 120B A12B - access through LLMTR
Nemotron 3 Super (120B total, 12B active parameters) is a balanced open-weight model with a 262K-token context window. It fits general chat, summarization, coding help, function calling, and product flows that expect JSON-formatted output. It is offered free with a daily usage quota.
Technical specifications
| Canonical ID | nvidia/nemotron-3-super-120b-a12b |
|---|---|
| Provider | NVIDIA |
| Context window | 262,144 tokens |
| Operations | CHAT_COMPLETIONS |
| Modalities | text |
Pricing
An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.
| Operation | Metric | Unit | Price |
|---|---|---|---|
| CHAT_COMPLETIONS | INPUT_TEXT | PER_1M_TOKENS | Not available |
| CHAT_COMPLETIONS | OUTPUT_TEXT | PER_1M_TOKENS | Not available |
Example usage
With existing OpenAI SDK flows, change only the base URL and model identifier.
curl https://llmtr.com/v1/chat/completions -H "Authorization: Bearer llmtr-your_key" -H "Content-Type: application/json" -d '{"model":"nvidia/nemotron-3-super-120b-a12b","messages":[{"role":"user","content":"Hello"}]}'
Related models
- Nemotron 3 Ultra 550B A55B nvidia/nemotron-3-ultra-550b-a55b
- Nemotron 3 Ultra 550B A55B 262K nvidia/nemotron-3-ultra-550b-a55b-262k
- Gemma 4 llmtr/gemma-4
- Qwen 3.6 35B-A3B llmtr/qwen3-6-35b
- Trendyol Asure 12B llmtr/trendyol-asure-12b
- Magibu 11B v8 llmtr/magibu-11b-v8
- Muse Glimmer 30B (Turkey) llmtr/muse-glimmer-30b-tr
- EmbeddingGemma 300M llmtr/embeddinggemma-300m