NVIDIA
7 params
NVIDIA Nemotron 3 Super 120B A12B parameters
These are the parameters modelparams.dev tracks for NVIDIA Nemotron 3 Super 120B A12B. Each row gives the type, default, valid range or values, and the conditions that gate it. It's the same data the JSON API serves.
Length
2 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Max tokens
max_tokens
|
integer (1…32768) | 16384 | Maximum number of tokens to generate. Generation stops when this limit is reached. | — |
|
Stop
stop
|
string | — | A string or list of strings where the API will stop generating further tokens. The returned text will not contain the stop sequence. | — |
Sampling
3 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Temperature
temperature
|
number (-∞…1) | 1 | Controls randomness. Lower values make outputs more focused; higher values make them more varied. Not recommended to modify both temperature and top_p in the same call. | — |
|
Top P
top_p
|
number (-∞…1) | 0.95 | Controls nucleus sampling by limiting generation to tokens within the selected cumulative probability. Not recommended to modify both temperature and top_p in the same call. | — |
|
Seed
seed
|
integer (0…18446744073709552000) | — | Best-effort deterministic sampling seed. Repeated requests with the same seed and parameters should return the same result. | — |
Reasoning
2 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Reasoning effort
reasoning_effort
|
enum (none | low | high) | "high" | Controls the reasoning mode. 'none' disables reasoning tokens, 'low' enables low-effort reasoning, and 'high' enables full reasoning. | — |
|
Reasoning budget
reasoning_budget
|
integer (-1…32768) | 16384 | Maximum number of tokens the model may use for internal reasoning before being forced to end the reasoning trace. Use -1 to disable budget enforcement. | — |
NVIDIA Nemotron 3 Super 120B A12B API parameters in brief
NVIDIA Nemotron 3 Super 120B A12B documents 7 API parameters, grouped by what they control:
-
Length:
max_tokensdefaults to 16384, range 1–32768;stop. -
Sampling:
temperaturedefaults to 1, maximum 1;top_pdefaults to 0.95, maximum 1;seedrange 0–18446744073709552000. -
Reasoning:
reasoning_effortdefaults to "high", accepts "none", "low", "high";reasoning_budgetdefaults to 16384, range -1–32768.
Frequently asked questions
- How many parameters does NVIDIA Nemotron 3 Super 120B A12B accept?
- NVIDIA Nemotron 3 Super 120B A12B accepts 7 API parameters: temperature, top_p, max_tokens, reasoning_effort, reasoning_budget, seed, and more.
- What is the default temperature for NVIDIA Nemotron 3 Super 120B A12B?
- The default temperature for NVIDIA Nemotron 3 Super 120B A12B is 1, with a maximum of 1.
- What is the default top_p for NVIDIA Nemotron 3 Super 120B A12B?
- The default top_p for NVIDIA Nemotron 3 Super 120B A12B is 0.95, with a maximum of 1.
- What is the default max_tokens for NVIDIA Nemotron 3 Super 120B A12B?
- The default max_tokens for NVIDIA Nemotron 3 Super 120B A12B is 16384, within a valid range of 1 to 32768.
Resources
Other NVIDIA models
GLiNER PII
4 params
View
Llama 3.1 NemoGuard 8B Topic Control
6 params
View
Llama 3.1 Nemotron Nano 8B v1
7 params
View
Llama 3.1 Nemotron Safety Guard 8B v3
1 param
View
Llama 3.1 Nemotron Ultra 253B v1
7 params
View
Llama 3.3 Nemotron Super 49B v1
7 params
View
Llama 3.3 Nemotron Super 49B v1.5
7 params
View
NemoGuard Jailbreak Detect
0 params
View
Nemotron 3 Nano 30B A3B
5 params
View
Nemotron 3 Ultra
Subscription
6 params
View
Nemotron 3 Ultra 550B A55B
7 params
View
Nemotron Content Safety Reasoning 4B
5 params
View
Nemotron Mini 4B Instruct
6 params
View
Riva Translate 4B Instruct v1.1
6 params
View
USDCode Llama 3.1 70B Instruct
4 params
View