NVIDIA Llama 3.1 NemoGuard 8B Topic Control API parameters
These are the API parameters modelparams.dev tracks for NVIDIA Llama 3.1 NemoGuard 8B Topic Control on Chat Completions with an API key — the settings you send in a request. Each row gives the type, default, valid range or values, and the conditions that gate it. It's the same data the JSON API serves.
After the parameter count instead — how many weights Llama 3.1 NemoGuard 8B Topic Control has? That's a different number, and we don't track it. Here's the difference.
Length
2 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Max tokens
max_tokens
|
integer (1…+∞) | 1024 | Maximum number of tokens to generate. Generation stops when this limit is reached. | — |
|
Stop
stop
|
string | — | A string or list of strings where the API will stop generating further tokens. The returned text will not contain the stop sequence. | — |
Sampling
4 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Temperature
temperature
|
number (0…2) | 0.5 | Controls randomness. Lower values make outputs more focused; higher values make them more varied. Not recommended to modify both temperature and top_p in the same call. | — |
|
Top P
top_p
|
number (-∞…1) | 1 | Controls nucleus sampling by limiting generation to tokens within the selected cumulative probability. Not recommended to modify both temperature and top_p in the same call. | — |
|
Frequency penalty
frequency_penalty
|
number (-2…2) | 0 | Penalizes new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. | — |
|
Presence penalty
presence_penalty
|
number (-2…2) | 0 | Positive values penalize new tokens based on whether they appear in the text so far, increasing the model's likelihood to talk about new topics. | — |
NVIDIA Llama 3.1 NemoGuard 8B Topic Control API parameters in brief
NVIDIA Llama 3.1 NemoGuard 8B Topic Control on Chat Completions documents 6 API parameters, grouped by what they control:
-
Length:
max_tokensdefaults to 1024, minimum 1;stop. -
Sampling:
temperaturedefaults to 0.5, range 0–2;top_pdefaults to 1, maximum 1;frequency_penaltydefaults to 0, range -2–2;presence_penaltydefaults to 0, range -2–2.
Frequently asked questions
- Which API parameters does NVIDIA Llama 3.1 NemoGuard 8B Topic Control support?
- NVIDIA Llama 3.1 NemoGuard 8B Topic Control accepts 6 API parameters in the request body: temperature, top_p, max_tokens, frequency_penalty, presence_penalty, stop.
- What is the default temperature for NVIDIA Llama 3.1 NemoGuard 8B Topic Control?
- The default temperature for NVIDIA Llama 3.1 NemoGuard 8B Topic Control is 0.5, within a valid range of 0 to 2.
- What is the default top_p for NVIDIA Llama 3.1 NemoGuard 8B Topic Control?
- The default top_p for NVIDIA Llama 3.1 NemoGuard 8B Topic Control is 1, with a maximum of 1.
- What is the default max_tokens for NVIDIA Llama 3.1 NemoGuard 8B Topic Control?
- The default max_tokens for NVIDIA Llama 3.1 NemoGuard 8B Topic Control is 1024, with a minimum of 1.
Resources
Other NVIDIA models
DeepSeek v4 Flash 0731
7 params
View
DeepSeek v4 Pro 0813
7 params
View
Gemma 4 31B IT
7 params
View
GLiNER PII
4 params
View
GLM-5.3 Flash
7 params
View
GPT-OSS 120B
7 params
View
GPT-OSS 20B
7 params
View
Kimi K3
7 params
View
Laguna Xs 2.1
7 params
View
Llama 3.1 Nemotron Nano 8B v1
7 params
View
Llama 3.1 Nemotron Safety Guard 8B v3
1 param
View
Llama 3.1 Nemotron Ultra 253B v1
7 params
View
Llama 3.3 Nemotron Super 49B v1
7 params
View
Llama 3.3 Nemotron Super 49B v1.5
7 params
View
MiniMax M3
7 params
View
Muse Glimmer 30B
7 params
View
NemoGuard Jailbreak Detect
0 params
View
Nemotron 3 Nano 30B A3B
5 params
View
Nemotron 3 Super 120B A12B
7 params
View
Nemotron 3 Ultra
Subscription
6 params
View
Nemotron 3 Ultra 550B A55B
7 params
View
Nemotron Content Safety Reasoning 4B
5 params
View
Nemotron Mini 4B Instruct
6 params
View
Riva Translate 4B Instruct v1.1
6 params
View
USDCode Llama 3.1 70B Instruct
4 params
View