NVIDIA
7 params
NVIDIA DeepSeek v4 Flash 0731 API parameters
These are the API parameters modelparams.dev tracks for NVIDIA DeepSeek v4 Flash 0731 — the settings you send in a request. Each row gives the type, default, valid range or values, and the conditions that gate it. It's the same data the JSON API serves.
After the parameter count instead — how many weights DeepSeek v4 Flash 0731 has? That's a different number, and we don't track it. Here's the difference.
Length
2 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Max tokens
max_tokens
|
integer (1…16384) | 4096 | Maximum number of tokens to generate. Generation stops when this limit is reached. | — |
|
Stop
stop
|
string | — | A string or list of strings where the API will stop generating further tokens. The returned text will not contain the stop sequence. | — |
Sampling
5 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Temperature
temperature
|
number (0…1) | 0.6 | Controls randomness. Lower values make outputs more focused; higher values make them more varied. Not recommended to modify both temperature and top_p in the same call. | — |
|
Top P
top_p
|
number (-∞…1) | 0.95 | Controls nucleus sampling by limiting generation to tokens within the selected cumulative probability. Not recommended to modify both temperature and top_p in the same call. | — |
|
Frequency penalty
frequency_penalty
|
number (-2…2) | 0 | Penalizes new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. | — |
|
Presence penalty
presence_penalty
|
number (-2…2) | 0 | Positive values penalize new tokens based on whether they appear in the text so far, increasing the model's likelihood to talk about new topics. | — |
|
Seed
seed
|
integer (0…18446744073709552000) | 0 | Best-effort deterministic sampling seed. Changing the seed produces a different response with similar characteristics. Fix the seed to reproduce results. | — |
NVIDIA DeepSeek v4 Flash 0731 API parameters in brief
NVIDIA DeepSeek v4 Flash 0731 documents 7 API parameters, grouped by what they control:
-
Length:
max_tokensdefaults to 4096, range 1–16384;stop. -
Sampling:
temperaturedefaults to 0.6, range 0–1;top_pdefaults to 0.95, maximum 1;frequency_penaltydefaults to 0, range -2–2;presence_penaltydefaults to 0, range -2–2;seeddefaults to 0, range 0–18446744073709552000.
Frequently asked questions
- Which API parameters does NVIDIA DeepSeek v4 Flash 0731 support?
- NVIDIA DeepSeek v4 Flash 0731 accepts 7 API parameters in the request body: temperature, top_p, max_tokens, frequency_penalty, presence_penalty, seed, and more.
- What is the default temperature for NVIDIA DeepSeek v4 Flash 0731?
- The default temperature for NVIDIA DeepSeek v4 Flash 0731 is 0.6, within a valid range of 0 to 1.
- What is the default top_p for NVIDIA DeepSeek v4 Flash 0731?
- The default top_p for NVIDIA DeepSeek v4 Flash 0731 is 0.95, with a maximum of 1.
- What is the default max_tokens for NVIDIA DeepSeek v4 Flash 0731?
- The default max_tokens for NVIDIA DeepSeek v4 Flash 0731 is 4096, within a valid range of 1 to 16384.
Resources
Other NVIDIA models
GLiNER PII
4 params
View
Llama 3.1 NemoGuard 8B Topic Control
6 params
View
Llama 3.1 Nemotron Nano 8B v1
7 params
View
Llama 3.1 Nemotron Safety Guard 8B v3
1 param
View
Llama 3.1 Nemotron Ultra 253B v1
7 params
View
Llama 3.3 Nemotron Super 49B v1
7 params
View
Llama 3.3 Nemotron Super 49B v1.5
7 params
View
NemoGuard Jailbreak Detect
0 params
View
Nemotron 3 Nano 30B A3B
5 params
View
Nemotron 3 Super 120B A12B
7 params
View
Nemotron 3 Ultra
Subscription
6 params
View
Nemotron 3 Ultra 550B A55B
7 params
View
Nemotron Content Safety Reasoning 4B
5 params
View
Nemotron Mini 4B Instruct
6 params
View
Riva Translate 4B Instruct v1.1
6 params
View
USDCode Llama 3.1 70B Instruct
4 params
View