NVIDIA
4 params
NVIDIA USDCode Llama 3.1 70B Instruct API parameters
These are the API parameters modelparams.dev tracks for NVIDIA USDCode Llama 3.1 70B Instruct — the settings you send in a request. Each row gives the type, default, valid range or values, and the conditions that gate it. It's the same data the JSON API serves.
After the parameter count instead — how many weights USDCode Llama 3.1 70B Instruct has? That's a different number, and we don't track it. Here's the difference.
Length
1 param
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Max tokens
max_tokens
|
integer (1…2048) | 1024 | Maximum number of tokens to generate. Generation stops when this limit is reached. | — |
Sampling
2 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Temperature
temperature
|
number (0…1) | 0.1 | Controls randomness. Lower values make outputs more focused; higher values make them more varied. Not recommended to modify both temperature and top_p in the same call. | — |
|
Top P
top_p
|
number (-∞…1) | 1 | Controls nucleus sampling by limiting generation to tokens within the selected cumulative probability. Not recommended to modify both temperature and top_p in the same call. | — |
Metadata
1 param
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Expert type
expert_type
|
enum (auto | code | knowledge | helperfunction) | "auto" | The type of expert to use. 'knowledge' answers with USD knowledge, 'code' responds with vanilla OpenUSD code, 'helperfunction' uses high-level helper functions, and 'auto' lets the LLM determine which expert to use. | — |
NVIDIA USDCode Llama 3.1 70B Instruct API parameters in brief
NVIDIA USDCode Llama 3.1 70B Instruct documents 4 API parameters, grouped by what they control:
-
Length:
max_tokensdefaults to 1024, range 1–2048. -
Sampling:
temperaturedefaults to 0.1, range 0–1;top_pdefaults to 1, maximum 1. -
Metadata:
expert_typedefaults to "auto", accepts "auto", "code", "knowledge", "helperfunction".
Frequently asked questions
- Which API parameters does NVIDIA USDCode Llama 3.1 70B Instruct support?
- NVIDIA USDCode Llama 3.1 70B Instruct accepts 4 API parameters in the request body: temperature, top_p, max_tokens, expert_type.
- What is the default temperature for NVIDIA USDCode Llama 3.1 70B Instruct?
- The default temperature for NVIDIA USDCode Llama 3.1 70B Instruct is 0.1, within a valid range of 0 to 1.
- What is the default top_p for NVIDIA USDCode Llama 3.1 70B Instruct?
- The default top_p for NVIDIA USDCode Llama 3.1 70B Instruct is 1, with a maximum of 1.
- What is the default max_tokens for NVIDIA USDCode Llama 3.1 70B Instruct?
- The default max_tokens for NVIDIA USDCode Llama 3.1 70B Instruct is 1024, within a valid range of 1 to 2048.
Resources
Other NVIDIA models
GLiNER PII
4 params
View
Llama 3.1 NemoGuard 8B Topic Control
6 params
View
Llama 3.1 Nemotron Nano 8B v1
7 params
View
Llama 3.1 Nemotron Safety Guard 8B v3
1 param
View
Llama 3.1 Nemotron Ultra 253B v1
7 params
View
Llama 3.3 Nemotron Super 49B v1
7 params
View
Llama 3.3 Nemotron Super 49B v1.5
7 params
View
NemoGuard Jailbreak Detect
0 params
View
Nemotron 3 Nano 30B A3B
5 params
View
Nemotron 3 Super 120B A12B
7 params
View
Nemotron 3 Ultra
Subscription
6 params
View
Nemotron 3 Ultra 550B A55B
7 params
View
Nemotron Content Safety Reasoning 4B
5 params
View
Nemotron Mini 4B Instruct
6 params
View
Riva Translate 4B Instruct v1.1
6 params
View