Cerebras
10 params
Cerebras Zai GLM-4.7 parameters
These are the parameters modelparams.dev tracks for Cerebras Zai GLM-4.7. Each row gives the type, default, valid range or values, and the conditions that gate it. It's the same data the JSON API serves.
Length
2 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Max completion tokens
max_completion_tokens
|
integer (1…+∞) | — | Maximum number of output tokens the model may generate, including reasoning tokens. | — |
|
Stop
stop
|
string | — | A string or list of strings where the API will stop generating further tokens. Cerebras accepts up to four stop sequences. | — |
Sampling
5 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Temperature
temperature
|
number (0…2 step 0.1) | — | Controls randomness. Lower values make outputs more focused; higher values make them more varied. Adjust this or top_p, not both. | — |
|
Top P
top_p
|
number (0…1 step 0.01) | — | Controls nucleus sampling by limiting generation to tokens within the selected cumulative probability. | — |
|
Frequency penalty
frequency_penalty
|
number (-2…2 step 0.1) | 0 | Penalizes tokens by how often they have appeared, reducing verbatim repetition. | — |
|
Presence penalty
presence_penalty
|
number (-2…2 step 0.1) | 0 | Penalizes tokens that have already appeared, encouraging the model to introduce new topics. | — |
|
Seed
seed
|
integer | — | Seed used for best-effort deterministic sampling when reproducible outputs are desired. | — |
Reasoning
2 params
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Reasoning effort
reasoning_effort
|
enum (none | low | medium | high) | — | Controls how much reasoning the model performs before answering. 'none' disables reasoning. | — |
|
Clear thinking
clear_thinking
|
boolean | true | When true, the model's thinking from previous turns is excluded from the conversation context; when false, it is preserved, which is useful for agentic workflows. | — |
Output
1 param
| Parameter | Type | Default | Description | Condition |
|---|---|---|---|---|
|
Response format
response_format.type
|
enum (text | json_object) | "text" | Forces the response into plain text or a JSON object. | — |
Cerebras Zai GLM-4.7 API parameters in brief
Cerebras Zai GLM-4.7 documents 10 API parameters, grouped by what they control:
-
Length:
max_completion_tokensminimum 1;stop. -
Sampling:
temperaturerange 0–2;top_prange 0–1;frequency_penaltydefaults to 0, range -2–2;presence_penaltydefaults to 0, range -2–2;seed. -
Reasoning:
reasoning_effortaccepts "none", "low", "medium", "high";clear_thinkingdefaults to true. -
Output:
response_format.typedefaults to "text", accepts "text", "json_object".
Frequently asked questions
- How many parameters does Cerebras Zai GLM-4.7 accept?
- Cerebras Zai GLM-4.7 accepts 10 API parameters: max_completion_tokens, temperature, top_p, frequency_penalty, presence_penalty, seed, and more.