modelparams.dev
Moonshot AI Deprecated 11 params

Moonshot AI Moonshot v1 128K API parameters

These are the API parameters modelparams.dev tracks for Moonshot AI Moonshot v1 128K — the settings you send in a request. Each row gives the type, default, valid range or values, and the conditions that gate it. It's the same data the JSON API serves.

This model is deprecated. Shutdown date: 2026-08-31. Suggested replacement: moonshot/kimi-k2.6.

After the parameter count instead — how many weights Moonshot v1 128K has? That's a different number, and we don't track it. Here's the difference.

Length 3 params
Parameter Type Default Description Condition
Max completion tokens
max_completion_tokens
integer (1…+∞) Maximum number of tokens to generate in the chat completion.
Stop sequence
stop
string Stops generation when this sequence is produced. Moonshot accepts up to 5 sequences of at most 32 bytes each.
Number of completions
n
integer (1…5) 1 How many chat completion choices to generate for the request.
Sampling 4 params
Parameter Type Default Description Condition
Temperature
temperature
number (0…1 step 0.1) 0.3 Controls randomness. Lower values make outputs more focused; higher values make them more varied.
Top P
top_p
number (0…1 step 0.01) 1 Controls nucleus sampling by limiting generation to tokens within the selected cumulative probability.
Presence penalty
presence_penalty
number (-2…2 step 0.1) 0 Penalizes tokens that have already appeared, encouraging the model to talk about new topics.
Frequency penalty
frequency_penalty
number (-2…2 step 0.1) 0 Penalizes tokens by how often they have appeared, reducing verbatim repetition.
Tools 1 param
Parameter Type Default Description Condition
Tool choice
tool_choice
enum (auto | none | required) "auto" Controls tool calling: auto lets the model decide, none blocks tool calls, and required forces one.
Output 1 param
Parameter Type Default Description Condition
Response format
response_format.type
enum (text | json_object) "text" Forces the response into plain text or a JSON object.
Observability 2 params
Parameter Type Default Description Condition
Log probabilities
logprobs
boolean false Controls whether the response includes log probabilities for the generated tokens.
Top log probabilities
top_logprobs
integer (0…20) Number of most likely tokens to return a log probability for at each position.
Only when logprobs = true

Moonshot AI Moonshot v1 128K API parameters in brief

Moonshot AI Moonshot v1 128K documents 11 API parameters, grouped by what they control:

Frequently asked questions

Which API parameters does Moonshot AI Moonshot v1 128K support?
Moonshot AI Moonshot v1 128K accepts 11 API parameters in the request body: max_completion_tokens, stop, temperature, top_p, n, presence_penalty, and more.
What is the default temperature for Moonshot AI Moonshot v1 128K?
The default temperature for Moonshot AI Moonshot v1 128K is 0.3, within a valid range of 0 to 1.
What is the default top_p for Moonshot AI Moonshot v1 128K?
The default top_p for Moonshot AI Moonshot v1 128K is 1, within a valid range of 0 to 1.

Resources

All Moonshot AI models Glossary Full catalog

Moonshot v1 128K — JSON

The full model definition as served by the API. Copy it or open the endpoint directly.

{
  "$schema": "https://modelparams.dev/api/v1/schema.json",
  "provider": "moonshot",
  "authType": "api_key",
  "model": "moonshot-v1-128k",
  "status": "deprecated",
  "replacement": "moonshot/kimi-k2.6",
  "shutdownOn": "2026-08-31",
  "params": [
    {
      "path": "max_completion_tokens",
      "label": "Max tokens",
      "description": "Maximum number of tokens to generate in the chat completion.",
      "group": "generation_length",
      "type": "integer",
      "range": {
        "min": 1
      }
    },
    {
      "path": "stop",
      "label": "Stop sequence",
      "description": "Stops generation when this sequence is produced. Moonshot accepts up to 5 sequences of at most 32 bytes each.",
      "group": "generation_length",
      "type": "string"
    },
    {
      "path": "temperature",
      "label": "Temperature",
      "description": "Controls randomness. Lower values make outputs more focused; higher values make them more varied.",
      "group": "sampling",
      "type": "number",
      "default": 0.3,
      "range": {
        "min": 0,
        "max": 1,
        "step": 0.1
      }
    },
    {
      "path": "top_p",
      "label": "Top P",
      "description": "Controls nucleus sampling by limiting generation to tokens within the selected cumulative probability.",
      "group": "sampling",
      "type": "number",
      "default": 1,
      "range": {
        "min": 0,
        "max": 1,
        "step": 0.01
      }
    },
    {
      "path": "n",
      "label": "Number of completions",
      "description": "How many chat completion choices to generate for the request.",
      "group": "generation_length",
      "type": "integer",
      "default": 1,
      "range": {
        "min": 1,
        "max": 5
      }
    },
    {
      "path": "presence_penalty",
      "label": "Presence penalty",
      "description": "Penalizes tokens that have already appeared, encouraging the model to talk about new topics.",
      "group": "sampling",
      "type": "number",
      "default": 0,
      "range": {
        "min": -2,
        "max": 2,
        "step": 0.1
      }
    },
    {
      "path": "frequency_penalty",
      "label": "Frequency penalty",
      "description": "Penalizes tokens by how often they have appeared, reducing verbatim repetition.",
      "group": "sampling",
      "type": "number",
      "default": 0,
      "range": {
        "min": -2,
        "max": 2,
        "step": 0.1
      }
    },
    {
      "path": "response_format.type",
      "label": "Response format",
      "description": "Forces the response into plain text or a JSON object.",
      "group": "output_format",
      "type": "enum",
      "default": "text",
      "values": [
        "text",
        "json_object"
      ]
    },
    {
      "path": "tool_choice",
      "label": "Tool choice",
      "description": "Controls tool calling: auto lets the model decide, none blocks tool calls, and required forces one.",
      "group": "tooling",
      "type": "enum",
      "default": "auto",
      "values": [
        "auto",
        "none",
        "required"
      ]
    },
    {
      "path": "logprobs",
      "label": "Log probabilities",
      "description": "Controls whether the response includes log probabilities for the generated tokens.",
      "group": "observability",
      "type": "boolean",
      "default": false
    },
    {
      "path": "top_logprobs",
      "label": "Top log probabilities",
      "description": "Number of most likely tokens to return a log probability for at each position.",
      "group": "observability",
      "applicability": {
        "only": {
          "logprobs": true
        }
      },
      "type": "integer",
      "range": {
        "min": 0,
        "max": 20
      }
    }
  ]
}

Other Moonshot AI models

How to use

Building with an AI agent? Hit Copy to grab this whole guide as Markdown and paste it in — or point your agent straight at /llms.txt.

modelparams.dev is an open, community-maintained catalog of model parameters. Each entry shows the knobs you can turn — type, default, range, and the conditions that gate it.

The same model accessed via an API key and via a subscription usually exposes a different set of parameters. We list both as separate entries so the data stays honest.

Catalog API

The full catalog is static JSON, CORS-enabled, served from the edge.

curl https://modelparams.dev/api/v1/models.json

Each entry is keyed by provider/model for API-key variants; subscription variants append -subscription.

If you only need the params for one model contract, use the providerless endpoint. Subscription contracts are model slugs with -subscription.

curl https://modelparams.dev/api/v1/models/openai/gpt-5.5.json
curl https://modelparams.dev/api/v1/models/openai/gpt-5.5-subscription.json

Single model

curl https://modelparams.dev/api/v1/models/anthropic/claude-opus-4-7.json
curl https://modelparams.dev/api/v1/models/anthropic/claude-opus-4-7-subscription.json

JSON Schema

Every entry validates against a JSON Schema you can use in your editor or pipeline.

curl https://modelparams.dev/api/v1/schema.json

Add this header to any YAML you author for autocomplete in VS Code:

# yaml-language-server: $schema=https://modelparams.dev/api/v1/schema.json

Logos

Provider logos are available at /assets/logos/{provider}.svg where {provider} is the provider slug. They use currentColor so they inherit your text color.

curl https://modelparams.dev/assets/logos/anthropic.svg

Logos are sourced from the models.dev repo (MIT) and used under nominative fair use.

Contribute

The data lives in YAML under models/{provider}/{model}-{auth}.yaml in the GitHub repo. Open a PR; CI validates against the schema and rebuilds.

Edit on GitHub MIT licensed