Skip to content
Draft
Changes from all commits
Commits
File filter

Filter by extension

Filter by extension

Conversations
Failed to load comments.
Loading
Jump to
Jump to file
Failed to load files.
Loading
Diff view
Diff view
67 changes: 32 additions & 35 deletions docs/providers/kimi-code.md
Original file line number Diff line number Diff line change
@@ -1,26 +1,27 @@
---
sidebar_label: Kimi Code
description: Use Kimi's coding models with your Kimi Code subscription (OAuth) or an API key, with configurable reasoning effort.
description: Connect Zoo Code to a Kimi membership or Kimi Code API key and choose between Kimi K3 and K2.7 Code models.
keywords:
- kimi code
- kimi subscription
- kimi k3
- moonshot
- kimi k2.7 code
- moonshot ai
- zoo code
- api provider
- oauth
- reasoning effort
---

# Kimi Code Provider

Use Kimi's coding models (Kimi K3) through your Kimi Code subscription with OAuth device-flow sign-in, or with a Kimi Code API key. Model metadata is discovered automatically after you authenticate.
The Kimi Code provider connects Zoo Code to the coding models included with a Kimi membership. You can sign in with your Kimi subscription or use a Kimi Code API key. Model metadata is discovered automatically after you authenticate.

:::info Setup Required
1. **Select "Kimi Code"** as your provider in Zoo Code settings
2. **Authenticate**:
- **Kimi Code subscription (OAuth)**: Click "Sign in", then approve the device code shown in Zoo Code at the Kimi authorization page
- **API key**: Paste your Kimi Code API key
3. **Pick a model**: The model list refreshes automatically once you're authenticated
1. Select **Kimi Code** as your provider in Zoo Code settings.
2. Choose an authentication method:
- **Kimi Code subscription (OAuth):** Click **Sign in**, then approve the device code at the Kimi authorization page.
- **API key:** Create a key in the [Kimi Code Console](https://www.kimi.com/code/console) and paste it into Zoo Code.
3. Pick a model available to your membership tier. The model list refreshes automatically after authentication.
:::

**Website:** [https://www.kimi.com/code](https://www.kimi.com/code)
Expand All @@ -29,42 +30,38 @@ Use Kimi's coding models (Kimi K3) through your Kimi Code subscription with OAut

## Available Models

Kimi Code's coding model offers a large context window and up to 32,768 max output tokens. The exact model list is fetched from your account after sign-in; use "Refresh Models" in the provider settings to update it.
| Model ID | Model | Context | Reasoning behavior |
| --- | --- | --- | --- |
| `k3` | Kimi K3 | Up to 1M tokens, depending on membership | `low`, `high`, or `max` reasoning effort; defaults to `high` |
| `k3-256k` | Kimi K3 256K | 256K tokens | `low`, `high`, or `max` reasoning effort; defaults to `high` |
| `kimi-for-coding` | Kimi K2.7 Code | 256K tokens | Thinking is always enabled and preserved across turns |
| `kimi-for-coding-highspeed` | Kimi K2.7 Code HighSpeed | 256K tokens | Same thinking behavior with faster output for eligible memberships |

---

## Configuration

### Authentication Method
- **Kimi Code subscription (OAuth)**: Device-flow sign-in with automatic token refresh
- **API key**: Direct key-based access
Zoo Code reads current model capacity from Kimi Code when available. K3 uses a 131K default output limit through the direct Kimi Code API; this differs from limits imposed by third-party routing providers.

### Reasoning Effort
:::tip Choosing a K3 model
Use `k3-256k` for routine coding tasks when you do not need a larger context window. Kimi reports that it provides the same results within 256K while consuming less membership quota than `k3`.
:::

Kimi K3 always thinks; you control how hard with the **Model Reasoning Effort** dropdown in the provider settings:
---

- **Low**: Faster responses with reduced reasoning
- **High**: Deeper reasoning
- **Max** (default): Maximum reasoning effort
## Reasoning and Thinking

The selected effort is sent as the `reasoning_effort` parameter on every request. Thinking cannot be turned off for Kimi K3, so there is no "None" option.
### Kimi K3

---
K3 always thinks and supports **Low**, **High**, and **Max** reasoning effort. Zoo Code defaults to **High** and sends the selected level as `reasoning_effort`. Changing the effort within a session can invalidate Kimi's context cache.

## Key Features
### Kimi K2.7 Code

- **OAuth 2.0 device flow**: Secure sign-in with automatic token refresh and transparent retry on expired tokens
- **API key support**: Alternative authentication for headless setups
- **Automatic model discovery**: Model list and capabilities fetched from your account
- **Configurable reasoning effort**: Low, High, or Max (default Max)
K2.7 Code uses preserved thinking rather than configurable reasoning effort. Zoo Code keeps thinking enabled and carries the model's reasoning context across tool calls and turns.

---

## Common Issues
## Notes

**"Not authenticated with Kimi Code"**
- Sign in from the Kimi Code provider settings (OAuth), or switch to the API key method
- Start a new task after switching model IDs to avoid carrying context cached for a different model.
- Model availability and maximum K3 context depend on your Kimi membership tier.
- Kimi Code is separate from Zoo Code's **Moonshot** provider, which uses Moonshot's pay-as-you-go API.
- With OAuth, Zoo Code refreshes expired tokens and retries an unauthorized request once automatically.

**"401 Unauthorized"**
- With OAuth, Zoo Code refreshes the token and retries once automatically
- If it persists, sign out and sign in again
For current model and membership details, see the [Kimi Code model documentation](https://www.kimi.com/code/docs/en/kimi-code/models.html).
Loading