Moonshot
Kimi API
Moonshot's Kimi models — a frontier tier and a coding-specialised one.
2
Modelle in der Reihe
Kimi splits by specialisation rather than by size: K3 is Moonshot's frontier model for long-context reasoning and agentic tool use, and K2.7 Code is tuned for repository-scale edits and agentic software tasks. The two are priced within about six percent of each other on input, so the choice is about what the model was trained for, not about budget.
Auf einen Blick
- Ab
- $1.48 / 1 Mio. Tokens
Modelle dieser Familie
Preise
| Modell | Option | Preis |
|---|---|---|
| Kimi K3 | input | $1.48/ 1 Mio. Tokens |
| Kimi K3 | output | $7.4/ 1 Mio. Tokens |
| Kimi K3 | cacheRead | $0.148/ 1 Mio. Tokens |
| Kimi K2.7 Code | input | $1.57/ 1 Mio. Tokens |
| Kimi K2.7 Code | output | $7.13/ 1 Mio. Tokens |
| Kimi K2.7 Code | cacheRead | $0.265/ 1 Mio. Tokens |
Welches Modell wählen
Kimi K3
Long-context reasoning and agentic tool use — the general frontier tier.
Kimi K2.7 Code
Repository-scale code edits and agentic software work, where a coding-tuned model beats a general one.
Stärken
- A coding-specialised tier that is not just the general model with a different prompt.
- Input prices sit within about six percent of each other, so specialisation does not cost extra.
- Prompt caching on both tiers.
- OpenAI-protocol compatible.
Einschränkungen
- K3's cache reads are discounted to a tenth of input while K2.7 Code only reaches about a sixth, so caching pays back differently per tier.
- Output costs four and a half to five times input on both tiers.
- Only two tiers, both at a similar price point — no cheap volume option.
Die Beschreibung des Anbieters zu dieser Reihe: Moonshot AI — docs