Kimi K2.7 Code API: Moonshot's Coding Model, Standard and High-Speed Tiers (2026)
K2.7 Code is Moonshot's programming-tuned model: 256K context, hardened instruction-following. Standard at ¥6.17/¥25.65 per M tokens, high-speed at double for interactive work. How it fits alongside K3 and Claude.
What K2.7 Code is good at
Two things: instruction-following that survives long context (models drift when you stuff a large project in — this one is tuned not to), and task completion rate — finishing the code rather than sketching it. Thinking-mode means it plans before writing, which suits agent-executor roles.
Is the high-speed tier worth 2x?
Depends where latency sits. Batch jobs and overnight pipelines: standard, half the price. Interactive coding in Cursor/Claude Code or multi-turn agent loops: high-speed — waiting time is developer cost, and doubled output speed usually beats doubled token price. Same model, switch per workload.
Setup and where it fits
OpenAI-compatible: base_url https://cocodot.co/api/ai/v1, model kimi-k2.7-code or kimi-k2.7-code-highspeed. A pragmatic stack: cheap tier for completions (Haiku/Sonnet 5/K2.7), Opus 5 for hard problems; K2.7 is the strongest Chinese-model alternative for code when you want a non-Anthropic lane — A/B it on your own workload with the same key.
Kimi coding lineup on cocodot (CNY per M tokens)
| Model | In / Out | Context | Position |
|---|---|---|---|
| K2.7 Code standard | ¥6.17 / ¥25.65 | 256K | coding workhorse |
| K2.7 Code high-speed | ¥12.35 / ¥51.3 | 256K | same model, fast serving |
| K3 (reference) | ¥20 / ¥100 (list) | 1M | repo-scale context |