Browse all Moonshot AI text models
Provider logo

Kimi K2.7 Code High-Speed

moonshotai/kimi-k2.7-code-highspeed
Back
Provider logo

Kimi K2.7 Code High-Speed

moonshotai/kimi-k2.7-code-highspeed
Back

Kimi K2.7 Code High-Speed is the accelerated coding-focused variant tuned for fast agentic software engineering. It targets roughly 180 tokens per second, with short-context responses reaching up to about 260 tokens per second for rapid coding iterations.

Added Jun 15, 2026

Context Window

262.1K

Max Output

65.5K

Input Price (Auto)

$1.90/1M

Output Price (Auto)

$8.00/1M

Cache Read (Auto)

$0.32/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

25.8

Better than 78% of models compared

Coding Index

60.8

Better than 74% of models compared

Agentic work

AutomationBench-AA

Workflow automation with guardrail penalties

24.5%

Better than 48% of models compared

Harvey LAB-AA

Legal agentic work criterion pass rate

85.0%

Better than 44% of models compared

AA-Briefcase

Agentic knowledge work (Elo)

854 Elo

Better than 39% of models compared

GDPval-AA v2

Economically valuable tasks (Elo)

1025 Elo

Better than 52% of models compared

Document reasoning

GDP.pdf

Professional PDF reasoning: all-pass rate

11.2%

Better than 46% of models compared

AA-LCR v1.1

Long context reasoning with updated grading

79.3%

Better than 85% of models compared

MLCR-AA

Medical long-context reasoning

16.1%

Better than 55% of models compared

Reasoning

HLE

Humanity's Last Exam

35.0%

Better than 85% of models compared

IFBench

Instruction-following benchmark

63.1%

Better than 74% of models compared

Coding

Terminal-Bench v4.0

Practical coding and terminal tasks

1.0%

Better than 44% of models compared

SciCode

Python programming for scientific computing

47.8%

Better than 50% of models compared

Legacy benchmarks

GPQA Diamond (legacy)

Graduate-level scientific reasoning

89.6%

Better than 88% of models compared

Terminal-Bench Hard (legacy)

Agentic coding and terminal use

44.7%

Better than 92% of models compared

T²-Bench Telecom (legacy)

Conversational AI agents in dual-control scenarios

90.1%

Better than 84% of models compared

AA-LCR (unversioned / legacy)

Long context reasoning evaluation

79.3%

Better than 85% of models compared

Last updated Sep 24, 2026

Artificial Analysis

Providers

Provider information for this model’s automatic routing. These routes cannot be selected individually.

Loading provider options…

Compare Kimi K2.7 Code High-Speed with similar models from the same provider or model family.

Kimi K2.7 Code

moonshotai/kimi-k2.7-code

Kimi K2.7 Code is Moonshot AI's coding-focused agentic model built for long-horizon software engineering workflows. It supports native image input, tool calling, and forced thinking mode; instant/non-thinking mode is not supported.

Kimi K3

moonshotai/kimi-k3

Kimi K3 is Moonshot AI's open-weight, always-thinking multimodal model for long-context reasoning, coding, tool use, and native image/video understanding.

Kimi Latest

moonshotai/kimi-latest

Compatibility alias that routes to the newest Kimi model. Currently routes to Kimi K3.

Kimi K2.6

moonshotai/kimi-k2.6

Kimi K2.6 is an open-source, native multimodal agentic model built for long-horizon coding, coding-driven design, and large-scale task orchestration. It can turn simple prompts and visual inputs into production-ready interfaces and full-stack workflows, and is designed to coordinate complex multi-agent plans with thousands of steps across code, documents, and spreadsheets.

Kimi K2.6 Thinking

moonshotai/kimi-k2.6:thinking

Kimi K2.6 Thinking is the reasoning-optimized K2.6 variant for deeper multi-step planning and execution. It is tuned for long-horizon coding and design workflows, including complex orchestration across many specialized sub-agents and autonomous end-to-end output generation.

Kimi K2.5

moonshotai/kimi-k2.5

Kimi K2.5 is Moonshot AI's native multimodal model built on Kimi K2 with ~15T mixed visual and text tokens, delivering strong general reasoning, visual coding, and agentic tool-calling. This route uses instant (non-thinking) mode for faster responses.