Kimi K2.7-Code cuts thinking tokens 30% — but practitioners say the benchmarks don't check out
Source:
ventureBeat
June 12, 2026 · 14:55
Moonshot AI released Kimi K2.7-Code this week, an open-source update to its K2 coding model family, claiming leaner reasoning and double-digit performance gains.K2.7-Code is built on the same trillion-parameter mixture-of-experts architecture as its predecessor K2.6, and drops in via an OpenAI-compatible API — which matters for teams already running K2.6 in production gateways.When K2.6 launched in April, it topped OpenRouter's weekly LLM leaderboard — a ranking based on actual API routing decisions by developers, not self-reported benchmark scores.Moonshot AI says K2.7-Code addresses what it calls "overthinking," reducing thinking-token usage by 30% compared to K2.6 — a number that would directly affect inference costs for teams running agentic workflows. Whether that efficiency gain hold…
The original article opens on the publisher's website.
More from Media
View topic →TikTok explores peer-to-peer payments via DMs, report says
techcrunch
Aug 18, 2026 · 13:03
Social media on trial as $200bn case against Facebook and Instagram begins
guardianTech
Aug 18, 2026 · 11:27
'Profits won.' The child safety trial against Meta kicks off in federal court
nprNews
Aug 18, 2026 · 10:39
‘Can’t enjoy your pint’: Wetherspoons customers welcome ban on loud phones
guardianTech
Aug 18, 2026 · 10:32
Bluesky says its recent outage was caused by another DDoS attack
techcrunch
Aug 18, 2026 · 08:59