Moonshot’s Kimi K2.7 Code Tops Open-Weight SWE-bench at 78.2% in Vals Benchmarks
Moonshot AI’s Kimi K2.7 Code scored 78.2% on SWE-bench and 67% on Terminal-Bench 2.1, making it the top open-weight model on both tests in ValsAI’s latest coding-benchmark evaluation. Vals said the next-closest open model on Terminal-Bench 2.1 scored 60.67%.
Vals also placed Kimi K2.7 Code third among open-weight models on Vibe Code Bench at 47.21%, up from 37.89% for Kimi K2.6, and second on ProgramBench with a 49.44% raw pass rate. The new scores add another public comparison point after ErdosBench earlier ranked the coding model second on a 14-problem smoke test behind Claude Fable-5-max and ahead of OpenAI’s GPT-5.5 xhigh.
From the sources (25 posts)
@kimi_moonshot🌘 Kimi-K2.7-Code, our latest coding model, is now released and open-sourced! 🔷 Improved coding & agent performance over K2.6: +21.8% on Kimi Code Bench v2, +11.0% on Program Bench, and +31.5% on MLS Bench Lite. 🔷 Reasoning efficiency: Less
@kimi_moonshot🎸 We're also launching the Kimi Code Beta Program today. Apply now if you'd like to try upcoming models and features before public release 👉
@stevibeKimi K2.7-Code!
@kimi_moonshot🔗 Weights & code:
@teortaxestexI want to see this compared with Composer 2.5 Like, really hard Cursor has a ton of proprietary data, a large head start, and threw a Colossus at RLing Kimi K2.5 checkpoint. What is the gap now?
@vanstriendanielRT @Kimi_Moonshot: 🔗 Weights & code:
@crystalsssupless overthinking 👀
@teortaxestexHUGE @htihle and others will get to test this soon I hope
@zephyr_z9Great work from KIMI
@eliebakouchRT @Kimi_Moonshot: 🌘 Kimi-K2.7-Code, our latest coding model, is now released and open-sourced! 🔷 Improved coding & agent performance over…
@code_starIt would be really cool if the top Chinese labs could properly benchmark and evaluate Fable/Mythos to show how it compares to their releases. I guess that isn’t possible though.
@thezachmuellerRT @Kimi_Moonshot: 🌘 Kimi-K2.7-Code, our latest coding model, is now released and open-sourced! 🔷 Improved coding & agent performance over…
@kimmonismusMoonshot just released Kimi-K2.7 code, a huge upgrade to Kimi-K2.6! Big jump over K2.6: +21.8% on Kimi Code Bench v2 +11.0% on Program Bench +31.5% on MLS Bench Lite It also uses 30% fewer reasoning tokens, follows instructions better, and
@andrewcurran_RT @Kimi_Moonshot: 🌘 Kimi-K2.7-Code, our latest coding model, is now released and open-sourced! 🔷 Improved coding & agent performance over…
@nrehiew_I think K2.6 is the third best model on the planet - albeit one of the slowest. With the overthinking fixed and overall performance improvements, can’t wait to use this
@vllm_project🎉 Congrats to @Kimi_Moonshot on Kimi K2.7-Code, a coding-focused agentic model built on K2.6. ✨ 1T-parameter Mixture-of-Experts, 32B active per token ✨ MLA attention with a 256K-token context window ✨ ~30% fewer thinking tokens than K2.6 f
@clementdelangueRT @Kimi_Moonshot: 🔗 Weights & code:
@kimi_moonshotRT @vllm_project: 🎉 Congrats to @Kimi_Moonshot on Kimi K2.7-Code, a coding-focused agentic model built on K2.6. ✨ 1T-parameter Mixture-of-…
@matthewbermanNearly frontier open source coding model. Congrats to the Kimi team. I would love to see this tested against DeepSWE
@thezachmuellerRT @elliotarledge: I benchmarked Kimi K2.7-Code (1T MoE, coding-specialized, just dropped) on KernelBench-Hard, against its general predece…
@threepointoneRT @michellechen: k2.7 code live on workers ai
@thdxrRT @opencode: Kimi 2.7 Code now available in Go text · image · optimized for coding similar pricing as 2.6
@tekniumKimi 2.7 Coder now available in Hermes Agent! No need to update - your model picker will pick it up automatically.
@ollama.@Kimi_Moonshot's kimi-k2.7-code is now available on Ollama's cloud! On Ollama's cloud, this model is hosted in the US on the latest NVIDIA B300 datacenter GPUs. Your data stays private and is never trained on. ❤️❤️❤️ Try the model: C
@testingcatalogKimi-K2.7-Code is now available on AI/ML API 👀 > Kimi K2.7 Code is the latest agentic coding model from Kimi AI that supports extended reasoning and tool use. > AI/ML API is a single gateway to Chat, Reasoning, Image, Video, Audio, Voice