QDNASales and integration of LLM inference and training platforms, on-premises or hybrid

Kimi K2.7 Code

Moonshot

Kimi K2.7 Code

Kimi K2.7 Code delivers stability in agentic context. It maintains reliable tool calls over long sessions, supports the MCP protocol, and achieves strong results on SWE-Bench. Its size is close to one trillion parameters.

Key points: Agentic · Reliable tool calls · MCP protocol · SWE-Bench · 1000B.

Architecture

Kimi K2.7 Code is a mixture-of-experts model close to one trillion parameters, tuned for stability in agentic context.

Strengths

It maintains reliable tool calls over long sessions, supports the MCP protocol, and achieves strong results on SWE-Bench.

Use cases

Kimi K2.7 Code serves coding agents that chain many tool calls without drifting, on long code tasks.

Deployment

Served by vLLM on a GPU server, or consumed via API. It integrates with OpenCode and Hermes through the LiteLLM gateway.

Moonshot unveiled Kimi K3 on 16 July 2026: 2,800 billion parameters, native vision, and always-on reasoning, aimed at enterprise AI. K2.7 Code remains the model of choice for coding assistants.

All models on the platform are interchangeable through a single gateway: switching models is a one-line configuration change, with no code to rewrite. The priority remains local execution.

Deploy this model on your infrastructure

On your own hardware, with your data staying in-house.

Book a call