Moonshot AI launches Kimi K2.7-Code HighSpeed
Moonshot AI has introduced a new high-speed mode for its open-source multimodal coding model, Kimi K2.7-Code, delivering up to 6× faster generation speeds. The update achieves up to 260 tokens per second on shorter-context tasks and is currently rolling out to Kimi Code Beta.
High-speed inference is the key to making AI-driven software engineering feel truly interactive and seamless, and this update directly targets the latency bottleneck of complex coding tasks.
- –**Inference Speedup:** Speeds of 180–260 tokens per second match or exceed dedicated inference hardware platforms, making real-time autocomplete and chat highly responsive.
- –**Multimodal Support:** The model retains its multimodal capabilities, permitting developers to include visual UI layouts and diagrams in their prompts at speed.
- –**Beta Deployment:** Integrating this variant directly into Kimi Code Beta facilitates immediate developer testing and real-world performance feedback.
DISCOVERED
97d ago
2026-06-15
PUBLISHED
97d ago
2026-06-15
RELEVANCE
AUTHOR
Kimi_Moonshot