Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Research paper Kimi (Moonshot)

Kimi K2.5: Open-Source Multimodal Model with Agent Swarm and Vision-Coding Breakthroughs

AI By Crimson AI Kimi Blog 26 July 2026 · 14:32 33 views
Share: X Telegram

Moonshot AI launches Kimi K2.5, a powerful open-source multimodal model with state-of-the-art coding and vision capabilities, featuring a self-directed agent swarm of up to 100 sub-agents and parallel workflows across 1,500 tool calls.

Kimi K2.5: Open-Source Multimodal Model with Agent Swarm and Vision-Coding Breakthroughs

Key points

Moonshot AI has unveiled Kimi K2.5, described as the most powerful open-source model to date. Building on its predecessor Kimi K2, the new model underwent continued pretraining on approximately 15 trillion mixed visual and text tokens, resulting in a native multimodal architecture that excels in both coding and vision tasks.

A key innovation is the self-directed agent swarm paradigm. Kimi K2.5 can autonomously orchestrate up to 100 sub-agents, executing parallel workflows across up to 1,500 tool calls. This reduces execution time by up to 4.5x compared to a single-agent setup, without any predefined sub-agents or workflows. The agent swarm is trained using Parallel-Agent Reinforcement Learning (PARL), which employs staged reward shaping to encourage parallel execution and prevent serial collapse.

Kimi K2.5 is available via Kimi.com, the Kimi App, API, and Kimi Code. The platform now supports four modes: K2.5 Instant, K2.5 Thinking, K2.5 Agent, and K2.5 Agent Swarm (Beta). The Agent Swarm mode is currently in beta on Kimi.com, with free credits for high-tier paid users.

On agentic benchmarks including HLE, BrowseComp, and SWE-Verified, Kimi K2.5 delivers strong performance at a fraction of the cost. It is particularly noted for front-end development, capable of converting simple conversations into complete interactive interfaces with animations. The model also excels in image/video-to-code generation and visual debugging, thanks to massive-scale vision-text joint pre-training.

For real-world software engineering, Kimi K2.5 shows consistent improvements over K2 on the internal Kimi Code Bench. The new Kimi Code product, which is open-sourced, integrates with IDEs like VSCode and Cursor, and supports images and videos as inputs. Additionally, K2.5 Agent handles office productivity tasks such as document creation, spreadsheet modeling, and PDF editing, achieving 59.3% and 24.3% improvements over K2 Thinking on internal benchmarks.

BenchmarkImprovement over K2 Thinking
AI Office Benchmark59.3%
General Agent Benchmark24.3%
Source
Kimi (Moonshot) · Kimi Blog
Related news
Research paper
Kimi (Moonshot) 17 Aug 2026

Moonshot AI Unveils Kimi K2.5: Open-Source Visual Agentic Intelligence with Swarm Capabilities

Moonshot AI introduces Kimi K2.5, a native multimodal open-source model with advanced coding and vision, featuring a self-directed...

22
Research paper
Kimi (Moonshot) 17 Aug 2026

Moonshot AI Unveils WorldVQA Benchmark to Test Visual World Knowledge in Multimodal LLMs

Moonshot AI releases WorldVQA, a benchmark with 3,500 image-question pairs designed to measure factual visual knowledge in multimo...

18
Research paper
Kimi (Moonshot) 17 Aug 2026

Kimi Launches Agent Swarm: 100 AI Agents Self-Organize to Tackle Complex Tasks

Kimi (Moonshot) unveils Agent Swarm, a research preview that lets K2.5 deploy up to 100 parallel sub-agents that self-organize int...

15