Learn About Moonshot AI's AI Products
Create a free account to access in-depth lessons on each tool and model.
Start Learning Free📋About Moonshot AI
Updated August 1, 2026Moonshot AI (月之暗面) is a Beijing-based frontier AI lab founded in March 2023 by Yang Zhilin, previously a co-founder of Zhipu AI. The company became China's fastest decacorn, reaching a $10 billion valuation in February 2026, and has raised through a rapid succession of rounds since — a $500 million Series C in January 2026 from IDG Capital, Alibaba, and Tencent, then a $2 billion round in May 2026 led by Meituan's Long-Z Investments arm with Tsinghua Capital, China Mobile, and CPE Yuanfeng participating. As of July 2026 the lab is reported to be raising again at a materially higher valuation.
Moonshot's product line runs on the Kimi brand: the Kimi chatbot on web and mobile, Kimi Code (a coding agent with terminal, VS Code, Cursor, and Zed integrations, plus MCP support), and Kimi Work, a desktop agent for macOS and Windows that runs the lab's Agent Swarm system on local hardware. The chatbot first gained prominence on long-context document analysis, and long context has remained a defining thread through every generation since.
The company's strategic identity is open weights at frontier scale. Where most labs treat open release as a tier below their flagship, Moonshot has released its frontier models publicly — the K2 family shipped under a Modified MIT License — and priced API access aggressively enough to exert real cost pressure on US frontier vendors. The lab is one of the few Chinese frontier labs to publicly disclose consumption figures, reporting $200 million in annualized recurring revenue in April 2026, and Kimi models have consistently ranked among the most-used on OpenRouter.
On July 16, 2026, Moonshot shipped Kimi K3, the clearest expression of that strategy to date. K3 is a 2.8 trillion parameter mixture-of-experts model activating just 16 of 896 experts per token, built on the lab's own Kimi Delta Attention and Attention Residuals work, with a 1 million token context window and native vision. It is the largest open-weights model any Chinese lab has produced, and it beats Claude Opus 4.8 on most of the published coding and agentic suite — most dramatically on FrontierSWE, 81.2 against 66.7. Moonshot concedes K3 still trails Claude Fable 5 and GPT-5.6 Sol overall, and names user experience rather than raw capability as the remaining gap. Moonshot published the full weights on July 27, 2026 — 96 safetensors shards spanning well over a terabyte — but under a custom Kimi K3 License rather than the Modified MIT terms the K2 family used. The new license permits download, modification, and commercial use with conditions: an inference or fine-tuning service whose revenue exceeds $20 million over 12 consecutive months must negotiate a separate agreement with Moonshot, and any product above 100 million monthly active users or $20 million in monthly revenue must display Kimi K3 branding in its interface. Internal use and access through Moonshot’s own products are exempt. The shift matters beyond the legal text: the lab whose strategic identity is open weights at frontier scale attached commercial strings at exactly the point its model became competitive with the proprietary frontier.
K3's pricing marks a shift in posture. At $15 per million output tokens, Moonshot is no longer undercutting the frontier the way earlier Chinese open models did — it is pricing like a frontier lab, with the aggressive rate reserved for cached input at 30 cents per million tokens. Taken with the launch's timing one day before Xi Jinping used the World AI Conference in Shanghai to make open source explicit Chinese national strategy, Moonshot has become the most concrete evidence available for a policy position China's head of state is now articulating directly.
🛠️Products & Tools (5)
Moonshot AI's open-source coding tool with integrations for terminals, VS Code, Cursor, and Zed. Companion to the Kimi K2.5 model.
Moonshot AI's current flagship — a 1 trillion parameter MoE model (32B active) with 262K context, native INT4 quantization, and an Agent Swarm system scaling to 300 sub-agents and 4,000 coordinated steps in 12-hour autonomous coding sessions. Open-weights under Modified MIT License. Ranks 2nd on OpenRouter.
Moonshot AI's flagship — a 2.8 trillion parameter mixture-of-experts model activating 16 of 896 experts per token, with a 1 million token context window and native vision. Beats Claude Opus 4.8 on most coding benchmarks; largest open-weights model from a Chinese lab. Weights due July 27, 2026.
Moonshot AI's 1 trillion parameter MoE model (32B active). 256K context, multimodal (15T mixed tokens). Beats GPT 5.2 on SWE-Bench Multilingual and Claude Opus 4.5 on VideoMMU.
Moonshot AI's AI assistant powered by Kimi K2.5 (1T MoE, 32B active). 256K context, multimodal (15T mixed tokens). Beats GPT 5.2 on SWE-Bench Multilingual. Kimi Code for IDE integration.
