COMPANY
Moonshot AI
companytopic-notechinese-aiopen-source
Overview
Moonshot AI is a Beijing-based Chinese AI lab and the developer of the Kimi model family, positioning itself as an open-weights frontier-adjacent competitor to closed US labs. The company entered the corpus on 2026-07-17 with the release of Kimi K3, a 2.8T-parameter mixture-of-experts model priced at Claude Sonnet 5-tier rates ($3/$15 per M tokens) with a 1M-token context — a headline pricing entry on the same axis where Anthropic‘s commodity tier competes.
Timeline
- 2026-07-17-AI-Digest — Moonshot AI released Kimi K3, a 2.8T-parameter mixture-of-experts open model with a 1M-token context window and $3 input / $15 output per M tokens (with a $0.30/M cache-hit discount) — same headline pricing as Anthropic‘s Claude Sonnet 5 and materially below the $5/$25 of Claude Opus 4.7. Active-parameter count undisclosed (which matters for cost-per-throughput reads against Inkling‘s 41B active). Simon Willison’s release-day post carries the disciplined framing that pelican-style microbenchmarks are saturated at the frontier and the honest test for K3 is agentic tool-calling and long-conversation reliability, not one-shot SVG generation. Corpus framing: pricing is the story, not raw scale — a claimed 3T-class open model at GPT-5.4 tier undercutting Opus 4.7 output by ~40% is the kind of drop that accelerates the OpenRouter Chinese-origin routed-token share (~46% vs US ~30%, down from ~70% in June ‘25). First Moonshot AI entry in the corpus.
- 2026-07-18-AI-Digest — Kimi K3 surfaces as one of three Bloomberg-named accelerants of the chip-stocks bear-market entry — SOX widening its drop to ~20% from the late-June record — alongside Samsung soft prelims and the second Netlist ITC probe. The digest holds the disciplined spark-on-dry-tinder framing: SOX had already shed ~7% on July 7 Samsung prelims and Applied Materials had shed ~10% before K3 shipped, and TNW literally frames the rout as “already loaded” when K3 landed — K3 provides the visible ignition point and the price-per-token comparison the market wanted for headlines, but the pre-drawdown posture on hyperscaler-capex durability was already set. Two benchmark corrections carry against early framing: K3 beats Claude Opus 4.8 and GPT-5.5 while trailing Claude Fable 5 and GPT-5.6 Sol on coding benchmarks per VentureBeat, and MXFP4 weights arrive 2026-07-27 (not launch), with full-precision self-host still ~1.4 TB storage and 8–16 nodes of 8×H100/B200.
- 2026-07-19-AI-Digest — Moonshot AI surfaces today via Kimi K3‘s role as the Sonnet-tier API-pricing anchor against Anthropic‘s Claude Fable 5 subscription cuts — Max/Team Premium capped at 50% of already-reduced weekly caps (~33% effective vs pre-cycle), Pro/Team Standard losing bundled access with a one-time $100 API credit and then paying list ($10/$50). K3 at $3/$15 per M is the sharpest headline-price comparator on the Pro-tier practitioner segment specifically: Pro subscribers pushed to Fable 5 API rates now weigh K3-at-Sonnet-pricing on the same axis, and the price-per-throughput comparison shifts materially in the open-weights direction. Also contextually adjacent to today’s UK AISI open-weight cyber-capability gap compression (6–10mo → 4–7mo) — the distribution-share thread from 2026-07-15-AI-Digest continues to compound on the capability axis as well as the price axis. No fresh Moonshot AI product action today; the corpus logs today as comparator + pricing-anchor framing rather than a new Moonshot thread.
- 2026-07-20-AI-Digest — Moonshot AI surfaces today as the 72-hour trigger for Alibaba‘s Qwen 3.8 preview: the digest frames Qwen 3.8’s Jul 19 announcement as “the second China-open-weights response to Kimi K3 in 72 hours,” positioning Moonshot AI‘s K3 release as the launch that opened the same-week counter-announcement cycle. No fresh Moonshot AI product action today. The corpus also notes neither Kimi K3 nor Claude Fable 5 nor GPT-5.6 Sol has posted Aider polyglot numbers — the fifth-consecutive-day freeze traces back to 2026-06-12-AI-Digest and continues to read as inclusion-lag, not plateau; K3’s absence remains the load-bearing K3-specific data point on the Aider board.
- 2026-07-21-AI-Digest — Moonshot surfaces today as the Bloomberg-framed centerpiece of market anxiety over Kimi K3 pricing, but the digest holds the disciplined reframe: the pricing move underneath is the actual story, not the parameter count. K3 shipped 2026-07-16 as a 2.8T-parameter open-weight at $3 / $15 per M tokens (
$0.30cached input) — identical to Sonnet 5‘s post-Sept 1 rate card and ~6× the K2.6 rate of$0.95 / $4. The “China ships cheap open weights” thread has now inverted for at least this release — K3 is priced at Sonnet-parity, not below it. Structural read the digest carries: OpenRouter Q2-2026 shows combined Chinese providers >45% of weekly-token share on the inference-volume battlefield, but Anthropic and OpenAI still hold the enterprise-integration and regulated-workload battlefields intact — K3’s Sonnet-parity pricing is Moonshot moving off the inference-volume playbook into the enterprise-margin one, not the other way around. Kimi Work HN thread (~465 pts) as adjacent Moonshot data point on the “Chinese labs shipping product surface, not just weights” line first named in 2026-07-19-AI-Digest‘s K3 pricing coverage.
- Kimi K3 Sonnet-Parity Pricing as Enterprise-Margin Move, Not Inference-Volume Play (July 21, 2026): The Bloomberg framing centres market anxiety and DeepSeek-style compute-moat reevaluation; the disciplined corpus reframe is that a top-of-market Chinese open-weight release chose Western-frontier rates rather than undercut. Combined Chinese providers >45% of OpenRouter weekly-token share on the inference-volume battlefield, but US closed labs still own enterprise-integration and regulated-workload lanes. Moonshot’s Sonnet-parity pricing on K3 is the shift from the inference-volume playbook to the enterprise-margin one — the pattern is not “China open-weights are winning” as a single-winner story, it’s a split.
Key Developments
- Kimi K3 Ships at Sonnet-Tier Pricing (July 17, 2026): 2.8T MoE with 1M-token context at $3/$15 per M tokens — identical headline pricing to Anthropic‘s Claude Sonnet 5 and roughly 40% below Claude Opus 4.7 on output. Active-parameter count undisclosed matters for cost-per-throughput reads against Inkling‘s 41B active. First Moonshot AI entry in the corpus and the sharpest headline-priced open frontier-adjacent entrant since Claude Sonnet 5 shipped. 60-day watch: K3’s Aider polyglot entry once submitted — a top-5 finish at Sonnet pricing would collapse the “cheap but weaker” default assumption; a lower placement re-anchors the price/performance-per-tier read.
Related
See also: Kimi K3, Kimi K2.5, MOC - Open Source Models, MOC - Major Companies.