Index Company
Moonshot AI
16 CÈ pieces since 9 May 2026: 11 intel items, 5 essays. Newest first. All companies →
-
Washington banned the chips. Chinese labs rent them in Thailand instead.
Washington will license the wire rather than the chip, and Commerce will do it by rule before the Senate ever votes.
-
The base is free, the moat is not: Harvey post-trains Kimi K3 for law
Harvey and Fireworks AI ran asynchronous reinforcement learning on Moonshot's open Kimi K3 across roughly 150 NVIDIA B300 GPUs for two months. Tenet nearly doubles completed tasks on Harvey's Legal Agent Benchmark. No weights, no API.
-
DeepSeek’s API price quadruples Sunday. Peak hours are Singapore’s working day.
V4-Pro-0813 shipped on 13 August with a one-million-token context and an MIT-licensed agent harness. Three days later the promotional rate that put it into OCBC’s toolset and Indosat’s Indonesian stack expires.
-
Alibaba will publish Qwen3.8-Max’s weights next week. Intelligence stops being the scarce input.
Alibaba shipped a 2.4-trillion-parameter model on 3 August and promised the weights for the week of 10 August, which reprices every Southeast Asian AI project around the one input nobody can download.
-
Agents hit 100% on builds and still fumble access rules: Supabase ships Evals under Apache-2.0
Supabase released Evals on August 1, an Apache-2.0 harness that runs coding agents against containerised Supabase stacks seeded from real support tickets. Opus 5 and Kimi K3 cleared the build stage at 100%, while row-level security policies stayed the weak surface.
-
The gap between open and closed models is now months: Moonshot ships Kimi K3's 2.8T weights
Moonshot published Kimi K3's full weights under a modified MIT licence, a 2.8 trillion parameter mixture-of-experts model with a million-token context that ranks first on LMArena's Frontend Code Arena.
-
The largest open-weight model ever ships today: Moonshot releases Kimi K3's 2.8 trillion parameters
Moonshot AI publishes K3's full weights under a modified MIT licence. The mixture-of-experts model fires 16 of 896 experts per token, roughly 50 billion active parameters, and the download runs about 1.4 terabytes.
-
When open weights stop crossing borders, who loses? Beijing weighs curbs with Alibaba and Zhipu
China's commerce ministry has consulted Alibaba, ByteDance, and Zhipu AI on restricting overseas downloads of advanced model weights and transfers of training data, leaving cloud and API access to the same models open. Nothing is final.
-
Beijing Will Not Wall Off Its AI (Even Now That It Leads)
Days after an open-weight Chinese model took the top of the coding charts, Beijing began consulting its champions on export-controlling AI, the same weapon it called economic bullying when Washington aimed it the other way.
-
Frontier coding at $0.30 per million tokens: Moonshot opens Kimi K3's 2.8 trillion weights
Moonshot AI releases Kimi K3's weights on July 27, a 2.8-trillion-parameter model built for long-running autonomous coding. Artificial Analysis places it just behind GPT-5.6 Sol and Claude Fable 5, while Arena.ai's frontend leaderboard ranks it above both.
-
Grok 4.5 cut coding-AI prices by roughly 80 percent. Where to point it first.
xAI, OpenAI, and Meta all cut inference prices in the same week, and the cheapest coding model is now roughly twice as willing to make things up as the one it replaced.
-
Chinese open-weight models pass 60 percent of OpenRouter's token volume
By May 2026, openly licensed models from Alibaba, DeepSeek, Moonshot, and Zhipu accounted for roughly 61 percent of tokens consumed on OpenRouter, with four of the five most-used models built in China.
-
Coinbase halves its AI bill by defaulting engineers to Chinese open models
Coinbase cut internal AI spending nearly 50 percent by defaulting engineers to open-weight GLM-5.2 and Kimi K2.7 Code, priced near $1.40 per million tokens against Anthropic Opus at $5.
-
Moonshot's open Kimi K2.7 Code undercuts GPT-5.5 and Claude twelvefold
Moonshot AI released Kimi K2.7 Code, a trillion-parameter Mixture-of-Experts coding model priced at $0.95 per million input tokens, roughly a twelfth of frontier API rates, while cutting thinking-token usage by about 30%.
-
Four Chinese labs ship frontier-parity coding models in three-week window
Alibaba's Qwen 3.7 Max, Moonshot's Kimi K2.7, MiniMax M3, and Zhipu's GLM-5.2 all launched in May and June 2026, matching Claude Opus 4.8 and GPT-5.5 on coding benchmarks at one-fifth to one-thirtieth the per-token cost.
-
Four Chinese labs ship frontier-grade open-weights coding models in 12-day window
Z.ai's GLM-5.1, MiniMax M2.7, Moonshot's Kimi K2.6 and DeepSeek V4 all landed within a fortnight, hitting near-frontier benchmarks on agentic coding at under a third of Claude Opus 4.7's inference cost.