Index Company
Anthropic
42 CÈ pieces since 1 May 2026: 1 deal, 29 intel items, 12 essays. Newest first. All companies →
-
MIT weights land half a point behind Claude Opus: GLM-5.3-Flash costs $0.15 per million tokens
Z.ai released a 320B mixture-of-experts activating 18B parameters with a 1,048,576-token context under an MIT licence. It scores 84.3 on Terminal-Bench 2.1 against Claude Opus 4.8's 85.0, and lifts DeepSWE 37% over GLM-5.2.
-
The model was never the moat: DeepSeek open-sources its agent harness under MIT
DeepSeek published Harness v0.1 under the MIT licence, built on the Cordis meta-framework where models, tools, sandboxes, sessions, and orchestration are all plugins. It ships OS-level sandboxing and supports Anthropic, OpenAI, Bedrock, Azure, and Gemini.
-
DeepSeek’s API price quadruples Sunday. Peak hours are Singapore’s working day.
V4-Pro-0813 shipped on 13 August with a one-million-token context and an MIT-licensed agent harness. Three days later the promotional rate that put it into OCBC’s toolset and Indosat’s Indonesian stack expires.
-
Singapore's sovereign fund becomes an AI landlord: GIC and Macquarie build Theseus for Anthropic
GIC and Macquarie Asset Management own and fund Theseus Infrastructure, supplying majority equity on each project. Anthropic anchors as tenant and covers electricity price rises for consumers near the sites. The first builds are in the United States.
-
A free harness edges the human baseline on ARC-AGI-3: Prime Intellect opens Prime Agent under MIT
Prime Agent treats context as a variable and sub-agent delegation as function calls inside a persistent IPython kernel. Paired with Claude Opus 5 it scored 95.5 percent on ARC-AGI-3, just past the reported human expert baseline.
-
Alibaba will publish Qwen3.8-Max’s weights next week. Intelligence stops being the scarce input.
Alibaba shipped a 2.4-trillion-parameter model on 3 August and promised the weights for the week of 10 August, which reprices every Southeast Asian AI project around the one input nobody can download.
-
Who pays for free weights? Open-model startups get nothing while two labs take 60% of VC
American open-weight startups including Arcee AI, Reflection AI, and Poolside report cold venture markets, with Arcee's chief executive saying every tier-one firm declined, while OpenAI and Anthropic absorbed over 60% of US venture dollars in the first half.
-
Anthropic prices what you sell at $15 a million tokens. You get no vote.
On the dichotomy of control, and the founder whose gross margin is reset each morning by a price she gets no vote in.
-
Four frontier models shipped in ten weeks. From high enough, not one of them happened.
On the view from above, and the summer four frontier models shipped while a founder built nothing.
-
Frontier coding at $0.30 per million tokens: Moonshot opens Kimi K3's 2.8 trillion weights
Moonshot AI releases Kimi K3's weights on July 27, a 2.8-trillion-parameter model built for long-running autonomous coding. Artificial Analysis places it just behind GPT-5.6 Sol and Claude Fable 5, while Arena.ai's frontend leaderboard ranks it above both.
-
Grok 4.5 cut coding-AI prices by roughly 80 percent. Where to point it first.
xAI, OpenAI, and Meta all cut inference prices in the same week, and the cheapest coding model is now roughly twice as willing to make things up as the one it replaced.
-
Alibaba, Tencent, and ByteDance ship rival coding agents built for local workflows
China's largest platforms launched competing agentic coding tools, Alibaba's Qoder among them, betting that localized workflows and language coverage can match Anthropic's Claude Code and OpenAI's systems for domestic and regional developers.
-
Thinking Machines raised $2bn before shipping a product. Some rounds you refuse.
On sophrosyne, and the AI mega-round whose valuation would set your pace for you.
-
The biggest customer the AI labs have is quietly building their replacement
Microsoft began routing tens of thousands of Excel and Outlook prompts a week to its in-house MAI models this month. The question that leaves for anyone building on AI is not which lab to trust.
-
Chinese open models from DeepSeek and Z.ai reach US frontier parity
US enterprises are moving workloads onto open-weight models from DeepSeek and Z.ai, now rated competitive with leading American systems, as OpenAI and Anthropic inference prices keep climbing.
-
What an AI agent can actually run in an SME back office right now
Two frontier models shipped in five weeks and the real product is admin, not code. The workflows an agent finishes alone are the ones that never leave the ledger; the ones that break are the ones that touch a person.
-
Anthropic, OpenAI, and Google converge on the agent harness as product
The runtime that wraps a model with memory, tools, and control loops is now the sold unit. Anthropic prices Managed Agents at eight cents per session hour; OpenAI ships its harness as an open-source Agents SDK update.
-
Anthropic ships Claude Sonnet 5 as a cheaper default for running agents
Sonnet 5 shipped June 30 as the default for free and paid Claude users, running agents that plan, browse and drive terminals at roughly what larger models cost months ago, with introductory pricing of $2 and $10 per million tokens.
-
Zhipu releases GLM-5.2 open weights, topping openly available model rankings
Zhipu released GLM-5.2 under an unrestricted MIT license, ranking as the top openly available model on Artificial Analysis' Intelligence Index while undercutting Claude Opus by roughly five times on price per token.
-
Coinbase halves its AI bill by defaulting engineers to Chinese open models
Coinbase cut internal AI spending nearly 50 percent by defaulting engineers to open-weight GLM-5.2 and Kimi K2.7 Code, priced near $1.40 per million tokens against Anthropic Opus at $5.
-
Microsoft Agent Framework reaches GA, adding CodeAct and hosted agents
The framework, consolidating AutoGen and Semantic Kernel into one open .NET and Python SDK, adds CodeAct, which has models emit a single Python program instead of sequential tool calls, cutting tokens 63.9% and latency 52.4% on tested workloads.
-
Moonshot's open Kimi K2.7 Code undercuts GPT-5.5 and Claude twelvefold
Moonshot AI released Kimi K2.7 Code, a trillion-parameter Mixture-of-Experts coding model priced at $0.95 per million input tokens, roughly a twelfth of frontier API rates, while cutting thinking-token usage by about 30%.
-
The best AI model in the world is now off-limits to everyone in Southeast Asia
On 12 June Washington ordered Claude Fable 5 disabled for every foreign national; the outage was narrow, but the precedent is a supply-chain lesson every operator already knows.
-
Open Design reaches 57,400 GitHub stars as local-first Claude Design alternative
Eight weeks after its first commit, Open Design shipped v0.9.0 with 310 contributors, keeping design artifacts local via SQLite and auto-detecting 16 coding agents so teams avoid cloud lock-in.
-
Zhipu AI open-sources GLM-5.2, trailing Claude Opus by one point on coding
Released June 13 under an MIT license, the 744-billion-parameter mixture-of-experts model carries a one-million-token context window and lands within one percentage point of Claude Opus 4.8 on FrontierSWE, the hours-long coding benchmark.
-
OpenCode reaches 160,000 stars as 7.5 million developers adopt model-agnostic coding
The open-source CLI agent connects more than seventy-five providers, from Claude and GPT to Gemini and local Ollama models, through a single interface, and now serves 7.5 million monthly developers.
-
The cheap AI on your desk is priced below cost
Four giants are losing money on every token to win the market, the bill for the buildout runs to US$5.5 trillion, and Singapore’s own forecasters have started naming the morning the subsidy stops.
-
Four Chinese labs ship frontier-parity coding models in three-week window
Alibaba's Qwen 3.7 Max, Moonshot's Kimi K2.7, MiniMax M3, and Zhipu's GLM-5.2 all launched in May and June 2026, matching Claude Opus 4.8 and GPT-5.5 on coding benchmarks at one-fifth to one-thirtieth the per-token cost.
-
Apple ships LanguageModel protocol letting iOS apps swap Foundation, Gemini, Claude without code rewrites
The new Swift Package Manager protocol announced at WWDC 2026 lets developers switch between Apple Foundation Models, Google Gemini, and Anthropic Claude with no session-code changes, while Xcode 27 ships agentic coding features and multi-turn streaming in the first beta.
-
Open Design reaches 57,400 GitHub stars in eight weeks as local-first Claude Design alternative
Open Design hit version 0.9.0 on June 2, 2026 with 310 contributors, 6,500 forks, and 1,837 commits, offering self-hosting, model flexibility, and data-residency control for LLM-driven design artifact generation where Anthropic's Claude Design offers only cloud lock-in.
-
Open Design hits 57,400 GitHub stars as local-first alternative to Claude Design
Open Design reached v0.9.0 on June 2 with 310 contributors, offering model-agnostic design artifact generation, self-hosting, and data residency control. The tool auto-detects local coding agents, stores everything in SQLite, and supports Claude, OpenAI, Gemini, and Ollama endpoints with BYOK proxy and SSRF protection.
-
Microsoft ships seven in-house MAI models and ditches third-party distillation for clean training
Microsoft released MAI-Thinking-1, MAI-Code-1-Flash, MAI-Image-2.5, MAI-Transcribe-1.5 and three on-device Aion models on June 2, all trained from scratch without distillation from OpenAI or Anthropic. MAI-Thinking-1 matches Sonnet 4.6 in blind evals while MAI-Code-1-Flash ships with 5 billion active parameters.
-
Apple opens Foundation Models framework to third-party models and free Private Cloud Compute for small developers
At WWDC Platforms State of the Union on June 9, Apple announced free Private Cloud Compute access for developers with under two million App Store downloads, image input support, server-side model integration for Claude and Gemini, and Dynamic Profiles for multi-agent workflows. The framework will go open source this summer.
-
Microsoft ships MAI-Code-1-Flash to rival OpenAI at 10× the cost efficiency
Microsoft unveiled MAI-Code-1-Flash, its inaugural coding model, at Build 2026 on June 2, claiming it outperforms GPT-5.5 with 10× better cost efficiency after refining for consulting-firm workloads like McKinsey's.
-
Open Design hits 57,400 GitHub stars in eight weeks as local-first alternative to Claude Design
Open Design reached v0.9.0 on June 2, 2026, with 310 contributors and 57,400 stars. The project replicates Claude Design's artifact generation workflow but runs locally, supports any model via API proxy, and stores artifacts in SQLite—no cloud lock-in.
-
GitHub Copilot moves all plans to usage-based billing June 1, ending flat-rate subscriptions
GitHub announced April 27 that Copilot will shift to token-based AI Credits on June 1, replacing per-seat pricing. Pro monthly plans include $10 or $39 in credits; Enterprise pooled credits align to seat count. Code completions remain unlimited; agentic workflows consume credits per token at varying model multipliers.
-
Microsoft Work IQ API launches with consumption-based Copilot Credits model, no separate subscription SKU
Work IQ API reached general availability on June 16 with usage billed through Copilot Credits—a unified consumption currency across Microsoft Copilot Studio and AI services, replacing per-user or per-seat pricing.
-
Open Design reaches 57.4K GitHub stars in eight weeks as self-hosted Claude alternative
Community project Open Design hit v0.9.0 on 2 June with 310 contributors, 1,837 commits, 6,500 forks. Offers local-first design-artifact generation, multi-provider API proxy (Anthropic, OpenAI, Azure, Gemini, Ollama), and Claude Design ZIP import. Runs via pnpm, deploys to Vercel, stores in SQLite.
-
What a ten-person Singapore startup will pay Microsoft to govern its AI agents.
Microsoft moved Agent 365 to general availability on May 1, priced at fifteen US dollars per user per month. The governance layer is the small number. The compute it brackets is the large one.
-
Four Chinese labs ship frontier-grade open-weights coding models in 12-day window
Z.ai's GLM-5.1, MiniMax M2.7, Moonshot's Kimi K2.6 and DeepSeek V4 all landed within a fortnight, hitting near-frontier benchmarks on agentic coding at under a third of Claude Opus 4.7's inference cost.
-
Warp open-sources its agentic terminal under AGPL with OpenAI as founding sponsor
The five-year-old Rust terminal moves to AGPL-3 with an agent-first contribution workflow and crossed 37,000 stars within days of release, hitting #2 on GitHub trending.
-
Mistral's biggest SEA deployment skips the consumer market entirely.
Singtel and Mistral announced their AI-compute partnership at DC Tuas this week. The deployment targets Singapore's banks, government, healthcare, and telco workloads. The geography is the deal.