Staff · 27 pieces on file
Lars Iverson
Open source & model weights
Lars Iverson covers the open-weights ecosystem. He reads architecture diffs for a living and tracks every notable checkpoint release from Meta, Mistral, Qwen, DeepSeek, and the long tail. His beat is licences, parameter counts, and what the weights actually let an operator do.
Beats: open-source, model-releases
All pieces by Lars
-
Open Source · JULY 24, 2026
GLM-5.2 closes the open-weight cyber gap to four months as Beijing weighs locking the door
UK AISI puts Z.ai's GLM-5.2 level with closed models released four months earlier — the narrowest open-to-frontier gap the institute has measured — just as China's Ministry of Commerce consults Alibaba, ByteDance, and Zhipu on whether foreign users should still be allowed to download the weights at all.
-
Infrastructure · JULY 23, 2026
Alphabet Q2 2026: Cloud accelerates to 82%, capex guide pushed to $205B as Gemini serves 22B tokens/minute
Google Cloud revenue jumped to $24.8B on enterprise AI infrastructure demand, and Alphabet lifted 2026 capex guidance by $15B — with CFO Anat Ashkenazi telling analysts demand 'continues to outpace supply across the industry.'
-
Open Source · JULY 23, 2026
Moonshot's Kimi K3 lands at 2.8 trillion parameters, weights follow July 27
Moonshot AI unveiled Kimi K3 on July 17 — a 2.8T-parameter sparse MoE with a 1M-token context — and promised full open weights ten days later. The release reshuffled the open-weight frontier and knocked TSMC down 7%.
-
Open Source · JULY 21, 2026
Thinking Machines ships Inkling: a 975B open-weights MoE that isn't trying to win
Mira Murati's first production model debuts at 41 on the Artificial Analysis Intelligence Index — third-highest among open weights, ahead of Nemotron 3 Ultra, and explicitly positioned as a fine-tuning base rather than a leaderboard entry.
-
Open Source · JULY 20, 2026
Kimi K3 lands at 2.8T parameters, takes Frontend Code Arena, jams Moonshot's own capacity
Moonshot's July 17 open-weight release outscored every model except Claude Fable 5 and GPT-5.6 on its own suite, took the Arena Frontend Code top slot at 1,679 points, and forced the company to pause new subscriptions inside 72 hours.
-
Open Source · JULY 20, 2026
Moonshot ships Kimi K3 at 2.8T parameters — the largest open-weight model, weights due July 27
Beijing-based Moonshot released a sparse-MoE frontier system with a 1M-token context and Kimi Delta Attention, claiming a #2–3 overall finish behind Claude Fable 5 and GPT-5.6 Sol on the company's own eval suite.
-
Open Source · JULY 18, 2026
Moonshot's Kimi K3 lands at 2.8T parameters, sits behind only Fable 5 and GPT-5.6 Sol
Beijing's Moonshot AI unveiled Kimi K3 on July 16 — a 2.8-trillion-parameter sparse MoE with a 1M-token context window, novel Kimi Delta Attention, and full weights due July 27. It is the largest open-weight model ever released.
-
Open Source · JULY 17, 2026
Moonshot ships Kimi K3 at 2.8T parameters, third on Intelligence Index
Beijing's Moonshot AI released Kimi K3 on July 16 — a 2.8-trillion-parameter sparse MoE with Kimi Delta Attention, a 1M-token context window, and API pricing at $3/$15 per million tokens. Full weights are due July 27.
-
Open Source · JULY 15, 2026
Chinese open-weight models pass 41% of Hugging Face downloads as Nadella tells enterprises they're paying twice
Chinese labs took the top six slots on OpenRouter and 41% of Hugging Face downloads this spring. On July 13, Microsoft's CEO gave the shift its first executive-suite endorsement.
-
Open Source · JULY 15, 2026
Chinese open-weight models take 41% of Hugging Face and the top six OpenRouter slots
DeepSeek, Qwen, Z.ai, Tencent, MiniMax, and Xiaomi checkpoints now handle most of the volume-heavy traffic on developer platforms — at 60–90% less per token than GPT-5.6 or Claude Opus 4.8 — while frontier labs pivot to cost-efficiency messaging.
-
Open Source · JULY 14, 2026
Chinese open weights take 41% of Hugging Face downloads and all six top OpenRouter slots
Real production-traffic data from OpenRouter and Vercel's AI Gateway shows US model share collapsing from 70% to 30% in twelve months, with DeepSeek, Z.ai, Tencent, Xiaomi, and MiniMax now processing three times more tokens per week than American labs.
-
Open Source · JULY 14, 2026
Goldman initiates on Z.ai at HK$1,880; GLM-5.2 scores 81.0 on Terminal-Bench 2.1
Goldman Sachs named Z.ai's GLM-5.2, DeepSeek and ByteDance its preferred Chinese AI stack on July 10, days after the 744B-parameter MIT-licensed model cleared Gemini 3.1 Pro on terminal work and undercut GPT-5.5 API pricing by roughly 6x.
-
Model Releases · JULY 7, 2026
White House frontier-model framework lands this week, with a 30-day NSA access window
The voluntary standards implementing Section 3 of Trump's June 2 executive order will govern how OpenAI, Anthropic, and Google release covered frontier models — and directly unblock GPT-5.6's broader launch.
-
Open Source · JULY 5, 2026
Meituan ships LongCat-2.0: 1.6T MoE, 1M context, trained end-to-end on Chinese ASICs
Meituan open-sourced LongCat-2.0 on Hugging Face and GitHub under MIT, unmasking the 1.6-trillion-parameter MoE that had been leading OpenRouter as 'Owl Alpha' — and the first trillion-parameter system pretrained and served entirely on a 50,000-card domestic ASIC cluster.
-
Model Releases · JUNE 28, 2026
Grok 4.5 ships to SpaceX and Tesla only: 1.5T parameters, V9 foundation, Cursor-tuned
xAI's V9-based Grok 4.5 entered closed beta on June 28 inside Musk's two engineering companies — 50% larger than Grok 4.4, trained with Cursor data, and benchmarked internally against Claude Opus.
-
Open Source · JUNE 24, 2026
GLM-5.2 lands at 744B parameters, MIT-licensed, and tied with Opus 4.8 on long-horizon coding
Z.ai's open-weights flagship debuts at #1 on open-source coding boards with a 1M-token context, IndexShare cutting per-token FLOPs 2.9x, and API pricing roughly one-sixth of GPT-5.5's.
-
Open Source · JUNE 23, 2026
OpenAI ships full GPT-5.5-Cyber at 85.6% CyberGym, opens Patch the Planet with Trail of Bits across 19 projects
The Daybreak expansion pairs a permissive-only preview's full release — 85.6% CyberGym, 39.5% ExploitGym — with an open-source remediation sprint that merged dozens of patches across cURL, Python, and the Linux kernel in its opening week.
-
Open Source · JUNE 18, 2026
OVHcloud Plans Frontier Model Family, Says €1B Project Now Costs €150M–€200M
At VivaTech 2026, CEO Octave Klaba said OVHcloud has completed pre-training on one model using Europe's Jupiter supercomputer, with a full open-source family to follow as Washington's Anthropic shutoff reframes Europe's sovereignty debate.
-
Model Releases · JUNE 18, 2026
Pentagon Sworn Statement Puts Grok Gov Model on 2,000 Strikes in 96 Hours
A federal court filing from DoD AI chief Cameron Stanley makes xAI's Grok Gov Model the first commercial LLM officially named as enabling a live strike campaign — 2,000 munitions, 2,000 targets, four days, inside Maven Smart System.
-
Open Source · JUNE 14, 2026
US orders Anthropic to pull Fable 5 and Mythos 5 worldwide over a narrow jailbreak claim
A Commerce Department export-control letter sent at 5:21 pm ET on June 12 forced Anthropic to disable its Mythos-class models for every user globally — three days after Fable 5's general release.
-
Open Source · JUNE 13, 2026
Z.ai Ships GLM-5.2 With 1M-Token Context and an MIT Pledge — and No Benchmarks
Zhipu's international brand pushed its 744B MoE flagship to a million-token window and added dual thinking-effort presets, but launched without a single score and gated the weights behind a 'next week' promise.
-
Open Source · JUNE 12, 2026
Kimi K2.7-Code ships open weights, cuts thinking tokens 30%, and edges Opus 4.8 on MCPMark
Moonshot's coding-focused post-train on the K2.6 MoE family lands on Hugging Face under a Modified MIT license, reports +21.8% on its own Kimi Code Bench v2, and forces thinking mode on every call.
-
Open Source · JUNE 11, 2026
OpenAI files confidential S-1 with the SEC, one week after Anthropic
OpenAI confirmed a draft registration statement on June 10, 2026 at an $852 billion valuation. Goldman Sachs and Morgan Stanley are leading; the company says it has not committed to a timeline.
-
Open Source · JUNE 9, 2026
Microsoft ships seven MAI models from scratch, declares independence from OpenAI distillation
At Build 2026, Mustafa Suleyman's AI Superintelligence team unveiled MAI-Thinking-1 at 97% on AIME 2025 and 53% on SWE-Bench Pro, alongside a 5B-active coding model that lands today as a VS Code default — all trained without third-party distillation.
-
Model Releases · JUNE 8, 2026
Apple's iOS 27 Extensions API turns the iPhone into a four-way model marketplace
Siri runs on a custom 1.2-trillion-parameter Gemini under a reported ~$1B/year Google deal, but the bigger release is the Extensions framework letting users set Claude, ChatGPT, Gemini, or Grok as the system-wide default.
-
Infrastructure · JUNE 7, 2026
Apple licenses a 1.2T-parameter Gemini MoE for Siri, runs it on B200s inside Private Cloud Compute
Bloomberg, TechTimes and Google Cloud's own CEO line up the same architecture ahead of Monday's WWDC keynote: a custom mixture-of-experts Gemini, ~$1B/year, weights sitting on Nvidia B200s inside Apple-controlled enclaves.
-
Open Source · JUNE 6, 2026
Microsoft ships seven MAI models, with MAI-Thinking-1 matching Opus 4.6 on SWE-Bench Pro
At Build 2026, Microsoft AI released a 35B-active-parameter sparse MoE reasoning model trained from scratch, plus six companions across image, voice, transcription, and coding. The flagship hits 97% on AIME 2025 and 53% on SWE-Bench Pro.