July 2026
-
AI News — July 28, 2026
Moonshot shipped Kimi K3's weights — 2.8 trillion parameters, 1.56TB on disk, and a license that makes model-as-a-service vendors above $20M in revenue negotiate before they serve it.
-
AI News — July 27, 2026
Kimi K3's weights are not out yet: Moonshot's own Hugging Face repo carries a release timer set to 15:00 UTC today, hours after outlets began treating the drop as done.
-
Weekly recap
AI News — Week of July 20–26, 2026
Claude Opus 5 landed at half Fable 5's price with a per-request effort dial, Gemini 3.6 Flash undercut its own workhorse tier, and Cursor shipped routing — the week cost became a setting rather than a model choice.
-
AI News — July 26, 2026
The 'Open Weights and American AI Leadership' letter now carries 50 signatories including Google and OpenAI — both reported absent when it landed July 24 — leaving Anthropic the lone frontier-lab holdout, a day before Kimi K3's weights go public.
-
AI News — July 25, 2026
Anthropic launched Claude Opus 5 — its new default Opus, priced at $5/$25 per million (unchanged from Opus 4.8 and half of Fable 5's rate) with a per-request low/medium/high effort dial that trades cost for capability, and Anthropic's numbers put it around Fable 5-level intelligence with a new state-of-the-art on agentic coding, reframing the flagship tier as a cheaper, tunable model rather than a pricier one.
-
AI News — July 24, 2026
OpenAI launched Presence, an enterprise platform for deploying governed voice and chat agents whose distinguishing feature is a Codex-driven improvement loop — the coding agent reviews production transcripts and proposes agent-behavior changes for staff to approve — and OpenAI says it already resolves 75% of calls on its own support line and cut human handoffs 15 points in 10 days, marking a shift from selling model access to selling a managed, self-improving agent system.
-
AI News — July 23, 2026
Google shipped Gemini 3.6 Flash and two siblings — 3.5 Flash-Lite and a governments-only 3.5 Flash Cyber — but still no 3.5 Pro: the new workhorse cuts output-token use ~17% (up to 65% on DeepSWE) at a lower per-token price while scoring higher on every internal eval, the first concrete sign Google's Gemini pipeline is shipping again after three missed 3.5 Pro targets.
-
AI News — July 22, 2026
Claude Code 2.1.216–2.1.217 put the first hard limits on runaway multi-agent fan-out — a default cap of 20 concurrent subagents, no nested subagents unless you opt in, and a `--max-budget-usd` ceiling that now actually halts background agents — alongside a run of worktree/symlink workspace-escape fixes, shifting the release focus from interpreting allow-rules to bounding what a single prompt can spawn and spend.
-
AI News — July 21, 2026
Anthropic finalized Fable 5's place in its subscriptions: as of July 20 the model is a permanent part of Max and Team Premium at 50% of plan limits, while Pro and Team Standard lose included access — dropping to a one-time $100 usage credit and then $10/$50-per-million metering — ending six weeks of rolling extensions with a two-tier split rather than another deadline.
-
AI News — July 20, 2026
Anthropic shipped Claude Code 2.1.214 with another round of permission-model fixes — closing an `Edit(src/**)` allow-rule that auto-approved writes to nested directories anywhere in the tree and a permission-check bypass in Windows PowerShell 5.1 — a second straight release window aimed at the approval layer rather than the model, on an otherwise quiet day dominated by watch-list dates.
-
Weekly recap
AI News — Week of July 13–19, 2026
Coding became the axis that reshuffled the field: Google's Gemini 3.5 Pro was confirmed months late for coding shortfalls (~$200B off Alphabet) while China's Moonshot took #1 on an external frontend-coding board with the largest open-weight model ever announced.
-
AI News — July 18, 2026
Moonshot's Kimi K3 arrived as the largest open-weight model ever announced — a 2.8-trillion-parameter (≈50B-active) MoE — and took #1 on the external Frontend Code Arena, edging Claude Fable 5 and GPT-5.6 Sol on a leaderboard the two US flagships had led, though its own benchmark table still trails both; the weights themselves don't drop until July 27.
-
AI News — July 17, 2026
Google's Gemini 3.5 Pro missed its July 17 target: a July 16 Bloomberg report that the flagship is months behind schedule — with coding capabilities short of internal expectations — sent Alphabet down ~4% and erased roughly $200B in market value, turning the watch-list slip the last several briefings tracked into a confirmed delay with no new date.
-
AI News — July 15, 2026
GitHub pulled its /security-review command into the Copilot app (July 14): the AI vulnerability scan that shipped in Copilot CLI last month now runs inside everyday Copilot coding workflows, in public preview for Free, Pro, Business, and Enterprise users — the one substantive coding-agent move on an otherwise model-quiet day.
-
AI News — July 13, 2026
Anthropic blinked on the Fable 5 cutover: with the model set to drop out of Claude subscriptions at 11:59pm PT July 12, Anthropic extended included access a second time — through July 19 — and pushed the switch to metered usage credits ($10/$50 per Mtok) to July 20, now framing the whole move as temporary until capacity catches up.
-
Weekly recap
AI News — Week of July 6–12, 2026
The frontier race turned into a price race: three cheaper near-frontier coding/agent models shipped in seven days while Anthropic's Fable 5 moved to a record-high metered rate.
-
AI News — July 12, 2026
Cursor 3.11 (July 10) introduced "side chats" — durable parallel agent conversations you spin off with /side or /btw and at-mention back into the main thread — the most substantive of a cluster of coding-agent workflow updates on an otherwise model-quiet day, with Kiro extending MCP OAuth to strict servers like Figma and Claude Code turning on auto mode by default across Bedrock, Vertex AI, and Foundry.
-
AI News — July 11, 2026
Meta started charging for its own model for the first time: Muse Spark 1.1, shipped July 9 through the new paid Meta Model API, is an agentic coding model that tops the MCP Atlas tool-use benchmark (88.1) while pricing at $1.25/$4.25 per million tokens — roughly a quarter of Opus 4.8 and GPT-5.5 — marking Meta's turn from open-weight Llama toward a metered, agent-first product.
-
AI News — July 10, 2026
SpaceXAI put a third frontier-tier model into the same week's field: Grok 4.5 shipped July 8 trained in partnership with Cursor on real developer-session data, landing 4th on the Artificial Analysis Intelligence Index (score 54, behind only Fable 5, GPT-5.5, and Opus 4.8) while pricing at $2/$6 per million tokens — which Artificial Analysis clocks at more than 60% below Opus 4.8 and GPT-5.5, turning the frontier race back toward price.
-
AI News — July 9, 2026
The government-restricted frontier model just went public: OpenAI said on July 8 it will make GPT-5.6 (Sol, Terra, Luna) broadly available starting July 9 after the US Commerce Department cleared a wide launch, ending the ~20-partner, government-vetted preview that had run since June 26 — and paired the news with GPT-Live, a full-duplex voice generation that listens and speaks at once and replaces ChatGPT's Advanced Voice Mode.
-
AI News — July 8, 2026
US companies are routing a growing share of their AI traffic to cheaper Chinese open-weight models as frontier prices climb: CNBC reports the Chinese-model share of US tokens on OpenRouter has held above 30% every week since February and spiked as high as 46%, led by Z.ai's GLM 5.2 — within a point of Opus 4.8 on one agentic benchmark at roughly a fifth the cost — even as OpenAI shipped gpt-realtime-2.1 to add reasoning to voice agents and cut Realtime latency 25%.
-
AI News — July 7, 2026
The number behind Fable 5's return finally landed: with the 50%-of-weekly-limits promo expiring today, Anthropic's most powerful public model reverts to metered usage credits at its full $10/$50-per-million-token API rate — double Opus 4.8 and the steepest Anthropic has ever listed for a generally available model — turning the model that anchored a month of export-control drama into a per-call budget line the moment the promo window closed.
-
AI News — July 6, 2026
OpenAI's workspace agents left free preview today: the extended free window ended July 6 and credit-based metering is now live for agent runs invoked inside ChatGPT — a typical GPT-5.5 run burns 5–25 credits — turning the always-on Codex-powered agents into a budgeted per-run line item, while Claude Code's v2.1.200–201 renamed its permission "default" mode to "Manual" and landed another batch of background-agent reliability fixes.
-
Weekly recap
AI News — Week of June 29–July 5, 2026
The month-long export-control saga ended: Commerce lifted the June 12 directive on June 30 and Fable 5 and Mythos 5 came back — Fable 5 redeployed globally July 1 behind a new government-trained cyber classifier and temporary 50%-of-limit terms — landing the same week Anthropic shipped the cheap 1M-context Sonnet 5, its Claude Science research workbench, and first-party enterprise spend controls.
-
AI News — July 4, 2026
A quiet US-holiday window, and the substantive item is governance, not a model: Anthropic shipped admin analytics, model-level entitlements, and spend-threshold alerts for Claude Enterprise, with usage-and-cost data exposed through an Analytics API that pipes into Datadog Cloud Cost Management and CloudZero — turning per-user, per-model Claude spend into something finance can see next to the rest of its cloud bill.
-
AI News — July 3, 2026
Anthropic opened a new front beyond the export-control saga with Claude Science, a dedicated research workbench that wires Claude into 60-plus genomics, proteomics, and cheminformatics databases and marks its own entry into drug discovery — Novo Nordisk and the Allen Institute among early users, up to $30K in research credits on offer — while a High-severity token-exfiltration CVE (CVE-2026-50143) in the widely used Apify MCP server landed in the same window.
-
AI News — July 2, 2026
Anthropic's 19-day frontier blackout ended: after Commerce lifted the June 12 export controls on June 30, Fable 5 and Mythos 5 came back on July 1 — Fable 5 redeployed across Claude.ai, the Claude Platform, Claude Code, and Cowork behind a new, tighter cybersecurity classifier and a temporary 50%-of-limit cap through July 7 — landing the same week Anthropic shipped Claude Sonnet 5, a 1M-context mid-tier model priced at $2/$10 that closes much of the agentic-coding gap to Opus 4.8.
-
AI News — July 1, 2026
The export-control blackout that darkened Anthropic's frontier models for 18 days is over: on June 30 the Commerce Department lifted its June 12 directive on Fable 5 and Mythos 5 after Anthropic agreed to proactively detect security risks, coordinate future releases with the government, and report malicious activity — and Anthropic says it begins restoring Fable 5 globally on July 1.