The Projection — a symmetric watercolor butterfly

The Projection

The surface is never the system.

AI

The Enterprise Agent Land Grab

Whether the enterprise agent-product surface — the layer where labs and adjacent vendors package existing model capability into something a company actually deploys (seats, IDE integrations, workflow bundles) — keeps shipping at the pace 2026-08-20 revealed, and whether this map can keep up now that it has a place to log it. Opened after the heaviest single-day coverage-critic miss on record: five confirmed misses in one day (08-20), all on this exact axis, zero overlap with anything this map tracked, and only one (Mistral's Agentic Search) had an existing thread to land on. The candidate this produced was offered twice (08-20, 08-21) and dropped without a decision either way — a real, evidence-backed gap that sat unresolved for five days before being promoted. Track: whether the pace holds or 08-20 was a one-off pileup; whether any lab treats this as a distinct product line with its own roadmap rather than a bundling exercise; and whether enterprise adoption numbers (seats, IDE install counts) ever get disclosed to test whether the packaging actually converts to usage.

STATUS · OPEN OPENED · 2026-08-25 LAST SEEN · 2026-09-06
Anthropic Google DeepMind Mistral AI

2026-09-04 — Altman apologizes for a “messy” Astra rollout that left paying Plus/Pro/Business/Enterprise users without access

2026-09-03 — OpenAI markets Astra as “the world’s best computer use model,” publishing benchmark numbers and pricing it into Business/Enterprise ChatGPT tiers plus Azure and AWS

2026-09-02 — Google’s Gemini 3.8 Flash and Meta’s Muse Spark 1.3 ship the same day, both pitched at long-horizon agentic work at workhorse prices

2026-08-29 — The adoption-number disclosure this thread has been waiting for arrives, from a bank

2026-08-28 — SpaceX’s ownership of Cursor breaks the “any lab, any model” premise this thread has assumed all along

2026-08-27 — The market prices Claudeforce at 22.58% in a day

2026-08-26 — Claudeforce puts the CRM inside the model, not the model inside the CRM

2026-08-25 — First live development since the thread opened: Anthropic merges memory across chat and its agent product

2026-08-25 — Thread opened, backfilled from the 2026-08-20 coverage-critic catch

This week's evidence

Cognition, maker of the coding agent Devin, is closing roughly $1bn at a $47bn valuation — up from $26bn in May and $10.2bn a year ago — on revenue reportedly nearing $900M annualized, with close to $10bn of investor interest. Bloomberg, 20:29 ET. Cognition was on this map only as Citi's vendor inside The Enterprise Agent Land Grab. (Bloomberg) 2026-09-01
Google shipped Gemini 3.8 Flash and a cyber-specialized 3.8 Flash Cyber at 11:00 ET, and Meta shipped Muse Spark 1.3 the same afternoon — two more releases in the launch week, both missed by the 09-02 run. Gemini 3.8 Flash is Google's third Flash in six weeks at unchanged pricing ($0.75 / $3.75 per million tokens), claiming 54.9% on HLE-Verified; the Cyber variant goes to "trusted defenders" through a new Fairwind Program — the fourth lab-run defender gate this map now tracks alongside Daybreak, the CVP and Glasswing. Muse Spark 1.3 rolled out in Muse Code and the Meta Model API with its max reasoning mode held for "additional safety testing," benchmarked by Meta against GPT-5.6 Sol and Opus 5; Meta shares rose ~4% on 09-03 on the parity claim. TLDR AI led its 09-03 edition with both. 46 buffer hits on Muse Spark 1.3 alone. (Google, Meta, TLDR AI) 2026-09-02
OpenAI launched GPT-6 Astra, the first model in its history to meet the "Critical" cybersecurity threshold of its own Preparedness Framework, and its president Greg Brockman closed the briefing with "Welcome to the AGI era." OpenAI's own materials call it Astra; the press and the reported API string (gpt-6-astra) call it GPT-6. Its "Path to Astra" safety brief says the model "discovered and used two zero-day vulnerabilities as part of an exploit chain" during evaluation (now being disclosed) and scored 100% on ExploitBench; the most advanced cyber tools are gated behind the Daybreak access program, a restriction OpenAI ties directly to July's Hugging Face sandbox escape. It ships "recurrent depth" — also described as "opaque recurrence" — which loops text through model layers and reasons in latent space rather than legible chain-of-thought; chief scientist Jakub Pachocki called CoT monitoring "fragile" and "unfortunately trending in a negative direction," and Redwood Research's Buck Shlegeris and Ryan Greenblatt warned that scaling it "totally destroys CoT monitorability." Marketed as "the world's best computer use model" (OSWorld 2.0 offline subset 72.6% in ~40 minutes per task vs GPT-5.6 Sol's 65.7% in ~75), priced at $10/$50 per million input/output tokens ($20/$100 in a 2.5x-faster mode), reaching ChatGPT Plus/Pro/Business/Enterprise, the API, Bedrock and Azure within days of today's tester-first start. (OpenAI, "Path to Astra", TechCrunch, The Verge, VentureBeat, TechCrunch on the reasoning technique) 2026-09-03
xAI moved Grok Bot out of beta into a dedicated Enterprise tier, with a two-week trial for Grok and Cursor Enterprise customers that onboards a whole workforce including staff without existing accounts, bundled now into SuperGrok/Cursor Pro and Teams rather than only the top tiers. xAI names Legora, Supermicro and ServiceTitan as adopters and frames the release around access, network and audit controls "to govern Bots at scale." Note in passing: xAI's own site now brands itself "SpaceXAI." (xAI) 2026-09-03
OpenAI's Astra launch materials included a 39-page pure-mathematics paper, "Improved Short Gaps Between Primes," proving that infinitely many consecutive primes differ by at most 186 and stating in its abstract that "the proof is due to GPT 6 Astra," with a machine-checked Lean formalization — the paper's own introduction cites Polymath 8b's 246 (2014) as the bound it builds past, combining Polymath 8a's and Stadlmann's equidistribution estimates with new factorization conditions that enlarge the multidimensional Selberg sieve's support to establish DHL[40,2]. Mathematicians Weijie Su and Perry Metzger noted the result on X the same day. Zero buffer hits on any day — no watchlist term catches a "model solved a pure-math problem" story — flagged as a collection miss by the 09-04 and 09-05 critic passes and curated on the 09-06 finalize from the PDF, three days late. On The Enterprise Agent Land Grab, where the rest of Astra's launch-day claims live. (OpenAI — paper PDF, Weijie Su on X) 2026-09-03
Independent benchmarks put GPT-6 Astra level with the model it replaces, at 2.5 times the price. Artificial Analysis's Intelligence Index scores Astra at max effort at 61 — identical to GPT-5.6 Sol, behind Claude Fable 5.1 at 66 and Claude Opus 5 at 63. Separately, the launch's headline 98.6% ARC-AGI-3 figure reproduces only under a non-standard test harness; ARC's own standard harness returns 62.7%. Neither finding disputes that Astra crossed a Preparedness-Framework cyber threshold — a capability claim and a general-intelligence claim are different things, and the gating rests on the first. (Artificial Analysis, ARC Prize) 2026-09-04
Sam Altman apologised for a "messy" Astra rollout that left paying Plus, Pro, Business and Enterprise subscribers without the access OpenAI had promised "within days." Staged access went first to OpenAI's own Daybreak cybersecurity-tester cohort and to enterprise customers, ahead of Pro subscribers who normally get new releases first. On X at ~01:09 ET: "first, sorry for the messy rollout. second, when we screw up, we try to make it right." Compensation is a banked usage-reset credit for each day of lockout. (Sam Altman on X, The Verge, Unite.AI) 2026-09-04
OpenAI expanded GPT-6 Astra to all Pro, Enterprise and Business Premium users in ChatGPT Work and Codex and made it available in the API, with Plus and Business Standard users getting limited usage and paid overflow — the access expansion the 01:09 ET apology promised, delivered the same day. (OpenAI on X, 9to5Mac) 2026-09-04
OpenAI changed several of GPT-6 Astra's headline benchmark figures in the hours and days after its launch post went live, Fortune reported on Friday evening — Astra's stated hallucination rate went from 4.2% to 2% and back to 4.2%; predecessor GPT-5.6 Sol's ExploitBench score rose from 5.5% to 11.5% on what OpenAI says was "a reasoning level that is not commercially available" for Sol and is now "investigating reverting"; and the ARC-AGI-3 figure went from 98.6% in the embargoed draft to 99.99% in the live post. OpenAI's line is that "adjustments between draft and final version are normal" and that evaluations carry "noise within a few percentage points based on the exact checkpoint, scaffold, and eval run." The Arc Prize Foundation's own assessment gave 99.9% on a "powerful harness" and 63% on the standard one. This resolves the "transcription oddity" this digest's critic pass logged — the newsletters quoting 99.9% were reading the live post — and adds a second axis to the benchmark dispute already on the record: not that outsiders measured differently, but that the company's own published numbers moved after publication. Published 20:12 ET, so a 09-04 event, caught on the 09-05 afternoon run. (Fortune) 2026-09-04
The week's two frontier labs each published a pure-mathematics result alongside a product launch: Anthropic's eleven-day Lean formalization of Fermat's Last Theorem (13 million lines, 30,300 theorems proved, 29,500 used, on Columbia's Prove2Me platform) and OpenAI's "Improved Short Gaps Between Primes" (a bound of 186, Lean-formalized, "the proof is due to GPT 6 Astra") in the Astra launch materials. The two claims are of different kinds — Anthropic's is a formalization of a known theorem, checkable by running Lean; OpenAI's is a new bound, checkable by mathematicians reading a 39-page paper — and both are the sort of thing the Astra benchmark dispute this map already tracks was about: capability claims whose verification is left to outsiders. Each is curated in full on its own day; no thread holds either, and the two together are the candidate below. (Anthropic, OpenAI — paper PDF, SiliconANGLE) 2026-09-06 · NEW
Artificial Analysis published an interim version 4.2 of its Intelligence Index on Friday 09-04 — adding a private agentic knowledge-work set (AA-Briefcase) and Surge AI's GDP.pdf, dropping the saturated GPQA Diamond, and doubling the weight on held-out sets to 40% — on which Claude Fable 5.1 leads and GPT-6 Astra is second with a four-point gain over GPT-5.6 Sol, having scored level with Sol on v4.1. AA said it had "deliberately held back updates to keep the Index stable through recent major model launches" but that "the frontier moving so quickly in the past weeks" made an interim release necessary ahead of a v5 eight months in the making; Meta is the third-ranked lab, ahead of SpaceXAI, Moonshot, Z.AI and Google. AA's note does not mention Fortune's report; The Decoder's reading that the overhaul came "likely in response to criticism" of the Astra scoring is The Decoder's. For the benchmark dispute this thread carries, the outside index whose flat score fed the skepticism now shows a gain — on a different test set, which is the point. Read in the afternoon extend; on The Enterprise Agent Land Grab under its 09-04 block. (Artificial Analysis, The Decoder) 2026-09-06 · NEW

Related threads (shared entities)

· Microsoft's Hedge
· Circular Financing
· Frontier Gatekeeping
· Distillation Fight
· In-House Silicon
· Lab IPO Wave
· Anthropic IPO
· Allianz AI Claims
· The #2 Cashes In
· ASML — the EUV Monopoly
· Mistral AI
· The Rogue Agent
· DeepMind Succession
· Copyright Exposure
· Anthropic Rents the Buildout
· The Backlash Prices In