The AI Brief

Vol. I · No. 93 · Wednesday, August 26, 2026

Today's brief:

  • OpenAI dismantled a Russian operation that used ChatGPT to fabricate a fake Israeli think tank and seed pro-Kremlin narratives across five platforms, revealing how AI lowers the cost of manufacturing institutional credibility from scratch.
  • Lambda's pursuit of a $12 billion-plus valuation tells operators one thing: institutional money still bets the GPU scarcity premium holds long enough to survive a public listing, so any multi-year cloud contract you sign today is priced against that assumption.
  • OpenAI's retirement of o3 from ChatGPT today marks its fourth model removal in three months, compressing the model lifecycle from 18 months to six and making pinned version IDs and fallback routing a baseline operational requirement, not a nice-to-have.
  • Update: Gemini 3.5 Pro reached day 99 unshipped, now the longest-delayed flagship model of 2026, with no model ID, no pricing, and Google listing Gemini 3.7 Flash as its best callable model in the gap.
  • Anthropic moved Claude Mythos 5 into Claude Security for all Enterprise customers and launched a $35 million Defender Advantage Fund for open-source vulnerability patching, giving defenders access to its most restricted model without exposing the model itself.

OpenAI Dismantles Russian Operation That Used ChatGPT to Manufacture a Fake Think Tank

Why it matters
The operation's significance is less what it reached than what it built: a self-reinforcing apparatus of plagiarized academic credentials, a fabricated sovereignty index, and AI-generated social copy that would have been prohibitively expensive to assemble without ChatGPT.
What's at stake
For most operators, this is context, not a decision. For any team consuming third-party research, monitoring vendors, or sourcing expert commentary, the case illustrates a new attack surface: AI-manufactured institutional authority that passes cursory source checks.
Detail

OpenAI banned a cluster of Russian ChatGPT accounts on August 25 after an investigation traced AI-generated social media content back to a coordinated campaign built around the International Burke Institute (IBI), a self-described "expert community" claiming a street address in Israel. OpenAI recently banned a cluster of ChatGPT accounts originating in Russia that were being used to promote the IBI, a self-described "expert community" based in Israel. The operators accessed ChatGPT via VPNs to circumvent OpenAI's Russia restrictions, routing their connections through VPNs and using ChatGPT to draft social media content that appeared on Substack, Telegram, X, Facebook, and LinkedIn.

In a review of a sample of 36 articles linked to experts on the IBI website and published between September 2025 and May 2026, 34 of the articles were copied from elsewhere on the internet. In one particularly glaring example, a supposed migration policy expert named in a misattributed paper was actually a food sciences professor from Australia. The site also ran a sovereignty index that relied on an internal IBI calculation that was uniformly critical of Western countries at odds with Russia over its invasion of Ukraine. OpenAI rated the campaign Category Three on the Brookings Breakout Scale.

While the immediate impact of the operation appeared limited, OpenAI said the "significance of the operation lies less in the audience it reached, however, than in the infrastructure it had built." Posts produced with ChatGPT steered readers toward what looked like a legitimate research body, a tactic OpenAI said demonstrates how AI tools can help bad actors create an appearance of institutional credibility, bury the true source of preferred narratives, and lay groundwork for operations that could grow considerably larger.


$12B+
Valuation target in Lambda's $3B pre-IPO raise talks

Lambda Seeks $12 Billion-Plus Valuation in Pre-IPO Round, Eyes 2027 Listing

Why it matters
A neocloud targeting $12 billion at over $1.5 billion projected 2026 revenue implies continued institutional conviction in GPU-rental infrastructure even as CoreWeave's post-IPO performance and AI demand forecasts have made the sector's exit window genuinely uncertain.
What's at stake
For operators evaluating multi-year GPU contracts, Lambda's capital-market positioning matters: a company accessing equity ahead of an IPO has different incentives on price, SLA, and contract flexibility than one funded by leveraged hardware debt alone.
Decode
Neocloud = a cloud infrastructure company that rents Nvidia GPU capacity to AI teams for training and inference, without the general-purpose compute stack of hyperscalers like AWS or Azure. CoreWeave is the best-known public example after its 2025 IPO; Lambda, Crusoe, and Nscale are in the same category.
Detail

Lambda Inc., an AI cloud-computing provider backed by Nvidia, is in talks to raise as much as $3 billion in a round that could tee it up to go public next year, according to people familiar with the efforts. The company has discussed raising the money at a valuation of as much as $12 billion or more, the people said, asking not to be identified because the discussions aren't public.

Lambda is expected to generate more than $1.5 billion in revenue in 2026, and the round would come roughly nine months after its $1.5 billion Series E, as rival neoclouds CoreWeave, Crusoe, and Nscale all chase similarly large raises or listings. On August 12, Lambda priced a $926 million senior secured term loan B facility, which it described as the first investment-grade-rated term loan B financing executed by a private neocloud. The new equity round, if it closes, would layer on top of that debt rather than replace it.

Nvidia is an investor, a supplier, and indirectly a source of demand, an arrangement it has replicated across the sector. The wider question for any neocloud is what happens when the capacity crunch eases, as these businesses earn their margins on scarcity, and the same customers renting GPUs today are building their own silicon or negotiating directly with hyperscalers. While Lambda would not be the first neocloud to launch an IPO, CoreWeave launched in 2025, it comes into a market with several analysts anticipating that AI demand will cool, and its IPO could serve as a canary in the coal mine.

Sources
NotePrimary source is Bloomberg; paywalled. Figures cited from Yahoo Finance/Bloomberg wire. · Yahoo Finance / Bloomberg: AI Cloud Provider Lambda in Talks for $3 Billion Pre-IPO Round · TechFundingNews: Lambda eyes $12B valuation with $3B pre-IPO raise

Update: OpenAI Retires o3 From ChatGPT Today, Marking Four Model Removals in Three Months

Why it matters
The pattern, not any single retirement, is the signal: OpenAI's model lifecycle has compressed from roughly 18 months to closer to six, and teams without pinned versioned IDs and a fallback routing layer face operational disruption on every deprecation cycle.
What's at stake
For teams using o3 in ChatGPT, the consumer interface cutoff is today; the API endpoint survives until December 11, 2026, replaced by GPT-5.6 Sol, a longer runway but a different cost profile, since Sol is frontier-priced and o3 handled a wide range of tasks that Terra at $2/$12 per million tokens covers at 40% of Sol's cost.
Detail

OpenAI officially retired the o3 model family from ChatGPT on August 26, 2026, ending a 90-day sunset period that began with the May 28 announcement. The model benchmarked at 96.7% on AIME 2024, 87.7% on GPQA Diamond, and a 2,727 Codeforces Elo and will quietly disappear from the model picker for paid subscribers. One notable carve-out: o3-pro, the higher-compute variant, will remain available in ChatGPT for Pro, Team, Enterprise, and Edu subscribers after today.

OpenAI has retired GPT-4o, GPT-4.1, o4-mini, GPT-4.5, and now o3 from ChatGPT, more models in 2026 than in all prior years combined, with the model lifecycle compressed from roughly 18 months to closer to six. The API shutdown of o3 on December 11, 2026, replaced by GPT-5.6 Sol, will be the true stress test for enterprise migration readiness. Note that today also marks the hard shutdown of the Assistants API, covered in Vol. I, No. 92.

OpenAI reinstated the five-hour rolling usage cap on Codex and ChatGPT Work for Plus subscribers on August 25, framing it as compute load-smoothing. Engineering lead Thibault "Tibo" Sottiaux said the cap lets OpenAI "smoothen the load on our compute," and confirmed Pro $100 and Pro $200 subscriptions will not have the five-hour limit re-enabled "for the upcoming months." The cap's return was cited by some builders as reason to re-evaluate Codex against Claude Code. First covered as an announcement in Vol. I, No. 92.


Update: Gemini 3.5 Pro Hits Day 99 With No Model ID, No Pricing, No Date

Why it matters
Google's best callable model is now Gemini 3.7 Flash, a Flash model, not its announced Pro flagship, and the gap between what Google promised at I/O and what it has shipped is becoming a procurement and planning liability for teams that built roadmaps around a June 2026 Pro release.
What's at stake
For most operators, this is context: Flash is capable and available. For teams that selected Gemini for its Pro-tier reasoning in multi-step agentic pipelines and built contract terms around a June window, the delay has already forced substitutions, and the substitute tier carries different pricing and benchmark ceilings.
Detail

Gemini 3.5 Pro is still unreleased as of August 23, 2026, more than three months after Google announced it at I/O on May 19. It has now missed three targets, June, mid-July, early August, and has no model ID, no pricing, and no launch date. The most recent entry in the Gemini API release notes is August 13, 2026, when Gemini 3.7 Flash went generally available. The official Google DeepMind models page lists Gemini 3.1 Pro as the current Pro model and carries a "3.5 Pro coming soon" marker.

The story behind the delay traces to engineering decisions made months ago. Reports cite persistent coding and reliability problems, senior researcher departures, and a possible complete retraining from the foundational pre-training phase due to a "structural problem." In a single week in June, Google lost Gemini co-lead Noam Shazeer to OpenAI and Nobel laureate and AlphaFold coinventor John Jumper to Anthropic. Google, which acquired DeepMind in 2014, has not unveiled a frontier model since early 2026.

First covered as an ongoing delay in Vol. I, No. 90, with daily updates since. The delay has now extended past the Made by Google hardware event (August 12), the period covered in No. 91 at day 97, and every other window prediction markets assigned meaningful probability. Google has not commented publicly on a revised launch date.


Anthropic Puts Its Most Restricted Model Into Claude Security, Funds $35M Open-Source Patch Drive

Why it matters
Anthropic's architecture, scan results rather than raw model access, every patch still requiring human approval, is a structural answer to the dual-use problem: Mythos 5's offensive cyber capabilities reach defenders without the model itself becoming available to attackers.
What's at stake
For Claude Enterprise customers who run codebases, Claude Security is now a Mythos-class scanner available at claude.ai/security with no separate add-on; the practical question is whether CWE-tagged findings with suggested patches justify the Enterprise contract relative to standalone scanning tools.
Decode
CWE (Common Weakness Enumeration) = a standardized taxonomy of software vulnerability types maintained by MITRE, used here so Claude Security's findings can be categorized and compared against existing security tooling rather than returned as free-form prose.
Detail

Anthropic said on August 21, 2026 that Claude Mythos 5, the cyber-capable model it has limited to vetted defenders since April 2026, is now running vulnerability scans in Claude Security for Enterprise customers. The same announcement lays out plans to bring Mythos 5 into security partners' products, commits $35 million in credits to a new fund for open-source defense, and previews an expansion of the company's Cyber Verification Program toward Mythos-class access.

The scanning process connects to a GitHub repository, traces data flows across files, and provides users with findings categorized by CWE, along with confidence and severity ratings and suggested patches. Anthropic explained: "Claude Security uses Mythos 5 to scan code you own, and returns detailed findings rather than raw outputs without exposing the model itself," meaning defenders can access Mythos 5's capabilities without the model becoming accessible to those who might misuse it. Anthropic also launched the Defender Advantage Fund (0xDAF), putting $35 million in Claude credits toward organizations that help open-source maintainers secure their projects.

Grants will focus on patching live vulnerabilities in widely used projects, automating scanning and patching in ways other projects can replicate, and helping projects pursue security approaches that make them resistant to whole classes of attack. Anthropic is starting with a small number of larger pilot grants to learn what works and scales best, and will share details on initial recipients in the coming weeks. Traffic on Mythos-class models carries a 30-day retention requirement, which Anthropic says is used for safety purposes and not training.

Disclosure: Claude, which generates this brief, is built by Anthropic.