The AI Brief

Vol. I · No. 74 · Friday, August 7, 2026

Today's brief:

  • OpenAI makes GPT-5.6 Luna the default for all free ChatGPT users and removes text rate limits, while updating Sol with 68% fewer factual errors for paid tiers, the biggest single-day change to the product's one-billion-user surface since launch.
  • Anthropic signs a $10 billion, six-year compute deal with Volta Infra, a months-old Nvidia-backed startup, for a Norway data center on Vera Rubin chips, marking a new phase in which neoclouds with no operating history can land frontier-scale contracts.
  • OpenAI launches "Sign in with ChatGPT" in beta with six developer-ecosystem partners, positioning OpenAI as an identity provider in direct competition with Google, Apple, and Microsoft for the login layer across the web.
  • Update: Gemini 3.5 Pro misses August 7, the date prediction markets gave a 73% probability of launch, with no official release, no model card, and a Made by Google hardware event on August 12 now the next watch point.
  • The UK AI Regulation and Safety Bill cleared the House of Commons, with Royal Assent expected by October, the first binding UK statute specifically targeting frontier AI developers, requiring mandatory safety data sharing.

OpenAI Removes Free-Tier Limits and Upgrades Sol Across ChatGPT's Entire User Base

Why it matters
OpenAI just handed every free user a meaningfully better model, GPT-5.6 Luna with unlimited text chat, while giving paid users a Sol that produces 68% fewer factual errors on financial, medical, and legal prompts and a first-ever manual reasoning slider, collapsing the previous Instant/Thinking split into one adjustable model.
What's at stake
For most operators, this is context, not a decision. For teams evaluating ChatGPT as an enterprise default or building on the ChatGPT API surface, the Sol accuracy gains and unified reasoning-effort control change the baseline they are testing against, and the Luna-as-free-default sets the new floor for competitive positioning across AI consumer products.
Detail

OpenAI updated GPT-5.6 Sol for Plus and Pro subscribers and pushed GPT-5.6 Luna to everyone on the free tier, while removing the rate limit on text conversations for free users. In internal evaluations across financial, medical, and legal prompts, responses containing at least one factual error were approximately 68% less common with GPT-5.6 Sol compared to GPT-5.5 Instant. Luna reduced factual errors by 62% on the same benchmarks versus GPT-5.5 Instant.

Plus and Pro subscribers gained a new slider letting them control how much reasoning GPT-5.6 Sol applies before generating a response, replacing the previous separate Instant and Thinking modes with one adjustable, unified experience. Users can choose from Instant, Medium, High, and Extra High based on the complexity of the prompt. Free users also get a new "Think" button built into the chat composer; tapping it prompts GPT-5.6 Luna to spend extra processing time on complex logic before answering.

The changes arrive as OpenAI says ChatGPT has now reached one billion people every week, prompting the company to rethink how much intelligence it makes available across every tier. OpenAI clarified that this update applies strictly to general ChatGPT conversations, leaving the versions powering Codex and ChatGPT Work untouched. Limits continue to apply to file uploads, images, voice chats, and image-generation tools, with the unlimited text rollout expected to reach all eligible users within the coming week.

Sources
CaveatAccuracy figures (68%, 62%) are from OpenAI's own internal evaluations, not independent benchmarks. · OpenAI: GPT-5.6 Sol and Luna update (primary) · Help Net Security: OpenAI drops ChatGPT text chat limits · Android Headlines: OpenAI unlocks unlimited ChatGPT text chats

$10B
Anthropic's six-year compute commitment to Volta Infra, a startup that had no logo when the deal closed

Anthropic Bets $10 Billion on a Neocloud Startup With Months of Operating History

Why it matters
When a company founded in January by former Brookfield Asset Management executives can sign a $10 billion, six-year deal with a frontier AI lab within months of incorporation, the compute-procurement market has structurally shifted: chips-plus-power bundlers can now outcompete established clouds on availability if they move fast enough, regardless of operating track record.
What's at stake
For most operators, this is context, not a decision. For AI infrastructure investors and enterprises building sovereign or dedicated AI capacity, Anthropic's diversification across Amazon, Google, SpaceX Colossus, AMD, and now a neocloud illustrates the multi-supplier hedging strategy that frontier labs now treat as table stakes, and the execution risk embedded in deals where the counterparty still has to build the facility.
Decode
Neocloud = a new class of AI-focused data-center operator that finances, builds, and rents GPU capacity without owning the underlying chip supply chain, distinct from hyperscalers (AWS, Google Cloud, Azure) and from colocation providers. Examples include CoreWeave, Lambda Labs, and Volta.
Detail

Anthropic struck a $10 billion deal for computing capacity from a months-old infrastructure startup; the AI developer signed a contract to use a data center managed by Nvidia-backed cloud startup Volta Infra Holdings. The deal will see Volta partner with crypto-miner-turned-data-center-builder Bitdeer to build a 133-megawatt facility in Norway, running on Nvidia's Vera Rubin chip architecture.

When it emerged from stealth, Volta said it had raised about $300 million across its seed and Series A rounds, reaching a $2.4 billion valuation. Volta describes itself as a vertically integrated AI infrastructure business set up to finance, build, and operate AI factories, operating on the rent-a-GPU model like Nscale and CoreWeave. Volta was founded in January by former executives from Brookfield Asset Management.

Anthropic has aggressively expanded its compute capacity over recent months; it struck a deal with SpaceX to use all of the compute capacity at the Colossus 1 data center, gaining access to more than 300 megawatts of new capacity. Anthropic also struck its largest compute deal yet with Google and Broadcom as its run rate reached $30 billion. The Volta agreement adds another 133 megawatts of dedicated capacity to an infrastructure stack that now spans multiple continents and hardware generations.

Disclosure: Claude, which generates this brief, is built by Anthropic.

Sources
NotePrimary source is Bloomberg; paywalled. Figures cited from Bloomberg reporting as reprinted by Yahoo Finance and confirmed by TechCrunch (Anthropic acknowledged the deal). · TechCrunch: Anthropic signs $10B deal with Volta · Techgenyz: Anthropic Volta $10 billion deal

OpenAI Enters the Identity Market With "Sign in with ChatGPT," Targeting the Login Layer

Why it matters
An AI company becoming an identity provider is a structural move, not a product feature: the login layer gives OpenAI a persistent signal of user behavior across every site and tool where the button appears, beyond what any single chat session reveals, and it puts OpenAI directly on the same competitive axis as Google and Apple SSO.
What's at stake
For most operators, this is context, not a decision. For developers evaluating whether to integrate "Sign in with ChatGPT" into their products, the operational question is what OpenAI retains at the authentication-infrastructure layer versus what gets shared with the partner app, a gap that TechTimes has noted the launch announcement does not foreground, and that an active class action alleging undisclosed ChatGPT data flows to Meta and Google makes materially relevant.
Decode
Identity provider (IdP) = a service that authenticates a user's identity and vouches for them to third-party applications via a standard protocol (here, OAuth/OIDC). "Sign in with Google" and "Continue with Apple" are the dominant consumer examples. Becoming an IdP gives the provider a cross-site authentication footprint independent of its primary product.
Detail

On August 2, OpenAI launched "Sign in with ChatGPT" as a live beta, the company's first cross-platform identity system, with six named developer-ecosystem partners: Airtable, GitLab, HubSpot, Notion, Supabase, and Vercel. The beta makes OpenAI not just an AI assistant but a digital identity provider, placing it in direct competition with Google, Apple, and Microsoft for the login layer across the web.

When a user connects a supported app, they can create or link an account in fewer steps; only name, email, and profile picture are shared with the partner application. Sign in with ChatGPT is available globally to authenticated ChatGPT users, including Enterprise organizations, and is available on OpenAI Academy and Codex Sites in addition to the partner-site rollout.

What the launch announcement does not foreground is the gap between what partner applications receive when a user signs in and what OpenAI itself retains as the operator of the authentication infrastructure. The feature launched on the same day that the EU AI Act's enforcement and fining powers became active, August 2, 2026, adding regulatory context to a product that handles personal data across international user bases. OpenAI is also currently named in a class action alleging the company embedded Facebook Pixel and Google Analytics tracking code inside ChatGPT, silently transmitting user identifiers to Meta and Google; the case has not been adjudicated and OpenAI has not publicly responded.


Update: Gemini 3.5 Pro Misses August 7, Prediction Markets' Top Target Date Passes Without a Launch

Why it matters
Gemini 3.5 Pro has now missed five publicly tracked windows, June, July 17, July 24, July 31, and August 7, while Google's CTO Koray Kavukcuoglu has taken over DeepMind day-to-day operations and the Gemini 4 pretraining run has already begun, raising the question of whether 3.5 Pro ships as a standalone model or gets absorbed into a faster push toward the next generation.
What's at stake
For enterprise buyers and developers who signed contracts with OpenAI, Anthropic, or xAI models during the delay window, the practical question is no longer whether Gemini 3.5 Pro is late, it is whether a belated launch can recapture stack positions that have already stabilized around GPT-5.6 Sol and Claude Opus 5.
Detail

First covered in Vol. I, No. 68 (2026-07-13). Kalshi data shows traders wagered more than $320,000 on when Google will release Gemini 3.5 Pro; August 7 carried a 46% chance of release, rising to 63% by August 21 and 76% by August 31. August 7 was the single most-favored individual date in a separate prediction market that had tracked the model's delays, priced at 73% as recently as mid-July.

Google will hold its Made by Google event in New York City on August 12, where the company plans to introduce the Pixel 11 smartphone series; some market participants expect Google to reveal Gemini 3.5 Pro alongside the new devices, though Google has not confirmed whether the model will appear at the event or confirmed any of its features, pricing, access plans, or rollout schedule.

Gemini 3.5 Pro has missed consecutive launch deadlines after Google DeepMind's rebuilt model failed key reliability standards, including frequent hallucinations, and fell short of GPT-5.6 in benchmark tests. On July 21, Google confirmed that 3.5 Pro still isn't ready and mentioned almost as a footnote that it had started pretraining an entirely new flagship called Gemini 4. As of today, no gemini-3.5-pro model ID appears in Google's public API documentation.


UK AI Regulation and Safety Bill Clears Commons, Setting October Royal Assent in View

Why it matters
The UK becomes the second major jurisdiction after the EU to put a binding AI statute on the path to law, and unlike the EU AI Act, whose frontier-model obligations are under the AI Office in Brussels, the UK bill specifically targets foundation model developers with mandatory safety-data-sharing requirements that take effect upon Royal Assent.
What's at stake
For most operators, this is context, not a decision. For frontier AI developers with UK users or UK-registered entities, a list that includes OpenAI, Anthropic, Google DeepMind, and Mistral, October Royal Assent triggers a compliance clock on safety-data-sharing obligations that will require new disclosure infrastructure, not just policy updates.
Detail

The UK AI Regulation and Safety Bill cleared the House of Commons, with Royal Assent expected by October. Upon receiving Royal Assent, the bill will trigger mandatory safety data sharing for foundation models. The bill originated in the House of Lords, introduced by Lord Holmes of Richmond, before advancing to the Commons.

The UK bill arrives in a crowded regulatory week. The EU AI Act's core obligations came into force on August 2, and China issued its first fines under its new companion AI rules. In the United States, the "Great American AI Act" stalled in the House over a state preemption clause, leaving California and Colorado enforcing local rules. The UK bill fills a gap that has made the country an outlier: the European Union passed an AI Act in 2024, and as of earlier this year the United Kingdom had no AI Bill before Parliament.

The bill's practical scope for operators is narrower than the EU AI Act's full Annex III regime but more targeted: foundation model developers, the subset of the AI industry the UK government has consistently described as its regulatory priority, face the most immediate obligations. Any frontier lab with a UK entity, UK cloud infrastructure, or UK developer-facing API endpoints will need to assess coverage before October.

Sources
NotePrimary source is the UK Parliament bills record; the House of Commons passage date is cited from an aggregator and has not been independently confirmed from a Hansard or official Parliament announcement. Figures and framing cited from Cubbbix AI Regulation update. · UK Parliament: Artificial Intelligence (Regulation) Bill · Cubbbix: AI Regulation August 2026 global update