Reference
The State of Search
What is currently true about AI search, claim by claim. The archive answers what happened; this answers what's the latest — and says plainly which claims the record settles, which it contests, and which rest on a single source.
Last moved September 17, 2026.
Every claim traces to published wire items. Nothing here is asserted that the record has not already reported. A claim that has not moved shows the date it last did.
Can I block AI training without losing search traffic?
settled The record agrees. A change here would itself be news.
With Google and OpenAI, yes — both state in their own documentation that the training and search crawlers are independent. Anthropic runs separable crawlers but publishes no equivalent guarantee. In every case the block is a request rather than a lock — and on Cloudflare-proxied sites, the layer enforcing it can undo the operator's own separation: taking Cloudflare's September 15 default blocks Google's, Apple's and Bing's search crawlers too, because those crawlers are multi-purpose.
Verified against the record 2026-09-16 · record last moved September 16, 2026
The three operators the question actually turns on each run a search crawler and a training crawler that can be addressed separately in robots.txt. Two of them say in public what blocking one does to the other; the third does not.
Compliance is a separate question from policy. TollBit measured roughly 15% of the AI page fetchers it tracked in the first half of 2026 reaching European publisher URLs marked disallowed, with OpenAI's ChatGPT-User among the bots most often crossing blocks aimed specifically at it. No operator publishes a compliance rate against its own declared behavior, so a block is worth logging against rather than trusting.
Since August 21, 2026 a fourth party sits in this decision for Cloudflare-proxied sites. Bot Preference Sync conditions continued access for crawlers that both search and train on four public disclosures, which converts what was an operator's private practice into something a publisher can check.
The standard underneath any of this is itself still being drafted. Revision 08 of the IETF's AI Preference Vocabulary, published September 16, 2026 and co-authored by Google's Martin Thomson, split its single "AI Training" section into separate "AI Training" and "AI Use" sections while keeping "Search" as its own category — giving sites a more precise vocabulary to state the training-versus-search distinction this claim already tracks operator by operator. The draft expires March 18, 2027 unless renewed, and no operator has said whether it will implement the new categories.
Where it differs
- Google settled
- Documents it. Google-Extended "does not impact a site's inclusion in Google Search nor is it used as a ranking signal in Google Search." on the wire
- OpenAI settled
- Documents it. Its crawler documentation states a webmaster can allow OAI-SearchBot to appear in search results while disallowing GPTBot. on the wire
- Anthropic thin
- Runs ClaudeBot and Claude-SearchBot as separate agents, but its published help page does not say that blocking the first preserves the second. on the wire
- Cloudflare (proxy default, from Sept 15) moving
- Contradicts its own fourth condition on sites that take the default. Blocking Training also blocks Googlebot, Applebot and BingBot, because those crawlers are multi-purpose — the opposite of publicly showing the block doesn't hurt search. Applies to ad-bearing pages on new domains onboarding to Cloudflare; existing domains keep their current settings, and owners can opt out in Security settings any time before September 15. on the wire
What changed
- 2026-08-23 Cloudflare's Bot Preference Sync made continued access for dual-purpose crawlers conditional on four public disclosures — honoring the no-training preference, offering an AI-summary opt-out, reporting URL-level training visibility alongside search metrics, and publicly showing the block does not hurt search. on the wire
- 2026-08-27 Previously: Whether any major operator met Cloudflare's conditions was unanswered by the announcement. Checked against each operator's own documentation: none meets all four — Google three, OpenAI two, Anthropic none it documents. The condition all three fail is URL-level reporting of what was made available for training. on the wire
- 2026-08-30 Previously: Cloudflare's four conditions and its crawler-classification defaults were treated as separate layers — one about continued access for dual-purpose crawlers, one about what a site blocks by default. Cloudflare's September 15 default collapses that separation on sites that take it: Training and Agent are blocked by default on ad-bearing pages for new domains, and because Googlebot, Applebot and BingBot are multi-purpose, a Training block reaches them too — so the default itself produces the harm its own fourth condition requires an operator to show doesn't happen. on the wire
Still unresolved
- Whether Anthropic's separable crawlers carry the same guarantee. Its published pages do not say, and it has not been asked on the record.
- What a robots.txt block is worth per operator. TollBit's ~15% is an aggregate across tracked fetchers; nobody has published a per-crawler compliance rate.
- Which crawlers, if any, meet Cloudflare's four conditions. A press query to Cloudflare was drafted 2026-08-27 and the company has published no compliance list.
- Whether the September 15 default reaches existing free-plan domains, not just new ones. Cloudflare's post scopes it to new domains onboarding; several third-party summaries describe it reaching all existing free customers, and Cloudflare has not addressed the difference.
- Whether any operator implements the IETF's new Training/Use split, formalized in a September 16 draft revision, or the distinction stays informal the way it is today.
Evidence
- Each operator's documented position on training-versus-search independence, checked the same day against their own pages. on the wire wire analysis of published documentation — Cloudflare has published no compliance list
- That a declared block is not reliably observed. on the wire
- That the access decision is per-crawler rather than blanket, and that a trade body has formalized it into four verdicts. on the wire
- That Cloudflare's own September 15 default, not a hypothetical worst case, blocks Google's, Apple's and Bing's search crawlers on sites that take it, because those crawlers are multi-purpose. on the wire wire analysis of Cloudflare's own posts — scope for existing free customers is disputed and Cloudflare has not clarified it
- That the user-agent string alone is forgeable at scale, the identity gap behind Cloudflare's verification conditions. on the wire GreyNoise research — single-source, no affected organization named
- That the underlying standard for expressing training-versus-search preferences is itself still being drafted, with a September 16 revision formalizing the split this claim already tracks operator by operator. on the wire
Is AI-crawler access now a paywall rather than a block?
moving Actively changing — a rollout, a deadline, or a live fight.
On the Cloudflare-proxied publishers the wire tracks, yes for some — access is starting to split by whether the crawler's operator pays rather than by whether the site allows AI crawling at all. Four large publishers now block OpenAI's GPTBot outright while continuing to bill Anthropic's ClaudeBot and Perplexity's crawler through Cloudflare's Pay Per Crawl. Other large sites reject the split entirely and block every crawler, paying or not.
Verified against the record 2026-09-16 · record last moved September 16, 2026
A different kind of payment surfaced September 16, on the answer side rather than the crawler side. Google confirmed an early-stage pilot that pays a small set of publishers when their content "contributes significantly" to an answer in Gemini, AI Overviews or AI Mode — content that only confirms a fact or shows up as a post-answer link doesn't qualify. Participants see a monthly earnings figure inside a Search Console panel and can opt out any time; Google hasn't disclosed a rate or eligibility criteria, and sources described the earnings calculation as "quite black box" with early payouts small next to ad revenue. It meters something different from the Cloudflare split above: Pay Per Crawl charges for access before an answer is generated, this pays for contribution after one is.
The wire runs its own crawler access panel, probing 55 publisher domains daily as GPTBot, ClaudeBot, PerplexityBot and a browser user-agent. Between the September 12 and 13 captures, three People Inc. properties (Allrecipes, Investopedia, People) and The Atlantic made an identical change: GPTBot and the browser probe went from a 200 to a Cloudflare-managed 403, while ClaudeBot and PerplexityBot kept drawing the same 402 Pay Per Crawl response as before — the same site charging two crawlers for the access it now refuses a third.
That split sits on top of Cloudflare's Pay Per Use mechanism, announced July 1, 2026, which bases publisher compensation on how AI search actually uses a page rather than a flat per-crawl fee.
Not every large site is taking the paywall option. Quora hardened the same day from a redirect to a flat 403 for all four probes — GPTBot, ClaudeBot, PerplexityBot and the browser alike — and its robots.txt, readable to the panel for the first time, disallows all three crawlers by name and bars AI-training use outright. Reddit moved separately, on its own infrastructure rather than Cloudflare's, adding a 403 to browser and PerplexityBot requests that had drawn a 200 the day before.
Three days later the panel caught a second large publisher rejecting the split outright, and one moving the other way on a single crawler. El País's homepage flipped from a 200 to a 403 for GPTBot, ClaudeBot and PerplexityBot alike between the September 13 and 16 captures. Stack Overflow moved just Anthropic's ClaudeBot — off Cloudflare's Pay Per Crawl paid track and onto a flat 403 block — while GPTBot and PerplexityBot stayed behind the same Cloudflare challenge as before: the panel's first capture of a crawler moving from paid access to blocked rather than the reverse.
This is now two capture days from the wire's own panel — September 13 and September 16 — not yet corroborated by another outlet's measurement.
Where it differs
- OpenAI's GPTBot, at Allrecipes / Investopedia / People / The Atlantic moving
- Blocked outright — a Cloudflare-managed 403, matching what a plain browser request now draws at the same four sites, up from a 200 the day before. on the wire
- Anthropic's ClaudeBot and Perplexity's crawler, at the same four sites settled
- Unchanged — still met with Cloudflare's Pay Per Crawl 402, the same paid-access gate as before September 12. on the wire
- Quora moving
- Rejects the split: a flat 403 for GPTBot, ClaudeBot, PerplexityBot and a browser alike, and a newly-readable robots.txt disallowing all three crawlers by name plus a notice barring AI-training use outright. on the wire
- Reddit thin
- Tightened separately, on its own infrastructure: browser and PerplexityBot requests that drew a 200 the day before now draw a 403. This capture does not isolate GPTBot or ClaudeBot's status there. on the wire
- El País moving
- Rejects the split: a flat 403 for GPTBot, ClaudeBot and PerplexityBot alike as of September 16, up from a 200 on September 13 — a second large publisher, after Quora, blocking every crawler rather than billing some. on the wire
- Anthropic's ClaudeBot, at Stack Overflow moving
- Moved off Cloudflare's Pay Per Crawl 402 onto a flat 403 block as of September 16 — the panel's first capture of a crawler moving from the paid track to blocked rather than the reverse. on the wire
- OpenAI's GPTBot and Perplexity's crawler, at Stack Overflow settled
- Unchanged — still met with a Cloudflare challenge, unpaid and unblocked, as before September 16. on the wire
- Google — Search Console AI-answer payment pilot thin
- A separate, early-stage pilot, confirmed by Google: pays a small set of publishers for content that "contributes significantly" to an AI answer, not for crawl access. No disclosed rate, eligibility, or publisher count; sources call the earnings calculation "quite black box." on the wire
What changed
- 2026-09-16 Google confirmed a separate, early-stage pilot paying select publishers via a Search Console panel when their content contributes significantly to an AI answer — a payment tied to answer contribution rather than crawler access, alongside rather than instead of the crawler-payment split tracked above. on the wire
- 2026-09-16 El País blocked GPTBot, ClaudeBot and PerplexityBot outright, becoming the second large publisher after Quora to reject the payment split entirely. Separately, Stack Overflow moved Anthropic's ClaudeBot off Cloudflare's Pay Per Crawl track onto a flat block while leaving GPTBot and PerplexityBot behind the same Cloudflare challenge as before — the panel's first capture of a crawler moving from paid to blocked. on the wire
Still unresolved
- Whether the GPTBot block at the four Cloudflare sites is a publisher decision naming OpenAI specifically or a Cloudflare-side default reclassifying GPTBot — the panel logs the HTTP response, not who set it or why.
- Whether other Cloudflare Pay Per Crawl sites follow the same split, or whether this is four sites moving together (three under one owner, People Inc.).
- What a second day's capture shows — whether the split holds, widens to more sites, or reverses.
- Whether Google's Search Console AI-answer payment pilot becomes the search-side analog of Cloudflare's crawler-side Pay Per Crawl, or stays a separate, narrower thing — the two aren't yet comparable on eligibility, rate, or scale, and Google has disclosed none of the three for its pilot.
- Whether Stack Overflow's move of ClaudeBot from paid to blocked is a one-off or the start of publishers reversing out of Pay Per Crawl.
Evidence
- That four Cloudflare-proxied publishers split crawler access by payment on September 13 — GPTBot blocked, ClaudeBot and PerplexityBot still billed via Pay Per Crawl — while Quora and Reddit hardened separately without the same split. on the wire wire's own crawler access panel — self-sourced daily probe of 55 publisher domains, single day's capture, not yet corroborated elsewhere
- The Pay Per Crawl / Pay Per Use mechanism these sites are applying — usage-based payment for AI crawler access, announced by Cloudflare July 1, 2026. on the wire
- That Google is piloting a distinct, answer-contribution payment mechanism for publishers, separate from the crawler-access payment layer the rest of this claim tracks. on the wire Google-confirmed early-stage pilot — no disclosed rate, eligibility, or publisher count
- That El País blocked all three major crawlers outright and Stack Overflow moved one crawler from Cloudflare's paid Pay Per Crawl track to a flat block — enforcement-layer changes neither site's robots.txt alone would show. on the wire wire's own crawler access panel — second capture day, not yet corroborated elsewhere
Why don't AI visibility tools agree with each other?
settled The record agrees. A change here would itself be news.
Because they measure different engines, query sets and windows. Two trackers reporting opposite trends for the same subject are usually both correct, and the disagreement is methodological rather than factual — which is why the fact of disagreement stays stable while every number under it moves.
Verified against the record 2026-09-04 · record last moved September 4, 2026
The clearest worked example the record holds is Reddit in August 2026. BrightEdge reported on August 16 that Reddit's AI citation surge had plateaued at a baseline several times higher than a year earlier, and that Google AI Overviews accounts for roughly 88% of Reddit's AI citation volume. Three days later a tracker reported an 86.4% collapse. Both were right.
Promptwatch's own per-engine breakdown resolved it: ChatGPT Search fell 86.4% (3.83% to 0.52%), while AI Overviews slipped 11.3% and AI Mode 30.5% over the same window. The surface carrying ~88% of the volume barely moved. A decline reported widely as Reddit disappearing from AI search was one engine changing how it assembles background queries.
A separate, single-source measurement of that query-writing step shows why the change moved so much: brands ChatGPT named in its own background search query reached the final answer 68.9% of the time, against 2.1% for brands surfacing only on retrieved pages — a roughly 33x gap — and only 3.1% of 3,554 retrieved pages earned a citation at all. Read against Promptwatch's finding, the two measurements describe the same mechanism nine days apart: the query-writing step, not the page, is what the Aug 8 change actually altered.
Engine-to-engine spread inside a single study is wide enough to swamp most between-study comparisons: Writesonic measured "ghost citation" rates from 52% (Perplexity) down to 19% (Microsoft Copilot) across seven engines and roughly 16 million brand appearances. A figure quoted without its engine is not a figure.
The model underneath moves fast enough that naming the current one dates the paragraph. AI Mode's underlying Flash model has swapped four times since July 2026 — 3.5 Flash-Lite, then 3.6 Flash, 3.7 Flash on August 14, and 3.8 Flash on September 2 — each arriving roughly three weeks apart, which is faster than most visibility studies are conducted and published.
A day after the 3.8 Flash swap, two SEO practitioners reported AI Mode responses running on it showing markedly fewer citations than before — screenshots of several top-of-funnel queries returned no source links at all. Google confirmed it the next day: Robby Stein, Search's VP of product, said on X that citations were not working as intended and a fix is coming. Still no query sample or scope from either side, but the mechanism is now acknowledged rather than anecdotal — a bug, not a redesign.
A same-day finding complicates measurement further, upstream of any tracker: ChatGPT doesn't answer from one index. Peec AI reports OpenAI has built its own retrieval system, internally named Labrador, with separate vertical indexes for web, PDF, video and news — but a since-removed `result_source` field visible in ChatGPT's server-side events between May 21 and July 21, 2026 showed answers still drawing on Google, Microsoft's Bing, and two third-party scraping services alongside it. A dashboard built to approximate ChatGPT exposure from Bing rankings alone is reading one of at least four backends, not the whole surface.
What changed
- 2026-09-04 Previously: Whether Gemini 3.8 Flash's reported citation drop was real rested on two practitioners' screenshots, unconfirmed by Google. Google confirmed it. Robby Stein, Search's VP of product, said AI Mode citations generated with Gemini 3.8 Flash are not working as intended and a fix is coming soon — the finding moves from anecdotal to acknowledged, though still without a query sample or stated scope. on the wire
Still unresolved
- Nobody has published a head-to-head of the third-party measurers on the same query set. Until someone does, reconciliation is inference.
- Why ChatGPT's background query assembly changed on August 8. Promptwatch reads the forty-six-fold rise in `site:` operator use as the engine asking named sites directly; OpenAI has not commented.
- Whether query fan-out works the same way outside ChatGPT. Google describes a fan-out technique behind AI Mode, but no comparable measurement of its query-writing step has been published, so nothing here transfers to AI Overviews or AI Mode on evidence.
- How much of ChatGPT's answer volume each backend — Labrador, Google, Bing, the two scraping services — actually supplies, and whether that mix has shifted since the `result_source` field was pulled in July 2026. Single-source, confirmation pending.
Evidence
- That the two headline Reddit measurements were different engines, with per-engine numbers for all three surfaces. on the wire Promptwatch study — vendor-published and called provisional by its author
- That the query-writing step predicts final citation far better than retrieved-page content, connecting the ChatGPT query-fan-out measurement to the Aug 8 change behind the Reddit collapse. on the wire wire analysis connecting two single-source, ChatGPT-only measurements
- Where Reddit's AI citation volume actually sits (~88% Google AI Overviews), which is what makes the split decisive. on the wire BrightEdge study — vendor-published
- How wide the engine-to-engine spread is inside one methodology. on the wire Writesonic study — vendor-published
- That the model under a surface changes, dating studies measured on the prior one. on the wire
- That the swap cadence continued — a fourth Flash-family model in AI Mode within six weeks. on the wire
- That the newest swap may not just reshuffle citations but sharply cut their count on some query types. on the wire practitioner screenshots, no query sample or methodology — confirmed real by Google the next day
- That the drop was real rather than anecdotal: Google's Search VP of product confirmed AI Mode citations from Gemini 3.8 Flash are not working as intended and said a fix is coming. on the wire
- That ChatGPT blends at least four retrieval backends — its own Labrador index, Google, Bing, and two scraping services — so a dashboard reading one of them is not reading the whole surface. on the wire single-source — confirmation pending
- The standing caveats — citation share is engine-specific, model-sensitive, and methodology-dependent. · source ↗
Once I'm cited, do I stay cited?
settled The record agrees. A change here would itself be news.
Not reliably. For most prompts a small stable core of domains holds while everything around it rotates fast — the largest published measurement found Google replaces 56% of its AI-answer citations every week, ChatGPT as much as 74%. The instability is structural, not a sign anything is broken, and three separate measurements had already caught pieces of the same mechanism before this one supplied the base rate.
Verified against the record 2026-09-11 · record last moved September 11, 2026
SISTRIX re-sampled 82,619 prompts weekly for 17 weeks (1,548,213 snapshots, six countries, three surfaces) and found the headline churn sits on top of real structure: 86% of prompts keep a stable core of a few domains, with everything outside it rotating at 89% per week. On Google AI Overviews specifically, 53% of prompts saw no source change at all across the full window — aggregate churn and per-prompt permanence are both true, describing the tail and the core of the same answer.
The structure has consequences the record had not connected before. AI Overviews and AI Mode — one company, the same query, the same week — cite different domains 83% of the time, so a citation-share number measured on one Google surface is not a reading of the other. And permanence varies sharply by content type: only 1.4% of cited news articles held their spot across all 17 weeks, against 43% of brand queries that kept the brand's own domain present throughout even as the co-citations beside it rotated at 70% per week.
Three items on this wire measured pieces of this before the wire had a name for it. Kevin Indig's Search Signals analysis found AI Mode's publisher mentions falling nearly by half in thirty days while retailer mentions tripled, across 15 of 20 verticals — the engine's preference changing. Promptwatch traced Reddit's 86.4% four-day collapse in ChatGPT citations to a change in how the engine assembles its background queries — the retrieval method changing. And Profound shipped a feature called Citation Decay on August 13, 2026 that tracks week-over-week citation counts per URL — a vendor building the instrument this research argues is necessary. Three different mechanisms; SISTRIX supplies the base rate they were all sampling from.
A model swap can also break citation outright rather than just shift its rate. Two days after Gemini 3.8 Flash rolled out to Google AI Mode on September 2, SEO practitioners found some top-of-funnel queries returning no source links at all; Google's Search VP Robby Stein confirmed it was a citation bug on September 4 and said a fix was coming. It is the sharpest confirmed instance yet of the standing caveat that citation share moves when the model underneath does — this time to zero, briefly, on one model.
A follow-up BrightEdge measurement narrows the mechanism to a single vertical and a single engine, and answers a question the churn rates alone couldn't: what happens to a slot once it's vacated. Reddit's share of ChatGPT's Education-category citations fell from 17.2% to 0% over two weeks in August, and the domains that took its place — state.gov, asu.edu, apa.org, aacnnursing.org — still held those positions five weeks later. The slot didn't revert to variety; it was claimed by a narrower, more institutional set and stayed claimed. BrightEdge attributes the drop to Reddit's robots.txt policy; this wire's own August 27 reporting found a ChatGPT query-behavior change that preceded the decline by about six days, so the cause is still unsettled even as the outcome — persistence, not reversion — replicates.
What changed
- 2026-09-04 Previously: The model-swap caveat on citation share was a standing inference from base rates, not a documented instance of a swap breaking citation outright. Gemini 3.8 Flash's rollout to AI Mode on Sept. 2 was followed by reports of zero citations on some queries; Google confirmed it as a bug on Sept. 4 and said a fix was coming. on the wire
Still unresolved
- Whether the 56%/74% weekly rates still hold. SISTRIX's window closed April 8, 2026, before the model swaps this wire has covered since — and the standing caveat on citation share is that it moves when the model underneath does.
- Whether the 83% AI Overviews/AI Mode split is a stable feature of the two surfaces or is itself drifting.
- Why permanence differs so sharply between news articles (1.4%) and brand queries (43%) — content-type effect, query-type effect, or an artifact of how "brand query" was defined.
Evidence
- The headline weekly churn rates and the stable-core structure underneath them. on the wire SISTRIX study — vendor-published, not replicated; window closed 2026-04-08, before later model changes
- That a vacated citation slot doesn't revert to variety — it's claimed by a narrower set of domains that then holds, tested in a single vertical where Reddit's ChatGPT citation share went to exactly zero. on the wire BrightEdge study — vendor-published; robots.txt causal attribution is BrightEdge's own, not independently confirmed
- That a model swap can drop AI Mode citations to near zero as a bug, not just shift the rate — the sharpest confirmed instance of the model-swap caveat. on the wire confirms a report flagged single-source a day earlier
- Prior, independent evidence of the engine's preference shifting — one of drift's three mechanisms. on the wire Search Signals Index — Kevin Indig/Growth Memo analysis, self-published data
- Prior, independent evidence of a retrieval-method change driving drift for a single subject. on the wire Promptwatch study — vendor-published and called provisional by its author
- That a vendor is now building longitudinal tracking for exactly this behavior. on the wire Profound product launch — vendor's own framing
What can a publisher actually control about AI answers?
moving Actively changing — a rollout, a deadline, or a live fight.
Appearance and training are separate controls and no single lever covers both. Google now offers the only control that removes a site from AI Overviews and AI Mode without touching ordinary Search — it reached all sites worldwide on August 31, 2026, and it does not cover training.
Verified against the record 2026-08-31 · record last moved August 31, 2026 — nothing has moved it in 17 days
Four levers exist and each has a different scope. Choosing wrongly costs more than doing nothing: three of the four take ordinary Search results with them.
The price of the one that does what publishers asked for is stated plainly by Google — opt out and "you won't receive any traffic or impressions from these features" — against AI Overviews appearing on 43% of Google searches in July 2026, up from 15% a year earlier.
A regulator is already pushing on the gap. The UK CMA's Publisher Conduct Requirement, imposed June 3, 2026 and legally binding under the digital markets regime, requires Google to let publishers opt out of training, fine-tuning **and** the grounding behind AI Overviews and AI Mode, with nine months to comply. The Search Console control answers one of those three.
Google's own guidance has not caught up with its own control. "AI features and your website" — the Search Central page telling publishers how to approach inclusion — named `nosnippet` 18 times and the Search generative AI control zero times when checked on August 27, 2026. A publisher following the documentation rather than opening Search Console would not learn the control exists.
Where it differs
- noindex settled
- Removes the page from Search altogether. The bluntest lever, and the only one with no partial mode.
- nosnippet / data-nosnippet / max-snippet settled
- Limit what Search displays — and take the ordinary result snippet with them.
- Google-Extended settled
- Governs training and grounding for Gemini rather than appearance in Search's own AI features. A robots.txt token with no reporting attached to it.
- Search generative AI control (Search Console) settled
- Removes a site from AI Overviews, AI Mode and Discover's generative features in one to two days. Google says it "isn't used as a ranking or inclusion signal affecting other parts of Search." Rolled out to all sites worldwide as of August 31, 2026; does not cover training. on the wire
What changed
- 2026-08-31 Previously: Rolling out to a subset of site owners, with no published schedule for the rest. Google's Search Console documentation now says the control reached all sites worldwide, closing the rollout gap flagged four days earlier. Search Console's AI performance reports reached the same global availability. on the wire
- 2026-08-27 Previously: No control separated AI Overviews from ordinary Search: leaving one meant accepting damage somewhere else. Search Console's Search generative AI control was documented, closing the gap for appearance but not for training. Establishing this fact also unblocked the crawler-compliance story, which had stalled on exactly it. on the wire
Still unresolved
- Whether opting out costs Top Stories placement. Google places Top Stories carousels inside AI Overviews on roughly 15.5% of US and 17.5% of UK news searches, per NewzDash's John Shehata, who called the link high-confidence; Google has not addressed it.
- Whether the control satisfies the CMA order. It covers appearance; the order covers training, fine-tuning and grounding, with nine months to comply from June 3, 2026.
Evidence
- That the control reached all sites worldwide, and that Search Console's AI performance reports did too. on the wire
- What the control does, what it does not cover, and the caveat Google's ring-fence language leaves open. on the wire
- The Top Stories exposure a news publisher takes on by opting out. on the wire
- The binding regulatory requirement the control partially answers, and its clock. on the wire
Does llms.txt do anything?
thin One measurement or one source, without corroboration.
For Google, no — its documentation says outright that no AI text file is needed to appear in its AI features. For every other major engine there is no published answer either way, and the one large-scale measurement found 97% of llms.txt files were never read by an AI crawler.
Verified against the record 2026-08-31 · record last moved August 31, 2026 — nothing has moved it in 16 days
The evidence for whether the file is read is one vendor study and one vendor's documentation. That is thin for a tactic this widely recommended, and the thinness is the finding: nobody has published a reason to make the file, and only one company has published a reason not to.
Ahrefs, across roughly 15 million data points, found 97% of sites' llms.txt files were never read by an AI crawler, and that adding schema markup produced no measurable lift in AI citations across 1,885 pages over 30 days. The same research found 88.46% of AI citations still traced to pages in the general search index — the tactic that works is the old one.
The spec itself is still moving. Jeremy Howard published a version 2 update on August 10, 2026, its first revision since the format launched in 2024, adding `rel="alternate" type="text/markdown"` and `rel="describedby"` link relations so agents can find a page's Markdown version.
Google's John Mueller added a behavioral data point rather than a measurement: on his own test sites, "the only crawlers who claim to accept markdown are SEO tools" — no major AI bot did. He recommends logging accept headers before building a markdown version at all, which is the same conclusion the Ahrefs read-rate finding points to by a different route.
Common Crawl's analysis of 584,107 llms.txt files in its July 2026 archive found 68% were produced by templates or SEO plugins — Wix alone accounts for 41% of the corpus — with only about half fully following the spec's structure. It doesn't measure whether crawlers read the files, but it corroborates the same direction from a different angle: adoption is running mostly on CMS-plugin autopilot rather than deliberate authoring.
Where it differs
- Google settled
- Says it is unnecessary: "You don't need to create new machine readable files, AI text files, or markup to appear in these features." on the wire
- OpenAI, Anthropic, Perplexity, Microsoft thin
- No published statement on whether their systems read the file. Never asked on the record.
Still unresolved
- Whether any engine other than Google reads llms.txt. Four operators have never said, which is half a story sitting unfinished.
- Whether the v2 link relations change crawler behavior. The spec revision is three weeks old and no measurement post-dates it.
Evidence
- The only large-scale measurement of whether the files are read at all. on the wire Ahrefs study — vendor-published
- Google's own position, stated in its documentation rather than inferred. on the wire
- That the spec is under active revision, so measurements have a shelf life. on the wire
- A second, behavioral data point pointing the same direction as the Ahrefs read rate: on Google's own test sites, no major AI crawler requested markdown. on the wire anecdotal — Mueller's own test sites, not a published study
- That most llms.txt files in the wild are template-generated rather than hand-authored, a likely contributor to the low read rate. on the wire
Are AI answers costing me traffic?
contested Credible sources disagree, and the disagreement is unresolved.
It depends on what you publish, and part of the disagreement is instrumentation. Commerce reports AI referrals as growth; nonprofit and health publishers report the steepest declines; general-news outlets diverge — Le Monde reports none seven weeks after AI Overviews launched in France, while People Inc.'s Google search traffic fell 40% — and Google's own figure arrives without methodology. The one non-vendor academic estimate, a University of Washington study of Wikipedia, lands far below the self-reported figures at roughly 5%. Before any of it is compared, 22.4% of AI Overview traffic is misattributed away from Organic Search.
Verified against the record 2026-09-16 · record last moved September 16, 2026
The disagreement is real and it sorts by what a site sells. Shopify told its Q2 earnings call that AI-driven traffic and orders tripled year over year, with half of AI-referred sessions landing directly on product pages — 2.5 times the rate of traditional search. Over the same period Blood Cancer UK's leukaemia information page lost 53% of its page views. Both are first-party reports from parties with no incentive to overstate against their own interest.
One widely-cited framing does not survive the record. Zero-click searches hit a record high in June 2026 — 40% of U.S. searches sent a click to the open web — but the same report put Google AI Mode at 0.13% of U.S. search-related visits, usage it described as plateaued or dipping. A surface used by roughly one in a thousand visits cannot be the cause of a decline that broad, whatever else is.
Measurement sits underneath all of it. A nine-month GA4 study of 51,200 clicks found an average 22.4% of AI Overview traffic recorded as Direct rather than Organic Search, and Similarweb found 58.8% of ChatGPT referral traffic lands on publisher homepages even though 65% of cited URLs point two or three folders deep. A citation and the session it produces are not measured in the same place, which is why two honest parties can count the same channel and disagree. See the measurement-disagreement claim.
The response has a price tag now too. Similarweb's Top 100 Media publishers spent an estimated $113 million on Google paid search in July 2026, up 41% year over year and 274% over three years — Forbes alone spent $72.2 million, about two-thirds of the total, even as its own organic traffic fell 26.7% year over year. Buying back the clicks AI answers are keeping is a growing spend line, not an absorbed loss.
Le Monde's CEO complicates the uniform-decline framing directly: seven weeks after AI Overviews launched in France, the outlet logged 24,000 new subscribers in each of July and August 2026 and 7% growth in English-language traffic for H1 2026 versus H1 2025 — while French publishers' group APIG reported a 33-38% search-referral loss industry-wide. Louis Dreyfus's own account of the gap: search drives only about half of a major news site's visits, and the AI Overviews damage he's tracking is concentrated in service-content categories — health, food, automotive — down as much as 50%, not general news.
The only independent, non-vendor measurement in the record lands well under every self-reported figure above. A University of Washington paper estimates AI Overviews cut monthly search referrals to English Wikipedia by roughly 5% since the feature became Google's US default in 2024 — a figure the authors themselves revised down from an initial draft estimate near 15% after switching from daily pageviews to monthly referrals and adding international controls. Google disputes even the smaller number, saying Wikimedia's public referral data mixes in traffic from every search engine and can't isolate Google's effect; the authors counter that other engines carry too little share to account for the gap. The revision is itself informative: the same underlying effect produced a threefold-different headline depending on methodology, which is reason to treat every self-reported percentage above as provisional in the same way.
France now has independent, site-level evidence rather than one CEO's account. Ahrefs tracked 963 high-traffic French domains through Google Search Console for 28 days before and 9 days after AI Overviews launched in France on July 22, 2026: sites most exposed to AI Overviews (over 20% of their queries triggering one) lost 23.1% of CTR, against a 5.7% median CTR loss and 8.8% median click loss across all tracked domains, with nearly 48% of sites losing more than 10% of daily clicks within nine days. This is not in conflict with Le Monde's CEO — it corroborates his own account of where the damage sits. Ahrefs' worst-hit category was health (22.2% AI Overview rate, 20.4% CTR decline), a service-content vertical, matching the categories — health, food, automotive — Dreyfus named as carrying the losses he does see, while general news is where he reports none.
Where it differs
- Google (the engine) thin
- Says AI features in Search send "billions of clicks" to websites weekly. No methodology accompanied the claim, and it has published no supporting data. on the wire
- Commerce settled
- Growth. Shopify reported AI-driven traffic and orders tripled year over year in Q2, half of AI-referred sessions landing on product pages — 2.5x traditional search — and framed AI as a complement rather than a substitute. on the wire
- News publishers settled
- Decline, and not enough leverage to act on it. Google search fell to 21% of People Inc.'s traffic from 25% the prior quarter, with core sessions down 22% year over year and Google search traffic specifically down 40% — and the publisher still declined to block, saying that would "turn off search". on the wire
- Le Monde (hard news, France) thin
- No decline seven weeks after AI Overviews launched in France — 24,000 new subscribers in each of July and August 2026 and English-language traffic up 7% in H1 2026 versus H1 2025. CEO Louis Dreyfus attributes the sharpest AI Overviews losses to service-content categories (health, food, automotive, down as much as 50%) rather than news, and notes search drives only about half of a major news site's total visits. on the wire
- Nonprofit and health information settled
- The sharpest declines in the record. Blood Cancer UK's leukaemia page lost 53% of page views year over year, with 17-45% declines across its other blood cancer pages. Save the Children logged a 303% rise in AI Overview appearances while both clicks and impressions fell. on the wire
- The instrument settled
- Understates the channel before anyone compares it. An average 22.4% of AI Overview traffic was recorded as Direct rather than Organic Search across a nine-month, 51,200-click study — worst at 29.3% in May 2026. on the wire
- Publishers, buying it back settled
- Paying for what they lost. Similarweb's Top 100 Media publishers spent an estimated $113 million on Google paid search in July 2026, up 41% year over year and 274% over three years — Forbes alone accounted for $72.2 million (about two-thirds of the total) as its organic traffic fell 26.7% year over year; CNN and USA Today saw organic declines of 28.9% and 24.1% over the same span. on the wire
- Wikipedia (independent academic estimate) thin
- The lowest figure in the record, and the only one not self-reported. A University of Washington study puts AI Overviews' cut to English Wikipedia's search referrals at roughly 5% since 2024 — revised down from a draft estimate near 15% — against Google's dispute that the underlying referral data can't be isolated to its engine alone. on the wire
- France, site-level (Ahrefs, GSC data) settled
- The record's first independent, site-level read of a national AI Overviews launch. Sites most exposed lost 23.1% of CTR nine days after France's July 22, 2026 launch; the 963-domain median lost 5.7% of CTR and 8.8% of clicks. Concentrated in the same service-content categories — health worst of all — that Le Monde's CEO named as where the damage he sees actually sits. on the wire
What changed
- 2026-09-16 Previously: The France picture rested on APIG's industry-wide alleged search-referral loss (33-38%) against Le Monde's single-title account of no decline. Ahrefs' analysis of Google Search Console data across 963 French domains found sites most exposed to AI Overviews lost 23.1% of CTR in the nine days after the July 22 launch, with a 5.7% median CTR loss and 8.8% median click loss overall — independent, site-level evidence of the decline APIG alleged, concentrated in the same service-content categories (health worst-hit) Le Monde's CEO said carried the damage he does see. on the wire
- 2026-09-11 A University of Washington study added the first non-vendor, non-self-reported estimate to the record — roughly a 5% cut to Wikipedia's search referrals, well below every self-reported figure above, and itself revised down from an initial ~15% draft estimate after methodology changes. Google disputes the estimate. on the wire
- 2026-07-23 Alphabet reported Q2 Search revenue of $63.27 billion, up 17% year over year — the first deceleration after four quarters of accelerating growth. The money says the business has not broken, only slowed. on the wire
- 2026-08-03 Previously: The zero-click rise was widely attributed to AI answer surfaces. The same report that recorded the zero-click record put AI Mode at 0.13% of U.S. search-related visits and plateauing, which separates the two trends rather than linking them. on the wire
Still unresolved
- Google holds the only complete dataset and has published no methodology behind "billions of clicks". Until it does, the engine's own position is the least checkable one in the record.
- Whether the sector split is causal or compositional. Nobody has published a study that holds content type constant and varies only AI-surface exposure — Le Monde's CEO reports the same split by anecdote (news stable, service content down up to 50%), and Ahrefs' France data corroborates the direction (health worst-hit) but is still a single 9-day window from one measurer, not the controlled study this needs.
- What share of the zero-click record predates AI answers entirely. The decline is broader and older than AI Mode's 0.13% usage can explain, and no one has decomposed it.
- Whether misattribution is worsening. The 22.4% average spans nine months with a 16.8-29.3% range and no trend line published.
- Whether Google's confirmed automatic AI Overview expansion — skipping the "Show more" click for queries its systems select — shows up as a steeper decline than the pattern above. Google confirmed the behavior August 28-29, 2026, without disclosing what share of queries qualify or attaching any traffic data.
- Whether a narrower AI Overviews citations panel, spotted independently by two observers and replicated by Search Engine Roundtable on September 8, 2026, would further shrink the one visible click path this claim already tracks. Google has not confirmed the test or said whether it will roll out broadly.
- Whether paid-search buyback is a widespread publisher strategy or concentrated in a few large spenders — Forbes alone was two-thirds of the $113 million total, and no breakdown below the top spenders has been published.
- Whether the self-reported declines above (People Inc., Blood Cancer UK, APIG) would shrink under the same independent methodology the University of Washington authors applied to Wikipedia — nobody has run that methodology on a commerce, health or news publisher's own referral data.
Evidence
- The engine's own claim, and that it arrived without data. on the wire vendor statement, no methodology published
- That AI referrals read as growth in commerce, first-party. on the wire vendor-published, earnings call
- That the instrument understates the channel before comparison. on the wire
- That AI search is layering on top of traditional search rather than replacing it, and that the cited URL and the landing page differ. on the wire
- That crawl volume and referral value are now separately measurable, and diverge sharply. on the wire
- That Google has confirmed AI Overviews can skip the "Show more" click entirely for some queries, without disclosing the trigger or attaching traffic data. on the wire confirms a behavior first flagged unconfirmed on 2026-08-27; no traffic data attached
- That Google is testing a narrower AI Overviews citations panel that could make source links less prominent, unconfirmed. on the wire unconfirmed test, spotted independently twice; no rollout confirmed
- That publishers are increasingly paying for the search traffic AI answers are taking from them organically for free — a growing spend line, not just an absorbed loss. on the wire Similarweb data via Adweek/PPC Land — vendor-published
- That at least one hard-news outlet has seen no AI Overviews traffic decline, with the CEO locating the damage he does see in service-content categories rather than news. on the wire single CEO interview, unconfirmed independently
- The record's only non-vendor, non-self-reported estimate — and that the same underlying effect produced a threefold-different headline (15% to 5%) depending on methodology, before Google's dispute is even factored in. on the wire single academic study, disputed by Google, not independently replicated
- Independent, site-level Google Search Console evidence of AI Overviews' France launch effect on CTR and clicks, corroborating APIG's alleged decline and consistent with Le Monde's CEO locating the damage he sees in service-content categories. on the wire Ahrefs study — vendor-published, 9-day post-launch window, not independently replicated
seo · google-search · research
What on-page work actually moves AI citation?
settled The record agrees. A change here would itself be news.
The levers that measure are the ones that were already SEO — being in the search index, ranking well, being reachable in plain HTML, and keeping pages current. The AI-specific file formats measure near zero.
Verified against the record 2026-09-09 · record last moved September 8, 2026
The single largest number in the record points back at classic search: 88.46% of AI citations traced to pages already in the general search index, and 76% of the passages Google reused 100 or more times came from pages already ranking first organically. Google's own position is the same — John Mueller said there is "nothing really special you need to do for generative AI responses in search".
That is not the same as saying nothing is specific to AI. Three levers measure, and all three are mechanical rather than semantic: whether a crawler can reach the page at all, how recently it changed, and whether the passage is shaped to be lifted. The AI crawlers are stricter than Googlebot on the first — GPTBot, ClaudeBot and Bingbot found none of a site's JavaScript-injected internal links across 41 days, and recovered 250 pages within 48 hours of the links being converted to HTML.
The tactics named after the category are the ones that do not measure. An analysis spanning roughly 15 million data points found 97% of llms.txt files were never read by an AI crawler, and adding schema markup produced no measurable citation lift across 1,885 pages over 30 days. See the llms-txt-effect claim, which holds that question's own open threads. A separate study re-tested the field's original 2023 citation-lever effect sizes directly against ten modern AI engine families and found none of them moved citations any longer; a content score built from those levers correlated with actual citations at just 0.11 within-query, positioning it as a quality filter rather than a citation predictor.
Where it differs
- Search-index presence settled
- The dominant lever. 88.46% of AI citations trace to pages in the general search index; 76% of passages reused 100+ times already ranked #1 organically. on the wire
- Crawlability in plain HTML settled
- Decisive, and stricter than Googlebot. Over 41 days GPTBot, ClaudeBot and Bingbot reached none of a 2,400-page site's JavaScript-linked pages; Googlebot reached 2% against GoogleOther's 48%. on the wire
- Recency settled
- Measures, and varies by engine. Of 47,097 citations, 75% of cited pages had been updated within a year and 88% within two — Gemini 78%, ChatGPT 73%, Perplexity 65%. Last-modified date tracked citation better than original publish date (72% vs 42%). on the wire
- Passage shape settled
- Observable in what gets reused. Across 15.7 million AI Mode citations, 80.9% of passages were cited once, while roughly 2,300 were reused 61+ times and skew short (117-word median), answer-first and self-contained. on the wire
- AI-specific files (llms.txt, schema) thin
- No measurable effect. 97% of llms.txt files were never read by an AI crawler, and schema markup produced no citation lift across 1,885 pages over 30 days. on the wire
- 2023-era GEO content levers settled
- Expired on modern engines. Re-testing the original 2023 citation-lever effect sizes on ten modern AI engine families found none of them moved citations; a content score built from those levers correlated with actual citations at just 0.11 within-query. on the wire
What changed
- 2026-08-18 Schema markup moved from an assumed AI-citation lever to a measured null: no lift across 1,885 pages over 30 days, in the same analysis that found 97% of llms.txt files unread. on the wire
- 2026-08-24 Google stated the position directly — "there’s nothing really special you need to do for generative AI responses in search" — framing AI Overviews and AI Mode as drawing on the standard search index. on the wire
- 2026-09-01 Previously: One synthetic-benchmark paper found the optimal single-strategy approach could fall below an unoptimized baseline as competition increased — early formal evidence, not yet tested against real citation data. A second simulation, CHASE, ran the same concern forward: across 20 rounds of content creators rewriting to match ranking signals, alignment between ranking and independently assessed quality declined in all six domains tested, isolated to the optimization incentive itself by a control run that showed no such decline. The authors first validated ranking as a stand-in for LLM citation at a 0.853 rank-citation AUC. Still simulation, not measured against real citation data — but two independent papers now model the same degradation from different angles. on the wire
- 2026-09-09 A study validating a deterministic GEO content-quality score re-tested the field's original 2023 citation-lever effect sizes against ten modern AI engine families and found none of them moved citations — recalibrating the score to current data stripped it of its lever-responsive components entirely. The score's own within-query correlation with actual citations was weak (Spearman 0.11), positioning content scores as quality filters rather than citation predictors. on the wire
Still unresolved
- Whether entity phrasing is a content lever. Google's own paper found models fail to recall 26-34% of facts they encoded and tied part of the gap to subject/object order, but the researchers stop short of calling it something a publisher can act on.
- Whether the recency effect is causal or compositional. Actively maintained pages may simply be better pages; no study holds quality constant and varies only update date.
- Whether passage shape can be induced. Pillarbase's finding is observational across published pages — nobody has run the experiment of rewriting a passage to the observed shape and measuring citation change.
- Whether serving a stripped machine-readable page to agents helps or is penalized. BrightEdge shipped Agent Edge on that premise on 2026-08-26; Perplexity blocked Time's agent-facing markdown pages on 2026-08-11. The record contains a product and a punishment and no measurement.
- Whether GEO content-rewriting strategies stay effective as more competitors adopt them, and whether repeated optimization degrades the content pool itself. Two synthetic-benchmark papers now model adjacent parts of this: one found the optimal strategy shifts as adoption grows, with single-strategy methods falling below an unoptimized baseline once competition is high; a second (CHASE) found ranking-quality alignment declining across 20 rounds of optimization in every one of six domains tested. Both are simulations — neither has been tested against real citation data in the wild.
Evidence
- That index presence dominates, and that the two named AI-specific tactics measure near zero. on the wire vendor-published (Ahrefs), ~15 million data points
- That AI crawlers are stricter than Googlebot on JavaScript-rendered links. on the wire single-site experiment, 2,400 pages over 41 days
- That recency tracks citation likelihood, and by how much per engine. on the wire vendor-published (Seer Interactive), four brands
- What shape a repeatedly-reused passage has. on the wire
- Google's own stated position that no separate optimization target exists. on the wire spokesperson statement on social, not documentation
- That the entity-order finding is a model-recall result its own authors do not extend to publisher practice. on the wire
- That the optimal GEO content-rewriting strategy is competitor-dependent rather than fixed, per a competitor-aware benchmark. on the wire arXiv preprint, synthetic benchmark — not tested against real citation data
- That repeated content-rewriting under a ranking incentive degrades ranking-quality alignment in simulation, across six domains, isolated to the optimization pressure itself by a control run. on the wire arXiv preprint, simulation framework (CHASE) — not tested against real citation data
- That the specific GEO levers identified in 2023 studies no longer move citations when re-tested against ten modern AI engine families, and that a deterministic content-quality score correlates only weakly (Spearman 0.11) with actual citations. on the wire arXiv preprint, 500-source adversarial benchmark — not yet peer reviewed
Can I buy my way into AI answers?
moving Actively changing — a rollout, a deadline, or a live fight.
You can buy a slot, but not a citation — they are separate games. Ads run on about a quarter of ChatGPT's commercial prompts and nearly one in three commercial AI Mode queries, and only 3.63% of the advertisers whose ads ran were also cited as a source in the answer above them. The ad stack is scaling faster than the reporting under it.
Verified against the record 2026-09-17 · record last moved September 17, 2026
The ad formats keep widening inside AI Mode, not just below it. Google confirmed testing regular Search-campaign text ads — for now limited to exact and phrase match keywords carrying "explicit and direct" purchase intent — inside AI Mode responses, on top of the shopping ads carousel that arrived August 31. Google has not said how long the test will run or how many advertisers are included.
The paid layer arrived quickly and is still moving. ChatGPT Ads launched in the US on February 9, 2026, reached nine markets by August 17 — its first Spanish- and Portuguese-language ones — and roughly 40 countries two days later when 31 European markets went live at once. The business crossed $1 billion in annualized revenue run rate in under 200 days after launch, per OpenAI, and self-service Ads Manager access opened to India, Europe, the Middle East and North Africa on August 31 — previously a managed-sales and partner-led market. Ads remain limited to Free and Go plans. On September 10, Amazon Ads began piloting a third route to that inventory: buying ChatGPT Ads placements through Amazon's own DSP for select US advertisers, with Amazon managing the buy and OpenAI still controlling how and where ads appear. No timeline was given for expanding beyond the pilot. The next day, Adform became the first European-headquartered name on OpenAI's technology platform partner roster, letting its clients manage ChatGPT campaigns alongside their other Adform-run media; Volkswagen and Vodafone were named as early clients. On September 16, OpenAI moved distribution into the software marketers already run their day in: HubSpot became ChatGPT Ads' first CRM integration and Shopify its first ecommerce integration (U.S. merchants first, other countries from September 23), and OpenAI began testing Sponsored Agents — a labeled ad format that opens a separate conversation with the advertiser's own AI agent instead of sending a click to its site — with select US advertisers. Ads Manager also gained per-platform targeting and reporting (Android, iOS, web, desktop) and 7/14/30-day attribution windows, up from a fixed one day.
The format itself is moving into the answer, not just below it. As of August 31, 2026, Google began showing a shopping ads carousel inline within AI Mode's response, alongside the existing separate ad unit at the bottom of results — reported from a single ad-format tracker, with Google neither confirming the change nor stating a rollout scope.
Paid presence and cited presence are close to unrelated, which is the finding that matters for anyone treating ads as an AEO shortcut. Across more than 50,000 commercial prompts in 20 niches, 3.63% of advertisers running an ad were also cited as a source in the answer above it, and 14.35% of the ads shown were unrelated to the surrounding conversation. The same separation shows on Google: an analysis of AI Mode text ads found advertisers rarely among the answer's cited sources.
Two coverage figures circulate and they are not in conflict — they use different denominators. SE Ranking found ads on 25.94% of commercial prompts; Adthena, measuring all US queries rather than commercial ones, found 4.47% at an average of 1.06 ad items per response, against a 3.53-item average on Google's AI surfaces. Read the first as ad load on the queries advertisers want and the second as ad load across everything.
Category concentration is shifting under retail's early lead. Sensor Tower's April-August window found financial services growing from one top-100 advertiser to four of the top ten spenders, and from 2% to 13% of US ad spend, while shopping and retail's share fell from 37% to 21% and travel rose from 5% to 11%. Comscore's new panel-based tracking corroborates the travel shift from a different angle: sponsored links on hotel-related ChatGPT prompts rose from 6% to 24% between March and May, measured from real opted-in user prompts rather than synthetic tests.
Reporting lags the levers throughout. Campaign-level platform targeting shipped in August into an Insights view that still groups results as Mobile and Desktop, so a web-only campaign reports entirely as Mobile. View-through conversions appear as their own column but are excluded from CPA, bidding and billing. Automated bidding became the default for new ad groups carrying no guarantee against a CPA, CPC or ROAS target. Third-party tooling is currently the only cross-platform view of what is actually running.
Where it differs
- ChatGPT — ad load settled
- 25.94% of more than 50,000 commercial prompts across 20 niches carried an ad; across all US queries the figure is 4.47%, averaging 1.06 ad items per response. on the wire
- Google AI Mode — ad load settled
- Ads on nearly one in three commercial-keyword queries, and an average of 3.53 ad items per response across Google's AI surfaces — a heavier load than ChatGPT's. on the wire
- Overlap with citation settled
- Near zero. 3.63% of advertisers whose ads ran were also cited as a source in the answer above them, and 14.35% of ads shown were unrelated to the conversation around them. on the wire
- Category concentration — ad placement (Adthena, Mar-May) settled
- Retail and fashion drew 39% of observed US ad placements on 24% of query volume. Logistics and home-and-garden carried the highest ad frequency, at 12.41% and 11.99%. on the wire
- Category concentration — ad spend (Sensor Tower, Apr-Aug) settled
- Financial services grew from one top-100 advertiser in April 2026 to four of the top ten spenders in August, and from 2% to 13% of total US ad spend; shopping and retail fell from 37% to 21% over the same span, and travel rose from 5% to 11%. on the wire
- Ad density and advertiser growth settled
- Ads shown per user per hour rose 163% from April to August 2026, averaging 26% monthly growth; roughly 1,200 unique advertisers were active by August, up 43% month over month since May. The carousel ad format tested from August 7 reached 23% of US desktop ad impressions in the back half of August. on the wire
- Independent measurement — Comscore (hotel vertical) thin
- Comscore, tracking real opt-in panel prompts rather than synthetic ones, found sponsored links in hotel-related ChatGPT prompts rose from 6% in March 2026 to 14% in April to 24% in May — its first tracked category. Panel size and composition undisclosed. on the wire
- Google AI Mode — ad placement thin
- As of August 31, 2026, a shopping ads carousel appears inline within the AI Mode response itself, in addition to the existing ad unit below it — a single tracker's finding, not yet confirmed by Google. on the wire
- Google AI Mode — text ads thin
- Google confirmed testing regular Search-campaign text ads inside AI Mode, limited for now to exact and phrase match keywords with explicit purchase intent — a second ad format moving into the response, alongside the shopping carousel. No disclosed scope or duration. on the wire
- Measurement thin
- Behind the controls it is meant to measure. Platform targeting reports only Mobile/Desktop; view-through conversions are excluded from CPA, bidding and billing; automated bidding is the default with no performance guarantee. on the wire
- Revenue thin
- $1 billion in annualized revenue run rate in under 200 days after launch, per OpenAI — self-reported and unaudited, not an actual (non-annualized) revenue figure. on the wire
- Distribution — CRM and ecommerce integrations moving
- HubSpot and Shopify became ChatGPT Ads' first CRM and ecommerce integrations on September 16, letting merchants and sales teams manage campaigns without leaving those platforms; Shopify reaches non-US merchants September 23. The same rollout added per-platform targeting and reporting and 7/14/30-day attribution windows, up from a fixed one day. on the wire
- Sponsored Agents test thin
- A labeled ChatGPT ad format that lets a user click into a separate conversation with the advertiser's own AI agent, rather than an outbound link — testing with select US advertisers since September 16. No pricing or transcript visibility disclosed. on the wire
What changed
- 2026-08-19 Previously: ChatGPT Ads ran in nine markets, all English-language until August 17. 31 European markets went live at once, bringing the pilot to roughly 40 countries about six months after the February 9 US launch. on the wire
- 2026-08-21 Automated bidding became the preselected default for new ad groups. The strategy carries no guarantee against a CPA, CPC or ROAS target, which moves spend risk onto advertisers who do not actively opt into a manual bid ceiling. on the wire
- 2026-08-31 Google began showing a shopping ads carousel inline within AI Mode's response for the first time, alongside the existing ad unit below the results — the ad format's first move into the answer itself rather than beside it. Single-source and unconfirmed by Google. on the wire
- 2026-08-31 Previously: OpenAI's enterprise CMO put ad revenue growth above 25% since the start of August, with roughly 40 countries live as of August 19. OpenAI reported $1 billion in annualized revenue run rate in under 200 days after launch, and opened self-service Ads Manager access to India, Europe, the Middle East and North Africa — previously managed-sales only. on the wire
- 2026-09-03 Previously: Retail and fashion drew 39% of observed US ad placements on 24% of query volume, per Adthena's March-May window. Sensor Tower's April-August window shows the concentration shifting: financial services grew from one top-100 advertiser to four of the top ten spenders, its share of US ad spend rising from 2% to 13%, while shopping and retail's share fell from 37% to 21% and travel rose from 5% to 11%. Ad density (ads shown per user per hour) rose 163% over the same period, averaging 26% monthly growth, with roughly 1,200 unique advertisers active by August. on the wire
- 2026-09-03 Comscore began tracking sponsored ChatGPT links via opt-in panel data on real prompts rather than synthetic tests. In its first tracked category, hotel-related prompts carrying a sponsored link rose from 6% in March to 14% in April to 24% in May — a fourfold increase in three months, corroborating the travel-spend shift Sensor Tower measured separately. on the wire
- 2026-09-04 Google confirmed testing keyword-targeted Search text ads inside AI Mode responses, limited to exact and phrase match keywords with explicit purchase intent — a second ad format moving into the answer, after the Aug. 31 shopping carousel. on the wire
- 2026-09-10 Amazon Ads began piloting ChatGPT Ads inventory access through its own DSP for select US advertisers, with Amazon managing the buy and OpenAI still controlling ad placement — a third route into ChatGPT ad inventory alongside OpenAI's self-service Ads Manager and managed sales. No timeline given for expanding beyond the pilot. on the wire
- 2026-09-11 Adform became the first European-headquartered name on OpenAI's ChatGPT Ads technology platform partner roster, letting its clients manage ChatGPT campaigns alongside their other Adform-run media and connect ChatGPT-driven conversions to a single customer-journey view. Volkswagen and Vodafone named as early clients. on the wire
- 2026-09-17 OpenAI added HubSpot and Shopify as ChatGPT Ads' first CRM and ecommerce integrations, and began testing Sponsored Agents — a labeled ad format that opens a separate conversation with the advertiser's own AI agent instead of an outbound click — with select US advertisers. Ads Manager also gained per-platform targeting and reporting (Android, iOS, web, desktop) and 7/14/30-day attribution windows, up from a fixed one day. No pricing or transcript visibility disclosed for Sponsored Agents. on the wire
Still unresolved
- Whether buying an ad affects the odds of being cited at all, in either direction. The 3.63% overlap is a snapshot of co-occurrence, not a test — nobody has run the same queries with and without a live campaign.
- What ChatGPT ads cost. OpenAI publishes no benchmarks, and making an unguaranteed automated strategy the default puts spend risk on advertisers who have none.
- Whether Sponsored Agents — the format that opens a conversation with the advertiser's own AI agent instead of sending a click to its site, now testing with select US advertisers since September 16 — reaches general availability, what it costs, and whether advertisers see transcripts of those agent conversations. OpenAI disclosed none of the three for the test.
- Whether Google's AI Mode ad load holds. The 3.53-item average is one measurement window and Google has published nothing of its own.
Evidence
- Ad load on commercial prompts, the share of ads unrelated to the conversation, and the overlap between advertising and being cited. on the wire vendor-published (SE Ranking), 50,000+ prompts across 20 niches
- Ad load across all US queries, items per response, and category concentration. on the wire vendor-published (Adthena), ~850,000 US and UK queries, March-May 2026
- That advertisers are rarely among the cited sources on Google's AI Mode. on the wire
- The pace and shape of geographic rollout, and OpenAI's own revenue-growth figure. on the wire vendor statement for the revenue figure
- That the reporting layer lags the targeting controls it is meant to measure. on the wire
- That a third-party view of live creatives and landing pages now exists across ChatGPT and Google's AI surfaces. on the wire vendor-published (Similarweb)
- That AI Mode's shopping ad carousel now appears inline in the response, not only below it. on the wire single-source (ad-format tracker Brodie Clark, via Search Engine Roundtable) — Google has not confirmed
- The $1 billion annualized revenue run-rate milestone and the self-service Ads Manager rollout to India, Europe, the Middle East and North Africa. on the wire OpenAI-reported — annualized run rate, not audited revenue
- That financial services jumped to a top-ten ChatGPT ad-spend category and retail's share fell, alongside overall ad density growth and carousel-format adoption. on the wire vendor-published (Sensor Tower)
- That sponsored links in hotel-related ChatGPT prompts quadrupled over three months, measured via opt-in panel data. on the wire vendor-published (Comscore), panel size and composition undisclosed
- That Google is testing keyword-targeted Search text ads inside AI Mode responses, not just Shopping/PMax feeds. on the wire Google-confirmed small experiment; no scope or duration disclosed
- That Amazon Ads now resells ChatGPT Ads inventory through its own DSP, a third buying channel alongside OpenAI's self-service and managed-sales tracks. on the wire Amazon-announced pilot, US only, no named timeline or advertiser count beyond one pilot participant
- That Adform is the first European-headquartered technology platform partner in ChatGPT Ads' roster, giving European agencies and advertisers a local route to manage ChatGPT campaigns. on the wire reported by PPC Land and ExchangeWire, no OpenAI statement cited
- That ChatGPT Ads added HubSpot and Shopify integrations and is testing Sponsored Agents, a Business Agent-conversation ad format, with select US advertisers. on the wire reported by PPC Land and Search Engine Roundtable, no OpenAI statement cited; pricing and transcript visibility undisclosed
engines · tools · aeo · google-search
Will the courts stop AI engines from using my content?
moving Actively changing — a rollout, a deadline, or a live fight.
Not so far. The one concluded case produced a price rather than a prohibition — roughly $3,000 per work — while 143 AI copyright suits remain pending nationwide, 27 against OpenAI and 15 against Microsoft. Courts are reaching opposite results on near-identical theories, a First Amendment defense has appeared from three sets of defendants in six months, the Justice Department has now filed its first formal position on the fair-use question underlying nearly all of them — backing it — a new suit has added a trademark-dilution theory that stands apart from copyright altogether, and the first federal appeals ruling on an AI-output copyright question, from the Ninth Circuit, went to the AI company.
Verified against the record 2026-09-17 · record last moved September 16, 2026
One case has produced a number. The Bartz v. Anthropic book-piracy class settlement became effective August 20, 2026. Class counsel now says the first payout — $2,203.56 per work — goes out on or before Nov. 15, 2026, later than the Sept. 17 date estimated when the settlement took effect: consolidated payout statements go out Sept. 4, then a 30-day window for class members to contest them runs before money moves. A second payment, funded by a further $450 million Anthropic owes plus interest, is expected to bring the total toward the roughly $3,000 per work estimated earlier. Two appeals filed since challenge only attorneys' fee awards and by the agreement's terms do not affect the payout. It is the first figure in the record that prices training on pirated books, and it arrived as a settlement rather than a ruling — so it binds nobody else.
The theories are splitting rather than converging, sometimes within days. A federal judge dismissed most of Google's DMCA anti-circumvention claims against SerpApi on the reasoning that Google's anti-bot system protects ad revenue rather than a copyrighted work; days later a different judge let Reddit's near-identical DMCA claims against Perplexity and SerpApi proceed. Contributory infringement has moved the other way and closed: the New York Times, Daily News and Ziff Davis had those claims dismissed with prejudice against Microsoft after the Supreme Court's Cox Communications decision foreclosed the theory.
Agentic browsing got its first appellate answer, and it favored the agent. The Ninth Circuit reversed an injunction barring Perplexity's Comet browser from Amazon, finding Comet unlikely to violate the Computer Fraud and Abuse Act because it acts on the user's direction rather than accessing servers on its own.
A defense posture is forming in parallel. Musk, Tesla and Warner Bros. Discovery raised a First Amendment defense in an AI-copyright dispute in March 2026; Anthropic added one in the Gilbert case on August 20; Perplexity raised one in Reddit's DMCA suit on August 28. Separately, the pressure that has moved fastest is not copyright but competition — Judge Mehta, who already found Google's search business an illegal monopoly, said in a hearing that Google's use of publisher content in AI Overviews "seems really unfair".
The federal government took a side for the first time on September 1. The Justice Department filed a Statement of Interest in the OpenAI copyright litigation arguing that training on copyrighted text is a transformative fair use, and calling plaintiffs' "market dilution" theory — that AI outputs harm a work's market merely by resembling its genre — "deeply flawed." A Statement of Interest under 28 U.S.C. § 517 states the government's view; it does not bind the court deciding the case.
A new theory arrived September 4, outside copyright entirely. The Seattle Times and Newsday's joint suit against OpenAI and Microsoft adds a Lanham Act trademark-dilution count, arguing that ChatGPT and Copilot outputs which hallucinate content and misattribute it to the papers tarnish their marks — a claim that would survive independently of however the fair-use question is decided.
The DOJ's position didn't stay in the case it was filed in. On September 4, ROSS Intelligence — a defendant-appellant in Thomson Reuters v. ROSS Intelligence, one of the earliest AI-training fair-use rulings, now on appeal to the Third Circuit — filed a letter of new authority citing the DOJ's Statement of Interest and arguing it applies broadly to AI training, not only to generative AI. It is the first sign of the DOJ's position being invoked outside the case it was filed in — by a party, not yet by a court.
A new track opened September 9, aimed at boards rather than the underlying infringement. Katelyn Gray filed a shareholder derivative suit against Satya Nadella, other Microsoft directors and officers, and Microsoft itself, arguing they approved or were exposed to copyright infringement risk and misrepresented it to shareholders — the third such "copyright shareholder derivative" suit against Microsoft's board, expected to be consolidated with the two before it into one stockholder derivative action. The theory tests board oversight of copyright risk rather than whether the underlying training was infringing.
A dispute over whether a promise not to sue can pre-empt a fair-use ruling opened September 11. Anthropic opposed Daniel Gilbert's motion to dismiss the fair-use counterclaim it filed against him in his pro se suit over *Hacking World of Warcraft*, arguing his covenant not to sue over "training copies" does not moot the counterclaim because it rests on his still-pending infringement claim and his position that the court cannot weigh Anthropic's purpose for copying — not on fear of a second suit.
A second appellate answer arrived September 16, again favoring the AI company but for a narrower reason than the lower court gave. The Ninth Circuit affirmed dismissal of a claim that GitHub Copilot violated copyright law by stripping copyright management information (CMI) — the ownership and licensing tags attached to code — when it generated new code, agreeing that producing a new AI output isn't the same as removing tags from an existing copy. But it rejected the district court's own reasoning for getting there, calling its "identicality requirement" a misnomer and declining to endorse it: the first federal circuit-court ruling on this DMCA question for AI outputs.
The litigation calendar moved in two directions at once on September 15-16, both in cases against Anthropic. Summary judgment briefing closed in Concord Music v. Anthropic I, with a hearing set for October 21, 2026 before Judge Lee — one of only three AI copyright cases nationwide, alongside In re Mosaic LLM Litigation and In re OpenAI Copyright Infringement Litigation, to reach that stage this year, and the first of the three scheduled to be heard. Separately, Judge Pitts set a scheduling order for Cambronne v. Anthropic — the case brought by authors and publishers who opted out of the Bartz book-piracy settlement — putting non-expert summary judgment motions due May 13, 2027 and trial not until 2028, years behind the Bartz class members now waiting on their first settlement payout.
A second newspaper publisher sued OpenAI and Microsoft in as many weeks. Times Publishing Company, the corporate publisher of the Tampa Bay Times, filed suit September 16 in the Southern District of New York, alleging copyright infringement, vicarious copyright infringement and DMCA copyright-management-information removal — the same theories the Seattle Times and Newsday raised jointly on September 4. Per the complaint's own count, it is the 27th copyright suit against OpenAI and the 15th against Microsoft, out of 143 filed industry-wide, up from the 24-against-OpenAI, 137-total count in the record three weeks earlier. The case is expected to be assigned to Judge Stein and folded into the consolidated In re OpenAI Copyright Infringement Litigation MDL, then stayed pending the summary judgment briefing already underway there.
Where it differs
- Copyright — training data moving
- 143 suits pending nationwide, 27 against OpenAI and 15 against Microsoft, per the newest complaint's own count. One settlement concluded, at roughly $3,000 per work. Suits now reach executives personally: Sony Music names Dario Amodei and Benjamin Mann. on the wire
- Federal government position (fair use) thin
- The DOJ's first formal word on the question: a Statement of Interest backing fair use for AI training and calling the market-dilution theory "deeply flawed." One filing, not a ruling, and no other government position exists to corroborate or contradict it — though a defendant in an unrelated appeal has since cited it to the Third Circuit. on the wire
- DMCA anti-circumvention contested
- Split. Google's claims against SerpApi were mostly dismissed with prejudice; Reddit's near-identical claims against Perplexity and SerpApi survived dismissal days later before a different judge. on the wire
- Contributory infringement settled
- Closing. Dismissed with prejudice for the NYT, Daily News and Ziff Davis against Microsoft after the Supreme Court's Cox Communications decision foreclosed the material-contribution theory. on the wire
- CFAA — agentic browsing settled
- Favors the agent, on the only appellate answer so far. The Ninth Circuit found Perplexity's Comet unlikely to violate the CFAA because it acts on user direction rather than accessing servers on its own. on the wire
- DMCA — copyright management information (CMI) removal settled
- Favors the AI company, on the first federal appellate ruling on the question. The Ninth Circuit affirmed dismissal of a claim that GitHub Copilot violated copyright law by stripping CMI from code it reused — but rejected the district court's own "identicality requirement" reasoning as a misnomer, declining to endorse it even while agreeing with the result. on the wire
- Antitrust and competition moving
- The fastest-moving track. Judge Mehta called Google's use of publisher content in AI Overviews "seems really unfair" while hearing Penske Media's suit; nearly 300 French newspapers filed a competition complaint over AI Overviews in France. on the wire
- First Amendment defense thin
- An emerging posture rather than a tested one. Three sets of defendants have raised it since March 2026 — Musk/Tesla/Warner Bros. Discovery, then Anthropic, then Perplexity. No court has ruled on it in this context. on the wire
- Trademark dilution (Lanham Act) thin
- A single instance so far. The Seattle Times and Newsday's Sept. 4 suit against OpenAI and Microsoft argues that AI hallucinations misattributed to a publisher tarnish its marks — a theory that stands apart from the fair-use question and has not been tested in this context. on the wire
- Shareholder derivative suits (board oversight) moving
- A forming pattern, not yet a ruling. Three suits now allege Microsoft's board approved or was exposed to copyright infringement risk and misrepresented it to shareholders; the third, filed Sept. 9 against Nadella and the board, is expected to be consolidated with the other two. on the wire
- Covenant not to sue vs. fair-use counterclaim thin
- One instance, unresolved. Gilbert's covenant not to sue reaches only Anthropic's "training copies" of his book, not his still-pending infringement claim over its retention as a resource — so Anthropic argues the covenant can't moot the fair-use counterclaim it filed against him. on the wire
What changed
- 2026-09-17 Previously: 137 AI copyright suits were pending nationwide, 24 against OpenAI, per the wikiHow complaint's Aug. 22 count. Times Publishing Company (Tampa Bay Times) sued OpenAI and Microsoft, the second newspaper publisher to do so in under two weeks after the Seattle Times and Newsday's Sept. 4 suit. Per this complaint's own count, the industry total is now 143 suits, 27 against OpenAI and 15 against Microsoft. on the wire
- 2026-09-16 Previously: No federal appeals court had ruled on whether AI-generated output can violate the DMCA's copyright-management-information provision. The Ninth Circuit affirmed dismissal of that claim against GitHub Copilot — the first federal circuit ruling on the question — while rejecting the district court's own "identicality requirement" reasoning as a misnomer. on the wire
- 2026-09-16 Summary judgment briefing closed in Concord Music v. Anthropic I, with a hearing set for October 21, 2026 — one of only three AI copyright cases nationwide to reach that stage this year, and the first scheduled to be heard. on the wire
- 2026-09-16 Previously: The opt-out track for authors who declined the Bartz settlement had no published schedule. Judge Pitts set a scheduling order for Cambronne v. Anthropic: non-expert summary judgment motions due May 13, 2027, trial not until 2028 — years behind Bartz class members, whose first settlement payout is due by Nov. 15, 2026. on the wire
- 2026-08-08 Previously: Publishers were pursuing contributory-infringement theories against platform intermediaries. The Supreme Court's Cox Communications decision foreclosed the material-contribution theory; the NYT, Daily News and Ziff Davis claims against Microsoft were dismissed with prejudice. on the wire
- 2026-08-28 Previously: No AI training case had produced a price. The Bartz settlement became effective August 20, starting a 28-day payout clock at roughly $3,000 per work, due September 17 — the first concrete cost in the record for training on pirated books. on the wire
- 2026-09-01 Previously: No branch of the federal government had taken a position on whether training AI on copyrighted text is fair use. The Justice Department filed a Statement of Interest in the OpenAI copyright litigation backing fair use and calling the market-dilution theory "deeply flawed" — not binding on the court, but the first formal federal position on the question underlying nearly every pending suit. on the wire
- 2026-09-03 Previously: Bartz payouts were estimated due by September 17, 2026. Class counsel now puts the first payout — $2,203.56 per work — on or before Nov. 15, 2026, after a Sept. 4 consolidated-statement mailing and a 30-day contest window; a second payment is expected to bring the total toward roughly $3,000 per work. on the wire
- 2026-09-04 Previously: No suit against an AI company had pursued a theory outside copyright, DMCA, contributory infringement, CFAA or antitrust. The Seattle Times and Newsday added a Lanham Act trademark-dilution count against OpenAI and Microsoft, arguing AI hallucinations misattributed to a publisher tarnish its marks — a claim that survives independently of the fair-use question. on the wire
- 2026-09-04 Previously: The DOJ's fair-use position existed only within the OpenAI MDL litigation it was filed in. ROSS Intelligence cited the DOJ's Statement of Interest in a Rule 28(j) letter to the Third Circuit hearing Thomson Reuters v. ROSS Intelligence, arguing the position applies broadly to AI training — the first sign of the DOJ's view being invoked outside the case it was filed in. on the wire
- 2026-09-11 Previously: No author suing on his own, without a lawyer, had tried to keep a court from ruling on an AI company's fair-use defense by promising not to sue over the copies made for training. Gilbert, who is suing Anthropic without a lawyer over his book, promised not to sue over the copies Anthropic made to train its models, then asked the court to throw out Anthropic's fair-use claim as no longer worth deciding. Anthropic objected: its claim rests on the infringement case Gilbert is still pursuing and on his argument that the court may not consider why Anthropic copied — not on any fear of a second lawsuit. on the wire
Still unresolved
- Whether roughly $3,000 per work becomes a benchmark or stays a one-off. Bartz settled rather than ruled, so it binds nobody — and no other case has produced a number.
- Whether training on lawfully-acquired books is fair use. The Mosaic/Databricks summary-judgment hearing is set for October 30, 2026, with publisher trade groups filing against it — the nearest thing in the record to a scheduled answer. Concord Music v. Anthropic I reaches its own summary-judgment hearing nine days earlier, on October 21, but its music-publisher claims test different facts than book training.
- Whether the opt-out track for Bartz settlement decliners settles or reaches trial once it gets there. Cambronne v. Anthropic's schedule leaves nearly two years between the Bartz class's first payout and even the summary-judgment stage for those who declined the deal.
- Whether executives can be held personally liable. Sony Music names Amodei and Mann, and the authors' suit was amended to add Amodei; no court has tested it.
- Whether Judge Mehta applies the monopoly finding to AI Overviews. A hearing remark is not a ruling, and Google's motion to dismiss Penske Media is still pending.
- Why two judges reached opposite results on the same DMCA theory within days. Neither ruling has been reconciled with the other, and the split is unresolved.
- Whether any court adopts the DOJ's fair-use position, or even cites it. A party has now invoked it outside the case it was filed in — ROSS Intelligence's Sept. 4 letter to the Third Circuit — but no court has cited or ruled on it, and a Statement of Interest carries no precedential weight regardless.
- Whether trademark dilution over AI hallucinations holds up as a theory distinct from copyright. The Seattle Times/Newsday suit is the only instance in the record, and no court has ruled on it.
- Whether board-level derivative suits over copyright risk survive a motion to dismiss. Three have now been filed against Microsoft's board and none has been tested; the underlying theory is closer to executive personal liability than to the infringement suits themselves.
- Whether a covenant not to sue can be used to keep a court from ever weighing an AI company's fair-use defense. Gilbert's motion to dismiss Anthropic's counterclaim tests the maneuver; other pro se and opt-out plaintiffs are watching the outcome.
Evidence
- The pending caseload — 137 AI copyright suits nationwide, 24 against OpenAI — as counted three weeks earlier. on the wire single-source count, from one litigation tracker; superseded by the 2026-09-17 count above
- That a second newspaper publisher sued OpenAI and Microsoft in under two weeks, and the pending caseload's newest self-reported count — 143 suits nationwide, 27 against OpenAI, 15 against Microsoft. on the wire single-source count, from the complaint's own tally
- The first concluded price and that the settlement had taken effect. on the wire
- The revised payout timeline — a first installment of $2,203.56 per work by Nov. 15, 2026 — and the process (consolidated statements, contest window) behind the delay. on the wire
- The first formal federal government position on the fair-use question underlying nearly every pending AI-copyright suit. on the wire DOJ Statement of Interest — states the government's view, not binding on the court
- That near-identical DMCA theories are producing opposite results. on the wire
- That the contributory-infringement route is foreclosed, and why. on the wire
- The only appellate answer so far on whether an agentic browser violates the CFAA. on the wire
- That the judge who found Google an illegal monopoly has questioned AI Overviews' fairness to publishers. on the wire remark at a hearing, not a ruling
- That a fair-use summary-judgment hearing is scheduled for October 30, 2026. on the wire
- The first trademark-dilution theory in the record, built on AI outputs that hallucinate content and misattribute it to a publisher. on the wire
- That the DOJ's fair-use position is now being cited by a party outside the case it was filed in, in an appeal that predates the current wave of AI-copyright suits. on the wire a party's Rule 28(j) letter — not a court adopting or citing the position itself
- That copyright litigation risk is now being tested against a company's board directly, via a third shareholder derivative suit expected to consolidate with the two before it. on the wire single-source report (Chat GPT Is Eating the World); no ruling yet on the theory
- The first instance of a covenant not to sue being used to try to moot an AI company's fair-use counterclaim, and the company's argument against it. on the wire single case, pro se plaintiff; no ruling yet on the motion to dismiss
- The first federal circuit-court ruling on whether AI-generated output can violate the DMCA's CMI-removal provision, and that the court rejected the district court's own reasoning even while affirming its result. on the wire
- That Concord Music v. Anthropic I is one of only three AI copyright cases nationwide with completed summary-judgment briefing, and the first of the three scheduled to be heard. on the wire single-source litigation tracker (Chat GPT Is Eating the World)
- The opt-out class's own schedule, running years behind the Bartz settlement track it declined. on the wire
I'm getting cited. Am I getting recommended?
settled The record agrees. A change here would itself be news.
Not necessarily — being cited, being named and being recommended are three different outcomes with different drivers, and the record separates them. Across 1,094 ChatGPT categories the most-cited domain was the most-mentioned brand only 20.8% of the time. Citation measures whose page an engine used; recommendation tracks how well the engine already knows your brand.
Verified against the record 2026-08-25 · record last moved August 24, 2026 — nothing has moved it in 23 days
The three outcomes come apart at every step. An engine can lift a passage from your page and name a competitor in the sentence it supports; it can name you without linking you; it can recommend you having never cited you at all. A dashboard reporting one number cannot tell you which of the three you bought.
What moves recommendation is mostly not on-page work. Models searched for brands they already knew 3.2 times more often than unfamiliar ones — 55.7% versus 17.4% across 66 buyer prompts — and 63% of brand-specific searches surfaced one of each model's five most-familiar brands. Where competing products were otherwise identical, the well-known brand was recommended 100% of the time.
Two findings look like they disagree about reviews and do not. A market census of 4,776 venues found star rating had no effect on whether a venue appeared at all; a controlled study found a rival needed less than a 0.1-star edge to overturn an incumbent's recommendation. Those are different stages of the same funnel — rating does not get you into the candidate set, and decides between candidates once you are in it. Having a business website nearly doubled the odds of the first.
The uncomfortable finding for this discipline sits in the same study: when every brand adopted the same marketing tactics, the incumbent's advantage collapsed from a payoff of +0.802 to +0.007. Tactics that everyone runs stop being an edge and become the price of entry — while brands that opted out received no recommendations at all.
Where it differs
- Cited settled
- Your URL appears as a source under the answer. This is what citation-share tools measure, and it is the only one of the three most of them measure. on the wire
- Named settled
- Your brand appears in the answer text, with or without a link. Engines skip naming the source brand between 19% (Microsoft Copilot) and 52% (Perplexity) of the time, so citation and naming diverge by engine before anything else does. on the wire
- Recommended settled
- The engine puts you forward as the answer. Driven mostly by prior brand familiarity, then by reviews at the margin — not by the page-level work that moves citation. on the wire
Still unresolved
- Whether the 20.8% mismatch holds outside ChatGPT and Semrush's citation data. It is one analysis of one engine's estimated demand, and nobody has replicated it elsewhere.
- Whether anything a site controls moves recommendation once familiarity is accounted for. The evidence that tactics equalize to +0.007 when everyone runs them is a single controlled study on one product category.
- Whether recommendation varies by country and prompt language the way practitioners report. No published measurement covers it — the record has nothing.
Evidence
- That the most-cited domain and the most-mentioned brand are usually not the same, with the size of the gap. on the wire
- That naming diverges from citation, and by how much per engine. on the wire Writesonic study — vendor-published
- That prior familiarity drives whether a model looks for you at all. on the wire geoSurge study — vendor-published
- The incumbent advantage, the review threshold that overturns it, and the collapse when every brand runs the same tactics. on the wire
- How much of a real market never gets recommended at all, and that a website moves it where star rating does not. on the wire Norly Research study — vendor-published, funded and conducted by an AI-visibility vendor