A new Pew Research Center study, built on nearly half a million English-language pages pulled from the Common Crawl archive and scored with Open Pangram's AI-detection technology, finds that 35% of web pages published after ChatGPT's November 2022 release show significant signs of AI authorship or heavy AI editing — versus roughly 10% across a random sample that includes older, pre-ChatGPT pages. The study lands the same week Cloudflare reported that bot web traffic has already overtaken human web traffic, sooner than the company had forecast.
Domain type matters a lot: .com pages show AI-authorship signals at roughly 10x the rate of .edu or .gov pages (both around 1%), while .org domains sit at 4.6%. Pew is careful to flag that AI-detection tools like Pangram can misclassify genuinely human-written pages, so the numbers are directional rather than exact — but at scale, researchers say the trend is unmistakable: much of the newer web is now written by AI, and increasingly read by AI crawlers rather than people.
The finding arrives one day after TechCrunch's own data-driven look at hardening AI-trust backlash (Aug 19's top story), reinforcing a consistent theme this week: as AI content and AI-mediated discovery both scale up simultaneously, questions about authenticity, trust, and who—or what—is actually consuming the content are converging into one storyline.
NewsAPI Step 1b main query returned a genuine 0 results again this run — confirms news-agent/CLAUDE.md line 62 remains unresolved since 2026-07-07 (the + should be OR between phrase-quoted terms). The competitor query (1c) returned 9 raw hits, 0 genuine (a celebrity-death drug probe, a Grateful Dead idiom explainer, Indian political commentary, a Nigerian energy-council story, a PETA campaign piece, an architecture listing, a Robert Downey Jr. quote piece, and a healthcare-accountability op-ed — no GEO-competitor signal at all today).
VentureBeat AI's RSS feed did not repeat its mislabeled-stale-content bug again today — today's fetch resurfaced only the already-logged Aug 19 Rob Strechay item, correctly dated, with no new post and no mislabeled old content standing in for breaking news.
shopify.dev's changelog feed.xml still returns HTTP 500 — direct page-render fallback confirms no new item since Aug 19's already-logged app-intent change; worth adopting the fallback by default until the feed is fixed.
Anthropic Newsroom RSS (rss.xml) still 404s — direct fetch of the newsroom page confirms no post beyond Aug 14's watermark FAQ, so no signal was lost.
Firecrawl: 3/3 scrapes attempted; all three returned a Cloudflare Turnstile challenge banner prepended to the markdown — the same recurring failure mode as recent runs; the full article text followed cleanly beneath the challenge boilerplate in all three cases (the Pew AI-authorship piece, the Google Preferred Sources piece, and the Ramp OpenAI-vs-Anthropic piece), so no content was lost.