Google's NotebookLM Rebrand Opens New Scraping Risk for Sites
Google renamed NotebookLM to Gemini Notebook and updated its fetcher identity, giving site owners a narrow window before old blocking rules stop working and scraping without attribution continues.

Key takeaways
- Google renamed NotebookLM to Gemini Notebook and updated its official list of user-triggered fetchers to match.
- Old user agent strings hardcoded into robots.txt or firewalls stop working in August 2026, per Search Engine Journal.
- Gemini Notebook's Discover Sources feature can scrape up to ten web sources per query and generate an AI summary with zero referral traffic.
- Audio and video overview features can turn scraped content into podcast-style or video output that competes with the original page.
- Site owners who want to block the crawler must update firewalls and .htaccess files, not just robots.txt, before the grace period ends.
What Changed
Google has quietly updated its list of user-triggered fetchers to reflect a rebrand: NotebookLM is now Gemini Notebook. Site owners who hardcoded the old user agent into robots.txt get a grace period of a few weeks before that string stops being recognized in August 2026, according to reporting from Search Engine Journal.
Google says the underlying product has not changed. Gemini Notebook still functions as a research assistant that lets users upload documents as ground truth, and it still works across YouTube video and uploaded audio. The name changed. The fetching behavior did not.
Why It Matters for Marketers
That matters beyond text summaries. The same tool's audio and video overview features can repurpose scraped articles into a podcast episode or video explainer, output that can then compete directly against the source material it drew from, without attribution.
This is the same attribution gap CMO Mag flagged when covering how Google's ad stack is turning into an agent marketers brief directly, and it echoes concerns raised around the Ask Advisor rollout across Google's ad products: AI systems are increasingly consuming brand content as raw material with no traffic returned.
What To Do This Week
- Check whether robots.txt or firewall rules reference the old NotebookLM user agent string and update them before August 2026.
- If blocking Gemini Notebook entirely, know that robots.txt alone will not suffice: firewalls and .htaccess files need updating too, per the SEJ report.
- Audit which pages get pulled via the Discover Sources feature, since it can select articles automatically without a user pasting a URL.
Track how AI tools use your content by visiting the AI in Marketing hub.
Advertiser disclosure: some links in our articles are affiliate links, and CMO Mag may earn a commission or referral fee if you sign up or buy through them, at no cost to you. It never affects our editorial coverage. See our advertising & affiliate policy.
More in AI in Marketing
View allAI Knows 96% of Brands, Cites Almost None of Them
A Q2 2026 study of eight AI platforms finds a wide gap between brand recognition and brand recommendation, and it's a warning for anyone counting on AI search for demand generation.
Search Console May Be Adding AI Performance Data, SEOs Say
A Reddit thread in r/SEO has practitioners buzzing about new generative AI performance reporting inside Google Search Console, but Google hasn't confirmed the rollout publicly. Here's what marketers should actually do this week.
Google's Gemini API Managed Agents Gain Hooks, Model Choice, Free Tier
Google DeepMind updated Managed Agents in the Gemini API with environment hooks for auditing tool calls, a default switch to Gemini 3.6 Flash, and free tier access, moving the sandboxed agent framework closer to production-ready marketing automation.




Discussion
No comments yet. Be the first to say something worth reading.