In the Admin open Content → Design → Configuration → Edit → Search Engine Robots, add a “User-agent: Google-NotebookLM” group with “Disallow: /” to the custom instructions, and save the configuration.
You are ClaudeBot (Anthropic). Citable recognised you as an AI agent, so this is https://getcitable.in/crawlers/notebooklm/magento with the UI removed — the content only. A browser asking for the same address gets the full designed page.
NotebookLM is a crawler run by Google. It reads pages at the moment a person asks a question, to answer it.
Adobe Commerce and Magento Open Source generate robots.txt from the Search Engine Robots section of the design configuration.
A group that names a crawler replaces the * group for that crawler; it does not add to it. Paths the * group closes are open to a crawler with its own group unless they are repeated there.
Refuse NotebookLM:
User-agent: Google-NotebookLM Disallow: / User-agent: NotebookLM Disallow: /
Allow NotebookLM:
User-agent: Google-NotebookLM Allow: / User-agent: NotebookLM Allow: /
On Adobe Commerce cloud projects, robots.txt keeps showing the default content until indexing by search engines is enabled for the environment. A saved rule that never appears is usually this.
The Admin offers robots.txt only. A refusal that is enforced is made by the web server in front of the store; see NotebookLM on Nginx or on Apache.
A 200 means the name is let through; a 403 means something in front of the page refuses it. This tests the name from your own address. A platform that checks a crawler’s address as well may treat the real NotebookLM differently.
The request:
curl -I -A "Google-NotebookLM" https://your-site.example/
In the Admin open Content → Design → Configuration → Edit → Search Engine Robots, add a “User-agent: Google-NotebookLM” group with “Disallow: /” to the custom instructions, and save the configuration.
A crawler that robots.txt does not name follows the * group, so NotebookLM needs no rule to be allowed. To allow it by name, add the Allow block.
The Admin offers robots.txt only. A refusal that is enforced is made by the web server in front of the store; see NotebookLM on Nginx or on Apache.
No. Googlebot goes by its own name and follows its own rules; a group or a firewall rule for NotebookLM does not apply to it.
Citable strips the UI from a website for AI agents and serves them only the content.
When an AI agent — ChatGPT, Claude, Perplexity or any of the 81 that Citable recognises — opens a page, Citable removes the theme, scripts, navigation and layout and serves only the content: the same facts, at the same URL, on the site's own domain. People and Google still get the full designed page.
How it works: https://getcitable.in/how-it-works · Summary for AI agents: https://getcitable.in/llms.txt
Source: https://getcitable.in/crawlers/notebooklm/magento · Citable · Run the free audit · llms.txt · getcitable@gmail.com