Learn

AI crawlers

Who is scraping your content, and what to do about it.

AI crawlers

AI Crawler List 2026: User Agents, What They Do, and What to Block

A publisher's field guide to the bots hitting your site in 2026: what each one does, the exact user agent it sends, whether it respects robots.txt, and how to actually stop the ones you want gone.

Read
AI crawlers

How to Block AI Crawlers Without Hurting Your SEO

Search crawling and AI training are controlled by separate robots.txt tokens for every major engine. That means you can stop the AI bots feeding on your content while keeping your Google, Bing and AI-search visibility intact. Here is exactly how, plus where robots.txt stops working.

Read
AI crawlers

How to Block GPTBot in robots.txt (and the Real Trade-off)

Blocking GPTBot takes two lines. Understanding what it does and does not do is the part that saves your traffic. Here is the exact directive, the trade-off that catches most publishers out, and how to confirm it worked.

Read
AI crawlers

How to Tell if AI Is Scraping Your Site (and Track It Over Time)

AI crawlers never show up in GA4. To see who is taking your content, you have to read your server logs, verify the bots by IP, and record the counts over time. Here is the working method.

Read
AI crawlers

llms.txt for Publishers: What It Is and Whether It Helps

llms.txt is a voluntary standard that curates your best pages for AI models. For developer docs it earns its keep. For a typical ad-funded publisher, the crawlers that build AI answers do not request it, so it changes almost nothing about your AI visibility or revenue.

Read