Skip to content

other

GPTBot

GPTBot is OpenAI's web crawler that fetches pages to gather training data for its AI models; site owners can block it via robots.txt.

Known aliases

  • GPTBot/1.2
  • OpenAI GPTBot

Relationships

No evidence-backed relationships are recorded.

Current stories

build1 publisher

One fetch as Googlebot exposed the empty shell behind 270 unindexed articles

Googlebot got 0 characters inside a client-rendered blog's root div, its developer found after Google crawled 270 of the posts and indexed none. A bot-only branch in an existing Netlify edge function now hands crawlers the rendered article and JSON-LD.

Publishers:dev.to

Reality

Evidence55
Adoption
Insufficient
Hype gap+10
Incentives
Insufficient
Confidence50
build1 publisher

A page that ranks can hand GPTBot an empty HTML shell

A dev.to post sets out two ways assistant retrieval fails while ordinary search keeps working, one in robots.txt rules written per user-agent and one in copy that appears only after JavaScript runs. It offers no measurement.

Publishers:dev.to

Reality

Evidence30
Adoption
Insufficient
Hype gap+15
Incentives
Insufficient
Confidence40
build1 publisher

48 startups, 4 known by name, 28 recommended by category

A one-afternoon test splits AI visibility in two: name recall lags funding and press by years, while category retrieval is already working for products the model cannot describe.

Publishers:dev.to

Reality

Evidence34
Adoption
Insufficient
Hype gap+28
Incentives55
Confidence38

Earlier coverage

  1. ChatGPT-User outfetched Googlebot for 34 days on one small site. Read your logs.

    Build · August 16, 2026 · 1 publisher

  2. Three files, three contracts: robots.txt, sitemap.xml and llms.txt are not rivals

    Build · August 15, 2026 · 1 publisher