Channel sheet · CH-23 · gain 3 min · logged October 10, 2026

Content & SEO in the AI EraDirect input

Google Warns Against Markdown-Only Sites Built for AI Crawlers

Google has cautioned publishers against maintaining separate Markdown-only versions of pages for AI crawlers, warning the practice resembles cloaking and falls under existing spam policies.

By Sophie Lindqvist3 min read521 words

Signal notes

  1. Google has warned against serving separate Markdown versions of websites to AI crawlers — Search Engine Journal reported the caution but did not name the specific Google representative.
  2. The practice of generating text-only mirrors for LLM bots has grown since 2023 as 'AI SEO' tooling multiplied.
  3. Google's longstanding policy treats serving different content to different user agents as cloaking, a spam-policy violation.
  4. Cloudflare introduced pay-per-crawl controls in 2025 amid broader publisher backlash to AI scraping.
  5. Search Engine Journal did not specify whether the warning came from a public forum, private communication, or documentation update.
Google Cautions Against Markdown Versions Of Websites For AI SEO - Search Engine Journal
Input monitorGoogle Cautions Against Markdown Versions Of Websites For AI SEO - Search Engine Journal — AI-generated

Google has cautioned website operators against maintaining separate Markdown-only versions of their pages as a tactic for AI search optimization, according to a report published by Search Engine Journal.

The warning targets a growing shortcut in the publishing world: serving stripped-down text files to AI crawlers while keeping the full visual site for human visitors. Operators hoping to be picked up by systems like OpenAI's GPTBot, Anthropic's ClaudeBot, or Google's own Gemini crawler have, in various implementations, generated .md files or .txt summaries alongside their HTML.

What is the practice?

Markdown is a lightweight formatting syntax — the same one used to write README files and most AI chatbot responses. Some publishers produce machine-readable summaries of every URL and expose them to crawlers that identify themselves as bots belonging to AI companies. The intent is to make content easier for large language models to ingest and cite.

Why is Google pushing back?

The company's longstanding position is that site owners should not differentiate content based on user agent. Rendering a different page for Googlebot than for an end user is treated as cloaking, a violation of Google's spam policies. By the same logic, a Markdown-only copy aimed at AI crawlers sits in the same category.

What does the warning change?

In practice, little — the policy is not new. What is new is the volume: as AI-driven search referrals become a measurable share of traffic, more operators have experimented with text-first duplicates. Google's caution signals that the search giant considers these workarounds inconsistent with its guidelines, regardless of whether the duplicate is intended for Google's own crawler or a third-party AI bot.

Who is most exposed?

Publishers who auto-produce llms.txt files, Markdown mirrors of product pages, or schema-style summaries of articles are the primary audience for this caution. Tools that promised "AI SEO" gains through simplified text outputs have proliferated since 2023; Google's stance puts pressure on that category.

What should teams actually do?

Operators who want AI systems to surface their content should rely on the same practices that work for traditional search: clean HTML, structured data where appropriate, fast load times, and crawlable links. There is no separate optimization track for AI crawlers inside Google's playbook.

The broader context

Search Engine Journal's report lands amid a year-long debate over how publishers should respond to AI ingestion. Cloudflare rolled out pay-per-crawl controls in 2025; The New York Times and other major publishers have sued OpenAI over scraping. Google's caution is the search-side voice in that conversation, and it leans toward uniformity: one site, one set of content, served the same way to every client.

What remains unclear

Search Engine Journal's write-up does not name the specific Google representative who issued the caution, nor specify whether the warning came in a public forum, a private communication, or a documentation update. Operators waiting to align their stack should monitor Google's official Search Central documentation and public forums for the underlying statement before changing their crawl directives.

via Google News — AI content and SEO (Source)

Filed under

  • google-search
  • ai-crawlers
  • cloaking
  • llms-txt
  • seo
Share this article:

More from Sophie Lindqvist

Sophie Lindqvist

Show full bio

Staff writer covering business strategy at Mart Signal.

104 articles

Bus out

‹ Previous article