MikeTrendsTrends right now

search

robots.txt

Trends

  1. 1
    Japanese sites block AI crawlers despite robots.txt permission●Trancoの日本のサイト300件を2026-10-07に確認。普通に読めた248件のうち23件で、robots.txtが許可しているAIクローラーのUser-Agentに403。内訳、形の違い、確かめ方、限界。 # geo # seo #MmastodonTechnologyAI22 d ago

    A check of 300 top Japanese websites from the Tranco list on 7 October 2026 found that 23 of the 248 pages that loaded normally returned HTTP 403 errors to AI crawler user-agents, even though the sites' robots.txt files permitted those bots. The writeup details which sites are affected, how the blocking patterns differ, how to verify it, and the limits of the method.

  2. 2

    A Bloomberg report argues that the web's underlying systems are unprepared for the surge of automated agents now browsing, shopping and acting online. These AI bots often ignore conventions like robots.txt and other rules sites use to govern automated access, straining infrastructure and undermining site owners' control. The piece highlights a widening gap between how the internet was built and how AI-driven traffic actually behaves.