User-agent: * # Block non-content paths from crawling (saves crawl budget for content pages) # Note: /css/ is NOT blocked — Google needs CSS to render and judge Core Web Vitals Disallow: /api/ Disallow: /watchlist Disallow: /watch/ # Block ALL pagination (page 2+) — saves crawl budget, page 1 is the canonical Disallow: /*?*page=[2-9] Disallow: /*?*page=[0-9][0-9] Disallow: /*?*page=[0-9][0-9][0-9] # Block server-specific watch URLs (duplicate content) Disallow: /*?*server= # Crawl-delay removed — Bing obeys this literally at 1 req/sec, crippling index speed # 2026-10-02: Meta-WebIndexer (Meta AI search) by name, same rules, no "Allow: /" line — it reads the # first matching line and had been crawling /watch/ anyway. Welcome on content pages. User-agent: Meta-WebIndexer Disallow: /api/ Disallow: /watchlist Disallow: /watch/ Disallow: /*?*page=[2-9] Disallow: /*?*page=[0-9][0-9] Disallow: /*?*page=[0-9][0-9][0-9] Disallow: /*?*server= # 2026-10-02: training-only crawlers — nothing to crawl here (owner's decision). Search engines and # citation bots (OAI-SearchBot, ChatGPT-User, Claude-SearchBot, PerplexityBot, Meta-WebIndexer…) are not affected. User-agent: GPTBot User-agent: ClaudeBot User-agent: meta-externalagent User-agent: Bytespider User-agent: CCBot User-agent: Google-Extended User-agent: Applebot-Extended Disallow: / # Sitemap location Sitemap: https://indexflix.org/sitemap.xml