# ReceiptSplit robots.txt # # /privacy/ and /terms/ (and pl variants) deliberately use noindex meta + # X-Robots-Tag headers, NOT robots.txt Disallow. They contain personal data and # must not appear in any search index, but Google must be allowed to crawl them # in order to read the noindex tag โ€” Disallow would block the crawl and leave # bare URL entries in the index forever. # # /support/ (and pl variants) is FAQ content with no personal data and stays # indexable โ€” its Q&A is high-value AEO surface area for support queries. # # AI crawlers (GPTBot, ClaudeBot, PerplexityBot, Google-Extended) are deliberately # allowed. Citation visibility in AI Overviews / ChatGPT Search / Perplexity # outweighs training-data concern for an open marketing site. Do not toggle this # without intent โ€” see docs/seo/SEO_PLAN_2026.md ยง1.1 for the rationale. User-agent: * Allow: / Sitemap: https://receiptsplit.work/sitemap.xml