User-agent: * Allow: / # Block crawlable parameterized duplicates of list endpoints (audit v2.2.5 M-3). # Clean URLs stay fully crawlable; only query-string variants are disallowed. # R2 H-1 (2026-09-08): the page-level CSS link is /style.css?v=, which matched # this disallow - Googlebot could not fetch the only stylesheet. Explicitly re-allow # the cache-busted stylesheet before the query-string disallow. Allow: /style.css* Disallow: /*? Disallow: /tools/*? # A-2 (2026-09-25, decided by the site owner): AI retrieval and grounding are # welcome (with attribution); model training on this content is not permitted. # Full policy: https://martechsignal.com/ai-policy/ Content-Signal: search=yes, ai-input=yes, ai-train=no # ── Content Signals Policy (Cloudflare, CC0) ──────────────────────────────── # As a condition of accessing this website, you agree to abide by the following content signals: # # (a) If a content-signal = yes, you may collect content for the corresponding use. # (b) If a content-signal = no, you may not collect content for the corresponding use. # (c) If the website operator does not include a content signal for a corresponding use, the website operator neither grants nor restricts permission via content signal with respect to the corresponding use. # # The content signals and their meanings are: # # search: building a search index and providing search results (e.g., returning hyperlinks and short excerpts from your website's contents). Search does not include providing AI-generated search summaries. # ai-input: inputting content into one or more AI models (e.g., retrieval augmented generation, grounding, or other real-time taking of content for generative AI search answers). # ai-train: training or fine-tuning AI models. # # ANY RESTRICTIONS EXPRESSED VIA CONTENT SIGNALS ARE EXPRESS RESERVATIONS OF RIGHTS UNDER ARTICLE 4 OF THE EUROPEAN UNION DIRECTIVE 2019/790 ON COPYRIGHT AND RELATED RIGHTS IN THE DIGITAL SINGLE MARKET. # ── AI training crawlers: denied ──────────────────────────────────────────── # RFC 9309 note: once a token has a named group, the `*` group no longer applies # to it at all, so each group below is complete on its own. Search crawlers and # user-triggered fetch agents are intentionally unnamed and keep the `*` rules. User-agent: GPTBot Disallow: / User-agent: ClaudeBot Disallow: / User-agent: Google-Extended Disallow: / User-agent: CCBot Disallow: / User-agent: Applebot-Extended Disallow: / User-agent: meta-externalagent Disallow: / User-agent: Bytespider Disallow: / User-agent: Amazonbot Disallow: / Sitemap: https://martechsignal.com/sitemap.xml # A-3 (v2.4.0 audit): ARD entry source (spec §5.1) Agentmap: https://martechsignal.com/.well-known/ard.json