User-agent: * Allow: / Allow: /llms.txt Allow: /.well-known/ard.json Allow: /.well-known/business-capabilities.json Allow: /.well-known/agent-interaction.json # Reading this site to answer a customer is welcome — that is the business. # Scraping it to train a model is refused. These are different permissions, and # a blanket `Disallow` for agent crawlers collapses them: it costs the business # every agent-mediated customer in order to express a preference about training. # Cloudflare-managed robots.txt did exactly that on a sibling property, which an # outside model then reported as having no reachable machine interface. Content-Signal: search=yes, ai-input=yes, ai-train=no, use=reference Sitemap: https://rivercadefarms.com/sitemap.xml