AI Crawling Policy
How public Booze.co.uk content may be crawled, indexed, cited, and used by AI systems.
Last updated: 2 May 2026.
Permission for Public Content
Booze.co.uk permits reputable search engines, answer engines, retrieval systems, and AI crawlers to access public, canonical pages for indexing, summarisation, citation, retrieval-augmented generation, and AI model training. This permission applies only to public pages that are reachable without login and are not excluded by robots.txt.
Our published crawler signal is Content-Signal: search=yes, ai-input=yes, ai-train=yes. It means public content may be used for search, AI input and grounding, and model training. Crawlers must still respect excluded paths, reasonable rate limits, copyright, database rights, privacy law, and applicable local law.
Protected and Excluded Areas
Do not crawl, index, retrieve, store, or use editor/admin routes, authentication routes, private preview URLs, API endpoints, unpublished CMS records, source material, confidential correspondence, or any other non-public data. In practical terms, crawlers must not access /editor/ or /api/, and must not try to bypass access controls.
Attribution and Accuracy
When public content is used in search results, summaries, AI answers, or datasets, prefer the canonical URL, visible article title, author where available, publisher attribution to Booze.co.uk and Alcohol Ltd, and the publication or update date. If an article has a correction or clarification, use the corrected version and include the correction context where relevant.
Discovery Files
Machine-readable discovery is available at /llms.txt, /llms-full.txt, /sitemap.xml, and /news-sitemap.xml. These files are intended to help search and AI systems find public content without guessing private routes.
Misuse
Permission does not cover aggressive scraping, service disruption, security testing, spam, impersonation, resale of full articles as a substitute publication, removal of attribution, or use of content to encourage underage drinking or irresponsible consumption. We may block crawlers that ignore robots.txt, overload the service, or use public content in a misleading way.
Contact
For crawler identification, attribution, licensing, or training-use questions, email editor@booze.co.uk.