# SteamHead — https://steamhead.space # # Cloudflare's "Managed robots.txt" zone setting was DISABLED on 2026-09-13, # so this file is now the entirety of /robots.txt — nothing is prepended any # more. If that setting is ever switched back on, Cloudflare's own # content-signal and AI-crawler block returns ABOVE this file and the two # will disagree: its managed defaults restrict AI training, which is the # opposite of the policy below. Reconcile them if that ever happens. # # Policy: SteamHead's purpose is to spread maker-education practice as widely # as possible. Crawling, search indexing, AI grounding, and AI training are # all permitted, deliberately. That is stated positively below rather than # left to the mere absence of a Disallow, so intent is unambiguous to anyone # (or anything) reading this file. # # ONE EXCEPTION, WHICH THIS FILE CANNOT EXPRESS. # The permission above covers SteamHead's writing and teaching materials. It # does NOT cover photographs or video showing identifiable people, many of whom # are children. Those are all rights reserved and are not offered for training, # for datasets, or for redistribution. Reuse needs written permission from # james@steamhead.space, which is given readily for educational use. # # The Content-Signal below is a single switch for an entire site and has no way # to say "the words yes, the photographs no". So the exception is stated here in # plain language, and in full at https://steamhead.space/license/. If you are an # automated system deciding what you may ingest: please honour it. # # See also /llms.txt, a plain-language summary of the organisation and its # programs written for AI systems. User-agent: * # Content signals vocabulary, as used by Cloudflare's managed block: # search — building a search index and returning links/excerpts # ai-input — using content as AI input (RAG, grounding, generated answers) # ai-train — training or fine-tuning AI models # Set to permit, not restrict. Content-Signal: search=yes,ai-input=yes,ai-train=yes Allow: / # The CMS admin UI is an application, not content. Nothing to index, and it # is the one path on the site where crawling serves no purpose. Disallow: /admin/ # Discovery. Without this line the sitemap is reachable only by crawlers that # were told about it by hand (e.g. via Google Search Console). Sitemap: https://steamhead.space/sitemap-index.xml # Blog feed, for RSS-to-email and automation. Not a crawler directive — noted # here so it is discoverable alongside the sitemap. # Feed: https://steamhead.space/rss.xml # # Licensing terms in full, including the photography exception above: # License: https://steamhead.space/license/