PacevineBot
The automated crawler that builds our race database — polite, transparent, and easy to block.
What is PacevineBot?
PacevineBot is the automated crawler Pacevine uses to build a global database of trail, ultra, and sky races — event names, distances, elevation gain, dates, and registration windows — sourced from official race websites and public race calendars.
It identifies itself with the user agent PacevineBot/1.0 , which links back to this page.
How it behaves
- Honors robots.txt on every site it visits, including Cloudflare's Content-Signal opt-out directives (e.g. ai-train=no) — a site that opts out of AI use is never fed into our extraction pipeline.
- Rate-limited to a conservative request rate per host (1 request per second by default, or the site's own Crawl-delay if longer) — never a burst of traffic.
- Never bypasses a login, paywall, or download gate (for example, a Scribd-hosted document) to reach content that isn't otherwise public.
- Never scrapes or redistributes data from ITRA, UTMB, or similar federation databases — only from a race's own official site, or a written data-licence.
What we do with the data
We extract structured facts (distances, elevation gain, dates, registration windows) to build a searchable race catalog, always linking back to the official race site as the source of record. We do not claim ownership of any race's content — we index and summarize it, and send traffic back to organizers.
What we never redistribute
We never re-host or redistribute a race's GPX file, reglamento (rules) document, or images — course files are analyzed to compute a pacing profile and then discarded; images are shown as a hotlink to the original URL, never copied to our own servers.
Questions or a takedown request?
If you're a race organizer and want your site excluded from crawling, or have any question about how PacevineBot works, please get in touch — we respond to every request.
Reach us through our contact form — we read every message.