For website owners

RaiseDiscoveryBot

The crawler Raise uses to find publicly listed job openings on company career pages.

What it does

Raise helps professionals follow the roles that matter to them. To do that, RaiseDiscoveryBot reads the public job listings that companies publish on their career pages and applicant-tracking systems. It reads listing data only: title, company, location, date and link back to the original posting.

How it behaves

  • It identifies itself. Every request carries the user agent RaiseDiscoveryBot/1.0 and a link to this page.
  • It reads your robots.txt before crawling a site, and stops if a path is disallowed. If robots.txt cannot be read, it does not crawl.
  • It is slow on purpose: at most one request per second, and one connection at a time, per host.
  • It backs off. On a 429 response it honours the Retry-After header and gives up for that run. It never tries to get around a 403 or a bot-protection page.
  • It only reads public pages. It never logs in, never fills in a form, and never reads anything behind authentication.
  • It does not store the full text of job descriptions in its catalogue. Each listing links back to your page.

How to block it

Add these lines to your robots.txt. The change is picked up within 24 hours.

User-agent: RaiseDiscoveryBot
Disallow: /

Contact

If the bot causes a problem, or you want a listing removed, write to us at:

contact@raisecareer.ai