Skip to main content

Replacement robots.txt Parser in pure Python

Project description

Replaces the built-in robotsparser with a RFC-conformant implementation that supports modern robots.txt constructs like Sitemaps, Allow, and Crawl-delay. Main features:

  • Memoization of fetched robots.txt

  • Expiration taken from the Expires header

  • Batch queries

  • Configurable user agent for fetching robots.txt

  • Automatic refetching basing on expiration

This is a patched fork of the last pure Python version that works on Python 2 and 3.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

reppy2-0.3.5.tar.gz (10.2 kB view hashes)

Uploaded Source

Built Distribution

reppy2-0.3.5-py2.py3-none-any.whl (12.1 kB view hashes)

Uploaded Python 2 Python 3

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page