Skip to main content

Library of web-related functions

Project description

https://secure.travis-ci.org/scrapy/w3lib.png?branch=master Coverage report

Overview

This is a Python library of web-related functions, such as:

  • remove comments, or tags from HTML snippets

  • extract base url from HTML snippets

  • translate entites on HTML strings

  • convert raw HTTP headers to dicts and vice-versa

  • construct HTTP auth header

  • converting HTML pages to unicode

  • sanitize urls (like browsers do)

  • extract arguments from urls

Requirements

Python 2.7 or Python 3.3+

Install

pip install w3lib

Documentation

See http://w3lib.readthedocs.org/

License

The w3lib library is licensed under the BSD license.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

w3lib-1.17.0.tar.gz (30.4 kB view details)

Uploaded Source

Built Distribution

w3lib-1.17.0-py2.py3-none-any.whl (19.9 kB view details)

Uploaded Python 2 Python 3

File details

Details for the file w3lib-1.17.0.tar.gz.

File metadata

  • Download URL: w3lib-1.17.0.tar.gz
  • Upload date:
  • Size: 30.4 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No

File hashes

Hashes for w3lib-1.17.0.tar.gz
Algorithm Hash digest
SHA256 d8c654827fcf92ba4d7111f8588d2eff8653c5580c27ca61b1bc7805c080506f
MD5 03f4d6160208c547e4c31a63486b9516
BLAKE2b-256 acb691ae356d48dd1d48732967eb79b2e41be4b2493b4e43a89be57b1f3be37d

See more details on using hashes here.

Provenance

File details

Details for the file w3lib-1.17.0-py2.py3-none-any.whl.

File metadata

File hashes

Hashes for w3lib-1.17.0-py2.py3-none-any.whl
Algorithm Hash digest
SHA256 562373c5f7aab03e742ddcca5a5eb37ebd324a9f5403b2eccf610c211fad1eac
MD5 c1d6488a926fbc4d56b8d3a090fd4efc
BLAKE2b-256 203eba9865b88c39edd09100a8c8df11722c8881bbf76aef0c0ae5b970eb42b7

See more details on using hashes here.

Provenance

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page