Skip to main content

Data visualization toolchain based on aggregating into a grid

Project description



Turn even the largest data into images, accurately

Build Status Linux/MacOS Build Status Windows Build status
Coverage codecov
Latest dev release Github tag
Latest release Github release PyPI version datashader version conda-forge version defaults version
Docs gh-pages site

What is it?

Datashader is a data rasterization pipeline for automating the process of creating meaningful representations of large amounts of data. Datashader breaks the creation of images of data into 3 main steps:

  1. Projection

    Each record is projected into zero or more bins of a nominal plotting grid shape, based on a specified glyph.

  2. Aggregation

    Reductions are computed for each bin, compressing the potentially large dataset into a much smaller aggregate array.

  3. Transformation

    These aggregates are then further processed, eventually creating an image.

Using this very general pipeline, many interesting data visualizations can be created in a performant and scalable way. Datashader contains tools for easily creating these pipelines in a composable manner, using only a few lines of code. Datashader can be used on its own, but it is also designed to work as a pre-processing stage in a plotting library, allowing that library to work with much larger datasets than it would otherwise.

Installation

Datashader supports Python 2.7, 3.6 and 3.7 on Linux, Windows, or Mac and can be installed with conda:

conda install datashader

or with pip:

pip install datashader

For the best performance, we recommend using conda so that you are sure to get numerical libraries optimized for your platform. The latest releases are avalailable on the pyviz channel conda install -c pyviz datashader and the latest pre-release versions are avalailable on the dev-labelled channel conda install -c pyviz/label/dev datashader.

Fetching Examples

Once you've installed datashader as above you can fetch the examples:

datashader examples
cd datashader-examples

This will create a new directory called datashader-examples with all the data needed to run the examples.

To run all the examples you will need some extra dependencies. If you installed datashader within a conda environment, with that environment active run:

conda env update --file environment.yml

Otherwise create a new environment:

conda env create --name datashader --file environment.yml
conda activate datashader

Developer Instructions

  1. Install Python 3 miniconda or anaconda, if you don't already have it on your system.

  2. Clone the datashader git repository if you do not already have it:

    git clone git://github.com/holoviz/datashader.git
    
  3. Set up a new conda environment with all of the dependencies needed to run the examples:

    cd datashader
    conda env create --name datashader --file ./examples/environment.yml
    conda activate datashader
    
  4. Put the datashader directory into the Python path in this environment:

    pip install --no-deps -e .
    

Learning more

After working through the examples, you can find additional resources linked from the datashader documentation, including API documentation and papers and talks about the approach.

Some Examples

USA census

NYC races

NYC taxi

About HoloViz

Datashader is part of the HoloViz ecosystem for making browser-based data visualization in Python easier to use, easier to learn, and more powerful. See holoviz.org for related packages that you can use with Datashader and status.holoviz.org for the current status of each HoloViz project.

Datashader is supported and maintained by Anaconda.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

datashader-0.9.0.tar.gz (30.2 MB view details)

Uploaded Source

Built Distribution

datashader-0.9.0-py2.py3-none-any.whl (15.5 MB view details)

Uploaded Python 2 Python 3

File details

Details for the file datashader-0.9.0.tar.gz.

File metadata

  • Download URL: datashader-0.9.0.tar.gz
  • Upload date:
  • Size: 30.2 MB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/2.0.0 pkginfo/1.5.0.1 requests/2.22.0 setuptools/42.0.2.post20191203 requests-toolbelt/0.9.1 tqdm/4.40.0 CPython/3.6.9

File hashes

Hashes for datashader-0.9.0.tar.gz
Algorithm Hash digest
SHA256 3a423d61014ae8d2668848edab6c12a6244be6f249570bd7811dd5698d5ff633
MD5 5b0d6fc0f9c3443f6b7bcb0ce33709c9
BLAKE2b-256 2c3735112f57d1a83d76ed2a6e084f6a1da8e24ef94f2982578d94434a1ccd2b

See more details on using hashes here.

Provenance

File details

Details for the file datashader-0.9.0-py2.py3-none-any.whl.

File metadata

  • Download URL: datashader-0.9.0-py2.py3-none-any.whl
  • Upload date:
  • Size: 15.5 MB
  • Tags: Python 2, Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/2.0.0 pkginfo/1.5.0.1 requests/2.22.0 setuptools/42.0.2.post20191203 requests-toolbelt/0.9.1 tqdm/4.40.0 CPython/3.6.9

File hashes

Hashes for datashader-0.9.0-py2.py3-none-any.whl
Algorithm Hash digest
SHA256 01e61b7d10c2ada06aa39d7486ab633c82faa56083250cda1007db1bb8558efb
MD5 a3ca1029f9ceb16377a2bf7a834f7c48
BLAKE2b-256 7f31755a2088567b2cfc0b1c41fb21062faf92d0e5c154f379c11a56a10b5559

See more details on using hashes here.

Provenance

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page