A package for converting time series data from e.g. electronic health records into wide format data.

These details have not been verified by PyPI

Project links

Project description

Timeseriesflattener

Time series from e.g. electronic health records often have a large number of variables, are sampled at irregular intervals and tend to have a large number of missing values. Before this type of data can be used for prediction modelling with machine learning methods such as logistic regression or XGBoost, the data needs to be reshaped.

In essence, the time series need to be flattened so that each prediction time is represented by a set of predictor values and an outcome value. These predictor values can be constructed by aggregating the preceding values in the time series within a certain time window.

timeseriesflattener aims to simplify this process by providing an easy-to-use and fully-specified pipeline for flattening complex time series.

🔧 Installation

To get started using timeseriesflattener simply install it using pip by running the following line in your terminal:

pip install timeseriesflattener

⚡ Quick start

import numpy as np
import pandas as pd

if __name__ == "__main__":

    # Load a dataframe with times you wish to make a prediction
    prediction_times_df = pd.DataFrame(
        {
            "id": [1, 1, 2],
            "date": ["2020-01-01", "2020-02-01", "2020-02-01"],
        },
    )
    # Load a dataframe with raw values you wish to aggregate as predictors
    predictor_df = pd.DataFrame(
        {
            "id": [1, 1, 1, 2],
            "date": [
                "2020-01-15",
                "2019-12-10",
                "2019-12-15",
                "2020-01-02",
            ],
            "value": [1, 2, 3, 4],
        },
    )
    # Load a dataframe specifying when the outcome occurs
    outcome_df = pd.DataFrame({"id": [1], "date": ["2020-03-01"], "value": [1]})

    # Specify how to aggregate the predictors and define the outcome
    from timeseriesflattener.feature_specs.single_specs import OutcomeSpec, PredictorSpec
    from timeseriesflattener.aggregation_fns import maximum, mean

    predictor_spec = PredictorSpec(
        timeseries_df=predictor_df,
        lookbehind_days=30,
        fallback=np.nan,
        aggregation_fn=mean,
        feature_base_name="test_feature",
    )
    outcome_spec = OutcomeSpec(
        timeseries_df=outcome_df,
        lookahead_days=31,
        fallback=0,
        aggregation_fn=maximum,
        feature_base_name="test_outcome",
        incident=False,
    )

    # Instantiate TimeseriesFlattener and add the specifications
    from timeseriesflattener import TimeseriesFlattener

    ts_flattener = TimeseriesFlattener(
        prediction_times_df=prediction_times_df,
        entity_id_col_name="id",
        timestamp_col_name="date",
        n_workers=1,
        drop_pred_times_with_insufficient_look_distance=False,
    )
    ts_flattener.add_spec([predictor_spec, outcome_spec])
    df = ts_flattener.get_df()
    df

Output:

	id	date	prediction_time_uuid	pred_test_feature_within_30_days_mean_fallback_nan	outc_test_outcome_within_31_days_maximum_fallback_0_dichotomous
0	1	2020-01-01 00:00:00	1-2020-01-01-00-00-00	2.5	0
1	1	2020-02-01 00:00:00	1-2020-02-01-00-00-00	1	1
2	2	2020-02-01 00:00:00	2-2020-02-01-00-00-00	4	0

📖 Documentation

Documentation
🎓 Tutorial	Simple and advanced tutorials to get you started using `timeseriesflattener`
🎛 General docs	The detailed reference for timeseriesflattener's API.
🙋 FAQ	Frequently asked question
🗺️ Roadmap	Kanban board for the roadmap for the project

💬 Where to ask questions

Type
🚨 Bug Reports	GitHub Issue Tracker
🎁 Feature Requests & Ideas	GitHub Issue Tracker
👩‍💻 Usage Questions	GitHub Discussions
🗯 General Discussion	GitHub Discussions

🎓 Projects

PSYCOP projects which use timeseriesflattener. Note that some of these projects have yet to be published and are thus private.

Project	Publications
Type 2 Diabetes		Prediction of type 2 diabetes among patients with visits to psychiatric hospital departments

Project details

These details have not been verified by PyPI

Project links

Release history Release notifications | RSS feed

2.4.0

Sep 27, 2024

2.3.0

Sep 27, 2024

2.2.6

May 23, 2024

2.2.5

May 22, 2024

2.2.4

May 17, 2024

2.2.3

May 7, 2024

2.2.2

May 3, 2024

2.2.1

May 2, 2024

2.2.0

Apr 30, 2024

2.1.2

Apr 18, 2024

2.1.1

Apr 18, 2024

2.1.0

Feb 27, 2024

2.0.2

Feb 27, 2024

2.0.1

Feb 26, 2024

2.0.0

Feb 26, 2024

1.36.2

Feb 23, 2024

1.36.1

Feb 22, 2024

1.36.0

Feb 22, 2024

1.35.0

Feb 22, 2024

1.34.0

Feb 22, 2024

1.33.0

Feb 22, 2024

1.32.0

Feb 22, 2024

1.31.3

Feb 22, 2024

1.31.2

Feb 19, 2024

1.31.1

Feb 19, 2024

1.31.0

Feb 19, 2024

1.30.0

Feb 19, 2024

1.29.0

Feb 19, 2024

1.28.0

Feb 19, 2024

1.27.0

Feb 16, 2024

1.26.0

Feb 16, 2024

1.25.1

Feb 16, 2024

1.25.0

Feb 16, 2024

1.24.0

Feb 15, 2024

1.23.0

Feb 14, 2024

1.22.0

Feb 14, 2024

1.21.1

Feb 14, 2024

1.21.0

Feb 14, 2024

1.20.1

Feb 13, 2024

1.20.0

Feb 13, 2024

1.19.0

Feb 13, 2024

1.18.1

Feb 13, 2024

1.18.0

Feb 13, 2024

1.17.0

Feb 13, 2024

1.16.0

Feb 12, 2024

1.15.0

Feb 12, 2024

1.14.0

Feb 12, 2024

1.13.0

Feb 12, 2024

1.12.0

Feb 12, 2024

1.11.0

Feb 9, 2024

This version

1.10.0

Jan 25, 2024

1.9.1

Jan 23, 2024

1.9.0

Jan 18, 2024

1.8.0

Nov 24, 2023

1.7.0

Oct 20, 2023

1.6.1

Oct 5, 2023

1.6.0

Aug 9, 2023

1.5.2

Aug 2, 2023

1.5.1

Aug 1, 2023

1.5.0

Aug 1, 2023

1.4.0

Jul 12, 2023

1.3.1

Jun 30, 2023

1.3.0

Jun 29, 2023

1.2.1

Jun 20, 2023

1.2.0

Jun 20, 2023

1.0.0

Jun 15, 2023

0.27.0

May 19, 2023

0.26.0

May 4, 2023

0.25.1

May 2, 2023

0.25.0

Apr 26, 2023

0.24.0

Apr 20, 2023

0.23.11

Mar 28, 2023

0.23.10

Mar 28, 2023

0.23.9

Mar 28, 2023

0.23.8

Mar 28, 2023

0.23.7

Mar 24, 2023

0.23.6

Mar 20, 2023

0.23.5

Mar 20, 2023

0.23.4

Mar 20, 2023

0.23.3

Mar 10, 2023

0.23.2

Mar 1, 2023

0.23.1

Feb 24, 2023

0.23.0

Feb 9, 2023

0.22.1

Dec 19, 2022

0.22.0

Dec 15, 2022

0.21.0

Dec 14, 2022

0.20.3

Dec 13, 2022

0.20.2

Dec 9, 2022

0.20.1

Dec 9, 2022

0.20.0

Dec 8, 2022

0.19.1

Dec 8, 2022

0.19.0

Dec 8, 2022

0.18.0

Dec 8, 2022

0.17.0

Dec 8, 2022

0.16.0

Dec 7, 2022

0.15.0

Dec 6, 2022

0.14.0

Dec 6, 2022

0.13.0

Dec 6, 2022

0.12.1

Dec 2, 2022

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

timeseriesflattener-1.10.0.tar.gz (5.4 MB view details)

Uploaded Jan 25, 2024 Source

Built Distribution

timeseriesflattener-1.10.0-py3-none-any.whl (4.3 MB view details)

Uploaded Jan 25, 2024 Python 3

File details

Details for the file timeseriesflattener-1.10.0.tar.gz.

File metadata

Download URL: timeseriesflattener-1.10.0.tar.gz
Upload date: Jan 25, 2024
Size: 5.4 MB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/3.8.0 pkginfo/1.9.6 readme-renderer/42.0 requests/2.31.0 requests-toolbelt/1.0.0 urllib3/2.1.0 tqdm/4.66.1 importlib-metadata/7.0.1 keyring/24.3.0 rfc3986/2.0.0 colorama/0.4.6 CPython/3.10.13

File hashes

Hashes for timeseriesflattener-1.10.0.tar.gz
Algorithm	Hash digest
SHA256	`3f94afc4b4e54a2a19fff345f5dd8b74426054aa6194a0d4e870dc8b5fbbc49f`
MD5	`a560cf2a4e20964b25532391eb48600d`
BLAKE2b-256	`358e0f7235a05362e11a0da7ed575670353371c96b3526173c054b7378d4496a`

See more details on using hashes here.

File details

Details for the file timeseriesflattener-1.10.0-py3-none-any.whl.

File metadata

Download URL: timeseriesflattener-1.10.0-py3-none-any.whl
Upload date: Jan 25, 2024
Size: 4.3 MB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/3.8.0 pkginfo/1.9.6 readme-renderer/42.0 requests/2.31.0 requests-toolbelt/1.0.0 urllib3/2.1.0 tqdm/4.66.1 importlib-metadata/7.0.1 keyring/24.3.0 rfc3986/2.0.0 colorama/0.4.6 CPython/3.10.13

File hashes

Hashes for timeseriesflattener-1.10.0-py3-none-any.whl
Algorithm	Hash digest
SHA256	`84f56302f670a2eeb0af25f0841366436fe10a9ab2eff0c460711371aeff625b`
MD5	`5fd3b29af92b94519ed42b4cb64c9852`
BLAKE2b-256	`52fefefd36db6c4a1a4b35028665684fd5213f926441957a99e86a72e209cb28`