A package to provide pathlib like access to zip & tar archives.
Project description
archive-path
A package to provide pathlib like access to zip & tar archives.
Installation
$ pip install archive-path
Usage
For reading zip (ZipPath
) or tar (TarPath
) files:
from archive_path import TarPath, ZipPath
path = TarPath("path/to/file.tar.gz", mode="r:gz")
sub_path = path / "folder" / "file.txt"
assert sub_path.filepath == "path/to/file.tar.gz"
assert sub_path.at == "folder/file.txt"
assert sub_path.exists() and sub_path.is_file()
assert sub_path.parent.is_dir()
content = sub_path.read_text()
for sub_path in path.iterdir():
print(sub_path)
For writing files, you should use within a context manager, or directly call the close
method:
with TarPath("path/to/file.tar.gz", mode="w:gz") as path:
(path / "new_file.txt").write_text("hallo world")
# there are also some features equivalent to shutil
(path / "other_file.txt").putfile("path/to/external_file.txt")
(path / "other_folder").puttree("path/to/external_folder", pattern="**/*")
Note that archive formats do not allow to overwrite existing files (they will raise a FileExistsError
).
For performant access to single files:
from archive_path import read_file_in_tar, read_file_in_zip
content = read_file_in_tar("path/to/file.tar.gz", "file.txt", encoding="utf8")
These methods allow for faster access to files (using less RAM) in archives containing 1000's of files.
This is because, the archive's file index is only read until the path is found (discarding non-matches),
rather than the standard tarfile
/zipfile
approach that is to read the entire index into memory first.
Windows compatibility
Paths within the archives are always read and written as being /
delimited.
This means that the package works on Windows,
but will not be compatible with archives written outside this package with \\
path delimiters.
Development
See CONTRIBUTING.md for details on how to contribute to this package.
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
Hashes for archive_path-0.3.0-py3-none-any.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | ac195760fb5807702986a4bf926110c40c3c27247290fe15a5ee749306890c84 |
|
MD5 | a7f4e3fd68047eb95fb32ece4f985b5a |
|
BLAKE2b-256 | ae7d289742434414423c2d1b77232f7b2b8be5824cf727481baf25b5ef087c17 |