Skip to main content

zipwire

Extract individual files from remote ZIP archives over HTTP - without downloading the whole thing.

A zip wire gets you straight to your destination. This library does the same: it uses HTTP range requests to fetch only the central directory and the specific entries you ask for, skipping everything else. A 10 KB file inside a 2 GB archive? zipwire downloads roughly 10 KB (plus a small overhead for the central directory), not 2 GB.

How it works

ZIP archives store a central directory at the end of the file that lists every entry with its offset and size. zipwire fetches that directory first (a single range request), then makes one additional range request per file you extract. The server must support Range requests (Accept-Ranges: bytes), which most CDNs, object stores, and static file servers do.

Key features

  • Selective extraction - download only the files you need, not the entire archive.
  • Streaming decompression - read_into decompresses in chunks, keeping memory usage low even for large entries.
  • Sync and async - SyncRemoteZip for synchronous code, AsyncRemoteZip with await/async with for asyncio.
  • Wheel metadata - SyncRemoteWheel / AsyncRemoteWheel read a Python wheel's .dist-info (METADATA, WHEEL, RECORD) straight from PyPI in a single adaptive tail request, without downloading the wheel.
  • ZIP64 - supports archives and entries larger than 4 GiB.
  • Pluggable backends - bring your own HTTP library (see below).
  • Local files - FileReader / AsyncFileReader open an archive on disk through the exact same API, no HTTP server required.

Installation and backends

The default installation includes the urllib3 backend. To use a different HTTP library, install the matching extra - for example httpx2 gives you both sync and async:

pip install zipwire[httpx2]
Backend Class Mode HTTP Install extra
urllib3 Urllib3Reader sync 1.1 (included)
httpx2 Httpx2SyncReader sync 1.1, 2 httpx2
httpx2 Httpx2AsyncReader async 1.1, 2 httpx2
requests RequestsReader sync 1.1 requests
aiohttp AiohttpReader async 1.1 aiohttp

Every HTTP backend accepts an optional pre-configured client or session so you can share connection pools, authentication, and retry configuration.

For archives that already live on the local filesystem, FileReader and AsyncFileReader (in zipwire.backends, no extra dependency) satisfy the same reader protocols - see the local-file example below.

Examples

Read Python wheel metadata without downloading the wheel

A common use case: fetch a wheel's METADATA, WHEEL, or RECORD from PyPI without downloading the (often huge) wheel itself. SyncRemoteWheel and AsyncRemoteWheel parse the wheel URL to locate the .dist-info directory and fetch an adaptive tail, so metadata entries are served from memory without extra HTTP requests:

from zipwire import SyncRemoteWheel
from zipwire.backends import Urllib3Reader

url = "https://files.pythonhosted.org/.../requests-2.32.3-py3-none-any.whl"
with SyncRemoteWheel(Urllib3Reader(url)) as whl:
    print(whl.read(whl.metadata_name).decode())

This optimization relies on the recommended wheel layout of placing .dist-info at the end of the archive; wheels built otherwise still work, falling back to a normal range request per entry.

Sync - list files and read one

from zipwire import SyncRemoteZip
from zipwire.backends import Urllib3Reader

reader = Urllib3Reader("https://archive.example/data.zip")
with SyncRemoteZip(reader) as rz:
    for info in rz.infolist():
        print(f"{info.filename}  {info.file_size} bytes")

    data = rz.read("path/to/file.txt")

Sync - stream a large file to disk

read_into decompresses in chunks so peak memory stays low:

from zipwire import SyncRemoteZip
from zipwire.backends import Urllib3Reader

reader = Urllib3Reader("https://archive.example/large.zip")
with SyncRemoteZip(reader) as rz:
    with open("output.bin", "wb") as f:
        rz.read_into("big-file.bin", f)

Async

import asyncio
from zipwire import AsyncRemoteZip
from zipwire.backends import AiohttpReader

async def main():
    reader = AiohttpReader("https://archive.example/data.zip")
    async with AsyncRemoteZip(reader) as rz:
        data = await rz.read("path/to/file.txt")
        print(data.decode())

asyncio.run(main())

Local archive - same API, no HTTP

FileReader opens an archive from disk through the same interface, accepting a path or a file:// URI. AsyncFileReader is the async counterpart. This lets code that already uses zipwire handle local and remote archives the same way, without special-casing either.

from zipwire import SyncRemoteZip
from zipwire.backends import FileReader

with SyncRemoteZip(FileReader("/path/to/archive.zip")) as rz:
    data = rz.read("path/to/file.txt")

License

Apache-2.0

Release files for zipwire 0.4.0

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for zipwire 0.4.0
File Size Uploaded
zipwire-0.4.0.tar.gz 166.8 kB Details

Built distribution (wheel)

Table of built distributions (wheels) for zipwire 0.4.0
File Interpreter ABI Platform
zipwire-0.4.0-py3-none-any.whl Python 3 none any Details

Total release size: 202.8 kB

Release files / zipwire-0.4.0.tar.gz

Download URL zipwire-0.4.0.tar.gz
Size 166.8 kB
Tags Source
SHA-256 checksum
How to use checksums
bb202af61e0a2c7d9f3073324258bf352711ebaf070bf540974d92ff992ec122
BLAKE2b-256 checksum
How to use checksums
6eb5964f2d343c93ab7e48e9d9e428d2c2b93e98c2156bac2b9e0cc832af0bff
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
Yes
Uploaded via twine/7.0.0 CPython/3.13.14

Provenance

Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.

PyPI Publish Attestation

PyPI verified that this artifact, at this checksum, originated from the publisher listed below.

Signed by GitHub Actions, verified by PyPI on Aug 29, 2026.

Transparency log

Release files / zipwire-0.4.0-py3-none-any.whl

Download URL zipwire-0.4.0-py3-none-any.whl
Size 36.0 kB
Tags Python 3
SHA-256 checksum
How to use checksums
09a621ef725c2afbc1bffcbf37618d07941640950ed82bb1168f0f552fc1fe29
BLAKE2b-256 checksum
How to use checksums
3a90b91c55fe8aa5c600c93b80b0c5f68f852f4b0adc9dca4f231ceb0121c2a0
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
Yes
Uploaded via twine/7.0.0 CPython/3.13.14

Provenance

Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.

PyPI Publish Attestation

PyPI verified that this artifact, at this checksum, originated from the publisher listed below.

Signed by GitHub Actions, verified by PyPI on Aug 29, 2026.

Transparency log

Release history Release notifications | RSS feed

This release

0.4.0 This release

2 release files

0.3.0

2 release files

0.2.0

2 release files

0.1.0

2 release files

0.0.1

2 release files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page