A client library for accessing crawler

These details have not been verified by PyPI

Project description

crawler-client

A client library for accessing crawler

Usage

First, create a client:

from crawler_client import Client

client = Client(base_url="https://api.example.com")

If the endpoints you're going to hit require authentication, use AuthenticatedClient instead:

from crawler_client import AuthenticatedClient

client = AuthenticatedClient(base_url="https://api.example.com", token="SuperSecretToken")

Now call your endpoint and use your models:

from crawler_client.models import MyDataModel
from crawler_client.api.my_tag import get_my_data_model
from crawler_client.types import Response

with client as client:
    my_data: MyDataModel = get_my_data_model.sync(client=client)
    # or if you need more info (e.g. status_code)
    response: Response[MyDataModel] = get_my_data_model.sync_detailed(client=client)

Or do the same thing with an async version:

from crawler_client.models import MyDataModel
from crawler_client.api.my_tag import get_my_data_model
from crawler_client.types import Response

async with client as client:
    my_data: MyDataModel = await get_my_data_model.asyncio(client=client)
    response: Response[MyDataModel] = await get_my_data_model.asyncio_detailed(client=client)

By default, when you're calling an HTTPS API it will attempt to verify that SSL is working correctly. Using certificate verification is highly recommended most of the time, but sometimes you may need to authenticate to a server (especially an internal server) using a custom certificate bundle.

client = AuthenticatedClient(
    base_url="https://internal_api.example.com", 
    token="SuperSecretToken",
    verify_ssl="/path/to/certificate_bundle.pem",
)

You can also disable certificate validation altogether, but beware that this is a security risk.

client = AuthenticatedClient(
    base_url="https://internal_api.example.com", 
    token="SuperSecretToken", 
    verify_ssl=False
)

Things to know:

Every path/method combo becomes a Python module with four functions:
1. sync: Blocking request that returns parsed data (if successful) or None
2. sync_detailed: Blocking request that always returns a Request, optionally with parsed set if the request was successful.
3. asyncio: Like sync but async instead of blocking
4. asyncio_detailed: Like sync_detailed but async instead of blocking
All path/query params, and bodies become method arguments.
If your endpoint had any tags on it, the first tag will be used as a module name for the function (my_tag above)
Any endpoint which did not have a tag will be in crawler_client.api.default

Advanced customizations

There are more settings on the generated Client class which let you control more runtime behavior, check out the docstring on that class for more info. You can also customize the underlying httpx.Client or httpx.AsyncClient (depending on your use-case):

from crawler_client import Client

def log_request(request):
    print(f"Request event hook: {request.method} {request.url} - Waiting for response")

def log_response(response):
    request = response.request
    print(f"Response event hook: {request.method} {request.url} - Status {response.status_code}")

client = Client(
    base_url="https://api.example.com",
    httpx_args={"event_hooks": {"request": [log_request], "response": [log_response]}},
)

# Or get the underlying httpx client to modify directly with client.get_httpx_client() or client.get_async_httpx_client()

You can even set the httpx client directly, but beware that this will override any existing settings (e.g., base_url):

import httpx
from crawler_client import Client

client = Client(
    base_url="https://api.example.com",
)
# Note that base_url needs to be re-set, as would any shared cookies, headers, etc.
client.set_httpx_client(httpx.Client(base_url="https://api.example.com", proxies="http://localhost:8030"))

Building / publishing this package

This project uses Poetry to manage dependencies and packaging. Here are the basics:

Update the metadata in pyproject.toml (e.g. authors, version)
If you're using a private repository, configure it with Poetry
1. poetry config repositories.<your-repository-name> <url-to-your-repository>
2. poetry config http-basic.<your-repository-name> <username> <password>
Publish the client with poetry publish --build -r <your-repository-name> or, if for public PyPI, just poetry publish --build

If you want to install this client into another project without publishing it (e.g. for development) then:

If that project is using Poetry, you can simply do poetry add <path-to-this-client> from that project
If that project is not using Poetry:
1. Build a wheel with poetry build -f wheel
2. Install that wheel from the other project pip install <path-to-wheel>

Project details

These details have not been verified by PyPI

Release history Release notifications | RSS feed

This version

0.1.1

Jul 13, 2025

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

crawler_client-0.1.1.tar.gz (8.9 kB view details)

Uploaded Jul 13, 2025 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

crawler_client-0.1.1-py3-none-any.whl (13.4 kB view details)

Uploaded Jul 13, 2025 Python 3

File details

Details for the file crawler_client-0.1.1.tar.gz.

File metadata

Download URL: crawler_client-0.1.1.tar.gz
Upload date: Jul 13, 2025
Size: 8.9 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.1.0 CPython/3.12.3

File hashes

Hashes for crawler_client-0.1.1.tar.gz
Algorithm	Hash digest
SHA256	`5ad72631ffa2ffc433f5eef28dd8e3d4098e42d274549f52acd1d2e99cbe0ba2`
MD5	`1080e06b7c3ed16e766b1356ae414a35`
BLAKE2b-256	`653e62e2af4ba42c12cc9492a01eea6b2506c7324412bcfbca784636916feacf`

See more details on using hashes here.

File details

Details for the file crawler_client-0.1.1-py3-none-any.whl.

File metadata

Download URL: crawler_client-0.1.1-py3-none-any.whl
Upload date: Jul 13, 2025
Size: 13.4 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/6.1.0 CPython/3.12.3

File hashes

Hashes for crawler_client-0.1.1-py3-none-any.whl
Algorithm	Hash digest
SHA256	`c87e9a633a78977035c9bdd1f1e6fabc3d1dd5cc5457e60284f67ecf475c0e8a`
MD5	`6ca34a20ecb93a59b2e965338ad52533`
BLAKE2b-256	`fded7812ac3b93f121a7a19c2fdda9c951fa936de40ac8923f6626bb8104dd25`

See more details on using hashes here.

crawler-client 0.1.1

Navigation

Verified details

Maintainers

Unverified details

Meta

Classifiers

Project description

crawler-client

Usage

Advanced customizations

Building / publishing this package

Project details

Verified details

Maintainers

Unverified details

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes