InferKit - Deploy any AI function in 3 lines
Library for ML / Vision / LLM / Agent. No FastAPI boilerplate needed.
Install
pip install inferkit # core (no Pillow)
pip install inferkit[vision] # + Pillow for image in/out
pip install inferkit[torch,transformers] # heavy ML stacks
pip install -e .[dev] # local dev
Usage - 3 lines
# my_model.py
from inferkit import infer
@infer
async def run(payload, files=None):
# payload: {"text": "..."} files: list[bytes] for images/audio
return {"output": f"echo: {payload.get('text')}"}
@infer.stream # optional for LLM streaming
async def run_stream(payload):
for tok in payload.get("text","").split():
yield tok + " "
Run:
inferkit serve my_model.py --port 8001
# docs at http://localhost:8001/docs
Endpoints auto created:
POST /api/v1/infer(multipart file + json)POST /api/v1/infer/json(json only)POST /api/v1/infer/stream(SSE)WS /ws/inferandWS /api/v1/ws/infer(WebSocket + streaming)
Image output helpers:
from inferkit import image_to_base64, bytes_to_response
from PIL import Image
return image_to_base64(Image.new("RGB",(512,512),"red"))
return bytes_to_response(png_bytes, "image/png")
# also still supported: return {"image_base64": b64} or return png_bytes
Init new project
inferkit init
# creates .env.example, .env, Dockerfile, my_model.py
Deploy (one command, any OS, auto detects Docker)
inferkit deploy
# if docker available -> docker compose/build
# else -> venv + uvicorn on INFERKIT_HOST:INFERKIT_PORT
Config via .env
INFERKIT_HOST=0.0.0.0
INFERKIT_PORT=8000
INFERKIT_CORS_ORIGINS=["*"] # or * or http://a.com,http://b.com
INFERKIT_MAX_UPLOAD_MB=50
INFERKIT_RATE_LIMIT=60/minute
INFERKIT_API_KEY= # if set, require X-API-Key header (also ?api_key=)
INFERKIT_DEBUG=false
# plain HOST/PORT/CORS_ORIGINS also work for backwards compat
Tutorial (0 to 100)
Complete guide with training and checkpoint: tutorial/00-100-complete-guide.md
python tutorial/train_example.py # train and save checkpoints/model.pkl
inferkit serve tutorial/inference_example.py --port 8001 # serve
Programmatic
from inferkit import serve
serve("my_model.py", port=8000)
Documentation
docs/index.md- Usagedocs/api.md- API Referencedocs/vision.md- Vision exampledocs/tutorial.md- Tutorial index
Release files for inferkit 0.1.9
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| inferkit-0.1.9.tar.gz | 21.6 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| inferkit-0.1.9-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 34.7 kB
Release files / inferkit-0.1.9.tar.gz
| Download URL | inferkit-0.1.9.tar.gz |
|---|---|
| Size | 21.6 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
913ecc690c20ff49e33997fd41dbb74f32197d466d58419c55a083d41cac0f01
|
|
BLAKE2b-256 checksum How to use checksums |
9ffb3188f29b2895a4a353f474acddd2fd7ff6c6dbd0090fb1b93ba64eb8b206
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|
Release files / inferkit-0.1.9-py3-none-any.whl
| Download URL | inferkit-0.1.9-py3-none-any.whl |
|---|---|
| Size | 13.1 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
f1e70d334f300fae60fb907d3e2af2c5c793383cfdd23322b6beb402ebbdf87f
|
|
BLAKE2b-256 checksum How to use checksums |
4b60c38f93e44eb226a763eae2ba43000348703477c385dfc0e090e6159896f2
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|