Model-agnostic generative vision abstractions (image/video) for the Abstract ecosystem

These details have been verified by PyPI

Project links

GitHub Statistics

Maintainers

lpalbou

These details have not been verified by PyPI

Project description

AbstractVision

Model-agnostic generative vision API (images, optional video) for Python and the Abstract* ecosystem.

What you get

A small orchestration API: VisionManager
A packaged capability registry (“what models can do”): VisionModelCapabilitiesRegistry backed by vision_model_capabilities.json
Optional artifact-ref outputs (small JSON refs): LocalAssetStore and RuntimeArtifactStoreAdapter
Built-in backends (execution engines): src/abstractvision/backends/
- OpenAI-compatible HTTP: openai_compatible.py
- Local Diffusers: huggingface_diffusers.py
- Local stable-diffusion.cpp / GGUF: stable_diffusion_cpp.py
CLI/REPL for manual testing: abstractvision
Optional static Playground UI (server-backed): playground/vision_playground.html (docs: playground/README.md)

How it fits together (diagram)

flowchart LR
  Caller[Python / CLI / AbstractCore] --> VM[VisionManager]
  VM --> BE[VisionBackend]
  BE --> VM
  VM -->|optional| Store[MediaStore]
  Store --> Ref[Artifact ref dict]
  VM -->|no store| Asset["GeneratedAsset (bytes + mime)"]

Status (current backend support)

Development status: Alpha (0.x). The public API is stable-by-design, but breaking changes may still happen and will be called out in CHANGELOG.md.
Built-in backends implement: text_to_image and image_to_image.
Video (text_to_video, image_to_video) is supported only via the OpenAI-compatible backend when endpoints are configured.
multi_view_image is part of the public API (VisionManager.generate_angles) but no built-in backend implements it yet.

Details: docs/reference/backends.md.

Installation

pip install abstractvision

Note (CUDA): on Windows/Linux, pip install abstractvision may install a CPU-only PyTorch build. If you want to use an NVIDIA GPU, install a CUDA-enabled PyTorch build first (see https://pytorch.org/get-started/locally/) and verify torch.cuda.is_available() is True.

Install optional integrations:

pip install "abstractvision[abstractcore]"

If you hit “missing pipeline class” errors for newer model families, see docs/getting-started.md. In that case you may need Diffusers from source (main):

pip install -U "abstractvision[huggingface-dev]"
pip install -U "git+https://github.com/huggingface/diffusers@main"

For local dev (from a repo checkout):

pip install -e .

Usage

Start here:

Getting started: docs/getting-started.md
FAQ: docs/faq.md
API reference: docs/api.md
Architecture: docs/architecture.md
Docs index: docs/README.md

Recommended default model (local / cross-platform)

The REPL defaults to a cache-only Diffusers setup using runwayml/stable-diffusion-v1-5 on auto device. Pre-download the model outside the REPL, then start generating:

huggingface-cli download runwayml/stable-diffusion-v1-5
export ABSTRACTVISION_BACKEND=diffusers
export ABSTRACTVISION_MODEL_ID=runwayml/stable-diffusion-v1-5
export ABSTRACTVISION_DIFFUSERS_DEVICE=auto
abstractvision repl

For a fresh cache, you can also permit the REPL to download missing files:

ABSTRACTVISION_DIFFUSERS_ALLOW_DOWNLOAD=1 abstractvision repl

More recommendations by VRAM: docs/getting-started.md.

Capability-driven model selection

from abstractvision import VisionModelCapabilitiesRegistry

reg = VisionModelCapabilitiesRegistry()
assert reg.supports("runwayml/stable-diffusion-v1-5", "text_to_image")

print(reg.list_tasks())
print(reg.models_for_task("text_to_image"))

Backend wiring + generation (artifact outputs)

The default install is “batteries included” (Torch + Diffusers + stable-diffusion.cpp python bindings), but heavy modules are imported lazily (see src/abstractvision/backends/__init__.py).

from abstractvision import LocalAssetStore, VisionManager, VisionModelCapabilitiesRegistry, is_artifact_ref
from abstractvision.backends import OpenAICompatibleBackendConfig, OpenAICompatibleVisionBackend

reg = VisionModelCapabilitiesRegistry()

backend = OpenAICompatibleVisionBackend(
    config=OpenAICompatibleBackendConfig(
        base_url="http://localhost:1234/v1",
        api_key="YOUR_KEY",      # optional for local servers
        model_id="REMOTE_MODEL", # optional (server-dependent)
    )
)

vm = VisionManager(
    backend=backend,
    store=LocalAssetStore(),         # enables artifact-ref outputs
    model_id="zai-org/GLM-Image",    # optional: capability gating
    registry=reg,                   # optional: reuse loaded registry
)

out = vm.generate_image("a cinematic photo of a red fox in snow")
assert is_artifact_ref(out)
print(out)  # {"$artifact": "...", "content_type": "...", ...}

png_bytes = vm.store.load_bytes(out["$artifact"])  # type: ignore[union-attr]

When installed next to AbstractCore, AbstractVision is also discovered as a llm.vision capability plugin. The plugin defaults to the same local Diffusers Stable Diffusion 1.5 setup as the REPL; set ABSTRACTVISION_BACKEND=openai and ABSTRACTVISION_BASE_URL when you want the plugin to call an OpenAI-compatible image endpoint instead.

Interactive testing (CLI / REPL)

abstractvision models
abstractvision tasks
abstractvision show-model runwayml/stable-diffusion-v1-5

abstractvision repl

Inside the REPL:

/t2i "a watercolor painting of a lighthouse" --width 512 --height 512 --steps 10 --open

For a newer but still relatively small local model, try black-forest-labs/FLUX.2-klein-4B after installing Diffusers from source (see docs/getting-started.md):

/backend diffusers black-forest-labs/FLUX.2-klein-4B mps float16
/t2i "a product photo of a matte black espresso machine" --steps 4 --guidance-scale 1.0 --open

OpenAI-compatible server example:

/backend openai http://localhost:1234/v1
/t2i "a watercolor painting of a lighthouse" --width 512 --height 512 --steps 10 --open

The CLI/REPL can also be configured via ABSTRACTVISION_* env vars; see docs/reference/configuration.md.

One-shot commands (OpenAI-compatible HTTP backend only):

abstractvision t2i --base-url http://localhost:1234/v1 "a studio photo of an espresso machine"
abstractvision i2i --base-url http://localhost:1234/v1 --image ./input.png "make it watercolor"

Local GGUF via stable-diffusion.cpp

If you want to run GGUF diffusion models locally, use the stable-diffusion.cpp backend (sdcpp). Start with a single-file Stable Diffusion model when possible; Qwen Image and FLUX GGUF component sets are heavier.

Recommended:

macOS (Apple Silicon / Metal): install sd-cli (stable-diffusion.cpp executable) from releases and use CLI mode for Metal acceleration.
Otherwise (pip-only convenience): pip install abstractvision already includes the stable-diffusion.cpp python bindings (stable-diffusion-cpp-python), but this may run CPU-only depending on the wheel build.

Alternative (external executable):

Install sd-cli: https://github.com/leejet/stable-diffusion.cpp/releases

In the REPL:

/backend sdcpp /path/to/sd-v1-5.gguf /path/to/sd-cli
/t2i "a watercolor painting of a lighthouse" --width 512 --height 512 --steps 10 --open

FLUX.2-klein-4B GGUF component example:

/backend sdcpp /path/to/flux-2-klein-4b-Q8_0.gguf /path/to/flux2_ae.safetensors /path/to/Qwen3-4B-Q4_K_M.gguf /path/to/sd-cli
/t2i "a product photo of a matte black espresso machine" --steps 4 --guidance-scale 1.0 --sampling-method euler --diffusion-fa --offload-to-cpu --open

Extra flags are forwarded via request.extra. In CLI mode they are forwarded to sd-cli; in python bindings mode, keys are mapped to python binding kwargs when supported and unsupported keys are ignored.

AbstractCore tool integration (artifact refs)

If you’re using AbstractCore tool calling, AbstractVision can expose vision tasks as tools:

from abstractvision.integrations.abstractcore import make_vision_tools

tools = make_vision_tools(vision_manager=vm, model_id="zai-org/GLM-Image")

AbstractFramework ecosystem

AbstractVision is part of the AbstractFramework ecosystem and is designed to compose with:

AbstractFramework (project hub): https://github.com/lpalbou/AbstractFramework
AbstractCore (orchestration + tool calling): https://github.com/lpalbou/abstractcore
AbstractRuntime (runtime services, including artifact storage): https://github.com/lpalbou/abstractruntime

In practice:

AbstractVision standardizes generative vision outputs (image/video) behind VisionManager.
AbstractCore can discover and use AbstractVision via the capability plugin (src/abstractvision/integrations/abstractcore_plugin.py) or you can expose vision tasks as tools (src/abstractvision/integrations/abstractcore.py).
Artifact refs returned by AbstractVision are designed to travel across processes; RuntimeArtifactStoreAdapter bridges to an AbstractRuntime-style artifact store (src/abstractvision/artifacts.py).

Project

Release notes: CHANGELOG.md
Contributing: CONTRIBUTING.md
Security: SECURITY.md
Acknowledgments: ACKNOWLEDGMENTS.md
Agent docs: llms.txt and llms-full.txt

Requirements

Python >= 3.9

License

MIT License - see LICENSE file for details.

Author

Laurent-Philippe Albou

Contact

contact@abstractcore.ai

Project details

These details have been verified by PyPI

Project links

GitHub Statistics

Maintainers

lpalbou

These details have not been verified by PyPI

Release history Release notifications | RSS feed

0.3.18

May 31, 2026

0.3.17

May 29, 2026

0.3.16

May 26, 2026

0.3.15

May 26, 2026

0.3.14

May 26, 2026

0.3.13

May 23, 2026

0.3.12

May 22, 2026

0.3.11

May 22, 2026

0.3.10

May 22, 2026

0.3.9

May 21, 2026

0.3.8

May 20, 2026

0.3.7

May 19, 2026

0.3.6

May 17, 2026

0.3.5

May 13, 2026

0.3.4

May 9, 2026

0.3.3

May 8, 2026

0.3.2

May 8, 2026

0.3.1

May 7, 2026

0.3.0

May 7, 2026

0.2.6

May 6, 2026

0.2.5

May 6, 2026

0.2.4

May 6, 2026

This version

0.2.3

May 6, 2026

0.2.2

May 6, 2026

0.2.1

Feb 5, 2026

0.1.0

Jan 9, 2026

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

abstractvision-0.2.3.tar.gz (145.7 kB view details)

Uploaded May 6, 2026 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

abstractvision-0.2.3-py3-none-any.whl (56.1 kB view details)

Uploaded May 6, 2026 Python 3

File details

Details for the file abstractvision-0.2.3.tar.gz.

File metadata

Download URL: abstractvision-0.2.3.tar.gz
Upload date: May 6, 2026
Size: 145.7 kB
Tags: Source
Uploaded using Trusted Publishing? Yes
Uploaded via: twine/6.1.0 CPython/3.13.12

File hashes

Hashes for abstractvision-0.2.3.tar.gz
Algorithm	Hash digest
SHA256	`7b224d8e2b9f5fc2bc76b74a0768047fe0d7b8e66f0aa1423a0e04c21ef8d0f5`
MD5	`3adf723201dbb0aa6ccc117c8ad1785f`
BLAKE2b-256	`db9fc28f073f175d4efed8659c9ac6bce422e65e2a81ec56aab4ee243ac114ce`

See more details on using hashes here.

Provenance

The following attestation bundles were made for abstractvision-0.2.3.tar.gz:

Publisher: release.yml on lpalbou/AbstractVision

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Statement:
- Statement type: https://in-toto.io/Statement/v1
- Predicate type: https://docs.pypi.org/attestations/publish/v1
- Subject name: abstractvision-0.2.3.tar.gz
- Subject digest: 7b224d8e2b9f5fc2bc76b74a0768047fe0d7b8e66f0aa1423a0e04c21ef8d0f5
- Sigstore transparency entry: 1451398873
- Sigstore integration time: May 6, 2026
Source repository:
- Permalink: lpalbou/AbstractVision@de8cb5167bbc5d6459283b061ae12e0a81043176
- Branch / Tag: refs/tags/v0.2.3
- Owner: https://github.com/lpalbou
- Access: public
Publication detail:
- Token Issuer: https://token.actions.githubusercontent.com
- Runner Environment: github-hosted
- Publication workflow: release.yml@de8cb5167bbc5d6459283b061ae12e0a81043176
- Trigger Event: push

File details

Details for the file abstractvision-0.2.3-py3-none-any.whl.

File metadata

Download URL: abstractvision-0.2.3-py3-none-any.whl
Upload date: May 6, 2026
Size: 56.1 kB
Tags: Python 3
Uploaded using Trusted Publishing? Yes
Uploaded via: twine/6.1.0 CPython/3.13.12

File hashes

Hashes for abstractvision-0.2.3-py3-none-any.whl
Algorithm	Hash digest
SHA256	`fe3b746333660fe75dfd7cea15052abaf5c25532c845890af9c1198969f22e10`
MD5	`cf5449febd8d8ab600486171595dfd84`
BLAKE2b-256	`935ad6495ce675e3a3b2a09a92364ab95ffb982c4f3bbde71ccf27cac2a62fe2`

See more details on using hashes here.

Provenance

The following attestation bundles were made for abstractvision-0.2.3-py3-none-any.whl:

Publisher: release.yml on lpalbou/AbstractVision

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Statement:
- Statement type: https://in-toto.io/Statement/v1
- Predicate type: https://docs.pypi.org/attestations/publish/v1
- Subject name: abstractvision-0.2.3-py3-none-any.whl
- Subject digest: fe3b746333660fe75dfd7cea15052abaf5c25532c845890af9c1198969f22e10
- Sigstore transparency entry: 1451399110
- Sigstore integration time: May 6, 2026
Source repository:
- Permalink: lpalbou/AbstractVision@de8cb5167bbc5d6459283b061ae12e0a81043176
- Branch / Tag: refs/tags/v0.2.3
- Owner: https://github.com/lpalbou
- Access: public
Publication detail:
- Token Issuer: https://token.actions.githubusercontent.com
- Runner Environment: github-hosted
- Publication workflow: release.yml@de8cb5167bbc5d6459283b061ae12e0a81043176
- Trigger Event: push

abstractvision 0.2.3

Navigation

Verified details

Project links

GitHub Statistics

Maintainers

Unverified details

Meta

Classifiers

Project description

AbstractVision

What you get

How it fits together (diagram)

Status (current backend support)

Installation

Usage

Recommended default model (local / cross-platform)

Capability-driven model selection

Backend wiring + generation (artifact outputs)

Interactive testing (CLI / REPL)

Local GGUF via stable-diffusion.cpp

AbstractCore tool integration (artifact refs)

AbstractFramework ecosystem

Project

Requirements

License

Author

Contact

Project details

Verified details

Project links

GitHub Statistics

Maintainers

Unverified details

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

Provenance

File details

File metadata

File hashes

Provenance