Z-Vision Generator
Local AI image and video generation — hassle-free and fun. No tangled node graphs, no cloud dependencies, just prompts and results. Runs on macOS (Apple Silicon / MLX) and on Windows and Linux with NVIDIA CUDA through diffusers.
Features
- Image generation — text-to-image with Z-Image and FLUX.2 Klein (4B/9B) model families, plus Ideogram 4 (FP8) on macOS/MLX
- Video generation — text-to-video and image-to-video with platform-specific LTX aliases across macOS, Windows, and Linux
- Cross-platform — automatic backend selection: MLX on macOS, diffusers/CUDA on Windows and Linux for images, and the shared diffusers/CUDA LTX backend on Windows and Linux for video
- Prompt system — YAML prompt files with variables, structured prompts, snippets, and batch runs
- Model store — central
~/.ziv/directory with bare-name resolution and HuggingFace fallback - LoRA support — single or stacked, configurable weights, bare-name resolution
- Image upscale — generate small → Lanczos → img2img refine → CAS sharpen
- Video upscale — 2× spatial upscaling through the platform LTX backend when supported by the selected runtime
- Reference images — img2img steering from any starting image
- Model variants — image quantization across supported image backends, plus macOS MLX video Q4/Q8 aliases
- Post-processing — contrast, saturation, and CAS sharpening (image only)
- Interactive controls — skip, quit, pause, and repeat during batch runs (image only)
Platform Support
| Platform | Image Generation | Video Generation |
|---|---|---|
| macOS (Apple Silicon) | ✅ Z-Image / FLUX / Ideogram 4 via mflux/MLX | ✅ LTX via MLX aliases (ltx-4, ltx-8) |
| Windows (NVIDIA GPU) | ✅ Z-Image / FLUX via diffusers/CUDA | ✅ LTX via diffusers/CUDA alias (ltx-2.3) |
| Linux (NVIDIA GPU) | ✅ Z-Image / FLUX via diffusers/CUDA | ✅ LTX via diffusers/CUDA alias (ltx-2.3) |
Installation
Requires Python 3.14+ and uv.
uv is required. This package cannot be installed with pip — some dependencies require uv-specific resolution that pip does not support. All commands below use uv.
# Install globally from PyPI
uv tool install z-vision-generator
# Install globally from repository
uv tool install -e git+https://github.com/knuthelge/ZVisionGenerator.git
# Development setup
git clone https://github.com/knuthelge/ZVisionGenerator && cd ZVisionGenerator
uv sync
Video generation requires ffmpeg. On Windows and Linux, image and video generation require an NVIDIA GPU with CUDA available to PyTorch.
The packaged Windows/Linux
ltx-2.3alias defaults to the configurable diffusers-converted repositorydg845/LTX-2.3-Diffusers. This is the diffusers-compatible layout required by the Windows/Linux video backend, not an official Lightricks alias. Override it in~/.ziv/config.yamlif you want to pointltx-2.3at a different compatible diffusers repository.
The macOS video aliases
ltx-4andltx-8are the shipped MLX Q4/Q8 presets. Windows and Linux use the diffusers-backedltx-2.3alias instead, so the Q4/Q8 naming does not carry across platforms.
Quick Start
# Generate an image (bare name from ~/.ziv/models/)
ziv-image -m my-model --prompt "a beautiful sunset"
# Generate from a HuggingFace model
ziv-image -m Tongyi-MAI/Z-Image-Turbo --prompt "a cat in a garden"
# Batch run from a prompts file
ziv-image -m my-model -p prompts.yaml -r 3
# Generate a video on macOS
ziv-video -m ltx-4 --prompt "A cat walking through a garden"
# Generate a video on Windows or Linux
ziv-video -m ltx-2.3 --prompt "A cat walking through a garden"
# Image-to-video on macOS
ziv-video -m ltx-4 --image photo.jpg --prompt "Camera zooms in slowly"
# Image-to-video on Windows or Linux
ziv-video -m ltx-2.3 --image photo.jpg --prompt "Camera zooms in slowly"
# Launch the Web UI
ziv ui
# Show command help and available subcommands
ziv
The Web UI is an explicit local launcher: use ziv ui or ziv-ui. Running bare ziv prints command-discovery help without starting a browser or server. If a local environment is missing required Web UI packages, the launcher reports how to repair the base install instead of failing with an import traceback.
The packaged Web UI is served from local static files and does not fetch cloud fonts or other font assets at runtime. It uses local system font fallbacks when Inter or JetBrains Mono are not installed.
Tip:
ziv image,ziv video, andziv modelare also available as subcommands of the unifiedzivparent command. Useziv -horziv --helpto print terminal help.
Documentation
Full documentation is available at knuthelge.github.io/ZVisionGenerator.
- Getting Started — installation, model store, quick start
- Image Guide — aliases, sizes, reference images, LoRA, upscaling, quantization
- Video Guide — T2V, I2V, upscale, audio, LoRA, constraints
- Prompts Guide — prompt files, variables, structured prompts, snippets
- Model & LoRA Guide — checkpoint conversion, LoRA import, asset listing
- CLI Reference — full argument tables for all commands
- Development — setup, testing, architecture
Contributing
Contributions are welcome! See CONTRIBUTING.md for guidelines.
License
This project is licensed under the GNU Affero General Public License v3.0 or later.
Metadata
Release files for z-vision-generator 0.12.0
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| z_vision_generator-0.12.0.tar.gz | 885.4 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| z_vision_generator-0.12.0-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 1.3 MB
Release files / z_vision_generator-0.12.0.tar.gz
| Download URL | z_vision_generator-0.12.0.tar.gz |
|---|---|
| Size | 885.4 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
e6f8f691e13c8def20192daa5f0daf2d3719c7fcb562bcb413ab4b1cd4646996
|
|
BLAKE2b-256 checksum How to use checksums |
1fe61d33e2e2bd11925d460a5bda240456f40464e895917ab4e9bb92edf9b10a
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Oct 1, 2026.
Transparency logRelease files / z_vision_generator-0.12.0-py3-none-any.whl
| Download URL | z_vision_generator-0.12.0-py3-none-any.whl |
|---|---|
| Size | 419.3 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
d8abf57486ebf595af1f15135fd59bcfa6bde97818c6e570db69161313b0da1c
|
|
BLAKE2b-256 checksum How to use checksums |
2ddbfbb6ce33c705fe9d06cae46ee39de9c7831c3ce80c87e6d971191bce0fc5
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Oct 1, 2026.
Transparency log