Skip to main content

Unified GUI + engine for GGUF models: local LLM server, image/video generation and GGUF editing/quantization, all on one bundled gguf.cpp engine

Project description

gguf-cpp

One package for working with GGUF models locally: an OpenAI-compatible LLM server, a diffusion image/video/audio generator and a GGUF metadata/tensor editor with a built-in quantizer — three panels on one GUI, powered by one unified gguf.cpp engine compiled in a single build with a shared set of ggml kernels.

Install

pip install gguf-cpp

The build compiles the bundled engine (CPU by default, Metal on macOS). GPU backends are opt-in at install time:

GGUF_CPP_CUDA=1 pip install gguf-cpp     # NVIDIA
GGUF_CPP_HIP=1 pip install gguf-cpp      # AMD ROCm
GGUF_CPP_VULKAN=1 pip install gguf-cpp   # Vulkan

Run

gguf-cpp                 # unified GUI — Server / Diffuser / Editor panels
python -m gguf_cpp       # same thing

screenshot

Each panel also runs on its own, exactly like the standalone gguf-server / gguf-diffusion / gguf-editor packages did:

gguf-cpp server          # LLM server GUI
gguf-cpp diffuser        # image generation GUI
gguf-cpp editor          # GGUF editor GUI

And the engines are directly scriptable from the CLI:

gguf-cpp server engine -- --model model.gguf --port 8888
gguf-cpp diffuser engine -- -m sd.gguf -p "a lighthouse at dusk" -o out.png
gguf-cpp editor quantize -m in.gguf -o out-q4_k.gguf --type q4_k
gguf-cpp editor devices

Layout

vendor/engine/           the unified gguf.cpp engine (one CMake build)
  kernels/               shared ggml kernels (CPU + optional GPU backends)
  src/ common/ mtmd/     GGUF LLM runtime
  app/                   the gguf-server HTTP server
  diffusion/             diffusion runtime + CLI
  quantizer/             quantizer shared library (its own quant kernels)
src/gguf_cpp/            the Python package
  server/ diffuser/ editor/   the three panels (backend + web frontend each)
  gui.py static/         the unified 3-panel GUI shell

screenshot

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

gguf_cpp-0.0.7.tar.gz (32.1 MB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

gguf_cpp-0.0.7-py3-none-win_amd64.whl (62.7 MB view details)

Uploaded Python 3Windows x86-64

File details

Details for the file gguf_cpp-0.0.7.tar.gz.

File metadata

  • Download URL: gguf_cpp-0.0.7.tar.gz
  • Upload date:
  • Size: 32.1 MB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.11.9

File hashes

Hashes for gguf_cpp-0.0.7.tar.gz
Algorithm Hash digest
SHA256 4aa386be4e887bf3fa482f0645366a1388492bd271167b61a4f8c2e15a579fc7
MD5 ae85eaee2bfc3ae57fac3370a1d5b90b
BLAKE2b-256 488695cfa4ab57ba2eb074bacf8ebfb5fc47781d12558496204f76455afdaeca

See more details on using hashes here.

File details

Details for the file gguf_cpp-0.0.7-py3-none-win_amd64.whl.

File metadata

  • Download URL: gguf_cpp-0.0.7-py3-none-win_amd64.whl
  • Upload date:
  • Size: 62.7 MB
  • Tags: Python 3, Windows x86-64
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.11.9

File hashes

Hashes for gguf_cpp-0.0.7-py3-none-win_amd64.whl
Algorithm Hash digest
SHA256 4180e022dd3d3b38cb66b07d895dac01b92cae0649fbf0d624b04524f13c956b
MD5 6517e2d42073989d14e12589960aa9aa
BLAKE2b-256 2384d775dc3b36bf02a86b312b3734bed68e11446afce3910b688a83247db28c

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page