Skip to main content

ONNX Runtime-based CPU/GPU inference plugin for VapourSynth with CUDA support

Project description

VapourSynth-MLRT-ORT

This package contains the ONNX Runtime backend implementation of the vs-mlrt plugin.

Installation

To install the standard CPU/DirectML/CoreML package:

pip install vapoursynth-mlrt-ort

To install the CUDA-enabled package:

pip install vapoursynth-mlrt-ort-cuda --extra-index-url https://jaded-encoding-thaumaturgy.github.io/vs-wheels/simple

Building from source

Requirements

  • C++ Compiler: C++20 compatible (e.g. MSVC 2019+, GCC, Clang)
  • Dependencies:
    • onnxruntime (ONNX Runtime SDK)
    • ONNX
    • Protobuf
  • Optional Backend Dependencies:
    • DirectML (Windows): Requires DirectML SDK. Define the DML_DIR environment/CMake variable to point to the SDK directory.
    • CUDA: Requires CUDAToolkit and cuDNN SDKs. Ensure CUDA_PATH, CUDNN_PATH / CUDNN_HOME are set correctly.

Compilation

By default, the package builds the CPU backend (with CoreML on macOS and optionally DirectML on Windows):

uv build --package vapoursynth-mlrt-ort

To build the CUDA-enabled version, the package definition must first be updated using the helper script:

# Update pyproject.toml package configuration to target CUDA
uv run --script scripts/cuda_pyproject.py pyproject.toml

# Compile the CUDA package
uv build --package vapoursynth-mlrt-ort-cuda

Detailed parameter information from the parent project follows.


VapourSynth ONNX Runtime

The vs-onnxruntime plugin provides optimized CPU & CUDA runtime for some popular AI filters.

Building and Installation

To build, you will need ONNX Runtime, protobuf, ONNX and their dependencies.

Please refer to ONNX Runtime Docs for installation notes. Or, you can use our prebuilt Windows binary releases from AmusementClub.

Please refer to our github actions workflow for sample building instructions.

If you only use the CPU backend, then you just need to extract binary release into your vapoursynth/plugins directory.

However, if you also use the CUDA backend, you will need to download some CUDA libraries as well, please see the release page for details. Those CUDA libraries also need to be extracted into VS vapoursynth/plugins directory. The plugin will try to load them from vapoursynth/plugins/vsort/ directory or vapoursynth/plugins/vsmlrt-cuda/ directory.

Usage

Prototype: core.ort.Model(clip[] clips, string network_path[, int[] overlap = None, int[] tilesize = None, string provider = "", int device_id = 0, int verbosity = 2, bint cudnn_benchmark = True, bint builtin = False, string builtindir="models", bint fp16 = False, bint path_is_serialization = False, bint use_cuda_graph = False])

Arguments:

  • clip[] clips: the input clips, only 32-bit floating point RGB or GRAY clips are supported. For model specific input requirements, please consult our wiki.
  • string network_path: the path to the network in ONNX format.
  • int[] overlap: some networks (e.g. CNN) support arbitrary input shape where other networks might only support fixed input shape and the input clip must be processed in tiles. The overlap argument specifies the overlapping (horizontal and vertical, or both, in pixels) between adjacent tiles to minimize boundary issues. Please refer to network specific docs on the recommended overlapping size.
  • int[] tilesize: Even for CNN where arbitrary input sizes could be supported, sometimes the network does not work well for the entire range of input dimensions, and you have to limit the size of each tile. This parameter specify the tile size (horizontal and vertical, or both, including the overlapping). Please refer to network specific docs on the recommended tile size.
  • string provider: Specifies the device to run the inference on.
    • "CPU" or "": pure CPU backend
    • "CUDA": CUDA GPU backend, requires Nvidia Maxwell+ GPUs.
    • "DML": DirectML backend
    • "COREML": CoreML backend
  • int device_id: select the GPU device for the CUDA backend.'
  • int verbosity: specify the verbosity of logging, the default is warning.
    • 0: fatal error only, ORT_LOGGING_LEVEL_FATAL
    • 1: also errors, ORT_LOGGING_LEVEL_ERROR
    • 2: also warnings, ORT_LOGGING_LEVEL_WARNING
    • 3: also info, ORT_LOGGING_LEVEL_INFO
    • 4: everything, ORT_LOGGING_LEVEL_VERBOSE
  • bint cudnn_benchmark: whether to let cuDNN use benchmarking to search for the best convolution kernel to use. Default True. It might incur some startup latency.
  • bint builtin: whether to load the model from the VS plugins directory, see also builtindir.
  • string builtindir: the model directory under VS plugins directory for builtin models, default "models".
  • bint fp16: whether to quantize model to fp16 for faster and memory efficient computation.
  • bint path_is_serialization: whether the network_path argument specifies an onnx serialization of type bytes.
  • bint use_cuda_graph: whether to use CUDA Graphs to improve performance and reduce CPU overhead in CUDA backend. Not all models are supported.
  • int ml_program: select CoreML provider.
    • 0: NeuralNetwork
    • 1: MLProgram

When overlap and tilesize are not specified, the filter will internally try to resize the network to fit the input clips. This might not always work (for example, the network might require the width to be divisible by 8), and the filter will error out in this case.

The general rule is to either:

  1. left out overlap, tilesize at all and just process the input frame in one tile, or
  2. set all three so that the frame is processed in tilesize[0] x tilesize[1] tiles, and adjacent tiles will have an overlap of overlap[0] x overlap[1] pixels on each direction. The overlapped region will be throw out so that only internal output pixels are used.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

vapoursynth_mlrt_ort_cuda-15.16.2.tar.gz (676.2 kB view details)

Uploaded Source

Built Distributions

If you're not sure about the file name format, learn more about wheel file names.

vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-win_amd64.whl (101.0 MB view details)

Uploaded Python 3Windows x86-64

vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-manylinux_2_27_x86_64.manylinux_2_28_x86_64.whl (94.2 MB view details)

Uploaded Python 3manylinux: glibc 2.27+ x86-64manylinux: glibc 2.28+ x86-64

vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-manylinux_2_27_aarch64.manylinux_2_28_aarch64.whl (92.2 MB view details)

Uploaded Python 3manylinux: glibc 2.27+ ARM64manylinux: glibc 2.28+ ARM64

File details

Details for the file vapoursynth_mlrt_ort_cuda-15.16.2.tar.gz.

File metadata

File hashes

Hashes for vapoursynth_mlrt_ort_cuda-15.16.2.tar.gz
Algorithm Hash digest
SHA256 48acdccbcaee2e2ea8e0c717beb10ee8bdaa5039256c1fcf73733947b35b92b2
MD5 afb42ebb213bfcefd56d6cc1c19312f8
BLAKE2b-256 7972b142ace07355f831c5584a379dc71cf0b3318d35d1c95ea4c71d8b3b15d7

See more details on using hashes here.

Provenance

The following attestation bundles were made for vapoursynth_mlrt_ort_cuda-15.16.2.tar.gz:

Publisher: cd-publish.yml on Jaded-Encoding-Thaumaturgy/vs-wheels

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-win_amd64.whl.

File metadata

File hashes

Hashes for vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-win_amd64.whl
Algorithm Hash digest
SHA256 bc1d844f51dee51756ec05205548943651d6891033fb7fd0d14aaed6f4223b72
MD5 9735fa2e743bed7fb44b78eaf6f46450
BLAKE2b-256 9b640e130ace99a15d1f3223cde0e99129be5f9b9e1b9a20c27b4b397d24bf4d

See more details on using hashes here.

Provenance

The following attestation bundles were made for vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-win_amd64.whl:

Publisher: cd-publish.yml on Jaded-Encoding-Thaumaturgy/vs-wheels

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-manylinux_2_27_x86_64.manylinux_2_28_x86_64.whl.

File metadata

File hashes

Hashes for vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-manylinux_2_27_x86_64.manylinux_2_28_x86_64.whl
Algorithm Hash digest
SHA256 eb8b678a005bef36f0367a7f036543031de370ffd5436a2d6d1217f3ee1bd494
MD5 7dd98f592f0f76d90fcdeb3ec36de9f5
BLAKE2b-256 27a6a6c4cc1af5f2266ddb25b8fb133af658cc4feb7ab9ad75001f9f2d3af8d2

See more details on using hashes here.

Provenance

The following attestation bundles were made for vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-manylinux_2_27_x86_64.manylinux_2_28_x86_64.whl:

Publisher: cd-publish.yml on Jaded-Encoding-Thaumaturgy/vs-wheels

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-manylinux_2_27_aarch64.manylinux_2_28_aarch64.whl.

File metadata

File hashes

Hashes for vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-manylinux_2_27_aarch64.manylinux_2_28_aarch64.whl
Algorithm Hash digest
SHA256 70e72a855c04541ac60a1faf2a297881b64f1feb3c838d42cc67c2586bc00fa7
MD5 e18e6ecaa3347f4fd2c44e7dd975933c
BLAKE2b-256 169a10759634b35d8c464c397e03a582230138e6701c7aaf1e5df33047d399a1

See more details on using hashes here.

Provenance

The following attestation bundles were made for vapoursynth_mlrt_ort_cuda-15.16.2-py3-none-manylinux_2_27_aarch64.manylinux_2_28_aarch64.whl:

Publisher: cd-publish.yml on Jaded-Encoding-Thaumaturgy/vs-wheels

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page