Skip to main content

A packaged version of the Silero VAD model

Project description

Mailing list : test Mailing list : test License: CC BY-NC 4.0

Open In Colab

header


Silero VAD


Silero VAD - pre-trained enterprise-grade Voice Activity Detector (also see our STT models).


Real Time Example

https://user-images.githubusercontent.com/36505480/144874384-95f80f6d-a4f1-42cc-9be7-004c891dd481.mp4


Key Features


  • Stellar accuracy

    Silero VAD has excellent results on speech detection tasks.

  • Fast

    One audio chunk (30+ ms) takes less than 1ms to be processed on a single CPU thread. Using batching or GPU can also improve performance considerably. Under certain conditions ONNX may even run up to 4-5x faster.

  • Lightweight

    JIT model is around one megabyte in size.

  • General

    Silero VAD was trained on huge corpora that include over 100 languages and it performs well on audios from different domains with various background noise and quality levels.

  • Flexible sampling rate

    Silero VAD supports 8000 Hz and 16000 Hz sampling rates.

  • Flexible chunk size

    Model was trained on 30 ms. Longer chunks are supported directly, others may work as well.

  • Highly Portable

    Silero VAD reaps benefits from the rich ecosystems built around PyTorch and ONNX running everywhere where these runtimes are available.

  • No Strings Attached

    Published under permissive license (MIT) Silero VAD has zero strings attached - no telemetry, no keys, no registration, no built-in expiration, no keys or vendor lock.


Typical Use Cases


  • Voice activity detection for IOT / edge / mobile use cases
  • Data cleaning and preparation, voice detection in general
  • Telephony and call-center automation, voice bots
  • Voice interfaces

Links



Get In Touch


Try our models, create an issue, start a discussion, join our telegram chat, email us, read our news.

Please see our wiki and tiers for relevant information and email us directly.

Citations

@misc{Silero VAD,
  author = {Silero Team},
  title = {Silero VAD: pre-trained enterprise-grade Voice Activity Detector (VAD), Number Detector and Language Classifier},
  year = {2021},
  publisher = {GitHub},
  journal = {GitHub repository},
  howpublished = {\url{https://github.com/snakers4/silero-vad}},
  commit = {insert_some_commit_here},
  email = {hello@silero.ai}
}

Examples and VAD-based Community Apps


  • Example of VAD ONNX Runtime model usage in C++

  • Voice activity detection for the browser using ONNX Runtime Web

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

silero-vad-fork-0.1.0.tar.gz (11.8 kB view details)

Uploaded Source

Built Distribution

silero_vad_fork-0.1.0-py3-none-any.whl (10.1 kB view details)

Uploaded Python 3

File details

Details for the file silero-vad-fork-0.1.0.tar.gz.

File metadata

  • Download URL: silero-vad-fork-0.1.0.tar.gz
  • Upload date:
  • Size: 11.8 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/4.0.2 CPython/3.11.5

File hashes

Hashes for silero-vad-fork-0.1.0.tar.gz
Algorithm Hash digest
SHA256 ac022c3d0b60a7c6959761049771cede294847b062ba4e4e597ef58c4bc16786
MD5 2fba0008bd18d9c6b9bf5102b8d398a4
BLAKE2b-256 cf58e8e511666274d0643a21c413038f28ccc515ec1d1d91b8eb8e183953f3b0

See more details on using hashes here.

File details

Details for the file silero_vad_fork-0.1.0-py3-none-any.whl.

File metadata

File hashes

Hashes for silero_vad_fork-0.1.0-py3-none-any.whl
Algorithm Hash digest
SHA256 c5326ff601f508a7716736065ca7bff80f2bb4e22b8a329e686e7500fdd1e95d
MD5 4e333af45fd333bdf7db0255b00c86ce
BLAKE2b-256 f1504f9df28d96b7ee2a72a3998772f5908a9d3ea98230d81693776510ea4709

See more details on using hashes here.

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page