A packaged version of the Silero VAD model

Project description

header

Silero VAD

Silero VAD - pre-trained enterprise-grade Voice Activity Detector (also see our STT models).

Real Time Example

https://user-images.githubusercontent.com/36505480/144874384-95f80f6d-a4f1-42cc-9be7-004c891dd481.mp4

Key Features

Stellar accuracy

Silero VAD has excellent results on speech detection tasks.
Fast

One audio chunk (30+ ms) takes less than 1ms to be processed on a single CPU thread. Using batching or GPU can also improve performance considerably. Under certain conditions ONNX may even run up to 4-5x faster.
Lightweight

JIT model is around one megabyte in size.
General

Silero VAD was trained on huge corpora that include over 100 languages and it performs well on audios from different domains with various background noise and quality levels.
Flexible sampling rate

Silero VAD supports 8000 Hz and 16000 Hz sampling rates.
Flexible chunk size

Model was trained on 30 ms. Longer chunks are supported directly, others may work as well.
Highly Portable

Silero VAD reaps benefits from the rich ecosystems built around PyTorch and ONNX running everywhere where these runtimes are available.
No Strings Attached

Published under permissive license (MIT) Silero VAD has zero strings attached - no telemetry, no keys, no registration, no built-in expiration, no keys or vendor lock.

Typical Use Cases

Voice activity detection for IOT / edge / mobile use cases
Data cleaning and preparation, voice detection in general
Telephony and call-center automation, voice bots
Voice interfaces

Get In Touch

Try our models, create an issue, start a discussion, join our telegram chat, email us, read our news.

Please see our wiki and tiers for relevant information and email us directly.

Citations

@misc{Silero VAD,
  author = {Silero Team},
  title = {Silero VAD: pre-trained enterprise-grade Voice Activity Detector (VAD), Number Detector and Language Classifier},
  year = {2021},
  publisher = {GitHub},
  journal = {GitHub repository},
  howpublished = {\url{https://github.com/snakers4/silero-vad}},
  commit = {insert_some_commit_here},
  email = {hello@silero.ai}
}

Examples and VAD-based Community Apps

Example of VAD ONNX Runtime model usage in C++
Voice activity detection for the browser using ONNX Runtime Web

Project details

Release history Release notifications | RSS feed

This version

0.1.0

Sep 19, 2023

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

silero-vad-fork-0.1.0.tar.gz (11.8 kB view hashes)

Uploaded Sep 19, 2023 Source

Built Distribution

silero_vad_fork-0.1.0-py3-none-any.whl (10.1 kB view hashes)

Uploaded Sep 19, 2023 Python 3

Hashes for silero-vad-fork-0.1.0.tar.gz

Hashes for silero-vad-fork-0.1.0.tar.gz
Algorithm	Hash digest
SHA256	`ac022c3d0b60a7c6959761049771cede294847b062ba4e4e597ef58c4bc16786`
MD5	`2fba0008bd18d9c6b9bf5102b8d398a4`
BLAKE2b-256	`cf58e8e511666274d0643a21c413038f28ccc515ec1d1d91b8eb8e183953f3b0`

Hashes for silero_vad_fork-0.1.0-py3-none-any.whl

Hashes for silero_vad_fork-0.1.0-py3-none-any.whl
Algorithm	Hash digest
SHA256	`c5326ff601f508a7716736065ca7bff80f2bb4e22b8a329e686e7500fdd1e95d`
MD5	`4e333af45fd333bdf7db0255b00c86ce`
BLAKE2b-256	`f1504f9df28d96b7ee2a72a3998772f5908a9d3ea98230d81693776510ea4709`