Silero-VAD
Lightweight, high‑accuracy voice activity detection model for real‑time speech/silence classification.
Silero VAD is a compact, production‑ready Voice Activity Detection model trained on a large multilingual corpus. It processes 32 ms audio chunks at 16 kHz and outputs a speech probability per chunk, enabling real‑time detection of speech segments in streaming or file‑based audio. The model uses an LSTM‑based architecture and is well‑suited for edge deployment on Qualcomm Snapdragon devices.
Not supported
This model is currently not supported on any IoT chipset.
To see performance metrics for this model on other chipsets, click the button below.
View for other chipsetsTechnical Details
Input resolution:512 samples (32 ms at 16 kHz)
Model size (float):2.27 MB
Number of parameters:0.24M
Supported sample rates:16000 Hz
Applicable Scenarios
- Smart Home
- Voice Assistants
- Accessibility
- Transcription Services
License
Model:MIT
Tags
- real-time
Supported IoT Devices
- Arduino VENTUNO Q
- Dragonwing IQ-9075 EVK
- Dragonwing IQ-X5121
- Dragonwing IQ-X7181
- Dragonwing Q-6690 MTP
- Dragonwing Q-7790
- Dragonwing Q-8750
- QCS8550 (Proxy)
Supported IoT Chipsets
- Qualcomm® Dragonwing™ Q-6690
- Qualcomm® QCS5121
- Qualcomm® Dragonwing™ IQ-X7181
- Qualcomm® Dragonwing™ Q-7790
- Qualcomm® Dragonwing™ IQ-8275
- Qualcomm® Dragonwing™ QCS8550 (Proxy)
- Qualcomm® Dragonwing™ Q-8750
- Qualcomm® Dragonwing™ IQ-9075
Related Models
See all modelsLooking for more? See models created by industry leaders.
Discover Model Makers









