-
Notifications
You must be signed in to change notification settings - Fork 1k
Video Transcription for the SubtitleSwitcher Plugin
Vosk is an offline speech recognition toolkit that supports multiple languages. The Vosk-transcriber is a Python package that utilizes the Vosk library to transcribe audio files. By installing Vosk-transcriber in Ubuntu, you can enable the SubtitleSwitcher plugin to automatically transcribe videos and audio. This will allow you to generate accurate subtitles for your videos in various languages without the need for an internet connection. Follow this step-by-step guide to install Vosk-transcriber in Ubuntu and enhance the functionality of the subtitleSwitcher plugin with speech recognition capabilities.
Before installing new packages, it's a good practice to update your system. Open a terminal and run the following command:
sudo apt update && sudo apt upgrade
The easiest way to install vosk API is with pip. You do not have to compile anything.
- Python version: 3.5-3.9
- pip version: 20.3 and newer.
pip3 install vosk
For more information check https://alphacephei.com/vosk/install
Vosk requires a language model to perform speech recognition. Download a pre-built language model from the Vosk website (https://alphacephei.com/vosk/models) or the GitHub repository (https://github.com/alphacep/vosk-api/blob/master/doc/models.md). For example, to download the English language model, run:
wget https://alphacephei.com/vosk/models/vosk-model-small-en-us-0.15.zip
Unzip the downloaded language model:
unzip vosk-model-small-en-us-0.15.zip
The Open Source Video Platform Solution
| Service | Description | Link |
|---|---|---|
| 🎯 | Professional Support - Direct assistance from core developers | Contact |
| ☁️ | AVideo CDN - High-performance video delivery network | Pricing |
AVideo Platform © 2024 - Self-hosted video streaming platform
Made with ❤️ by WWBN and the open source community