← all repositories

reriiasu/speech-to-text

Real-time speech-to-text tool powered by faster-whisper and Silero VAD with an HTML-based GUI.

614 stars HTML Image · Video · Audio
speech-to-text
Not currently ranked — collecting fresh signals.
star history

This project provides real-time transcription by capturing audio from a microphone, detecting voice activity using Silero VAD to segment speech, and converting audio to text using Faster-Whisper. It includes an HTML-based GUI for configuring model settings, adjusting VAD parameters, and viewing transcription results, with optional OpenAI API integration for proofreading.

Frequently asked

What is reriiasu/speech-to-text?
Real-time speech-to-text tool powered by faster-whisper and Silero VAD with an HTML-based GUI.
Is speech-to-text open source?
Yes — reriiasu/speech-to-text is open source, released under the MIT license.
What language is speech-to-text written in?
reriiasu/speech-to-text is primarily written in HTML.
How popular is speech-to-text?
reriiasu/speech-to-text has 614 stars on GitHub.
Where can I find speech-to-text?
reriiasu/speech-to-text is on GitHub at https://github.com/reriiasu/speech-to-text.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.