A fully local and private Speech-To-Text app, offering multiple model backends, diarization & calendar mode - Available for Windows, macOS & Linux
-
Updated
Aug 21, 2026 - TypeScript
A fully local and private Speech-To-Text app, offering multiple model backends, diarization & calendar mode - Available for Windows, macOS & Linux
Revisiting End-to-End Speech-to-Text Translation From Scratch
Real-time voice to text for Windows. Transcribe microphone input to text in any app. Free offline speech recognition tool for Windows 10/11
CLI that turns a YouTube URL into a text transcript: yt-dlp for the audio, Whisper for the transcription, running 100% locally with GPU support.
A specialized RAG-inspired localization pipeline that leverages Whisper-v3-Turbo for high-fidelity Rioplatense transcription and Llama-3.1-8B for deep sociolinguistic decoding of Argentinian "porteño" Spanish dialects.
Speech recognition translation tool using Google Apps Script
Attribution framework for analyzing audio–text context mixing in Speech-to-Text Translation models.
Add a description, image, and links to the speech-to-text-translation topic page so that developers can more easily learn about it.
To associate your repository with the speech-to-text-translation topic, visit your repo's landing page and select "manage topics."