Command line interface for the built-in speech recognition and transcription capabilities in macOS.
-
Updated
Aug 30, 2026 - Objective-C
Command line interface for the built-in speech recognition and transcription capabilities in macOS.
💬 Fast, cross-platform CLI and GUI for batch transcription, translation, speaker annotation and subtitle generation using OpenAI’s Whisper on CPU, Nvidia GPU and Apple MLX.
一个可以让 AI 帮你把台本转换为带时间轴的字幕的工具。
OCTRA is a web-application for the orthographic transcription of audio files.
Offline macOS menu bar app for speech-to-text transcription using AI models (NVIDIA Parakeet, Whisper). 100% local & private.
🎵 Complete offline audio transcription system with speaker diarization using OpenAI Whisper and PyAnnote. Features automatic audio cleaning, precise timestamps, multiple output formats (JSON/TXT/Markdown), and support for 20+ audio formats. No external APIs required - works entirely offline.
Modular tool for Digital Humanities: IIIF downloader + Studio environment. Supports PDF import, hybrid OCR/HTR, side-by-side manual correction, and global library search. 🛠️📜
French audio transcription using gradio
WhisperVoice is a browser extension that converts speech to text in real-time using speech recognition APIs. It’s perfect for quick transcriptions, note-taking, and accessibility, supporting multiple languages and customizable settings for a tailored experience.
Speakr — Free open source alternative to Wispr Flow. Desktop voice dictation with global hotkey, Groq Whisper transcription, and auto-paste into any app.
The Whisper Subtitle Generator leverages OpenAI's Whisper model to generate subtitles from audio and video files. This Python-based tool supports multiple languages and employs advanced audio processing techniques to ensure high accuracy in transcription.
AI-Video-Transcriber is an intelligent, open-source tool that automatically transcribes video and audio files using advanced artificial intelligence. It supports multiple languages, accurate speech recognition, and provides easy-to-read text transcripts for content creators, educators, and businesses.
Self-hosted AI tool to transcribe and summarize meetings. Upload audio files, transcribe with Whisper, and generate structured summaries using OpenAI GPT or Google Gemini.
Audio transcription and practice tool with waveform visualization, loop regions, timeline tags, and pitch-preserving playback for musicians.
Dictator – Supercharge Cursor Chat with voice-to-text, custom AI prompts, and workflow automation. Speak your ideas, inject templates instantly, and code faster with AI-powered assistance.
字幕一站式视频处理工具:字幕转换 / 轨道提取封装(MKVToolNix)/ AI 翻译(术语表·断点续传·润色)/ Whisper 转写(双后端)/ 烧录压制,Flutter Windows 桌面版
Open Video Transcribe - Open-source video transcription tool that emphasizes the primary use case: transcribing video files to text with support for multiple model types.
AI-powered audio transcription with post-processing, supports OpenAI Whisper/GPT and Google Gemini with automatic fallback.
VoicePad is a terminal-first voice note and transcription tool built with Python and Whisper. It records audio, generates local AI-powered transcriptions, and streamlines voice-based workflows through an ergonomic CLI and developer-friendly automation features.
Convert audio and video to text directly on your own machine - 100% local, completely offline, and fully private.
To associate your repository with the transcription-tool topic, visit your repo's landing page and select "manage topics."