Offline GPU-accelerated audio ML pipeline. faster-whisper + cross-encoder reranking + FFmpeg. Runs on 6GB VRAM.
-
Updated
Dec 26, 2025 - Python
Offline GPU-accelerated audio ML pipeline. faster-whisper + cross-encoder reranking + FFmpeg. Runs on 6GB VRAM.
A fully local OpenCode + llama.cpp + Ornith 1.5 coding-agent setup tuned for 128K context on a tight 6 GB VRAM budget.
To associate your repository with the 6gb-vram topic, visit your repo's landing page and select "manage topics."