Speakrail releases a fully-local, low-latency voice assistant running on a single RTX 4090
Read the original at www.reddit.com→tl;dr: I created a fully-local open-source full-duplex voice agent that rivals GPT-Live on some benchmarks. It uses Voxtral Realtime with a turn-taking head, a microturn-finetuned Gemma 4 12B and Breeze TTS 2 under...
Original headline: "Speakrail - a low-latency fully-local voice assistant that runs on a single RTX 4090"
Coverage timeline
- Oct 5, 15:50 UTC r/LocalLLaMA lead source Speakrail - a low-latency fully-local voice assistant that runs on a single RTX 4090