Chat UIs with native audio input for multimodal models; discussion of direct audio file submission to models without separate STT layer
Read the original at old.reddit.com→I've been running Gemma 4 E4B with oMLX and I can't find any chat interfaces that directly send the audio file to the model instead of running the audio through a separate STT layer. I can confirm the audio layers...
Original headline: "Chat UIs with native audio input for multimodal models?"
Coverage timeline
- Aug 10, 09:08 UTC r/LocalLLaMA lead source Chat UIs with native audio input for multimodal models?