Offline-first voice AI (<1 s latency, MLX, Apple Silicon)
Offline-first voice AI (<1 s latency, MLX, Apple Silicon)
Hey HN, I built an offline-first voice AI that runs entirely on Apple Silicon using MLX + FastAPI. It achieves <1s end-to-end latency for speech-to-speech conversations, with a minimal UI. Repo: https://github.com/shubhdotai/offline-voice-ai Would love feedback on performance, model choices, and other ideas...
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Incorrect prediction on native model
Similar products
Talk to your Mac offline – sub-second Voice AI (Apple Silicon and MLX)
I made Android boot on Apple Silicon
MLX (Apple Silicon tensor library) bindings for Erlang
Offline encrypted vault for Apple Silicon Mac.
Silicon Feelings
Where's the latency in my API?
Voice AI for websites
Write in your own voice with AI
ClassicTunes – a from-scratch remake of iTunes 7-10 for Apple Silicon
Roq – C++ HFT on Crypto Exchanges with μs Latency