Talk to your Mac offline – sub-second Voice AI (Apple Silicon and MLX)
Talk to your Mac offline – sub-second Voice AI (Apple Silicon and MLX)
I wanted a voice assistant that feels realtime but runs completely offline. This prototype uses MLX + FastAPI on Apple Silicon to hit sub-second latency for speech-to-speech conversations. Repo: https://github.com/shubhdotai/offline-voice-ai It’s fast, minimal, and hackable — would love feedback on latency tricks, model swaps, or use-cases you’d like to see next.
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Incorrect prediction on native model
Similar products
Offline-first voice AI (<1 s latency, MLX, Apple Silicon)
Offline encrypted vault for Apple Silicon Mac.
Sub-Second Analytics with ClickHouse and Apache Kafka at DoubleCloud
Apple – Every Second
I made Android boot on Apple Silicon
MLX (Apple Silicon tensor library) bindings for Erlang
Fastly Pub/Sub
Sub-Modals in BootStrap
Dedicated Apple silicon Mac minis, rented by the month.
Enlist AI: Sub-second interview coaching with persistence