I trained a 9M speech model to fix my Mandarin tones
I trained a 9M speech model to fix my Mandarin tones
Built this because tones are killing my spoken Mandarin and I can't reliably hear my own mistakes. It's a 9M Conformer-CTC model trained on ~300h (AISHELL + Primewords), quantized to INT8 (11 MB), runs 100% in-browser via ONNX Runtime Web. Grades per-syllable pronunciation + tones with Viterbi forced alignment. Try it here: https://simedw.com/projects/ear/
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Correct prediction on native model
Similar products
New Cartesia Text-to-Speech Model
Conversational speech model that achieves voice presence
The next-generation multilingual text-to-speech model
The most expressive Text to Speech model ever
ThiruvalluvarGPT – GPT model trained using Tamil Tirukkural Poems
Real-time text-to-speech model you can self-host
I trained an AI to understand and fix command-line errors
Sense2vec model trained on all 2015 Reddit comments
On the 50th anniversary of MLK's speech, here is our tribute
Speech.is interop for Namecoin