New 'Whistle' model brings on-device speech-to-text in 16.9 MB
A new speech recognition model called Whistle has been released for mobiles, wearables, robots, smart home devices, automotive systems and microcontrollers. It ships as a single 16.9 MB file that runs entirely on-device via CPU with no dependencies, handling transcription, word-level timestamps and speech embeddings for seven languages without sending audio off the device.