Cactus Whistle
Cactus released Whistle, a 16.9 MB on-device speech-recognition model for CPU deployment. It transcribes up to 30 seconds of audio in seven languages and also returns word timestamps and speech embeddings without sending audio off-device.
For buildersTest it on short, representative clips on the target device, measuring word errors, timestamps, latency, and performance on names before using transcripts to trigger actions.