Whistle: Tiny 16.9MB Speech-to-Text Model for Edge Devices
· via Hacker News
CactusCompute released Whistle, a lightweight speech recognition model designed for mobile, wearable, and edge devices. The 16.9MB model runs on CPU with no dependencies, offering transcription, word timestamps, and speech embedding directly on the device. It supports multiple languages and features efficient decoding with beam search and keyword biasing.
Read the full article
Continue reading at Hacker News →This is an AI-generated summary. Read the original for the full story.