YOU ARE VIEWING ONE ITEM FROM THE AICRIER FEED

Whistle Packs Multilingual Speech-to-Text into 16.9 MB.

AICrier tracks AI developer news across Product Hunt, GitHub, Hacker News, YouTube, X, arXiv, and more. This page keeps the article you opened front and center while giving you a path into the live feed.

// WHAT AICRIER DOES

7+

TRACKED FEEDS

24/7

SCRAPED FEED

Short summaries, external links, screenshots, relevance scoring, tags, and featured picks for AI builders.

Whistle Packs Multilingual Speech-to-Text into 16.9 MB.
OPEN LINK ↗
// 1h agoMODEL RELEASE

Whistle Packs Multilingual Speech-to-Text into 16.9 MB.

Cactus Compute released Whistle, an open CPU speech-recognition model that transcribes seven languages, supports 30-second inputs, word timestamps, speech embeddings, and on-device processing. It shares the Needle runtime and weighs just 16.9 MB. [Cactus launch post](https://www.cactuscompute.com/blog/whistle)

// ANALYSIS

Whistle makes local voice interfaces substantially more practical for constrained devices, though its real-world accuracy still needs independent testing.

  • –Runs offline across phones, wearables, robots, vehicles, browsers, and microcontrollers
  • –Combines transcription with keyword biasing, silence handling, timestamps, and embeddings
  • –Cactus reports faster inference and lower size than Whisper base on several benchmarks
  • –Published results are company-reported, with Whisper still ahead on some datasets
  • –Shared runtime enables a direct path from speech input to local structured tool calls
// TAGS
whistlespeechsttedge-aiinferenceopen-sourcelocal-first

DISCOVERED

1h ago

2026-10-04

PUBLISHED

1h ago

2026-10-04

RELEVANCE

8/ 10

AUTHOR

AI Search