Inflect-Micro-v2 ships standalone 9.36M parameter TTS
Inflect-Micro-v2, released by developer Owen Song on Hugging Face, is an open-weight text-to-speech (TTS) model designed for efficient local execution. Packing a complete end-to-end speech generation pipeline into 9.36 million parameters (~37.5 MB FP32), it generates 24 kHz English speech directly on CPU or CUDA hardware without external vocoders.
Shrinking complete end-to-end speech generation models under 10 million parameters unlocks real-time, privacy-focused voice synthesis on low-power edge hardware.
- –High Efficiency: At just 9.36M parameters, the model enables rapid local inference on standard CPUs and GPUs with minimal memory consumption.
- –Standalone Architecture: Integrates text processing, timing prediction, and waveform generation without requiring external vocoder models.
- –Open & Accessible: Published under an open license on Hugging Face with an interactive space for immediate browser testing.
DISCOVERED
1h ago
2026-07-26
PUBLISHED
5h ago
2026-07-26
RELEVANCE
AUTHOR
nateb2022