Back to feed
Reddit r/LocalLLaMA·

I released Inflect-Nano, an ultra-extreme tiny 4.63m parameter TTS model.

Signal
72
Hype
25
In three linesInflect-Nano-v1, a 4.63M parameter TTS model, is the 2nd smallest publicly released speech synthesis model. Comprises acoustic model (3.46M) and vocoder (1.17M), generates 24 kHz English audio. ~17x smaller than Kokoro, ~108x smaller than Chatterbox. Runs locally via PyTorch, suited for embedded devices and offline voice assistants.
Read source
Your take?
VoiceOpen sourceToolsInfrastructure

Summary generated by Claude — human-verified