Inflect-Micro-v2: complete voice in 9.36M parameters
Inflect-Micro-v2 is a text-to-speech synthesis model that delivers complete voice generation in under 10 million parameters (9.36M), enabling local deployment on CPU or CUDA with deterministic, reproducible outputs. The model achieved a 66.2% human preference rate against comparable compact TTS systems while maintaining a small 37.53 MB footprint and real-time performance capabilities. The independently-built project supports long-text handling and fixed-voice English synthesis, with potential for future expansion to include additional languages and voices.
Read Full Article →