Inflect-Micro-v2: A Complete Voice Model Implemented with 9.36M Parameters
Key point
Inflect-Micro-v2, a highly efficient TTS model capable of local execution with fewer than 10 million parameters, has been released.
Details
Inflect-Micro-v2, a complete local TTS (Text-to-Speech) model that converts text into waveforms, has been released. This model boasts high efficiency using only 9.36M (approximately 9.36 million) parameters.
Key features are as follows:
- Two sizes available: It consists of a quality-focused Micro model (under 10M) and a size-focused Nano model (under 4M).
- Efficient inference: Inference is possible on both CPU and CUDA environments, and it supports 24 kHz mono output.
- Control features: It has fixed-voice implementation through deterministic seeds and long-text processing capability.
- Lightweight: The model size is very light at approximately 37.53 MB in FP32.
This model is a suitable solution for developers looking to implement stable speech synthesis in a local environment without a high-performance server.
This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.