0.4.2
edwko/OuteTTS0.4.2May 19, 2025by edwko
AI Summary
This release introduces fade-in/fade-out audio decoding to eliminate clipping artifacts at segment boundaries, adds high-throughput batched decoding interfaces with multiple backend support (EXL2 Async, VLLM, llama.cpp Async), and adds support for the new OuteTTS 1.0 0.6B model.
Key Highlights
- Fade-in/fade-out audio decoding to eliminate clipping artifacts
- Batched decoding interfaces (EXL2 Async, VLLM, llama.cpp Async Server)
- Single-stream llama.cpp server endpoint
- OuteTTS 1.0 0.6B model support
- Enhanced pre-prompt normalization pipeline
New Features
- Fade-in/fade-out audio decoding for decoded audio chunks
- Batched inference with EXL2 Async backend
- Batched inference with VLLM backend
- Batched inference with llama.cpp Async Server endpoint
- Single-stream decode endpoint for llama.cpp server
- OuteTTS 1.0 0.6B model compatibility
- Batched interface configuration parameters
- Pre-prompt normalization improvements
Full Release Notes
# OuteTTS v0.4.2 * **Fade-in / Fade-out Audio Decoding** Introduced quick fade-in and fade-out on decoded audio chunks to eliminate clipping artifacts at segment boundaries. * **Batched Decoding Interfaces** Added support for high-throughput, batched inference via three new backends: * **EXL2 Async**: Asynchronous batch processing using the EXL2. * **VLLM**: Asynchronous batch decoding with VLLM. (Experiment support) * **llama.cpp Async Server Endpoint**: Connects to a continuously-batched llama.cpp server for async inference. * **Single-Stream Decoding** * **llama.cpp Server Endpoint**: single-stream decode endpoint for llama.cpp server. * **OuteTTS 1.0 0.6B Model Support** Compatibility with the new [OuteTTS-1.0-0.6B](https://huggingface.co/OuteAI/OuteTTS-1.0-0.6B), including config defaults. * **Batched Interface Parameters** New configuration options to control batched interface. * Enhanced pre-prompt normalization pipeline. * Documentation Updates, expanded the batched interface usage.