0.4.2

zumerlab/snapdom0.4.2May 19, 2025by edwko

AI Summary

This release focuses on improving audio quality and throughput. It introduces fade-in/out audio decoding to prevent clipping at segment boundaries, adds high-throughput batched decoding interfaces (EXL2 Async, VLLM, llama.cpp), and supports the new OuteTTS 1.0 0.6B model.

Key Highlights

  • Fade-in / Fade-out Audio Decoding to eliminate clipping artifacts
  • Batched Decoding Interfaces (EXL2 Async, VLLM, llama.cpp Async Server)
  • Single-stream decoding endpoint for llama.cpp server
  • OuteTTS 1.0 0.6B Model Support

New Features

  • Fade-in / Fade-out audio decoding
  • Batched decoding interfaces (EXL2, VLLM, llama.cpp)
  • Single-stream decoding
  • OuteTTS 1.0 0.6B model compatibility
  • Enhanced pre-prompt normalization
  • Documentation updates for batched interfaces

Full Release Notes

# OuteTTS v0.4.2

* **Fade-in / Fade-out Audio Decoding**
  Introduced quick fade-in and fade-out on decoded audio chunks to eliminate clipping artifacts at segment boundaries.

* **Batched Decoding Interfaces**
  Added support for high-throughput, batched inference via three new backends:

  * **EXL2 Async**: Asynchronous batch processing using the EXL2.
  * **VLLM**: Asynchronous batch decoding with VLLM. (Experiment support)
  * **llama.cpp Async Server Endpoint**: Connects to a continuously-batched llama.cpp server for async inference.

* **Single-Stream Decoding**

  * **llama.cpp Server Endpoint**: single-stream decode endpoint for llama.cpp server.

* **OuteTTS 1.0 0.6B Model Support**
  Compatibility with the new [OuteTTS-1.0-0.6B](https://huggingface.co/OuteAI/OuteTTS-1.0-0.6B), including config defaults.

* **Batched Interface Parameters**
  New configuration options to control batched interface.

* Enhanced pre-prompt normalization pipeline.

* Documentation Updates, expanded the batched interface usage.