v1.3.1
stanfordnlp/dspyv1.3.1Feb 6, 2026by denizsafak
AI Summary
This release introduces a major architectural expansion with a Flask-based Web UI and an EPUB 3 packaging pipeline, alongside advanced audio generation features like multi-voice support and GPU-accelerated TTS.
Key Highlights
- Massive contribution (>55k lines) enabling Web UI and EPUB 3 pipeline.
- New Flask-based Web UI (`abogen-web`) for Docker and headless deployments.
- Supertonic TTS engine support with GPU acceleration.
- Multi-voice 'theatrical' audiobook support with speaker/role assignment.
- Integration with Calibre OPDS and Audiobookshelf.
New Features
- EPUB 3 packaging pipeline
- Flask-based Web UI
- Supertonic TTS engine (GPU)
- Entity analysis and pronunciation override
- Speaker/role assignment for multi-voice
- Calibre OPDS and Audiobookshelf integration
Full Release Notes
- Special thanks to [@jeremiahsb](https://github.com/jeremiahsb) for his [massive contribution](https://github.com/denizsafak/abogen/pull/120) (>55k lines!) that brought the Web UI, EPUB 3 pipeline, and core architectural improvements to life. - Added an EPUB 3 packaging pipeline that builds media-overlay EPUBs from generated audio and chunk metadata. - Persisted chunk timing metadata in job artifacts and exercised the exporter with automated tests. - Added Flask-based Web UI (`abogen-web`) for Docker and headless server deployments. - Reorganized codebase to support both PyQt6 desktop GUI and Web UI from a shared core. - Added Supertonic TTS engine support with GPU acceleration. - Added entity analysis and pronunciation override system for proper nouns. - Added speaker/role assignment for multi-voice "theatrical" audiobooks. - Added Calibre OPDS and Audiobookshelf integration. Update: ``` uv tool update abogen ```