v2.0
clusterzx/paperless-aiv2.0Mar 31, 2026by lipku
AI Summary
Major refactoring of the digital avatar and audio feature code structure, migrating to a plugin-based architecture and adding RTMP output support.
Key Highlights
- Refactored code structure moving models and audio features to `avatars` directory
- Migrated digital avatar, TTS, and transport methods to plugin-based architecture
- Added RTMP output support
- Integrated common inference and audio feature code
New Features
- Plugin-based architecture for avatar and audio components
- RTMP output stream support
- Consolidated common inference code into BaseAvatar
- Consolidated audio feature splitting code into BaseAsr
Full Release Notes
1,重构代码,数字人模型和音频特征代码移到avatars目录下 2,数字人模型、tts、传输方式改成plugin方式接入 3,传输方式整理成单独的类,放到streamout目录下,添加rtmp输出 4,数字人推理和回贴公共部分代码整合到BaseAvatar中,各模型只需要实现自己特有部分代码 5,音频特征的切分代码整合到BaseAsr中 6,tts添加阿里云qwentts