v2.0

Fictionarry/TalkingGaussianv2.0Mar 31, 2026by lipku

AI Summary

Major architectural refactoring separating code into specific directories and adopting a plugin-based architecture for models and TTS.

Key Highlights

  • Refactored code structure with models and audio features moved to avatars directory
  • Switched models, TTS, and transport methods to plugin architecture
  • Added RTMP output in streamout directory
  • Integrated Aliyun QwenTTS
  • Unified inference and reply code in BaseAvatar class

New Features

  • Plugin architecture for models and TTS
  • RTMP output support
  • Aliyun QwenTTS integration

Full Release Notes

1,重构代码,数字人模型和音频特征代码移到avatars目录下
2,数字人模型、tts、传输方式改成plugin方式接入
3,传输方式整理成单独的类,放到streamout目录下,添加rtmp输出
4,数字人推理和回贴公共部分代码整合到BaseAvatar中,各模型只需要实现自己特有部分代码
5,音频特征的切分代码整合到BaseAsr中
6,tts添加阿里云qwentts