v3.59

jianchang512/pyvideotransv3.59Feb 28, 2025by jianchang512

AI Summary

Integrated ElevenLabs for speech recognition and added Minimaxi TTS support.

Key Highlights

  • Added Elevenlabs.io as a new speech recognition channel.
  • Added Minimaxi TTS support in the custom API section.
  • Fixed ElevenLabs.io TTS functionality.

New Features

  • Elevenlabs.io speech recognition
  • Minimaxi TTS support

Full Release Notes

> 预打包版仅适用于Windows10/11,MacOS和Linux请源码部署
> MacOS/Linux升级:重新拉取源码覆盖,然后执行
>  `pip3 install --upgrade openai-whisper elevenlabs` 
>  `pip3 install --no-deps  --force-reinstall "faster-whisper @ https://github.com/SYSTRAN/faster-whisper/archive/refs/heads/master.tar.gz"`

## Change

- Feat: 在自定义TTS API 中接入 Minimaxi 配音,文档 https://pvt9.com/minimaxi
- Fix: #751 #671 
- Fix: elevenlabs.io TTS
- Feat: 语音识别新增`Elevenlabs.io`渠道,来自 [11ElevenLabs scribe_v1模型](https://elevenlabs.io/docs/api-reference/speech-to-text/convert)


## Win v3.59 完整包下载

> 如果未安装过旧版本,请在此下载完整版,如需cuda加速,需有英伟达显卡并且安装cuda12.x及cudnn9

百度网盘下载地址(含tiny/medium模型):  https://pan.baidu.com/s/1oh9oxrMlGIlTywbAtIHIxQ?pwd=sa9j

GitHub地址: https://github.com/jianchang512/pyvideotrans/releases/download/v3.59/win-videotrans-v3.59-tiny.7z

 

## v3.63 补丁包190MB

> 如果已安装过3.x版本,可下载补丁包后解压在sp.exe所在目录,覆盖已有sp.exe和文件夹

百度网盘下载地址: https://pan.baidu.com/s/1F3hmCpPxSs0X8qAnWfI_UA?pwd=im7f

GitHub地址: https://github.com/jianchang512/pyvideotrans/releases/download/v3.59/win-PatchUpdate-3.63.7z


> 若补丁覆盖后打开提示`需下载完整包`,请升级显卡驱动、升级cuda到12.x、升级cudnn到cudnn9,并重新下载完整包


----

### 所有模型下载地址

为避免压缩包体积过大,预打包版只内置最小模型 tiny,识别效果不佳,效果更好的模型请点击下载 ,建议至少使用medium模型,推荐large-v2

https://github.com/jianchang512/stt/releases/tag/0.0