3.3.0
pyannote/pyannote-audio3.3.0Jun 14, 2024by hbredin
Full Release Notes
## TL;DR `pyannote.audio` does [speech separation](https://hf.co/pyannote/speech-separation-ami-1.0): multi-speaker audio in, one audio channel per speaker out! ```bash pip install pyannote.audio[separation]==3.3.0 ``` ## New features - feat(task): add `PixIT` joint speaker diarization and speech separation task (with [@joonaskalda](https://github.com/joonaskalda/)) - feat(model): add `ToTaToNet` joint speaker diarization and speech separation model (with [@joonaskalda](https://github.com/joonaskalda/)) - feat(pipeline): add `SpeechSeparation` pipeline (with [@joonaskalda](https://github.com/joonaskalda/)) - feat(io): add option to select torchaudio `backend` ## Fixes - fix(task): fix wrong train/development split when training with (some) meta-protocols ([#1709](https://github.com/pyannote/pyannote-audio/issues/1709)) - fix(task): fix metadata preparation with missing validation subset ([@clement-pages](https://github.com/clement-pages/)) ## Improvements - improve(io): when available, default to using `soundfile` backend - improve(pipeline): do not extract embeddings when `max_speakers` is set to 1 - improve(pipeline): optimize memory usage of most pipelines ([#1713](https://github.com/pyannote/pyannote-audio/pull/1713) by [@benniekiss](https://github.com/benniekiss/))