v0.2.24
withastro/astrov0.2.24Jul 11, 2026by QuentinFuxa
AI Summary
This release introduces a new streaming LLM translation backend called AlignAtt and refreshes benchmark figures. It also fixes issues with empty captions in mlx-whisper and adds loud failure modes for the ASR backend.
Key Highlights
- Added AlignAtt translation backend for streaming LLM translation.
- Refreshed benchmark figures measured on H100.
- Fixed mlx-whisper producing empty captions with torch >= 2.13.
- Added loud failure handling for ASR backend startup errors.
New Features
- AlignAtt translation backend
Full Release Notes
Streaming LLM translation, refreshed benchmarks and loud failures. ## Added - AlignAtt translation backend (`--translation-backend alignatt`): streaming LLM translation through an [Alignatt4LLM](https://github.com/QuentinFuxa/Alignatt4LLM) sidecar. The model drafts ahead over the unstable ASR tail and commits only target words whose attention lands on committed source words, so translations are append-only and released the instant the ASR commits. See [docs/translation-alignatt.md](https://github.com/QuentinFuxa/WhisperLiveKit/blob/main/docs/translation-alignatt.md). - Benchmark figures re-measured on an H100 with the current backends (the README scatter plots dated from March and misrepresented the qwen3 backends). Raw results in `benchmarks/h100_scatter/`. ## Fixed - mlx-whisper produced empty captions with torch >= 2.13 (#383, thanks @RobertBartelds-FleetEnergies). - The server now fails loudly when the ASR backend produces nothing: warmup errors abort startup with the cause, and a watchdog logs an explicit error if audio flows but no text is ever produced.