v0.2.23
pathwaycom/pathwayv0.2.23Jul 9, 2026by QuentinFuxa
AI Summary
The release adds support for Qwen3-ASR backends, provides prebuilt Docker images for easier deployment, and introduces API token authentication and per-session language settings.
Key Highlights
- Qwen3-ASR backends (PyTorch and vLLM) with reduced compute requirements
- Prebuilt Docker images on GHCR
- Optional API token authentication
- Per-session language and translation target configuration
New Features
- Qwen3-ASR backends (qwen3-streaming, qwen3-vllm, qwen3-vllm-metal)
- Prebuilt Docker images on GHCR
- Optional API token auth (--api-token)
- Per-session language and translation target via WebSocket query params
Full Release Notes
Qwen3-ASR backends, latency work and bug fixes. ## Added - Qwen3-ASR backends: `qwen3-streaming` (PyTorch, CUDA/MPS/CPU), `qwen3-vllm` (CUDA), `qwen3-vllm-metal` (Apple Silicon), via the new [qwen3-asr-causal](https://github.com/QuentinFuxa/Qwen3-ASR-causal) package. The causal streaming mode uses 3x less compute, at constant cost per audio second (English only). - Prebuilt Docker images on GHCR: `ghcr.io/quentinfuxa/whisperlivekit`. - Optional API token auth (`--api-token`). - Per-session language and translation target via WebSocket query params. ## Fixed - `mode=full` sessions no longer lose early lines after 5 minutes (#372). - REST transcription of long files no longer times out silently (#374). - qwen3-vllm word timestamps and CJK spacing (#375, thanks @samx81). - Sortformer memory growth on long sessions (#349). - Local Sortformer weights: `--sortformer-model-path` (#378, thanks @felixmr1).