v0.5.18

met4citizen/TalkingHeadv0.5.18Jul 3, 2026by decolua

AI Summary

This release enhances usage tracking with cached token costs and introduces support for new NVIDIA models and the ClinePass provider. It also fixes critical bugs related to streaming, cross-IdP authentication, and provider-specific tool handling.

Key Highlights

  • Usage tracking now correctly calculates input/output/cache costs.
  • Added support for new NVIDIA models and capabilities.
  • Added support for the ClinePass provider.
  • Fixed cross-IdP account overwrites and non-SSE stream pipe crashes.

New Features

  • Usage tracking for cached tokens.
  • Reset credit expiry details for Codex.
  • New NVIDIA models.
  • ClinePass provider support.
  • Region selector and key validation for Xiaomi-tokenplan.

Full Release Notes

Features:
- Usage: track cached tokens + correct input/output/cache cost (#2209)
- Codex: show reset credit expiry details (#2290)
- NVIDIA: add new models and capabilities
- ClinePass: add provider support

Fixes:
- Usage: dedupe streaming request-details log entries
- Claude: drop foreign thinking signatures in passthrough
- Prevent non-SSE stream pipe crash and cross-IdP account overwrites (#2244)
- Kiro: route IdC auth to regional CodeWhisperer surface (#2297)
- Kiro: add Claude Sonnet 5 model support (#2264)
- Xiaomi-tokenplan: region selector, key validation, multi-connection (#2251)
- Translator: strict Anthropic content block compliance (#2225)
- Kimchi: strip reasoning_content echo to bound multi-turn input tokens
- Kimchi: bump User-Agent to kimchi/0.1.40 (#2256)
- Codebuddy-cn: strip empty tool_calls arrays to preserve reasoning
- Antigravity: preserve Claude tool delta index (#2223)
- MITM: generate root CA on server startup (#2228)