v1.0.1
EricLBuehler/mistral.rsv1.0.1Jan 6, 2026by asdek
AI Summary
This update focuses on improving AI agent reliability through enhanced error diagnostics, a stable DuckDuckGo search implementation, and updated OpenAI provider configurations to handle content filters and instability.
Key Highlights
- Enhanced Error Diagnostics with stop reason handling and automatic max_tokens adjustment
- DuckDuckGo Search Stability via new API migration and HTML parsing
- Provider Guardrails Bypass with explicit authorization framework
- Switched OpenAI model to o4-mini due to prompt evaluation instability
New Features
- Authorization framework to all agent prompts
- Customer interaction protocol for AskUser tool
- Enhanced message formatting in vector store communications with document match scores
Full Release Notes
## 🐛 Bug Fixes & Improvements ### Enhanced Error Diagnostics - Added stop reason to error messages when LLM fails to generate tool calls - If stop reason is `length`, increase `max_tokens` parameter for the affected agent in provider settings - Improves troubleshooting and configuration optimization ### DuckDuckGo Search Stability - Migrated to new DuckDuckGo API with HTML response parsing - Added comprehensive test coverage with real-world search scenarios - Significantly improved reliability and result quality ### Provider Guardrails Bypass - Added explicit authorization framework to all agent prompts - Prevents blocking by OpenAI, Anthropic, and Google Gemini content filters - Clarified penetration testing context as pre-authorized activity ### OpenAI Configuration Updates - Temporarily switched from `gpt-5` to `o4-mini` for primary agent and assistant due to OpenAI prompt evaluation instability - Increased `max_tokens` limits across multiple agents for better output capacity - **Recommendation**: Enable Human-in-the-loop mode (`ASK_USER=true` in `.env`) when using OpenAI provider for improved stability ### Additional Improvements - Enhanced message formatting in vector store communications with document match scores - Improved clarity in generator and refiner prompts for user task interpretation - Added customer interaction protocol for AskUser tool --- **Full Changelog**: https://github.com/vxcontrol/pentagi/compare/v1.0.0...v1.0.1