v1.0.1

EricLBuehler/mistral.rsv1.0.1Jan 6, 2026by asdek

AI Summary

This update focuses on improving AI agent reliability through enhanced error diagnostics, a stable DuckDuckGo search implementation, and updated OpenAI provider configurations to handle content filters and instability.

Key Highlights

  • Enhanced Error Diagnostics with stop reason handling and automatic max_tokens adjustment
  • DuckDuckGo Search Stability via new API migration and HTML parsing
  • Provider Guardrails Bypass with explicit authorization framework
  • Switched OpenAI model to o4-mini due to prompt evaluation instability

New Features

  • Authorization framework to all agent prompts
  • Customer interaction protocol for AskUser tool
  • Enhanced message formatting in vector store communications with document match scores

Full Release Notes

## 🐛 Bug Fixes & Improvements

### Enhanced Error Diagnostics
- Added stop reason to error messages when LLM fails to generate tool calls
- If stop reason is `length`, increase `max_tokens` parameter for the affected agent in provider settings
- Improves troubleshooting and configuration optimization

### DuckDuckGo Search Stability
- Migrated to new DuckDuckGo API with HTML response parsing
- Added comprehensive test coverage with real-world search scenarios
- Significantly improved reliability and result quality

### Provider Guardrails Bypass
- Added explicit authorization framework to all agent prompts
- Prevents blocking by OpenAI, Anthropic, and Google Gemini content filters
- Clarified penetration testing context as pre-authorized activity

### OpenAI Configuration Updates
- Temporarily switched from `gpt-5` to `o4-mini` for primary agent and assistant due to OpenAI prompt evaluation instability
- Increased `max_tokens` limits across multiple agents for better output capacity
- **Recommendation**: Enable Human-in-the-loop mode (`ASK_USER=true` in `.env`) when using OpenAI provider for improved stability

### Additional Improvements
- Enhanced message formatting in vector store communications with document match scores
- Improved clarity in generator and refiner prompts for user task interpretation
- Added customer interaction protocol for AskUser tool

---

**Full Changelog**: https://github.com/vxcontrol/pentagi/compare/v1.0.0...v1.0.1