chore: release api v1.0.0
- Event-based streaming for task agent - Retry logic when LLM responds without calling tools - Hardened prompts to enforce tool use - Working directory context in all agent prompts - Project paused: local LLMs not capable enough for agentic use Co-Authored-By: Claude Opus 4.5 <noreply@anthropic.com>
This commit is contained in:
@@ -7,6 +7,29 @@ and this project adheres to [Semantic Versioning](https://semver.org/spec/v2.0.0
|
||||
|
||||
## [Unreleased]
|
||||
|
||||
## [1.0.0] - 2026-01-15
|
||||
|
||||
### Added
|
||||
- Event-based streaming for task agent (`StreamEvent` objects instead of raw text)
|
||||
- New `tools_streaming.py` with all tools emitting structured events
|
||||
- Event types: `tool_start`, `tool_done`, `thinking`, `response`, `error`, `done`
|
||||
- Retry logic when LLM responds without calling tools (max 2 retries)
|
||||
- Tracks `tools_called` counter on TaskContext
|
||||
- Stronger retry prompt forces tool use
|
||||
- Working directory context injected into all agent prompts
|
||||
|
||||
### Changed
|
||||
- Hardened system prompts to enforce tool use before responding
|
||||
- Added "CRITICAL RULE" section requiring tool calls first
|
||||
- Made "MANDATORY WORKFLOW" more emphatic
|
||||
- Updated explore and plan agents with `_build_prompt_with_context()` method
|
||||
|
||||
### Fixed
|
||||
- Agent path hallucination - now explicitly communicates working directory to LLM
|
||||
|
||||
### Note
|
||||
- Project paused: Local LLMs (Mistral Nemo 12B on available hardware) are not capable enough for reliable agentic tool use. Models frequently hallucinate responses instead of calling tools, even with prompt hardening and retry logic. Would require larger models (70B+) or cloud API integration to continue.
|
||||
|
||||
## [0.4.2] - 2026-01-11
|
||||
|
||||
### Added
|
||||
|
||||
Reference in New Issue
Block a user