# Self-Improving AI Agents

> Status: Exploring

Exploring how agents improve across attempts through execution feedback, verification, memory, and evaluation.

## The system around the model

The question is whether coding and research agents get more reliable by treating their own output as something to execute, inspect, evaluate and revise, rather than returning the first plausible result.

The interesting work sits around the model rather than inside it: better verification, better retry decisions, execution feedback, memory worth keeping, search and branching, and lessons that survive a failure.

## Attempt, verify, improve

1. **Task**: A bounded goal with its constraints attached.
2. **Attempt**: The agent works through tools, code, and search.
3. **Verify**: Execution and tests decide, not a second opinion.
4. **Improve**: A failure becomes feedback rather than a retry.
5. **Retain**: Lessons worth keeping go to memory; the rest does not.

## Open questions

- What feedback should survive between attempts, and what should not?
- What belongs in persistent memory rather than temporary task context?
- When is a deterministic verifier better than another model reviewing the work?
- How should an agent decide whether to retry, branch, search, or stop?
- How do you measure improvement without over-fitting to a weak evaluator?
- How much can improve without touching model weights?
- Where does human review belong in the loop?

## Current scope

The current scope is agent-system design around feedback, verification, memory, evaluation, retry decisions, and human review.

## Topics

- Verification
- Feedback
- Evaluation
