Files
Doomsday_Survival_Manual/customized_skills/self-improving-agent/references/appendix.md

132 lines
3.1 KiB
Markdown

# Appendix
## Self-Validation
### Validation Report Template
```markdown
## Validation Report Template
**Date**: [YYYY-MM-DD]
**Scope**: [skill(s) validated]
### Checks
- [ ] Examples compile or run
- [ ] Checklists match current repo conventions
- [ ] External references still valid
- [ ] No duplicated or conflicting guidance
### Findings
- [Finding 1]
- [Finding 2]
### Actions
- [Action 1]
- [Action 2]
```
## Memory File Structure
```
~/.claude/memory/
├── semantic/
│ └── patterns.json
├── episodic/
│ ├── 2025/
│ │ ├── 2025-01-11-prd-creation.json
│ │ └── 2025-01-11-debug-session.json
│ └── episodes.json
├── working/
│ ├── current_session.json
│ ├── last_error.json
│ └── session_end.json
└── index.json
```
## Automatic Workflow Integration
```
Any Skill Run
-> workflow-orchestrator
-> self-improving-agent (background)
-> create-pr (ask_first)
-> session-logger (auto)
```
## Continuous Learning Metrics
```json
{
"metrics": {
"patterns_learned": 47,
"patterns_applied": 238,
"skills_updated": 12,
"avg_confidence": 0.87,
"user_satisfaction_trend": "improving",
"error_rate_reduction": "-35%",
"self_corrections": 8
}
}
```
## Human-in-the-Loop
### Feedback Collection
```markdown
## Self-Improvement Summary
I've learned from our session and updated:
### Updated Skills
- `debugger`: Added callback verification pattern
- `prd-planner`: Enhanced UI/UX specification requirements
### Patterns Extracted
1. **state_monitoring_over_callbacks**: Use usePrevious for state-driven side effects
2. **ui_ux_specification_granularity**: Explicit visual specs prevent rework
### Confidence Levels
- New patterns: 0.85 (needs validation)
- Reinforced patterns: 0.95 (well-established)
### Your Feedback
Rate these improvements (1-10):
- Were the updates helpful?
- Should I apply this pattern more broadly?
- Any corrections needed?
```
### Feedback Integration
```yaml
User Feedback:
positive (rating >= 7):
action: Increase pattern confidence
scope: Expand to related skills
neutral (rating 4-6):
action: Keep pattern, gather more data
scope: Current skill only
negative (rating <= 3):
action: Decrease confidence, revise pattern
scope: Remove from active patterns
```
## Templates
| Template | Purpose |
|----------|---------|
| `templates/pattern-template.md` | Adding new patterns |
| `templates/correction-template.md` | Fixing incorrect guidance |
| `templates/validation-template.md` | Validating skill accuracy |
## References
- [SimpleMem: Efficient Lifelong Memory for LLM Agents](https://arxiv.org/html/2601.02553v1)
- [A Survey on the Memory Mechanism of Large Language Model Agents](https://dl.acm.org/doi/10.1145/3748302)
- [Lifelong Learning of LLM based Agents](https://arxiv.org/html/2501.07278v1)
- [Evo-Memory: DeepMind's Benchmark](https://shothota.medium.com/evo-memory-deepminds-new-benchmark)
- [Let's Build a Self-Improving AI Agent](https://medium.com/@nomannayeem/lets-build-a-self-improving-ai-agent-that-learns-from-your-feedback-722d2ce9c2d9)