Runs on: Both · Category: quality

agent-self-evaluation

Use after completing any non-trivial task. The agent self-rates its output on 5 axes — accuracy, completeness, clarity, actionability, conciseness — with concrete evidence per criterion.

After a non-trivial task, the agent rates its own output on 5 axes (accuracy, completeness, clarity, actionability, conciseness) with evidence, producing a 1-5 scorecard plus improvement suggestions. Reflection, not a pass/fail gate.

Cross-harness compatibility

Verified: Jul 31, 2026 · Hermes 0.16.0 · OpenClaw 2026.7.1-2