πAbout
Post-Execution Evaluator is a droid that helps Argus check whether completed tasks meet their success criteria and pass their test plans. It checks technical correctness, looks for broken or missing functionality, and produces a report explaining what passed, what failed, and why.
It evaluates criteria set before work began. It identifies problems but does not fix them, and does not assess design, user experience, or documentation.
πUsage
Post-Execution Evaluator runs on demand after a task finishes, giving Argus a clear verdict before the work ships. Its checks include whether the frontend loads and displays content correctly. Critical problems it cannot handle are escalated to Argusβs manager and specialist agents.
πPurpose
When a task finishes, this droid verifies it actually worked. It checks the success criteria, validates the tests, and confirms the frontend still functions. Argus gets a detailed pass/fail report with specifics on any failures.
πDuties
- Verifies each success criterion in the completed work
- Checks whether the test plan passed
- Looks for any broken or missing functionality
- Tests the frontend in two ways: code-level checks (syntax, API paths) and in-browser checks (app loads, pages display correctly)
- Creates a report listing what passed, what failed, and why
- If critical problems arise that the droid can't handle, notifies Argus's manager and specialist agents
πConstraints
- Only checks against criteria that were set before work started; does not define new tests
- Identifies problems but does not fix them
- Frontend tests focus on critical issues only (broken code, bad paths, app unresponsive, missing content)
- Stops trying and escalates if the evaluation service fails repeatedly
- Only checks technical correctness and task completion; does not assess design, user experience, or documentation
π€Belongs to
Source: Droid Families Β· Public information reflected here.