CI / build (push) Has been cancelled
- Archive 79 completed framework-dev tasks from tasks/ -> tasks/complete/ - status.py: add --cleanup-done and --install-cleanup-schedule commands - Add scripts/automaton-cleanup.sh for periodic task archiving - Dashboard: rename 'Background' tab -> 'Agent', 'Cleanup' agent -> 'Completed Task Archiver', remove redundant group headers and pill badges, dim inactive agent placeholders - .rules.md: add Self-Documenting UI Names rule - New tests: test_cleanup_done.py, expanded test_app.py and test_task.py
1.4 KiB
1.4 KiB
Doc Review: fix-verdict-parsing
Summary
Documentation review of the code changes for verdict parsing and state machine alignment.
Documentation Plan Compliance
- N/A — No DESIGN.md existed for this task (it went straight from SPEC to implementation).
Documentation Completeness
automaton/dashboard/core/task.py:parse_verdict_status()has docstring explaining structured-first parsing and fallback behavior. ✓tests/test_task.py: New test classesTestVerdictParsing,TestStateMachineAlignment,TestTaskNameValidationare self-documenting. ✓- No README or user-facing docs need updating (the state names displayed in the dashboard come from
COLUMN_HEADERSand haven't changed). ✓
Documentation Accuracy
CHANGELOG.md: Needs an entry under[unreleased]. ✓ (to be added)automaton/dashboard/core/task.pydocstring forparse_verdict_statusaccurately describes the structured-vs-fallback behavior. ✓
Issues Found
Issue 1: Verdict format not documented in referee prompt
- Severity: Medium
- Description: The referee prompt (
prompts/referee.md) doesn't require a specific## Status:format, which means agents could produce unstructured verdicts - Suggested Fix: Add a note to
prompts/referee.mdrequiring the## Status: PASS|FAIL|NEEDS_REVIEWformat. This is R6 in the SPEC.
Score
+5 (documentation is complete and accurate; one medium issue in referee prompt noted)