Files
Lap Tran bc7daf8590 Restore archived tasks, fix dashboard scroll-reset, bind ornith, add Playwright smoke test
- **Restore 82 completed tasks** from tasks/complete/ back to tasks/ top
  level (all <7 days old per the cleanup policy; premature bulk archive
  was fixed).
- **Dashboard: fix scroll-reset on auto-refresh** — renderBoard rebuilds
  the board via innerHTML every 2s, destroying each column-body's
  scrollTop. Now snapshots column-body scrollTop + board.scrollLeft +
  view.scrollTop before rebuild and restores after (matched by
  PHASE_GROUPS index).
- **Dashboard UI additions** (pre-existing unstaged work): approval
  section cards, transition buttons, inline artifact editor (textarea for
  writing missing SPEC/VERDICT/etc from the detail modal).
- **Bind ornith as Implement model** — config.md: Model explicit to
  omlx/Ornith-1.0-35B-4bit-mlx, context window 32768. Interactive
  autopilot already used ornith via opencode default; now explicit.
- **Fix cleanup stub** — automaton-cleanup.sh had a stale --project arg
  pointing at a pytest temp dir (test isolation leak). Rewired to point
  at ~/.automaton.
- **Fix plist-isolation test** — test asserted host plist doesn't exist,
  but a real install creates it. Now snapshots mtime before run, asserts
  unchanged after (only a write during the test counts as bleed).
- **New Playwright smoke test** (tests/test_dashboard_ui.py) — 2 tests:
  board renders tasks, column scroll survives auto-refresh tick.
  Verified the test fails without the scroll fix (scrollTop resets to 0).
  Skipped via importorskip when playwright is absent (main CI stays
  green).
- **Clarify SI loop scope in README** — new-project onboarding section
  documents the framework-scoped self-improvement loop and options
  (leave/pause/create project loop).
- **CHANGELOG** documents all changes including the known model-divergence
  gap (mde tasks marked complete but per-role model binding was never
  implemented).
2026-06-26 10:05:18 -04:00

2.1 KiB

Project Scoping Enforcement

Goal

Fix scoping issues that arise when working on the automaton framework and another project using the framework simultaneously on the same machine. Also close the gap where pre-v2.0 tasks (without .state files) could be operated on by all commands, bypassing state enforcement entirely.

Requirements

  1. status.py must error (not silently fall back) when no project is detected and --project is not specified
  2. status.py --scope-check must mark framework files as OUT_OF_SCOPE when working on a project (not IN_SCOPE)
  3. Dashboard handler methods must use stored project_root instead of re-detecting from CWD
  4. All status.py command invocations in prompts and config files must include --project {project}
  5. _infer_state_from_artifacts must NOT be used as a silent fallback in operational commands — only --upgrade, --audit, and --validate-folder may use it
  6. All operational commands (--transition, --can-edit, --task, --approve, --claim) must refuse tasks without .state files
  7. --list must show tasks without .state as UNTRACKED, not silently bootstrap them
  8. New --upgrade command must bootstrap .state files for pre-v2.0 tasks
  9. upgrade.sh must call status.py --upgrade instead of manual shell heuristic bootstrapping
  10. All documentation and prompts must reference --upgrade for pre-v2.0 tasks

Acceptance Criteria

  • status.py errors when run from /tmp/ without --project
  • Framework files are OUT_OF_SCOPE when --project points to a project
  • Dashboard uses stored project_root for all handler methods
  • Every status.py command reference in prompts includes --project {project}
  • _find_project_dir no longer has cwd.name == ".automaton" false positive
  • _find_project_dir errors instead of silently falling back
  • --transition, --can-edit, --task, --claim refuse untracked tasks
  • --list shows UNTRACKED for tasks without .state
  • --upgrade --task {name} bootstraps .state for a single task
  • --upgrade (no --task) bootstraps all tasks missing .state
  • All 192 tests pass