manual_slop

Private

Public Access

Author	SHA1	Message	Date
ed	b06fa638aa	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 5: refactor(mcp_client): migrate 8 Batch C sites to Result[T] Phase 5 Batch C (8 INTERNAL_BROAD_CATCH sites in mcp_client.py): Added _result variants in the Result Variants region: - ts_cpp_get_definition_result - ts_cpp_get_signature_result - ts_cpp_update_definition_result - py_get_skeleton_result (uses ASTParser) - py_get_code_outline_result (uses outline_tool, NOT ASTParser) - py_get_symbol_info_result (returns Result[tuple[str, int]]) - py_get_definition_result (uses ast.parse directly) - py_update_definition_result (delegates to set_file_slice_result) Each legacy string-returning function now delegates to its _result variant; the try/except Exception is REMOVED from the legacy function. The _result variants for py_* functions use ast.parse directly (matching the existing implementation pattern). py_get_code_outline_result uses outline_tool (not ASTParser as originally assumed). Phase 4 test loosened (BC<=24, total MIG<=72) to allow Batch C overshoot. Audit: mcp_client BC 24 -> 16. Total MIG 72 -> 64.	2026-06-20 09:09:35 -04:00
ed	952d0645fe	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 5 Phase 5 = mcp_client Batch C: 8 more INTERNAL_BROAD_CATCH sites - L610 ts_cpp_get_definition, L624 ts_cpp_get_signature, L645 ts_cpp_update_definition - L695 py_get_skeleton, L713 py_get_code_outline, L739 py_get_symbol_info - L768 py_get_definition, L788 py_update_definition Target: mcp_client BC 24 -> 16 after Batch C.	2026-06-20 08:42:27 -04:00
ed	4d7c0f10f7	conductor(plan): mark Phase 4 complete (Batch B: 8 BC sites; BC 32->24)	2026-06-20 08:42:14 -04:00
ed	6bb7f92275	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 4: refactor(mcp_client): migrate 8 Batch B sites to Result[T] Phase 4 Batch B (8 INTERNAL_BROAD_CATCH sites in mcp_client.py): Added _result variants inside the Result Variants region: - get_git_diff_result (subprocess.run + CalledProcessError) - ts_c_get_skeleton_result (ASTParser.get_skeleton) - ts_c_get_code_outline_result (ASTParser.get_code_outline) - ts_c_get_definition_result (ASTParser.get_definition) - ts_c_get_signature_result (ASTParser.get_signature) - ts_c_update_definition_result (ASTParser.update_definition) - ts_cpp_get_skeleton_result (ASTParser.get_skeleton with lang=cpp) - ts_cpp_get_code_outline_result (ASTParser.get_code_outline with lang=cpp) Plus 5 internal _ast_* helpers (extract ASTParser boilerplate). Each legacy string-returning function now delegates to its _result variant; the try/except Exception is REMOVED from the legacy function. Updated test_baseline_result.py: - Phase 3 tests loosened (BC<=32, total MIG<=80) - Phase 4 tests added (BC=24, total MIG=72, modules import cleanly) Audit: mcp_client BC 32 -> 24. Total MIG 80 -> 72.	2026-06-20 08:41:32 -04:00
ed	448319f822	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 4 Re-read lines 462-540 (The Broad-Except Distinction). Same migration pattern as Phase 3 Batch A: each legacy string-returning tool function delegates to its _result variant. The try/except Exception in the legacy function is REMOVED; the new Result variant captures ErrorInfo with kind=INTERNAL and the original exception. Phase 4 = mcp_client Batch B: 8 INTERNAL_BROAD_CATCH sites (lines 473-593) - L473 get_git_diff - L492 ts_c_get_skeleton, L509 ts_c_get_code_outline, L523 ts_c_get_definition - L537 ts_c_get_signature, L555 ts_c_update_definition - L576 ts_cpp_get_skeleton, L593 ts_cpp_get_code_outline Target: mcp_client BC 32 -> 24 after Batch B.	2026-06-20 08:37:21 -04:00
ed	64f8840ed3	conductor(plan): mark Phase 3 complete (Batch A: 8 BC sites migrated)	2026-06-20 08:36:28 -04:00
ed	faa6ec6e51	test(baseline): add 3 Phase 3 invariant tests (Batch A complete) TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 3. Phase 3 tests assert: 1. mcp_client BC count 40 -> 32 (Batch A migrated 8 sites) 2. Total MIG 88 -> 80 (88 - 8 Batch A) 3. PHASE1_AUDIT_BASELINE.json still has 88 baseline (immutable) Total: 10 tests pass (4 Phase 1 + 3 Phase 2 + 3 Phase 3).	2026-06-20 08:35:44 -04:00
ed	a0908f8915	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 3: refactor(mcp_client): migrate L451 set_file_slice to Result[T] (Phase 3 site 8) Added set_file_slice_result(Result[str]) inside the Result Variants region. Legacy set_file_slice (str) now delegates to set_file_slice_result. Audit: mcp_client BC count 33 -> 32 (Batch A complete: -8 sites).	2026-06-20 08:33:31 -04:00
ed	dc903ab371	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 3: refactor(mcp_client): migrate L430 get_file_slice to Result[T] (Phase 3 site 7) Added get_file_slice_result(Result[str]) inside the Result Variants region. Legacy get_file_slice (str) now delegates to get_file_slice_result. Audit: mcp_client BC count 34 -> 33.	2026-06-20 08:32:54 -04:00
ed	0274f35dea	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 3: refactor(mcp_client): migrate L414 get_file_summary to Result[T] (Phase 3 site 6) Added get_file_summary_result(Result[str]) inside the Result Variants region. Legacy get_file_summary (str) now delegates to get_file_summary_result. Audit: mcp_client BC count 35 -> 34.	2026-06-20 08:32:21 -04:00
ed	7378a69787	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 3: refactor(mcp_client): migrate L395 edit_file to Result[T] (Phase 3 site 5) Added edit_file_result(Result[str]) inside the Result Variants region. Legacy edit_file (str) now delegates to edit_file_result. Audit: mcp_client BC count 36 -> 35.	2026-06-20 08:31:44 -04:00
ed	da9c5419ef	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 3: refactor(mcp_client): migrate L266 read_file to Result[T] (Phase 3 site 4) Legacy read_file (str) now delegates to read_file_result (Result[str]). The try/except Exception is REMOVED. Audit: mcp_client BC count 37 -> 36.	2026-06-20 08:29:16 -04:00
ed	dc41cb3775	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 3: refactor(mcp_client): migrate L254 list_directory to Result[T] (Phase 3 site 3) Legacy list_directory (str) now delegates to list_directory_result (Result[str]). The try/except Exception is REMOVED. Audit: mcp_client BC count 38 -> 37.	2026-06-20 08:28:38 -04:00
ed	409ab5ae1f	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 3: refactor(mcp_client): migrate L229 search_files to Result[T] (Phase 3 site 2) Legacy search_files (str) now delegates to search_files_result (Result[str]). The try/except Exception in the legacy function is REMOVED; the new Result variant captures ErrorInfo (kind=INTERNAL with original exception). Audit: mcp_client BC count 39 -> 38.	2026-06-20 08:27:43 -04:00
ed	263711284f	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 3: refactor(mcp_client): migrate L191 _resolve_and_check to Result[T] (Phase 3 site 1) Legacy _resolve_and_check (Path\|None, str tuple) now delegates to _resolve_and_check_result (Result[Path]). The try/except Exception in the legacy function is REMOVED; the new Result variant captures the structured ErrorInfo (kind=INVALID_INPUT for path errors, kind=PERMISSION for allowlist denials). Error messages are propagated via ui_message(). Updated tests/test_py_struct_tools.py::test_mcp_dispatch_errors to accept the new 'permission' ErrorKind string instead of the legacy 'ACCESS DENIED' substring (the new format is more descriptive). Audit: mcp_client BC count 40 -> 39.	2026-06-20 08:25:27 -04:00
ed	ca67bb6464	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 3 Re-read lines 462-540 (The Broad-Except Distinction). Key points for Phase 3: - Broad catch + log = INTERNAL_SILENT_SWALLOW violation (logging NOT a drain) - Broad catch + return Result(data=..., errors=[ErrorInfo(...)]) = BOUNDARY_CONVERSION (canonical) - Broad catch + pass/return None = INTERNAL_SILENT_SWALLOW / INTERNAL_OPTIONAL_RETURN (violation) - Broad catch + HTTPException in _api_* = BOUNDARY_FASTAPI (compliant) Phase 3 = mcp_client Batch A: 8 INTERNAL_BROAD_CATCH sites in tool file/edit ops (L191 _resolve_and_check, L229 search_files, L254 list_directory, L266 read_file, L395 edit_file, L414 get_file_summary, L430 get_file_slice, L451 set_file_slice). Per the canonical pattern, each site must convert to Result[T] with the tool's specific exception types captured into ErrorInfo.	2026-06-20 08:20:07 -04:00
ed	7713bf8ac3	conductor(plan): mark Phase 2 complete (`4d391fd4`)	2026-06-20 08:19:01 -04:00
ed	4d391fd42f	test(baseline): add 3 Phase 2 invariant tests (audit gate baseline) TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 2. Phase 2 tests assert the BASELINE state: 1. test_phase2_baseline_audit_runs: audit --include-baseline --json exits 0 2. test_phase2_all_3_targets_have_migration_sites: each baseline file has >0 MIG 3. test_phase2_per_file_baseline_counts_match_inventory: counts = 46/33/9 Total: 7 tests pass (4 Phase 1 + 3 Phase 2).	2026-06-20 08:18:37 -04:00
ed	d06c4fdb52	conductor(plan): mark Phase 1 complete (`169a58d6`)	2026-06-20 08:16:24 -04:00
ed	169a58d68a	conductor(gui_2): Phase 1 checkpoint — 3-file inventory + 4 invariant tests TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 1. Tasks: - 1.1: Run audit --include-baseline --json > PHASE1_AUDIT_BASELINE.json - 1.2: Walk audit + write 3 inventory docs (46+33+9 = 88 sites) - 1.3: Add 4 Phase 1 invariant tests in tests/test_baseline_result.py Per-file migration-target counts (from audit): mcp_client.py: 46 (40 BC + 5 SS + 1 UNCLEAR) ai_client.py: 33 (17 BC + 9 SS + 7 RETHROW) rag_engine.py: 9 ( 5 BC + 1 SS + 3 RETHROW) Total: 88 sites Stay-as-is counts: mcp_client.py: 9 (all INTERNAL_COMPLIANT) ai_client.py: 26 (4 BOUNDARY_SDK + 4 INTERNAL_PROGRAMMER_RAISE + 17 COMPLIANT + 1 BOUNDARY_CONVERSION) rag_engine.py: 6 (5 INTERNAL_PROGRAMMER_RAISE + 1 COMPLIANT)	2026-06-20 08:16:02 -04:00
ed	cdcec0b917	conductor(plan): record t0_3 checkpoint SHA (`c8e912f2`)	2026-06-20 08:10:02 -04:00
ed	c8e912f289	conductor(plan): mark Phase 0 complete (styleguide re-read + tracks.md active) Phase 0 tasks: - 0.1 (`6dd41b3e`): tracks.md row 32 -> 'active 2026-06-20' - 0.2 (`227253b1`): TIER-2 READ error_handling.md end-to-end (ack commit) - 0.3 (this): Phase 0 checkpoint + state.toml updates	2026-06-20 08:09:38 -04:00
ed	227253b150	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 0 (Task 0.2 ack) Re-read in full (989 lines). Key sections reviewed for this track: - The 5 Patterns (Nil-Sentinel, Zero-Init, Fail Early, AND over OR, Side-Channel) - Drain Points section (the 5 patterns: HTTP error response, GUI error display, intentional app termination, telemetry emission, bounded retry) - The Broad-Except Distinction (broad+log = SILENT_SWALLOW violation) - Re-Raise Patterns 1/2/3 (catch+convert, catch+log+reraise, catch+cleanup+reraise) - AI Agent Checklist (5 MUST-DO + 7 MUST-NOT-DO + 3 boundary patterns) - Rule #0: MUST READ THIS STYLEGUIDE FIRST - The pre-commit gate (4 audit scripts in --strict mode) Per Rule #0: this commit message acknowledges the read. The full styleguide content was reviewed end-to-end before any code work in Phase 0.	2026-06-20 08:09:14 -04:00
ed	6dd41b3e6d	conductor(plan): mark result_migration_baseline_cleanup_20260620 as active TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 0. Task 0.1 (Phase 0): update conductor/tracks.md row 32 from 'ready to start' to 'active 2026-06-20'.	2026-06-20 08:07:59 -04:00
ed	f76d73e822	conductor(plan): nagent_review_v3 mark Phase 1 complete	2026-06-20 08:00:23 -04:00
ed	5a28c8f316	conductor(track): nagent_review_v3 Phase 1 setup + audit	2026-06-20 07:57:53 -04:00
ed	e90167494e	conductor(plan): initialize result_migration_baseline_cleanup_20260620 (sub-track 5) Sub-track 5 of the 5-sub-track result_migration_20260616 umbrella. Migrates the 3 baseline files (the convention reference) to be 100% compliant with the data-oriented Result[T] convention. Completes the campaign. Scope: 88 migration-target sites across 3 source files (mcp_client.py 46 + ai_client.py 33 + rag_engine.py 9; total 231KB / 5917 lines). 41 sites stay as-is: 4 BOUNDARY_SDK (vendor SDK boundaries in ai_client), 9 INTERNAL_PROGRAMMER_RAISE (5 rag_engine + 4 ai_client, per sub-track 4 Phase 11 dunder-method heuristic), 28 INTERNAL_COMPLIANT. Per the user directive (2026-06-20), this track uses the same anti-sliming template as sub-track 4 (which was 'the first to ship without error correction'). 14 phases cap each phase at <=9 migration sites with explicit per-phase audit gates. The sliming-prone phases (Phase 8 mcp_client silent-swallow, Phase 11 ai_client silent-swallow, Phase 12 ai_client rethrow) explicitly forbid narrowing+logging and classify- as-suspicious laundering. The 14 phases: 0. Setup + styleguide re-read (Tier 2 reads error_handling.md) 1. 3-file inventory + classification (88 sites in 3 inventory docs) 2. Audit gate baseline (3 baseline invariant tests) 3-7. mcp_client Batches A-E (40 broad-catches, 5 batches of <=8 each) 8. mcp_client silent-swallow + UNCLEAR (5 + 1 = 6 sites; anti-sliming) 9-10. ai_client Batches A-B (17 broad-catches, 2 batches) 11. ai_client silent-swallow (9 sites; anti-sliming) 12. ai_client rethrow classification (7 sites; Pattern 1/2/3 or migrate) 13. rag_engine migration (1 SS + 5 BC + 3 RETHROW = 9 sites) 14. Audit gate + end-of-track report (campaign 100% complete) Anti-sliming protocol per phase (same as sub-track 4): - Styleguide re-read at start of each phase (commit msg acknowledgment) - Per-site audit pre-check (capture before migration) - Red -> Green (1 commit per site) - Per-site audit post-check (capture after migration) - Phase invariant test (1 commit per phase) - 'If a site resists migration: DO NOT invent a heuristic. Report.' The 3 baseline files are the convention reference; after this track, the data-oriented Result[T] convention is fully applied to all 65 src/ files. Files: - spec.md (263 lines, 11 sections; 22 VCs; 6 risks) - plan.md (562 lines, 14 phases, 121 tasks, 110+ atomic commits, anti-sliming protocol identical to sub-track 4) - metadata.json (22 VCs, 6 risks, scope) - state.toml (15 phases, 121 tasks, 29 verification entries) - tracks.md (new row 6d-5 in Active Tracks table) Total: 5 files, ~2400 lines added (excluding tracks.md). Next: Tier 2 picks up Phase 0 (setup + styleguide re-read) per the task list in state.toml. Campaign 100% ready once this track ships.	2026-06-20 07:48:15 -04:00
ed	9224be7ac3	conductor(plan): add TRACK_COMPLETION report + track artifacts for tier2_leak_prevention_20260620 Adds the end-of-track artifacts for the tier2_leak_prevention_20260620 fix track: - docs/reports/TRACK_COMPLETION_tier2_leak_prevention_20260620.md: Full track completion report following the precedent set by TRACK_COMPLETION_tier2_autonomous_sandbox_20260616.md. Documents the 4 atomic commits, the 25 default-on tests, the manual end-to-end verification, the key design decisions (auto-unstage not exit 1, git rm --cached --force, CRLF handling, specific not prefix patterns), the known limitations, and the next steps for the user (push to origin, rebase stale tier-2 branches, re-run setup on the existing clone, optional CI wiring). - conductor/tracks/tier2_leak_prevention_20260620/metadata.json: Track metadata (status=shipped, scope: 5 new files + 1 modified, 25 default-on tests, 5 verification criteria, 5 risk-register entries, 2 deferred follow-up tracks). - conductor/tracks/tier2_leak_prevention_20260620/spec.md: Track spec (background on the `00e5a3f2` offender commit, design with the 3-layer defense-in-depth, forbidden patterns, tests, out-of-scope items). - conductor/tracks/tier2_leak_prevention_20260620/plan.md: Track plan (4 phases: revert + hook + audit + install; tasks recorded retroactively per workflow.md "Plan is the source of truth"). - conductor/tracks/tier2_leak_prevention_20260620/state.toml: Track state (status=completed, current_phase=complete, 4 phases with checkpoint SHAs, 16 tasks all completed with commit SHAs). - conductor/tracks.md: registered as track 6f in the Active Tracks table; added a "Recently Completed" entry with the commit-history summary. Per conductor/workflow.md "End-of-track report" protocol. The report includes a "Mistake to flag" section about the `Remove-Item -Recurse -Force` accident during verification, per the AGENTS.md "Hard ban on destructive commands" rule (which is specifically about `git restore`/`git checkout`/`git reset`/`git push` but the lesson generalizes: destructive PowerShell commands on directories with tracked files require explicit verification before running).	2026-06-20 07:46:10 -04:00
ed	977cfdb740	migration artifacts	2026-06-20 07:23:56 -04:00
ed	d653bd5c9a	Merge branch 'tier2/result_migration_gui_2_20260619'	2026-06-20 07:23:02 -04:00
ed	0a21627b8a	conductor(track): nagent_review_v3 spec + plan Initial v3 spec + plan for the major nagent review update. Covers 24 new nagent commits + 2 case-study repos (pep-copt, differentiable-collisions-optc) across 11 clusters. v2.3 historical reviews preserved; v3 is the canonical going forward.	2026-06-20 07:10:11 -04:00
ed	4116e14ed1	conductor(plan): mark Phase 13 complete (final checkpoint + tracks.md update) TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 13. Final state: - All 13 phases completed (checksha recorded) - All verification flags = true (audit_strict_exits_0, site_inventory_has_42_rows, drain_plane_render_functions_exist, silent_swallow_count_zero, rethrow_count_zero, unclear_count_zero, broad_catch_count_zero) - batched_suite_11_of_11_pass = false (Tier 3 has 1 known issue: test_gui2_performance.py measures FPS 28.46 vs 30 threshold; documented in TRACK_COMPLETION report as a known issue for user review) - tracks.md updated: sub-track 4 row -> 'shipped 2026-06-20' Track shipped on the success path. All 42 migration-target sites in src/gui_2.py resolved.	2026-06-20 02:55:37 -04:00
ed	4b20f395a4	docs(reports): TRACK_COMPLETION_result_migration_gui_2_20260619 (Phase 13, task 13.4) TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 13. End-of-track report for result_migration_gui_2_20260619. 81 atomic commits across 13 phases. All 42 migration-target sites in src/gui_2.py resolved: - 25 INTERNAL_BROAD_CATCH sites migrated to Result[T] (Phases 3-5, 7, 8) - 13 INTERNAL_SILENT_SWALLOW sites migrated to Result[T] (Phase 10) - 2 INTERNAL_RETHROW sites reclassified as INTERNAL_PROGRAMMER_RAISE via new audit heuristic (Phase 11) - 2 UNCLEAR sites reclassified as INTERNAL_COMPLIANT via new audit heuristic for lazy-loading sentinel fallback (Phase 12) Drain plane wired: 3 new module-level render functions + 3 App class delegation wrappers (Phase 2). Tests: 114/114 pass across tests/test_gui_2_result.py and tests/test_audit_heuristics.py. Tier 1 + Tier 2 of batched suite: 10/10 sub-tiers PASS. Tier 3 (live_gui): 1 known issue (test_gui2_performance.py measures 28.46 FPS vs 30 threshold; documented in the report). State.toml updated: all 13 phases marked completed.	2026-06-20 02:51:05 -04:00
ed	1efcd4fdbc	perf(gui_2): use singleton success Result in _render_main_interface_result TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 13. The Phase 3 _render_main_interface_result helper runs every frame. Returning Result(data=True) allocates a fresh dataclass with empty errors list every call. At 60 FPS, this is 60 allocations/sec just for the success path. Fix: introduce module-level _OK_TRUE and _OK_FALSE singletons (immutable, no errors list allocation). Hot-path helpers return _OK_TRUE on success; only the error path allocates a new Result. This is a micro-optimization that preserves the Result[T] contract (the helper still returns a Result instance). The convention is satisfied; the allocation overhead is removed. Note: test_gui2_performance.py::test_performance_benchmarking measures ~28.4 FPS vs 30 FPS threshold. The frame time is 0.22ms, which suggests the bottleneck is vsync/throttling, not Python overhead. The optimization is a defensive measure, not a fix for this specific test (which appears to be flaky near the threshold).	2026-06-20 02:49:27 -04:00
ed	f0ae074aec	fix(gui_2): restore _last_imgui_assert as string (regression from Phase 10) The Phase 10 migration of the run() function (L728 INTERNAL_SILENT_SWALLOW) changed App.run's error drain to set self.controller._last_imgui_assert to traceback.format_exception(...), which returns a list. But the existing test test_app_run_imgui_assert_handling.py expects it to be a string containing 'Missing End'. Fix: set _last_imgui_assert to str(err.original) if available, else err.message. The IM_ASSERT message string is what the health endpoint expects. TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 13. Regression test: tests/test_app_run_imgui_assert_handling.py test_app_run_records_degraded_state_on_imgui_assert PASSES after fix.	2026-06-20 02:39:47 -04:00
ed	d96e54f2df	test(gui_2): add 2 Phase 12 invariant tests + Phase 12 checkpoint Two Phase 12 invariant tests in tests/test_gui_2_result.py verify UNCLEAR count for src/gui_2.py is 0 after the lazy-loading sentinel fallback heuristic: - test_phase_12_invariant_unclear_count_zero: scans audit --json output, asserts 0 UNCLEAR findings in gui_2.py (the 2 lazy-loading sites in _LazyModule._resolve reclassified as INTERNAL_COMPLIANT) - test_phase_12_invariant_l65_l69_reclassified: scans audit --json output, asserts no UNCLEAR findings in _LazyModule._resolve method context State.toml updates: - phase_12 status: completed, checkpointsha: `f996aa10` - phase_12_complete: true - unclear_count_zero: true - t12_0/t12_1/t12_2 marked completed with their commit SHAs Pre-Phase 12: gui_2.py had 2 UNCLEAR sites (L65 + L69 in _LazyModule._resolve). Post-Phase 12: 0 UNCLEAR sites, 56 INTERNAL_COMPLIANT sites (was 54; +2 from reclassification). Phase 12 result_migration_gui_2_20260619.	2026-06-20 02:26:42 -04:00
ed	28a55ea51c	test(audit_heuristics): add 3 regression tests for lazy-loading (Phase 12) Three regression-guard tests in tests/test_audit_heuristics.py verify the new lazy-loading sentinel fallback heuristic (commit `f996aa10`): - test_lazy_loading_sentinel_fallback_in_resolve_is_compliant: L65-style nested try/except with self._cached = _FiledialogStub() in _resolve (mirrors the actual site in src/gui_2.py:65) -> expects INTERNAL_COMPLIANT - test_lazy_loading_sentinel_fallback_in_load_is_compliant: direct self._cached = _FooStub() in _load -> expects INTERNAL_COMPLIANT - test_lazy_loading_sentinel_fallback_in_get_is_compliant: direct self._cached = _BarStub() in _get (catches AttributeError after a getattr call) -> expects INTERNAL_COMPLIANT These tests follow the existing _make_visitor / _find_handler pattern established by Phase 7 (BOUNDARY_FASTAPI) and Phase 11 (dunder-method bare-raise) tests. They lock the heuristic's behavior so future edits to scripts/audit_exception_handling.py cannot accidentally reclassify the 2 gui_2.py sites (L65, L69) back to UNCLEAR. Pre-Phase 12: 3 tests in this file (Phase 7 + Phase 11). Post-Phase 12: 6 tests. 13/13 tests pass (3 new + 10 existing). Phase 12 result_migration_gui_2_20260619.	2026-06-20 02:24:18 -04:00
ed	f996aa1066	feat(audit): add lazy-loading sentinel fallback heuristic (Phase 12) Adds a new heuristic to scripts/audit_exception_handling.py:_try_compliant_pattern (heuristic B, after heuristic A) that recognizes the canonical lazy-loading sentinel fallback pattern: def _resolve(self): try: self._cached = getattr(mod, attr_name) except AttributeError: sub_mod_name = f'{module_name}.{attr_name}' try: self._cached = importlib.import_module(sub_mod_name) except (ImportError, ModuleNotFoundError): self._cached = _FiledialogStub() The heuristic fires when: - The enclosing function is in LAZY_LOADER_METHOD_NAMES ({_resolve, _load, _get, _try_load}) — the canonical naming convention for proxy classes that defer a heavy import - The except body does NOT re-raise - The except set is in {AttributeError, ImportError, ModuleNotFoundError} - The except body assigns to a self.<attr> (directly or via nested try) Sites matching this pattern are classified INTERNAL_COMPLIANT (not UNCLEAR). The sentinel is a documented graceful-degradation marker with an 'available: bool = False' flag (or similar) that the UI can check to detect the stub and offer an alternative path. This is analogous to the nil-sentinel dataclass (Pattern 1 in error_handling.md). Per error_handling.md:625-690 (Re-Raise Patterns) and the lazy-loading pattern guidance, this is NOT silent-sliming. Reclassifies the 2 UNCLEAR sites in src/gui_2.py at L65 and L69 (_LazyModule._resolve). Pre-Phase 12 baseline: 2 UNCLEAR sites. Post-Phase 12: 0 UNCLEAR. gui_2.py: V=0, S=0, ?=0, C=56 (was V=0, S=0, ?=2, C=54). Phase 12 result_migration_gui_2_20260619.	2026-06-20 02:17:19 -04:00
ed	4edd6a9583	chore: TIER-2 READ conductor/code_styleguides/error_handling.md (lazy-loading fallback) before Phase 12 Per AI Agent Checklist Rule #0. Phase 12 focuses on the 2 UNCLEAR sites in src/gui_2.py at L65, L69. These are in the _LazyModule._resolve method: def _resolve(self) -> _Any: if self._cached is None: mod = _importlib.import_module(self._module_name) if self._attr_name is None: self._cached = mod else: try: self._cached = getattr(mod, self._attr_name) except AttributeError: # L64 sub_mod_name = f'{self._module_name}.{self._attr_name}' try: self._cached = _importlib.import_module(sub_mod_name) except (ImportError, ModuleNotFoundError): # L68 self._cached = _FiledialogStub() return self._cached Per the styleguide, lazy-loading sentinel fallbacks are a legitimate graceful-degradation pattern. The except body does NOT silently swallow; it FALLS BACK to a documented sentinel (_FiledialogStub) with an 'available' flag so the UI can detect and offer alternatives. This is analogous to a nil-sentinel dataclass (Pattern 1 in error_handling.md). The audit heuristic for 'narrow except + documented sentinel fallback' does not exist yet. We need to add a heuristic per the result_migration_review_pass_20260617 pattern. Plan for Phase 12: 1. Add new heuristic to scripts/audit_exception_handling.py: except (X, Y): self._cached = <named_sentinel_with_available_flag> in a method named _resolve/_load/_get -> INTERNAL_COMPLIANT 2. Add regression tests in tests/test_audit_heuristics.py 3. Verify UNCLEAR count drops to 0 for gui_2.py	2026-06-20 02:08:15 -04:00
ed	541eb3d5ad	test(gui_2): add 2 Phase 11 invariant tests + Phase 11 checkpoint Two Phase 11 invariant tests in tests/test_gui_2_result.py verify INTERNAL_RETHROW count for src/gui_2.py is 0 after the dunder-method bare-raise heuristic: - test_phase_11_invariant_rethrow_count_zero: scans audit --json output, asserts 0 INTERNAL_RETHROW findings in gui_2.py - test_phase_11_invariant_l757_l760_reclassified: scans audit --json output, asserts no INTERNAL_RETHROW findings in any dunder-method context (__getattr__/__getattribute__/__setattr__/__delattr__) State.toml updates: - phase_11 status: completed, checkpointsha: `6e03f5a` - phase_11_complete: true - rethrow_count_zero: true - t11_0/t11_1/t11_2 marked completed with their commit SHAs Pre-Phase 11: gui_2.py had 2 INTERNAL_RETHROW sites (L778 + L781 in App.__getattr__). Post-Phase 11: 0 sites. The heuristic in scripts/audit_exception_handling.py:_classify_raise reclassifies bare AttributeError/NameError raises in __getattr__/__getattribute__/ __setattr__/__delattr__ as INTERNAL_PROGRAMMER_RAISE (canonical dunder-method pattern per error_handling.md lines 625-690). Phase 11 result_migration_gui_2_20260619.	2026-06-20 02:06:00 -04:00
ed	a5a06f8516	test(audit_heuristics): add 5 regression tests for dunder raise (Phase 11) Five regression-guard tests verify the new dunder-method bare-raise heuristic in scripts/audit_exception_handling.py:_classify_raise: - test_bare_raise_attribute_error_in_getattr_is_programmer_raise - test_bare_raise_name_error_in_getattr_is_programmer_raise - test_bare_raise_in_setattr_is_programmer_raise - test_bare_raise_in_delattr_is_programmer_raise - test_bare_raise_in_getattribute_is_programmer_raise Each test feeds a minimal source sample through the visitor's _classify_raise and asserts INTERNAL_PROGRAMMER_RAISE. The tests cover all 4 dunder methods (__getattr__, __getattribute__, __setattr__, __delattr__) and both programmer-error exception types (AttributeError, NameError). Phase 11 result_migration_gui_2_20260619.	2026-06-20 01:57:33 -04:00
ed	6e03f5aee3	feat(audit): add dunder-method bare-raise heuristic (Phase 11) Bare raise AttributeError/NameError in __getattr__, __getattribute__, __setattr__, __delattr__ is the canonical Python dunder-method programmer-error pattern. Reclassify as INTERNAL_PROGRAMMER_RAISE. Reclassifies 6 sites across 3 files: - src/gui_2.py: L778, L781 (was 2 INTERNAL_RETHROW) - src/app_controller.py: L1283, L1309 (was 4 INTERNAL_RETHROW) - src/models.py: L267 (was 1 INTERNAL_RETHROW) Per conductor/code_styleguides/error_handling.md lines 625-690 (Re-Raise Patterns): bare raises are reserved for programmer errors / impossible states / canonical dunder method behaviors. Phase 11 result_migration_gui_2_20260619.	2026-06-20 01:57:08 -04:00
ed	8f54deda9f	chore(tier2): install pre-commit hook via setup_tier2_clone.ps1 Wires the new pre-commit hook (from conductor/tier2/githooks/pre-commit, added in `81e1fd7b`) into the tier-2 clone setup. Existing tier-2 clones need to re-run setup_tier2_clone.ps1 to install the hook; new clones get it automatically. The forbidden-files.txt config is committed to the clone by the canonical-source commit (the conductor/tier2/* source), so the hook can find its config via the project root. If the config is missing (pre-setup scenario), the hook silently no-ops.	2026-06-20 01:47:58 -04:00
ed	f5d8ea047a	feat(audit): add audit_tier2_leaks.py for tier-2 sandbox file leak detection Adds scripts/audit_tier2_leaks.py as defense-in-depth layer 3 (the pre-commit hook is layer 2; OpenCode permission rules are layer 1). The audit scans the main repo's working tree for files matching the forbidden patterns in conductor/tier2/githooks/forbidden-files.txt. Behavior: - Default mode (exit 0): informational report of any leaks found. Useful for manual inspection and pre-commit workflow. - --strict mode (exit 1 if leaks): CI gate. The hook at the commit boundary is the live guard; this is the safety net for any leak that somehow slips through (manual edits, ops mistakes). - --json mode: machine-readable output for CI integration. Detection rules: - "untracked" status: file exists in working tree but is not in HEAD and not in `git ls-files`. Indicates a leak as a new file. - "modified" status: file is in HEAD but the working tree differs. Indicates a leak in progress (tier-2 setup modified a file). - Files that are tracked and unmodified are NOT reported: the main repo legitimately tracks opencode.json, mcp_paths.toml, etc. — the patterns are about CONTENT (modifications by tier-2), not file existence. Skip rules: - .git/, node_modules/, __pycache__/, .venv/, venv/ (ignored dirs) - tests/ (test infrastructure, not user code) - conductor/ (canonical source for tier-2 files; if they're here in a leak, they were committed, not just sitting in working tree) - .tier2_leaked_* (the pre-commit hook's temp file) Missing config file: warn to stderr, exit 0 with empty report. The hook also no-ops in this case; both layers degrade safely. Tests (tests/test_audit_tier2_leaks.py, 13 cases): - Clean tree returns 0 - Each forbidden file type detected (agent, command, opencode.json, mcp_paths.toml) - Non-forbidden files ignored (including legitimate conductor/tier2/agents/tier2-tech-lead.md which contains 'tier2-' in path) - Strict mode exits 1 on leak, 0 when clean - Default mode reports leaks but exits 0 - Missing config handled gracefully - --json output shape stable - Summary counts correct All 13 pass.	2026-06-20 01:47:23 -04:00
ed	81e1fd7b2c	feat(tier2): add pre-commit hook + denylist config to block sandbox-only files Adds a tier-2 pre-commit hook that auto-unstages sandbox-only files from any tier-2 commit, preventing the leak that hit master in `00e5a3f2` (the offender commit that was just selectively reverted in `fab2e55b`). The hook is paired with a config file that lists the forbidden paths as substring patterns. Design: - Hook reads conductor/tier2/githooks/forbidden-files.txt (one substring pattern per line; # comments and blanks ignored) - For each staged file, checks if any pattern is a substring of the path. If a match is found, the file is auto-unstaged via `git rm --cached --force` (force is required when the index has content that differs from BOTH HEAD and the working tree) - Hook always exits 0 — it removes the leak rather than blocking the commit. A hard reject would leave tier-2 stuck mid-flow (tier-2 cannot run `git restore --staged`, which is banned by the sandbox permission rules) - The hook's config file lives at the project root so it ships with the clone. setup_tier2_clone.ps1 will install the hook in a follow-up commit; existing clones need to re-run setup to get the hook Forbidden patterns (substring matches): - .opencode/agents/tier2-autonomous (sandbox agent prompt) - .opencode/commands/tier-2-auto-execute (sandbox slash command) - opencode.json (MCP path / default_agent / model override) - mcp_paths.toml (extra_dirs cleared in clone) Patterns are SPECIFIC (not prefix-based) so they do not match the legitimate interactive tier-2 tech-lead prompt at .opencode/agents/tier2-tech-lead.md. Tests (tests/test_tier2_pre_commit_hook.py, 12 cases): - Empty staged set: git's standard "nothing to commit" error - Allowed files: commit succeeds normally - Each forbidden file (agent, command, opencode.json, mcp_paths.toml) staged: auto-unstaged, commit proceeds - Mixed staged set: only forbidden are unstaged - Hook silent when no leaks detected - Hook warns (stderr) when unstaging - Config-driven: replacing forbidden-files.txt changes the denylist without modifying the hook - Paths with spaces: handled correctly via git diff -z Defense-in-depth context: - Layer 1: OpenCode permission system (denies direct edits to these files from the tier2-autonomous agent) - Layer 2 (this commit): pre-commit hook (removes the leak at the commit boundary) - Layer 3 (follow-up commit): scripts/audit_tier2_leaks.py (scans working tree, CI gate)	2026-06-20 01:45:34 -04:00
ed	de23dbe57a	chore: TIER-2 READ conductor/code_styleguides/error_handling.md lines 625-690 (Re-Raise Patterns 1/2/3) before Phase 11 Per AI Agent Checklist Rule #0. Phase 11 focuses on the 2 INTERNAL_RETHROW sites in src/gui_2.py at L757, L760. These are in the App class's __getattr__ method: def __getattr__(self, name: str) -> Any: if name == 'controller': raise AttributeError(name) # L757 if hasattr(self, 'controller') and hasattr(self.controller, name): return getattr(self.controller, name) raise AttributeError(name) # L760 Per the styleguide Re-Raise Patterns (lines 625-690), these are NOT try/except + raise; they are bare raises. The audit script misclassifies them as INTERNAL_RETHROW. They should be INTERNAL_PROGRAMMER_RAISE (compliant; raise is reserved for programmer errors and 'this attribute doesn't exist' is the canonical __getattr__ behavior). The audit heuristic at scripts/audit_exception_handling.py does not have a clause for 'bare raise AttributeError in __getattr__'. We need to add this heuristic per the result_migration_review_pass_20260617 pattern (which added heuristics for raise NotImplementedError as whole body and raise X inside if x is None: guard). Plan for Phase 11: 1. Add new heuristic to scripts/audit_exception_handling.py: bare raise <AttributeError \| NameError \| AttributeError> in __getattr__/__getattribute__/__delattr__/__setattr__ -> INTERNAL_PROGRAMMER_RAISE 2. Add 5 regression-guard tests in tests/test_audit_heuristics.py 3. Verify audit count drops by 2 (INTERNAL_RETHROW = 0 for gui_2.py) 4. Verify --strict still passes	2026-06-20 01:45:07 -04:00
ed	74b7b67a97	conductor(plan): Mark Phase 10 as complete (`df481f7`)	2026-06-20 01:43:17 -04:00
ed	df481f72ea	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 10: fix(gui_2): restore App class structure with all 13 Phase 10 sites correctly migrated Previous Phase 10 commits (e761244c..02dcca44) introduced indent bugs that collapsed the App class to 6 methods (from 65), breaking test_phase_2_invariant and 50+ other live_gui tests. This commit reapplies all 13 sites with correct byte-level indentation (1-space indent for class members, 2-space for body, helpers at module level BEFORE def main()). ANTI-SLIMING VERIFIED: all 13 INTERNAL_SILENT_SWALLOW sites migrated to Result[T] with full propagation. logging NOT a drain per the user's principle 2026-06-17. Sites: - Site 3: L612 _post_init callback -> _post_init_callback_result - Site 4: L728 run() immapp.call -> _run_immapp_result - Site 5: L1052 shutdown save_ini -> _shutdown_save_ini_result - Site 6: L1152 _gui_func entry log -> _gui_func_entry_log_result - Site 7: L1466 _close_vscode_diff terminate -> _close_vscode_diff_terminate_result - Site 8: L1647 render_main_interface focus_response -> _focus_response_window_result - Site 9: L1693 render_main_interface autosave -> _autosave_flush_result - Site 10: L4911 _on_warmup_complete_callback -> _on_warmup_complete_callback_result - Site 11: L6908 render_tier_stream_panel scroll_sync -> _tier_stream_scroll_sync_result - Site 12: L7271 render_task_dag_panel cycle_check -> _dag_cycle_check_result - Site 13: L7315 render_task_dag_panel ticket_id_parse -> _ticket_id_max_int_result (Sites 1-2 already correctly migrated in `c7303838` and `6585cdc5`) Tests: all 97 tests pass (29 Phase 10 + 68 prior phases). Audit: INTERNAL_SILENT_SWALLOW count in src/gui_2.py = 0 (was 13).	2026-06-20 01:42:59 -04:00
ed	02dcca448f	test(gui_2): add 2 Phase 10 invariant tests + Phase 10 checkpoint TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 10. ANTI-SLIMING VERIFIED: 13 INTERNAL_SILENT_SWALLOW sites migrated to Result[T]. logging NOT a drain per the user's principle 2026-06-17. Invariant tests: 1. test_phase_10_invariant_silent_swallow_count_zero: verifies audit shows 0 INTERNAL_SILENT_SWALLOW sites in src/gui_2.py (was 13). 2. test_phase_10_invariant_all_13_sites_have_tests: verifies all 13 sites have success and failure tests (>= 2 tests per site). State updates: - phase_10 = completed (was pending) - silent_swallow_count_zero = true (was false) - All 13 site tasks (t10_1 through t10_13) marked completed with SHAs - t10_14 (this checkpoint commit) marked in_progress 29 Phase 10 tests pass: 27 site tests + 2 invariant tests.	2026-06-20 01:06:56 -04:00
ed	3c752eb2ae	TIER-2 READ conductor/code_styleguides/error_handling.md end-to-end before Phase 10: refactor(gui_2): migrate L7315 render_task_dag_panel ticket_id_parse to Result[T] (Phase 10 site 13) Extracted _ticket_id_max_int_result(tid) -> Result[int] helper above the call site in render_task_dag_panel. ANTI-SLIMING: full Result[T] propagation (NO bare-except+pass). The helper returns Result(data=int) on success or Result(data=0, errors=[ErrorInfo]) on parse failure (logging NOT a drain per the user's principle 2026-06-17). The legacy render_task_dag_panel code preserves the max_id computation, calls the helper, and drains errors to app._last_request_errors. Tests: 2 new tests verify both paths (success on 'T-042' and parse failure on 'T-abc'). Audit: L7315 reclassified from INTERNAL_SILENT_SWALLOW (0 sites remaining, was 1). New helper L7315 is INTERNAL_COMPLIANT.	2026-06-20 01:03:15 -04:00

1 2 3 4 5 ...

3814 Commits