Protocol update log

Last updated: September 24, 2026. Current: 0.1-rc.4-candidate.3. MCP 0.2.0rc3; ai-ready Skill/contract 0.2.0-rc.3.

This is a public record of changes, not a validation claim. Earlier versions remain available. No completed external pilot, adoption, endorsement, or performance improvement is claimed.

September 24, 2026 — candidate.3: challenge the pass, not only improve the presentation

A deeper audit compared the candidate with the structure and evidence boundaries of WCAG, NIST AI RMF, human-AI interaction guidelines and the SSDF AI profile. It also reproduced three false positives in candidate.2: an altered acceptance criterion, an unrelated proposed-state destination, and an empty action-boundary inventory could each preserve PROCEED_TO_ENGINEERING.

Why rc.4 changed

The practitioner themes remain requirements-first evaluation, traceability (Hillel Glazer, attribution approved), measured current work, realistic review burden and actual end-user use. This revision deepens those themes through reproducible software findings; it does not claim new practitioner approval or external validation.

There are now four principles and fifteen criteria. Existing IDs are retained; C19–C25 in the manifest map the audit changes to acceptance tests. rc.3 and both earlier candidates stay frozen. This remains a candidate with no recorded external pilot.

Read the deep audit, claims and governance, worksheet, current test record, change-to-test manifest, and release status.

Preserved candidate.2 package.

September 24, 2026 — candidate.2: make the review usable and inspectable

The first candidate corrected the decision boundary but still asked too much of the reader: technical records were more developed than the actual guided experience. Its current-state summary did not require a connected step map; tests were linked to artifacts without explicitly recording the revision tested; preparation effort could remain hidden.

Why rc.4 changed

Feedback theme Concrete change Evidence boundary
Requirements-first, risk-scaled evaluation Requirements and acceptance criteria precede tests; risk routes before the quick review; operational deployment remains separate. Correspondence informed design, not validation.
Hillel Glazer: reference material and form/fit/function traceability Inspectable requirement/artifact/test chains, exact tested revisions, reference material, and expandable details. Attribution approved; no endorsement implied.
SMB current-state and adoption Connected current-workflow map, six plain questions, one next action, visible stops, and preparation/session time. The 15-minute target and usability remain untested.
Review/correction burden Gross and net benefit remain separate; nonpositive benefit needs an explicit investment rationale. Scenario estimates are not measured operating performance.
Actual end-user discussion and use Distinct work/discussion origins; general references do not count; a short voluntary pilot packet captures actual use and comprehension. Outreach and interest are not pilots.

Other correspondents' identities and comments remain private. A separate permission-controlled audit maps the original emails to these public themes.

What is new beyond the first candidate

The criteria structure takes inspiration from WCAG's principles/criteria/techniques hierarchy. It is not a W3C standard, accessibility certification, or proof of universal usability.

What still needs evidence

Actual bounded end-user/advisor use, accessible-use testing, observed completion and preparation time, evidence-gathering burden, and whether the decision/next action is understood. Controlled effectiveness and the proposed visual-fidelity experiment require separate designs. A passing software suite cannot answer these questions.

See the test record, change-to-test manifest, and release status.

September 24, 2026 — first rc.4 candidate

Introduced the current-state gate, six mandatory gates, risk-tier depth, separated review/operating costs, engineering-only decisions, permission-aware feedback, pilot contracts, and synchronized MCP/Skill logic. This was a software-tested candidate, not an empirically validated release.

Preserved first-candidate package.

Historical rc.3

Original protocol and earlier tools remain frozen. The historical DOI identifies rc.3, not any rc.4 candidate. Research Harness v3 is a separate project.

Return to the review guide.