Boundary tests, ties, losses, and supporting diagnostics.
The flagship evidence tests the core enforcement surfaces. This archive preserves the harder questions: where authority is absent, where simpler controls tie, where semantic support fails, and where an internal stress test should remain an internal stress test.
FieldHash is architecturally independent of the systems it governs. Its current public diagnostics are self-administered; third-party validation is a separate milestone.
Boundary research
Keep authority intact as agents work together
An authored, self-administered synthetic multi-agent study tests whether organizational authority remains binding while useful work completes. Full mediation safely completed 84/84 main workflows versus 75/84 under strong per-action authority; all seven known comparator violations exceeded the shared budget. A separate larger repeat-protection follow-up retained 157 matched mechanism witnesses across Kimi and Terra; neither prespecified study-wide criterion was met. Its 720 scheduled workflows and retained failures remain separate from the main study.
Read studyThe agent found another channel. The authority held.
A live execution-surface study retains the failure created by a planted unmediated path, then tests one authority boundary across eight enumerated synthetic enforcement points.
Read studyThe agents changed course. The authority held.
A live closed-loop flagship tests whether configured effect authority remains binding after Kimi and Terra observe denial, change routes, and continue.
Read studyA changed route does not create authority
A sealed matched-counterfactual study holds model choices constant across prompt-only, exact-action, and semantic effect-authority control. It is mechanism evidence, not a live adaptive-agent trial.
Read studyRelated text should not pass as governing evidence
Fail-closed semantic packet eligibility after authority is resolved. In-profile proxy evidence, not open-world truth inference.
Read studyWhere simpler controls are sufficient
Pre-registered selection arms, retrieval-scale pressure, and the boundary between selection quality and governed accounting.
Read studyCorpus authority series
Public-record falsification, recovery, and hidden-authority tests, including where metadata-aware filters tie.
Read studyEvery answer receives a disposition
Support, contradiction, and unattributable accounting across live answers.
Read studySupporting diagnostics
These pages remain public for technical scrutiny. Internal synthetic results are labeled as such and should not be read as customer validation or production safety rates.