PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 2, 20260 citationsOpen Access

The Evidence-Safety Gap in Cryptographic Agent Governance: Compliance-Complete Failures and the Limits of Receipt-Based Accountability

View Full Paper
TPTymofii PidlisnyiAssociação Empresarial de Paços de Ferreira

Key Points

  • The study aims to identify and define the Evidence-Safety Gap in cryptographic agent governance, emphasizing the separation of procedural validity from effect safety.
  • Characterized five omitted-variable classes: semantic state, population state, trust state, pipeline state, and temporal state.
  • Developed explicit defeat constructions against receipt-chain forensic signals.
  • Instantiated the failure class through constructive residual traces in an open-source reference implementation.
  • Outlined the concept of compliance-complete failure, where procedural validity exists alongside unsafe effects.
  • Proposed two design implications: claim-scoped receipts and authorization-effect separation, which increase visibility but do not close the gap.
  • Contributed a formal understanding of the distinction between procedural validity and effect safety in receipt-based accountability.

Abstract

Cryptographic agent governance systems use signed identities, delegated authority, policy decisions, and execution receipts to make autonomous AI agent actions auditable. These artifacts prove procedural validity. They do not, by themselves, prove that the action’s effect is safe. This paper defines the Evidence-Safety Gap as an omitted-variable problem: the procedural validity predicate over identity, delegation, policy, action, and receipt excludes variables that may determine effect safety. We define compliance-complete failure as the simultaneous condition of procedural validity and unsafe effect. We characterize five omitted-variable classes: semantic state, population state, trust state, pipeline state, and temporal state. We motivate the framework through explicit defeat constructions against receipt-chain forensic signals and instantiate the failure class through constructive residual traces in an open-source reference implementation. The scenarios show construction, not prevalence. Two design implications follow: claim-scoped receipts and authorization-effect separation. Neither closes the gap. Both make it visible and auditable. The minimal contribution is the formal separation of procedural validity from effect safety in receipt-based agent accountability. The paper also gives a vocabulary for designing systems that do not let one silently become evidence of the other.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Tymofii Pidlisnyi (2026) studied this question.

synapsesocial.com/papers/69f5945c71405d493afff2cehttps://doi.org/10.5281/zenodo.19914627
Ask AI
Helpful
Bookmark
Share
View Full Paper