Off-support barrier: semantic safety constraints are not learning-problem invariants and implications for design, containment, and verification
Read the original at arxiv.org→arXiv:2608.11243v1 Announce Type: new Abstract: We argue that a single structural fact organizes a wide range of phenomena in contemporary AI safety: a semantic safety constraint (e.g., the agent does not escape its...
Original headline: "The Off-Support Barrier: Why Semantic Safety Constraints Are Not Learning-Problem Invariants, and What Follows for Prior Design, Containment, and Verification"
Coverage timeline
- Aug 13, 04:00 UTC arXiv cs.AI lead source The Off-Support Barrier: Why Semantic Safety Constraints Are Not Learning-Problem Invariants, and What Follows for Prior Design, Containment, and Verification