White paper WP-013
AI Safety at the Infrastructure Layer: The Cryptographic Distress Guardrail
Respond to crisis signals without building a database of the worst things users say.
At a glance
- Paper
- WP-013
- Topic
- AI Safety at the Infrastructure Layer: The Cryptographic Distress Guardrail
- Format
- PDF + web summary
- Signatures
- ML-DSA-65 post-quantum (NIST FIPS 204)
- Sandbox
- Reproducible at affix-io.com/sandbox
- Company
- AffixIO, Wales, UK
AI Safety at the Infrastructure Layer: The Cryptographic Distress Guardrail is an AffixIO technical paper. Respond to crisis signals without building a database of the worst things users say.
Safety tooling usually means storing the conversations you most need to protect. AffixIO detects distress tiers at the policy enforcement layer, returns human-reviewed crisis resources, and logs only hashed safety events anchored in Merkle proofs.
Summary
Safety tooling usually means storing the conversations you most need to protect. AffixIO detects distress tiers at the policy enforcement layer, returns human-reviewed crisis resources, and logs only hashed safety events anchored in Merkle proofs.
Download the full PDF for technical detail, diagrams, and reproduction steps. Public sandbox: affix-io.com/sandbox.
Related reading
Frequently asked questions
Can you detect distress without storing messages?
Pattern tiers trigger at the infrastructure layer; only hashed identifiers and tier classifications enter the audit record.
How does this relate to the Online Safety Act?
Age assurance and harm mitigation duties can be met with privacy-preserving eligibility and safety proofs rather than full content retention.
What happens on Level C events?
Escalation alerts fire with hashed session context for human review within 15 minutes, without exposing message plaintext.