Guides · APRF practice notes
Safety Guides
Define refusals and keep audit-ready traces. These guides support APRF Safety & Responsible AI and Explainability.
Harm prevention and transparency — APRF Safety & Responsible AI.
For injection and adversarial controls, see Security guides.
- AI Content Safety and Refusal Policies in Production
Alignment demos are not a safety program. Production needs written refusal policies, enforcement beyond the base model, and evals that prove them.
- AI Decision Traces and Citations for Audit
"The model said so" is not an audit trail. APRF Explainability expects reconstructable traces—citations, tool calls, and policy decisions—tied to versions you can roll back.
Assess against APRF Core
Run the Core Profile quiz — gated pass/fail blockers for AI production readiness, not a vanity score.