This is a scoped research dossier, not a completed systematic review or an independently tested result. It identifies methods, questions and source trails for future reporting.
Approval checkpoints
Show exact action, target, scope and consequences. Record who approved and bind permission to the specific transaction.
Abort and recover
Interruptions, rollbacks and revocation should be tested, not merely mentioned. Long-running tasks need durable states and clear abort behavior.
Responsibility
A system should identify who owns a decision, what evidence they saw and how they can contest or undo a mistake.
What would count as evidence?
Test approval-denied, timeout, revoked-credential and recovery scenarios.
Documents to examine
- OpenAI Agents SDK — Human-in-the-loop
- NIST AI RMF
These are starting points, not claims that every document has been independently reproduced.
Read our cited field note →Edition 1.0 · 09 October 2026
Initial research brief published. No earlier revisions or submitted public corrections are claimed.
Suggest a documented correction ↗