research-document RFR-007
Risk-triggered human escalation study
Risk-triggered human escalation study
Opportunity: Compare fixed approval gates, risk-triggered escalation, and end-only review.
Unknowns: Best intervention timing, reviewer burden, missed-risk rate, autonomy loss, and trust calibration.
Origin and evidence: AI-Knowledge-Gap-Analysis.md human-agent collaboration row; AI-Research-Roadmap.md Phase 5; 12-research-and-engineering-roadmap.md Workstream D; 10-threat-model.md.
Dependencies: RFR-004 and RFR-008; ethically bounded tasks and trained reviewers.
Method: Within-task policy comparison with explicit risk strata; measure verified outcomes, preventions, false alarms, review minutes, and recovery.
Outputs and success: Escalation policy, burden/risk frontier, failure analysis. Success requires lower consequential-risk exposure without unacceptable verified-outcome loss.
Recommended agent: Human-factors researcher with safety reviewer. Effort: large. Expected gain: evidence for controlled delegation.