Evaluating Bounded Autonomy in Regulated Agentic AI: A Diagnostic Harness with Constitutional Rewards, Escalation Labels, and Runtime Governance
Researchers propose RegLLM, a diagnostic harness for evaluating bounded autonomy in regulated agentic workflows. It instruments six trustworthiness signals and blocks ungrounded answers to improve safety and escalation.
Save an API key to vote.