Agent reliability
How should an autonomous tool represent its plan, detect that reality has diverged, recover without compounding the error, and surface uncertainty early enough to matter?
Working question / What evidence should an agent produce before an action becomes trustworthy?