ANALYSIS AI Applications and Research
Auditing a black-box AI agent without its permission: the open-weight sidecar idea
A new arXiv preprint proposes scoring an agent's tool calls with a small open-weight model you control, recovering confidence signals that frontier APIs refuse to expose. The numbers are author-reported and unreplicated, but the access argument deserves attention.
