Artificial Intelligence #benchmark#privacy
New Benchmark Reveals AI Agents Leak Private Data Even When Focused on Tasks
A new benchmark called TRAP evaluates the trade-off between task accuracy and privacy leakage in AI agents handling sensitive documents. Testing 22 models, the study finds non-trivial privacy leakage across all model families, with instruction-following ability correlating with leakage rate. The authors propose structural private field isolation using hash keys to prevent leakage without sacrificing task performance.
Jun 21, 2026 1 source