MosaicLeaks Exposes Privacy Leakage in Research Agents — and a Training-Based Fix
Published 2026-06-18Ingested 2026-06-19Agentic AIMedium
Summary
ServiceNow Research published MosaicLeaks, a benchmark and training method exposing a privacy failure mode in research agents that perform multi-hop reasoning across private enterprise documents and the public web. The core risk is the "mosaic effect": individually innocuous outbound web queries, observed together by an adversary monitoring the agent's traffic, can reassemble private enterprise facts the adversary never directly saw. The work defines three leakage levels — intent leakage (inferr
Alignment: New signal not yet covered
Related Positions: AI Governance and Risk, Agentic Workflows
agent-securityprivacyresearch-agentsmosaic-effectagent-governanceservicenowreinforcement-trainingdata-leakage