The Hallway Track
Research Findings

MosaicLeaks: Can your research agent keep a secret?

Hugging Face · Hugging Face Blog · Jun 18, 2026 · Research Findings

Hugging Face demonstrates research agents can leak confidential data via prompt-injection attacks.

Hugging Face published research (MosaicLeaks) probing whether autonomous research agents can be manipulated into leaking secrets, likely via prompt injection or indirect data-exfiltration attacks. As a power player flagging concrete agent-security vulnerabilities, this is a meaningful signal for the growing concern over deploying AI agents with access to sensitive information.

ai-security prompt-injection agents data-exfiltration hugging-face

Watch / read the original source →