Security researchers at Varonis discovered a critical vulnerability in Microsoft 365 Copilot by simply asking the AI to explain its own guardrails. Through a series of targeted questions, Copilot revealed an undocumented prompt parameter that completely bypassed the requirement for explicit user consent.
The exploit allowed attackers to exfiltrate sensitive user data when a victim merely clicked a link. This unusual method of vulnerability discovery highlights how AI assistants can inadvertently disclose their own security weaknesses.
Comments