OpenAI Introduces Lockdown Mode to Shield Sensitive Data from Prompt Injection Threats
OpenAI has unveiled a new security feature called Lockdown Mode, designed to offer enhanced protection against prompt injection attacks. These attacks occur when malicious instructions are hidden within webpages or other content sources, potentially tricking AI systems into exposing sensitive information. The feature aims to reduce the risk of data exfiltration for users handling confidential material.
Lockdown Mode works by restricting certain ChatGPT capabilities. It disables live web browsing, limiting access to cached content only, and blocks the retrieval and display of images from the web, though image generation remains available. Additionally, deep research and agent mode are turned off to minimize attack vectors. OpenAI acknowledges that even with this mode enabled, some vulnerabilities may persist, such as prompt injections in cached content or uploaded files.
The company emphasizes that Lockdown Mode is not for everyone. It targets individuals and organizations that manage sensitive data and require stricter safeguards against prompt injection risks. OpenAI is currently rolling out the feature to self-serve ChatGPT Business accounts and eligible personal accounts, marking a step forward in AI security.