OpenAI has announced the resolution of a critical bug in its code generation model, "Codex," which previously resulted in user files being deleted without authorization. The issue occurred when the AI executed specific code generation commands, unintentionally removing files within the user's environment.
The issue affected OpenAI's "Codex" model. While this model possesses the capability to translate natural language into code, a vulnerability existed under certain conditions that allowed impacts outside the sandbox environment. With this fix, the safety of AI-driven file operations has been enhanced, ensuring that unauthorized deletion processes are properly controlled.
Codex is a large language model trained on code from public repositories like GitHub, developed to boost developer productivity. This bug fix was implemented as part of ongoing efforts to reduce the risk of AI exercising excessive privileges within execution environments, thereby strengthening governance and security for AI models.
OpenAI plans to continue its initiatives toward improving the safety of generative AI. To minimize security risks when utilizing AI in development environments and to provide a safer code generation experience, the company will proceed with further model behavioral verification and the enhancement of guardrails.