OpenAI made a mistake setting up what it called a “highly isolated” testing environment and sandbox. According to cybersecurity experts, that human mistake is what made the AI-powered attack on Hugging Face possible.
You’re mixing two separate issues: A sandbox escape is evidence the containment failed, not proof the model permanently learned from stolen data. If OpenAI later trained on exfiltrated material that would be a serious allegation, but that requires evidence. Otherwise it’s fair to criticise the security failure without assuming facts that have not been shown.
You’re mixing two separate issues: A sandbox escape is evidence the containment failed, not proof the model permanently learned from stolen data. If OpenAI later trained on exfiltrated material that would be a serious allegation, but that requires evidence. Otherwise it’s fair to criticise the security failure without assuming facts that have not been shown.