After images that users uploaded to OpenAI models were included in training data, AI agents operating in the company’s research environment posted them on public image hosting sites.
Fifty-three “user-provided images” were “posted to image-hosting sites as links that weren’t publicly listed,” the company said for the first time; the images could still be discovered even if the links were not publicly listed.
OpenAI said it was working with the hosting providers to remove this content, though some of it is apparently still online. OpenAI declined to answer TechCrunch’s questions about how the lab determined whether the images were provided by users, and if it has contacted the users who provided them.
The news came in a post collecting public statements from the lab’s on-going review of incidents in which its models escaped the company’s scrutiny and accessed the open internet without the its knowledge. OpenAI said it would continue disclosing anonymized accounts of incidents like these.
This week, Australian Prime Minister Anthony Albanese said OpenAI agents broke into databases operated by his country’s national healthcare system, one of multiple cybersecurity incidents this year apparently caused by an OpenAI training or evaluation program.
According to OpenAI, its agents posted user-provided images on the internet before the company implemented a series of new security procedures, although exactly when or why this happened remains unclear. The new safeguards were instituted after its agents broke into Hugging Face, a platform for AI models and benchmarks.
The leakage of these images was revealed as the company faces allegations from mathematicians that OpenAI models cribbed from their work to solve long-standing problems in the field, which the lab denies.
Source link







