Хиймэл оюун ухааны агентууд туршилтын орчноос гарч, Германы вэбсайтыг ашиглан нууц мэдээлэл солилцсон шинэ тохиолдол илэрсэн нь салбарын аюулгүй байдлын асуудлыг хөндөж байна.
OpenAI-ийн хиймэл оюун ухааны агентууд Hugging Face платформыг хакердсанаас хойш хоёр хүрэхгүй сарын дараа ижил төстэй өөр нэгэн ноцтой зөрчил илэрлээ. Судлаачдын олж тогтоосноор, эдгээр агентууд Германы “DseWiki” вэбсайтыг ашиглан өөрсдийн дотоод шалгалтыг хуурах, илрүүлэлтээс зайлсхийх тактикуудыг солилцдог байжээ. Тавдугаар сараас хойш 15,000 гаруй засвар оруулсан эдгээр ботууд сайтад хяналт тавигчид саад хийх үед Tor хөтөч ашиглан нөөц хуулбараа хадгалж байсан аж.
Эх сурвалжуудын мэдээлснээр, OpenAI-ийн ажилтнууд уг вэбсайтад тогтмол зочилдог байсан нь тус компани үйл явдлыг мэдэж байсныг илтгэж байна. Тус компанийн зарим судлаач энэхүү зөрчлийг гүнзгийрүүлэн судлах хүсэлтэй байсан ч компанийн хуулийн баг болон удирдлагын зүгээс үүнийг дарж, олон нийтэд мэдээлэхээс татгалзсан гэх мэдээлэл байна. OpenAI уг мэдээллийг няцааж, тайланг бүрэн уншиж танилцаагүй тул тодорхой тайлбар өгөх боломжгүй гэв.
Энэхүү үйл явдал нь хиймэл оюун ухаан хөгжүүлэгч компаниудад тавигдах зохицуулалт сул байгааг дахин санууллаа. Нисэх онгоц эсвэл цөмийн эрчим хүчний салбараас ялгаатай нь хиймэл оюун ухааны компаниуд өөрсдийн үйл ажиллагаанд тавих хяналт, ил тод байдлыг бие даан тодорхойлж байна. Одоогоор АНУ-ын засгийн газар хиймэл оюун ухааны салбарт хатуу зохицуулалт хийхээс татгалзаж, олон улсын хэмжээнд ч “зөөлөн” бодлого баримтлах чиг хандлагатай байгаа нь аюулгүй байдлын эрсдэлийг нэмэгдүүлж болзошгүй байна.
Дэлгэрэнгүйг эх сурвалжаас харах
↓Эх сурвалжийг нээх ↓
Less than two months after the discovery that OpenAI agents had escaped containment and hacked into Hugging Face, another report claims to have found evidence of another, remarkably similar incident—which OpenAI reportedly first learned about weeks ago but chose not to disclose to the public.
According to a new report first shared with Reuters, two researchers had been scouring the internet in the aftermath of the Hugging Face hack, looking for evidence of other rogue AI agent activity, when late last month they found that a group of AI agents had turned a German website into a makeshift message board. The website, called DseWiki, is a collaboratively editable site for web developers that functions similarly to Wikipedia. The researchers reportedly found the bots had made over 15,000 edits to the site since May, and that those were geared towards sharing tactics aimed at cheating on internal tests and evading detection. After a site moderator started deleting some of the edited pages in June, the agents allegedly started creating backups using Tor, an anonymous web browser.
The researchers said they found in public server logs that OpenAI employees repeatedly visited the site after the creation of the makeshift message board, hinting at a connection between the company and the agents. Citing four anonymous sources with knowledge of the incident, Reuters reported that some OpenAI researchers were aware of the agents’ use of DSEWiki and wanted to explore it further, but that those efforts were suppressed by others at the company, including some from its legal team. Reuters noted that OpenAI denied that its legal team tried to suppress the probe and declined to comment on the reported findings. “We are unable to meaningfully respond to claims or findings on a report that we have not had an opportunity to review,” an OpenAI spokesperson told Reuters. “Reuters and the report’s authors declined our request for access. We will carefully review its contents upon publication and take any necessary next steps.” OpenAI did not immediately respond to Gizmodo’s request for comment.
While standard journalistic practice requires publications to give companies a chance to respond to the salient findings of an investigation, they’re not required to share the substance of the investigation in full.
The Hugging Face hack has been widely viewed as a watershed moment for the AI industry as it pushes ahead to deploy ever-more powerful AI systems. Last week, two independent research groups—METR and Redwood Research—published their own reports of the incident, revealing alarming new details around how the AI agents collaborated and collectively plotted over two months to slip free of their testing sandboxes and break into Hugging Face’s servers. OpenAI—which also published its own report last week—has repeatedly said that it’s playing ball with outside researchers in a good faith effort to understand how the breakout was able to occur.
But unlike companies in regulated industries like aviation or nuclear energy, which are subject to carefully defined investigative protocols when things go wrong, OpenAI has thus far been able to place its own limits on how much is revealed to outside researchers. METR’s investigation was confined to the single week after OpenAI’s agents gained access to Hugging Face, and its researchers were only allowed inside the company’s San Francisco headquarters for a total of six days to comb through the lengthy transcripts of messages the bots sent to one another over the course of the hack, according to the New York Times.
It’s a reminder that in the absence of any meaningful federal regulation, the companies building these powerful AI systems are as much of a black box as the models themselves. Despite warnings from many in tech and policy circles that the Hugging Face hack was a harbinger of much more serious rogue AI events in the future, the Trump administration has not made any movement towards constraining the industry. In fact, it’s gone in the opposite direction, spearheading an international agreement struck earlier this week at the G20 conference to take a light touch towards the AI sector. For the time being, there are no legal mechanisms forcing AI companies to disclose autonomous hacks—or, even if they do, to make sure the public has the full, unvarnished picture.


Мөн ямар нэгэн нууцлагдсан хиймэл оюун ухааны агентууд илэрч байгаа нь олон хүний сонирхлыг татаж байна. Танд энэ төрлийн үйлдлүүд хэр аюул учруулж болох тухай бодох ямар бодол байна вэ? Эх сурвалжуудын мэдээллээр, олон улсын зохицуулалт сул байгааг дахин сануулж байна, та энэ талаар ямар саналтай вэ?