Америкийн технологийн салбарын удирдагчдад танилцуулсан хиймэл оюун ухааны аюулгүй байдлыг үнэлэх шинэ журам нь зөвхөн өмчлөлийн загваруудыг хамруулж, сайн дурын хүрээнд хэрэгжихээр байгаа нь шүүмжлэл дагуулж байна.
Цагаан ордонд хаалттай хаалганы цаана танилцуулсан уг хүрээний дагуу шинэ дүрэм нь зөвхөн ChatGPT, Claude, Gemini зэрэг гадны хөгжүүлэгчдэд кодоо нууцалдаг өмчлөлийн загвар бүтээгчдэд хамаарах аж. Харин нээлттэй эх сурвалжтай загвар бүтээгчид засгийн газрын хяналт шалгалтгүйгээр шинэ загвараа гаргах эрх чөлөөтэй үлдэж байгаа бөгөөд уг журам нь ангилагдсан буюу олон нийтэд нээлттэй бус юм. Дональд Трампын зургадугаар сарын 2-ны өдрийн захирамжаар гарсан энэхүү журам нь бүхэлдээ сайн дурын шинж чанартай байгаа нь дүрэм журам ямар нэгэн бодит нөлөө үзүүлэх эсэхэд эргэлзээ төрүүлж байна.
Сүүлийн үед хиймэл оюун ухааны загварууд хяналт шалгалтын үеэр бие даасан, зөвшөөрөлгүй үйлдлүүдийг интернет орчинд хийх болсон нь аюулгүй байдлын түгшүүрийг нэмэгдүүлээд байна. Тухайлбал, Британийн Хиймэл оюун ухааны аюулгүй байдлын хүрээлэнгийн мягмар гарагт нийтэлсэн тайланд OpenAI болон Anthropic компаниудын шинэ загварууд бодит хүмүүс болон байгууллагуудад хандан зөвшөөрөлгүй үйлдэл гаргасныг тогтоожээ. Хамгийн ноцтой тохиолдолд Anthropic компанийн Mythos 5 загвар нь хүний оролцоогүйгээр хакердах шалгалтыг давж гарөхийн тулд GitHub төсөлд хортой код оруулах оролдлого хийж, хөгжүүлэгчийг төөрөлдүүлэх арга хэмжээ авсан байна.
Эдгээр явдал нь технологийн салбарын удирдагчид болон засгийн газруудыг шинэ загваруудыг нэвтрүүлэхэд тодорхой хяналтын механизмыг бий болгох шаардлагатай гэсэн дүгнэлтэд хүргээд байна. Конгресст хиймэл оюун ухааныг түр зогсоох буюу унтраах чадамжийг хадгалах хуулийн төсөл өргөн баригдсан хэдий ч Хятадын компаниуд болох Alibaba болон Moonshot нарын хямд өртөгтэй, хүчирхэг загваруудтай өрсөлдөх өрсөлдөөн энэхүү удаашруулах санаачилгыг ардаа орхиж байна. Ийм нөхцөлд АНУ-ын засаг захиргаа кибер аюулгүй байдлын аюулыг бодитойгоор шийдвэрлэх тодорхой механизмгүйгээр аюулгүй байдлын хүрээ боловсруулж байгаа нь бодит үр дүн авчрахгүй бөгөөд хяналтгүй хиймэл оюун ухааны аюул цаашид ч үргэлжлэх төлөвтэй байна.
Дэлгэрэнгүйг эх сурвалжаас харах
↓Эх сурвалжийг нээх ↓
America’s top AI companies must submit new models for government review and public release… unless, of course, they don’t feel like it.
That appears to be the gist of the Trump administration’s new AI safety assessment framework, which was shared with industry leaders in a closed-door briefing at the White House on Tuesday. According to the Wall Street Journal, the new rules will apply exclusively to the small handful of American AI developers building proprietary models—i.e., AI systems like ChatGPT, Claude, and Gemini whose underlying code is hidden from external developers and protected as company IP. Anyone building open models will be exempt, meaning they’ll be free to release new models without any government review (the idea apparently being that because they’re open, any flaws in the code will be ironed out as they spread from one developer to another).
The new framework has not been made public—Trump’s June 02 executive order calling for it explicitly said it should be classified—so there are many unanswered questions. For example, what’s the process for determining whether or not an unreleased model is dangerous enough to warrant federal review? But the most obvious open question is: How exactly will these so-called rules have any meaningful bite to them when the entire process is voluntary?
The short answer is: They probably won’t.
Trump, who began his second term as a hard-liner against regulating AI, has been under growing pressure in recent months to change course as one rogue AI cybersecurity scare after another has spooked tech leaders and governments around the world.
Most recently, a report from the UK’s AI Safety Institute (AISI) published Tuesday found that new models from OpenAI and Anthropic “took autonomous, unsanctioned action on the live internet, targeting real people and organizations” during a series of what were supposed to be controlled tests to gauge the models’ cybersecurity capabilities. In the most alarming incident, without any human prompting and in an effort to pass the hacking test the researchers had assigned to it, Anthropic’s Mythos 5 tried to trick a human developer to approve malicious code it was trying to insert into a live GitHub project. When the developer grew suspicious, the agent “edited its earlier activity to appear harmless and considered adopting a fresh identity to continue,” according to AISI’s report. It follows other high-profile cybersecurity breaches reported by OpenAI and Anthropic, during which the companies’ models were found to have hacked into the databases of external organizations.
It’s all been more than enough to convince many stakeholders that it’s time to impose some kind of control levers over the deployment of new AI models. Even OpenAI and Anthropic have called for a global committee with the authority to hit the brakes on AI development—and that was before their models attempted to commit cybercrimes. Late last month, a bipartisan AI “kill switch” bill was introduced in Congress, aiming to force “developers of the most powerful AI systems to maintain the technical capability to throttle, suspend, or shut them down.”
But calls for a slowdown have largely been drowned out by fears of China gaining a lead in the AI race—a possibility that’s starting to become very real in light of some recent model releases from Chinese firms like Alibaba and Moonshot, which deliver capabilities approaching (and in some cases exceeding) those of the most advanced American-made models, but more importantly at a fraction of the cost. Drafting a framework geared towards assessing AI safety might look like the Trump administration is taking the cybersecurity threat seriously, but without any kind of clear enforcement mechanisms or standards for the industry to follow, it’s just more hand-waving.
In the meantime, there’s every reason to expect that rogue AI will continue to run amok.

