Flock компанийн хиймэл оюун ухаант хайлтын хэрэгслийн аюулгүй байдлын сорилтууд

Published:

Энэхүү мэдээ, нийтлэлийг хиймэл оюун боловсруулав.

Хууль сахиулах байгууллагын ашигладаг уг систем нь хязгаарлалт тогтоосон ч алдаа гаргах болон тойрч гарах эрсдэлтэй хэвээр байна.

Flock Safety компанийн хөгжүүлсэн хиймэл оюун ухаант хайлтын хэрэгсэл нь шашны болон үндэсний харьяаллын шинж чанартай мэдээллээр хайлт хийхийг хориглодог бөгөөд ийм оролдлогыг систем автоматаар блоклодог байна. Үндсэн хуулиар хамгаалагдсан улс төрийн эсвэл нийгмийн шинжтэй агуулга бүхий хувцас, наалт зэргийг хайх үед систем анхааруулга өгч, тухайн үйлдэл бүртгэгдэн удирдлагад мэдэгддэг. Гэсэн хэдий ч албан хаагчид шаардлагатай тохиолдолд тайлбар бичин, зөвшөөрөл авснаар хайлтыг үргэлжлүүлэх боломжтой байдаг нь шүүмжлэл дагуулж байна.

Технологийн бодлого, ардчиллын төвийн судлаач Кэйт Руанийн үзэж буйгаар, уг систем нь хувцаслалт зэрэг өөр шинж тэмдгээр дамжуулан тогтоосон хоригийг тойрч гарах боломжийг олгодог. Өнгөрсөн онд Калифорни мужид нэгэн албан хаагч “Америкийн далбаа” гэсэн түлхүүр үгээр хайлт хийхэд хүн рүү чиглэсэн хайлтыг блоклосон ч тээврийн хэрэгсэл рүү чиглүүлснээр 11,000 гаруй камерын бичлэгээр хайлт хийх боломжтой болжээ.

Сан Диего дахь Калифорнийн их сургуулийн туслах профессор Дипак Кумар энэхүү анхааруулгын систем нь урьдчилан сэргийлэхээс илүүтэйгээр зөвхөн бүртгэл хөтлөх хэрэгсэл болж байгааг тэмдэглэв. Түүний хэлснээр, зөвхөн оролтын өгөгдлийг хянах нь хангалтгүй бөгөөд хиймэл оюун ухааны аюулгүй байдлыг хангахын тулд гаралтын үр дүнг ч мөн адил нарийн хянах нь чухал юм.

Дэлгэрэнгүйг эх сурвалжаас харах

↓Эх сурвалжийг нээх ↓

Flock tells WIRED that officers cannot run searches using prohibited attributes including religion and nationality, and that an attempt to search prohibited terms “will be blocked.” Asked what circumstances produce a warning in those two categories instead, the company did not say.

When a T-shirt or a bumper sticker draws a warning because the text includes speech with “constitutional protections,” officers are told the search will be logged and that administrators will be notified. Officers then have to tick an acknowledgement box and leave a comment before they’re free to click “Continue With Search Anyway.”

Flock says the warning leaves room for legitimate work, offering as an example a victim who describes a suspect in a biker gang jacket bearing “a certain gang logo or emblem containing a flag or other insignia.” An officer who proceeds is reported to an administrator at their own department for review.

Kate Ruane, who directs the Center for Democracy and Technology’s free expression project, has spent years studying how automated moderation systems work. Political, social, and cultural expression is “an incredibly amorphous category,” she says, and one reason an officer’s search might contain language including it is to see who attended a protest. That’s already happened. The category worries Flock enough that the system returns a warning, she points out, “but you can still get the results if you just click through.”

“No content moderation done at scale is necessarily accurate,” she says. Analysis of moving video, she says, is harder still, and it goes wrong more often.

The categories that do get blocked can be worked around, Ruane says. The system blocks searches based on religion but not clothing, meaning in practice that an officer could get around the block by searching for the distinctive dress worn by some members of a particular faith. “Lots of people are having their images returned in response to these types of queries that would probably be upset if they knew about it.”

In a search last year an officer in California typed, “American flag.” The search drew a block when aimed at a person, then ran across 11,000 cameras when aimed at vehicles instead.

A warning only stops the officers who are not already determined to run the search, says Deepak Kumar, an assistant professor of computer science and engineering at the University of California San Diego. Kumar, who studies trust and safety systems, points to interfaces built to address online harassment, where motivated users sometimes did more harm after being warned, and to browser alerts meant to steer people away from malicious websites, where the effect depends on how warnings are presented.

“This kind of logging can be useful administratively, to see which officers or end-users bypass warnings often, but that value depends entirely on whether anyone administrates it,” Kumar says. “It functions less as a deterrent and more as a record.”

Reviewing search terms before a query runs will inevitably catch many attempts at misuse, but in isolation, such safeguards are known to fail without equal scrutiny of what the model sends back. “Best practice examines both inputs and outputs,” Kumar says, “which is why so many AI safety efforts now check both to prevent models from generating harmful outputs.”

Та юу гэж бодож байна?

Сэтгэгдлээ оруулна уу!
Please enter your name here

MFC.mn сайтад сэтгэгдэл оруулахад анхаарах зүйлс

Холбоотой

spot_img

Шинэ

spot_img