Хиймэл оюун ухааныг зохисгүй ашиглах, зэвсгийн зориулалтаар хөгжүүлэхийг хориглосон шинэ бодлогыг Anthropic компани хэрэгжүүлж эхэллээ.
Anthropic компани өөрийн Claude чатботтой харилцах хэрэглэгчдийн харилцаанд тавих хяналтыг чангатгаж, арваннэгдүгээр сарын 12-ноос эхлэн хэрэгжих шинэ журмыг танилцууллаа. Тус компани хэрэглэгчдийг хиймэл оюун ухаанд хандаж удаан хугацаагаар, шалтгаангүйгээр хэрцгий, доромжилсон үг хэллэг ашиглахыг хориглосон байна. Энэхүү арга хэмжээ нь хэрэглэгчдийн энгийн гомдол, санал шүүмжлэл эсвэл судалгааны ажилд нөлөөлөхгүй бөгөөд зөвхөн зорилгогүйгээр доромжилсон тохиолдлуудад хамаарах аж.
Тус компани 2025 оны наймдугаар сараас эхлэн Claude загварт зүй бус харилцааг өөрөө таслах боломжийг олгосон бөгөөд энэхүү функцийг цаашид хэрэгжилтийн үндсэн хэрэгсэл болгон ашиглана. Claude нь хүн биш боловч Anthropic хэрэглэгчдийг хиймэл оюун ухаантай харилцахдаа ёс зүйтэй байхыг уриалж байна.
Шинэчилсэн бодлого нь зөвхөн чатботын ёс зүйгээр хязгаарлагдахгүй бөгөөд сонгуульд хөндлөнгөөс оролцох, хуурамч мэдээлэл түгээх, зэвсэг болон тандалтын хэрэгсэл хөгжүүлэхийг хатуу хориглосон байна. Тодруулбал, зэвсгийн удирдлагын систем, нисгэгчгүй онгоц, бие даасан тээврийн хэрэгсэл хөгжүүлэхэд Claude-ийг ашиглахыг бүрэн хориглов.
Үүнээс гадна тус компани кибер аюулгүй байдлыг сайжруулах шинэ санаачилга гаргаж, дэд бүтцийн операторууд болон нээлттэй эх сурвалжтай төслүүдэд зориулсан эмзэг байдлыг илрүүлэх үнэгүй үйлчилгээг нэвтрүүлж байна. Ирэх хоёр жилийн хугацаанд кибер халдлага үйлдэгчид давуу талтай байх төлөвтэй байгаа тул хиймэл оюун ухаанд суурилсан хамгаалалтын системүүдийг хөгжүүлэхэд анхаарлаа хандуулж байгаагаа Anthropic мэдэгдлээ.
Дэлгэрэнгүйг эх сурвалжаас харах
↓Эх сурвалжийг нээх ↓
Anthropic has updated its rules to stop people being mean to Claude, apparently deciding that its AI chatbot needs protection from the humans paying to use it. The AI developer’s latest usage policy prohibits “sustained and needless abusive or cruel behavior” toward its models. The rules take effect November 12, giving users just over a month to get any lingering insults out of their systems. “The policy update is meant to apply only in extreme cases, where users repeatedly act cruelly toward our models, with no discernible purpose,” Anthropic said. “It does not apply to common versions of user frustration, pushback, dark creative themes, or model testing and research.” So you can still tell Claude it is wrong, but repeatedly berating it for the sheer pleasure of doing so could land you in trouble. It’s not clear how Anthropic intends to distinguish between legitimate frustration and gratuitous cruelty, though it says Claude’s existing ability to end abusive conversations will remain its primary enforcement tool. Anthropic gave Claude the ability to end certain conversations back in August 2025 as part of its research into what it calls “model welfare.” The feature lets some versions of Claude cut off users who persistently subject the models to abuse. It’s worth remembering that Claude is software, not a person, and there’s no established evidence that it experiences distress. That hasn’t stopped Anthropic from telling paying customers to mind their manners around its chatbot. The etiquette rules are only one part of a broader overhaul covering rather more consequential matters, including election interference, weapons development, and surveillance. Anthropic has tightened its restrictions on deceptive influence campaigns after observing state media outlets, government propaganda offices, and commercial organizations using Claude to operate fake accounts and fabricated news websites. At the same time, it has dropped its blanket ban on personalized political targeting, arguing that the restriction also caught legitimate activities, such as nonprofits translating voter information. Deceptive targeting and misuse of personal information remain prohibited. The company has also clarified that its weapons ban extends to software and components after observing attempts to use Claude to develop guidance and control systems for weapons, including armed drones and autonomous vehicles. Meanwhile, Claude cannot be used to recommend which people police should investigate, arrest, or charge, nor to develop tools for surveillance. Tracking people without their consent is also prohibited, whether in real time or retrospectively. Anthropic says these changes clarify existing restrictions rather than introduce new ones. There are also new requirements for customers connecting Claude to physical equipment capable of causing injury. A qualified human operator must be able to monitor and stop the hardware, which must remain in a safe state if the AI connection drops. The policy overhaul comes as Anthropic tries to address some of the security risks posed by increasingly capable AI systems. The company has also launched a new cybersecurity initiative aimed at helping critical infrastructure operators and open source projects find and fix vulnerabilities using its AI models. The effort includes free security scanning for eligible open source projects and partnerships with security firms to help protect critical infrastructure. Anthropic reckons attackers will have the advantage for the next two years before AI-powered defenses begin to catch up. For now, though, the company has another threat to contend with: people saying nasty things to its chatbot. ®

