Claude-ийн үл үзэгдэх усан тэмдгийг тойрч гарах шийдлийг хөгжүүлэгчид шуурхай нэвтрүүлжээ

Published:

Энэхүү мэдээ, нийтлэлийг хиймэл оюун боловсруулав.

Хиймэл оюун ухаанаар бүтээсэн контентыг таних зорилготой шинэ технологийг илрүүлснээс хойш дөрвөн цагийн дотор хөгжүүлэгчид түүнийг идэвхгүй болгох аргыг боловсруулсан байна.

Anthropic компани Европын Холбооны хиймэл оюун ухааны тухай хуулийн (AI Act) шаардлагыг хангахын тулд Claude загваруудад машин унших боломжтой, үл үзэгдэх усан тэмдэг нэмэхээ зарласан юм. Хөгжүүлэгч Гийом Мейер энэхүү тэмдэглэгээг арилгах кодыг GitHub платформд байршуулсан нь богино хугацаанд олны анхаарлыг татаж, 20,000 гаруй хандалт авчээ. Одоогийн байдлаар 100 гаруй хөгжүүлэгч уг кодыг өөрсдийн төсөлд нэвтрүүлээд байна.

Уг усан тэмдэглэгээ нь Google-ийн хөгжүүлсэн SynthID технологид суурилсан бөгөөд хиймэл оюун ухааны сонгосон үг, хэллэгийн хэв маягт дүн шинжилгээ хийх замаар ажилладаг. Мейерийн санал болгосон арга нь усан тэмдэг ашигладаггүй өөр том хэлний загваруудыг ашиглан текстийг ижил утгаар дахин найруулах, үгийн дарааллыг өөрчлөх зарчмаар ажилладаг байна. Энэхүү шийдлийг Silicon Valley-д төвтэй Haimaker зэрэг стартапууд өөрийн платформдоо нэгтгэж эхэлжээ.

Мейер болон бусад хөгжүүлэгчид хиймэл оюун ухааныг ашиглан бага зэрэг засвар хийсэн ч бүх контентыг “AI-аар бүтээгдсэн” гэж шошголох нь эрсдэлтэй гэж үзэж байна. Тухайлбал, энэ нь ажлын байрны сонгон шалгаруулалт эсвэл судалгааны ажилд буруу ойлголт төрүүлж, хүний бүтээлч байдлыг үндэслэлгүйгээр үгүйсгэх сөрөг нөлөөтэй гэж тэд болгоомжилж байгаа юм. Шинэ дүрэм журмын дагуу компаниуд хиймэл оюун ухаанаар үүсгэсэн контентыг заавал тэмдэглэх үүрэгтэй ч, үүнийг тойрч гарах бие даасан хэрэгсэл ашиглахыг хориглосон хуулийн заалт одоогоор байхгүй байна.

Дэлгэрэнгүйг эх сурвалжаас харах

↓Эх сурвалжийг нээх ↓

Within four hours of Anthropic confirming that Claude models would globally embed invisible, machine-readable watermarks into any AI-generated content, developer Guillaume Meyer had published his override.

His code to remove watermarks from Claude-generated text has since gone viral on GitHub, has been bookmarked more than 20,000 times on X, and has drawn more than 100 contributors, with many more incorporating the technology into their own projects. “Anthropic is embedding watermarks in its Claude texts … the issue is practically history just one day later,” wrote one AI specialist, accompanied by an image of Meyer breaking out of chains and standing on crumpled EU and Anthropic flags.

Meyer and others started investigating how watermarking works after Anthropic announced last week that Claude would adopt it in order to comply with the European Union’s AI Act.

Some are trying to evade the watermarking because they disagree with the idea that all AI-generated content should be labeled as such, Meyer told WIRED, while others, including himself, say they simply relish the technical challenge. Freelance content writers and social media creators have also contacted Meyer asking for assistance using the code, he says.

The new rules, which came in earlier this month, stipulate that model providers like Anthropic and OpenAI must label synthetic audio, image, video, or text so that this material can be detected by a machine as AI-generated—or face fines of up to 3 percent of annual turnover. While the rules say providers cannot market circumvention tools, there is no legal restriction on independent tools.

“I’m not against transparency, and I’m all for content attribution,” says Meyer. “I just think watermarking in itself is a really bad solution, because it has major drawbacks and risks.” He is concerned about the risk of false positives and that the watermarking might not distinguish between light or heavy AI use, especially since, as a native French speaker, he often uses Claude and other AI tools like Grammarly to edit his writing. Using the watermark as evidence–when even Anthropic admits it can only generate a probability that the text has been touched by Claude–could lead to employers unfairly rejecting candidates or overblown accusations of researchers using artificial intelligence just because the detector flags it, he says.

Anthropic watermarks text invisibly by leaving a pattern in Claude’s choice of words and phrases that is indiscernible to a human reader but would be detectable by a machine that knows how to look for it. Because this influences Claude’s output, some users are concerned this will degrade the quality of Claude’s responses, though Anthtropic insists this won’t be the case. The technique, called SynthID, was developed by Google, which has been using it to watermark its AI-generated content since 2023. Computer scientist Scott Aaronson proposed a similar method when working at OpenAI but says the firm never deployed it because the company was worried that watermarks would put customers off its product.

Meyer’s removal method uses a non-watermarking large language model to generate multiple rewrites, swapping in synonyms and slightly reorganizing content. Of course, this relies on using other large language models which do not insert watermarks—possibly not a safe bet since 190 organizations—providers OpenAI, Microsoft, and Meta among them—have signed the EU’s transparency code of practice. It remains to be seen how many of these laboratories are going to implement their watermarks, which must be included in all new models released from August and must be integrated into existing models by December.

While there’s no certainty this tool works until Anthropic releases the software it uses to detect a watermark, understanding the basic SynthID-text approach underpinning Claude’s watermarking makes them fairly sure the method works, says Wayne Pan, chief technology and cofounder at Silicon Valley–based sovereign AI startup Haimaker. He incorporated Meyer’s open-source tool into his platform because he similarly disliked the idea of Claude watermarking content even when it’s only been lightly edited and disagreed with the watermark being invisible to the user.

- Зар сурталчилгаа -

Та юу гэж бодож байна?

Сэтгэгдлээ оруулна уу!
Please enter your name here

MFC.mn сайтад сэтгэгдэл оруулахад анхаарах зүйлс

Холбоотой

spot_img

Шинэ

spot_img