Францын хиймэл оюун ухааны лаборатори Mistral AI нь “Le Chonk” хэмээн нэрлэгдсэн Mistral Large 4 загвараа танилцуулж, салбарын өрсөлдөөнд томоохон байр суурь эзлэхийг зорьж байна.
Mistral AI-ийн шинэ загвар болох Mistral Large 4 (ML4) нь нэг их наяд параметр бүхий том хэмжээний мультимодаль загвар юм. Одоогийн байдлаар уг загварыг олон нийтийн хандалтын цэгээр дамжуулан ашиглах боломжтой байгаа бөгөөд аюулгүй байдлын шалгалтын дараа гурван долоо хоногийн дотор нээлттэй жингийн (open-weight) хувилбарыг гаргахаар төлөвлөж байна. Тус компани энэхүү загвараа Америк болон Хятадын ижил төрлийн бүтээгдэхүүнүүдтэй өрсөлдөхүйц, хиймэл оюун ухааны салбарт “гуравдагч зам” болно хэмээн тодорхойлжээ.
ML4-ийг бүтээхдээ 4,000 ширхэг NVIDIA GPU ашигласан нь өрсөлдөгч компаниудтай харьцуулахад хамаагүй бага тооцоолох хүчин чадал зарцуулсныг илтгэж байна. Mistral AI-ийн Шинжлэх ухааны дэд ерөнхийлөгч Пьер Стокын мэдээлснээр, тус загвар нь кибер аюулгүй байдал, санхүү, чипний дизайн зэрэг салбарт илүү үр дүнтэй ажиллахаар оновчлогдсон байна.
Энэхүү шинэ загвар нь ASML болон Samsung зэрэг стратегийн түншүүдийн сонирхлыг татаж байгаа бөгөөд саяхан Samsung-ийн тэргүүлсэн санхүүжилтийн дараа тус стартапын үнэлгээ 21 тэрбум евро буюу ойролцоогоор 24.39 тэрбум ам.долларт хүрсэн юм. Mistral AI нь өөрийн загваруудыг аюулгүй ашиглах нөхцөлийг бүрдүүлэх зорилгоор засгийн газрууд болон итгэмжлэгдсэн түншүүдтэйгээ хамтран ажиллахаа мэдэгдлээ.
Дэлгэрэнгүйг эх сурвалжаас харах
↓Эх сурвалжийг нээх ↓
The race between open and closed AI is on, and Europe is still in the mix. On Tuesday, French AI lab Mistral AI released Mistral Large 4 (ML4), a new large multimodal model aiming to leapfrog both American and Chinese rivals — following what French president Macron described as “a third way in AI.”
Amid a growing divide between closed models that can be unplugged and open models that are often made in China, Mistral is positioning ML4 as an alternative to both. Nicknamed Le Chonk in reference to its 1 trillion parameters, ML4 is definitely not small; but is not an open-weight model yet. For the time being, it can only be accessed via a public guardrail endpoint, but Mistral plans to make its weights available in just three weeks, after safety testing is complete.
“In the meantime, we’ll work with trusted partners and governments to make sure that the open source weights can be used to defend, but not to [perform] malicious attacks,” Mistral VP Science Pierre Stock told TechCrunch.
Security concerns have been mounting in recent months, particularly among Mistral’s core audience — enterprises and institutions. At the same time, an open-weight model is easier to audit, Stock said.
Another important behind-the-scenes aspect is that ML4 was trained entirely on Mistral’s compute; using only 4,000 Nvidia GPUs “which is two to three times less than our Chinese competitors, and significantly less than the closed source competitors,” Stock said.
With benchmark results still pending, Mistral hopes ML4 will be best in class among open-weight models, especially outside of China, but not only, Stock said. Thanks to focused training, it could also outperform closed models in specific areas that are key to its customers, and where multimodal capabilities can add value.
According to Stock, ML4’s optimized use cases include cybersecurity and finance, but also chip design, which is core to two of Mistral’s main backers — Dutch giant ASML, which led its Series C, and Samsung, which led its Series D last month at a €21 billion valuation (about $24.39 billion).
At the time, the company tried to convey that its decision to host Chinese models wasn’t a pivot into becoming a mere inference provider. With Le Chonk in its corner, the company believes it should still be considered a frontier lab as well.
When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

