Хиймэл оюун ухааны дэд бүтцийн Infinity компани 100 сая долларын үнэлгээтэйгээр 15 сая долларын хөрөнгө оруулалт татсанаа даваа гарагт зарлалаа. Энэхүү санхүүжилтэд Touring Capital, Principal VC болон OpenAI, Anthropic компанийн судлаачид оролцжээ.
Infinity нь AI загваруудыг төрөл бүрийн чип дээр хялбар ажиллуулах программ хангамж бүтээж байна. Тус компани Nvidia-гийн CUDA программ хангамжийн нэгэн адил бүх төрлийн чип, тухайлбал GPU, SRAM болон утасны чипүүдтэй ажиллах боломжтой программ хангамжийн стек хөгжүүлж байгаа аж. Энэ нь хөгжүүлэгчдэд Nvidia-гаас хамааралгүйгээр AI аппликейшн бүтээх боломжийг олгох зорилготой юм.
Google Brain-ийн судлаач асан Жереми Никсон өнгөрсөн жил тус стартапыг үүсгэн байгуулжээ. Түүний хөгжүүлсэн “Ignition” хэмээх AI агент нь AI дүгнэлт (inference) хийхэд шаардлагатай доод түвшний кодыг автоматаар бичиж, алдааг засварлан, гүйцэтгэлийг сайжруулдаг. Энэхүү систем нь өөрөө суралцаж, чипийн архитектурт дасан зохицох чадвартай юм.
Infinity-гийн шийдлийг одоогоор D-Matrix зэрэг чип үйлдвэрлэгчид ашиглаж байгаа бөгөөд бусад томоохон үүлэн тооцооллын компаниудтай хэлэлцээ хийж байна. Стартап нь урьдчилсан төлбөр авахын оронд ажиллагааны хурд болон зардлын хэмнэлтэд суурилсан төлбөрийн загварыг баримталдаг. Нийт 26 хүний бүрэлдэхүүнтэй тус баг хүний оролцоотойгоор олон сар, жилээр үргэлжлэх ажлыг хэдхэн цаг эсвэл өдрөөр хэмжигдэх хугацаанд гүйцэтгэх боломжийг бүрдүүлж байна.
Дэлгэрэнгүйг эх сурвалжаас харах
↓Эх сурвалжийг нээх ↓
AIinfrastructure company Infinityannounced a $15 million raise at a $100 million valuation on Monday from investors including Touring Capital, Principal VC, andresearchersfrom companiessuch as OpenAI and Anthropic.
The startup is building software to make it easier for AI chips to run AI models. One big reason Nvidia became the top player is not just its high-performance chips, but also its CUDA software (Compute Unified Device Architecture), which allows its GPUs (originally designed to run graphics) to act as general-purpose processing CPUs.The largest AI development frameworksPyTorchand TensorFlow have been built on top ofCUDA. Thisallows developers to write their appsin popular languages likePython, usethose major AIframeworks andtheir appswill, by default,run on Nvidia chips.
Most of these app-level startups wouldn’t have the resources or know-how to write their own kernels — the low-level software that operates chips — and port their apps to other AI chips.SoInfinity is trying to buildCUDA-alternative kernelsoftware that works withany type of chip, like SRAM, GPUs, phone chips, and Systolic Arrays.Infinity ispart of a new wave of startups that areattempting, product by product,tochip away at Nvidia’s market dominance.
Infinity is attempting to build a universal inference library to run on all chips, allowing these chips to automate replicating state-of-the-artresearch results.
Infinitywas launched last year by Jeremy Nixon, once a researcher at Google Brain and creator of the hacker networkcommunityAGI House.Nixon told TechCrunch he decided to launch this company because he was obsessed with the idea of “automated invention” — the belief that “AI systems can actually be ametatechnology.” He himself had invented a machine learning algorithm called Omega, he said, whichessentially creatednew machine learning algorithms and automatically evaluated them in a feedback loop.
That success got him thinking about other cases where this approach could work, and he turned tohardware,believingthat automatedsystems could also generate the low-level code, like the kernels and so forth, needed to help run chips more effectively.
Infinity’s AI research agent Ignition is intended to write the low-level code needed for AI inference on Nvidia-alternative chips. It tests, debugs, and measures how fast the hardware performs with thecode, andautomatically rewrites the code if needed to improve performance. The system is self-optimizing, meaning it continuously learns and improves itself. It also adapts to different chip architectures, regardless of proprietary designs, Nixon says. The result is what Infinity claims is a CUDA-level software stack.
Customers include the AI chip maker (and would be Nvidia challenger)D-Matrix, and Infinity is in talks with other big chip and cloud companies, Nixon said.
Humans are in the loop, however, providing high-level direction while the agent does more of the tedious grunt work.Inone case study, the startup found theagent works much faster than a human alone, reducing what could have been a years- or months-long process to hours or days.Infinitydoesn’tcharge an upfront license fee; instead,it takes a cut of performance gains and cost savings, measuring changes in tokens per second.
Right now, Infinity has 26 employees, including those in design, operations, and engineering.
When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

