AI infrastructure company Infinity introduced a $15 million increase at a $100 million valuation on Monday from traders together with Touring Capital, Principal VC, and researchers from firms resembling OpenAI and Anthropic.
The startup is constructing software program to make it simpler for AI chips to run AI fashions. One large motive Nvidia turned the highest participant isn’t just its high-performance chips, but in addition its CUDA software program (Compute Unified System Structure), which permits its GPUs (initially designed to run graphics) to behave as general-purpose processing CPUs. The most important AI growth frameworks PyTorch and TensorFlow have been constructed on prime of CUDA. This enables builders to put in writing their apps in widespread languages like Python, use these main AI frameworks and their apps will, by default, run on Nvidia chips.
Most of those app-level startups wouldn’t have the assets or know-how to put in writing their very own kernels — the low-level software program that operates chips — and port their apps to different AI chips. So Infinity is making an attempt to construct CUDA-alternative kernel software program that works with any sort of chip, like SRAM, GPUs, telephone chips, and Systolic Arrays. Infinity is a part of a brand new wave of startups which might be making an attempt, product by product, to chip away at Nvidia’s market dominance.
Infinity is making an attempt to construct a common inference library to run on all chips, permitting these chips to automate replicating state-of-the-art analysis outcomes.
Infinity was launched final yr by Jeremy Nixon, as soon as a researcher at Google Mind and creator of the hacker community group AGI Home. Nixon informed TechCrunch he determined to launch this firm as a result of he was obsessive about the concept of “automated invention” — the assumption that “AI methods can really be a meta know-how.” He himself had invented a machine studying algorithm known as Omega, he mentioned, which primarily created new machine studying algorithms and mechanically evaluated them in a suggestions loop.
That success obtained him excited about different instances the place this method may work, and he turned to {hardware}, believing that automated methods may additionally generate the low-level code, just like the kernels and so forth, wanted to assist run chips extra successfully.
Infinity’s AI analysis agent Ignition is meant to put in writing the low-level code wanted for AI inference on Nvidia-alternative chips. It exams, debugs, and measures how briskly the {hardware} performs with the code, and mechanically rewrites the code if wanted to enhance efficiency. The system is self-optimizing, that means it repeatedly learns and improves itself. It additionally adapts to completely different chip architectures, no matter proprietary designs, Nixon says. The result’s what Infinity claims is a CUDA-level software program stack.
Clients embody the AI chip maker (and would-be Nvidia challenger) D-Matrix, and Infinity is in talks with different large chip and cloud firms, Nixon mentioned.
People are within the loop, nevertheless, offering high-level path whereas the agent does extra of the tedious grunt work. In one case study, the startup discovered the agent works a lot sooner than a human alone, lowering what may have been a years- or months-long course of to hours or days. Infinity doesn’t cost an upfront license payment; as a substitute, it takes a reduce of efficiency features and value financial savings, measuring modifications in tokens per second.
Proper now, Infinity has 26 workers, together with these in design, operations, and engineering.
If you buy by means of hyperlinks in our articles, we may earn a small commission. This doesn’t have an effect on our editorial independence.

