AI infrastructure startup Infinity has raised $15 million to develop software that allows AI models to run on various chips, aiming to rival Nvidia's market position.
New Delhi, India Jul 20, 2026 ALN: AI infrastructure company Infinity announced a $15 million raise at a $100 million valuation on Monday from investors including Touring Capital, Principal VC, and researchers from companies such as OpenAI and Anthropic.
The startup is building software to make it easier for AI chips to run AI models. One significant reason Nvidia became the top player is not just its high-performance chips, but also its CUDA software (Compute Unified Device Architecture), which allows its GPUs (originally designed to run graphics) to act as general-purpose processing CPUs. The largest AI development frameworks PyTorch and TensorFlow have been built on top of CUDA. This allows developers to write their apps in popular languages like Python, use those major AI frameworks, and their apps will, by default, run on Nvidia chips.
Most of these app-level startups wouldnât have the resources or know-how to write their own kernels â the low-level software that operates chips â and port their apps to other AI chips. So Infinity is trying to build CUDA-alternative kernel software that works with any type of chip, like SRAM, GPUs, phone chips, and Systolic Arrays. Infinity is part of a new wave of startups that are attempting, product by product, to chip away at Nvidiaâs market dominance.
Infinity is attempting to build a universal inference library to run on all chips, allowing these chips to automate replicating state-of-the-art research results. The significance of this effort cannot be overstated, especially in a technology landscape that is increasingly reliant on AI and machine learning. As AI applications proliferate across various sectors, the demand for versatile and efficient processing solutions has never been higher. Infinityâs approach could democratize access to AI capabilities by enabling a broader range of hardware to support advanced AI models.
Nvidia's dominance in the AI chip market is a result of several factors, including its early investments in AI-specific architecture and software ecosystems. CUDA has become a cornerstone of AI development, creating a significant barrier to entry for competitors. By developing an alternative software stack, Infinity aims to lower the dependency on Nvidiaâs technology and foster a more competitive environment in the AI hardware market.
Infinity was launched last year by Jeremy Nixon, once a researcher at Google Brain and creator of the hacker network community AGI House. Nixon stated he decided to launch this company because he was obsessed with the idea of âautomated inventionâ â the belief that âAI systems can actually be a meta technology.â He himself had invented a machine learning algorithm called Omega, which essentially created new machine learning algorithms and automatically evaluated them in a feedback loop. This innovation reflects a broader trend in AI research where self-improving algorithms are becoming increasingly feasible.
That success got him thinking about other cases where this approach could work, and he turned to hardware, believing that automated systems could also generate the low-level code, like the kernels and so forth, needed to help run chips more effectively. This belief is rooted in the understanding that the hardware-software interface is critical for maximizing the performance of AI models. By automating the generation of low-level code, Infinity seeks to eliminate the bottlenecks that developers currently face when trying to optimize their applications for different hardware.
Infinityâs AI research agent Ignition is intended to write the low-level code needed for AI inference on Nvidia-alternative chips. It tests, debugs, and measures how fast the hardware performs with the code, and automatically rewrites the code if needed to improve performance. The system is self-optimizing, meaning it continuously learns and improves itself. It also adapts to different chip architectures, regardless of proprietary designs, Nixon says. The result is what Infinity claims is a CUDA-level software stack. This adaptability is crucial in a rapidly evolving tech landscape where new chip architectures are frequently developed, often with unique features and capabilities.
Customers include the AI chip maker (and would-be Nvidia challenger) D-Matrix, and Infinity is in talks with other big chip and cloud companies, Nixon said. The interest from established companies underscores the potential market demand for alternatives to Nvidia's offerings. As businesses increasingly look to integrate AI into their operations, having access to a broader range of hardware options could lead to more innovative applications and solutions.
Humans are in the loop, however, providing high-level direction while the agent does more of the tedious grunt work. In one case study, the startup found the agent works much faster than a human alone, reducing what could have been a years- or months-long process to hours or days. This highlights the potential for AI to augment human capabilities rather than completely replace them. By allowing humans to focus on strategic decisions and creative problem-solving, Infinity aims to enhance productivity in the AI development process.
Infinity doesnât charge an upfront license fee; instead, it takes a cut of performance gains and cost savings, measuring changes in tokens per second. This business model aligns the interests of Infinity with its customers, as both parties benefit from improved performance. It also reduces the financial risk for startups and smaller companies that may be hesitant to invest heavily in new technology without a proven return on investment.
Right now, Infinity has 26 employees, including those in design, operations, and engineering. The companyâs growth trajectory will likely depend on its ability to attract talent in a competitive job market, especially as demand for skilled professionals in AI and machine learning continues to rise. As Infinity develops its technology and expands its customer base, it will be crucial for the company to maintain a culture of innovation and adaptability.
The implications of Infinity's funding and technological advancements extend beyond just the company itself. If successful, Infinity could play a pivotal role in reshaping the AI hardware landscape, fostering greater competition and innovation. This shift could lead to more affordable and accessible AI solutions, ultimately benefiting a wide range of industries, from healthcare to finance to entertainment. As AI technology continues to evolve, the importance of diverse and flexible hardware solutions will only grow, making Infinity's mission increasingly relevant.
To learn more about the latest developments in Funding & Investments, stay updated with our exclusive reports and analyses on AiLensNews.