TMCnet News
Infinity Raises $15 Million in Seed Funding to Build the Software Layer That Makes Any AI Chip Inference-ReadyInfinity.inc (Infinity Artificial Intelligence Institute), an early-stage AI infrastructure research company building the software layer that makes any AI chip inference-ready, today announced it has raised $15 million in seed funding at a $100 million post-money valuation. The round included significant participation from Touring Capital, along with Principal VC, executives at major chip companies, researchers from OpenAI and Anthropic, and other prominent angel investors. The funding will be used to scale Infinity's automated AI research platform and its autonomous AI agent, Ignition, which writes inference code for any new AI chip; expand Infinity's engineering team; and accelerate its work with chip partners, including d-Matrix. Infinity is already generating millions of dollars in ARR from its chip design partnerships, as it solves a key challenge for established and new silicon companies alike: creating the software that runs AI inference on their chips. "The AI industry has operated under an artificial constraint that only a handful of chips could run AI well, because only NVIDIA spent decades building the software to make them work optimally. Ignition eliminates that constraint," said Jeremy Nixon, founder and CEO of Infinity. "We believe the next era of AI will be defined not just by who makes the best chip, but by who can make any chip run state-of-the-art models at blazing speeds. For hardware providers, the difference between having that optimized software stack and not is the difference between mere potential and true performance. Our role is to ensure every partner reaches that potential." NVIDIA's CUDA ecosystem has long anchored the AI accelerator market, with NVIDIA holding an estimated ~80% share of data-center AI accelerators. While hardware performance from providers like AMD, Qualcomm, and AWS often rivals or exceeds industry standards, the challenge has historically been the time-intensive development of production-grade inference stacks. Infinity bridges this gap, allowing these providers to rapidly deploy their hardware at scale. Addressing this development cycle is a top priority across the semiconductor industry, as inference is on track to represent two-thirds of all AI compute spending in 2026. Ignition: Any Chip. Any Model. In Days, Not Years. Infinity's answer is Ignition, an AI research agent that automatically generates, tests and optimizes the low-level compute kernels that determine how efficiently a chip runs AI models. Ignition empowers human engineers to focus on high-level architecture rather than manual tuning by utilizing critical human steering and collaborative input throughout the optimization process. The system iterates continuously, with performance data feeding each successive generation of code. Infinity works shoulder to shoulder with partner engineering teams, treating their hardware and software stack as a strategic foundation that Ignition supercharges. Ignition's core features and benefits include:
"Every major technology platform ultimately creates the most value when it becomes broadly accessible. Infinity is tackling one of the defining challenges of the AI era: making powerful AI affordable and available at scale," said Songyee Yoon, founder and Managing Partner, Principal Venture Partners. "We are proud to back a team building the infrastructure needed to extend the benefits of this revolution to many more people." "As neural network architectures evolve at a rapid pace, AI hardware companies are caught in a constant cycle of keeping their software stacks current with state-of-the-art models and approaches," said Samir Kumar, General Partner at Touring Capital. "Infinity solves this fundamental software bottleneck. Its AI agents automate kernel development and build optimized inference libraries tailored to each target hardware platform, dramatically shortening the time it takes new silicon to achieve its intended peak performance. We're excited to support Jeremy and his team as they change the game in how chip companies maximize the real-world performance of their silicon." Infinity was founded by Nixon, a former Google Brain researcher who co-founded AGI House, a San Francisco-based Artificial General Intelligence community and hacker network that has launched hundreds of startups and projects. Nixon is a widely recognized voice on the future of AI, with his commentary covered in The New York Times and Forbes. Infinity is an environment built for engineers who aspire to 100x impact and is currently hiring engineers to help build the future of RSI systems for discovery. Contact [email protected] to join the Infinity.inc team in unlocking the true potential of AI hardware. INFINITY.INC FREQUENTLY ASKED QUESTIONS: Who founded Infinity.inc? Infinity.inc was founded by Jeremy Nixon, a former Google Brain researcher, co-founder of the AGI Houses, a San Francisco and Hillsborough-based AI community and hacker network that has launched hundreds of startups and projects and a widely recognized voice on the future of AI, with his work and commentary covered by The New York Times and Forbes. What does Infinity.inc do? Infinity builds the software layer that enables any AI chip to run inference workloads. Its autonomous AI agent, Ignition, automatically generates and optimizes the inference software stack for new silicon in days, a process that has historically taken engineering teams months or years. What is Ignition and how does it work? Ignition is Infinity's AI research agent that autonomously writes, tests and optimizes the low-level compute kernels that determine how efficiently a chip runs AI models. It operates without human kernel engineers and continuously improves performance through a real-world feedback loop. Why can't new AI chips compete with NVIDIA out of the box? NVIDIA's CUDA ecosystem represents two decades of inference software development that rival chipmakers cannot quickly replicate. Most new chips fail not because of weak hardware, but because they lack the software stack to run AI models efficiently. Infinity's platform solves this by automating the software development entirely. How fast can Infinity bring a new chip to production-ready AI inference? In a published case study with chip partner d-Matrix, Infinity reached 92% of a new chip's theoretical peak performance within 10 hours of first hardware access and had three frontier models running end-to-end within 10 days. That same process has historically taken engineering teams months or years. What AI inference performance gains has Infinity demonstrated? Using Ignition, Infinity improved inference throughput on a Qwen3-8B model from roughly 1,400 tokens per second to more than 20,000 tokens per second, a 14x gain achieved in a single day, outperforming the widely used vLLM framework by more than 34%. Who invested in Infinity and how much has it raised? Infinity raised $15 million in seed funding at a $100 million post-money valuation, with significant participation from Touring Capital, along with executives at major chip companies, researchers from OpenAI and Anthropic, and other prominent angel investors, including Infinity founder and CEO Jeremy Nixon. Which chip companies is Infinity working with? Infinity's first live design partnership is with d-Matrix, whose Corsair chip now runs production-quality inference via the Infinity d-Matrix Cloud. Infinity is currently in active partnership discussions with other major chip companies. Is Infinity.inc hiring? Hiring new talent is one of Infinity's key goals after raising its $15M seed funding round. The company plans to expand its engineering team. Current open roles include research engineers working on hardware enablement and global inference library engineers maintaining Infinity's Infy library. About Infinity.inc Infinity.inc (Infinity Artificial Intelligence Institute) is an early-stage AI infrastructure research company building the software layer that makes any chip competitive for AI inference. Its AI research agent, Ignition, automatically generates, tests and optimizes the low-level compute kernels that determine how efficiently a chip runs AI models, compressing inference software development from years into days. Infinity's live design partnership with d-Matrix powers the Infinity d-Matrix Cloud, and is currently in active partnership discussions with other major chip companies. Founded in August 2025 and headquartered in San Francisco, Infinity is a privately held company backed by Touring Capital and prominent angel investors, including executives at major chip companies, researchers from OpenAI and Anthropic, and Infinity CEO and founder, Jeremy Nixon. Follow Infinity.Inc on X and LinkedIn and Jeremy Nixon on X and LinkedIn.
View source version on businesswire.com: https://www.businesswire.com/news/home/20260720324097/en/ |

