Infinity Raises $15M to Scale AI Chip Infrastructure

AIVC

Sofya Zhamoitina

Venture Reporter at The Top Voices

July 23, 20261 min

Article hero imageImage credit: Infinity

Key Takeaways:

  • Infinity raised $15M at a $100M post-money valuation.
  • Ignition automates inference software development for AI chips.
  • Funding will scale the platform, engineering team, and chip partnerships.

Infinity.inc, an early-stage AI infrastructure research company developing software that makes AI chips inference-ready, raised $15 million in seed funding at a $100 million post-money valuation. Touring Capital participated alongside Principal VC, chip industry executives, OpenAI and Anthropic researchers, and angel investors. The funding will scale Infinity’s automated research platform and Ignition AI agent, expand engineering operations, and accelerate chip partnerships including d-Matrix.

Making AI Chips Inference-Ready

Ignition automatically generates, tests and optimizes low-level compute kernels required to run AI models efficiently across different chip architectures. The platform addresses a major software bottleneck for semiconductor companies by reducing inference stack development from lengthy manual processes to automated workflows completed within days.

The AI industry has operated under an artificial constraint that only a handful of chips could run AI well, because only NVIDIA spent decades building the software to make them work optimally. Ignition eliminates that constraint. We believe the next era of AI will be defined not just by who makes the best chip, but by who can make any chip run state-of-the-art models at blazing speeds. For hardware providers, the difference between having that optimized software stack and not is the difference between mere potential and true performance. Our role is to ensure every partner reaches that potential.” — Jeremy Nixon, founder and CEO of Infinity

Accelerating Chip Performance

Infinity demonstrated a 34% improvement in inference throughput for a Qwen3-8B model during one day of automated optimization. In partnership with d-Matrix, AI agents reached up to 92% of a new chip’s theoretical peak performance and enabled three frontier models to run end-to-end within 10 days.

The approach to AI-driven model enablement has the potential to significantly shorten the time required to bring new AI compute architectures into production, helping accelerate deployment and time-to-first-revenue. We're excited about Infinity's mission to help unlock the full potential of the next generation of AI compute.” — Sid Sheth, founder and CEO of d-Matrix

Expanding AI Infrastructure

The investment will support development of Ignition, engineering team growth and broader collaboration with semiconductor companies seeking faster deployment of AI models across new hardware architectures.

1483 views

Stay Ahead in Tech & Startups

Get monthly email with insights, trends, and tips curated by Founders

Join 3000+ startups

The Top Voices newsletter delivers monthly startup, tech, and VC news and insights.

Dismiss