Inception Raises 50 Million To Build Diffusion Models For Code And Text
Captured source
source ↗Inception raises $50 million to build diffusion models for code and text | TechCrunch
–:–:–:–
The first StrictlyVC of 2026 hits SF on April 30. Tickets are going fast. Register now.
Founder Summit ticket savings of up to $190 end June 26. Join 1,000+ founders and VCs for all-day bootcamp. REGISTER NOW.
Close
AI
Inception raises $50 million to build diffusion models for code and text
Russell Brandom
5:00 AM PST · November 6, 2025
With so much money flooding into AI startups, it’s a good time to be an AI researcher with an idea to test out. And if the idea is novel enough, it might be easier to get the resources you need as an independent company instead of inside one of the big labs.
That’s the story of Inception , a startup developing diffusion-based AI models that just raised $50 million in seed funding . The round was led by Menlo Ventures, with participation from Mayfield, Innovation Endeavors, Microsoft’s M12 fund, Snowflake Ventures, Databricks Investment, and Nvidia’s venture arm NVentures. Andrew Ng and Andrej Karpathy provided additional angel funding.
The leader of the project is Stanford professor Stefano Ermon, whose research focuses on diffusion models — which generate outputs through iterative refinement rather than word-by-word. These models power image-based AI systems like Stable Diffusion, Midjourney, and Sora. Having worked on those systems since before the AI boom made them exciting, Ermon is using Inception to apply the same models to a broader range of tasks.
Together with the funding, the company released a new version of its Mercury model, designed for software development. Mercury has already been integrated into a number of development tools, including ProxyAI, Buildglare, and Kilo Code. Most importantly, Ermon says the diffusion approach will help Inception’s models conserve on two of the most important metrics: latency (response time) and compute cost.
“These diffusion-based LLMs are much faster and much more efficient than what everybody else is building today,” Ermon says. “It’s just a completely different approach where there is a lot of innovation that can still be brought to the table.”
Understanding the technical difference requires a bit of background. Diffusion models are structurally different from auto-regression models, which dominate text-based AI services. Auto-regression models like GPT-5 and Gemini work sequentially, predicting each next word or word fragment based on the previously processed material. Diffusion models, trained for image generation, take a more holistic approach, modifying the overall structure of a response incrementally until it matches the desired result.
The conventional wisdom is to use auto-regression models for text applications, and that approach has been hugely successful for recent generations of AI models. But a growing body of research suggests diffusion models may perform better when a model is processing large quantities of text or managing data constraints . As Ermon tells it, those qualities become a real advantage when performing operations over large codebases.
Diffusion models also have more flexibility in how they utilize hardware, a particularly important advantage as the infrastructure demands of AI become clear. Where auto-regression models have to execute operations one after another, diffusion models can process many operations simultaneously, allowing for significantly lower latency in complex tasks.
“We’ve been benchmarked at over 1,000 tokens per second, which is way higher than anything that’s possible using the existing autoregressive technologies,” Ermon says, “because our thing is built to be parallel. It’s built to be really, really fast.”
Topics
AI , diffusion , inception , menlo ventures , Startups
When you purchase through links in our articles, we may earn a small commission . This doesn’t affect our editorial independence.
Russell Brandom
AI Editor
Russell Brandom has been covering the tech industry since 2012, with a focus on platform policy and emerging technologies. He previously worked at The Verge and Rest of World, and has written for Wired, The Awl and MIT’s Technology Review. He can be reached at russell.brandom@techcrunch.com or on Signal at 412-401-5489.
View Bio
November 4
Boston
Last chance to save up to $190 on TechCrunch Founder Summit. Join 1,000+ founders and VCs at all stages for real-world scaling insights and connections that move the needle.
Savings end June 26, 11:59 p.m. PT .
REGISTER NOW
Most Popular
Trump administration proposes axing brake-pedal requirement for AVs in a boost for Tesla
Sean O'Kane
Former Infosys chief has a new startup that wants to challenge the IT services world
Jagmeet Singh
Here’s why Slate changed the battery in its cheap EV truck
Tim De Chant
OpenAI unveils its first custom chip, built by Broadcom
Russell Brandom
Slate Auto’s radically simple electric truck starts at $24,950
Sean O'Kane
HaloBraid raises $7M from Seven Seven Six to end the six-hour hair salon appointment
Dominic-Madori Davis
WhatsApp gets new chief as Meta taps India’s CRED founder Kunal Shah and invests $900M in startup
Jagmeet Singh
Loading the next article
Error loading the next article
Notability
notability 7.0/10Significant $50M funding for diffusion model startup