Intelligence on the Edge: Liquid AI's Ramin Hasani on the Search for Device-Native Foundation Models artwork

Intelligence on the Edge: Liquid AI's Ramin Hasani on the Search for Device-Native Foundation Models

"The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis

July 4, 2026

Liquid AI co-founder and CEO Ramin Hasani joins Nathan to make a technically grounded case against the idea that scale alone defines the future of AI.
Speakers: Nathan Labenz, Ramin Hasani
**Nathan Labenz** (0:00)
Hello, welcome back to The Cognitive Revolution, and happy Fourth of July to everyone in the United States. Today, my guest is Ramin Hasani, CEO of Liquid AI, a company founded by MIT researchers that's developing device-native foundation models. I'll say upfront that just before recording, I encouraged Ramin to go deep into the weeds on the technical details of Liquid AI's work. And as you'll hear, he did a truly excellent job, demonstrating a mix of technical sophistication, differentiated vision, and a contagious passion that, in my humble opinion, makes this episode an instant classic.
We start with an overview of the team's research into tiny, biologically-inspired, differential-equation-based neural networks that Ramin and team developed at MIT, and which inspired them to start the company. Some of the capabilities they demonstrated, such as parking a car with a control module that consisted of just 12 liquid neurons, still sound a bit like science fiction today. And while those systems haven't scaled up to today's capabilities frontier, the company has maintained the liquid philosophy, which today means taking a neutral, empirical approach to designing and optimizing neural networks to perform under all sorts of exotic constraints, including most commonly the need to run on edge devices with limited memory and processing power.
Considering the fact that the global smartphone and laptop market is roughly $800 billion per year, a number which the global AI data center buildout is only now surpassing, the demand for inference threatens to price much of the world out of the frontier model market, and that so many enterprises and individuals value privacy and the ability to control their own information, this is an absolutely massive market opportunity unto itself. And Liquid has serious proof points, including holding the number five spot on the hugging face United States downloads leaderboard, plus notable partnerships with companies such as Shopify and Mercedes-Benz.
Anyone who doubts can do a quick download and demo of Liquid's Apollo app, which shows in my experience that even a one billion parameter model, which combines a small number of attention layers with a very simple gated learned convolution, while admittedly far from the frontier, can run fast enough on an iPhone to be a real option for basic use cases, such as privately searching through and classifying one's own local documents.
Perhaps most interesting is the network architecture search process that Liquid uses to develop networks for particular use cases and runtime environments. Having found that proxy metrics too often lead the process astray, they now evaluate models on real downstream tasks, on the actual target hardware that their customers intend to use. Ramin shares a lot more detail on their findings, but in short, while attention-based architectures continue to generalize better than any known alternative and therefore continue to dominate the frontier, the more specific your use case and the more limited the compute resources you have available, the more likely their search process is to land on an exotic architecture. This, for now, is where architectures like Mamba and other subquadratic innovations really shine.
Toward the end, Ramin teases a platform that Liquid will soon be introducing to allow customers to fine-tune small models for their own use cases on a self-serve basis. For multiple reasons, including the potential to ease demand for frontier models and improve access to AI globally, I, for one, will be very excited to see that come online. And so, without further ado, I hope you enjoy this high-energy look at how Liquid AI is squeezing as much intelligence as possible out of any given computational resource. With co-founder and CEO Ramin Hasani.
The Cognitive Revolution is brought to you by Mercury, the fintech that more than 300,000 ambitious companies and individuals trust to run their finances. I've wired AI into nearly every corner of my life. My email, my messages, my calendar. I even gave Mercury virtual cards to my agents, with low limits and category and merchant restrictions, for their autonomous use. But still, my AI's access to my financial data has remained limited. With a normal bank, I might export a bunch of statements and have my assistant process them for me. But for real-time, up-to-date information, and certainly for taking any action, trying to get your agent to use the bank via the browser is just too hard, too slow, and too error-prone to be worth it. And that's why Mercury's new conversational interface, Command, is such a big deal. It's built directly into Mercury, which means you get natural language access to your finances without exposing anything outside of your bank account. No exports, no spreadsheets, no pasting your transactions into third-party tools. I really think a lot of people are going to prefer it this way. And it can already help you take actions too, with everything bound by the permissions and approval policies that you've already set up in your account. I am genuinely impressed to see this level of AI integration in banking in 2026 And so I invite you to join me in the future. Visit mercury.com to learn more and apply online in minutes. Mercury is a fintech company, not an FDIC-insured bank. Banking services provided through Choice Financial Group and Column NA, members FDIC. Thank you to Mercury for supporting the Cognitive Revolution. And now on with the show. Ramin Hasani, co-founder and CEO at Liquid AI. Welcome to The Cognitive Revolution.

84 more minutes of transcript below

Feed this to your agent

Try it now — copy, paste, done:

curl -H "x-api-key: pt_demo" \
  https://spoken.md/transcripts/1000651996090

Works with Claude, ChatGPT, Cursor, and any agent that makes HTTP calls.

From $0.10 per transcript. No subscription. Credits never expire.

Using your own key:

curl -H "x-api-key: YOUR_KEY" \
  https://spoken.md/transcripts/1000775422931