Topics: Investing, Business, Technology
**Patrick O'Shaughnessy** (0:00)
Ramp is the only platform built to make your finance team leaner, faster, and better, saving businesses 5% annually on average, so you can stay focused on growth. Ramp customers grow revenue 3.2 times faster than the average American business. Visa, Vercel, Cursor, Stripe, Notion, 11Lab, Shopify, and 70,000 other businesses all now run on Ramp. Mine does too, and so should yours. Learn more at ramp.com/invest.
OpenAI, Cursor, Anthropic, Perplexity, and Vercel all have something in common. They all use Work OS. To achieve enterprise adoption at scale, you have to deliver on core capabilities like SSO, SCIM, RBAC, and Audit Logs. Instead of spending months building these mission-critical capabilities yourself, you can just use Work OS APIs to gain all of them on day zero. That's why so many of the top AI teams you hear about already run on Work OS. Work OS is the fastest way to become enterprise-ready and stay focused on what matters most, your product. Visit workos.com to get started.
Felix by Rogo is a personal finance agent that turns a single prompt into finished client-ready work using your firm's own templates, context, and standards. Send Felix an e-mail like, take these comments and turn them for me, or update my tracker with the context of these e-mails, and Felix sends back finished PowerPoint decks, Excel models, and sourced research. Felix works the way your team already does, delivering work quickly and accurately around the clock. Learn more at rogo.ai/felix.
Hello and welcome everyone. I'm Patrick O'Shaughnessy, and this is Invest Like the Best. This show is an open-ended exploration of markets, ideas, stories, and strategies that will help you better invest both your time and your money. If you enjoy these conversations and want to go deeper, check out Colossus, our quarterly publication with in-depth profiles of the people shaping business and investing. You can find Colossus along with all of our podcasts at colossus.com.
**SPEAKER_2** (1:52)
Patrick O'Shaughnessy is the CEO of PositiveSum. All opinions expressed by Patrick and podcast guests are solely their own opinions and do not reflect the opinion of PositiveSum. This podcast is for informational purposes only and should not be relied upon as a basis for investment decisions. Clients of PositiveSum may maintain positions in the securities discussed in this podcast. To learn more, visit psum.vc.
**Patrick O'Shaughnessy** (2:19)
My guest today is Neil Movva, the founder of SAIL Research. SAIL is building what Neil calls a token factory, an inference company designed for a specific kind of future, one where AI agents run in the background for hours or days at a time, rather than answering a human in real time. In that world, latency matters less and cost matters much more. Neil has built the entire company around driving the cost of a token as low as it can possibly go.
What makes this conversation special is it is one of the most detailed tours I've ever run through the full stack of intelligence, the software, the chips, the power, and how the three connect. Along the way, we cover the trade-off between speed and cost that lives inside of every GPU, his scavenger strategy for buying the chips and power no one else wants, his contrarian view on Nvidia, and why the premium the Frontier Labs charge for being three to six months ahead may not last. Please enjoy my conversation with Neil Movva.
I think it's important early in these conversations to just say the thing, literally what you're building and what it does today. Maybe just orient us there with a brief description, like literally what the system is that you're building and why it should exist.
**Neil Movva** (3:21)
Sail Research is a Token Factory. We have an API where anyone can send us requests, where they can use large language models, open source large language models for any task they want. We will serve those tokens to them at a price that is unbeatable in the market. We also support their ability to build agents on top of this. We host what we call Sailboxes, which are long-running agent virtual machines hosted in the cloud that are designed for agents that run for hours, days or weeks.
**Patrick O'Shaughnessy** (3:45)
So I should think about you as a peer company to others that serve different kinds of inference. You're serving one specific kind of inference. And your goal is to be the absolute cheapest provider and enabler of a certain kind of use of intelligence.
**Neil Movva** (3:58)
Exactly. The theme of our company is abundance. We want to deliver this new commodity of intelligence to as many people as possible at a cost that is sustainable for almost every industry. We think that whenever you make something 10 times cheaper, it's a new product category. We aspire to do that for tokens. We think it's so profound that the machine can think, and now our job is to make as many machines as possible in the world work towards thinking.
84 more minutes of transcript below
Thousands of transcripts fetched by people building searchable podcast archives
Try it now — copy, paste, done:
curl -H "x-api-key: pt_demo" \
https://spoken.md/transcripts/1000651996090
Works with Claude, ChatGPT, Cursor, and any agent that makes HTTP calls.
From $0.10 per transcript. No subscription. Credits never expire. Prices exclude VAT, added at checkout for EU customers. Not what you expected? Email us within 14 days with 20 or fewer credits used and we refund the pack in full.
Using your own key:
curl -H "x-api-key: YOUR_KEY" \
https://spoken.md/transcripts/YOUR_EPISODE_ID