Context windows, computer constraints, and energy consumption with Sarah and Elad artwork

Context windows, computer constraints, and energy consumption with Sarah and Elad

No Priors: Artificial Intelligence | Technology | Startups

May 9, 2024

This week on No Priors hosts, Sarah and Elad are catching up on the latest AI news. They discuss the recent developments in AI like Meta’s new AI assistant and the latest in music generation, and if you’re interested in generative AI music, stay tuned for next week’s interview!
Speakers: Sarah, Elad
**Sarah** (0:05)
Hey, listeners, you are here for another episode of No Priors just with me and Elad. And there's been a lot going on in earnings and in the technical world. So I think we will start with maybe just one fun thing that seems to have taken flight in terms of music generation.
Elad, what do you make of the popularity that Suno and Udo have found?

**Elad** (0:31)
When you said fun things, I thought you were going to talk about my hats. I have two hats.

**Sarah** (0:35)
We can talk about your hats.

**Elad** (0:36)
I have my Bitcoin having hat that I got from Coinbase, as you see, Bitcoin having. And then Zayde has these Make AI Great, again, hats that he's actually selling. I think it's going to fund his data labeling habit.

**Sarah** (0:49)
That's a lot of hats.

**Elad** (0:50)
I know. Yeah, that's all I got. That's all I got left, Sarah.

**Sarah** (0:53)
If anybody has read Jared Kushner's book, there's a great bit about how much money they're making from all the MAGA swag. And so, listeners, Elad and I are making hats and tequila for the guests.

**Elad** (1:07)
Yeah.

**Sarah** (1:07)
But for the low, low price of one H100 GPU, we will give each of those to you too.

**Elad** (1:14)
Or a Bitcoin, either way.

**Sarah** (1:15)
Okay, I'll take the Bitcoin instead. Yeah.

**Elad** (1:17)
Me too.
I think we all would at this point. We can check in again in a couple of months when the B100 has come out, or whatever time period.
So, as you know, there's been some really interesting things happening on the music generation side. And so, there's both Suno and UDO, and both seem to be kind of taking people by storm in terms of really interesting music-based models. And it feels like one of those things where it's early, but it's really giving a glimpse of what's coming in terms of the ability to create other types of content. Obviously, the very first content wave in some sense, was simple text-based things on GPT-3, like Jasper. And then we hit an image genwave, and that was mid-journey and stable diffusion and things like that. And then we had obviously chat come out as sort of a new type of format and interaction modality, and then we had video with things like Pika. And so it just feels like sequentially we're hitting these different formats.
And then obviously, Suno from OpenAI, and now we have these really interesting music models where you can specify the type of music that you want. You can write the lyrics, it'll add vocals. And so these really seem to be the two models initially, at least the people are really adopting. And so it just seems like an interesting moment in time from the perspective of look at all these different creative things that people are now empowered to do, and look at the different ways to engage. And of course, you could imagine going forward in time and saying, okay, at some point, there'll be voice cloning where, and I think it was Drake who put out a song, right, where he had two or three other rappers that he just voice cloned in. And you can imagine a world where you could use anybody's voice, assuming there's permissions and everything else, to generate your own songs and content and all things. So it just seems like a very exciting future world between UDOS and some of these other companies.

**Sarah** (2:59)
Yeah, I think one of the things that's not obvious here is in media platforms in general, the ratio varies, right, but there are a lot more readers on X or consumers, like people who scroll a feed on TikTok, or not TikTok anymore, I suppose, but whatever it is, than creators.
And so I think one thing that is just unknown is how many people actually want to create music if you make it a lot easier to create something that's any good, right? And if the music we get changes. And so I was talking to one of these founders and he was like, everybody should have a personalized soundtrack for their life, but in the voice of Taylor Swift, in the style of Taylor Swift.

**Elad** (3:47)
Yeah, I already have one of those, but yeah.

**Sarah** (3:50)
So what is yours?

**Elad** (3:51)
I can't really share publicly, but we can talk about it later.

**Sarah** (3:55)
Yeah. It's going to be a little bit, I think it's going to be a little bit of a personal thing.

**Elad** (3:59)
Yeah, that's true.

**Sarah** (4:00)
What else is going on?

**Elad** (4:02)
I don't know. I mean, what about local LLMs and the Apple release? Do you want to talk about that?

24 more minutes of transcript below

Feed this to your agent

Try it now — copy, paste, done:

curl -H "x-api-key: pt_demo" \
  https://spoken.md/transcripts/1000651996090

Works with Claude, ChatGPT, Cursor, and any agent that makes HTTP calls.

From $0.10 per transcript. No subscription. Credits never expire.

Using your own key:

curl -H "x-api-key: YOUR_KEY" \
  https://spoken.md/transcripts/1000655037921