NEW GPT 5.5 Instant Just Dropped artwork

NEW GPT 5.5 Instant Just Dropped

AI News Today | Julian Goldie Podcast

June 26, 2026

GPT 5.5 Instant: Faster Replies, Better Intent Understanding + How to Use It in Hermes AgentThe script introduces OpenAI’s newly released GPT 5.5 Instant, describing it as faster and better at understanding intent, following complex constraints, and improving shopping and local recommendations.
Speakers: Julian Goldie
**Julian Goldie** (0:00)
There's a brand new version of GPT 5.5 today, which is GPT 5.5 Instant. This is dropped. I'm going to show you what a test show about it and how it's worked so far. If you've never used this before, basically straight from OpenAI, just dropped a few hours ago.
So the model should be better at understanding the intent behind your question and adapting to it also handles complex constraints more reliably. So you can basically give it a few rules and it keeps up, right?
And shopping and local recommendations also got more useful too. It's rolling out today. Today is only for subscribers, but tomorrow it will actually roll out to free members as well. Bear in mind, you can use this with Hermes Agent or any sort of agent they actually use. So we've already been testing it. We've created a separate profile for GPT 5.5 Instant and it is a lot faster. And you can just log in with OAuth. So if you already have a subscription to ChatGPT, you can log in with OAuth into Hermes. You can also use it as the image generation model as well. Now, if you want to indicate how much faster is it actually, let me show you side by side. I actually just tested out a second ago comparing GPT 5.5 Instant, which you can see on the agent profile right here inside the agent operating system. And we've compared that versus the default. So this is like the normal model we would use and this is GPT 5.5 Instant. So what we can do is we can actually see side by side, which one performs the fastest and which one replies quicker. So let's actually see if they're faster side by side. So we've got Groq Build built into Hermes over here. And then we have GPT 5.5 Instant over here. So we're going to say, okay, tell me a joke. And we'll plug those into both of them at the same time. So this is GPT 5.5 Instant. This is Groq Build. Let's compare them side by side, see how they perform.
All right, and you can see the countdown here.
So this replied pretty much instantly. As you can see, this one is a lot slower. So that took about nine seconds. This one took about three seconds to reply, which I think if you're using Hermes or any sort of agent day to day, it does genuinely make a big difference and it helps as well. And then also if we look at the outputs here side by side, so tell me a joke, why did the developer run out? Because he used up all his cash.
That's from Groq Build. Let's have a look at GPT 5.5 Instant. Why did the scarecrow win an award? Because he was outstanding in his field. Both pretty bad jokes, but the point is this is a lot faster and you can just use your existing subscription with it as well.
So it's designed to be like intuitive, smart and very fun to chat with. He's running out to ProFest and Plus and free users tomorrow.
They did an actual example screenshot here. So they just typed in goat birthday. And you can see this, if you're talking about the goat, then today is Lionel Messi's birthday. And then it adds some details here. It says happy birthday to Leo with a goat emoji. So the response is pretty nice on the example right here. It's probably an optimized example, but you get the point. Something to note as well here is like GPT 5.6 was expected. So Mark, you can see here posted on X about this, that it might be kind of like a stop gap measure to make up for the fact that GPT 5.6 is delayed, which is pretty interesting. Rarely do I use a model without thinking as well. I mean, that's a really good point. I think what this model would actually be good for is like if you have, if you're using Hermes, but you just use it as kind of like a chat that's custom trained to you, and you don't really use the agentic task or you don't do anything complex with it, it will probably be great for that. So for example, little back and forth side this that you can just do on your phone or when you're out and about, or if you're trying to delegate tasks to subagents, that could be another option too.
When it comes to actually like building something, obviously, you're not going to build like a full agent operating system with ChatGPT Instant. So that's something to bear in mind as well. So why is this important? Well, it means you can probably be a bit lazier and shorter, and you will understand what you mean. So it understands intent better. You can actually stack a few rules and it will follow them.

4 more minutes of transcript below

Feed this to your agent

Try it now — copy, paste, done:

curl -H "x-api-key: pt_demo" \
  https://spoken.md/transcripts/1000651996090

Works with Claude, ChatGPT, Cursor, and any agent that makes HTTP calls.

From $0.10 per transcript. No subscription. Credits never expire.

Using your own key:

curl -H "x-api-key: YOUR_KEY" \
  https://spoken.md/transcripts/1000774360413