China’s Kimi K3 is SCARY GOOD! artwork

China’s Kimi K3 is SCARY GOOD!

AI News Today | Julian Goldie Podcast

July 18, 2026

China’s Kimi K3 is SCARY GOOD!
Speakers: Julian Goldie
**Julian Goldie** (0:00)
K3 from China is absolutely amazing. I wanna show you some examples of what we've built with it, but so far, let's recap on what we've got. So Kimi K3 just dropped yesterday, about 24 hours ago, 2.8 trillion parameter model, 1 million context window, really powerful, open source, comes out from China. And also you can use it on kimi.com, Kimiwork, KimiCode, KimiAPI. And it is right up there in terms of performing with Kimi versus Fable 5 versus GPT 5.6 Sol. And I'll show you some comparisons in a second, but essentially you can see that it's doing very, very well on the benchmarks. This is not just like an open source model. This is not like GLM 5.2 where it came out. It's like, oh yes, it's kind of right up there with Opus 4.8. No, no, no. This is Kimi K3 on certain benchmarks outperforming GPT 5.6 Sol and Fable 5 So this is not hype. I'm going to show you exactly what we tested out as well with this.
So on Goldie Bench, we tested out 50 different tasks. And if you check out the quality of this stuff, like it can create really nice 3D games. It's got a nice ambiance and feel to everything that we created. Here's like a full open world game that we created with Kimi K3. Just looks absolutely awesome.
We actually generated like a kind of copy of the style of Skyrim. So we use it for inspiration. And you can see if we refresh the page here, like even the opening video looks beautiful. Like this is fully custom made from K3, less than 24 hours ago. So it didn't take a long time to create, but the outputs of this stuff is unbelievable. It just looks beautiful. And this is like a huge open world game where, you know, the details, the ambiance, the vibe of it, just looks awesome. I can't see any bugs in it so far as well. It just feels really smooth when we look around. It looks fantastic. And we generated this in like what, one single prompt. So it's unbelievable what you can do with it. Now, you might be wanting to k-hounce it before, versus something like Fable 5 and GPT 5.6 in real benchmarks. Let me show you what we tested so far. So we ran a bunch of tests of GPT 5.5, sorry, GPT 5.6 versus Fable 5 versus Kimi K3. And if we have a look, for example, this is the output for like a Minecraft style game from Kimi K3. Super nice, works perfectly. It's even got water there and the graphics work nicely. If we compare that versus Fable 5 and GPT 5.6, this is the output from GPT 5.6, sorry, Fable 5 It looks okay, but it's a little bit buggy in places. And then also the controls are not that nice. And it doesn't have that water vibe that we had second ago. And then GPT 5.6 came nowhere near. Look at this output. It just doesn't look anywhere near as good. Let's have a look at a Dragon Realm game that we created. So I've already shown you this example from Kimi K3. If we have a look at Fable 5, it's not bad, but it just doesn't look as good. It doesn't feel as good. There's not as much detail, et cetera. And the same with the ambience inside GPT 5.6. This actually feels like almost like an AI generational gap between K3 and GPT 5.6 when we compare them. Like the graphics, the detail, the vibe, the colors, nothing is quite as good as Kimi K3 when we're building with it. Now for me personally, like I like to use all these models and I'll have them side by side running inside our agent operator system, which you can get link in the comments description or go to the aripathwarm.com. And with this system, you know, you've got Kimi code, you've got, for example, Claude, you have GPT 5.6 all working alongside each other. However, what we've also done is plugged in Kimi K3 into Hermes agent and it runs tool halls pretty nicely as well. So it responds pretty quickly when you're using it and you can use the coding plan to plug this into Hermes agent to use it identically and then for example, we tested it with a tool hall like forward slash learn, and then plugged in a guide here as you can see, and actually read the page, detailed all the lessons that it learned from that particular topic, and then actually gave us a link to the skill that it created so that it can remember how to use that topic forever. And it works really, really smoothly.

11 more minutes of transcript below

Feed this to your agent

Try it now — copy, paste, done:

curl -H "x-api-key: pt_demo" \
  https://spoken.md/transcripts/1000651996090

Works with Claude, ChatGPT, Cursor, and any agent that makes HTTP calls.

From $0.10 per transcript. No subscription. Credits never expire.

Using your own key:

curl -H "x-api-key: YOUR_KEY" \
  https://spoken.md/transcripts/1000777288762