**Julian Goldie** (0:00)
There's a brand new free model with Hermes Agent. I'm going to show you exactly how to use it, how to set up and how it works. This is Nemotron Ultra 3, and you can just see that they've announced it. So Nemotron 3 Ultra is available for free with Hermes Agent on the News Portal, which means you can start building with it and create an awesome stuff with it. So let me show you exactly how it works, what it means and whether you should actually use it. So this is the new model announcement. And literally Nemotron 3 Ultra just dropped today. So this is a brand new model you can start using straight away. And one of the biggest improvements with this is that it's five times faster, right? So it's designed for tasks and it's five times faster, which is great. Now, if you want to use this, you can go to your Mission Control, like you see right here, and then inside the model section, you can change it. So you can see, for example, we've selected News, which is News Portal, with Nvidia Nemotron 3 Ultra. So this is free for two weeks and you can have that as a main model. Now, once you've done that, if you go over to the chat, let's just test this out and you can see here it's working. So we can start using this straight away. Now, if you're wondering, okay, what is the importance of this? How does it work, etc. So this is a 550 billion parameter frontier model, and it's open source, and it's designed and built for long running agents. News Research is giving it away free for two weeks on News Portal. You can plug it into Hermes, you've got a frontier grade coding and reasoning brain, running your AI agents for free.
Here's an example of what we built with it, just for fun really, but you can see here basically we asked, and created a living galaxy that's been built by Nemotron with Hermes Agent. You might be wondering how to see it perform. So this is a frontier model. It's designed specifically for AI agents like, for example, Hermes. Most models have really built for chat. Nemotron-3 Ultra is built to work, so it can run for hours as an AI agent, planning, using tools, recovering from failures, deciding what to do next. Now it's a giant mixture of experts models of 550 billion total parameters, but it only fires the actual steps that it needs for each. So it's faster, right? And so it's a mixture of experts model. What that means essentially is if you ask it for a task, it will just use certain parts of the model, not the whole 550 billion parameters. It's also five times faster, which makes it a lot quicker. And obviously with AI agents, like you're going back and forth with it, you're asking it to do stuff, etc.
So Nemotron-3 Ultra is great for that. That also means it uses 30% less tokens on agentic tasks and it's open source, which means that if you had an amazing setup, you could run this locally. But if you don't have an amazing setup, no problem, you can set up News Portal, like I've shown you today. And it's a joint release from Nvidia and News Research, post trained specifically for agent setups, which is exactly why it slots straight into this stack really nicely. It's an engine built to run AI agents for hours without losing the threat. Now, how do you set it up? So three simple steps.
You sign up at News Portal, which is free. Then you connect it inside your terminal. Or I would recommend these days it's easier just to go to the manage section of your AI agent, go to the dashboard, go to models and then just change over there. Then you pick the model Nvidia Nemotron 3 Ultra. You also get a lot of other free models inside News Portal. So if you haven't checked this out, it's a great way to just avoid using and get free models. So step 3.7 Flash is also available for free. And then you can use it. You can chat or you can hand Hermes along task. And it now runs on a 550 billion parameter frontier brain for free for two weeks.
We've already put it inside the agent operating system and it works pretty nicely. So this is an example of what we created with it. We also tested it on some reasoning tasks like you can see right here. And it performed pretty well. So it worked the whole thing. Now, interestingly, here's something that makes it really good. So it thinks in terms of long horizons so it can plan, it can do multi-step tasks. And some people say 550 billion Frontier model is too expensive, but this is free. You might say free models are weak models. But actually, this is an open Frontier model from Nvidia and it's tuned for AI agents. And you might also say, I don't want to swap the model inside Hermes because it's too much work. But I've shown you how quick and easy it is to do. You just go to Manage, then you go to Models and then you change out. This is because of the new Mission Control dashboard that just came out from Hermes Agent yesterday. So with that, everything is easier in terms of changing it. You don't need the terminal anymore, which is great as well. The other thing I think this could be very powerful for is the Goal Mode. So with Goal Mode, obviously Hermes Agent, it works at a task. You give it a big task and a big mission. And then it will have a go at that task for 20 turns before stopping. And there's a judge that basically analyzes each turn, each attempt from Hermes Agent to see if it's actually been completed. If it's not completed, then Hermes Agent just keeps going until it finally gets a job done. And the judge judges it's actually completed. And so because this is a big model and because it is a model that works on log horizons, you could give it a task, like, for example, build a website or complete my SEO strategy. And it could just go for 20 turns or 50 turns or whatever you say and complete that for hours without you having to do anything in between, which is pretty powerful. It's also designed to excel at complex tasks like coding and deep research. So long running agents can spend the time planning, using tools, recovering from failures, deciding what to do next. You might say, how does it perform in terms of the benchmarks for this sort of stuff? So if you look at the benchmarks here, this is Nemotron 3 Ultra versus Glm 5.1, KimiK 2.6 and Qwen 3.5. And you can see in terms of agent productivity benchmarks, it's right at the top there, outperforming Glm and Qwen 3.5. In terms of instruction following is everything. If you look at long context, it's outperforming everyone as well. And then also, if you look at long horizon planning, it's outperforming KimiK 2.6 and Qwen 3.5. It's not outperforming Glm 5.1, but that's just something to bear in mind. So AXCEL's professional work tasks, long context, instruction following and agent productivity. And you can see here, they actually post trained it for AI agent harnesses. So for example, OpenCore is another model you can plug this into. It will be pretty powerful for it. And it's available on Hugging Face as well. If you wanted to host it locally, you can see it right here. So that's basically how to use it, what it is, how to get it for free, how to run Hermes Agent for free, how to set it up as well. If you want to get my full agent operating system for Hermes Agent, OpenCore and everything else, you can get that inside the AI Profit Boarding. Link in the comments description or go to the AI Profit boarding.com. This is my AI Automation community that's focused on helping you save time, grow and scale with AI automation. Inside the community, you can ask questions, get help and support. I answer the questions personally inside here. Plus, there's always people online 24 7, which means you get help from everyone inside the community. Inside the classroom, you get access to all of my new daily tutorials and best training. So if you're a complete beginner, you can go from beginner to expert here. If you want new daily updates, you can get out of here. And also, if you want my agent operating system for managing all this stuff, you can get it and we update this daily. Plus, you get a zip file you can just quickly install. Inside the calendar, you get four weekly coaching course. You can get on these coaching course, get help and support in real time. We also have a map where you can meet people in your local city using AI automation and AI agents like Hermes and OpenCourt too. And this is all inside the AI Profit Boarding. Hope to see you inside there. Cheers for watching. Bye bye.
Try it now — copy, paste, done:
curl -H "x-api-key: pt_demo" \
https://spoken.md/transcripts/1000651996090
Works with Claude, ChatGPT, Cursor, and any agent that makes HTTP calls.
From $0.10 per transcript. No subscription. Credits never expire.
Using your own key:
curl -H "x-api-key: YOUR_KEY" \
https://spoken.md/transcripts/1000771312048