**Julian Goldie** (0:00)
Today, I'm going to show you how to build a local powered Hermes engine, which is a team of AI agents powered by local free private AI. And you can basically automate whatever you want with it. So this is something I call the local Hermes agent engine. You can give it a goal, that could be a voice or text. Hermes breaks it down into goals and steps. Then it runs the tools with commands and creates the project for you, and it's built and verified. Plus you can come back to it inside the workspace, right? And the great thing about this is, number one, it's private, so your data is not going to the cloud. Number two, it's free and you can have these agents working as a team building stuff 24-7.
And number three, it is completely running locally, which means you don't need the Wi-Fi or anything like that to run it. So you can give it a goal, it plans, it runs the tools, it builds the thing, every step on your own machine. Pretty amazing. How does this work? So you can see an example right here. This is the agent Kanban Board, local agents. This is a team of local offline agents with Hermes working on a live Kanban Board, separate Kanban Board just for them. And so I could say, for example, build out an SEO blog about OpenClaw, which we've typed in here, and then it assembles the board. So what it actually does is it comes up with the ideas and what it needs to build for this website. Now if we click on, for example, run the team, what that's going to do is start building this step by step to create whatever we want, which is pretty amazing in itself. Now, if you want to indicate like, what does this actually look like in action? Well, you can see them building now. So these are AI agents running locally with Hermes to build and automate whatever we want. And then if we go over to the workspace here, we can see the stuff we built previously. We can also see the prompt. We can open this up full screen. You can see an example. These are just simple examples just to show you what's possible. I literally built this this morning into our agent operating system. And you can see it's now creating the landing page and everything else using our AI agents together. Now we can also open this up. We can give it feedback. We can tell it to refresh it. We can ask it to improve everything that's built. But that's basically running as a team, as a loop together, which is pretty amazing. So you can see some examples of what we've built with them. And each of these just comes from a single sentence. So this was built by a model running entirely on a Mac, no internet, no API costs, just Hermes agents, teams of agents working with local models. Now, if you're running, for example, Llama with Hermes, you can get this set up pretty quickly. So there's a couple of options here. Number one is you can actually go to Llama, just make sure you have Llama running in the background. And then from here, you would go inside whichever model you want to run locally, for example, like this one. And you can just copy and paste the command to run this inside your terminal with Hermes.
However, if you want teams of agents, that's why we've created this Kanban Board here. So we can have a team of agents working together to build or to make whatever we want, right? And so the great thing about this as well is like we can speak to our agents inside this section as well, which is great. So we've got like for every API, we can change the model and we can create a separate agent profile. We can also talk to Hermes here. We have Hermes Jarvis, which is a voice activated version of Hermes. We have a studio where we can generate images, video and voice. Bear in mind, you can generate images for local models like Ernie from Baidu. Then you got your session section and the workspace here with everything you've built with each model broken down, as you can see. So we can see everything that we've built previously with Hermes agent as well. Now, if you go to the manage section here and we go to models, you can also change the model this way. So for example, you can see the LM studio ready to go, which we could run local models with. And then we can also run it with NVIDIA. NVIDIA is another way to run local models. And of course, Ollama. Ollama is another way to run local models. So we've got Ollama ready to go with two different models. Over here, as you can see, you could run GMO4 with this, for example, and you can have teams of AI agents working together. Now, if you want to visualize the teams working together, you could have a Kanban Board here. So if we go back to this system, you can see everything is built now, which is pretty nice. And we can just see everything that was built inside the workspace here, but we can also assign new tasks and get them working together again on new tasks and just keep building these teams out. Right. So picture this, you can have a team running 24-7, building new useful stuff on a Kanban Board. You can see exactly what's built. You can view it inside the workspace, and this is all free, private, and offline.
10 more minutes of transcript below
Try it now — copy, paste, done:
curl -H "x-api-key: pt_demo" \
https://spoken.md/transcripts/1000651996090
Works with Claude, ChatGPT, Cursor, and any agent that makes HTTP calls.
From $0.10 per transcript. No subscription. Credits never expire.
Using your own key:
curl -H "x-api-key: YOUR_KEY" \
https://spoken.md/transcripts/1000773489692