← Back to search
Hermes Desktop + Ollama is Insane (FREE!)
AI News Today | Julian Goldie Podcast · 2026-06-11 · 8 min
Show full episode description
Hermes Desktop + Ollama Update: One-Command Setup, Multi-Agent Workflows, and Free Local Models This episode covers a new Hermes Desktop update that adds one-click integration with Ollama, letting you launch Hermes agents via a simple terminal command ("ollama launch Hermes Desktop") after downloading and updating Ollama from ollama.com. It demonstrates using Hermes Desktop for multi-agent workflows by spawning parallel sub-agents with isolated contexts, switching between agent profiles and sessions, managing skills, messaging, artifacts, files, and settings, and quickly changing models. The script explains why Hermes Desktop is easier than running Hermes in the terminal and highlights benefits of Ollama such as local and free cloud models, privacy, offline use, and fast model swapping (e.g., Minimax-M3, Gemma 4, Qwen, GLM, Nemotron). It also shows connecting Hermes to channels like Telegram, Discord, Slack, and WhatsApp, and mentions AI Profit Boardroom for the broader Agent OS system and coaching/community resources.
✨ Episode Outline — click any point to jump to it in the episode
Problem solved
How a new
Hermes Desktop update lets you run Hermes agents with
Ollama using free local and cloud models in a one-command setup.
Benefits
- One-command setup with Ollama launch hermes desktop
- Run free, private, offline local models
- Spawn parallel sub-agents with isolated contexts
- Manage skills, messaging and artifacts in one place
- Conversations sync across all platforms and channels
Use cases
- Spawn parallel sub-agents to research several topics at once with clean contexts
- Connect Hermes agent to Telegram, Discord or Slack via the messaging section
- Quickly switch models to Minimax M3 or Gemma 4 as a backup model in Hermes Desktop
- Run AI offline on a flight using free local Ollama models
- A 'cancel the outreach' chat from yesterday's agent OS synced into Desktop
KPIs / results
- One terminal command to launch and set up
- Free models: Nemotron 3 Ultra, Step 3.7 Flash, Minimax M3
📑 Chapters — tap a time to jump there
00:00
Ollama Integration Update
00:27
Multi Agent Workflows
- Spawn parallel sub-agents to research topics and run batches
00:43
Connect to Messaging Apps
01:39
Why Use Ollama Models
- Plug in local and free cloud models like Nemotron 3 Ultra, Minimax M3
01:59
Desktop vs Terminal
- Desktop manages everything; terminal hides context and previews
02:51
Settings and Model Switching
- Change model, chat and appearance in settings
- Switch to Minimax or Gemma 4
05:14
Parallel Agents and Contexts
- Parallel sub-agents keep contexts clean across sessions
05:38
Why Run Local Models
- Local models are free, private, work offline and fun
06:03
Recap and Sync Benefits
- Recap
- Conversations sync across all platforms into Desktop
06:55
Agent OS and Community Pitch
- Get the agent OS and join the AI Profit Boardroom
So today we have a brand new update for Hermes Desktop which allows you to use and you can basically integrate your AI agents with using this one-click setup here. So what you do is you make sure you have Ollama downloaded, then you go to Ollama launch Hermes Desktop and from there you can plug it into your Hermes Desktop app. And so from here what you can do is you can run it for multi-agent workflows so you can spawn parallel sub-agents with isolated contexts from the desktop. You can research several topics at once, run batches and stuff. Here's an example. And then this is also designed for using directly with Hermes which makes it super easy. It can use all the same stuff. So what is also cool about this is you can also plug in Hermes agent to multiple channels as well. So if we go to the messaging section here powered by Ollama, we can then connect this to Telegram or to Discord or to Slack or whatever we want using this system which is pretty cool. And so Hermes agent is number one easier to manage using Hermes Desktop but number two you can easily set this up. So how do you get this set up? Basically the first thing you do is go to Ollama.com and make sure you download this as you can see right here. So you can download it with one terminal command or you can click download Ollama. This is free to use. Once you've done that, you're then going to make sure you have it open like you can see here. And if you've already used it before, make sure you've got it updated. Then once you've done that, you can just run this terminal command which is Ollama launch Hermes Desktop. And this will launch your AI agents with Ollama. So it's pretty simple and easy to set up right there. Now the benefit of using, you might be wondering, is that basically you can plug in number one local models and number two free cloud models to it. So with Ollama, they have many different models, for example, like Nemetron 3 Ultra. You've got, for example, Minimax M3 and you can plug each of these models directly into Hermes Desktop. Now you might also wonder, okay, why would you use Hermes Desktop? It basically is better than using the terminal. It's not as good as an agent operating system. Like compare these side by side so you can see the difference. A Hermes agent is decent inside Hermes Desktop. It's much better than terminal, right? And the reason for that is because you can manage everything in one place. If you look at the terminal, the old way of doing this, right, if we go to the terminal here, and we type in Hermes. This is the old way of using Hermes, but look at that. It's just in the terminal. You can't manage anything. You can't see what you've built previously. It's pretty hard to see context and you can't preview anything that you build inside the terminal either. And so what we want to do instead is have some place to manage it. Now, Hermes Desktop is pretty easy to manage one version of Hermes and it's a lot nicer than terminal. So you can manage your skills, the messaging, the artifacts. You can see everything that you build, your images, your files, links, etc. You can also go over to your settings here and manage it here so you can change the model. You can set up the chat. You can change the appearance. And you can run this on local models because Olama is now available for this, right? Personally, for me, I don't like local models that well. But what I do think is good for this is if I want to switch model quickly with Hermes Desktop, I can just switch over to Minimax or I can switch over to Gemma 4. I could use Gemma 4 as a backup model, etc. And that's the way that I would get the most out of it. So that's basically it. Now, if we look at this as well, I'm just going to change this over. I don't really like the setup there. Change that. There we go. There we go. Also, I'd recommend using dark mode inside Hermes Desktop. And that's basically how it works. So that's how to set up with Olama, how to use free local models with it. If you're wondering, okay, which free local models we recommend for using with Olama. So if I was looking through the list here, I'd be like, okay, QEN 3.6, pretty good. GLM 5.1, that's a pretty good model. Minimax M3, obviously frontier level agentic model. It obviously depends on your setup as well. If you're just running on a super lightweight setup, then you can use something like Gemma 4 and you can plug this into Hermes Agent as well. If you need something quick and light, then you can use Gemma 4. Now, this depends which version of the model you use as well. So for example, if you use like the smaller model, 7.2, it's not going to be as intelligent or as powerful as something like 20 gigabytes. So depending on what your setup is, you can run bigger models and they'll be better. But if you've got a worse local setup, then you use a smaller model, but you're going to get less quality out of that. So that's the sort of balance that you have to make between them. And so it's one command and then you can run, I'm at launch Hermes desktop. So it's pretty easy to just plug a new model into this system as well. And some people say, well, I'm not technical enough for local AI. Really it's just one command inside a terminal. So it's easier than it's ever been essentially. And also what you can do is we've got this set up inside the operating system, as you can see. So what we can do is you can go to manage inside your agent, go to models, and then you can change the model over pretty quickly that way as well. So we could always switch to Ollama inside the set of model section here. We could change between all these different providers. As you can see, we could use M3 or M2.7, whatever we want, whenever we want. And also something to note here is like, if you like free models, you'll probably like news portal because they've got step 3.7 flash and also Nvidia and Amitron 3 Ultra for free. So two free models that you can use right there. Now, obviously you can run parallel sub-agents with this as well. So your context stays clean. So Hermes desktop can create parallel sub-agents with isolated contexts. What that means essentially is like you can switch between different agents and use them like you can see here. So at the bottom of the page, we can switch between all these different agents and see our sessions and see what we've created and see what we've built, etc. And then we can have them with different contexts, with different skills, with different use cases, different expertise. And also one of the best things about Hermes was like, it just gets better the more you use it. Now, obviously you might wonder, okay, why would you use local models? Number one, they're free. Number two, they're private. Number three, they work offline. So if you have a flight or something like that, you can still use AI. And number four, it's fun to use. And also even if you have a paid API that you prefer to use, you can always have a local model from Ollama set up as a sub-agent as well. So just to recap in terms of what you gained, right? You learned how to install this in one click using Hermes desktop and Olama. You learned how to create parallels by just clicking between these different profiles. And then you can add agents with different sessions here. So for example, if you look at this one, we can create a new session. But then if we click on this one, we've already got loads of other sessions up here, right? So we can switch between all these different agents. And bear in mind, this is synced across different platforms. So for example, here where I've typed cancel the outreach, that is actually from a chat that I had yesterday with Hermes inside the agent operating system. So all the conversations you have with your Hermes profile across all the different platforms and channels get synced inside desktop as well, which is pretty cool. You also learned how to stop app switching with the agent OS system. And you learned how to get free models and basically how to have one agent with one memory on every surface. Now, if you want my system for the agent operating system and you want to make Hermes part of your operating system, then you can get that inside the AR profit boardroom. Basically what this is, it's a powerful system to manage Hermes in one place. So we've got all the best features plugged into different tools, as you can see. So for example, we can chat with it, we can talk with it, we can use Hermes Jarvis and use it with voice activation. We can generate images and videos and voice. We can create new sessions, manage everything that we've created in one place. We have MCPs, we have the manage section, the control room and the goal mode, along with all of our other agents. Like we just plugged Claude 5 Fable into this. We're an agent group chat with our agents. We've got teams, so we have organizations with four teams of different agent profiles. And then we also have a pipeline where you can generate ideas and then plug those ideas into a system you implement. So feel free to get that link in the comments description or go to the end.com. This is my community for helping you save time and scale with AI automation. You can ask questions inside the community. You can see inside the classroom all of our new lessons. You can see the agent OS system and when it was last updated with the zip file to install as well. Inside the calendar, you can jump on weekly coaching calls. We have four weekly coaching calls. In the map, you can meet people in your local area. And that's all available inside the AI Profit Boardroom. Link in the comments description or go to the AIProfitboard.com. Thanks for watching. Cheers. Bye-bye.