← Back to search
Hermes Agent Just Did Something That Scared Me
AI News Today | Julian Goldie Podcast · 2026-05-17 · 15 min
Show full episode description
Hermes Agent + Grok OS Upgrade: Real‑Time X Search, Image/Video Gen & Voice (Step‑by‑Step Setup)The video explains how a new Hermes Agent update combined with Grok OS (xAI/Grok via OAuth) enables real-time X (Twitter) search, image generation, video generation, and text-to-speech inside one AI agent system, often at no added cost if you already have an X subscription. It walks through the setup steps: update Hermes, select the XAI Grok OAuth model via the terminal command (Hermes model) and log in, then enable and configure tools in Hermes tools (including Xsearch, video generation, image generation, and TTS), noting that some features may require reconfiguring providers (e.g., switching video from PHAL to X). The script shows troubleshooting by importing documentation so Hermes understands tool calls, demonstrates generating an X trend roundup, images, a short video, and a voice note, and then integrates these capabilities into an Agent OS “studio” dashboard alongside other agents and an Obsidian-based memory context system.
✨ Episode Outline — click any point to jump to it in the episode
Problem solved
AI tools like
Claude,
OpenClaw and
Hermes run in separate tabs, never sharing context or remembering past work.
Benefits
- One dashboard coordinating all AI agents
- Runs locally so business data stays private
- Compounding memory layer via Obsidian Vault
- Built with no code in plain language
- Free to run Hermes on OpenRouter/Nous Portal
Use cases
- Memory vault stored over 1,200 memories after about a month of use
- Obsidian knowledge base holding 1,300 notes feeding agent context
- 3,000 members already building this stack in the AI Profit Boardroom
- Built the entire Agent OS dashboard in one Claude session in about one hour
- Hermes runs free on OpenRouter using Step 3.5 Flash
KPIs / results
- 1,200+ memories stored
- 1,300 notes in vault
- 3,000 community members
- Built in ~1 hour, single Claude session
Tools / build
- AgentOS mission control dashboard
- Obsidian Vault memory integration
- OMI background screen/mic recorder
- Hermes Agent Kanban board
- Claude/OpenClaw/Hermes coordinated stack
📑 Chapters — tap a time to jump there
03:49
Enable X Search Tool
- Enabling the X (Grok) search tool for research
06:56
Personalized Automation Ideas
- Personalized automation ideas drawn from vault context
09:13
Video Generation Setup
- Setting up video generation in the stack
12:50
Image Generation Setup
- Setting up image generation
14:41
Text To Speech Voice
- Configuring text-to-speech voice output
16:02
Why Agent OS Matters
- Why an operating system beats isolated chatbots
19:51
Memory Vault With Omi
- Memory vault powered by OMI feeding Obsidian
21:09
Build Studio Dashboard
- Building the Build Studio dashboard
23:52
Mission Operator Mindset
- Becoming the mission operator, not the engineer
26:12
Wrap Up And Resources
- Wrap up and AI Profit Boardroom resources
Most people using Claude right now are leaving 90% of its power on the table. And I'm going to show you exactly why and exactly how to fix it in one session. This is AgentOS, a full operating system for your AI agents. Claude, OpenClaude, Hermes, all running together, all talking to each other, all in one single dashboard. And by the end of this, you're going to know how to build your own. But here's what I really want to talk about because the dashboard is just the surface, the thing that makes this genuinely different, the thing that makes your AI go from a generic chatbot to something that actually knows you, your business, your goals, that's deeper. And I'm going to walk you through every single layer of it. Let me start with the problem you've got right now. Maybe you've got Claude open in one tab. Maybe you've got ChatGPT in another. OpenClaude running somewhere in Terminal. Hermes doing its own thing in Terminal as well. None of them talk to each other. None of them remember really what you did yesterday. But if you're doing one thing in Claude, it's not going to connect over to, for example, ChatGPT. And you're doing the same stuff every single time, manually copying and pasting between tools, re-explaining yourself, rebuilding context from scratch. That's not really an AI setup. It's just a mess, right? Think about your phone. Before iOS or Android, every app just sat there. They couldn't share data. They couldn't talk to each other. You'd have to do everything manually. Then an operating system came along and connected them all. Now one tab opens maps, pulls your calendar, checks the weather, and routes you before you've even put your seatbelt on. And that's what AgentOS does for your AI tools. So you have one dashboard, one screen, all your AI agents running together, as you can see over here. So Claude, OpenClaude, Hermes, running together, coordinated, sharing context, and working for you. And the more you use it, the smarter it gets. Now I built this in a single Claude session, and I'm going to show you exactly how. So here's what AgentOS actually looks like when you're inside it. You've got a mission control dashboard at the top. You can see every agent, whether it's online, what it's doing, how many sessions are running, and Claude shows up here live. Hermes, it shows up here live. OpenClaude shows up here as your agent gateway. Okay, routing tasks between everything, right? And you can chat with any of them directly from the dashboard. So there's no separate tabs, there's no switching windows. You click the agent you want in the control room, and you're working inside a full interface built around the agent, its model, its provider, its previous session history, its skills, and plugins. All visible, all accessible right now. There's also a Kanban board wired in. So tasks you create inside the chat go into your board automatically with Hermes agent, right? You don't have to manually move things around. The agent just logs it. Analytics are built in too. So you can see how many sessions, how many tool calls, how many tokens you've used, which models you're hitting the most. Peak activity hours, your most productive days. It sounds like a small thing, but once you see it, you realize that you've been running blind. And here's one thing that separates this from every other setup. We've got it running locally, right? Your data stays on your machine. There's no round trips to some server. No sharing your business context with a cloud you don't control. It's faster. It's more private. It's more reliable. And now this is where it gets important. Before I show you how to build it, I want to tell you about the part most people skip entirely. And this part makes this whole thing way more powerful. And it is the memory layer. Right now, inside the AI Profit Boardroom, we've actually got a full 30-day agent OS roadmap step-by-step, how to set up every layer, how to connect your memory system, how to wire OpenClaw and Hermes into your dashboard. So they're running together. And every week on four live coaching calls, we go deep into setups like exactly this one, how to use it in your business, how to get it running about tech headaches, how to make your agents actually useful for generating leads and saving time. And there's 3,000 members already in their building with this stuff. So if you want the full system plus the prompts we use to build it, go to the AIProfitBoardroom.com or the link is in the comments in the description. Back to the memory layer. Every time you talk to Claude, it doesn't remember the last session really unless you manually paste the context back in. Now, it does have a memory, but it's not very good. And we all know that it doesn't track things the way you want them to be tracked with a memory. So every time you run, for example, OpenClaw or Hermes, etc., it has a tiny little understanding of what you do. But it's nowhere near the full context and memory that you require to get good outputs. It really doesn't have much idea on what you worked on yesterday. It doesn't remember everything, right? And that's the ceiling that most people hit and they never get past it. Agent OS breaks through that ceiling with an Obsidian Vault integration. Now, here's how it works. Obsidian is a notes app that stores everything locally on your machine. You can build a full personal knowledge base in it. Your goals, your projects, your team, your clients, your notes, everything about your business. It's just folders and files, but organized in a way that creates a web of connected information. When you wire that into Agent OS, your agents get full access to all this information with all of the notes and everything else, right? And so when Claude is helping you plan your content for the week, for example, already knows your brand voice, your audience, your previous campaigns, what worked and what didn't. When Hermes is running a workflow, it can pull in your SOPs, your client notes, your preferences. But when OpenClaw is calling it a task, it knows your priorities for the month. The agents stop being generic tools and start acting like a team that actually knows your business. And here's the part that compounds over time. Every conversation your agents have gets logged back into Obsidian automatically. So the more you use Agent OS, the richer the memory gets. It builds on itself every single day. I set mine up about a month ago, and right now it has over 1,200 memories stored. Let me show you an example. So you can see it takes notes on everything, even, for example, trips to Japan or people in my life, etc. We can go to the memory section here, and it stores all of this information. So we have 1,300 notes set up, as you can see. So basically, your agents know your goals, your team, what I'm working on, what I struggled with last Tuesday, for example, which projects are live, which clients we're onboarding. I can ask Hermes, based on everything in my vault, what should I automate today? And then if we go back to Hermes, as an example of that, you can see it right here. It actually takes in the information across all of my context from my Obsidian vault, and it comes back with specific, useful, actionable suggestions built around my actual business, right? There's no generic advice. It's just stuff I actually use. And that's the difference between a calculator and a system. Now, I want to walk you through the four layers of what I'm calling the Goldie Mission Stack, because this is a framework for how the whole thing fits together. So the first one is the intelligence layer. So that's Claude. Claude sits at the top. It's the thinking engine. That's full tool access. MCPs connected. Can write and run code. Analyze documents. Plan strategies. Inside Agent OS, you can see Claude's status live on dashboard over here. So one click can open its control room, which means there's no separate tabs. There's no switching between stuff. Claude is just always there, always ready, always wired into the rest of your system. Now, layer two is the execution layer, and that's OpenClaw. And this is important because OpenClaw isn't just another chat user interface, right? It's one of the most powerful AI agents in the world. It's your agent gateway. It can route tasks between agents. It manages sessions. It handles multi-agent coordination. So you want to think of it like a router in your house or your iPhone, your laptop, your TV. They all connect to the internet through the router. OpenClaw is a router for your agents. Everything passes through it. And in Agent OS, you can see exactly how many agents are running, how many sessions are alive, whether the gateway is healthy, all on the dashboard, as you can see right here. So we can look at OpenClaw. We can pull that up. We can have a look at the control room, and we can see what's going on over here, right? Now, layer number three is the research layer. This is Hermes Agent. Now, Hermes Agent can run deep tool calls. It can implement task workflows. It can use skills and plugins to go out into the world and actually do things. If you need a research competitor, if you need to, for example, run a multi-step process on a schedule, pull data from multiple sources, Hermes handles it, and it's free to run. The agent itself is open source, and you can use it with a free API on OpenRouter, like our alpha, or on Noose Portal. They're free APIs right there. We're running it with Step 3.5 Flash, which is free as well. So if you're watching this thinking, oh, it's going to be expensive, it doesn't have to be, right? Now, layer four is the self-layer, and that is the Obsidian Vault, as you can see right here. And this is one that most people never build, but it's actually the most powerful one, right? Because here's the truth. Every efficiency improvement you get from the other three layers, it compounds when layer four is running underneath it, right? Your agents have context. They understand you. They know what you're building, who you're building it for, and why it matters. That turns generic outputs into genuinely useful work. And this layer stays valuable no matter what else changes. New models come out every few weeks. Platforms change. Tools update. But your personal knowledge base in Obsidian, that just gets richer over time. The earlier you build it, the more powerful your agents get. Now, on top of all four layers, there's also one more thing that makes Agent OS work, which is OMI. OMI runs in the background in your machine. Now, you can set up the permissions you want that you're comfortable with, so you don't have to do it like me. For me, I record my screen in my microphone, and it's taking notes on me automatically. Not in a creepy way. I can control the permissions. But what it does is build a continuous record of what I'm working on, who you talk to, and what decisions you made. And it exports those notes directly into Obsidian Vault every day. So you can just click Export right here. So it creates tasks from what I'm working on as well. It looks at what I'm working on day to day. And then it creates tasks, as you can see right here. And it links those tasks to my existing notes. It's all automatic. It's all running in the background whilst I was doing actual work. So when my agents need context, they don't have to ask me what I did yesterday. They already know, right? Now, here's the part I want to address directly as well. Because I know what some of you are thinking. You're thinking, this sounds complicated. This sounds like it's built for developers. I've never set up anything like this before. The dashboard you're looking at was built to work like a phone, right? You see the agents. You click the one you want. You open it. You work. All the code, all the technical complexity is hidden underneath. Your job is to be the operator, not the engineer. And I built the entire thing in one Claude session. You can see an example of that right here. Literally in one hour, we built the whole thing. I typed what I wanted in plain language inside Claude desktop. Claude asked a few questions. I answered them. I built the whole thing. And I didn't write a single line of code, right? If you can text a friend, you can build this. Now, some people will say, I don't need AI to know about me. I just need it to answer questions. And I get that. That's how most people start with AI. But here's what I found. The single biggest performance unlock you can get from an AI agent isn't a better model, right? It's better context. When your agent knows your goals, your business, your voice, your clients, the output stops being generic and starts being actually useful. That's the gap between people who are getting real results from AI right now and people are just still playing around with it. Now, some people say, I already use Claude. That's enough. But here's the thing. Going into Claude desktop and typing a message is a tool. It's not a system. A tool you open and close never builds on itself. It starts fresh every time. Agent OS builds on itself, right? Every session adds to the memory. That's a compounding advantage that grows every single day you run it. Now, some people say, I need multiple AI tools. It gets overwhelming. That's exactly what Agent OS solves. You've got one screen with everything visible, which is great, right? So if we want to use Claude, we just click on Claude over here. If we want to use Hermes, we just click on Hermes over here, right? We can switch between them pretty easily. And it doesn't take a long time. Everything is coordinated. So you're not switching between five apps. You're running one operating system. Now, some people say, I need to be watching this 24-7. You actually don't. You can have your agents running whilst you sleep. You set the direction. You check the results in the morning. You scale what's working. And that's a shift from being the worker in your business to being the operator, running the system that does the work for you. So let me give you a real example of what this looks like in practice. I opened Hermes Agent OS, and it can run tasks in the background, right? So I can say, for example, okay, go off and build this out. And it will just do it in the background whilst I'm doing something else, right? So I can check the dashboard the next day. And everything can be healthy. Just review the outputs. And then I can say, for example, as well, based on my vault, based on my Obsidian vault, what should I automate this week? For example, for my agency, it can pull contacts from my Obsidian notes, my team setup, my active clients, my current bottlenecks. And then I can get a specific list like this with seven different things to improve, right? And the cool thing about this is it's not generic. So it's not like you should automate your email follow-ups. It literally says, based on your agency, your situation, your current client onboarding process, here's what you need to automate, right? Here's the tools to use. And here's the order to use it. Ready to implement, right? So it's like a junior operator who's read every single note I've ever written and knows exactly what I need. Now, here's the forward trajectory on all of this, right? Right now, we're at day one of what operating agents and systems are going to become, right? The version you can build today is already powerful. But as the models get better, as Claude gets more capable, as OpenClaw adds more features, as Hermes keeps expanding its tool library, all of that improvement flows directly into your agent OS, right? The system gets smarter because the tools inside it get smarter, right? And here's the thing most people aren't thinking about yet. We're heading into a world where every serious business operator is going to have a system like this running for them, not just using AI when they need it, but running AI as the infrastructure of how their business operates every single day. And the people who build the system now, who build the memory, who connect the agents, who start logging context today, they're going to have a six-month or 12-month head start on everyone else because the memory compounds, the context gets richer. You can't fast forward that. You just have to start. So here's what to do right now if you want to build this. Open Claude Desktop. Create a new chat. Say, build me a beautiful local operating system dashboard for managing Claude. OpenClaw and Hermes include a mission control view, agent status, direct chat, et cetera. And then it will start building the whole thing. Whilst that's running, go to Obsidian, download it, and start building out your vault. You can create one file called aboutme.md, for example. Write down your name, your business, your goals in the next 90 days, your team, and your main bottlenecks. That's it, right? That file alone will make your agents immediately more useful the moment you set it up. You can also download OMI and get that set up, and then you can start exporting to Obsidian every day. And that's the stack. That's the intelligence layer, right? Execution, research, self-layer. Your Claude, OpenClaw, Hermes, Obsidian. One dashboard, tying it all together. And this is the difference between someone using AI and someone running an AI operating system. The gap between those two things is about to get very large. Now, if you just want the whole setup from me, based on what I've built and all the prompts I personally use, you can get that inside the AI Profit Boardroom. Just go to the classroom, and then you can find it right here. We've also got the full memory system here as well. So if you want the infinite context engine, which is a memory system I've created, you can get that here as well, right? Full guides, full video tutorials, et cetera, right? Now, inside the AI Profit Boardroom, we've built the complete agent OS setup guide, right? Every prompt you need to build this from scratch. A full 30-day roadmap taking you from zero to a fully running mission control, connecting Claude, OpenClaw, and Hermes to your Obsidian Memory Vault. We've got four weekly coaching calls every week. We walk you through agent OS setups like this live. There's real members in there, 3,000 members. Lots of businesses asking questions about their specific setup. 3,000 members. A lot of them already running multi-agent setups and automating real work every single day. We also have a member map so you can connect with people locally who are building the same kind of systems and learn from them in person. And also, we have the full zip file with the actual files that set up are used personally if you just want to give that to Claude and ask it set up from there, right? And the question isn't whether AI is going to change how businesses run. It already has. The question is whether you're going to be running the system or just using the tools. This is a system. If you want to check it out, link in the comments description or go to the AI profile.