Si has estado atento a los últimos episodios del podcast, ya te habrás dado cuenta de que estoy completamente enfocado en exprimir la inteligencia artificial local y el software libre. En concreto, hay dos herramientas que se han convertido en mis compañeras inseparables de fatigas en el día a día: OpenCode, que me ayuda a programar de una forma increíble, y Hermes Agent, un asistente digital del que hoy te lo quiero contar absolutamente todo. El dilema de la instalación: ¿Docker o en tu propia máquina? Como ya me conoces, sabes bien lo mucho que me gusta a mí levantar "al rico contenedor" y solucionar cualquier despliegue con Docker. Sin embargo, en mis pruebas con Hermes Agent he preferido dar un paso atrás y realizar una instalación directa sobre el sistema operativo, utilizando un entorno virtual de Python. El peligro de la ventana de contexto y la sangría de tokens Aquí está uno de los grandes secretos que casi nadie te explica al principio. Cuando ejecutas el asistente de configuración inicial de Hermes Agent, te entran ganas de activar absolutamente todas las características que te ofrece: herramientas de visión, utilidades del sistema, navegación web, traducción... ¡todo suena fantástico! Pero hay una trampa invisible en la que es muy fácil caer. El superpoder de los perfiles aislados (Profiles) La solución definitiva a este problema de consumo y rendimiento tiene un nombre: perfiles. Hermes Agent te permite crear tantos perfiles aislados como consideres oportuno. Modelando el Alma y la Memoria de tu Agente En el podcast te detallo cómo dar personalidad a tu agente a través del archivo de alma. A mi asistente personal, que he bautizado como Chloe, le he configurado un tono sarcástico, irónico y burlón. Me encanta interactuar con ella de esta manera porque rompe completamente con la clásica respuesta robótica y aburrida de otras inteligencias artificiales comerciales; se siente como hablar con un colega de verdad. Eso sí, te doy pautas para redactar este archivo con cuidado, ya que un "alma" demasiado extensa también te comerá espacio de contexto útil de forma innecesaria. Ampliando fronteras: MCP, Telegram y automatizaciones automáticas Por último, abordamos el fantástico protocolo MCP (Model Context Protocol), que nos permite dotar de "manos y ojos" a nuestro agente. Y para rematar la jugada, la integración con Telegram y Matrix. Es una auténtica delicia poder ir caminando, mandarle un audio desde el móvil a mi bot de Telegram, que este use Whisper en local para transcribir mi voz, procese lo que le pido y me conteste con otro audio sintetizado a la velocidad que yo le he configurado de antemano. Todo ello combinado con tareas programadas (Cron) y un tablero de Kanban interno con el que el propio agente se organiza y ejecuta flujos de trabajo de forma completamente autónoma. Te invito a que te prepares un buen café, te pongas los auriculares y disfrutes de este viaje de configuración avanzada de 0 a 100. CAPÍTULOS DEL AUDIO:
✨ Episode Outline — click any point to jump to it in the episode
Problem solved
Basic Hermes Agent tutorials skip the real configuration steps needed for a useful, token-efficient assistant.
Benefits
Isolated profiles keep memory and skills uncontaminated
Custom skills teach the agent repeatable workflows
Creando "Skills" personalizadas y configurando API Keys
00:08:15
Perfiles aislados (Profiles): Qué son y por qué los necesitas
00:11:00
Cómo clonar y gestionar tus perfiles sin romper nada
00:13:35
soul.md: Diseñando el "Alma" y el tono de tu asistente
00:15:28
memory.md: El gran desafío de la memoria y el RAG en Rust
00:17:38
Expandiendo capacidades con MCP y conversión de voz
00:20:47
Llevando tu agente a Telegram con Cron y Kanban integrado
00:27:18
Reglas de oro para optimizar tu contexto y despedida
🌐 This transcript was automatically translated to English from the original.
Hello, I'm Lorenzo and this is Atareao con Linux, episode number 807. If you've followed the previous episodes a bit, especially what I'm talking about lately about intelligence models, about artificial intelligence in general, you'll have already noticed that I'm very focused on Hermes Agent. Actually, I am currently using two fundamental tools. On the one hand I am using OpenCode for programming and on the other hand I am using Hermes Agent as an assistant. And an assistant who is helping me with practically anything you can imagine. From reviewing the episodes or rather, reviewing the Atareao.es tutorials, which I will soon begin to publish, or rather, to republish the previous ones and to publish new ones, all of this helped with Hermes, as the fundamental part of programming. As you well know, one of the things I like the most is programming, so I am taking advantage of this help that I have from artificial intelligence to program using, as I said, OpenCode. But I'll talk to you about this later because I want to dedicate some specific episodes again to OpenCode with everything I've been getting. The reality is that if you follow the tutorials on the internet a little about installing Hermes Agent, they usually tell you the most basic things. And there are some details that people or in general do not mention. It is very likely that if you review all the episodes that are out there or all the YouTube videos that there are, you will put it all together and find more or less what I am going to tell you here. But hey, regardless of whether it is this or not, I want to give you a little insight into what I have done with Hermes Agent, how far I have come and what things I am implementing. And I start a little with the installation. Installing Hermes Agent is not much of a mystery. You can install it directly with Pip. You have two options, which is either install it using Docker or, as I said, install it directly on the machine. Until now what I have done has been to install it directly on the machine for a simple reason. Because the time I tried to install it in Docker, then when it came to configuring tools I had problems. So, since I want to get the most out of it, at least for the moment what I'm doing is, as I said, installing directly on the machine. Once you have it installed on the machine, simply by running Hermes Chat you can start talking to it. But what you are going to find is a generic agent, without personality, without memory, without connections, with absolutely nothing. The basic installation is what you are going to find. There's not much, but there's a lot more that can be done. As I told you, or as I was telling you, once you have it installed you already have the Hermes theme. You can install it with Docker. You configure the providers. You can also set up a local provider. Right now I am using it directly with Open Router and with the Deep Seek model, because it is the most economical of all. Although I'm blowing it up, because as I say, I'm reviewing all the articles on atareo.es, which is becoming a little crazy. Once you have everything installed, the essential thing is to run Hermes Setup. The only thing it does is follow a configuration that you will go through those things that you want to activate or deactivate or whatever you consider. With Hermes Chat you start the first conversation and then there is a very interesting tool that is Hermes Doctor, which is really your friend. And that's what's going to tell you what's not working well. And it will also help you determine what the problems are and how you can solve them. In that sense, when you run Hermes Setup, I would recommend that you install everything the first time and try it. Well, more or less everything with a little knowledge. I mean that the tools that are for macos, if you are using Linux, do not install them. It doesn't make sense. But if not, install everything without any problem. But with one observation that I am going to tell you. And everything you install, everything you configure at the beginning, all the tools you configure, what they do is use up the context window. And you will quickly realize it because you will observe how the tokens are eaten as if there were no tomorrow. It is a very exaggerated thing. In fact, what I have created has been, and I am going to tell you about this below, some profiles. I'll tell you this. The directory structure that you will find directly with Hermes is a very simple structure. A configuration called config.yaml. Then, a file that is soul.md, which is the one that gives you the personality of the people. A memory.md, which is where it stores the memory. And we will see this a little later because we have to turn this around. Another point where the keys are. And then there are several directories. The sessions directory, which is where the conversation history is saved. The skills director, this is essential because this is where he has all the skills, all the abilities installed. The cron, the cron directory, which is where you have the scheduled tasks. And the plugins directory, which is where the active plugins are. The skills directory is essential because what it will allow you is for your agent to learn. And not only that he learns, but that you teach him what he has to do and how he has to do it. In this sense, to review all the tutorials that I have on AtariAo.es, well, I have not reviewed all of them, but in the ones that I have reviewed using Hermes, what I have done is create a specific skill. A specific skill where I have indicated step by step what you had to do. What do you have to check? You have to add references to everything you indicate. How to review the documents. What parts does it have to include? What parts do not have to be included. Anyway, all this type of things is with skills. Just like this, in a previous episode I also told you that when I went for a run, what I was doing lately was taking a screenshot of exactly the training results and passing it on to him. What I have done is create a skill so that it knows exactly the data to extract from that screenshot. Which is much more to the point. Golden rule. The API keys, everything that is keys, tokens, etc., etc., have to go in a .env file. I don't really touch the .env file or the .config.yaml file at all. I don't touch any of these files. Forget about touching them. My recommendation is that you don't touch anything, but that you directly tell Hermes or your agent, whatever name you have given it, in my case I have named Chloe, tell your agent exactly what you want to do, where to keep it, if necessary, tell them. But it's not worth touching anything because that's what your assistant is already there for. Ok, at this point, you have the possibility to create commands. I mean, sorry, the possibility of creating... Yes, sorry, exactly. You have several commands to work with. One that is Hermes Chat, which is directly to talk to your agent. Hermes Doctor, we have seen him a little at the beginning, which is to make a diagnosis. Hermes Logs, which is to see all the logs, to... you can extract from there if you have any problems. Hermes Skill List, what it will do is list all the skills. And Hermes Config Show, what it will do is show you the active configuration. With these commands you can now survive. With this you can start things. But this is where the important part comes, in the profiles. Because you can do... That is, you can create as many profiles as you consider. With the great advantage that what Hermes does is make each of the profiles independent. With its own configuration, its own memory, its own skills and its own personality. Basically it's like having multiple assistants, but on the same machine. Each one specialized in something different. And with the great advantage that depending on how you configure it, the memory of one profile does not contaminate that of another. That is, you can have one profile to work, another profile to chat, another profile to research. This is as you want. And this is where I link it with what I mentioned practically at the beginning. When you start with all this, when you start with the profiles, it is very interesting to create a profile just to talk to him. Without doing anything else. Simply to tell you that... Well, for example, a profile like the one I made. A profile for simply when I come back from a run or when I do normal operations, a profile that has nothing. Because? Well, because what I have done is deactivate all the tool sets that you can configure at the beginning. The vision, I take all that away. And the reason for removing all this is because once you have removed it, the great advantage it has is that in the context window what you send each time is much smaller. Simply for that. Because if not, all the skills you have loaded, all the MCPs, all the plugin information, all the tool information, all of that is sent every time. And all of that consumes a lot of tokens. And then it doesn't make sense to have specific skills to do research when all you want is to talk to him. Then you just switch profiles, switch to the chat profile and talk. You tell him your things and he keeps his memory there without contaminating the researcher's memory. In addition, the researcher will have a series of skills and a series of MCPs specific to him, so that he can only investigate. Thus, to create a profile it is as simple as doing Hermes Profile, create chat and you already have the chat profile. And in the same way you can create a researcher, a writer and configure each of them with those tools, those tool sets that you need to work. And in this way you will optimize, as I told you, the consumption of tokens. Each profile generally generates its own command. That is, chat, chat, researcher, chat, writer, chat. For now I don't know why, I don't know if because I have an old version or whatever, the commands are not being created. So what I do is Hermes, that is, I execute Hermes, chat and script script profile and the name of the profile that I want to execute. Then, if you want to clone any of the profiles, because it basically has what you need, but you want to add some specific skills, you can use the clone option, which all it does is clone your soul, the soul of your agent. If you want to clone everything, simply clone all. And if you want to clone from a specific profile, you can also do a clone fraud. All these types of operations, as you can already imagine, you will find directly in the episode notes. So you don't have to worry about anything. Useful to experiment without any worries, you clone your productive profile, try all the changes you want and if something fails, you go back to the original. As I say, my recommendation is that if you want one simply to chat with him, I would remove absolutely everything. If you want, maybe you leave the text speech or the speech to text or whatever you consider, but I would leave it as clean as possible and then I would add little by little those things that you consider it necessary or that you need. In this way, regarding profile management, you have a series of commands such as profile list, which shows you all the profiles that you have enabled, enabled, sorry, the show chat that shows you all the details of your chat profile, configure your profile to use it by default, rename a profile, remove a profile, you can do all these types of operations perfectly and they are very convenient, especially when working with profiles. Within the .hermes directory you have another directory that is profiles and within that directory you will have all your profiles identified with their configuration, their soul, their memory. What's more, for example, for a researcher, maybe you need a certain... a certain... now I'm not going to get it... certain special characteristics that you don't need in the others. Well, you will have it there perfectly identified. Well, about soul and memory. Regarding soul and memory, basically what soul does is define who the agent is. It is not a prompt. In the end, if you have already taken the great leap of moving from GPT chat or James Minay to something with more substance, such as Hermes, soul, what we say, is the basic prompt, what gives the personality, the tone, the limit to your agent. In fact, I have already told it on more than one occasion, my agent, because I have told Chloe to be ironic, sarcastic, acidic, mocking, so that she has that personality and that she finds me as if there was someone real behind her. The great advantage of soul is that it is injected into every interaction. It is not a context that is lost. That always goes. With which you have to take into account a fundamental detail. And if you make a very large soul and have a very small context window, it is very likely that some of the information will be lost. This is why it is interesting that you define a soul as specific as possible. For example, define the soul voice. I want short, blunt answers, two sentences better than a paragraph, dry, sarcastic humor, your work, natural conversation and your rules. Define the rules you want. For example, only in Spanish. What are the dangers? A little bit of what I was telling you. If you have a soul that is too long, the agent ignores parts. If you have too generic a soul, your agent has no personality. And if there are contradictions within your soul, your agent's behavior is going to be completely erratic. And then there is another part which is memory. This is very important. I'm doing some tests, and here I'm already getting into trouble, I'm doing some tests, creating my own memory. A memory that I have called Hmemory, from HermesMemory, where I have defined a backend in Rast with a Rack so that it can do semantic search on the database. I'm testing it. In fact, I have created a different memory. I'm trying one, I'm trying another, to see how it works. There are many. I mean, there are many plugins that you can install directly to Hermes to work with it. And my recommendation is that you install some memory. Working with Smartdown, well, it's good, but it has many limitations. In fact, memory has a limitation of 2,000 characters, I seem to remember what it was. So, from there you have to jump to other memories. It is worth using one of the plugins instead to get the most out of your agent. And in the end, one of the most important parts of your agent is the issue of memory. If your agent doesn't remember things, he's a very stupid assistant in the end. And I tell you that he is a very stupid assistant because you have to keep repeating exactly the same thing over and over again. And this makes absolutely no sense. In fact, right now I have not only given it semantic search, but I have added tags to it, I have added a series of things to see how it works. But as I say, I'm trying. It has a layer where you remember recent conversations, a consolidation layer, a deep layer. In short, he has different things to play with him. But currently I wouldn't recommend you try it because it's a bit crazy that I'm trying with three layers of memory. You can imagine that I don't know. I would basically recommend that you try everything there is in memory and choose the one that best suits your needs. Then, the next block, once we have talked about memory, is the MCPs, the speech text and the speech to text. The MCP, I have already talked about this in other episodes, are the Mobile Context Protocols, which are basically tools for your agent. Each MCP is a server that exposes tools that Hermes can use, such as searching the web, accessing YouTube, querying databases, reading files, interacting with APIs. You can do all that. And not only that, you also have native Hermes MCPs. In my case, I have created my own MCPs that are available in the GitHub repositories, which will allow you to do web searches, but also with very specific tools where it will force you to install SearchNG and then also Invidius to do on YouTube. The great advantage of these two is that they are free because you are going to set up your two search engines, and I already dedicated an episode to this, and what is brutal is that it allows you to search on YouTube and find practically everything. Not only this, but I have also added, an MCP is now available so you can directly ask your agent what to do, that is, you can connect your agent directly with atareo.es to ask him questions. So, that's as far as we've come. Setting up an MCP is quite simple. It's simply Hermes, MCP, AT MCP, or something like that, I don't remember exactly, you'll find it in the episode notes, and there you define exactly all the MCPs you want and the configuration you need. Both HTTP type MCPs and STDio type MCPs, or any of them, you can do it. And then there is the speech text part and speech to text, which is basically converting from text to speech and from speech to text. There is a tool that is a provider, which is Edge, which comes by default, which is from Microsoft, and what it allows you to use is for free. What I created was my own tool because I wanted to modify the speed of speech, as well as the pitch. Those two things. And I have created a tool that is also available, of course, on GitHub, that you can download and use. Use cases. Well, here you can already imagine it. The first and most obvious use case is for it to read a long article to you while you cook, for example, or even to tell you the recipe you have to cook, to record the weather summary, to send you an audio with the news of the day. All of that is worth it and it is very interesting. And in the same way you have speech to text, that is, you directly speak, for example, to Telegram and what it does is convert that audio into text. For this it uses a provider which is Whisper and Whisper does it directly locally. You don't need anything else. You can choose the model and how it works. You can also use, of course, the OpenAI API, but hey, there you have to add the API Key and start spending. The jewel in the crown. The jewel in the crown and you have surely seen this in thousands of videos is how to configure it to use it with Telegram. I am currently using it with Telegram and Matrix, but I have to tell you that what works best so far is with Telegram. Although I tell you, I also have to tell you that I have changed the Matrix backend because the one I have currently been using, well, Conduit, was giving me some problems and I wanted to go to the newer ones. The great advantage of having Telegram or Matrix or whatever you want is that you directly speak to the cell phone as if it were another contact, it does its speech to text, understands what you are saying and directly answers you, either by voice, exactly the same as you, or directly by text. That's as you need. Another super interesting feature is Chrome. You can create Chrome to, for example, give you a summary of the weather for today. That every day. Or directly send you every day a summary of the most important news in relation to artificial intelligence. You can do all this and you don't need to go directly into the Telegram settings or the settings of whatever you want or the Hermes settings. You simply have to tell Hermes to schedule a task that is to do what he has to do and send it to you every day on Telegram. You tell him like this, by voice, and he creates the Chrome that needs to be created for you and does what it needs to do without you having to program anything at all. And I have to tell you how fundamental skills are, Chrome, which is a bit of what got you ahead, and Kanban. The skills at the end of the day and the most important thing and I have already talked to you about this in an episode and above all I have referred you to the Web Reactiva website of my friend Daniel Primo where he explains perfectly how the skills work. Well, the great advantage of all this and what I am using, for example to review all the documents that I have made, is to clearly specify what he has to do and how he has to do it. Insist that it is essential that he add references to all the chapters of this tutorial and he will tell you exactly that, review, review this tutorial or review this chapter of the tutorial, if you find any errors, correct them, add references. You say, apply this skill to all the docker tutorial documents that you find in atareado.es. He downloads each of the chapters of the atareado.es tutorial. He applies the conditions that I have said and leaves it ready to produce a skill. It is really very simple and I have already commented on this in other episodes. In the end it is a name, a description and a version and it has the great advantage that Hermes discovers the skills completely automatically when Hermes starts, he reviews all the skills and he already knows when or more or except when you have to use each skill to do what you have to keep in mind that it is not deterministic, that is, it depends on Hermes deciding to use the skill, so if you give a very clear description of exactly what the skill is for, when with examples of when you have to use it, etc., etc., it is very likely that in those cases the skill you need will be found. Some of the skills I am using is one, for example, which is avoid AI writing, which eliminates AI patterns from the content. to the tutorials and has a task or podcast script that what it generates is the script of a certain episode, I give it the conditions, I review it, in short, they are interactively, how we are evolving and how we are generating that script, but you have some basic conditions there to Let him start working as I tell you. Apart from these skills, I have created a dozen specialized skills to review certain chapters of the tutorial. Then what I have told you about the programmed tasks is that you simply have to tell him the task you want and he will program it without you doing anything. You can tell him that every day he looks for the latest news on the Internet. What he is going to do is connect to the Internet search MCP to extract the latest and with that generate a document that he sends to you every day if you want to tell him. send me a document in PDF because he sends you that document in PDF every day and then there is another super interesting thing which is the Kanban, what Kanban does is basically it opens this to you, you have to, well, it can be consulted on the web, you don't have to see it directly on the site, you see it on the site and directly what it shows you is tasks, all the tasks but in Kanban mode of to do doing and do those three columns with those three columns you are going to see all the tasks that you have to do and them all those tasks that you have registered in the Kanban will be launched as they are needed, in fact, there may be tasks that depend on other tasks and until a certain task is completed, others will not be launched. All this Kanban stuff is super interesting, it works super well and then apart from this you have plugins, plugins that I have already told you a little about, for example, in addition to the one I am implementing, you also have eight, you have memzero super memory and then you have others, such as to generate images, others of metrics and traces of observability, you have different tools and options, you also have them. For visual things, well, some configuration tips, I would tell you that the most important thing, of course, is the fewer tools you have configured, the more space each toolset that you do not use has to do things is the context that you are occupying. unnecessarily, with which you create the profiles that you need exactly with the right tools, neither more nor less. I would say that rather less and as you require more tools, the more toolset you incorporate them and in the same way, if in a certain profile you are going to stop using skills, you either uninstall them or delete them or save them or archive them so that those skills that do not give you absolutely anything are not constantly being loaded and this is a little bit of what I wanted to tell you in this episode a little bit of what It is the installation like this, above all, a first chat, a structure, then I have talked to you and I think it is super important the profiles that, as I told you, were or are isolated agents that have different personalities or you can configure them with different personalities, then the soul.md file, which is a super important file and that gives the personality to your agent. In the fourth blog that I have told you about the mcps, the test to speech, the speech to text and the telegram part, and finally I have been talking to you about the skills, the cron, the kanban, the plugins, in short, all this type of things that are super interesting to me, the most important thing is that you program it, that is, that you dedicate time to it, start using it and little by little you realize that there are certain things that are more interesting to configure in one way or another according to your needs. This is why it is super interesting to create different profiles depending on what you are going to do with each one of them and nothing more. This is a bit of what I wanted to tell you in this new episode, in this episode number 807, simply a configuration of Hermes from 0 to 100 with all the details so that you can get the most out of it as much as you want, I have to dedicate more episodes to Hermes but it will be much later because I want to talk to you a little about Kanban, show it to you in more detail and also see how the memory evolves, the memory is that I have integrated into it. or where I can also do semantic search and more detailed search and little more to tell you, remind you that this is good, remember the first thing that if you can, a rating on Evox Apple Podcast Spotify, a message, a comment on YouTube is fantastic, especially to tell me if you like it, if you don't like it, if you prefer me to do other types of things, I am completely open, remind you that this is a podcast from the usual suspects podcast network where you can find fantastic and wonderful podcasts, you can participate in that network by entering Telegram in Wintablet Info and the In the same way, if you want to participate in the busy with Linux group, you can always search for busy with Linux on Telegram and there you will find us talking about Hermes, Rack, MCP, Docker, Podman or anything else related to the world of open source and little more than to say, remind you that life is two days and one has already passed, so enjoy as if there were no tomorrow, that's how it can be with Linux and in this case with Hermes, best of all, greetings and see you next Thursday, see you later. Thank you!
Hola, soy Lorenzo y esto es Atareao con Linux, episodio número 807. Si has seguido un poco los episodios anteriores, sobre todo esto que estoy hablando últimamente de los modelos de inteligencia, de la inteligencia artificial en general, pues ya te habrás dado cuenta que estoy muy enfocado en Hermes Agent. Realmente, actualmente estoy utilizando dos herramientas fundamentales. Por un lado estoy utilizando OpenCode para lo que es la programación y por el otro lado estoy utilizando Hermes Agent como un asistente. Y un asistente que me está ayudando prácticamente a cualquier cosa que te puedas imaginar. Desde revisar los episodios o mejor dicho, revisar los tutoriales de Atareao.es, que dentro de poco empezaré a publicar, mejor dicho, a republicar los anteriores y a publicar nuevo, todo ello ayudado con Hermes, como la parte fundamental que es a programar. Como bien sabes, una de las cosas que más me gustan es programar, así que estoy aprovechando esta ayuda que tengo de la inteligencia artificial para programar utilizando, como te digo, OpenCode. Pero de esto ya te hablaré más adelante porque quiero dedicarle algunos episodios concretos otra vez a OpenCode con todo lo que he ido sacando. La realidad es que si sigues un poco los tutoriales que hay en internet sobre la instalación de Hermes Agent, pues normalmente te cuentan lo más básico. Y hay algunos detalles que la gente o en general no se mencionan. Es muy probable que si revisas todos los episodios que hay por ahí o todos los vídeos de YouTube que hay, vayas juntándolo todo y encuentres más o menos lo que te voy a contar aquí. Pero bueno, con independencia de que sea esto o que no sea esto, yo quiero darte un poco el enfoque de lo que he hecho con Hermes Agent, hasta dónde he llegado y qué cosas estoy implementando. Y empiezo un poco por la instalación. Instalar Hermes Agent no tiene mucho misterio. Directamente con Pip lo puedes instalar. Tienes dos opciones, que es o bien instalarlo utilizando Docker o bien, como te digo, instalarlo directamente en la máquina. Yo hasta ahora lo que he hecho ha sido instalarlo directamente en la máquina por un simple hecho. Porque la vez que lo he intentado instalar en Docker, luego a la hora de configurarle herramientas he tenido problemas. Entonces, como lo quiero exprimir al máximo, por lo menos por el momento lo que estoy haciendo es, como te digo, instalando directamente en la máquina. Una vez lo tienes instalado en la máquina, simplemente con ejecutar Hermes Chat ya empiezas a hablar con él. Pero lo que te vas a encontrar es un agente genérico, sin personalidad, sin memoria, sin conexiones, sin absolutamente nada. La instalación básica es lo que te vas a encontrar. No hay gran cosa, pero se puede hacer mucho más. Como te digo, o como te estaba comentando, una vez lo tienes instalado ya tienes el tema de Hermes. Lo puedes instalar con Docker. Le configuras los proveedores. También puedes configurar un proveedor local. Yo ahora mismo lo estoy utilizando directamente con Open Router y con el modelo de Deep Seek, porque es el más económico de todos. Aunque yo lo estoy reventando, porque como te digo, estoy revisando todos los artículos de atareado.es, lo cual se está volviendo una pequeña locura. Una vez lo tienes todo instalado, lo fundamental es que ejecutes Hermes Setup. Lo único que hace es seguir una configuración que vas a ir pasando por aquellas cosas que quieres activar o desactivar o lo que tú consideres. Con Hermes Chat empiezas la primera conversación y luego hay una herramienta muy interesante que es Hermes Doctor, que es realmente tu amigo. Y es lo que te va a decir qué es lo que no funciona bien. Y te va a ayudar también a determinar cuáles son los problemas y cómo puedes resolverlo. En ese sentido, cuando ejecutes Hermes Setup, yo te recomendaría que la primera vez instales de todo y vayas probando. Bueno, de todo más o menos con un poco de conocimiento. Quiero decir que las herramientas que son para macos, si estás utilizando Linux, no las instales. No tiene sentido. Pero si no, instálalo todo sin ningún tipo de problema. Pero con una observación que te voy a decir. Y es que todo lo que instalas, todo lo que configuras al principio, todas las herramientas que le configuras, lo que hacen es gastar ventana de contexto. Y te darás cuenta rápidamente porque observarás cómo los tokens se los come como si no hubiera un mañana. Es una cosa exageradísima. De hecho, yo lo que he creado ha sido, y esto te lo voy a contar a continuación, algunos profiles. Esto te lo cuento. La estructura de directorios que vas a encontrar directamente con Hermes es una estructura muy sencilla. Una configuración que se llama config.yaml. Luego, un archivo que es soul.md, que es el que te da la personalidad de la gente. Un memory.md, que es donde guarda la memoria. Y esto lo veremos un poco más adelante porque hay que darle una vuelta a esto. Otro punto en que dónde están las claves. Y luego ya hay varios directorios. El directorio de sessions, que es donde se guarda el historial de conversaciones. El director de skills, que esto es fundamental porque esta es donde tiene todas las skills, todas las habilidades instaladas. El cron, el directorio cron, que es donde tienes las tareas programas. Y el directorio de plugins, que es donde están los plugins activos. El directorio de skills es fundamental porque lo que te va a permitir es que tu agente vaya aprendiendo. Y no solamente que vaya aprendiendo, sino que vayas tú enseñándole qué es lo que tiene que hacer y cómo lo tiene que hacer. En este sentido, para revisar todos los tutoriales que tengo en AtariAo.es, bueno, todos no los he revisado, pero en los que he revisado utilizando Hermes, lo que he hecho ha sido crearle un skill específico. Un skill específico donde le he indicado paso a paso qué era lo que tenía que hacer. Qué es lo que tiene que revisar. Que tiene que añadir referencias a todo lo que indique. Cómo tiene que revisar los documentos. Qué partes tiene que incluir. Qué partes no tienen que incluir. En fin, todo este tipo de cosas es con los skills. Igual que esto, en un episodio anterior también te dije que cuando salía a correr, lo que estaba haciendo últimamente era hacer una captura de pantalla exactamente de los resultados del entrenamiento y se lo pasaba. Lo que he hecho ha sido crearle un skill para que sepa exactamente los datos que tiene que extraer de esa captura de pantalla. Con lo cual va mucho más al grano. Regla de oro. Las API keys, todo lo que es claves, tokens, etcétera, etcétera, tienen que ir en un archivo .env. Realmente yo no toco para nada ni el archivo .env, ni el archivo .config.yaml. No toco ninguno de estos archivos. Olvídate de tocarlos. Mi recomendación es que no toques nada, sino que directamente le digas a Hermes o a tu agente, el nombre que le hayas puesto, en mi caso yo le he puesto Chloe, a tu agente que le digas exactamente qué es lo que quieres hacer, dónde lo tiene que guardar, si hace falta decírselo. Pero no vale la pena que toques nada porque para eso ya está tu asistente. Vale, llegados a este punto, tienes la posibilidad de crear comandos. Digo, perdón, la posibilidad de crear... Sí, perdón, exactamente. Tienes varios comandos para trabajar. Uno que es Hermes Chat, que es directamente para hablar con tu agente. Hermes Doctor, que lo hemos visto un poco al principio, que es para hacer un diagnóstico. Hermes Logs, que es para ver todos los logs, para... de ahí puedes extraer si tienes algún problema. Hermes Skill List, que lo que va a hacer es listarte todas las skills. Y Hermes Config Show, que lo que te va a hacer es mostrarte la configuración activa. Con estos comandos ya puedes sobrevivir. Con esto ya puedes empezar la cosa. Pero aquí es donde viene la parte importante, en los perfiles. Porque tú puedes hacer... O sea, te puedes crear tantos perfiles como consideres. Con la gran ventaja de que Hermes lo que hace es independizar cada uno de los perfiles. Con su propia configuración, su propia memoria, sus propios skills y su propia personalidad. Básicamente es como tener varios asistentes, pero en la misma máquina. Cada uno especializado en algo distinto. Y con la gran ventaja de que según lo configures, la memoria de un perfil no contamina la de otro. Es decir, puedes tener un perfil para trabajar, otro perfil para chatear, otro perfil para investigar. Esto ya como tú quieras. Y aquí es donde lo enlazo con lo que he comentado prácticamente al principio. Cuando empiezas con todo esto, cuando empiezas con los perfiles, es muy interesante crearte un perfil solamente para hablar con él. Sin que no haga nada más. Simplemente para decirle que... Pues por ejemplo, un perfil como el que he hecho yo. Un perfil para simplemente cuando vuelvo de correr o cuando hago las operaciones normales, un perfil que no tenga nada. ¿Por qué? Pues porque lo que he hecho ha sido desactivarle todas las tool sets que puedes configurar al principio. La de visión, todo eso lo quito. Y la razón de quitar todo esto es porque una vez lo has quitado, la gran ventaja que tiene es que en la ventana de contexto lo que envías cada vez es mucho menor. Simplemente por eso. Porque si no, todos los skills que tengas cargados, todos los MCPs, toda la información de los plugins, toda la información de los tools, todo eso se envía cada vez. Y todo eso consume una gran cantidad de tokens. Y luego tampoco tiene sentido tener skills específicos para hacer investigación cuando lo único que quieres es hablar con él. Entonces simplemente cambias de perfil, te cambias al perfil de charlar y hablas. Le comentas tus cosas y él va guardando ahí su memoria sin contaminar la memoria del investigador. Que además el investigador va a tener una serie de skills y una serie de MCPs específicos para él, para que únicamente investigue. Así, para crear un perfil es tan sencillo como hacer Hermes Profile, create charla y ya tienes el perfil charla. Y de la misma manera puedes crear un investigador, un escritor y configurarle a cada uno de ellos aquellas herramientas, aquellos tool sets que necesitas para trabajar. Y de esta manera vas a optimizar, como te digo, el consumo de tokens. Cada perfil en general genera su propio comando. Es decir, charla, chat, investigador, chat, escritor, chat. Yo por ahora no sé por qué, no sé si porque tengo una versión antigua o lo que sea, no se me están creando los comandos. Entonces lo que hago es Hermes, o sea, ejecuto Hermes, chat y guión guión profile y el nombre del profile que quiero ejecutar. Luego, si quieres clonar alguno de los perfiles, porque básicamente tiene lo que necesitas, pero quieres añadirle algunos skills específicos, puedes utilizar la opción clone, que lo único que hace es clonarte el soul, el alma de tu agente. Si lo que quieres clonarlo todo es simplemente un clone all. Y si quieres clonar desde un perfil específico, también puedes hacer un clone fraud. Todo este tipo de operaciones, como te puedes imaginar ya, lo vas a encontrar directamente en las notas del episodio. O sea que no te tienes que preocupar por nada. Útil para experimentar sin ningún tipo de preocupaciones, clonas tu perfil productivo, pruebas todos los cambios que quieras y si falla algo, pues vuelves a lo original. Como te digo, mi recomendación es que si quieres uno simplemente para chatear con él, yo le quitaría absolutamente todo. Si quieres, a lo mejor le dejas el texto speech o el speech to text o lo que tú consideres, pero lo dejaría lo más limpio posible y luego le iría añadiendo poco a poco aquellas cosas que consideras que fuera o que necesitaras. De esta manera, sobre la gestión de perfil, pues tienes una serie de comandos como es profile list, que te muestra todos los perfiles que tienes habilitados, habilitados, perdón, el show charla que te muestra todos los detalles de tu perfil charla, configurar tu perfil para utilizarlo por defecto, renombrar un perfil, quitar un perfil, todo este tipo de operaciones las puedes hacer perfectamente y son muy cómodas, sobre todo a la hora de trabajar con perfiles. Dentro del directorio .hermes tienes otro directorio que es profiles y dentro de ese directorio vas a tener todos tus profiles identificados con su configuración, su soul, su memory. Es más, por ejemplo, para un investigador, a lo mejor necesitas un determinado... un determinado... ahora no me va a salir... unas determinadas características especiales que no necesitas en los otros. Bueno, pues lo vas a tener ahí perfectamente identificado. Bueno, sobre soul y memory. Sobre soul y memory, básicamente soul lo que hace es definir quién es el agente. No es un prompt. Al final, si ya has dado el gran salto de pasar de chat GPT o de James Minay algo con más enjundia, como puede ser Hermes, el soul lo que digamos es el prompt básico, lo que le da la personalidad, el tono, el límite a tu agente. De hecho, yo lo he contado ya en más de una ocasión, mi agente, pues a Chloe le he puesto que sea irónica, sarcástica, ácida, burlona, para que tenga esa personalidad y que me encuentre como si detrás hubiera alguien real. La gran ventaja de soul es que se inyecta en cada interacción. No es un contexto que se pierde. Eso siempre va. Con lo cual tienes que tener en cuenta un detalle fundamental. Y es que si haces un soul muy grande y tienes una ventana de contexto muy pequeño, es muy probable que parte de la información se pierda. Por esto es interesante que definas un soul lo más específico posible. Por ejemplo, definir la voz del soul. Quiero respuestas cortas, sin rodeos, dos frases mejor que un párrafo, humor seco, sarcástico, tu trabajo, charla natural y tus reglas. Define las reglas que quieres. Por ejemplo, solo en español. ¿Cuáles son los peligros? Un poco lo que te comentaba. Si tienes un soul demasiado largo, el agente ignora partes. Si tienes un soul demasiado genérico, tu agente no tiene personalidad. Y si dentro de tu soul hay contradicciones, el comportamiento de tu agente va a ser completamente errático. Y luego hay otra parte que es memory. Esto es muy importante. Yo estoy haciendo unas pruebas, y aquí ya me estoy metiendo en camisa de once varas, estoy haciendo unas pruebas creándole mi propia memoria. Una memoria que le he llamado Hmemory, de HermesMemory, donde le he definido un backend en Rast con un Rack para que pueda hacer búsqueda semántica sobre la base de datos. Estoy probándolo. De hecho, he creado otra memoria distinta. Estoy probando una, estoy probando otra, para ver cómo funciona. Hay muchas. Quiero decir, hay muchos plugins que le puedes instalar directamente a Hermes para trabajar con él. Y yo mi recomendación es que te instales alguna memoria. Trabajar con la Smartdown, bueno, está bien, pero tiene muchas limitaciones. De hecho, memory tiene una limitación de 2.000 caracteres, creo recordar que era. Con lo cual, a partir de ahí ya tienes que saltar a otras memorias. Vale la pena que en lugar de esto, utilices alguno de los plugins para sacarle el máximo jugo posible a tu agente. Y es que al final, una de las partes más importantes que tiene tu agente es el tema de la memoria. Si tu agente no recuerda las cosas, al final es un asistente muy tonto. Y te digo que es un asistente muy tonto porque le tienes que estar repitiendo una y otra vez exactamente lo mismo. Y esto no tiene absolutamente ningún sentido. De hecho, yo ahora mismo no solamente le he dado la búsqueda semántica, sino que le he añadido etiquetas, he añadido una serie de cosas para ver cómo funciona. Pero como te digo, estoy probando. Tiene una capa donde recuerda las conversaciones recientes, una capa de consolidación, una capa profunda. En fin, tiene distintas cosas para jugar con él. Pero actualmente yo no te recomendaría que lo probaras porque es un poco una locura que estoy probando con tres capas de memoria. Ya te puedes imaginar que no sé. Yo te recomendaría básicamente que probaras todo lo que hay de memorias y eligieras la que más adecua a tus necesidades. Luego, el siguiente bloque, una vez ya hemos hablado de la memoria, es los MCPs, el texto speech y el speech to text. El MCP, de esto ya te he hablado en otros episodios, son los Mobile Context Protocol, que son básicamente herramientas para tu agente. Cada MCP es un servidor que expone herramientas que Hermes puede utilizar, como puede ser buscar en web, hacer acceso a YouTube, consultar bases de datos, leer archivos, interactuar con APIs. Todo eso lo puedes hacer. Y no solamente eso, también tienes MCPs nativos de Hermes. Yo, en mi caso, me he creado mis propios MCPs que están disponibles en los repositorios de GitHub, que lo que te va a permitir es hacer búsquedas web, pero además con unas herramientas muy específicas donde te va a obligar a instalarte SearchNG y luego también Invidius para hacer en YouTube. La gran ventaja de estos dos es que son gratuitos porque vas a montar tú tus dos buscadores, y esto ya le dediqué un episodio, y lo que resulta brutal es que te permite buscar en YouTube y encontrarlo prácticamente todo. No solamente esto, sino que además he añadido, ya está disponible un MCP para que puedas preguntarle directamente a tu agente que haga, o sea, puedes conectar tu agente directamente con atareado.es para hacerle consultas. Así, hasta ahí hemos llegado. Configurar un MCP es bastante sencillo. Simplemente es Hermes, MCP, AT MCP, o una cosa así, no recuerdo exactamente, en las notas del episodio lo vas a encontrar, y allí defines exactamente todos los MCPs que quieras y la configuración que necesites. Tanto MCPs de tipo HTTP como MCPs de tipo STDio, o de cualquiera puedes hacerlo. Y luego está la parte del texto speech y el speech to text, que básicamente es convertir de texto a voz y de voz a texto. Existe una herramienta que es un provider, que es Edge, que viene por defecto, que es el de Microsoft, y lo que te permite es utilizarlo de forma gratuita. Yo lo que he creado ha sido mi propia herramienta porque quería modificar la velocidad del habla, así como el pitch. Esas dos cosas. Y he creado una herramienta que también está disponible, por supuesto, en GitHub, que la puedes descargar y la puedes utilizar. Casos de uso. Pues aquí ya te lo puedes imaginar. El primer caso de uso y el más evidente es que te lea un artículo largo mientras cocinas, por ejemplo, o incluso que te cuente la receta que tienes que cocinar, que te grabe el resumen del tiempo, que te mande un audio con las noticias del día. Todo eso te vale y es muy interesante. Y de la misma manera tienes el speech to text, es decir, que tú directamente le hablas, por ejemplo, a Telegram y él lo que hace es convertir ese audio en texto. Para eso utiliza un provider que es Whisper y Whisper lo hace directamente en local. No necesita nada más. Puedes elegir el modelo y cómo funciona. También puedes utilizar, por supuesto, la API de OpenAI, pero bueno, ahí ya tienes que meterle la API Key y empezar a gastar. La joya de la corona. La joya de la corona y esto seguro que lo has visto en miles de vídeos es cómo configurarlo para utilizarlo con Telegram. Yo actualmente lo estoy utilizando con Telegram y con Matrix, pero te tengo que decir que con lo que mejor funciona hasta el momento es con Telegram. Aunque te digo, también te tengo que decir que he cambiado el backend de Matrix porque el que he estado utilizando actualmente, pues, Conduit, estaba dándome algún problemilla y he querido irme a los que son más nuevos. La gran ventaja de tener Telegram o Matrix o lo que tú quieras es que directamente le hablas al móvil como si fuera un contacto más, él hace su speech to text, entiende lo que le estás diciendo y directamente te contesta, ya sea por voz, exactamente igual que tú, o directamente mediante texto. Eso como necesites. Otra de las características súper interesantes es los Chrome. Te puede crear Chrome para, por ejemplo, que te haga un resumen del tiempo para hoy. Eso todos los días. O que directamente te mande todos los días un resumen de las noticias más importantes en relación a la inteligencia artificial. Todo esto lo puedes hacer y no hace falta que entres directamente en la configuración de Telegram o en la configuración de lo que tú quieras o en la configuración de Hermes. Tú a Hermes simplemente le tienes que decir que programe una tarea que sea hacer lo que tenga que hacer y que todos los días te la envíe a Telegram. Se lo dices así, de voz y él te crea el Chrome que tenga que crearte y hace lo que tenga que hacer sin que tú tengas que programar absolutamente nada. Y me queda contarte lo fundamental también que son los skills, el Chrome que es un poco lo que te ha adelantado y el Kanban. Los skills al fin y al cabo y lo más importante y esto ya te ha hablado en algún episodio y sobre todo te he remitido a la página web de Web Reactiva del amigo Daniel Primo donde él explica perfectamente cómo funcionan los skills. Bueno, pues la gran ventaja de todo esto y para lo que yo estoy utilizando por ejemplo para revisar todos los documentos que he hecho es especificarle claramente qué es lo que tiene que hacer y cómo lo tiene que hacer insistirle es fundamental que añada referencias a todos los capítulos de este tutorial y él te igual que te digo eso revisa revisa este tutorial o revisa este capítulo del tutorial si encuentras algún error corrígelo añade referencias todo este tipo de cosas se las puedes contar exactamente y él te crea el skill con todo eso necesario y a partir de ahí tú le dices aplica este skill a todos los documentos de el tutorial sobre docker que encuentras en atareado.es él descarga cada uno de los capítulos del tutorial de atareado.es le aplica las condicionantes que yo he dicho y nos lo deja listo para producir una skill es realmente muy sencillito y esto ya lo he comentado en otros episodios al final es un nombre una descripción y una versión y tiene la gran ventaja de que Hermes descubre las skills de forma completamente automática cuando arranca Hermes él revisa todas las skills y ya sabe cuándo o más o menos cuándo tiene que utilizar cada skill para hacer qué cosa hay que tener en cuenta que no es determinista es decir depende de que Hermes decida utilizar la skill con lo cual si le pones una descripción muy clara de exactamente para qué vale la skill cuándo con ejemplos de cuándo la tiene que utilizar etcétera etcétera pues es muy probable que en esos casos se encuentre la skill que tú necesitas algunas de las skills que estoy utilizando es una por ejemplo que es avoid AI writing que elimina patrones de IA del contenido ha tarea house styles que lo que hace es darle mi personalidad a los tutoriales y ha tarea o podcast script que lo que genera es el guión de un determinado episodio yo le doy las condiciones lo reviso en fin son de forma interactiva cómo vamos evolucionando y cómo vamos generando ese ese guión pero tienes ahí unas condiciones básicas para que empiece a trabajar como te digo aparte de estos skills he creado pues una docena de skills particularizados para hacer revisiones de determinados capítulos del tutorial luego lo que te he contado de las tareas programadas que simplemente le tienes que decir la tarea que quieras y él la va a programar sin que tú hagas nada tú le puedes decir que todos los días busque las novedades que hay en internet lo que va a hacer es conectarse a al MCP de búsqueda en internet extraer lo más novedoso y con eso generarte un documento que te lo envía todos los días si le quieres decir envíame un documento en PDF pues él te envía todos los días ese documento en PDF y luego está otra cosa súper interesante que es el Kanban el Kanban lo que hace es básicamente te abre esto lo tienes que bueno se puede consultar por web no lo tienes que ver directamente en el sitio lo ves en el sitio y directamente lo que te muestra es tareas todas las tareas pero en modo Kanban del to do doing y done esas tres columnas con esas tres columnas tú vas a ver todas las tareas que tienes que hacer y ellas todas esas tareas que las hayas dado de alta en el Kanban se van a ir poniendo en marcha conforme se vayan necesitando es más pueden haber tareas que dependan de otras tareas y hasta que una determinada tarea no sea completado no se pone en marcha otras todo esto del Kanban es súper interesante funciona súper bien y luego aparte de esto tienes plugins plugins que ya te he hablado un poco del de memoria por ejemplo además del que yo estoy implementando también tienes oncho tienes memzero super memory y luego tienes otras como pueden ser para generar imágenes otras de métricas y trazas de observability tienes distintas herramientas y opciones igual tienes para ya cosas visuales bueno algunos consejos de configuración yo te diría que lo más importante desde luego es cuanta menos herramientas tengas configuradas más espacio tiene para hacer cosas cada toolset que no utilizas es contexto que estás ocupando innecesariamente con lo cual crea los profiles los perfiles que necesites exactamente con las herramientas justas ni más ni menos yo te diría que más bien menos y conforme le vaya requiriendo más herramientas más toolset las vayas incorporando y de la misma manera si en un determinado perfil vas a dejar de utilizar skills o bien los desinstalas o bien los eliminas o los guardas o los archivas para que no esté continuamente cargándose esos skills que no te aportan absolutamente nada y esto es un poco lo que quería contarte en este episodio un poco lo que es la instalación así por encima un primer chat una estructura luego te he hablado y me parece súper importante los perfiles que como te decía eran o son agentes aislados que tienen personalidades distintas o los puedes configurar con personalidades distintas luego el archivo soul.md que es un archivo súper importante y eso te da la personalidad a tu agente en el cuarto blog que te he hablado de los mcps el test to speech el speech to text y la parte de telegram y por último te he estado hablando de los skills el cron el kanban los plugins en fin todo este tipo de cosas que me resultan súper interesantes para mí lo más importante es que program o sea que le dediques tiempo que empieces a utilizarlo y te des cuenta poco a poco de que hay determinadas cosas que pues es más interesante configurarla de una manera o de otra según tus necesidades por esto es súper interesante la creación de distintos perfiles dependiendo de qué es lo que vayas a hacer con cada uno de ellos y nada más esto es un poco lo que quería contarte en este nuevo episodio en este episodio número 807 simplemente una configuración de Hermes de 0 a 100 con todos los detallitos para que lo puedas exprimir al máximo todo lo que quieras le tengo que dedicar más episodios a Hermes pero será mucho más adelante porque quiero hablarte un poco del Kanban mostrártelo con más detalle y también pues ver cómo evoluciona la memoria la memoria esta que le he integrado donde además puedo hacer búsqueda semántica y búsqueda más detallada y poco más que decirte recordarte que este es bueno recordate lo primero que si puedes una valoración en Evox Apple Podcast Spotify un mensaje un comentario en YouTube me viene fantástico sobre todo para indicarme si te gusta si no te gusta si prefieres que haga otro tipo de cosas esto estoy completamente abierto recordarte que este es un podcast de la red de podcast de sospechosos habituales donde puedes encontrar fantásticos y maravillosos podcasts puedes participar en esa red entrando en Telegram en Wintablet Info y de la misma manera si quieres participar en el grupo de atareado con Linux siempre puedes buscar atareado con Linux en Telegram y allí nos vas a encontrar hablando de Hermes de Rack de MCP de Docker de Podman o de cualquier otra cosa relacionada con el mundo del open source y poco más que decirte recordarte que la vida son dos días y uno ya ha pasado así que disfruta como si no hubiera mañana así puede ser con Linux y en este caso con Hermes mejor que mejor un saludo y nos escuchamos el próximo jueves hasta luego ¡Gracias!