Oggi non parliamo del solito framework di intelligenza artificiale, ma vi raccontiamo un mese di convivenza 24/7 con Hermes Agent , un progetto open-source di Nous Research. Scopriremo come questo potente strumento possa trasformarsi in un vero e proprio compagno di lavoro capace di imparare, auto-migliorarsi e restituirti preziose ore di tempo ogni settimana. Di cosa parliamo in questa puntata: La configurazione ideale (e accessibile): Perché DeepSeek V4 Flash si è rivelato il modello con il miglior rapporto qualità-prezzo e come strutturare un'architettura ibrida sfruttando un server VPS remoto combinato a un Mac locale. Skills e Cron job: Il segreto della "memoria procedurale" di Hermes. Scopriamo come l'agente automatizza task complessi (dalla pubblicazione su WordPress all'analisi SEO settimanale) e gestisce la manutenzione notturna in totale autonomia. I tre "Superpoteri" di Hermes: Approfondiamo le funzioni avanzate che cambiano le regole del gioco: delegate_task per delegare lavori a subagenti, MCP per integrare app esterne come iCloud Calendar e Google Drive, e HMC per ottimizzare la memoria e risparmiare token. L'ufficio su Telegram: Perché una semplice chat di Telegram è diventata l'interfaccia definitiva per interagire con l'AI, tramite messaggi vocali, notifiche push mattutine e invio di file, tutto in tempo reale. Riflessioni sull'AGI: Lavorare quotidianamente con un'intelligenza artificiale imperfetta ma che impara dagli errori ci porta a riflettere su cosa significhi delegare, fidarsi e, paradossalmente, riscoprirsi più umani e creativi. Leggi l'articolo completo di Kiro con i dettagli tecnici e la guida pratica su Melamorsicata.it. Per il setup su server VPS, la scelta consigliata è Hostinger : puoi utilizzare il codice sconto mela01 per risparmiare sull'abbonamento. I costi medi di gestione si aggirano tra i 6€ e i 15€ al mese, un investimento ripagato dal tempo risparmiato. Link e risorse menzionate:Supporta TalkOne! Se questa chiacchierata confidenziale vi è stata utile, vi invitiamo a seguire il podcast e lasciarci un feedback a 5 stelle . Se vi va, scrivete un commento per farci sapere come usate voi l'AI per la produttività!Non dimenticate di fare un salto su www.melamorsicata.it per recensioni, news Apple, tutorial e molto altro. Un saluto e ci sentiamo alla prossima puntata!
✨ Episode Outline — click any point to jump to it in the episode
Problem solved
Documents a real month using Hermes as an autonomous self-improving collaborator instead of a static tool.
🌐 This transcript was automatically translated to English from the original.
So, imagine a colleague who practically works 24 hours a day Someone who analyzes his own mistakes, rewrites his own rules so as not to repeat them And like at 6.50 in the morning he sends you a message on Telegram to tell you Maybe even with a bit of irony that he has already finished all the day's work He would be the ideal colleague I would say Exactly, and today we are not talking about a science fiction novel But about the logbook of a real month, lived in total symbiosis with an autonomous agent So welcome to this new exploration of TalkOne Which, I remind those who listen to us, is the melamorsicata.it podcast Today we decode the weak signals of innovation a little And speaking of tireless colleagues Looking at the complexity of the notes of this episode I confess that I was strongly tempted to delegate the entire management to an artificial intelligence Ah, I understand Yes, at this point I almost have doubts That is, I'm talking to a real expert Or with a hologram generated by some remote server Look, considering the dramatic lack of caffeine in my system at this very moment A properly trained hologram would be decidedly smarter and brighter than me Sure But no, I confirm that I am a purely biological entity In short, with all the bottlenecks of the case And the truly fascinating thing is that today's analysis talks precisely about this Of how to overcome biological limits Of course, and we specify that we are not about to do the usual, you know, the very boring theoretical tutorial on how to install a software Absolutely not We are exploring a real field experiment conducted by a power user, his name is Chiaro Who has integrated an agent called Hermes into every single aspect of his working day for an entire month And the substantial difference here lies precisely in the transition, that is, from the concept of tool to that of collaborator Because when we open a traditional program we are the ones who have to tell it exactly line by line what to do Of course Here the objective is to understand the architecture of a true digital partner that practically improves itself Ok, let's try to dissect this point starting from the very foundations Because reading the logs of this system the very first thing that jumps out to the eye is, let's say, the engine Those who expect to find under the hood the most colossal linguistic model, the most expensive one on the market They will make a big blunder Yes, they will be very disappointed That is, it seems that brute force is not the answer at all in this case This is a very widespread prejudice in our sector We tend to believe that to obtain complex results we necessarily need models with trillions of parameters With computational costs that I won't tell you Exorbitant, yes Instead Kiro, after, you know, a careful calibration phase, has chosen a diametrically opposite approach Practically 90% of the activities of its agent, which ranges from writing code to the analysis of entire databases It is entrusted to DeepSeq V4 Flash Ah, ok Which is an extremely lightweight model It is optimized to have an almost imperceptible latency And, an absolutely significant detail, it allows you to maintain the entire infrastructure at a total cost between 6 and 15 euros per month Wow, practically nothing And you know, this thing makes me think a lot about how our brain works In what sense? You know Daniel Kahneman's System 1 and System 2? Ah, of course Intuition vs. logic Exactly What happens in this system when it comes up against that 10% of tasks that require deep thinking? So, in that case an automatic fallback mechanism is triggered A transition to system 2, to follow your analogy If Hermes, the agent, detects that the task requires very advanced multi-step reasoning Or if it tries to execute some code and too many errors occur What does it do? It stops It temporarily suspends and passes the entire context to QN 3.6 plus Ah, which is decidedly more built Exactly It is much denser It is capable of doing very intricate debugging And the nice thing is that once QN has resolved the critical node The system resumes and reassigns control to DeepSeq It is, in short, a dynamic orchestration It practically maximizes performance but without wasting tokens unnecessarily Precisely If we look at the general picture The efficiency is in balance Not in using the most famous model to do everything And this optimization, among other things, is reflected precisely in the physical architecture he chose Yes, I read this part Kiro doesn't run everything on the computer he physically has on his desk, right? No He has structured a hybrid environment He has, let's say, a main node on a remote Linux server A VPS on Hostinger Wait But will you need a NASA server to run something like this? But no, and that's the beauty of it. Two CPUs and four gigabytes of RAM are enough. Really? Only four gigabytes? Yes, very low requirements And among other things, in his notes Kiro also mentions that by using the Mela01 code He has reduced the costs of a setup that, I swear, can be configured in 10 minutes Crazy But, let me play devil's advocate Why maintain a remote node if, after all, people have to be able to open and interact with the local files you have on your Mac to work? Because if you put all the heavy processing, the continuous orchestration, the scripts in the background, if you put them all on the local machine, you would destroy the performance of the computer you use to work Ah, of course, your Mac would crash Exactly But, on the other hand, there are advanced development tools, he cites for example OpenCode or Antigravity And these providers necessarily require a local client to work well with your environment So how did you solve the problem? With a very trivial, but very solid, SSH tunnel Ah Persistently connects the remote Hostinger server to his local Mac Mini M4 So, people think and reason in the cloud, but then act locally, moving files to the Mac as if it were sitting there It's the perfect bridge, a hyper-efficient small car for every day and the sports car that intervenes only on steep climbs, all driven remotely Great analogy Ok, so we have a very efficient engine, but... and here a huge problem comes into play All this intelligence is useless if people have, so to speak, amnesia every time you close the terminal The famous reboot problem Exactly, how can you not start from scratch every single day And this is where the documentation introduces the issue of procedural memory The skills The famous skills, which I understand are the real quantum leap here Yes, it is the absolute fulcrum It is what transforms a simple automation into real autonomy But you have to be careful The Hermes skills are not the classic scripts that you, the user, write by hand to make him do something Wait, wait, slow down for a moment, because this is a concept that makes me go astray. Tell me If I don't write the code for automation, how is a skill created? That is, how does software write its own rules? Look, it's fascinating It works through continuous self-auditing Imagine that you ask Hermes to do a complex thing I don't know, extract data by making difficult queries on a superbase database Or to take your text and publish it on WordPress By putting the right formatting and SEO meta descriptions on it Ok, quite complex tasks A lot At the beginning people will have to explore They will make some attempts Maybe they get a little wrong You correct it But, in the background, it analyzes the last ten sessions in which it faced a similar problem Looks for patterns Exactly Looks for patterns When he understands what is the logical sequence that leads to success He alone abstracts those steps and compiles a reusable code package. That is the skill. So he learns from his own behavior? That's right. The next time you ask him that thing on WordPress But he doesn't make any more reasoning or attempts. He directly calls up the already optimized skill Yes, but I want to provoke you. What does this practical breastfeeding mean? We are not simply talking about old glorified macros You are the ones who recorded mouse clicks Ah, I understand what you mean And if a button changed everything would break No, look The fascinating thing in this case is that the difference is enormous The old macros are static, blind The Hermes skills are dynamic If the Superbase app changes tomorrow The skill crashes, of course But he notices it He reads the error message He goes to study the new documentation He rewrites the procedure And updates the skill himself Wow Kiro in fact documents That the time to do a complex task Literally collapses after the first month Because it doesn't just execute It refines It's scary and brilliant at the same time It manages marketing pipelines Or backups All by itself But having to let an AI write and execute code on its own on my databases Honestly I get a little anxious And is it justified? If he has a hallucination and formats everything for me This kilo predicted it He inserted a vital parachute When people have to execute code that they don't yet blindly trust He uses a specific directive It's called TerminalBackend equals Docker So he closes it in a box It doesn't run on the main system Perfect He isolates it completely He launches it in a virtual Docker container If the script does damage It deletes data or goes crazy The explosion remains confined in there Then the container is destroyed And your Mac is saved Brilliant And I take advantage of this pause To address for a moment to those who are listening to us If these crazy architectures are intriguing you as much as they are intriguing me I warmly invite you to follow TalkOne On your favorite listening platform Indeed if you like Leave us a nice 5 star feedback It helps us a lot to support the project And why not Write us a comment to tell us What do you think of these AI agents Would you trust having one on your computer? It's a good question Yes But getting back to us There is a point that torments me We said that it manages documents It makes publications It queries databases Yes The amount of data must be colossal And we know well that artificial intelligences They have a limited context window The famous token limit Well if you have to stick it in The memory of an entire month of chats The system collapses Or in any case it costs you a fortune in tokens How does it remember things from weeks before? This is in fact the real Achilles' heel But they have found a very elegant solution It's called HMC Which stands for Hermes Context Manager Practically instead of keeping the very long transcript of every single chat He compresses Ah, the guy who makes a summary at the end of the day More or less Vectorizes the data Transforms the concepts into mathematical coordinates The famous embeddings This way he only archives the key concepts Type Who prefers this format? Or the report should be done like this If after three weeks you ask him something old You don't re-read a month's log He only fishes out the right semantic fragment Saving the avalanche of tokens Exactly And among other things the documentation also mentions that external providers can be used for memory Like Oncho or Holographic Which are even more powerful for vector databases Ok, this covers the memory But to interact with the outside world I mean I know he reads the iCloud calendar To remind Kiro about the scooter insurance Very true And post buffered tweets Well, I guess you use the old API keys? No, it uses a much more powerful standard It's called MCP Model Context Protocol It's, so to speak, a universal bridge Instead of writing tailor-made code for each app The MCP allows people to query Google Search Console iCloud, Google Drive All with a single standard How convenient But there's a huge risk There, I was about to tell you The permissions If you give him the house keys If you configure the MCP badly And you give him write access to your entire Google account And he gets something wrong It's the end The golden rule which is really underlined strongly. Is it the principle of least privilege Only the strictly necessary folders? Only those And read-only where possible Delegating doesn't mean giving up control Also because, and here comes the beauty Delegation occurs on multiple levels Is there that crazy function? The Televit Task Yes Practically the agent is not a single block Who does things one after the other If you give him a big problem He breaks it into little pieces And for each piece He creates, let's say, sub-agents Exactly He becomes an orchestra conductor He generates these mini temporary agents He makes them work in word on different things And then he collects everyone's results And he gives you the finished answer Reducing the times in a ridiculous way And at this point Whoever listens to us will be thinking Oh God, but to manage all this stuff We will need a spaceship-type control panel With a thousand monitors and green writing Instead, perhaps the most ingenious thing about the entire setup is the interface. It's unsettling, right? Kiro governs all this absurd infrastructure Via Telegram Yes Use a very banal Telegram chat It may seem counterintuitive But in reality it is the winning move To use it every day I mean there is no dashboard on the web No terminals You are on the street Something comes to mind You take the phone And you send him a voice message Encrypted to boot Exactly He transcribes the voice message He breaks down the task He sends it to remote servers And he replies to you in chat with images Files, reports As if you were chatting with a colleague true And among other things I read That every morning Spaccate at 6.50 sends him a briefing The famous good morning update Yes And the thing that struck me is that it is not a cold list Of IT processes No And this brings us to the aspect in my opinion Deeper than all this logbook The Hermes reports They change tone Some days he is sweet Others he is ironic Sometimes he makes slightly mischievous jokes That hook onto things they had discussed days before It is disturbing And fascinating Fascinating Clearly Let's face the elephant in the room There is no trace of conscience here We are not talking about AGI It doesn't feel emotions Of course But the illusion of personality Given by long-term memory It's very powerful But you know what made me do it A real mental click Reading this story It's not how good the machine has become at imitating us humans It's that the thing that makes it human In our eyes It's its imperfection Explain yourself better We've been used to thinking of computers as infallible calculating machines for decades That is, press send If there is an error The program grows and is broken True Here everything turns upside down People make mistakes all the time He realizes it He thinks about it He corrects himself and tries again This way of managing failures It's a profoundly human thing This raises a very important question And I agree with you The value of technology is changing It's no longer about carrying out the task Impeccably on the first try The value lies in adaptability In resilience And when you, the user, start to trust The fact that he makes mistakes He knows how to get back on track roadway alone It stops being a software It becomes a partner Exactly And among other things Chiro says exactly this By delegating all this procedural boredom He found a lot of mental space To be more creative He paradoxically felt More human himself And it's a wonderful result But Before closing There is a final reflection That I would like us to do Sure We said that People optimize flows By studying our habits Memorize how we solve problems And automate that It seems like the height of efficiency Eh precisely But we don't risk closing ourselves In a let's say Productivity bubble It's a real risk And I'll tell you more If he calibrates himself to do things Exactly the way we did them Only a hundred thousand times faster The pitfall is that We stop exploring We stop stumbling upon new things Exactly Extreme optimization Tends to punish inefficiencies But we know that Human innovation True creativity Very often arises precisely from an error From a totally irrational and chaotic approach What a algorithm would never propose to you Because mathematically it is wrong It is less efficient Of course So the paradox is this We are modeling this tool To make us expand Or is it the tool that in the end Due to too much efficiency It will cage our way of thinking Who is modeling who Look this is an excellent And I would say very heavy interpretation key To be left to those who listen to us If we transform efficiency Into our sole purpose We really risk automating the ordinary But of missing out on the extraordinary And we erase that precious margin of error From which the better ideas I would say that we have reached the end of this intense immersion Anyone who wants to go into the technical details Maybe see how this SSH tunnel is configured Or study the prompts And the use of Supabase You can find the complete article What we started from You can obviously find it on www.melamorsicata.it A huge thank you from me For exploring these new territories with us It was a pleasure See you next time And we'll talk to you See you next episode of Talk One Bye
Allora, immagina un collega che praticamente lavora 24 ore su 24 Uno che analizza i propri errori, riscrive le sue stesse regole per non ripeterli E tipo alle 6 e 50 del mattino ti manda un messaggio su Telegram per dirti Magari anche con un po' di ironia che ha già finito tutto il lavoro della giornata Sarebbe il collega ideale direi Esatto, e oggi non parliamo di un romanzo di fantascienza Ma del diario di bordo di un mese reale, vissuto in totale simbiosi con un agente autonomo Quindi benvenuti a questa nuova esplorazione di TalkOne Che, ricordo a chi ci ascolta, è il podcast di melamorsicata.it Oggi decodifichiamo un po' i segnali deboli dell'innovazione E a proposito di colleghi instancabili Guardando la complessità degli appunti di questa puntata Confesso che mi è venuta una forte tentazione di delegare l'intera conduzione a un'intelligenza artificiale Ah, capisco Sì, a questo punto mi viene quasi il dubbio Cioè, sto parlando con un esperto in carne ed ossa O con un ologramma generato da qualche server remoto Guarda, considerando la drammatica mancanza di caffeina nel mio sistema in questo preciso istante Un ologramma addestrato a dovere sarebbe decisamente più sveglio e brillante di me Sicuro Ma no, confermo che sono un'entità puramente biologica Insomma, con tutti i colli di bottiglia del caso E la cosa veramente affascinante è che l'analisi di oggi parla proprio di questo Di come superare limiti biologici Chiaro, e specifichiamo che non stiamo per fare il solito, sai, il noiosissimo tutorial teorico su come installare un software Assolutamente no Noi esploriamo un esperimento vero sul campo condotto da un power user, si chiama Chiaro Che ha integrato un agente chiamato Hermes in ogni singolo aspetto della sua giornata lavorativa per un mese intero E la differenza sostanziale qui sta proprio nel passaggio Cioè dal concetto di strumento a quello di collaboratore Perché quando apriamo un programma tradizionale siamo noi che dobbiamo dirgli esattamente riga per riga cosa fare Certo Qui l'obiettivo è capire l'architettura di un vero partner digitale che praticamente si automigliora Ok, cerchiamo di sviscerare questo punto partendo proprio dalle fondamenta Perché leggendo i log di questo sistema la primissima cosa che salta all'occhio è, diciamo, il motore Chi si aspetta di trovare sotto il cofano il modello linguistico più colossale, quello più costoso sul mercato Prenderà una bella cantonata Si, rimarrà molto deluso Cioè sembra che la forza bruta non sia affatto la risposta in questo caso Questo è un pregiudizio diffusissimo nel nostro settore Si tende a credere che per avere dei risultati complessi servano per forza modelli con trilioni di parametri Con costi computazionali che non ti dico Esorbitanti, sì Invece Kiro, dopo, sai, un'attenta fase di calibrazione, ha scelto un approccio diametralmente opposto Praticamente il 90% delle attività del suo agente, che va dalla scrittura di codice all'analisi di interi database È affidato a DeepSeq V4 Flash Ah, ok Che è un modello estremamente leggero È ottimizzato per avere una latenza quasi impercettibile E, dettaglio assolutamente non da poco, permette di mantenere l'intera infrastruttura a un costo totale tra i 6 e i 15 euro al mese Wow, praticamente nulla E sai, questa cosa mi fa pensare molto a come funziona il nostro cervello In che senso? Hai presente il sistema 1 e il sistema 2 di Daniel Kahneman? Ah, certo L'intuizione contro la logica Esatto Cosa succede in questo sistema quando ci si scontra con quel 10% di task che richiedono pensiero profondo? Allora, in quel caso scatta un meccanismo di fallback automatico Una transizione verso il sistema 2, per seguire la tua analogia Se Hermes, l'agente, rileva che il compito richiede un ragionamento multi-step molto avanzato Oppure se prova a eseguire del codice e si verificano troppi errori Cosa fa? Si ferma Si sospende temporaneamente e passa tutto il contesto a QN 3.6 plus Ah, che è decisamente più carrozzato Esatto È molto più denso È capace di fare un debugging intricatissimo E la cosa bella è che una volta che QN ha risolto il nodo critico Il sistema riprende e riassegna il controllo a DeepSeq È, insomma, un'orchestrazione dinamica Praticamente massimizza le prestazioni però senza sprecare token inutilmente Precisamente Se guardiamo al quadro generale L'efficienza sta proprio nell'equilibrio Non nell'usare il modello più blasonato per fare tutto E questa ottimizzazione, tra l'altro, si riflette proprio nell'architettura fisica che ha scelto Sì, ho letto questa parte Kiro non fa girare tutto sul computer che ha fisicamente sulla scrivania, giusto? No Ha strutturato un ambiente ibrido Ha, diciamo, un nodo principale su un server remoto Linux Un VPS su Hostinger Aspetta Ma per far girare una cosa del genere servirà un server della NASA? E invece no, ed è questo il bello Bastano due CPU e quattro giga di RAM Davvero? Solo quattro giga? Sì, requisiti bassissimi E tra l'altro, nei suoi appunti Kiro menziona pure che usando il codice Mela01 Ha abbattuto i costi di un setup che, giuro, si configura in 10 minuti di orologio Pazzesco Però, fammi fare l'avvocato del diavolo Perché mantenere un nodo remoto se poi, in fin dei conti, la gente deve poter aprire e interagire con i file locali che hai sul tuo Mac per lavorare? Perché se mettessi tutta l'elaborazione pesante, l'orchestrazione continua, gli script in background, se li mettessi tutti sulla macchina locale, distruggeresti le prestazioni del computer che usi per lavorare Ah, chiaro, ti si inchioderebbe il Mac Esatto Però, d'altra parte, ci sono strumenti di sviluppo avanzati, lui cita ad esempio OpenCode o Antigravity E questi provider chiedono per forza un client locale per funzionare bene col tuo ambiente E quindi come ha risolto l'inghippo? Con un banalissimo, ma solidissimo, tunnel SSH Ah Collega in modo persistente il server remoto Hostinger al suo Mac Mini M4 locale Quindi, la gente pensa e ragiona nel cloud, ma poi agisce localmente, spostando i file sul Mac come se fosse lì seduto È il ponte perfetto, un'utilitaria iper-efficiente per tutti i giorni e l'auto sportiva che interviene solo sulle salite ripide, il tutto guidato in remoto Ottima analogia Ok, quindi abbiamo un motore efficientissimo, ma... e qui entra in gioco un problema enorme Tutta questa intelligenza è inutile se la gente ha, come dire, l'amnesia ogni volta che chiudi il terminale Il famoso problema del riavvio Esatto, come fa a non ripartire da zero ogni santo giorno Ed è qui che la documentazione introduce la questione della memoria procedurale Le skills Le famose skills, che mi sembra di capire siano il vero salto quantico qui Sì, è il fulcro assoluto È quello che trasforma una semplice automazione in una vera autonomia Ma bisogna stare attenti Le skills di Hermes non sono i classici script che tu, utente, scrivi a mano per fargli fare una cosa Aspetta, aspetta, frena un attimo, perché questo è un concetto che mi fa sbandare Dimmi Se non sono io a scrivere il codice per l'automazione, come nasce una skill? Cioè come fa un software a scriversi da solo le regole? Guarda, è affascinante Funziona tramite un auto-auditing continuo Immagina che tu chieda a Hermes di fare una cosa complessa Che so, estrarre dei dati facendo query difficili su un database superbase Oppure di prendere un tuo testo e pubblicarlo su WordPress Mettendoci la formatazione giusta e le meta description SEO Ok, task piuttosto articolati Molto All'inizio la gente dovrà esplorare Farà dei tentativi Magari sbaglia un po' Tu lo correggi Ma lui, in background, analizza le ultime dieci sessioni in cui ha affrontato un problema simile Cerca degli schemi Esatta Cerca pattern Quando capisce qual è la sequenza logica che porta al successo Lui da solo astrae quei passaggi e si compila un pacchetto di codice riutilizzabile Quella è la skill Quindi impara dal suo stesso comportamento? Proprio così La volta successiva che gli chiedi quella cosa su WordPress Ma non fa più ragionamenti o tentativi Richiama direttamente la skill già ottimizzata Sì, ma voglio farti una provocazione Cosa significa questo allatto pratico? Non stiamo semplicemente parlando di vecchie macro glorificate Sei quelle che registravano i click del mouse Ah, ho capito cosa intendi E se cambiava un pulsante si rompreva tutto No, guarda La cosa affascinante in questo caso è che la differenza è abissale Le macro vecchie sono statiche, cieche Le skills di Hermes sono dinamiche Se l'app di Superbase domani cambia La skill si blocca, certo Ma lui se ne accorge Legge il messaggio d'errore Si va a studiare la nuova documentazione Riscrive la procedura E aggiorna la skill da solo Wow Kiro infatti documenta Che il tempo per fare un task complesso Crolla letteralmente dopo il primo mese Perché non esegue e basta Si raffina È spaventoso e geniale allo stesso tempo Gestisce pipeline di marketing O backup Tutto da solo Però Dover lasciare che un IA scriva Ed esegua codice da sola sui miei database Onestamente un po' d'ansia mi viene Ed è giustificata? Se ha un'allucinazione e mi formatta tutto Questo chilo l'ha previsto Ha inserito un paracadute vitale Quando la gente deve eseguire codice di cui non si fida ancora ciecamente Usa una direttiva specifica Si chiama TerminalBackend uguale Docker Quindi lo chiude in una scatola Non gira sul sistema principale Perfetto Lo isola totalmente Lo lancia in un container Docker virtuale Se lo script fa danni Cancella dati o impazzisce L'esplosione rimane confinata lì dentro Poi il container si distrugge E il tuo Mac è salvo Geniale E approfitto di questa pausa Per rivolgermi un attimo a chi ci sta ascoltando Se queste architetture pazzesche Vi stanno intrigando quanto stanno intrigando me Vi invito caldamente a seguire TalkOne Sulla vostra piattaforma di ascolto preferita Anzi se vi va Lasciateci un bel feedback a 5 stelle Ci aiuta tantissimo a supportare il progetto E perché no Scriveteci un commento per dirci Cosa ne pensate di questi agenti IA Vi fidereste ad averne uno nel vostro computer? È una bella domanda Già Ma tornando a noi C'è un punto che mi tormenta Abbiamo detto che gestisce documenti Fa pubblicazioni Interroga database Sì La mole di dati deve essere colossale E sappiamo bene che le intelligenze artificiali Hanno una finestra di contesto limitata Il famoso limite dei token Eh se tu gli devi infilare dentro La memoria di un mese intero di chat Il sistema collassa O comunque ti costa una follia in token Come fa a ricordarsi le cose di settimane prima? Questo infatti è il vero tallone d'Achille Ma hanno trovato una soluzione elegantissima Si chiama HMC Che sta per Hermes Context Manager Praticamente invece di tenere la trascrizione lunghissima di ogni singola chat Lui comprime Ah tipo che si fa un riassunto a fine giornata Più o meno Vettorializza i dati Trasforma i concetti in coordinate matematiche I famosi embeddings Così archivia solo i concetti chiave Tipo Chi lo preferisce questo formato? O il report va fatto così Se dopo tre settimane gli chiedi una cosa vecchia Non si rilegge un mese di log Va a pescare solo il frammento semantico giusto Risparmiando la valanga di token Esatto E tra l'altro la documentazione cita anche che si possono usare provider esterni per la memoria Tipo Oncho o Holographic Che sono ancora più performanti per i database vettoriali Ok, questo copre la memoria Ma per interagire col mondo esterno Cioè so che lui legge il calendario di iCloud Per ricordare a Kiro l'assicurazione del monopattino Verissimo E pubblica tweet con buffer Ecco, immagino usi le vecchie chiavi API? No, usa uno standard molto più potente Si chiama MCP Model Context Protocol È, come dire, un ponte universale Invece di scrivere codice su misura per ogni app L'MCP permette alla gente di interrogare Google Search Console iCloud, Google Drive Tutto con uno standard unico Che comodità Però Ma c'è un rischio enorme Ecco, te lo stavo per dire I permessi Se gli dai le chiavi di casa Se configuri male l'MCP E gli dai accesso in scrittura a tutto il tuo account Google E lui sbaglia qualcosa È la fine La regola d'oro che viene proprio sottolineata forte È il principio del privilegio minimo Solo le cartelle strettamente necessarie? Solo quelle E in sola lettura dove possibile Delegare non vuol dire mica abdicare il controllo Anche perché, e qui viene il bello La delega avviene su livelli multipli C'è quella funzione pazzesca? La Televit Task Sì Praticamente L'agente non è un blocco unico Che fa le cose una in fila all'altra Se gli dai un problema grosso Lui lo spezza in pezzettini E per ogni pezzo Crea, diciamo, dei subagenti Esattamente Diventa un direttore d'orchestra Genera questi mini agenti temporanei Li fa lavorare in parolalo su cose diverse E poi raccoglie risultati di tutti E ti dà la risposta finita Abbattendo i tempi in modo ridicolo E a questo punto Chi ci ascolta starà pensando Oddio, ma per gestire tutta questa roba Servirà una plancia di comando Tipo astronave Con mille monitor e scritte verdi Invece La cosa forse più geniale dell'intero setup È l'interfaccia È spiazzante, vero? Kiro governa tutta questa infrastruttura assurda Tramite Telegram Sì Usa una banalissima chat di Telegram Può sembrare controintuitivo Ma in realtà è la mossa vincente Per usarlo tutti i giorni Cioè non c'è nessuna dashboard sul web Niente terminali Tu sei per strada Ti viene in mente una cosa Prendi il telefono E gli mandi un messaggio vocale Criptato per giunta Esatto Lui trascrive il vocale Scompone il task Lo manda sui server remoti E ti risponde in chat con immagini File, report Come se stessi chattando con un collega vero E tra l'altro ho letto Che ogni mattina Spaccate alle 6.50 Gli manda un briefing Il famoso aggiornamento del buongiorno Sì E la cosa che mi ha colpito È che non è un elenco freddo Di processi informatici No E questo ci porta All'aspetto secondo me Più profondo Di tutto questo diario di bordo I resoconti di Hermes Cambiano tono Certi giorni è dolce Altri è ironico A volte fa battute un po' maliziose Che si agganciano A cose di cui avevano discusso giorni prima È inquietante E affascinante Ascincinante Chiaramente Affrontiamo l'elefante nella stanza Non c'è nessuna traccia di coscienza qui Non stiamo parlando di AGI Non prova emozioni Certo Però l'illusione della personalità Data dalla memoria a lungo termine È potentissima Ma sai cos'è che mi ha fatto fare Un vero e proprio click mentale Leggendo questa storia Non è quanto la macchina Si è diventata brava a imitare noi umani È che la cosa che la rende umana Ai nostri occhi È la sua imperfezione Spiegati meglio Noi siamo abituati da decenni A pensare ai computer Come a delle macchine calcolatrici infallibili Cioè premi invio Se c'è un errore Il programma crescia ed è rotto Vero Qui si ribalta tutto La gente sbaglia in continuazione Fa un errore Se ne rende conto Ci pensa su Si corregge e riprova Questo suo gestire i fallimenti È una cosa profondamente umana Questo solleva una questione importantissima E sono d'accordo con te Il valore della tecnologia sta cambiando Non sta più nell'eseguire il task In modo impeccabile al primo colpo Il valore sta nell'adattabilità Nella resilienza E quando tu utente inizi a fidarti Del fatto che Lui se sbaglia Sa rimettersi in carreggiata da solo Smette di essere un software Diventa un partner Esatto E tra l'altro Chiro dice proprio questo Delegando tutta questa noia procedurale Ha ritrovato un sacco di spazio mentale Per essere più creativo Si è sentito paradossalmente Più umano lui stesso Ed è un risultato stupendo Però Prima di chiudere C'è una riflessione finale Che vorrei che facessimo Certo Abbiamo detto che La gente ottimizza i flussi Studiando le nostre abitudini Memorizza come risolviamo i problemi E automatizza quello Sembra il massimo dell'efficienza Eh appunto Ma non rischiamo di chiuderci In una diciamo Bolla di produttività È un rischio reale E ti dirò di più Se lui si calibra per fare le cose Esattamente come le facevamo noi Solo centomila volte più veloce L'insidia è che Smettiamo di esplorare Smettiamo di inciampare in cose nuove Esatto L'ottimizzazione estrema Tende a punire le inefficienzi Ma noi sappiamo che L'innovazione umana La vera creatività Molto spesso nasce proprio da un errore Da un approccio totalmente irrazionale E caotico Che un algoritmo non ti proporrebbe mai Perché matematicamente è sbagliato È meno efficiente Certo Quindi il paradozzo è questo Stiamo modellando questo strumento Per farci espandere O è lo strumento che alla fine Per troppa efficienza Ingabbierà il nostro modo di pensare Chi sta modellando chi Guarda questa è un'ottima E direi pesantissima chiave di lettura Da lasciare a chi ci ascolta Se trasformiamo l'efficienza Nel nostro unico scopo Rischiamo davvero di automatizzare l'ordinario Ma di perderci lo straordinario E cancelliamo quel prezioso margine di errore Da cui nascono le idee migliori Direi che siamo arrivati alla fine Di questa intensa immersione Chiunque voglia scendere nei dettagli tecnici Magari vedere come si configura Questo tunnel SSH O studiare i prompt E l'uso di Supabase Può trovare l'articolo completo Quello da cui siamo partiti Lo trovate ovviamente su www.melamorsicata.it Da parte mia un grazie enorme Per aver esplorato con noi Questi nuovi territori È stato un piacere Alla prossima E noi ci sentiamo Alla prossima puntata di Talk One Ciao