← Back to search

China’s GLM 5.2 is TERRIFYING!

AI News Today | Julian Goldie Podcast · 2026-06-15 · 13 min
relevance 66 2711 words Episode page ↗ Audio ↗
Show full episode description
GLM 5.2 vs Claude Opus 4.8: I Ran 5 Build Tests (China Won 4) + The Catch The video compares China’s new GLM 5.2 (by ZAI, released June 13, 2026) against Anthropic’s Claude Opus 4.8 (released May 28) using the same five build prompts. GLM 5.2 produces better results in four tests—a running game, an Apple-style landing page, a liquid-sloshing animation, and a neon arcade game—while Claude wins the fifth test with a cleaner moving solar system map. The script highlights pricing differences (Claude’s per-token API costs vs GLM’s flat monthly coding plan) and argues the best approach is to run both: use GLM for cheap “building” tasks and Claude for deeper thinking inside a unified “agent operating system.” It notes key catches: no independent benchmarks yet, and GLM’s API/chatbot/open weights are coming next week.
✨ Episode Outline — click any point to jump to it in the episode
Problem solved
Whether China's GLM 5.2 beats Claude Opus 4.8, and why orchestrating both in Agent OS beats picking one model.
Benefits
  • GLM 5.2 won 4 of 5 build tests against Claude Opus 4.8
  • Flat monthly price instead of per-use API costs
  • GLM 5.2 plugs easily into AI agents
  • Run cheap and smart models together in one dashboard
  • Swap models in and out without rebuilding
Use cases
  • GLM 5.2 beat Claude on 4 of 5 builds: runner game, Apple-style landing page, liquid bowl animation, neon arcade game
  • Claude Opus 4.8 won test 5, the moving solar-system map
  • This very video was fully edited with Hermes agents plugged into GLM 5.2 inside Agent OS
  • GLM 5.2 spins up 5 landing pages overnight; Claude picks the best in the morning, replacing a week of work
  • AI Profit Boardroom helped 3,500+ business owners with 181 pages of testimonials
KPIs / results
  • GLM 5.2 won 4 of 5 build tests
  • Claude Opus 4.8: $5 per million words in, $25 out; GLM coding plan ~$18/month
  • GLM 5.2 released June 13, 2026; Claude Opus 4.8 released May 28; ZAI shipped 4 flagship models in ~4 months
  • Opus 4.8 came 41 days after Opus 4.7
Tools / build
0:00 / 0:00
📑 Chapters — tap a time to jump there
00:00
GLM vs Claude Setup
01:28
Build Test 1 Runner Game
  • Test 1 runner game: GLM built the more fun one
02:01
Build Test 2 Landing Page
  • Test 2 landing page: GLM 5.2 nailed Apple-style, Claude flat
02:26
Build Tests 3 and 4 Animations
  • Tests 3 and 4 animations: GLM wins liquid bowl and neon arcade
03:48
Test 5 and Price Shock
  • Test 5 Claude wins solar system; GLM ~$18/mo vs Claude per-use
04:54
Stop Picking Sides
  • Stop picking sides; cheap Chinese models aren't junk
05:26
Agent OS Workflow
  • Run GLM and Claude together in Agent OS mission control
07:01
The Catch and Missing Proof
  • Catch: no official benchmarks; API and open weights coming
08:14
Agent OS Offer Break
  • Agent OS offer: zip install, 30-day plan, coaching calls
09:13
Daily Business Use Example
11:13
Why Agents Beat Model Chasing
  • Agents beat model chasing; swap new models in easily
12:28
Final Verdict Run Both
  • Verdict: run both, GLM cheap/fast, Claude hard thinking
China's GLM DESTROYS CLAUDE That's the claim flying around right now. So instead of guessing, I ran the same tests myself. Same prompts, same jobs, GLM 5.2 from China on one side, and then we have CLAUDE Opus 4.8 from America on the other side. And here's what happened. On four out of five build tests, the Chinese model made the better thing, the more fun game. The smoother page. I'm going to show you exactly what we're built today with it and how it works and how this outperforms CLAUDE 4.8 on many of the tests. So the crazy thing about this is CLAUDE actually lost most of the tests. You might be thinking, now hold on, you know, before you think this is some China good America bad video, it's not. It's way more useful than that because by the end, you'll know which one to use, when to use it, and the smarter move that beats picking sides at all. And there's a catch with GLM 5.2 that almost nobody is talking about. So I'll get to that. It changes everything. But first, the tests. Let me set the scene. GLM just came out on June the 13th, 2026. It's built by a Chinese company called ZAI. It used to be Jirpu. And CLAUDE Opus 4.8 is built by Anthropic in America. It came out May the 28th. Both brand new. Both top tier. So I gave them the same five jobs and just watched. The first test was a running game. A simple running game. Just dodge blocks, grab coins, speed picks up as you go. And GLM built the fun one. You can see it right here. So this is the test. And we used the same prompt on all of these. But this was the most fun one that we actually created. And you can see how cool this is. We actually tested out Kimmy K215, which came nowhere near. But here's the crazy thing. CLAUDE actually lost. I mean, look how boring this game is compared to the other one. GLM 5.2 actually beat it. And it created a way more fun situation. Second test, a landing page. So we've got GLM 5.2 here. We have Opus 4.8 here. I asked Beau for a clean Apple-style launch page for a made-up product. GLM 5.2 absolutely nailed it. Let me show you this. It looked premium. The menu worked. The scroll felt nice. CLAUDE's page had less on it and actually felt flat when we tested out. This is CLAUDE's page right here. It just feels a little bit flat, a little bit boring in comparison to GLM 5.2's version. Third test, liquid sloshing in a bowl. And Kimmy K217 and Claude Opus made something really boring here. But look at the quality of GLM 5.2. This is GLM 5.2's output. You can change the theme. You can make it more interesting. It just looks way cooler. Sounds a bit silly to test, but it's a real test of how well the model handles tricky animation. GLM 5.2 made the most fun one. You could change the colors. It feels alive. CLAUDE actually fades out really fast when you try this out. Seems super boring. Kind of works, but you can only change two colors and they're both really boring. So China won round three. And then there was a fourth test, a neon arcade game. So you can see that right here. Now, if we have a look at GLM 5.2's version, it built something genuinely fun and wild. Look how cool that is, right? That's quite a cool, fun game to play. And genuinely fun. CLAUDE's worked, but it feels a little bit buggy when you use it. And also, it's just a little bit boring if you have a look at this. Like, this is pretty standard run-of-the-mill stuff compared to GLM 5.2. So China actually won round four. That's four wins for China's AI model, GLM 5.2. Now, here's where I have to be straight with you, because I'm not here to hype. CLAUDE didn't lose everything. On the fifth test, a moving map of the solar system, CLAUDE Opus 4.8 actually made the best version. Cleaner, better looking. So it's not a clean sweep. CLAUDE is still excellent at certain jobs, but the pattern is hard to ignore. A model from China that came out this week literally just dropped within the last 24 hours, beat one of the best American models on most of the things that I threw at it. And here's the part that should make you sit up as well, is the price. CLAUDE Opus 4.8 costs $5 per million words in, $25 per million words out. That adds to fast if you're using it all day on the API. GLM 5.2 runs on a coding plan that starts around $18 per month. And that's not per use. That's a whole month, including plugging into your AI agents. Now, as of June the 13th, 2026, GLM 5.2 is available to all users of the GLM coding plan across the light, pro and max tiers. So you've got a model from China beating CLAUDE on most tests for a flat price. That's a fraction of the cost. And you see why people are losing it. Now, let me break a belief you might be holding right now. A lot of business owners think the cheap Chinese models must be junk. Knockoffs, toys. That belief is costing them money. Because the thing I just showed you wasn't a toy. It built a better landing page than most of the expensive models on Earth. The old way was pay the most, get the best. The new way means that is not true anymore. The new way is really just test them and use the right one for the job. And that brings me to the real lesson here. Because picking aside China or America doesn't really make sense. That is the wrong game. Let me show you what I mean. This morning, for example, I didn't pick one. I had GLM 5.2 building amazing stuff out inside the agent operating system. And I had CLAUDE doing the deep thinking work. And I had them both running inside one dashboard. I call the agent operating system. One screen, all my agents talking to each other, working together. And that's a move nobody is talking about. You don't choose China or American models. You run both. You let the cheap one do the heavy lifting. And the smart one handle the hard parts. And you orchestrate them from one place. Think about what that means for your business. You can have GLM 5.2 building all your landing pages for almost nothing. Whilst CLAUDE handles your customer emails. You can have both running at the same time. That's the whole point of the AgentOS. It wires your agents together into one mission control. Your CLAUDE, your GLM, your free agents, all in one spot. You're not jumping between 10 tabs. You're not paying for one expensive model to do everything. You give each job to the cheapest agent that can do it well. And here's a key detail that most people are missing. You can plug GLM 5.2 into your AI agents. You can't do that as easily with CLAUDE inside other agents. With GLM 5.2 on that coding plan, you just drop it into your setup and it just runs. So we've already plugged it into Hermes Agent, for example. Now, day one support with GLM 5.2 includes CLAUDE, Code, Klein, OpenCode, RueCode, Goose, Crush, OpenClaw, and KeyloadCode. So your whole agent army can run on a flat monthly price instead of bleeding you per use. And that's the unlock. Not which model runs or wins, but how do I run them all for cheap in one single place? Now, let me show you the catch I promised as well, because this matters before you go all in on GLM 5.2. There are no official benchmarks. Not that I've seen anyway. No independent benchmark results were published at launch. The company says it's better than past GLM versions and long coding tasks, but third party checks are still pending. This is brand new. So when someone says GLM 5.2 beats CLAUDE on the numbers, they're guessing. The numbers don't exist yet unless they've created their own benchmarks, which again, wouldn't be an independent test. So what I showed you is real. I've tested with my own eyes, but that's my test, not a lapse. So I've got to be honest about that. The hype online is running way ahead of the proof. Two more things that you should know as well. The API, the chatbot, and the open waits are all coming next week. Not yet. So for example, if you actually go to ZAI, you can't use it directly. What I can say is it's built some amazing stuff. So for example, this video that you're watching right now was fully edited with our Hermes agents plugged into GLM 5.2 inside our agent operating system. In the open version, the one you can run free on your own computer, that isn't out yet. It's prioritized and promised, but not shipped. So I'm telling you this because most videos won't. They tell you it's a clawed killer and stop there. The truth is more useful. It's incredibly good at building things. It's cheap, but the proof is still coming. And some of the best parts haven't even shipped yet. Quick thing before I show you the rest. If you want the exact agent operating system I'm running right now with GLM 5.2, with clawed plugged in, with Hermes and GLM 5.2 and everything else. We've even got Hermes Jarvis, which is a voice activated version of Hermes. And this is all plugged in and talking to each other. That's inside the AR platform. You get the agent operating system as a zip file. You can just install, drop in your GLM 5.2, your clawed, your free agents, all-in-one mission control, plus a 30-day plan on exactly how to implement all this stuff into your actual business. And that walks you through wiring GLM 5.2 into your agent step-by-step. So you're building landing pages and content for almost nothing instead of paying per use. We've got four coaching calls every week inside there where you can share your screen. We'll help you set up live. And inside the community, loads of members are already running GLM and clawed together for client work, content, lead generation. They help you 24-7. So link in the comment description or go to theairprofitable.com. Now back to it. Let me show you the smartest way to actually use this day-to-day in a real business. Here's how I run it. So I've clawed desktop, operating my agent operating system and just building stuff and wiring it in there. So it hands jobs out to the other agents. GLM 5.2 does the building. The cheaper agents do the bulk work. It all flows beautifully into the agent operating system where I can see everything and everything that I've created inside here beautifully. And you can see how cool it is and how easy it is to just preview this stuff and see what we've created. I think that's absolutely amazing and the perfect way to use it. And so you can have like a live activity feed. You can have each agent its own workspace. It's a shared memory. So they all remember the same stuff because we have Obsidian plugged into this as well. And so that last part is big, like a shared memory. It means your agents don't just start from scratch every time. They remember your business, your customers, your brand, your voice. So the work gets better the longer you run it. And here's a real example anyone can use. You could have GLM 5.2 spin up five different landing pages for your offer overnight, then have Claude pick the best one in the morning. With one sentence, you just replaced a week of work in a designer's bill. Now, let me break the second belief that's probably stopping you right now. A lot of people think this agent stuff is just for coders. I'm not technical, for example, or this isn't for me. Wrong. Dead wrong. Because inside the Air Profit Boarding, we've actually helped over 3,500 business owners. As you can see with all these testimonials, we've got 181 pages of testimonials of people doing exactly this. And the load of them have never touched AI before they joined. I'm not saying that to boast, but I've seen agency owners, freelancers, shop owners, not coders. And they're running these agents right now to get more leads and get more customers. So if they can do it, so can you. And here's the thing about timing as well. We're early on all of this. This literally just dropped today. Bear in mind, it's pretty much the same day that Fable 5 got banned. So most people are still doing everything by hand. And the ones who learn to run agents like this now, whilst it's new, get a head start that's really hard to catch. And that's the whole game right now. If you move early, you get the advantage. Now, let me zoom out and tell you what this race actually means for you. Because China is shipping fast, like really fast. ZAI shipped four flagship models in roughly four months. That's a new top model every few weeks. And America is shipping just as fast. You know, Claude Opus 4.8 came out just 41 days after Opus 4.7. That's the shortest gap between Opus releases yet. So what does that mean for you? Well, it means the model you pick today probably won't be the best one next month. There's always a newer, cheaper, better one coming. If you build your whole business around one model, you're probably stuck. When the next one drops, you're just starting over again. That's the real reason the agent operating system matters. Because we can just swap models in and out. If a new model comes out, no problem. We just plug it in. You don't have to rebuild everything. And whilst everyone is panicking about which model was best this week, you just slot the new one in and keep going. You stop chasing models. You start running a system. And let me break the last belief because it's the one that keeps most people stuck as well. People think, I'll wait until this improves or until there's a clear winner. Then I'll learn it. The people winning aren't the ones who waited. They're the ones who started learning the system whilst it was messy. So that every model just makes them stronger and better. Waiting isn't safe. Waiting is how you fall behind. So there's where we land. So did China's GLM 5.2 destroy Claude? On most of my build tests, yes. It made the better thing for a fraction of the price. That's really when it's a big deal. But the smart play isn't to ditch Claude and run to China. The smart pet play is to run both. Let GLM 5.2 build cheap and fast. Let Claude handle the hard thinking. And together, you can tie them so that they work as one team whilst you sleep 24-7. And that's not picking a side. That's building a machine that uses whatever's best, whenever, for almost anything. If you want that exact machine, the agent operating system I'm running with GLM 5.2 and Claude, both plugged in. It's inside the AR Profit boardroom. You get the full agent operating zip file to install, the 30-day roadmap for wiring GLM 5.2 into your agents to pump out landing pages and content cheaply and automate whatever you want. You get four weekly coaching calls where we set up with you live daily tutorials, including how to use GLM with Hermes and Claude together. And 3,600 business owners who are already running these agents to get more leads and more customers ready to help you at any time. We update agent operating systems as new models like GLM 5.2 drop as well as you always run in the latest. Link in the description or go to theairprofitableroom.com. Test them both, run them together, and start now whilst it's early. That is a move.