← Back to search
New Chinese AI Model Is INSANE! (FREE & Open Source)
AI News Today | Julian Goldie Podcast · 2026-06-13 · 5 min
Show full episode description
N2 Pro: Free Open-Source Chinese Coding Model Beating Claude & GPT on Benchmarks (260K Context) The video introduces N2, a new free Chinese open-source model with open weights and a 260K-token context window, designed for coding and agentic workflows with strong tool/function calling. It highlights benchmark results where N2 outperforms paid APIs like Claude Opus 4.7, GPT 5.5, and DeepSeek V4 Pro on some tests (including SWE Bench), and notes it is based on Qwen 3.5’s architecture. The model is available on Hugging Face and via a free API on OpenRouter, and can be used inside agents such as Kilocode, Hermes Agent, Claude Code (via an open-source “Free Claude Code” project), OpenClaude, and Pi. The creator demonstrates configuring N2 in Hermes Agent using model profiles, compares N2 Pro and N2 Mini, and shows N2 responding faster than a Gemma 4 setup. The video ends by promoting training and resources in the AI Profit Boardroom.
✨ Episode Outline — click any point to jump to it in the episode
Problem solved
Showcasing N2, a free open-source Chinese agentic model that rivals paid frontier APIs for coding.
Benefits
- Free, local and open-source with open weights on Hugging Face
- 260k token context window built for coding and agentic use
- Strong tool calling and function calling
- Runs inside Hermes Agent, Claude Code, OpenClaw and more
- N2 Pro and faster N2 Mini variants available
Use cases
- N2 beats Claude Opus 4.7 and GPT 5.5 on some benchmarks; higher than ChatGPT on SWBench
- Built a to-do list app, animations and projects for free using N2
- Voice-controlled building with Free Claude Code plugged into N2 Pro
- Side-by-side in Hermes: N2 replies faster than Gemma 4 profile
- Beats paid DeepSeek V4 Pro on benchmarks while remaining free
KPIs / results
- 260k token context window
- Beats Claude Opus 4.7, GPT 5.5 on some benchmarks
- Higher than ChatGPT on SWBench
- Based on Quen 3.5 architecture
📑 Chapters — tap a time to jump there
00:00
Free N2 Model Reveal
- Free local open-source Chinese N2 model revealed
- Available free via API2
00:29
Benchmarks and Specs
- Beats Claude Opus 4.7 and GPT 5.5 on some benchmarks
- 260k context, based on Quen 3.5 architecture
00:58
Where to Get It Free
- Free on Hugging Face with open weights and OpenRouter
01:24
Setup in Hermes Agent
- Configure in Hermes via manage models, switch to OpenRouter
- New profile section for dedicated N2 agent
02:16
What You Can Build
- Built to-do app, animations; voice-controlled building
02:58
Use Cases and Variants
- Deep search, terminal agent, OpenClaw, maths reasoning
- Two variants: N2 Pro and N2 Mini
03:28
Speed Test vs Gemma
- N2 replies faster than Gemma 4 profile in Hermes
- Can run locally with good setup
04:18
Wrap Up and Next Steps
- Easy free model setup; can build with Pi too
04:30
AI Profit Boardroom Pitch
- AI Profit Warden offers training and Hermes Agent OS setup
05:17
Final Thanks
- Closing thanks; link in description / theairprofitable.com
There's a brand new free Chinese model that's local and open source plus you can get it for free on API2 that is an absolute powerhouse when it comes to benchmarks. You can see for example NEX is actually beating Claude on the benchmarks right here. Now this is Claude Opus 4.7 but again it's a free API you can use and it's also beating GPT 5.5 on some benchmarks as well as you can see here. So for example on SWBench it's scoring higher than ChatGPT. And it's an agentic model designed to be used for AI agents. Now you can see here it's available for free and you can get it on Hugging Face which is pretty amazing. So you can see the open weights right here and this has a context window of 260k tokens. It's free to use and it's pretty powerful for coding. This is what it's designed for, coding and agentic features. It's also good at tool calling, function calling which is perfect for agentic use. Now the other thing to note here as well is actually based on Quen 3.5's architecture that's how they built it and we can get it for free also on OpenRouter as you can see right here. In fact the number of tokens being used per day is increasing a lot because more and more people are switching to this. So if we take a look at this for example this can be used for free inside KeyloadCode, HermesAgent, ClaudeCode, OpenClaw, Py, all your favorite AI agents. In fact we've already plugged it in to HermesAgent as you can see right here. We were double checking it working and it actually works which is pretty powerful stuff. Now the way you can configure this inside Hermes is you can go to manage their models in your dashboard and then from here just change this over to OpenRouter and then 2. Now there's also, and this is pretty useful, there's a new profile section here inside HermesAgent where you can actually plug it in to an AI agent for example, a separate agent with the model specifically for N2. So that's what we've done with this model right here and then when we go to chat with our AI agents we can select this one and start talking to it straight away. Now you can also use this with ClaudeCode as well. So the way you can use this is you can use something called Free ClaudeCode that's an open source project. Plug Free ClaudeCode into N2 Pro and then you can build with it for free as well. So if you want to see some stuff that were built for example, you can see some examples of stuff we've built with N2 here like a to-do list app, some cool animations, some interesting stuff right here and we built this all for free using N2. We can even control it with our voice and then get it to build something out which is amazing when you think about it. It's also interesting to see this is a Chinese AI model and it's already overtaken DeepSeek. So if you compare this to DeepSeek V4 Pro which is a paid API, bear in mind like these are all paid APIs right? ClaudeOpus 4.7, GPT 5.5, DeepSeek V4 Pro, paid APIs you would normally pay for. But if you're using Nex N2 Pro, well that's better on benchmarks versus a lot of these APIs and it's free to use. Now a couple of things that I'll say right here is that it's pretty good at deep search. So you can see for example here was a question that was tested with N2. They can do a lot of deep research. You can also use it as a terminal agent. You can use it with OpenClaw, here's an example. Web buildings are building out stuff and then also maths reasoning too. And there's also two different models. You've got N2 Pro and N2 Mini. So N2 Pro is a lot more powerful as you can see. But even N2 Mini isn't too bad and obviously that will be faster to respond as well. Now we can actually compare these side by side. So we've got a different profile for Hermes here versus on N2. Let's try these out and see which one responds the fastest. And so we've got both of those working side by side as you can see here. So this is another model that we're using. And this is using N2 and you can see it actually replies faster than using the normal profile that we've got for Hermes. In fact, by a lot more. Now if we actually go to the manage section here and we check what model the other profile was on, that was actually running with Gemma 4. So Gemma 4, which is another free model, this is outperforming it. So if you have a choice between the two, bear in mind this can be run locally as well. If you have a good setup, then you could run this and it would be faster and better to respond as well, which is pretty cool. And then of course, you could go directly inside the chat or you could build out, for example, Pi with this as well. Pi is another AI agent that you could use with N2 Pro. By the way, you've got a free choice for a new model and it's easy to use, easy to set up and a pretty powerful agentic setup. So that's basically it. Quick video. Just wanted to show you the power of N2, how it works, et cetera. You can get all of my best training on this stuff inside the AI Profit Warden. If you want my setup, including the Hermes Agent OS system that runs with free code as well. You can get that inside the AI Profit Warden along with a video tutorial, the last update dates. You can see we update it daily and a zip file with the installation. And then we add new daily tutorials. As you can see right here inside the community, you can ask questions, get help and support whenever you need it. You can see, for example, how inspired some people are by this group, which is amazing. Inside the classroom, you get access to all my best training. So if you're a complete beginner, you can go from beginner to expert using this course. And you also get new daily updates in the H&O S over here. Inside the calendar, you can drop a week of coaching calls, get help and support in real time. And inside the map, you can actually meet people in your local area who are using Hermes and other AI agents like you. So feel free to get that link in the comments description or go to the theairprofitable.com. Thanks for watching.