The video introduces MiniMax M3, a new agentic and coding model from China, and shows how to connect it to Hermes Agent to create a free AI worker that can run autonomously for up to 12 hours, controlling apps, searching the web, scheduling tasks, and sending reports. It walks through setup using Hermes Agent with Ollama running in the background and selecting MiniMax M3, then demonstrates live agent tasks like opening ChatGPT in a browser and creating a note in the local Notes app. It notes potential speed limits when using Ollama’s cloud model and suggests using MiniMax’s coding plan for faster performance via OAuth. The video also previews a cloud-hosted Hermes option (Max Hermes) powered by M3 and explains the “brain (M3) + body (Hermes)” pairing, pricing, benchmarks, and agentic capabilities.
00:00 Minimax M3 Launch
00:36 Benchmarks and Setup
01:09 Install Hermes and Ollama
01:21 Connect Model in Terminal
02:12 Live Agent Task Demos
02:52 Speed Tips and Plans
03:21 Max Hermes Cloud Option
04:15 Brain vs Body Explained
05:33 Performance and Open Weights
06:04 Wrap Up and Resources
06:41 Community and Full Guide
07:27 Final Thanks
Full transcript
Minimax M3 just dropped out of China. When you plug it into Hermes A2, you get a free AI worker that can run on your computer for 12 hours straight. We're talking about an AI that can open up your app, schedule tasks, search the web, send reports, all on its own using this powerful agentic model that literally just dropped. And in this video, I'm gonna show you how to set up, test it live on real tasks.
And there's one thing you absolutely have to do to unlock the full power of this. Plus, I'll show you something really cool at the end, which is a cloud-hosted version of this you can access from anywhere. Let's get into it. So this is Minimax, it's just come out from China.
Minimax M3, brand new coding and agentic model. And we can check out the benchmarks here as an example. So you can see here that it's performing pretty well on benchmarks. And what you can actually do is plug it into Hermes using these commands.
And you can actually test out for free using these commands too. So the way that you can get started with this if you want to plug Minimax M3 directly into Hermes Agent is we can use the commands right here. So the first thing that we need to do is make sure that we have Hermes Agent installed. Make sure that you have Olama downloaded as well.
And once you've done that, make sure you have Olama running in the background like you can see. Now, we can run this simple command. So Olama launch Hermes and then pick the model which is Minimax M3, of course. You can run that inside a terminal and boom, shakalaka.
We now have Hermes Agent connected to Minimax M3 Cloud. Perfect. Now, if you're on the free plan on Olama, obviously there's token limits. Most people won't hit those.
But also the cool thing about this is in one single click, you can also use Minimax M3 inside Cloud Code, inside Codex app, inside OpenClaw, Codex, OpenCode, et cetera. So what you can also see here is this 500,000 token context window and the input is text and image as well. Now, if we go into my Agent OS, you can see that it's automatically synced to Minimax M3 Cloud, which is perfect. So if we go and test this out now, we can try it over here.
And just for a quick agentic task, we'll say go to chatgpt.com and let's see how that goes. So another test I've run here is open up the Notes app locally and say hello. And you can see it says note created and the Notes app is now in the background. With the note typed in, here's what it did.
And then if we go over to our notes, we can see hello from Hermes, right? So it can talk to us through the computer. It can navigate to local files. It can open up apps and also it can do that in the background whilst we're talking to it directly, which is perfect.
To say schedule afternoon tea at 3 p.m. daily, you can just see how it responds to more agentic tasks, to tasks running on a schedule. The one thing that I'll say here is if you are using it through Olama, it can be a little bit slow because it's running through the cloud models, right? If you want a faster version, what I'd actually recommend is you either use OpenRouter or even better, you can actually use M3 directly with the coding plan and then use that to use Hermes, right?
And the good thing about that is once you signed up to the coding plan, then you don't need to use APIs because you can use OAuth to log into Minimax directly. Now also something that's interesting here and 99% of people don't even know about this, but you can actually get a cloud hosted version of Hermes called Max Hermes. And since this release of Minimax M3, it's now powered by M3 as well. So if you're using Max Hermes in the cloud, you can deploy it in one single click and then you can use Hermes and also MaxClaw, which is a version of OpenClaw in the cloud too, right?
So if you wanted to deploy this, you can just click on start now, log in, and then you just need to make sure you have the plan like so. So this is what it looks like once it's deployed. Honestly, for me personally, you know, you could use a terminal. Do you have the full conversation history?
No, you could use Max Hermes in the cloud, but it still doesn't look that nice. I personally prefer using my AgentOS system. It's a lot nicer, it's a lot easier to navigate and the UI is just beautiful, right? The other cool thing about this is everything that you've created, you can preview live as well, which is pretty cool.
So at this point, you might be saying, okay, what is Minimax M3? Well, the way that I look at it is that Minimax M3 is a brain, right? So it's a brand new AI model and you wanna think of it like the thinking engine. So it's a brain that you can plug into your AI tools and it can run for 12 hours autonomously on different tasks.
It can work on long horizon tasks and also it can actually improve itself on its own. It's a self learning improving model and it's designed for agentic capabilities. Now, you might say, okay, what is Hermes? So Hermes is the agent that you can plug it into.
So if Minimax is a brain, then Hermes agent is a body. It's a part that can actually do things. It was created by a new research and it's one of the few agents that has an autonomous self learning loop, which means the more you use it, the better it gets. And so why is this so good working together?
Well, the brain Minimax M3 is free if you use it on a llama, it was long to seek within token limits, and the body itself Hermes agent is free and open source too. And so when you stick them together, you basically have a very smart helper that can write your content, it can search the web, it can send emails and reports, it can control apps locally, it can run tasks on a schedule whilst you sleep, it can learn your business over time, right? The other cool thing about all of this is you don't need to be up to code to use this, you can just use it inside the chat here. Now you might say, okay, how does it perform on benchmarks?
So we can see M3, is it as high performing as Opus 4.7? Not on SWE bench, but it is comparable, and obviously it's a lot cheaper than using something like Opus 4.7 as an API. The other thing here as well is that it's pretty good at, for example, genetic tasks, so it can use MCPs and stuff like that. And also it's an open weights model too, that's the thing.
So for frontier level, they've actually said they will soon be fully open sourced on Hugging Face, which is pretty cool too. You now know how to use it, how to set it up, how to use Minimax inside Hermes, we've tested out for different workflows. It connected to Obsidian pretty nicely too as well, and it's very responsive inside the chat. If you want it to be faster, I would actually recommend using the Minimax coding plan instead, but if you want to try it out for free just to test out, then you can use something like Olama with the cloud model as well.
The other thing that I would say about this is it's not quite at the level of Opus 4.7, or for example, GPT 5.5, but at the same time, you get it a lot cheaper and it's designed specifically for a genetic task. So thanks so much for watching. If you want my full guide on Minimax M3, how to use it inside Hermes, some of the cool stuff you can do with this, a 30-day roadmap, 100 prompts, and also my full agent operating system with all of my agents plugged in to a beautiful mission control that you can just install as a zip file, you can get that inside the AI Profit Board, and this is an amazing community with 8,300 business owners inside here. You can ask questions, I actually answer them with daily video tutorials.
You can also get access to all of my new training. So I update this daily and you can get my agent operating system as well as the Obsidian setup inside here. And additionally, if you want to know how to use Hermes for free, we've got a full tutorial right there. We do four weekly coaching calls where you can ask questions, get help and support, and this is all inside the AI Profit Boarding.
Link in the comments description or just go to the AIProfitBoarding.com to get access. Thanks for watching.
More episodes