AI News Today
← All episodes
Episode 1 · May 9, 2026 · 07:09

Hermes Agent + LM Studio is INSANE!

Run Hermes AI Agent Locally for Free with LM Studio


Learn how to run the powerful Hermes agent locally and for free using LM Studio. This step-by-step guide shows you how to set up a private, offline AI system using high-performance models like Quen and Gemma on your own computer.


00:00 - Intro: Free Local AI Agents

00:41 - Setting Up LM Studio

01:09 - Choosing the Best Local Models

02:03 - Starting the Local Server

02:41 - Connecting Hermes to LM Studio

03:36 - Switching Models & Providers

05:06 - Why Use Local AI Agents?

06:01 - Implementation & Next Steps

Full transcript

Today we are going to be testing out Hermes Agent and how to use it for free with local models with LM Studio. So LM Studio can run with Hermes Agent now which means that you can run Hermes Agent and use it for free with local models. Bear in mind Hermes Agent itself is a free open source model and then if you can plug in a pre LM Studio local running AI model then you can use it for free as well. So it has a free sort of way to run it inside Hermes.

So let's get started with this right now and you can see that it's quite easy to integrate these two together. So the first thing that we need to do is run the server from LM Studio. So we're going to open up LM Studio. If you don't have LM Studio already you can get it at lmstudio.ai.

It's a free app you can download it and then you can install local models right. So for example you can set up a local server and load a model with a local server and this means you can run local models with LM Studio. So for example if we search inside the model section here we can install something like for example Gemma4 and then run that with our AI models right. It's pretty simple and easy to set up.

So for example like this and depending on your setup LM Studio will actually tell you what's the best model to run. So you can see for example here it doesn't recommend using these models because it says likely too large but you could run a more lightweight model like for example this and the other thing is that it's important to note here is with LM Studio you get quantized versions of the same model which means that you get more lightweight versions of the model that are still powerful for running as a local AI and then you can also use Hermes offline as well. That's the other benefit So we can download Gemma4 like you can see here and this is a super small lightweight model but I just want to show you as an example of how you can get that set up and then what we can do from there is we need to start LM Studio right like this. So all we do to do that is we can go into LM Studio, close that, go to local server and then we just have to run it right.

So if we click on running now LM Studio server is now running locally right all we did was we just toggle that bit right here right so we toggle that on. Once we've done that we just need to load a model onto it as well so if we load a model here we'll just get that model from Gemma4 which is downloading right so then we can get this started. Now once you've started the local server as you can see right here we then need to run Hermes agent with LM Studio as a model provider. How do we do that?

So we're going to go to Hermes setup inside terminal here as you can see and then what you want to do is complete the LM Studio setup from there. So you see how we've got LM Studio on the model so we can select we can go from there but mind as well you can also use this with Olama which runs locally too so for example we've got Gemma4 with Olama set up locally too. You can also plug in the API key if you want to add an API key and you can just go with the defaults if you want as well then we can restart the gateway to pick up the changes that's basically how we can set this up. Now you can see that Gemma4 is now downloaded so we can use that inside a new chat and if we go to the local server here we just have to load the model which is Gemma4.

I wouldn't recommend using this version of Gemma4 but it's just as an example and that's basically how you can get this working with LM Studio and that's basically how you can set this up and then from here if we go back to our terminal click on done we can launch Hermes in chat, select model right so we can switch model by type in model and then we can switch to LM Studio right there right so that's how you can switch to LM Studio and then we would select Google Gemma4 inside the studio there pretty easy to set up with Gemma4. It will also tell you what the context window is the provider the model you switch to you can have multiple different models with LM Studio too. Now if we wanted to run this locally with Olama as well let me show you how to do that and again like inside LM Studio I don't recommend using that version of Gemma4 just wanted to show you a quick way of how you can pick the model load it in the server and then from there connect that to LM Studio right that's an example but some of the best options you can do here for this sort of stuff is like GLM 4.7 flash is pretty good and there are actually some local models designed specifically for Hermes right so for example you can see a bunch of like Noose Research models here. Noose Research is the founder of Hermes Agent which means these are some of the best models you can run locally with these agents right also you could run something like a Quen.

Quen is pretty good as a local model I would use Quen 3.5 for that and also have some lightweight model versions as you can see right here so Quen 3.59b and 35b if you've got a good setup etc right so that's basically it and then if you wanted to configure this with Hermes using Olama as a local model setup you can go over here Olama cloud you can plug in the API key if you want to use the cloud method and then from there you're good to go that's how you can run it. So just to recap on this whole process you can now run powerful free AI agents like Hermes on your own computer with no internet right using LM Studio. LM Studio is a free app you can install it lets you download and run powerful AI models you have to have a good setup for this to work but it will work on Windows, Mac and Linux. I do this on a Mac Studio honestly I prefer like cloud-based models but if you do want to know how to use LM Studio this is where you can do it and you might be wondering okay how does this work together so LM Studio acts as an engine it runs the AI model on your computer and Hermes engine agent acts as the driver so it tells the AI what to do and gets tasks done for you so together they create a local fully private fully free AI agent system.

Now why should you care it's free it's private it's fast as well right and it works offline right so if you're like on a plane or something like that you can still use Hermes and you can still get the most out of it which is pretty cool. So some examples of models you could use that could be like the Quen 3 series could be DeepSea, Coda, Lama as well as another one and if you want to set it up we've got a step-by-step guide right here and a full roadmap on how to implement it and what I'll do is I'll put that inside the AI Profitable Boarding and we'll add the video tutorial as well inside there but that's the full guide on how to use it how it works why you care how to set up plus a 30-day roadmap for implementing it into your business. Now inside the AI Profitable Boarding we add like new daily tutorials full step-by-step guides on how to use all this stuff which is pretty amazing. You also get all of my best trainings on AI SEO, how to go from beginner to expert with AI automation, my best trainings here, all the systems I use inside my business and a full course for agencies on how to get clients as well so feel free to check that out link in the comments description or go to the AI Profitable Boarding.

This is my community that's focused on helping you grow and scale with AI automation. Inside the community you can ask questions get help and support whenever you need to. Inside the classroom you get my trainings in the calendar you can jump on weekly coaching calls and get help and support whenever you need to and also inside the map you can connect with people in your local city who are also using AI agents just like you. So feel free to check that out.

Thanks for watching I'll see you in the next one. Cheers, bye-bye.

More episodes

Browse all episodes →