The video demos Qwable 27B Coder, a new open-source local coding model on Hugging Face (updated end of June) built on a Qwen 3.6 27B base, showing projects made locally like a polished animated landing page and working games. It’s run on a Mac Studio (Apple M4 Max, 36GB) using Apple MLX (not available via Ollama) and is integrated into an agent operating system to generate live previews and save builds in a workspace. The presenter says Qwable tops their local leaderboards (Goldie Bench/Cody Bench comparisons) and outperforms recently tested local models like Gemma 4 12B Coder, Quifos 9B, and Onif 1.0 on the same tasks, though it runs noticeably slower than smaller models such as Quifos 9B.
00:00 Meet Qwable 27B
00:56 Demos Landing Pages
01:23 Model Specs Setup
02:04 Benchmarks Versus Locals
02:40 Running On Mac
03:03 Agent OS Live Previews
04:24 Speed Tradeoffs
05:00 Real Task Comparisons
05:49 Local AI Is Accelerating
06:05 Agent OS Features Tour
06:40 Community Courses Pitch
07:28 Wrap Up Thanks
Full transcript
So today I'm going to be showing you Quable 27b Codo which is a new coding local model and I'm going to show you exactly how it works. Now we've actually built some interesting stuff with this far more than what we usually create with local models so that impressed me in the first place. Here's for example a landing page that we created locally using Quable and it was pretty easy and simple to create. It moves in the background, it looks better than like 99% of those sort of local websites you usually create.
I wouldn't say it's anywhere near like frontier level but at the same time like you can create some pretty cool stuff better than what some of the things I've built with other models and I'll show you some example comparisons in a second as well and we'll look at how it compares versus other local models I've tested recently like for example Onyf or Quifos and I do think these local models are getting better and better and this is one of the best ones I've seen recently so this is a cool little game we've built out. Here's another one. The thing that I will say here is that if you look at this it kind of feels like last year's frontier if that makes sense so it's not like going to start competing with Opus 4.8 or Fable 5 tomorrow but it can build some pretty cool stuff and this is actually better than I expected it to be when it comes to to build it out with these local models. So how does this work?
Well essentially what we built here is Quable 5.27b Coda. It's open source, it's free to run and it just dropped on Hugging Face. Pretty good so far, just got updated end of June and it's got a Quen 3.6.27b base so that is the base for creating this. So it's free to use, free to use locally and we actually ran it with Apple MLX which is a free open source setup just for using local models so we can run LLMs from Hugging Face using the system and that's how we ran it.
You can't get it through Olama and if you're wondering how it performed on the benchmarks here, yeah it did pretty good right. I would say in terms of local models as far as they go it's right at the top of the leaderboard from everything we've tested out recently. So for example recently we've tested out Gemma 4.12b Quifos 9b or NIF 1.0 and I would genuinely say like the quality of stuff that we got from our NIF is nowhere near the same level as the quality of stuff we got from Quable. Loving these names by the way.
So it's actually my favorite local model so far. I'm on a Mac Studio, Apple M4 Max, 36 gigabytes of memory. When I actually run and create stuff with this model I can definitely feel like the whole setup runs a bit slower but at the same time it can run in the background whilst I do other things so it's not like gonna completely slow down your whole setup. So we can still for example like run Cloud Desktop whilst this was running before.
Now also something that we did is we plugged it into our agent operating system so that we can generate live previews whilst we're building with this stuff. So for example if we give it a command well our local engine inside the agent operating system runs with Quable so we can run this on three models now like Quable 5 and then when we say build something out it will actually preview it so we can see what we've created and then open up the preview and everything that we create is plugged into our workspace so for example that 3D dragon game that we just talked about that is available to preview inside a workspace right here and then we can come back to everything that we've created which is pretty cool. Also inside the workspace we can open it up inside a new tab or we can get the code from it directly but that's basically a really cool way to build with local models, preview what you've created and then run it on three models as well. Now obviously Quable kind of a reference to Fable 5 but that kind of oversells it so it's based on Quen 3.6 27b which is a strong base model from Alibaba and it's got a fine tune as well so it's basically Quen 3.6 27b with a flashy Quable coder jacket and then if you're wondering how to run it so you can't run it from Olama you would run it with Apple MLX if you're running it on a Mac but yeah the stuff that I built was pretty nice the only problem was that it's a lot slower so if you were comparing it to for example like Quifos 9b, Quifos 9b is way way faster than Quable 27b so that's something to be aware of as well it's like yes it builds better stuff but it's going to slow you down however it's top of our leaderboards when it comes to GoldieBench and the local models we're testing out and this is something I'm just going to keep building out over time so you can see the comparisons and see how they perform.
The cool thing as well like you can run your AgentOS now on free local private models you don't need Wi-Fi to use these as well and they're just ready to go whenever you need them. Now if we have a look for example at the same task with Jemma 4, Jemma 4 12b coder totally failed on that task so if we click on this for example this is the same game and it just it just didn't work it didn't work at all the same for example with Quifos so Quifos 9b we tested the same prompt and it just didn't work right it totally failed so out of everything I've tested this is the one where stuff we've created actually works full-time and you can see the prompts here on GoldieBench if you want to see how they compare even like the landing page itself looks pretty nice pretty smooth to use pretty clean doesn't look like generic AI slop and the moving background is quite nice too so it is an impressive model from everything I've tested so far. So that's basically it the thing that I would say with all this is that local AI is moving fast models are changing all the time there's better and better stuff that I'm seeing coming out pretty much every single day so I feel like it's ramping up right now especially with open source models coming out like all these new updates from China which is pretty cool so Quable is one model but if you want the AgentOS which is the operating system that runs any model local frontier from one dashboard here's what you get inside so you get the local Hermes agent engine which can run free local models you got Agent Kanban in there you got GoldieBench style testing we actually have token efficiency playbooks as well so if you're using paid APIs you can make sure that's more efficient and use less tokens we have every CLI plugged into there a workspace with every build saved and a memory of your whole business built with Obsidian inside the AI profit boardroom so if you want to get the full system for us for the agent operating system you can get that inside our community link in the comments description or just go to the AIprofitboard.com inside the community you can ask questions get help and support I personally answer every single question every single day with video tutorial and then inside the classroom you can get access to all of our new courses and free trainings as you can see here so for example if you want to go from beginner to expert you can get that inside this section if you want to get our agent operating system you can get that over here if you want to learn more about the new stuff we have new videos tutorials and guides released every single day and we date them so you know that they're actually new and then inside the calendar you can jump on weekly coaching calls get help and support in real time inside the map you can meet people in your local area who are building with AI agents that's all available link in the comments description or just go to the AIprofitboard.com thanks for watching
More episodes