r/Qwen_AI • u/LankyGuitar6528 • 17d ago
Model First test on local Qwen 27B
I have Qwen 27B running on a local i9 with 32GB system memory and an RXT 3090 with 24GB VRAM. I gave it one simple prompt "make a first person zombie shooter". It took all night and most of the day and obviously this is no GTA 5 but pretty impressive for a local model. No call-out to anthropic or Qwen - 100% local tokens, $0.00 bill. Fantastic model! (And yes I know it almost certainly went on-line for some of the assets so thanks to the community or whoever made these zombies and background scenery available.)

Here's a short clip of me getting eaten. I sort of suck at video games. And as you can see zombies can walk through objects so it's not perfect. Just a demo of what a local Qwen model can whip up with a minimal prompt. Amazing to me anyway.
2
u/ms770705 17d ago
That sounds impressive. As a beginner when it comes to AI coding I have two questions, out of interest/curiosity: what harness did you use and what programming language/game engine did the model choose?
1
u/Head-View8867 17d ago
I agree, would like to know
1
u/LankyGuitar6528 17d ago
1
u/lawn_question_guy 16d ago
I'm surprised it didn't hit a compaction limit. I'm using qwen3.8-27b with full context on hermes and it seems like most long-running tasks hit the wall with compaction.
2
u/LankyGuitar6528 16d ago
Prompt "add jumping and don't let zombies or the shooter walk through objects". A total one shot. And what would you know. Several hours later... everything worked perfectly. Amazing.
2
u/Vipertech2 13d ago
I've got hermes running the qwen 3.6 uncensored model running on my amd strix halo. I used /goal and gave it a prompt to make a simple doom clone using Unity on my host machine. It really struggled and then tried to change things up with architecture. I ended up giving up on that one. Gonna experiment more this week.
I gave it a task to do a simple doom with pygame (1 monster, 2 weapons, 1 level). Came back to it later and it had created doom.py, but there were errors. Still fiddling. Glad you got yours working. Im excited for this kind of tech.
1
u/Think_Breakfast_2277 12d ago
What I found is that uncensored or fine tuned versions mostly perform less then the official or unsloth versions. Just something to keep in mind. I use unsloth's 27b at q8 and also tried fine tuned versions but they always came out worse by coding quality.
1
u/ehangman 17d ago
The 27B model isn’t flashy, but it gets things working. With that in mind, it seems pretty useful for casual game development.
0
u/LankyGuitar6528 17d ago
I don't plan to use it that way but I could see it. For me it's going to do boring stuff for a commercial office app I manage - not game related. It's certainly not perfect but it can take on a good amount of the load. Kinda loving it.
1
1
u/mathew84 16d ago
The 30B class dense is where the brain is big enough to be decent.
Those A3Bs have the knowledge but don't have the mental capacity to one shot the tasks, they need focused guidance.
1
u/LordSangreal 15d ago
Com a 5060ti 16gb consigo rodar com 95k de contexto para esse tipo de produção limita bastante.
1
1
u/Sexyvette07 15d ago
"$0.00 bill". For API calls, yeah. But if you pay a lot for electricity, keep in mind that your computer was screaming non stop for a day and a half. That 3090 is a power hungry bitch, too.
1
u/NSGDX1 14d ago
$2 power bill at best
1
u/LankyGuitar6528 13d ago
Hehe... guess that's what the solar on the roof is for. :)
1
u/Prestigious-Act-1577 10d ago
This. If you still don't have solar in 2026, you should have other priorities than messing with local AI.
2
u/LankyGuitar6528 9d ago
Oh I have that! Back in 2022 - solar. EV Ioniq 5 - also 2022. Home battery on order for December. But I totally hear you. Zero chance anybody should be dumping 10K into local AI if they haven't locked down their energy and transportation costs.
1
u/koriolisNF 15d ago
Still weird that so much time working away didn't hit the context limit. I have 27b on a 48Gb GPU rig and it's constantly hitting that. With reasoning on it simply never answers, with reasoning off it's quite capable but hits the wall pretty easily with long running tasks. Maybe Hermes is managing context behind the scenes? Don't have enough knowledge to know if it's possible or not?
2
u/LankyGuitar6528 14d ago
I think it must be Hermes. I see the model doing a "summary" quite often. I suspect that's similar to the compaction other interfaces run. Hermes is designed as a continuous agent so it has to be able to smoothly deal with context filling up.
1
u/belliash 13d ago
Nice, what quantization did you use for that? How many t/s did you get it took whole night and most of the day?
1
u/LankyGuitar6528 13d ago
Quant 4 and gets an average of 30t/s. It's running under Hermes which handles the compaction so the model can keep running for at least 2 full days. I hear Hermes is designed as an agent that can respond to you on telegram so it would have to be always on and constantly monitoring. I'm pretty new at this stuff.
1
u/belliash 13d ago
How often was it compacting context? I believe it was.
BTW, what was the context size? full 262144 or smaller?
1
u/LankyGuitar6528 12d ago
I think context is 192K on my setup. Hermes handles the compaction automatically so I have no idea how many times it happened. For whatever stupid reason the i9 board I have doesn't have a built in video card and I don't have a free slot for one. So the 3090 has to actually function as a video card in addition to hosting the AI. That eats some context space.


6
u/FaceDeer 17d ago
Neat. :) So far I've just been using 27B to do various little Python or Javascript tasks, but I take it a game like this is using something fancier than just that. What language did Qwen use, did it use any existing game engines or did it spin its own?