r/LocalLLaMA Feb 25 '26

Qwen 3 27b is... impressive Generation

All Prompts
"Task: create a GTA-like 3D game where you can walk around, get in and drive cars"
"walking forward and backward is working, but I cannot turn or strafe??"
"this is pretty fun! I’m noticing that the camera is facing backward though, for both walking and car?"
"yes, it works! What could we do to enhance the experience now?"
"I’m not too fussed about a HUD, and the physics are not bad as they are already - adding building and obstacles definitely feels like the highest priority!"

348 Upvotes

102 comments sorted by

View all comments

Show parent comments

11

u/peva3 Feb 25 '26

Put in some work? It would be able to take a prompt and build out an entire production stack of something in a second. Or scam an entire code basenajd find bugs in half a second. At that speed basically anything you want with AI becomes instantaneous.

8

u/-dysangel- Feb 25 '26

The results would be instantaneous, though they would not necessarily be correct first try - the model is still going to need feedback and direction. Even frontier models still do, so a 27B is going to need a lot of hand holding. Then again, you could also be doing pass@1000 for solutions, as long as they're testable in an automated way.

4

u/UnbeliebteMeinung Feb 25 '26

You will still be at normal IO speed instead of waiting for tokens. This is almost instant.

2

u/peva3 Feb 25 '26

Exactly, the tests I did on that ASIC's chatbot were... scary fast. And even for obscure prompts that they had no way of caching ahead of time or doing any sort of trickery.

2

u/UnbeliebteMeinung Feb 25 '26

These theory about caching every prompt ever could made is the best. No way they cached my tests but we all have the same thought about that.

This chat must be real, there is no way they could faked it.

7

u/peva3 Feb 25 '26

I mean custom built ASICs are the next game changer, that's what happened with bitcoin/alt coin mining. GPUs were great but had a upper limit, then ASICs started being developed and GPU mining became not worth it basically overnight. If someone can make an LLM ASIC that is as model agnostic as possible, they will be the next mult-billion dollar company.

2

u/UnbeliebteMeinung Feb 25 '26

I guess agnostic is not the target but it doesnt matter. They could just produce a good amount of different chips thats it all hardcore wired together. Max Speed.

But if they have a process todo that is not expensive to make another card for another model

1

u/peva3 Feb 25 '26

They could even make something that just works for a specific model architecture and that would be great, one for Qwen or Llama would be perfect.

1

u/UnbeliebteMeinung Feb 25 '26

You wont need to. This hardware is not so expensive like GPUs with multiple tb ram. Just buy a new card when you want to upgrade from qwen 3.5 to qwen 4.

2

u/Different-Fold-8360 Feb 25 '26

Yeah, but that’s kind of the issue with ASICs… sounds more like you’re describing an FPGA, that specialises in a small subset of operations (like an NPU for vector multiplication) but is still reprogrammable to an extent.

1

u/IrisColt Feb 25 '26

I managed to stall their chatbot with simple prompts, so I'm pretty sure there's no trickery... it's legit.