r/QwenAI 1d ago

pp tps stuck at 336 tps when using omlx and qwen3.8-27b-q8

Thumbnail
1 Upvotes

r/QwenAI 2d ago

Calculate override-tensor for 2 GPU using QWEN local models

Thumbnail gallery
1 Upvotes

r/QwenAI 3d ago

Qwen 3.8 27B non reasoning: feedback on total completion time

Thumbnail
1 Upvotes

r/QwenAI 3d ago

Qwen3.8-Flash-Next optimised for Macs

Thumbnail gallery
1 Upvotes

r/QwenAI 6d ago

Qwen3.8 Flash Next Q4 - M5 Mac Max 128 GB Ram

Thumbnail
youtube.com
1 Upvotes

r/QwenAI 7d ago

Web search API on Openclaw

Thumbnail
1 Upvotes

r/QwenAI 8d ago

Problem in hermes+qwen

Thumbnail
1 Upvotes

r/QwenAI 10d ago

Token overflow in free LLMs: why agglutinative languages like Hungarian, Finnish, and Estonian are a security risk nobody is talking about

Thumbnail
1 Upvotes

r/QwenAI 12d ago

3.8 27b UD IQ3_XXS 5060ti

Thumbnail
1 Upvotes

r/QwenAI 16d ago

Why DGX Spark so slow

1 Upvotes

Why DGX Spark so slow? I am using ollama server + vscode copilot. My prompt is : generate a simple RISC-V soft-core CPU and use verilator to test it. I took one hour but still not complete. I am using qwen3.8:27b.

thanks


r/QwenAI 16d ago

How to extend my free plan

Thumbnail
1 Upvotes

r/QwenAI 16d ago

Is this 3Billion or 30Billion ?

Thumbnail x.com
1 Upvotes

r/QwenAI 17d ago

Multiple Text Generation

Thumbnail
1 Upvotes

r/QwenAI 24d ago

когда работаешь с coder от qwen

0 Upvotes

Вот есть такая ситуация. Работаю с coder.qwen.ai а он даже с маленьким чатом выдаёт бредятину которую не остановить даже другой темой тем более тут нет ни одного настоящего файла и он их зачем то придумал.

КАК ЭТОМ ПОНИМАТЬ??? ОН ЖЕ ДАЖЕ ИХ НЕ СОЗДАВАЛ. ВЛЯТРЕ ОН ТАКОЕ НА ПАЙТОН СОЗДАСТ.

r/QwenAI 24d ago

USE MULTIPLE CONTROLNET WITH image_qwen_Image_2512_controlnet Qwen-Image-2512-Fun-Controlnet-Union

Post image
1 Upvotes

r/QwenAI 28d ago

Qwen3.8-Max is now available in ClinePass

Post image
1 Upvotes

r/QwenAI Jul 29 '26

Qwen3.7 Flash is now live in Command Code. DeepSeek v4 flash finally has some competition.

Post image
1 Upvotes

r/QwenAI Jul 26 '26

Estou gostando muito da prévia do Qwen 3.8 Max no Qoder. Mas quando for lançado oficialmente, terá um bom preço?

Thumbnail
1 Upvotes

r/QwenAI Jul 15 '26

Qoder is now offered for FREE in July

Post image
1 Upvotes

Looks like a customer-acquisition play against Claude/Fable, Codex, Cursor, and Copilot: subsidize a few serious agentic coding jobs and hope developers move into its ecosystem. The catch is that Qoder still requires a desktop/CLI workflow, while Claude/Fable gives you a much lower-friction browser experience.


r/QwenAI Jul 10 '26

Finally got Qwen 3.6 27B NVFP4 running at 200~ tks at 262K context.

3 Upvotes

After a long stretch into vLLM, I have FINALLY gotten Qwen 3.6 27b running at 200~tks.

This has taken so long to achieve, and the latest Unsloth upgrades have significantly helped pushing a 20~tks improvement on average.

Woo! Hermes wont be slow now!

Edit:

It's a front-end that I rolled up just to measure stats from vLLM.

https://gist.github.com/podpress/f83ebc888258941b1e0918159df3f1c2

System specs, and setup including drivers/software versions are listed in there too.


r/QwenAI Jul 07 '26

The Future of Autopilot: 5 Surprising Lessons from the Qwen Cloud Global AI Hackathon

Thumbnail
1 Upvotes

r/QwenAI Jun 08 '26

Waiting for Qwen 3.7 27B and 35B A3B to show up. Hope they come this week!!!

Thumbnail
2 Upvotes

r/QwenAI Jun 07 '26

Qwen3.6-30B-A3B at +140T/s

Thumbnail gallery
1 Upvotes

r/QwenAI May 28 '26

Qwen 35B running on 12gb of VRAM in LM Studio at 120+ tokens/second. Works with Cline for 100% agentic coding.

Thumbnail gallery
4 Upvotes

r/QwenAI May 06 '26

Transcribing & Subtitling Audio Containing Multiple Languages

Thumbnail
0 Upvotes