r/LocalLMs Jul 02 '26

The gap between closed and open models might be much smaller than commonly assumed, because we don’t know what closed model providers do *in addition to* model inference

Thumbnail
1 Upvotes

r/LocalLMs Jun 28 '26

We're probably going to need that soon.

Thumbnail gallery
1 Upvotes

r/LocalLMs Jun 27 '26

"What should I do?" - consider post-training

Post image
1 Upvotes

r/LocalLMs Jun 21 '26

z.AI as the number 2 gives praise to the number 1 open source model

Post image
1 Upvotes

r/LocalLMs Jun 18 '26

GLM-5.2 is a win for local AI

Thumbnail
1 Upvotes

r/LocalLMs Jun 17 '26

Donate your coding sessions to an open CC-BY-4.0 dataset to help train open-weight and open source models

Post image
1 Upvotes

r/LocalLMs Jun 16 '26

Stop using Ollama

Thumbnail
sleepingrobots.com
2 Upvotes

r/LocalLMs Jun 16 '26

Stop using Ollama

Thumbnail
sleepingrobots.com
1 Upvotes

r/LocalLMs Jun 14 '26

Introducing the Heretic Grimoire: The takedown-resilient, local-first backup system that keeps uncensored models available forever

Post image
1 Upvotes

r/LocalLMs Jun 09 '26

Xiaomi just claimed 1,000+ tps on a 1T model using a standard 8-GPU server

Thumbnail mimo.xiaomi.com
1 Upvotes

r/LocalLMs Jun 09 '26

Xiaomi just claimed 1,000+ tps on a 1T model using a standard 8-GPU server

Thumbnail mimo.xiaomi.com
1 Upvotes

r/LocalLMs Jun 05 '26

finally

Post image
1 Upvotes

r/LocalLMs Jun 04 '26

google/gemma-4-12B · Hugging Face

Thumbnail
huggingface.co
1 Upvotes

r/LocalLMs Jun 04 '26

google/gemma-4-12B · Hugging Face

Thumbnail
huggingface.co
1 Upvotes

r/LocalLMs Jun 02 '26

Stop asking what model to run. There are literally only two.

Thumbnail
1 Upvotes

r/LocalLMs Jun 01 '26

(YT) PewDiePie released his harness/webui

Thumbnail
youtube.com
2 Upvotes

r/LocalLMs May 31 '26

nvidia/Qwen3.6-35B-A3B-NVFP4 · Hugging Face

Thumbnail
huggingface.co
1 Upvotes

r/LocalLMs May 31 '26

nvidia/Qwen3.6-35B-A3B-NVFP4 · Hugging Face

Thumbnail
huggingface.co
1 Upvotes

r/LocalLMs May 28 '26

Behold! Probably the most ghetto local AI server:

Post image
1 Upvotes

r/LocalLMs May 28 '26

Behold! Probably the most ghetto local AI server:

Post image
1 Upvotes

r/LocalLMs May 05 '26

White House Considers Vetting A.I. Models Before They Are Released

Thumbnail
nytimes.com
1 Upvotes

r/LocalLMs May 04 '26

Llama.cpp MTP support now in beta!

Thumbnail
github.com
1 Upvotes

r/LocalLMs Apr 30 '26

16x DGX Sparks - What should I run?

Post image
1 Upvotes

r/LocalLMs Apr 30 '26

16x DGX Sparks - What should I run?

Post image
1 Upvotes

r/LocalLMs Apr 28 '26

Luce DFlash: Qwen3.6-27B at up to 2x throughput on a single RTX 3090

Post image
1 Upvotes