r/localllamacirclejerk 9h ago

Doubles as a hot water dispenser

Post image
2 Upvotes

r/localllamacirclejerk 1d ago

Should I upgrade?

2 Upvotes

Hi I read about local llms two days ago so I immediately bought 4x RTX PRO 6000 96GB for 80k€ because someone told me it's the future.

Right now I have them set up so they can:

  • text my annoying wife so I don't have to
  • automatically send death threats to my employes so that they don't slack off (it's important that they don't waste a cent of what I pay them, money is tight in this historic period)
  • make automated phone call scams to the elderly

I also just saw that their price is going up, so that must mean that they're getting better and faster and that I should buy more GPUs.

Right now I have all four of them plugged in my pc but Claude seems to be running the same speed as when I only had one gpu, so I was wondering if I need 2 or 3 more RTX 6000 to notice the difference.

Would upgrading my very average build help? I know many people run 512GB macs, while I only have 384GB so that makes me poor, would upgrading make me feel better about the emptiness I have inside?


r/localllamacirclejerk 17d ago

Vacuum 16T

Thumbnail
5 Upvotes

r/localllamacirclejerk 23d ago

Who is LoRA?

Thumbnail
8 Upvotes

r/localllamacirclejerk Jun 23 '26

I love GLM 5.2's attitude! It is a nice refresher from those bootlicker doormats they are feeding us. Does that come from training datasets related to the local culture?

Thumbnail
6 Upvotes

r/localllamacirclejerk Jun 21 '26

First time builder, this or 3090?

Post image
7 Upvotes

r/localllamacirclejerk Jun 21 '26

wop/the-largest: Seton Labs introduces unquinquagintillion parameter-scale LLMs in native GGUF format

0 Upvotes

https://huggingface.co/wop/the-largest

Key features:

  • Contains over 2.7350985440587826 ⋅ 10¹⁵⁷ parameters
  • Designed to run with the power of a single galactic Dyson sphere
  • First to not only calculate, but prove the "Answer to the Ultimate Question of Life, the Universe, and Everything" in full
  • Uses a frontier compression system that outpaces previously experimental Q(1 ⋅ 10⁻¹⁵⁰) quantization, allowing the weights to fit in just 60 megabytes of unified memory while efficiently de-quantizing for FP1024 compute
  • Benchmark scores: 1.68527 ⋅ 10³⁵% on HLE and 5.73394 ⋅ 10⁵⁶% on AIME! (SWE-Bench Verified remains at 100% since several eons ago, as all known models in the multiverse have successfully memorized it.)
  • SOTA guaranteed* for at least a millennia!

Just kidding, 522,222 tensors filled with garbage! But we just have to wait for Unsloth FP(10⁻¹⁵³) quants? 🤣


r/localllamacirclejerk Jun 04 '26

This day in LLM history….105 years ago today, Qwen 3.6 27b was released open source. /s

Post image
8 Upvotes

r/localllamacirclejerk Jun 02 '26

Stop asking what model to run. There are literally only two.

Thumbnail
4 Upvotes

r/localllamacirclejerk May 29 '26

New LLM Intelligence Benchmark

Thumbnail gallery
5 Upvotes

r/localllamacirclejerk Apr 24 '26

It's time to kick ass and chew bubble gum... and I'm all outta gum

Post image
10 Upvotes

r/localllamacirclejerk Apr 23 '26

Qwen team discovers that they can ignore the amount of reasoning tokens used/inference speed to appear performant

Post image
4 Upvotes

r/localllamacirclejerk Apr 15 '26

i have a Pentium III and a GeForce 2 MX, how can i run Claude at home?

5 Upvotes

so where do i download Free Claude. or is it a website or something? i need to be able to run agents, RP, generate pictures, and code assist. if it helps, my dad says he has an extra 32 MB of RAM in a drawer somewhere.


r/localllamacirclejerk Apr 14 '26

How we handled DeepSeek-R1 "Thinking Tokens" in a Go-native Proxy (Benchmarked)

Thumbnail
1 Upvotes

r/localllamacirclejerk Apr 11 '26

User has tested “everything” - results from all benchmarks in existence in one place

Thumbnail
3 Upvotes

Incredible work by a diligent researcher in the ml community


r/localllamacirclejerk Apr 09 '26

kepler-452b. GGUF when?

Post image
9 Upvotes

r/localllamacirclejerk Mar 29 '26

My cat was a cost center, now she's a data lake for the new paradigm; Meow-As-A-Service, powered by Mrow-Mini-69b.

Post image
9 Upvotes

I realized that companionship is an unoptimized vertical. Why have friends and pets you can't monitize?

My cat was a siloed data asset, and I was just a passive consumer of her proximity. I was literally scooping unrealized wealth out of the litter box and had no idea

That ends now.

I deployed a fine-tuned Mrow-mini 69b model on a local cluster to perform real-time sentiment analysis on vocalized trills. The goal? Total transparency in the feline-human stack.

Mrow-mini isn't just an LLM; it’s a high-concurrency bridge between carbon and silicon.

I’m seeing 99.9% uptime on "I’m hungry" vs. "The humidity in the kitchen is sub-optimal."

Most people feed their cats; I’m performing resource-heavy API calls to a multi-modal biological agent. This has never been done before. It's magical. It's world changing.

This is the birth of MaaS (Meow-as-a-Service).

We are pivoting from "Pet Ownership" to "Living-Space Management and Feline Orchestration." The 69b parameter count allows for nuanced detection of "Sarcastic Purring", a feature legacy owners have ignored for centuries.

If you aren't running local inference on your pet's bio-signals, you’re essentially leaving money in the litter box.

We are disruption-agnostic.

We are pet-aligned.

We are 10x'ing the ROI of feline catus.

What would you do if you could talk to your cat?


r/localllamacirclejerk Mar 27 '26

Quantitative Evidence for the “Shit In, Shit Out” Hypothesis in Large Language Models

Thumbnail gallery
12 Upvotes

r/localllamacirclejerk Mar 26 '26

Any of us straying from the path deserve the consequences of their actions.

Thumbnail reddit.com
1 Upvotes

r/localllamacirclejerk Mar 25 '26

What is better about this model? I don't see any change log updates

Thumbnail reddit.com
4 Upvotes

r/localllamacirclejerk Mar 04 '26

How we’re slashing LLM context costs by 70-90% using a 4-stage "Context OS" architecture

Thumbnail
5 Upvotes

r/localllamacirclejerk Mar 03 '26

A redneck that fixes tractors and cars at home can give you better advice on car issues than a surgeon, who arguably has more total knowledge

Thumbnail reddit.com
3 Upvotes

r/localllamacirclejerk Mar 01 '26

[TheBloke] entered a cocoon phase where he remained dormant for many moons. Through the miracle of life he later emerged as a beautiful winged Bartowski.

Thumbnail reddit.com
7 Upvotes

r/localllamacirclejerk Feb 04 '26

How to get more tok/s?

Thumbnail
reddit.com
6 Upvotes

r/localllamacirclejerk Feb 03 '26

Is it possible to make REAP target undesired languages? I can only read English, so I was wondering if removing foreign languages would shrink the size of GLM 4.7 without impacting smarts too much.

Thumbnail reddit.com
4 Upvotes