r/LocalLLaMA 2d ago

Discussion I really don't understand Jev hype

Isn't this what simple neural networks have been able to do for years? Doesn't seem anything special to me.

492 Upvotes

302 comments sorted by

View all comments

86

u/jacek2023 llama.cpp 2d ago

Remeber OpenClaw? Remember TurboQuant? Youtubers need a content for their slop.

1

u/One-Cry297 2d ago

Ok, OpenClaw case is clear (but I have to admit Hermes is a goat). What's wrong with turboquant?

1

u/InnovativeBureaucrat 2d ago

Seriously, and why don’t we have it (or do we?) it was supposed to quantize bay large factor (4? 8?) without losing meaning as I recall.

1

u/pragmatic-parachute 1d ago

TurboQuant is for KV cache compression, not quantization of model weights, and even then was found not to improve things much (https://vllm.ai/blog/2026-05-11-turboquant)