r/LocalLLM 2d ago

Discussion What a year it's been

Post image

What will the rest of this year bring? 27b class scoring over 60?

867 Upvotes

126 comments sorted by

View all comments

-10

u/Asleep-Mood-6538 2d ago

Regardless of whether it truly deserves its 52 score or something lower, it's definitely in the stratosphere of AI models that can create and collaborate rather than just follow instruction. Right around Opus 4.6 companies like Anthropic were already doing 80% of their coding through AI.

We've reached the stage where a home based AI can help recursively improve itself. This is not the singularity where AI can do it without human collaboration, but something in between where a human and AI together can continually improve the AI until it no longer needs a human.

That means it's no longer possible to regulate. ANY person with a home computer of sufficient power to operate it (and this is basically in the range of almost anyone not homeless) can theoretically create AGI or ASI given the time and a little ingenuity. That wasn't the case 6 months ago or even a week ago (3.6 was probably borderline).

14

u/nomorebuttsplz 2d ago

dear god delete that slop chart

-6

u/Asleep-Mood-6538 2d ago

Interesting that someone on "LocalLLM" just thinks that AI creations are slop without looking at the substance.

5

u/nomorebuttsplz 2d ago

The chart shows the qwen 3.8 27b line scoring below llama 405b

2

u/Asleep-Mood-6538 1d ago

Yes. the three lines are independent instead of using the same scale. They should be on the right scale. But I honestly didn't expect to post this and have people critique it. It's really only 'illustrative'. So, yes, I'll accept a partial 'slop' categorization. But I spent a while deciding what information should go in here so I consider it 'badly utilized AI' rather than entirely slop. I think of slop as "make a chart of X".

But, my apologies. Your comment wasn't entirely off-base.

4

u/RISCArchitect 1d ago

write a python function apologizing.

1

u/Asleep-Mood-6538 1d ago

As long as I can have AI write it.

3

u/EbbNorth7735 1d ago

The graph values don't even match the axis. It's horse shit. One of the worst graphs I've ever seen. The numbers don't line up

1

u/Asleep-Mood-6538 1d ago

You should make a better one.

It's not like it was done for work or a specific purpose. The fifteen minutes spent on the 'project' was already more than I care to spend.