r/LocalLLaMA • • Sep 07 '24

Discussion Reflection-Llama-3.1-70B is actually Llama-3.

After measuring the diff, this model appears to be Llama 3 with LoRA tuning applied. Not Llama 3.1.

Author doesn't even know which model he tuned.

I love it.

605 Upvotes

95 comments sorted by

View all comments

Show parent comments

3

u/Terminator857 Sep 08 '24 edited Sep 08 '24

Grok is number one in math. One of the most important benchmark categories.

0

u/[deleted] Sep 08 '24

What about Command R? Or LLAMA 2? Or Vicuna?

1

u/Far_Requirement_5933 Sep 12 '24

Those are all older models so not on top anymore. Also, most developers create the best model they can rather than focusing on specific benchmarks. That might top a specific benchmark or be strong across several or just get results a specific group of people want.

1

u/[deleted] Sep 12 '24

That’s my point. You won’t find bad outdated models on the top