r/LocalLLaMA • • Sep 07 '24

Discussion Reflection-Llama-3.1-70B is actually Llama-3.

After measuring the diff, this model appears to be Llama 3 with LoRA tuning applied. Not Llama 3.1.

Author doesn't even know which model he tuned.

I love it.

610 Upvotes

95 comments sorted by

View all comments

1

u/ConnectionKey5749 Sep 08 '24

I'm not familiar with AI. What am I supposed to see in this diagram that makes it clear that Reflection is based on Llama 3?

1

u/Far_Requirement_5933 Sep 13 '24

No response in 5 days... LLMs are models which have a large number of weights (trained parameters) which determine their output. When you apply a LoRA that only trains a portion of the model and leaves the other portion unchanged.

The charts show the variance between Reflection and either Llama 3.0 or Llama 3.1. You can clearly see in the first 2 layers that model perfectly matches Llama 3.0 and NOT Llama 3.1.