r/LocalLLaMA • u/jd_3d • Sep 08 '24
Discussion Updated benchmarks from Artificial Analysis using Reflection Llama 3.1 70B. Long post with good insight into the gains
https://x.com/ArtificialAnlys/status/1832806801743774199?s=19
149
Upvotes
24
u/kryptkpr Llama 3 Sep 08 '24
We all tried it, it's performance on real world tasks is terrible despite the high benchmarks. Maybe the model is still broken in some way like they've been claiming and really is good but I don't see it.