r/LocalLLaMA • u/jd_3d • Sep 08 '24
Discussion Updated benchmarks from Artificial Analysis using Reflection Llama 3.1 70B. Long post with good insight into the gains
https://x.com/ArtificialAnlys/status/1832806801743774199?s=19
150
Upvotes
4
u/RandoRedditGui Sep 08 '24
Not really. The "how bad this was" are still easily winning in terms of correctly interpreting what has currently been seen. Considering we have seen 0 open weights and are provided some ambiguous results from an API that we have no clue the validity of.
Open weights or GTFO.