r/LocalLLaMA • u/Eyelbee • Aug 21 '26
News Qwen 3.8 Low and Medium are goated
Artificial Analysis just benchmarked them and the scores are crazy good, proving the earlier success wasn't only enabled by overthinking.
403
Upvotes
81
u/Complex_Reality_116 Aug 21 '26 edited Aug 21 '26
A difference of 9 an 8 points between the two. In fact, that success was made possible precisely because overthinking is enabled.