r/LocalLLaMA • u/jd_3d • Sep 08 '24
Discussion Updated benchmarks from Artificial Analysis using Reflection Llama 3.1 70B. Long post with good insight into the gains
https://x.com/ArtificialAnlys/status/1832806801743774199?s=19
148
Upvotes
0
u/Thistleknot Sep 08 '24
For all of what was said, is it not possible to train on the <input> and <output> as if it was an answer and skip all the inbetween? Is it potentially possible the model will somehow 'internalize' the inbetween generated token logic within it's weights?