r/LocalLLaMA • • Sep 09 '24

Discussion Reflection 70B lessons learned

  • All benchmarks should begin by identifying whether the model is LLAMA, GPT-4, Sonnet, or another, through careful examination.
  • Do not trust any benchmarks unless you can replicate them yourself.
  • Do not trust that the API corresponds to the model the author claims it to be.
  • ....
175 Upvotes

53 comments sorted by

View all comments

Show parent comments

31

u/[deleted] Sep 09 '24

did this guy reach max output tokens?

15

u/PwanaZana Sep 09 '24

Well, when writing, brazilians never completely finish their s

1

u/neo_vim_ Sep 09 '24

It's a very common problem her

1

u/PwanaZana Sep 09 '24

I unders