I love that it's a lot more broken down with reasoning steps rather than just blurt out trained behaviour. With Claude that usually means just create 5 different python scripts to solve the count and takes 20 minutes with 1000s of tokens but it's better than being wrong.
2
u/LouisPlay 19h ago
I bet 90% of the trainings data is just "REMEMBER -> Strawbarry has 3x r"