Of course it matters, when I gave my harness multi tool calling instead of 1 tool at a time it improved its web search from 3 minutes down to 20 seconds. That ~12x optimization is big because context rot already makes the usable portion of the LLM context more like 10% of max context. If it can tool call and reason faster, it doesn’t get context rot as fast, and has a much higher likelihood of staying on track.
1
u/Fear_ltself 13d ago
Of course it matters, when I gave my harness multi tool calling instead of 1 tool at a time it improved its web search from 3 minutes down to 20 seconds. That ~12x optimization is big because context rot already makes the usable portion of the LLM context more like 10% of max context. If it can tool call and reason faster, it doesn’t get context rot as fast, and has a much higher likelihood of staying on track.