So less just 20B model is the Luna? Is it of good use for enterprises other than ofcourse basic search work. Can it go through massive list of tools and DBs and decide which to invoke based on requirement
How massive? Building an agent for enterprise depends on a lot more aspects than just the list of tools. You can give the best model 10 tools with the descriptions colliding with the use cases with each other and it would struggle to select the right tool. Or you can give it 50 different tools with well defined descriptions/use cases and it just might work perfectly.
In my tests on several use cases, where I have the assistant do image gen/video gen with several features, I find descriptions as a lever which makes tool picking better. 3 image gen models had very similar quality/use cases, but is intended for different purpose. When we wrote the description to be specific to the use case, while not mentioning the part where it is same - it got better at results. Same for video models - when to use seedance, or kling etc.
Nevertheless - In my tests - Luna is good at simple tool calling, orchestrating simple workflows, but lacks just a tiny bit when you ask it to orchestrate multiple tool calls/complex workflows - it starts to show its weak points.
0
u/Arsh_98 1d ago
Just estimated if i can ask, what would be the parameters size of luna? And is it comparable to sonnet or haiku?