Hacker News new | past | comments | ask | show | jobs | submit
It seems to me that the model struggles to have enough general intelligence, knowledge, or reasoning capacity for arbitrary prompted tool calling. At this size, not surprising.

I am VERY interested in seeing how it could perform with some fine-tuning for a specific family of tools/tasks. That would be a great addition to the demo.

i would assume a model this size would require finetuning tbh. even functiongemma recommends that.