6 ms·
Perhaps try a different model? Just from anecdotal experience, I find that the Gemma models smaller than 31B do not tool call as often as they should. Some of
by anana_ 3mo ago
Perhaps try a different model? Just from anecdotal experience, I find that the Gemma models smaller than 31B do not tool call as often as they should.
Some of the benchmarks appear to back this up [0]
Of course, a lot depends how you are using it (inference parameters, harness, prompting, etc.), but the model is quite important too.
[0]: https://artificialanalysis.ai/models/open-source/small?models=qwen3-6-27b%2Cqwen3-6-35b-a3b%2Cgemma-4-31b%2Cqwen3-6-27b-non-reasoning%2Cqwen3-5-9b%2Cgemma-4-31b-non-reasoning%2Cqwen3-6-35b-a3b-non-reasoning%2Cgemma-4-26b-a4b%2Cqwen3-5-35b-a3b-non-reasoning%2Cgemma-4-12b#intelligence-evaluations https://artificialanalysis.ai/models/open-source/small?model...