Skip to content
FrameworkHunt

Research · paper · added 3 Apr 2026

arXiv:2604.02155 — Brief Is Better: Non-Monotonic CoT Budget Effects in Function-Calling Agents

A research note, not a framework evaluation. It carries no architecture score and is not part of the directory or the comparison table.

The paper delivers an unexpected but well-supported finding: function-calling agents should think briefly, not deeply. The optimal CoT budget for tool selection is 8–16 tokens — approximately one sentence identifying the function and key arguments. Beyond that, reasoning quality degrades through a documented "dual failure" mechanism where extended thinking causes both function hallucination (the model generates names outside the candidate set) and wrong-function selection (the model talks itself out of the correct choice).

← All research