Don't Fine-Tune, Decode: Syntax Error-Free Tool Use via Constrained Decoding (2310.07075v3)

Published 10 Oct 2023 in cs.CL and cs.AI

Abstract: Instruction-tuned LLMs excel at many tasks but often fail to use external tools due to complicated and unfamiliar syntax constraints. While extensive fine-tuning and prompting can mitigate the issue, these approaches are expensive and hard to generalize. Furthermore, because syntax constraints are only learned implicitly during fine-tuning, models still make frequent syntax errors. Motivated by the fact that these constraints can be better satisfied explicitly with constrained decoding, we propose TOOLDEC, a decoding algorithm using finite state machines to force LLMs to follow tool syntax. Our experiments show that TOOLDEC eliminates all syntax errors, achieving significantly better performance on various base models and benchmarks. More surprisingly, when applied to generalist out-of-the-box LLMs such as Mistral-Instruct, TOOLDEC improves its accuracy in tool use from the initial 0% to an impressive 52%, matching the performance of specialized fine-tuned models such as ToolLLM.

References (28)

Citations (4)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

GitHub

GitHub - chenhongqiao/ToolDec: Syntax Error-Free and Generalizable Tool Use for LLMs via Finite-State Decoding (27 stars)

Don't Fine-Tune, Decode: Syntax Error-Free Tool Use via Constrained Decoding (2310.07075v3)

Summary

Related Papers

GitHub