Acting Less is Reasoning More! Teaching Model to Act Efficiently

Tool-integrated reasoning (TIR) augments large language models (LLMs) with the ability to invoke external tools during long-form reasoning, such as search engines and code interpreters, to solve tasks beyond the capabilities of internal reasoning. While reinforcement learning (RL) has shown promise in training such agents, most of existing approaches typically optimize only for final correctnes…

Paper

Similar papers

© 2026 NYSGPT2525 LLC