For the Fastest Reasoning-Based Tool Calling
Learning Domain-Specific Latent Action Spaces for Multi-Step Tool Use
Tool-calling systems are increasingly capable, but they still rely heavily on repeated autoregressive reasoning.
A typical agent loop looks like:
(q_t, s_t) \rightarrow