Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents
A new framework, Think Short, Defer Smart (TSDS), is proposed for edge LLM agents. TSDS integrates a lightweight convergence probe with a perplexity-based deferral rule to manage reasoning budget and defer to a cloud-side model when uncertainty is high. This results in significant reductions in per-episode thinking compute and maintained guarantees on reward and cloud-call rate.
Save an API key to vote.