Skip to main content
Aggregate arXiv cs.AI 人工智能 17 Aug 2026 - 14:00

Second Thought: Reasoning in Parallel as LLM Agents Act and Observe

RSS 官方收录 · 可信分层展示

关键摘要

arXiv:2608.…

  • 13667v1 Announce Type: new Abstract: LLM agents in the ReAct paradigm …
  • We identify this recurring interval for Action and Observation as a re…
  • Therefore, we propose Second Thought, a training-free inference framew…

摘要引擎:抽取

正文提要

arXiv:2608.13667v1 Announce Type: new Abstract: LLM agents in the ReAct paradigm alternate between reasoning, acting, and observing, but deliberate reasoning is confined to the Thought phase: while the agent serializes an action and waits for the environment, its reasoning is frozen. We identify this recurring interval for Action and Observation as a reasoning idle window and ask whether it can host additional reasoning in parallel that serves future turns. Therefore, we propose Second Thought, a training-free inference framework that forks four auxiliary branches the instant each Thought phase concludes, decodes them concurrently with the main loop, and merges the generated thoughts back when the environment observation arrives. In this way, Second Thought relocates the added reasoning off the main thread's sequential decoding path. Across three agentic benchmarks and three reasoning LLMs, Second Thought lowers the average turn count in all nine (model,benchmark) pairs and reduces main thread decoding in six of them by up to 43% (roughly 20% on average among those settings), while leaving it essentially unchanged in a seventh; Pass@1 shows no significant change in seven of nine pairs and the two significant differences are +12.4 and +10.2 points. Against a compute-matched control that forces an equivalent budget onto the main thread's own reasoning, it attains strictly higher Pass@1 with 1.3 to 3.2 less sequential decoding in all four settings where the control applies.

来源:https://arxiv.org/abs/2608.13667

打开官方原文 站点原文页 可信分区 本信源更多 今日简报 分享图 RSS 稍后再看列表