Back to library

AI / Technology

ReActSynergizing Reasoning and Acting in Language Models

Three key questions about this paper

What problem does ReAct: Synergizing Reasoning and Acting in Language Models address?

Authors: Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao Published: ICLR 2023 · arXiv v3, 10 March 2023 Primary source: arXiv:2210.03629 · PDF Reading note: Results below come from the paper's reported experiments. Product implications are explicitly marked as interpretation.

What evidence supports the main claim in ReAct: Synergizing Reasoning and Acting in Language Models?

The knowledge results are more nuanced. On FEVER, ReAct scored 60.9% accuracy, above chain-of-thought at 56.3%. On HotpotQA, however, ReAct scored 27.4 exact match, below chain-of-thought at 29.4 and self-consistent chain-of-thought at 33.4. Hybrid fallbacks performed best in the table: ReAct followed by self-consistent chain-of-thought reached 35.1 on HotpotQA, while the reverse order reached 64.6% on FEVER.

What limitation should readers know about ReAct: Synergizing Reasoning and Acting in Language Models?

The experiments used older, very large language models, narrow text interfaces, small numbers of prompt examples, greedy decoding in most cases, and benchmark environments rather than open-ended deployment. Prompting smaller PaLM models performed poorly; the authors' initial fine-tuning experiment used 3,000 model-generated correct trajectories and suggests potential, but does not establish broad real-world reliability. Source: paper pp. 5–9, Sections 3.3–5.

2 new free reports left todaySubscribe to Pro for unlimited reading and 10 new paper explanations each month.Upgrade to Pro