Exploring Reversal Mathematical Reasoning Ability for Large Language Models

Pei Guo, WangJie You, Juntao Li, Yan Bowen, Min Zhang


Abstract
Large language models (LLMs) have presented remarkable capabilities in the wide range of natural language understanding and reasoning tasks. Despite their success, a few works indicate that LLMs suffer from the “reversal curse”, in which LLMs can’t employ the inverted structure “B is A” when they are trained based on “A is B”. To explore the effect of the “reversal curse” for LLMs on complex mathematical reasoning tasks, we present two reversal datasets upon GSM8K and MathQA and verify that LLMs also struggle to solve reversal mathematical problems. We analyze the potential reason and attribute it to the insufficient modeling of the relationship between reasoning steps caused by the left-to-right objective. Consequently, based on the characteristics of multi-step reasoning, we design a novel training method to improve the general and reversal reasoning abilities. Finally, we conduct experiments on four mathematical datasets, and the results demonstrate that our method significantly improves the general reasoning capacities and alleviates the reversal problem. Our datasets and codes are available at https: //github.com/AllForward/ReversalMath.
Anthology ID:
2024.findings-acl.811
Volume:
Findings of the Association for Computational Linguistics ACL 2024
Month:
August
Year:
2024
Address:
Bangkok, Thailand and virtual meeting
Editors:
Lun-Wei Ku, Andre Martins, Vivek Srikumar
Venue:
Findings
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
13671–13685
Language:
URL:
https://aclanthology.org/2024.findings-acl.811
DOI:
Bibkey:
Cite (ACL):
Pei Guo, WangJie You, Juntao Li, Yan Bowen, and Min Zhang. 2024. Exploring Reversal Mathematical Reasoning Ability for Large Language Models. In Findings of the Association for Computational Linguistics ACL 2024, pages 13671–13685, Bangkok, Thailand and virtual meeting. Association for Computational Linguistics.
Cite (Informal):
Exploring Reversal Mathematical Reasoning Ability for Large Language Models (Guo et al., Findings 2024)
Copy Citation:
PDF:
https://aclanthology.org/2024.findings-acl.811.pdf