RLKGF: Reinforcement Learning from Knowledge Graph Feedback Without Human Annotations

Lian Yan; Chen Tang; Yi Guan; Haotian Wang; Songyuan Wang; Haifeng Liu; Yang Yang; Jingchi Jiang

doi:10.18653/v1/2025.findings-acl.344

RLKGF: Reinforcement Learning from Knowledge Graph Feedback Without Human Annotations

Lian Yan, Chen Tang, Yi Guan, Haotian Wang, Songyuan Wang, Haifeng Liu, Yang Yang, Jingchi Jiang

Abstract

Reinforcement Learning from Human Feedback (RLHF) has been shown to effectively align large language models (LLMs) with human knowledge. However, the lack of human preference labels remains a significant bottleneck when applying RLHF to a downstream domain. Humans in RLHF play a critical role in injecting reasoning preferences into LLM, and we assume the reasoning process underlying human assessments may potentially be replaced by reasoning pathways derived from Knowledge Graphs (KGs). Inspired by this assumption, we propose Reinforcement Learning from Knowledge Graph Feedback (RLKGF), a novel method that leverages KG semantics and structure to derive RL rewards in the absence of manual annotations. Unlike Reinforcement Learning from AI Feedback (RLAIF), RLKGF directly integrates human priors encoded in KGs as the reward model, aligning LLM responses with expert knowledge without additional preference labeling or reward model training. RLKGF structures context-relevant facts into knowledge subgraphs and defines rewards by simulating information flow across semantic and logical connections between question and candidate response entities. Experiments on three public and one private medical dialogue dataset demonstrate that RLKGF significantly outperforms the competitive RLAIF in improving LLM diagnostic accuracy. The code is available at https://github.com/YanPioneer/RLKGF.

Anthology ID:: 2025.findings-acl.344
Volume:: Findings of the Association for Computational Linguistics: ACL 2025
Month:: July
Year:: 2025
Address:: Vienna, Austria
Editors:: Wanxiang Che, Joyce Nabende, Ekaterina Shutova, Mohammad Taher Pilehvar
Venue:: Findings
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 6619–6633
Language:
URL:: https://aclanthology.org/2025.findings-acl.344/
DOI:: 10.18653/v1/2025.findings-acl.344
Bibkey:
Cite (ACL):: Lian Yan, Chen Tang, Yi Guan, Haotian Wang, Songyuan Wang, Haifeng Liu, Yang Yang, and Jingchi Jiang. 2025. RLKGF: Reinforcement Learning from Knowledge Graph Feedback Without Human Annotations. In Findings of the Association for Computational Linguistics: ACL 2025, pages 6619–6633, Vienna, Austria. Association for Computational Linguistics.
Cite (Informal):: RLKGF: Reinforcement Learning from Knowledge Graph Feedback Without Human Annotations (Yan et al., Findings 2025)
Copy Citation:
PDF:: https://aclanthology.org/2025.findings-acl.344.pdf

PDF Cite Search Fix data