Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Ruiyang Ke

Adaptive Reinforcement Learning Planning: Harnessing Large Language Models for Complex Information Extraction

Jun 17, 2024

Zepeng Ding, Ruiyang Ke, Wenhao Huang, Guochao Jiang, Yanda Li, Deqing Yang, Yanghua Xiao, Jiaqing Liang

Figure 1 for Adaptive Reinforcement Learning Planning: Harnessing Large Language Models for Complex Information Extraction

Figure 2 for Adaptive Reinforcement Learning Planning: Harnessing Large Language Models for Complex Information Extraction

Figure 3 for Adaptive Reinforcement Learning Planning: Harnessing Large Language Models for Complex Information Extraction

Figure 4 for Adaptive Reinforcement Learning Planning: Harnessing Large Language Models for Complex Information Extraction

Abstract:Existing research on large language models (LLMs) shows that they can solve information extraction tasks through multi-step planning. However, their extraction behavior on complex sentences and tasks is unstable, emerging issues such as false positives and missing elements. We observe that decomposing complex extraction tasks and extracting them step by step can effectively improve LLMs' performance, and the extraction orders of entities significantly affect the final results of LLMs. This paper proposes a two-stage multi-step method for LLM-based information extraction and adopts the RL framework to execute the multi-step planning. We regard sequential extraction as a Markov decision process, build an LLM-based extraction environment, design a decision module to adaptively provide the optimal order for sequential entity extraction on different sentences, and utilize the DDQN algorithm to train the decision model. We also design the rewards and evaluation metrics suitable for the extraction results of LLMs. We conduct extensive experiments on multiple public datasets to demonstrate the effectiveness of our method in improving the information extraction capabilities of LLMs.

Via

Access Paper or Ask Questions