Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:RuleR: Improving LLM Controllability by Rule-based Data Recycling

Jun 22, 2024

Ming Li, Han Chen, Chenguang Wang, Dang Nguyen, Dianqi Li, Tianyi Zhou

Figure 1 for RuleR: Improving LLM Controllability by Rule-based Data Recycling

Figure 2 for RuleR: Improving LLM Controllability by Rule-based Data Recycling

Figure 3 for RuleR: Improving LLM Controllability by Rule-based Data Recycling

Figure 4 for RuleR: Improving LLM Controllability by Rule-based Data Recycling

Share this with someone who'll enjoy it:

Abstract:Large language models (LLMs) still lack delicate controllability over their responses, which is critical to enhancing their performance and the user experience. However, curating supervised fine-tuning (SFT) datasets to improve LLM controllability usually relies on human experts or proprietary LLMs, which requires additional costs. To bridge this gap, we propose Rule-based Data Recycling (RuleR), a data augmentation method incorporating multiple constraints into the original data samples according to predefined rules, which creates new training tasks to consolidate the controllability of LLMs. Instead of creating new data from scratch, RuleR ``recycles'' existing data by simply applying rule-based edits to their responses and appending the rule-instructions in their original instructions. Experimental results demonstrate RuleR's effectiveness in improving LLM controllability while maintaining general instruction-following capabilities. The code will be released on https://github.com/MingLiiii/RuleR.

View paper on

Share this with someone who'll enjoy it:

Title:RuleR: Improving LLM Controllability by Rule-based Data Recycling

Paper and Code