Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Context-dependent Instruction Tuning for Dialogue Response Generation

Nov 13, 2023

Jin Myung Kwak, Minseon Kim, Sung Ju Hwang

Figure 1 for Context-dependent Instruction Tuning for Dialogue Response Generation

Figure 2 for Context-dependent Instruction Tuning for Dialogue Response Generation

Figure 3 for Context-dependent Instruction Tuning for Dialogue Response Generation

Figure 4 for Context-dependent Instruction Tuning for Dialogue Response Generation

Share this with someone who'll enjoy it:

Abstract:Recent language models have achieved impressive performance in natural language tasks by incorporating instructions with task input during fine-tuning. Since all samples in the same natural language task can be explained with the same task instructions, many instruction datasets only provide a few instructions for the entire task, without considering the input of each example in the task. However, this approach becomes ineffective in complex multi-turn dialogue generation tasks, where the input varies highly with each turn as the dialogue context changes, so that simple task instructions cannot improve the generation performance. To address this limitation, we introduce a context-based instruction fine-tuning framework for each multi-turn dialogue which generates both responses and instructions based on the previous context as input. During the evaluation, the model generates instructions based on the previous context to self-guide the response. The proposed framework produces comparable or even outstanding results compared to the baselines by aligning instructions to the input during fine-tuning with the instructions in quantitative evaluations on dialogue benchmark datasets with reduced computation budget.

* Work in Progress

View paper on

Share this with someone who'll enjoy it:

Title:Context-dependent Instruction Tuning for Dialogue Response Generation

Paper and Code