Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Ahmed Tawfik

Efficient Intent-Based Filtering for Multi-Party Conversations Using Knowledge Distillation from LLMs

Mar 21, 2025

Reem Gody, Mohamed Abdelghaffar, Mohammed Jabreel, Ahmed Tawfik

Figure 1 for Efficient Intent-Based Filtering for Multi-Party Conversations Using Knowledge Distillation from LLMs

Figure 2 for Efficient Intent-Based Filtering for Multi-Party Conversations Using Knowledge Distillation from LLMs

Figure 3 for Efficient Intent-Based Filtering for Multi-Party Conversations Using Knowledge Distillation from LLMs

Figure 4 for Efficient Intent-Based Filtering for Multi-Party Conversations Using Knowledge Distillation from LLMs

Abstract:Large language models (LLMs) have showcased remarkable capabilities in conversational AI, enabling open-domain responses in chat-bots, as well as advanced processing of conversations like summarization, intent classification, and insights generation. However, these models are resource-intensive, demanding substantial memory and computational power. To address this, we propose a cost-effective solution that filters conversational snippets of interest for LLM processing, tailored to the target downstream application, rather than processing every snippet. In this work, we introduce an innovative approach that leverages knowledge distillation from LLMs to develop an intent-based filter for multi-party conversations, optimized for compute power constrained environments. Our method combines different strategies to create a diverse multi-party conversational dataset, that is annotated with the target intents and is then used to fine-tune the MobileBERT model for multi-label intent classification. This model achieves a balance between efficiency and performance, effectively filtering conversation snippets based on their intents. By passing only the relevant snippets to the LLM for further processing, our approach significantly reduces overall operational costs depending on the intents and the data distribution as demonstrated in our experiments.

Via

Access Paper or Ask Questions

Score Combination for Improved Parallel Corpus Filtering for Low Resource Conditions

Nov 16, 2020

Muhammad N. ElNokrashy, Amr Hendy, Mohamed Abdelghaffar, Mohamed Afify, Ahmed Tawfik, Hany Hassan Awadalla

Figure 1 for Score Combination for Improved Parallel Corpus Filtering for Low Resource Conditions

Figure 2 for Score Combination for Improved Parallel Corpus Filtering for Low Resource Conditions

Figure 3 for Score Combination for Improved Parallel Corpus Filtering for Low Resource Conditions

Figure 4 for Score Combination for Improved Parallel Corpus Filtering for Low Resource Conditions

Abstract:This paper describes our submission to the WMT20 sentence filtering task. We combine scores from (1) a custom LASER built for each source language, (2) a classifier built to distinguish positive and negative pairs by semantic alignment, and (3) the original scores included in the task devkit. For the mBART finetuning setup, provided by the organizers, our method shows 7% and 5% relative improvement over baseline, in sacreBLEU score on the test set for Pashto and Khmer respectively.

* Accepted at WMT20 (EMNLP 2020 Fifth Conference on Machine Translation)

Via

Access Paper or Ask Questions

Synthetic Data for Neural Machine Translation of Spoken-Dialects

Nov 28, 2017

Hany Hassan, Mostafa Elaraby, Ahmed Tawfik

Figure 1 for Synthetic Data for Neural Machine Translation of Spoken-Dialects

Figure 2 for Synthetic Data for Neural Machine Translation of Spoken-Dialects

Figure 3 for Synthetic Data for Neural Machine Translation of Spoken-Dialects

Figure 4 for Synthetic Data for Neural Machine Translation of Spoken-Dialects

Abstract:In this paper, we introduce a novel approach to generate synthetic data for training Neural Machine Translation systems. The proposed approach transforms a given parallel corpus between a written language and a target language to a parallel corpus between a spoken dialect variant and the target language. Our approach is language independent and can be used to generate data for any variant of the source language such as slang or spoken dialect or even for a different language that is closely related to the source language. The proposed approach is based on local embedding projection of distributed representations which utilizes monolingual embeddings to transform parallel data across language variants. We report experimental results on Levantine to English translation using Neural Machine Translation. We show that the generated data can improve a very large scale system by more than 2.8 Bleu points using synthetic spoken data which shows that it can be used to provide a reliable translation system for a spoken dialect that does not have sufficient parallel data.

Via

Access Paper or Ask Questions