Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Atsushi Fujii

University of Library and Information Science

Effects of Language Modeling on Speech-driven Question Answering

Jul 10, 2004

Tomoyosi Akiba, Atsushi Fujii, Katunobu Itou

Figure 1 for Effects of Language Modeling on Speech-driven Question Answering

Figure 2 for Effects of Language Modeling on Speech-driven Question Answering

Figure 3 for Effects of Language Modeling on Speech-driven Question Answering

Abstract:We integrate automatic speech recognition (ASR) and question answering (QA) to realize a speech-driven QA system, and evaluate its performance. We adapt an N-gram language model to natural language questions, so that the input of our system can be recognized with a high accuracy. We target WH-questions which consist of the topic part and fixed phrase used to ask about something. We first produce a general N-gram model intended to recognize the topic and emphasize the counts of the N-grams that correspond to the fixed phrases. Given a transcription by the ASR engine, the QA engine extracts the answer candidates from target documents. We propose a passage retrieval method robust against recognition errors in the transcription. We use the QA test collection produced in NTCIR, which is a TREC-style evaluation workshop, and show the effectiveness of our method by means of experiments.

* Proceedings of the 8th International Conference on Spoken Language Processing (ICSLP 2004), pp.1053-1056, Oct. 2004
* 4 pages, Proceedings of the 8th International Conference on Spoken Language Processing (to appear)

Via

Access Paper or Ask Questions

Unsupervised Topic Adaptation for Lecture Speech Retrieval

Jul 10, 2004

Atsushi Fujii, Katunobu Itou, Tomoyosi Akiba, Tetsuya Ishikawa

Figure 1 for Unsupervised Topic Adaptation for Lecture Speech Retrieval

Figure 2 for Unsupervised Topic Adaptation for Lecture Speech Retrieval

Figure 3 for Unsupervised Topic Adaptation for Lecture Speech Retrieval

Abstract:We are developing a cross-media information retrieval system, in which users can view specific segments of lecture videos by submitting text queries. To produce a text index, the audio track is extracted from a lecture video and a transcription is generated by automatic speech recognition. In this paper, to improve the quality of our retrieval system, we extensively investigate the effects of adapting acoustic and language models on speech recognition. We perform an MLLR-based method to adapt an acoustic model. To obtain a corpus for language model adaptation, we use the textbook for a target lecture to search a Web collection for the pages associated with the lecture topic. We show the effectiveness of our method by means of experiments.

* Proceedings of the 8th International Conference on Spoken Language Processing (ICSLP 2004), pp.2957-2960, Oct. 2004
* 4 pages, Proceedings of the 8th International Conference on Spoken Language Processing (to appear)

Via

Access Paper or Ask Questions

Summarizing Encyclopedic Term Descriptions on the Web

Jul 10, 2004

Atsushi Fujii, Tetsuya Ishikawa

Figure 1 for Summarizing Encyclopedic Term Descriptions on the Web

Figure 2 for Summarizing Encyclopedic Term Descriptions on the Web

Figure 3 for Summarizing Encyclopedic Term Descriptions on the Web

Figure 4 for Summarizing Encyclopedic Term Descriptions on the Web

Abstract:We are developing an automatic method to compile an encyclopedic corpus from the Web. In our previous work, paragraph-style descriptions for a term are extracted from Web pages and organized based on domains. However, these descriptions are independent and do not comprise a condensed text as in hand-crafted encyclopedias. To resolve this problem, we propose a summarization method, which produces a single text from multiple descriptions. The resultant summary concisely describes a term from different viewpoints. We also show the effectiveness of our method by means of experiments.

* Proceedings of the 20th International Conference on Computational Linguistics (COLING 2004), pp.645-651, Aug. 2004
* 7 pages, Proceedings of the 20th International Conference on Computational Linguistics (to appear)

Via

Access Paper or Ask Questions

Test Collections for Patent-to-Patent Retrieval and Patent Map Generation in NTCIR-4 Workshop

Apr 10, 2004

Atsushi Fujii, Makoto Iwayama, Noriko Kando

Figure 1 for Test Collections for Patent-to-Patent Retrieval and Patent Map Generation in NTCIR-4 Workshop

Abstract:This paper describes the Patent Retrieval Task in the Fourth NTCIR Workshop, and the test collections produced in this task. We perform the invalidity search task, in which each participant group searches a patent collection for the patents that can invalidate the demand in an existing claim. We also perform the automatic patent map generation task, in which the patents associated with a specific topic are organized in a multi-dimensional matrix.

* Proceedings of the 4th International Conference on Language Resources and Evaluation (LREC-2004), pp.1643-1646, May. 2004.
* 4 pages, Proceedings of the 4th International Conference on Language Resources and Evaluation (to appear)

Via

Access Paper or Ask Questions

A Cross-media Retrieval System for Lecture Videos

Sep 13, 2003

Atsushi Fujii, Katunobu Itou, Tomoyosi Akiba, Tetsuya Ishikawa

Figure 1 for A Cross-media Retrieval System for Lecture Videos

Figure 2 for A Cross-media Retrieval System for Lecture Videos

Figure 3 for A Cross-media Retrieval System for Lecture Videos

Abstract:We propose a cross-media lecture-on-demand system, in which users can selectively view specific segments of lecture videos by submitting text queries. Users can easily formulate queries by using the textbook associated with a target lecture, even if they cannot come up with effective keywords. Our system extracts the audio track from a target lecture video, generates a transcription by large vocabulary continuous speech recognition, and produces a text index. Experimental results showed that by adapting speech recognition to the topic of the lecture, the recognition accuracy increased and the retrieval accuracy was comparable with that obtained by human transcription.

* Proceedings of the 8th European Conference on Speech Communication and Technology (Eurospeech 2003), pp.1149-1152, Sep. 2003

Via

Access Paper or Ask Questions

Building a Test Collection for Speech-Driven Web Retrieval

Sep 12, 2003

Atsushi Fujii, Katunobu Itou

Figure 1 for Building a Test Collection for Speech-Driven Web Retrieval

Figure 2 for Building a Test Collection for Speech-Driven Web Retrieval

Figure 3 for Building a Test Collection for Speech-Driven Web Retrieval

Figure 4 for Building a Test Collection for Speech-Driven Web Retrieval

Abstract:This paper describes a test collection (benchmark data) for retrieval systems driven by spoken queries. This collection was produced in the subtask of the NTCIR-3 Web retrieval task, which was performed in a TREC-style evaluation workshop. The search topics and document collection for the Web retrieval task were used to produce spoken queries and language models for speech recognition, respectively. We used this collection to evaluate the performance of our retrieval system. Experimental results showed that (a) the use of target documents for language modeling and (b) enhancement of the vocabulary size in speech recognition were effective in improving the system performance.

* Proceedings of the 8th European Conference on Speech Communication and Technology (Eurospeech 2003), pp.1153-1156, Sep. 2003

Via

Access Paper or Ask Questions

Speech-Driven Text Retrieval: Using Target IR Collections for Statistical Language Model Adaptation in Speech Recognition

Jun 24, 2002

Atsushi Fujii, Katunobu Itou, Tetsuya Ishikawa

Figure 1 for Speech-Driven Text Retrieval: Using Target IR Collections for Statistical Language Model Adaptation in Speech Recognition

Figure 2 for Speech-Driven Text Retrieval: Using Target IR Collections for Statistical Language Model Adaptation in Speech Recognition

Figure 3 for Speech-Driven Text Retrieval: Using Target IR Collections for Statistical Language Model Adaptation in Speech Recognition

Figure 4 for Speech-Driven Text Retrieval: Using Target IR Collections for Statistical Language Model Adaptation in Speech Recognition

Abstract:Speech recognition has of late become a practical technology for real world applications. Aiming at speech-driven text retrieval, which facilitates retrieving information with spoken queries, we propose a method to integrate speech recognition and retrieval methods. Since users speak contents related to a target collection, we adapt statistical language models used for speech recognition based on the target collection, so as to improve both the recognition and retrieval accuracy. Experiments using existing test collections combined with dictated queries showed the effectiveness of our method.

* Anni R. Coden and Eric W. Brown and Savitha Srinivasan (Eds.), Information Retrieval Techniques for Speech Applications (LNCS 2273), pp.94-104, Springer, 2002

Via

Access Paper or Ask Questions

Language Modeling for Multi-Domain Speech-Driven Text Retrieval

Jun 24, 2002

Katunobu Itou, Atsushi Fujii, Tetsuya Ishikawa

Figure 1 for Language Modeling for Multi-Domain Speech-Driven Text Retrieval

Figure 2 for Language Modeling for Multi-Domain Speech-Driven Text Retrieval

Figure 3 for Language Modeling for Multi-Domain Speech-Driven Text Retrieval

Figure 4 for Language Modeling for Multi-Domain Speech-Driven Text Retrieval

Abstract:We report experimental results associated with speech-driven text retrieval, which facilitates retrieving information in multiple domains with spoken queries. Since users speak contents related to a target collection, we produce language models used for speech recognition based on the target collection, so as to improve both the recognition and retrieval accuracy. Experiments using existing test collections combined with dictated queries showed the effectiveness of our method.

* IEEE Automatic Speech Recognition and Understanding Workshop, Dec. 2001

Via

Access Paper or Ask Questions

PRIME: A System for Multi-lingual Patent Retrieval

Jun 24, 2002

Shigeto Higuchi, Masatoshi Fukui, Atsushi Fujii, Tetsuya Ishikawa

Figure 1 for PRIME: A System for Multi-lingual Patent Retrieval

Figure 2 for PRIME: A System for Multi-lingual Patent Retrieval

Abstract:Given the growing number of patents filed in multiple countries, users are interested in retrieving patents across languages. We propose a multi-lingual patent retrieval system, which translates a user query into the target language, searches a multilingual database for patents relevant to the query, and improves the browsing efficiency by way of machine translation and clustering. Our system also extracts new translations from patent families consisting of comparable patents, to enhance the translation dictionary.

* Proceedings of MT Summit VIII, pp.163-167, Sep. 2001

Via

Access Paper or Ask Questions

Applying a Hybrid Query Translation Method to Japanese/English Cross-Language Patent Retrieval

Jun 24, 2002

Masatoshi Fukui, Shigeto Higuchi, Youichi Nakatani, Masao Tanaka, Atsushi Fujii, Tetsuya Ishikawa

Figure 1 for Applying a Hybrid Query Translation Method to Japanese/English Cross-Language Patent Retrieval

Figure 2 for Applying a Hybrid Query Translation Method to Japanese/English Cross-Language Patent Retrieval

Figure 3 for Applying a Hybrid Query Translation Method to Japanese/English Cross-Language Patent Retrieval

Figure 4 for Applying a Hybrid Query Translation Method to Japanese/English Cross-Language Patent Retrieval

Abstract:This paper applies an existing query translation method to cross-language patent retrieval. In our method, multiple dictionaries are used to derive all possible translations for an input query, and collocational statistics are used to resolve translation ambiguity. We used Japanese/English parallel patent abstracts to perform comparative experiments, where our method outperformed a simple dictionary-based query translation method, and achieved 76% of monolingual retrieval in terms of average precision.

* ACM SIGIR 2000 Workshop on Patent Retrieval, July, 2000

Via

Access Paper or Ask Questions