A reliable FAQ retrieval system using a query log classification technique based on latent semantic analysis

  • Kim, Harksoo
  • Lee, Hyunjung
  • Seo, Jungyun
Citations

WEB OF SCIENCE

12
Citations

SCOPUS

16

초록

To obtain high performances, previous works on FAQ retrieval used high-level knowledge bases or handcrafted rules. However, it is a time and effort consuming job to construct these knowledge bases and rules whenever application domains are changed. To overcome this problem, we propose a high-performance FAQ retrieval system only using users' query logs as knowledge sources. During indexing time, the proposed system efficiently clusters users' query logs using classification techniques based on latent semantic analysis. During retrieval time, the proposed system smoothes FAQs using the query log clusters. In the experiment, the proposed system outperformed the conventional information retrieval systems in FAQ retrieval. Based on various experiments, we found that the proposed system could alleviate critical lexical disagreement problems in short document retrieval. In addition, we believe that the proposed system is more practical and reliable than the previous FAQ retrieval systems because it uses only data-driven methods without high-level knowledge sources. (c) 2006 Elsevier Ltd. All rights reserved.

키워드

FAQ retrievallexical disagreement problemquery log clusterslatent semantic analysis
제목
A reliable FAQ retrieval system using a query log classification technique based on latent semantic analysis
저자
Kim, HarksooLee, HyunjungSeo, Jungyun
DOI
10.1016/j.ipm.2006.07.018
발행일
2007-03
유형
Article; Proceedings Paper
저널명
Information Processing and Management
43
2
페이지
420 ~ 430