Learning to rank for<Emphasis Type="Italic"> why</Emphasis>-question answering期刊界 All Journals 搜尽天下杂志传播学术成果专业期刊搜索期刊信息化学术搜索

按检索

Learning to rank for<Emphasis Type="Italic"> why</Emphasis>-question answering

Authors:	Suzan Verberne Hans van Halteren Daphne Theijssen Stephan Raaijmakers Lou Boves

Institution:	(1) Centre for Language and Speech Technology, Radboud University, Nijmegen, Netherlands;(2) Department of Linguistics, Radboud University, Nijmegen, Netherlands;(3) TNO Information and Communication Technology, Delft, Netherlands

Abstract:	In this paper, we evaluate a number of machine learning techniques for the task of ranking answers to why-questions. We use TF-IDF together with a set of 36 linguistically motivated features that characterize questions and answers. We experiment with a number of machine learning techniques (among which several classifiers and regression techniques, Ranking SVM and SVM ^map) in various settings. The purpose of the experiments is to assess how the different machine learning approaches can cope with our highly imbalanced binary relevance data, with and without hyperparameter tuning. We find that with all machine learning techniques, we can obtain an MRR score that is significantly above the TF-IDF baseline of 0.25 and not significantly lower than the best score of 0.35. We provide an in-depth analysis of the effect of data imbalance and hyperparameter tuning, and we relate our findings to previous research on learning to rank for Information Retrieval.

Keywords:
本文献已被 SpringerLink 等数据库收录！