Discovering links between lexical and surface features in questions and answers
DSpace at IIT Bombay
View Archive InfoField | Value | |
Title |
Discovering links between lexical and surface features in questions and answers
|
|
Creator |
CHAKRABARTI, S
|
|
Description |
Information retrieval systems, based on keyword match, are evolving to question answering systems that return short passages or direct answers to questions, rather than URLs pointing to whole pages. Most open-domain question answering systems depend on manually designed hierarchies of question types. A question is first classified to a fixed type, and then hand-engineered rules associated with the type yield keywords and/or predictive annotations that are likely to match indexed answer passages. Here we seek a more data-driven approach, assisted by machine learning. We propose a simple log-linear model over a pair of feature vectors, one derived from the question and the other derived from the a candidate passage. Features are extracted using a lexical network and surface context as in named entity extraction, except that there is no direct supervision available in the form of fixed entity types and their examples. Using the log-linear model, we filter candidate passages and see substantial improvement in the mean rank at which the first answer is found. The model parameters distill and reveal linguistic artifacts coupling questions and their answers, which can be used for better annotation and indexing.
|
|
Publisher |
SPRINGER-VERLAG BERLIN
|
|
Date |
2011-10-23T19:41:35Z
2011-12-15T09:10:45Z 2011-10-23T19:41:35Z 2011-12-15T09:10:45Z 2006 |
|
Type |
Article; Proceedings Paper
|
|
Identifier |
ADVANCES IN WEB MINING AND WEB USAGE ANALYSIS,3932,116-134
3-540-47127-8 0302-9743 http://dspace.library.iitb.ac.in/xmlui/handle/10054/15219 http://hdl.handle.net/100/1642 |
|
Source |
6th International Workshop on Knowledge Discovery on the Web (WEBKDD 2004),Seattle, WA,AUG 22-25, 2004
|
|
Language |
English
|
|