TY - GEN
T1 - Measuring term informativeness in context
AU - Wu, Zhaohui
AU - Giles, C. Lee
N1 - Publisher Copyright:
© 2013 Association for Computational Linguistics.
PY - 2013
Y1 - 2013
N2 - Measuring term informativeness is a fundamental NLP task. Existing methods, mostly based on statistical information in corpora, do not actually measure informativeness of a term with regard to its semantic context. This paper proposes a new lightweight feature-free approach to encode term informativeness in context by leveraging web knowledge. Given a term and its context, we model contextaware term informativeness based on semantic similarity between the context and the term's most featured context in a knowledge base, Wikipedia. We apply our method to three applications: core term extraction from snippets (text segment), scientific keywords extraction (paper), and back-of-The-book index generation (book). The performance is state-of-Theart or close to it for each application, demonstrating its effectiveness and generality.
AB - Measuring term informativeness is a fundamental NLP task. Existing methods, mostly based on statistical information in corpora, do not actually measure informativeness of a term with regard to its semantic context. This paper proposes a new lightweight feature-free approach to encode term informativeness in context by leveraging web knowledge. Given a term and its context, we model contextaware term informativeness based on semantic similarity between the context and the term's most featured context in a knowledge base, Wikipedia. We apply our method to three applications: core term extraction from snippets (text segment), scientific keywords extraction (paper), and back-of-The-book index generation (book). The performance is state-of-Theart or close to it for each application, demonstrating its effectiveness and generality.
UR - http://www.scopus.com/inward/record.url?scp=84889563700&partnerID=8YFLogxK
UR - http://www.scopus.com/inward/citedby.url?scp=84889563700&partnerID=8YFLogxK
M3 - Conference contribution
AN - SCOPUS:84889563700
T3 - NAACL HLT 2013 - 2013 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Proceedings of the Main Conference
SP - 259
EP - 269
BT - Proceedings of the 2013 Conference of the North American Chapter of the Association for Computational Linguistics
PB - Association for Computational Linguistics (ACL)
T2 - 2013 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL HLT 2013
Y2 - 9 June 2013 through 14 June 2013
ER -