Difference between revisions of "Wikipedia and How to Use It for Semantic Document Representation"
(+ Embed) |
(Infobox) |
||
Line 1: | Line 1: | ||
{{Infobox work | {{Infobox work | ||
| title = Wikipedia and How to Use It for Semantic Document Representation | | title = Wikipedia and How to Use It for Semantic Document Representation | ||
− | | date = | + | | date = 2010 |
| authors = [[Ian H. Witten]] | | authors = [[Ian H. Witten]] | ||
− | | doi = 10. | + | | doi = 10.1007/978-3-642-16248-0_5 |
− | | link = http:// | + | | link = http://dl.acm.org/citation.cfm?id=1929344.1929350 |
}} | }} | ||
− | '''Wikipedia and How to Use It for Semantic Document Representation''' - scientific work related to [[Wikipedia quality]] published in | + | '''Wikipedia and How to Use It for Semantic Document Representation''' - scientific work related to [[Wikipedia quality]] published in 2010, written by [[Ian H. Witten]]. |
== Overview == | == Overview == | ||
− | + | Wikipedia is a goldmine of information; not just for its many readers, but also for the growing community of researchers who recognize it as a resource of exceptional scale and utility. It represents a vast investment of manual effort and judgment: a huge, constantly evolving tapestry of concepts and relations that is being applied to a host of tasks. This talk focuses on the process of "wikification"; that is, automatically and judiciously augmenting a plain-text document with pertinent hyperlinks to [[Wikipedia]] articlesas though the document were itself a Wikipedia article. Author first describe how Wikipedia can be used to determine semantic [[relatedness]] between concepts. Then Author explain how to wikify documents by exploiting Wikipedia's internal hyperlinks for relational information and their anchor texts as lexical information. Data mining techniques are used throughout to optimize the models involved. | |
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− | |||
− |
Revision as of 07:46, 25 February 2021
Authors | Ian H. Witten |
---|---|
Publication date | 2010 |
DOI | 10.1007/978-3-642-16248-0_5 |
Links | Original |
Wikipedia and How to Use It for Semantic Document Representation - scientific work related to Wikipedia quality published in 2010, written by Ian H. Witten.
Overview
Wikipedia is a goldmine of information; not just for its many readers, but also for the growing community of researchers who recognize it as a resource of exceptional scale and utility. It represents a vast investment of manual effort and judgment: a huge, constantly evolving tapestry of concepts and relations that is being applied to a host of tasks. This talk focuses on the process of "wikification"; that is, automatically and judiciously augmenting a plain-text document with pertinent hyperlinks to Wikipedia articlesas though the document were itself a Wikipedia article. Author first describe how Wikipedia can be used to determine semantic relatedness between concepts. Then Author explain how to wikify documents by exploiting Wikipedia's internal hyperlinks for relational information and their anchor texts as lexical information. Data mining techniques are used throughout to optimize the models involved.