Fissaha Adafre Sisay, Jijkoun Valentin & de Rijke Maarten (2007). « Fact Discovery in Wikipedia ». In 2007 IEEE/WIC/ACM International Conference on Web Intelligence, novembre 2007.
Categories: 4. interfaces et modes de consultation
Keywords: extraction d'information, interface, recherche d'information, veille éditoriale, Wikipedia
Creators: Fissaha Adafre, Jijkoun, de Rijke
Collection: 2007 IEEE/WIC/ACM International Conference on Web Intelligence

We address the task of extracting focused salient information items, relevant and important for a given topic, from a large encyclopedic resource. Specifically, for a given topic (a Wikipedia article) we identify snippets from other articles in Wikipedia that contain important information for the topic of the original article, without duplicates. We compare several methods for addressing the task, and find that a mixture of content-based, link-based, and layout-based features outperforms other methods, especially in combination with the use of so-called reference corpora that capture the key properties of entities of a common type.
