BioSnowball: automated population of Wikis

13 years 3 months ago
BioSnowball: automated population of Wikis
Internet users regularly have the need to find biographies and facts of people of interest. Wikipedia has become the first stop for celebrity biographies and facts. However, Wikipedia can only provide information for celebrities because of its neutral point of view (NPOV) editorial policy. In this paper we propose an integrated bootstrapping framework named BioSnowball to automatically summarize the Web to generate Wikipedia-style pages for any person with a modest web presence. In BioSnowball, biography ranking and fact extraction are performed together in a single integrated training and inference process using Markov Logic Networks (MLNs) as its underlying statistical model. The bootstrapping framework starts with only a small number of seeds and iteratively finds new facts and biographies. As biography paragraphs on the Web are composed of the most important facts, our joint summarization model can improve the accuracy of both fact extraction and biography ranking compared to d...
Xiaojiang Liu, Zaiqing Nie, Nenghai Yu, Ji-Rong We
Added 18 Aug 2010
Updated 18 Aug 2010
Type Conference
Year 2010
Where KDD
Authors Xiaojiang Liu, Zaiqing Nie, Nenghai Yu, Ji-Rong Wen
Comments (0)