Title
WebSum: Enhanced SumBasic algorithm for Web site summarization.
Abstract
Due to the rapid increase of information in the World Wide Web, there exists an explosion of information on the Web that may overwhelm the common Web user. The Web user may find it quicker or more efficient to browse the Web by reading summaries of Web sites. This paper proposes WebSum to compress Web site content into a summary. WebSum is an enhancement of the SumBasic algorithm, that was mainly used for multi-document summarization. In the case of Web sites, we find that several Web characteristics such as title and keywords can be used to extract sentences that may represent the overall topic of the Web site. Initial results show that WebSum is able to reveal sentences relate to the concept of the Web site. WebSum is then evaluated against the original algorithm of SumBasic.
Year
DOI
Venue
2012
10.1109/DMO.2012.6329812
DMO
Keywords
Field
DocType
Web sites,document handling,SumBasic algorithm,Web site content compression,Web site summarization,WebSum,World Wide Web,keywords,multidocument summarization,sentence extraction,title
Static web page,Web development,Data mining,Web page,Computer science,Web modeling,World Wide Web,Information retrieval,Web standards,Data Web,Algorithm,Web navigation,Web service
Conference
Citations 
PageRank 
References 
0
0.34
7
Authors
3
Name
Order
Citations
PageRank
Jason Yong-Jin Tee111.06
Lay-Ki Soon22211.04
Choo-Yee Ting39013.19