WebSum: Enhanced SumBasic algorithm for Web site summarization


Tee, J.Y.-J. and Soon, L.-K. and Ting, C.-Y (2012) WebSum: Enhanced SumBasic algorithm for Web site summarization. Conference on Data Mining and Optimization. pp. 137-142. ISSN 21556938

[img] PDF
Restricted to Repository staff only

Download (0B)


Due to the rapid increase of information in the World Wide Web, there exists an explosion of information on the Web that may overwhelm the common Web user. The Web user may find it quicker or more efficient to browse the Web by reading summaries of Web sites. This paper proposes WebSum to compress Web site content into a summary. WebSum is an enhancement of the SumBasic algorithm, that was mainly used for multi-document summarization. In the case of Web sites, we find that several Web characteristics such as title and keywords can be used to extract sentences that may represent the overall topic of the Web site. Initial results show that WebSum is able to reveal sentences relate to the concept of the Web site. WebSum is then evaluated against the original algorithm of SumBasic.

Item Type: Article
Subjects: T Technology > TK Electrical engineering. Electronics Nuclear engineering > TK5101-6720 Telecommunication. Including telegraphy, telephone, radio, radar, television
Depositing User: Users 1102 not found.
Date Deposited: 07 Jan 2013 01:56
Last Modified: 07 Jan 2013 01:56
URII: http://shdl.mmu.edu.my/id/eprint/3780


Downloads per month over past year

View ItemEdit (login required)