Abstract
Due to the rapid increase of information in the World Wide Web, there exists an explosion of information on the Web that may overwhelm the common Web user. The Web user may find it quicker or more efficient to browse the Web by reading summaries of Web sites. This paper proposes WebSum to compress Web site content into a summary. WebSum is an enhancement of the SumBasic algorithm, that was mainly used for multi-document summarization. In the case of Web sites, we find that several Web characteristics such as title and keywords can be used to extract sentences that may represent the overall topic of the Web site. Initial results show that WebSum is able to reveal sentences relate to the concept of the Web site. WebSum is then evaluated against the original algorithm of SumBasic.
Original language | English |
---|---|
Title of host publication | Proceedings - 2012 4th Conference on Data Mining and Optimization, DMO 2012 |
Pages | 137-142 |
Number of pages | 6 |
DOIs | |
Publication status | Published - 2012 |
Externally published | Yes |
Event | Conference on Data Mining and Optimization 2012 - Langkawi, Malaysia Duration: 2 Sept 2012 → 4 Sept 2012 Conference number: 4th https://ieeexplore.ieee.org/xpl/conhome/6322848/proceeding (Proceedings) |
Publication series
Name | Conference on Data Mining and Optimization |
---|---|
ISSN (Print) | 2155-6938 |
ISSN (Electronic) | 2155-6946 |
Conference
Conference | Conference on Data Mining and Optimization 2012 |
---|---|
Abbreviated title | DMO 2012 |
Country/Territory | Malaysia |
City | Langkawi |
Period | 2/09/12 → 4/09/12 |
Internet address |