Web page clustering : a hyperlink-based similarity and matrix-based hierarchical algorithms


Autoria(s): Hou, Jingyu; Zhang, Yanchun; Cao, Jinli
Contribuinte(s)

Zhou, Xiaofang

Zhang, Yanchun

Orlowska, Maria

Data(s)

01/01/2003

Resumo

This paper proposes a hyperlink-based web page similarity measurement and two matrix-based hierarchical web page clustering algorithms. The web page similarity measurement incorporates hyperlink transitivity and page importance within the concerned web page space. One clustering algorithm takes cluster overlapping into account, another one does not. These algorithxms do not require predefined similarity thresholds for clustering, and are independent of the page order. The primary evaluations show the effectiveness of the proposed algorithms in clustering improvement.<br />

Identificador

http://hdl.handle.net/10536/DRO/DU:30005059

Idioma(s)

eng

Publicador

Springer

Relação

http://dro.deakin.edu.au/eserv/DU:30005059/hou-webpageclustering-2003.pdf

http://dx.doi.org/10.1007/3-540-36901-5_22

Direitos

2003, Springer-Verlag Berlin Heidelberg

Tipo

Conference Paper