Mark Last

Senior Academic

Clustering and classification of web documents using a graph model

Adam Schenke, Horst Bunke, Mark Last, Abraham Kandel

In this chapter we provide a summary of our previous work concerning the application of traditional machine learning techniques to data represented by graphs. We show how the fc-means clustering algorithm and the A;-nearest neighbors classification algorithm can easily and intuitively be extended from dealing with vector representations to graph representations. We present some of our experimental results, which confirm that the addition of structural information, not present in vector representations, improves both clustering and classification performance when dealing with web documents.

Publication language English
Pages 287-302
Publication status Published - 01.01.2005

ASJC Scopus subject areas

General Computer Science
Access to Document
10.1142/9789812775320_0016
Other files and links
Link to publication in Scopus