Skip to main navigation Skip to search Skip to main content

XML Tree Classification on Evolving Data Streams1

  • University of Waikato
  • Universidad Politecnica de Catalunia

Research output: Chapter in Book/Report/Conference proceedingChapterpeer-review

Abstract

Nowadays, advanced analysis of data streams is quickly becoming a key area of data mining research, as the number of applications demanding such processing increases. Online mining when such data streams evolve over time, that is, when concepts drift or change completely, is becoming one of the core issues. At the same time, closure-based mining on relational data has recently provided some interesting algorithmic developments as well as practical uses. In this chapter we show how to use closure-based mining to reduce drastically the number of attributes in XML tree classification tasks. Moreover, using maximal frequent trees, we reduce even more the number of attributes needed in tree classification, in many cases without losing accuracy. We show a general framework to classify XML trees using subtree occurrence, composing a Tree XML Closed Frequent Miner with a classifier algorithm. We present specific methods that can adaptively mining closed patterns from data streams that change over time.

Original languageEnglish
Title of host publicationXML Data Mining
Subtitle of host publicationModels, Methods, and Applications
PublisherIGI Global
Pages199-218
Number of pages20
ISBN (Electronic)9781613503577
ISBN (Print)9781613503560
DOIs
Publication statusPublished - 1 Jan 2011
Externally publishedYes

Fingerprint

Dive into the research topics of 'XML Tree Classification on Evolving Data Streams1'. Together they form a unique fingerprint.

Cite this