Abstract
Nowadays, advanced analysis of data streams is quickly becoming a key area of data mining research, as the number of applications demanding such processing increases. Online mining when such data streams evolve over time, that is, when concepts drift or change completely, is becoming one of the core issues. At the same time, closure-based mining on relational data has recently provided some interesting algorithmic developments as well as practical uses. In this chapter we show how to use closure-based mining to reduce drastically the number of attributes in XML tree classification tasks. Moreover, using maximal frequent trees, we reduce even more the number of attributes needed in tree classification, in many cases without losing accuracy. We show a general framework to classify XML trees using subtree occurrence, composing a Tree XML Closed Frequent Miner with a classifier algorithm. We present specific methods that can adaptively mining closed patterns from data streams that change over time.
| Original language | English |
|---|---|
| Title of host publication | XML Data Mining |
| Subtitle of host publication | Models, Methods, and Applications |
| Publisher | IGI Global |
| Pages | 199-218 |
| Number of pages | 20 |
| ISBN (Electronic) | 9781613503577 |
| ISBN (Print) | 9781613503560 |
| DOIs | |
| Publication status | Published - 1 Jan 2011 |
| Externally published | Yes |
Fingerprint
Dive into the research topics of 'XML Tree Classification on Evolving Data Streams1'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver