Practical methods for data mining with massive data sets
Files
Date
Authors
Journal Title
Journal ISSN
Volume Title
Publisher
Abstract
The increasing size of data sets has necessitated advancement in exploratory techniques. Methods that are practical for moderate to small data sets become infeasible when applied to massive data sets. Advanced techniques such as binned kernel density estimation, tours, and mode-based projection pursuit will be explored. Mean-centered binning will be introduced as an improved method for binned density estimation. The density grand tour will be demonstrated as a means of exploring massive high-dimensional data sets. Projection pursuit by clustering components will be described as a means to find interesting lower-dimensional subspaces of data sets.
Description
Advisor
Degree
Type
Keywords
Citation
Salch, John David. "Practical methods for data mining with massive data sets." (1998) Diss., Rice University. https://hdl.handle.net/1911/19307.