Abstract:
In this position paper, we discuss how Exploratory Data Analysis (EDA) and Machine Learning (ML) can work together in large-scale data analysis environments. In particula...Show MoreMetadata
Abstract:
In this position paper, we discuss how Exploratory Data Analysis (EDA) and Machine Learning (ML) can work together in large-scale data analysis environments. In particular, we describe how applying EDA techniques and ML methods in a complementary fashion can be used to address some of the challenges faced when applying ML techniques to large, real world data sets, and discuss tools that help do the job. This iterative approach is demonstrated with a simple example of how extracting events from a historical sensor data set was enabled by iteratively identifying and filtering various types of erroneous data.
Published in: 2013 IEEE International Symposium on Parallel & Distributed Processing, Workshops and Phd Forum
Date of Conference: 20-24 May 2013
Date Added to IEEE Xplore: 31 October 2013
Electronic ISBN:978-0-7695-4979-8