☆ 4.5 Article

An Incremental Kernel Density Estimator for Data Stream Computation

COMPLEXITY (2020)

Journal

COMPLEXITY

Volume 2020, Issue -, Pages -

Publisher

WILEY-HINDAWI

DOI: 10.1155/2020/1803525

Keywords

Funding

National Key R&D Program of China [2017YFC0822604-2]
Scientific Research Foundation of Shenzhen University for Newly Introduced Teachers [2018060]
National Training Program of Innovation and Entrepreneurship for Undergraduates [201910590017]

Ask authors/readers for more resources

Protocol

Community support

Reagent

Community support

Abstract

Probability density function (p.d.f.) estimation plays a very important role in the field of data mining. Kernel density estimator (KDE) is the mostly used technology to estimate the unknown p.d.f. for the given dataset. The existing KDEs are usually inefficient when handling the p.d.f. estimation problem for stream data because a bran-new KDE has to be retrained based on the combination of current data and newly coming data. This process increases the training time and wastes the computation resource. This article proposes an incremental kernel density estimator (I-KDE) which deals with the p.d.f. estimation problem in the way of data stream computation. The I-KDE updates the current KDE dynamically and gradually with the newly coming data rather than retraining the bran-new KDE with the combination of current data and newly coming data. The theoretical analysis proves the convergence of the I-KDE only if the estimated p.d.f. of newly coming data is convergent to its true p.d.f. In order to guarantee the convergence of the I-KDE, a new multivariate fixed-point iteration algorithm based on the unbiased cross validation (UCV) method is developed to determine the optimal bandwidth of the KDE. The experimental results on 10 univariate and 4 multivariate probability distributions demonstrate the feasibility and effectiveness of the I-KDE.

An Incremental Kernel Density Estimator for Data Stream Computation

Journal

COMPLEXITY

Publisher

WILEY-HINDAWI

Keywords

Categories

Funding

Ask authors/readers for more resources

Protocol

Reagent

Authors

I am an author on this paper

Reviews

Primary Rating

Secondary Ratings

Novelty

Significance

Scientific rigor

Rate this paper

Recommended

An Incremental Kernel Density Estimator for Data Stream Computation

Journal

COMPLEXITY

Publisher

WILEY-HINDAWI

Keywords

Categories

Funding

Ask authors/readers for more resources

Protocol

Reagent

Authors

I am an author on this paper

Reviews

Primary Rating

Secondary Ratings

Novelty

Significance

Scientific rigor

Rate this paper

Recommended

Export Citation

Share Paper