most citedAn Optimization Model for Outlier Detection in Categorical Data

58 citations · 205 across the 10 of their papers we have counts for

collaborators

10 papers

cs.AI200531 cited

K-Histograms: An Efficient Clustering Algorithm for Categorical Dataset

Zengyou He, Xiaofei Xu, Shengchun Deng +1

Clustering categorical data is an integral part of data mining and has attracted much attention recently. In this paper, we present k-histogram, a new efficient algorithm for clust…

cs.AI200552 cited

Clustering Mixed Numeric and Categorical Data: A Cluster Ensemble Approach

Zengyou He, Xiaofei Xu, Shengchun Deng

Clustering is a widely used technique in data mining applications for discovering patterns in underlying data. Most traditional clustering algorithms are limited to handling datase…

cs.DB200515 cited

A Fast Greedy Algorithm for Outlier Mining

Zengyou He, Xiaofei Xu, Shengchun Deng

The task of outlier detection is to find small groups of data objects that are exceptional when compared with rest large amount of data. In [38], the problem of outlier detection i…

cs.DB20053 cited

A Unified Subspace Outlier Ensemble Framework for Outlier Detection in High Dimensional Spaces

Zengyou He, Xiaofei Xu, Shengchun Deng

The task of outlier detection is to find small groups of data objects that are exceptional when compared with rest large amount of data. Detection of such outliers is important for…

cs.DB200558 cited

An Optimization Model for Outlier Detection in Categorical Data

Zengyou He, Xiaofei Xu, Shengchun Deng

The task of outlier detection is to find small groups of data objects that are exceptional when compared with rest large amount of data. Detection of such outliers is important for…

cs.DB200530 cited

Data Mining for Actionable Knowledge: A Survey

Zengyou He, Xiaofei Xu, Shengchun Deng

The data mining process consists of a series of steps ranging from data cleaning, data selection and transformation, to pattern evaluation and visualization. One of the central pro…