Search CORE

171 research outputs found

Non-convex clustering using expectation maximization algorithm with rough set initialization

Author: Mitra Pabitra
Pal Sankar K.
Siddiqi Md Aleemuddin
Publication venue: 'Elsevier BV'
Publication date: 01/03/2003
Field of study

An integration of a minimal spanning tree (MST) based graph-theoretic technique and expectation maximization (EM) algorithm with rough set initialization is described for non-convex clustering. EM provides the statistical model of the data and handles the associated uncertainties. Rough set theory helps in faster convergence and avoidance of the local minima problem, thereby enhancing the performance of EM. MST helps in determining non-convex clusters. Since it is applied on Gaussians rather than the original data points, time required is very low. These features are demonstrated on real life datasets. Comparison with related methods is made in terms of a cluster quality measure and computation time

Abstracts Of Papers Presented At The Third International Conference Of Business, Economics and Management Disciplines 2006

Author
Publication venue: CEDIMES Institute USA
Publication date: 01/12/2006
Field of study

University of New Brunswick: Centre for Digital Scholarship Journals

Advances in Data Mining Knowledge Discovery and Applications

Author
Publication venue: 'IntechOpen'
Publication date: 20/04/2021
Field of study

Advances in Data Mining Knowledge Discovery and Applications aims to help data miners, researchers, scholars, and PhD students who wish to apply data mining techniques. The primary contribution of this book is highlighting frontier fields and implementations of the knowledge discovery and data mining. It seems to be same things are repeated again. But in general, same approach and techniques may help us in different fields and expertise areas. This book presents knowledge discovery and data mining applications in two different sections. As known that, data mining covers areas of statistics, machine learning, data management and databases, pattern recognition, artificial intelligence, and other areas. In this book, most of the areas are covered with different data mining applications. The eighteen chapters have been classified in two parts: Knowledge Discovery and Data Mining Applications

Directory of Open Access Books (DOAB)

Anonymization procedures for tabular data: an explanatory technical and legal synthesis

Author: Aufschläger Robert
Buchner Benedikt
Folz Jakob
Guggumos Johann
Heigl Michael
März Elena
Schramm Martin
Publication venue
Publication date: 01/09/2023
Field of study

In the European Union, Data Controllers and Data Processors, who work with personal data, have to comply with the General Data Protection Regulation and other applicable laws. This affects the storing and processing of personal data. But some data processing in data mining or statistical analyses does not require any personal reference to the data. Thus, personal context can be removed. For these use cases, to comply with applicable laws, any existing personal information has to be removed by applying the so-called anonymization. However, anonymization should maintain data utility. Therefore, the concept of anonymization is a double-edged sword with an intrinsic trade-off: privacy enforcement vs. utility preservation. The former might not be entirely guaranteed when anonymized data are published as Open Data. In theory and practice, there exist diverse approaches to conduct and score anonymization. This explanatory synthesis discusses the technical perspectives on the anonymization of tabular data with a special emphasis on the European Union’s legal base. The studied methods for conducting anonymization, and scoring the anonymization procedure and the resulting anonymity are explained in unifying terminology. The examined methods and scores cover both categorical and numerical data. The examined scores involve data utility, information preservation, and privacy models. In practice-relevant examples, methods and scores are experimentally tested on records from the UCI Machine Learning Repository’s “Census Income (Adult)” dataset

OPUS Augsburg

Directory of Open Access Journals