Sarka Brodinova scite author profile

Sarka Brodinova

4Publications

54Citation Statements Received

74Citation Statements Given

How they've been cited

How they cite others

Affiliations

TU Wien, Statistics Austria

Publications

Order By: Most citations

Robust and sparse k-means clustering for high-dimensional data

Brodinova

Filzmoser

Ortner

et al. 2019

Adv Data Anal Classif

View full text Add to dashboard Cite

In real-world application scenarios, the identification of groups poses a significant challenge due to possibly occurring outliers and existing noise variables. Therefore, there is a need for a clustering method which is capable of revealing the group structure in data containing both outliers and noise variables without any pre-knowledge. In this paper, we propose a k-means-based algorithm incorporating a weighting function which leads to an automatic weight assignment for each observation. In order to cope with noise variables, a lasso-type penalty is used in an objective function adjusted by observation weights. We finally introduce a framework for selecting both the number of clusters and variables based on a modified gap statistic. The conducted experiments on simulated and real-world data demonstrate the advantage of the method to identify groups, outliers, and informative variables simultaneously.

show abstract

Guided projections for analysing the structure of high-dimensional data

Ortner¹,

Filzmoser²,

Zaharieva³

et al. 2017

Preprint

View full text Add to dashboard Cite

Clustering of imbalanced high-dimensional media data

Brodinova

Zaharieva

Filzmoser

et al. 2017

Adv Data Anal Classif

View full text Add to dashboard Cite

Media content in large repositories usually exhibits multiple groups of strongly varying sizes. Media of potential interest often form notably smaller groups. Such media groups differ so much from the remaining data that it may be worthy to look at them in more detail. In contrast, media with popular content appear in larger groups. Identifying groups of varying sizes is addressed by clustering of imbalanced data. Clustering highly imbalanced media groups is additionally challenged by the high dimensionality of the underlying features. In this paper, we present the imbalanced clustering (IClust) algorithm designed to reveal group structures in high-dimensional media data. IClust employs an existing clustering method in order to find an initial set of a large number of potentially highly pure clusters which are then successively merged. The main advantage of IClust is that the number of clusters does not have to be pre-specified and that no specific assumptions about the cluster or data characteristics need to be made. Experiments on real-world media data demonstrate that in comparElectronic supplementary material The online version of this article (https://doi.org/10.1007/s11634-017-0292-z) contains supplementary material, which is available to authorized users. B Šárka Brodinová

show abstract

Towards an Agile Framework for Business Intelligence Projects

Prouza

Brodinova

Tjoa

2020

View full text Add to dashboard Cite

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

hi@scite.ai

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Sarka Brodinova

Robust and sparse k-means clustering for high-dimensional data

Guided projections for analysing the structure of high-dimensional data

Clustering of imbalanced high-dimensional media data

Towards an Agile Framework for Business Intelligence Projects

Contact Info

Product

Resources

About