Roel Bertens scite author profile

Roel Bertens

5Publications

16Citation Statements Received

122Citation Statements Given

How they've been cited

How they cite others

122

Affiliations

Utrecht University

Publications

Order By: Most citations

Efficiently Discovering Unexpected Pattern-Co-Occurrences

Bertens¹,

Vreeken

Siebes³

2017

View full text Add to dashboard Cite

Our world is filled with both beautiful and brainy people, but how often does a Nobel Prize winner also wins a beauty pageant? Let us assume that someone who is both very beautiful and very smart is more rare than what we would expect from the combination of the number of beautiful and brainy people. Of course there will still always be some individuals that defy this stereotype; these beautiful brainy people are exactly the class of anomaly we focus on in this paper. They do not posses intrinsically rare qualities, it is the unexpected combination of factors that makes them stand out.In this paper we define the above described class of anomaly and propose a method to quickly identify them in transaction data. Further, as we take a pattern set based approach, our method readily explains why a transaction is anomalous. The effectiveness of our method is thoroughly verified with a wide range of experiments on both real world and synthetic data.

show abstract

Characterising Seismic Data

Bertens¹,

Siebes

2014

View full text Add to dashboard Cite

When a seismologist analyses a new seismogram it is often useful to have access to a set of similar seismograms. For example if she tries to determine the event, if any, that caused the particular readings on her seismogram. So, the question is: when are two seismograms similar?To define such a notion of similarity, we first preprocess the seismogram by a wavelet decomposition, followed by a discretisation of the wavelet coefficients. Next we introduce a new type of patterns on the resulting set of aligned symbolic time series. These patterns, called block patterns, satisfy an Apriori property and can thus be found with a levelwise search. Next we use MDL to define when a set of such patterns is characteristic for the data. We introduce the MuLTi-Krimp algorithm to find such code sets.In experiments we show that these code sets are both good at distinguishing between dissimilar seismograms and good at recognising similar seismograms. Moreover, we show how such a code set can be used to generate a synthetic seismogram that shows what all seismograms in a cluster have in common.

show abstract

Keeping it Short and Simple

Bertens

Vreeken

Siebes

2016

View full text Add to dashboard Cite

We study how to obtain concise descriptions of discrete multivariate sequential data. In particular, how to do so in terms of rich multivariate sequential patterns that can capture potentially highly interesting (cor)relations between sequences. To this end we allow our pattern language to span over the domains (alphabets) of all sequences, allow patterns to overlap temporally, as well as allow for gaps in their occurrences.We formalise our goal by the Minimum Description Length principle, by which our objective is to discover the set of patterns that provides the most succinct description of the data. To discover highquality pattern sets directly from data, we introduce DITTO, a highly efficient algorithm that approximates the ideal result very well.Experiments show that DITTO correctly discovers the patterns planted in synthetic data. Moreover, it scales favourably with the length of the data, the number of attributes, the alphabet sizes. On real data, ranging from sensor networks to annotated text, DITTO discovers easily interpretable summaries that provide clear insight in both the univariate and multivariate structure.

show abstract

Discretisation Effects in Naive Bayesian Networks

Bertens

Gaag

Renooij

2012

View full text Add to dashboard Cite

Keeping it Short and Simple: Summarising Complex Event Sequences with Multivariate Patterns

Bertens¹,

Vreeken²,

Siebes³

2015

Preprint

View full text Add to dashboard Cite

show abstract

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

hi@scite.ai

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Roel Bertens

Efficiently Discovering Unexpected Pattern-Co-Occurrences

Characterising Seismic Data

Keeping it Short and Simple

Discretisation Effects in Naive Bayesian Networks

Keeping it Short and Simple: Summarising Complex Event Sequences with Multivariate Patterns

Contact Info

Product

Resources

About