Hanna Wecker scite author profile

Hanna Wecker

2Publications

3Citation Statements Received

19Citation Statements Given

How they've been cited

How they cite others

Affiliations

Publications

Order By: Most citations

ClusterDataSplit: Exploring Challenging Clustering-Based Data Splits for Model Performance Evaluation

Wecker¹,

Friedrich²,

Adel³

2020

View full text Add to dashboard Cite

This paper adds to the ongoing discussion in the natural language processing community on how to choose a good development set. Motivated by the real-life necessity of applying machine learning models to different data distributions, we propose a clustering-based data splitting algorithm. It creates development (or test) sets which are lexically different from the training data while ensuring similar label distributions. Hence, we are able to create challenging cross-validation evaluation setups while abstracting away from performance differences resulting from label distribution shifts between training and test data. In addition, we present a Python-based tool for analyzing and visualizing data split characteristics and model performance. We illustrate the workings and results of our approach using a sentiment analysis and a patent classification task.

show abstract

Zufriedenheit mit der Gesundheitsversorgung: Gibt es strukturelle Unterschiede?

Miranda¹,

Prosi²,

Schmidt

et al. 2018

View full text Add to dashboard Cite

This study examines structural differences in the subjective quality of health care in Germany using a newspaper survey. We find that there are significant differences between urban and rural areas as well as between public and private insurance. In rural areas, the provision of general practitioners, specialists and hospitals are considered as worse than in cities. In particular, public insured individuals asses the provision of specialized doctors and hospitals as lower than private insured and criticize long waiting times for appointments and lacking coverage of health care costs by the statutory health insurance.

show abstract

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

customersupport@researchsolutions.com

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Hanna Wecker

ClusterDataSplit: Exploring Challenging Clustering-Based Data Splits for Model Performance Evaluation

Zufriedenheit mit der Gesundheitsversorgung: Gibt es strukturelle Unterschiede?

Contact Info

Product

Resources

About