Textual Data Augmentation for NER in Geosciences with LLMs
Elie Maze,
Reza Farahbakhsh,
Pierre-Emmanuel Barrallon
et al.
Abstract:Named Entity Recognition (NER) is a classic natural language processing task which aims to extract relevant and domain-specific information (e.g. meta data) from textual data. To that end, the key is to have enough labeled data for the entities we are interested to extract, and labeling sufficient domain-specific is challenging and costly especially in geoscience domains. One of the promising techniques to increase the volume of the labels is to rely on data augmentation techniques and there are several approa… Show more
Set email alert for when this publication receives citations?
scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.