Building and Evaluating Universal Named-Entity Recognition English corpus

This article presents the application of the Universal Named Entity framework\nto generate automatically annotated corpora. By using a workflow that extracts\nWikipedia data and meta-data and DBpedia information, we generated an English\ndataset which is described and evaluated. Furthermore, we conducted a set of\nexperiments to improve the annotations in terms of precision, recall, and\nF1-measure. The final dataset is available and the established workflow can be\napplied to any language with existing Wikipedia and DBpedia. As part of future\nresearch, we intend to continue improving the annotation process and extend it\nto other languages.\n

Paper

Similar papers

© 2026 NYSGPT2525 LLC