Pith. sign in

REVIEW

ANEA: Automated (Named) Entity Annotation for German Domain-Specific Texts

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2112.06724 v1 pith:OBNKGC3A submitted 2021-12-13 cs.CL

classification cs.CL
keywords nameddomain-specificaneaentitytextsautomatedcategoriesentities
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Named entity recognition (NER) is an important task that aims to resolve universal categories of named entities, e.g., persons, locations, organizations, and times. Despite its common and viable use in many use cases, NER is barely applicable in domains where general categories are suboptimal, such as engineering or medicine. To facilitate NER of domain-specific types, we propose ANEA, an automated (named) entity annotator to assist human annotators in creating domain-specific NER corpora for German text collections when given a set of domain-specific texts. In our evaluation, we find that ANEA automatically identifies terms that best represent the texts' content, identifies groups of coherent terms, and extracts and assigns descriptive labels to these groups, i.e., annotates text datasets into the domain (named) entities.

Discussion (0). Sign in to comment.

Pith tools