Named-entity recognition (NER) (also known as entity identification and entity extraction) is a subtask of information extraction that seeks to locate and classify atomic elements in text into predefined categories such as the names of persons, organizations, locations, expressions of times, quantities, monetary values, percentages, etc.Most research on NER systems has been structured as taking an unannotated block of text, such as this one:
:Jim bought 300 shares of Acme Corp. in 2006.
And producing an annotated block of text, such as this one:
:Jimbought300shares ofAcme Corp. in 2006.
In this example, the annotations have been done using so-called ENAMEX tags that were developed for the Message Understanding Conference in the 1990s.State-of-the-art NER systems for English produce near-human performance. For example, the best system entering MUC-7 scored 93.39% of F1 score|F-measure while human annotators scored 97.60% and 96.95%.