CoNLL#: Fine-grained Error Analysis and a Corrected Test Set for CoNLL-03 English
Modern named entity recognition systems have steadily improved performance in the age of larger and more powerful neural models. However, over the past several years, the state-of-the-art has seemingly hit another plateau on the benchmark CoNLL-03 English dataset. In this paper, we perform a deep di…