Pith. sign in

REVIEW 2 cited by

FLERT: Document-Level Features for Named Entity Recognition

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2011.06993 v2 pith:3L4T6NUA submitted 2020-11-13 cs.CL

classification cs.CL
keywords document-levelfeaturescontextentityexperimentsmodelnamedrecognition
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Current state-of-the-art approaches for named entity recognition (NER) typically consider text at the sentence-level and thus do not model information that crosses sentence boundaries. However, the use of transformer-based models for NER offers natural options for capturing document-level features. In this paper, we perform a comparative evaluation of document-level features in the two standard NER architectures commonly considered in the literature, namely "fine-tuning" and "feature-based LSTM-CRF". We evaluate different hyperparameters for document-level features such as context window size and enforcing document-locality. We present experiments from which we derive recommendations for how to model document context and present new state-of-the-art scores on several CoNLL-03 benchmark datasets. Our approach is integrated into the Flair framework to facilitate reproduction of our experiments.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. NER4all or Context is All You Need: Using LLMs for low-effort, high-performance NER on historical texts. A humanities informed approach

    cs.CL 2025-02 conditional novelty 5.0 of 10

    With context-rich prompts and persona modeling, ChatGPT-4o outperformed off-the-shelf spaCy and flair on named entity recognition for a 1921 German travel guide.

  2. GerPS-Compare: Comparing NER methods for legal norm analysis

    cs.CL 2024-12 conditional novelty 4.0 of 10

    On a German legal norm corpus, a fine-tuned XLM-RoBERTa outperforms a rule-based system and a prompted LLM (LeoLM) in macro F1 for ten annotation classes.

Pith tools