REVIEW 4 cited by
FiNER-ORD: Financial Named Entity Recognition Open Research Dataset
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Over the last two decades, the development of the CoNLL-2003 named entity recognition (NER) dataset has helped enhance the capabilities of deep learning and natural language processing (NLP). The finance domain, characterized by its unique semantic and lexical variations for the same entities, presents specific challenges to the NER task; thus, a domain-specific customized dataset is crucial for advancing research in this field. In our work, we develop the first high-quality English Financial NER Open Research Dataset (FiNER-ORD). We benchmark multiple pre-trained language models (PLMs) and large-language models (LLMs) on FiNER-ORD. We believe our proposed FiNER-ORD dataset will open future opportunities to use FiNER-ORD as a benchmark for financial domain-specific NER and NLP tasks. Our dataset, models, and code are publicly available on GitHub and Hugging Face under CC BY-NC 4.0 license.
Forward citations
Cited by 4 Pith papers
-
Named-Entity Recognition in the Crime Domain (CrimeNER): Case Study and Dataset
CrimeNER-db is a new, publicly released 1,568-document manually annotated corpus for crime-domain NER with a coarse/fine label hierarchy and zero-/few-shot benchmark results.
-
SusGen-GPT: A Data-Centric LLM for Financial NLP and Sustainability Report Generation
Small fine-tuned models on SusGen-30K are reported to nearly match GPT-4 on financial and ESG tasks, with a new TCFD-Bench benchmark, though the comparison is biased.
-
Financial Named Entity Recognition: How Far Can LLM Go?
Generic LLMs score lower than fine-tuned BERT and RoBERTa on financial NER, but few-shot prompting narrows the gap and reveals five recurring error types.
-
Open FinLLM Leaderboard: Towards Financial AI Readiness
The paper presents an open, continuously updated FinLLM leaderboard that aggregates existing financial benchmarks and demos for comparing models.
Discussion (0). Continue with ORCID to comment.