REVIEW 5 cited by
Unified Structure Generation for Universal Information Extraction
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Information extraction suffers from its varying targets, heterogeneous structures, and demand-specific schemas. In this paper, we propose a unified text-to-structure generation framework, namely UIE, which can universally model different IE tasks, adaptively generate targeted structures, and collaboratively learn general IE abilities from different knowledge sources. Specifically, UIE uniformly encodes different extraction structures via a structured extraction language, adaptively generates target extractions via a schema-based prompt mechanism - structural schema instructor, and captures the common IE abilities via a large-scale pre-trained text-to-structure model. Experiments show that UIE achieved the state-of-the-art performance on 4 IE tasks, 13 datasets, and on all supervised, low-resource, and few-shot settings for a wide range of entity, relation, event and sentiment extraction tasks and their unification. These results verified the effectiveness, universality, and transferability of UIE.
Forward citations
Cited by 5 Pith papers
-
Small Language Model Makes an Effective Long Text Extractor
A smaller span-based NER model with a compressed plus-shaped attention mechanism extracts long entities from very long texts with less memory than prior span-based methods.
-
DIRECT: Direct Decoding for Efficient and Aligned Sequence Labeling with Large Language Models
DPO after SFT plus constrained template-filling decoding improves LLM sequence-labeling accuracy and cuts inference time by reusing KV cache for non-label tokens.
-
Visual Information Extraction from Documents via Classification-Guided Large Vision-Language Models
Classification-guided dynamic prompts improve zero-shot visual information extraction from 16 certificate types, reaching 86.43 F1 without LVLM fine-tuning on a private bidding dataset.
-
The Joint Entity-Relation Extraction Model Based on Span and Interactive Fusion Representation for Chinese Medical Texts with Complex Semantics
A joint entity-relation extraction model with cross-attention feature fusion is proposed and evaluated on a new Chinese drug-drug interaction dataset and on CoNLL04.
-
Large Language Model for Extracting Complex Contract Information in Industrial Scenes
Clustering contracts, LLM-based labeling, augmentation, and LoRA fine-tuning improve Chinese industrial contract field extraction over traditional TF-IDF/TextRank/SNOWNLP/KeyBERT baselines.
Discussion (0). Continue with ORCID to comment.