REVIEW 2 cited by
exBERT: A Visual Analysis Tool to Explore Learned Representations in Transformers Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Large language models can produce powerful contextual representations that lead to improvements across many NLP tasks. Since these models are typically guided by a sequence of learned self attention mechanisms and may comprise undesired inductive biases, it is paramount to be able to explore what the attention has learned. While static analyses of these models lead to targeted insights, interactive tools are more dynamic and can help humans better gain an intuition for the model-internal reasoning process. We present exBERT, an interactive tool named after the popular BERT language model, that provides insights into the meaning of the contextual representations by matching a human-specified input to similar contexts in a large annotated dataset. By aggregating the annotations of the matching similar contexts, exBERT helps intuitively explain what each attention-head has learned.
Forward citations
Cited by 2 Pith papers
-
LXMERT: Learning Cross-Modality Encoder Representations from Transformers
LXMERT pretrains a three-encoder Transformer on image-sentence pairs with five tasks and achieves state-of-the-art VQA, GQA, and NLVR2 results after fine-tuning.
-
Lightweight Safety Classification Using Pruned Language Models
Intermediate-layer features of small LLMs plus a penalized logistic regression classifier achieve high F1 scores on content safety and prompt injection classification with very few labeled examples, per the paper's ex...
Discussion (0). Continue with ORCID to comment.