REVIEW 2 cited by
The Language Interpretability Tool: Extensible, Interactive Visualizations and Analysis for NLP Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We present the Language Interpretability Tool (LIT), an open-source platform for visualization and understanding of NLP models. We focus on core questions about model behavior: Why did my model make this prediction? When does it perform poorly? What happens under a controlled change in the input? LIT integrates local explanations, aggregate analysis, and counterfactual generation into a streamlined, browser-based interface to enable rapid exploration and error analysis. We include case studies for a diverse set of workflows, including exploring counterfactuals for sentiment analysis, measuring gender bias in coreference systems, and exploring local behavior in text generation. LIT supports a wide range of models--including classification, seq2seq, and structured prediction--and is highly extensible through a declarative, framework-agnostic API. LIT is under active development, with code and full documentation available at https://github.com/pair-code/lit.
Forward citations
Cited by 2 Pith papers
-
A Unified Approach to Interpreting Knowledge Distillation for Large Language Models via Interactions
KD sparsifies complex interactions in student LLMs; a Complex Interaction Penalty that enforces this improves multiple KD methods on in- and out-of-domain benchmarks.
-
Exploring the Landscape of Fairness Interventions in Software Engineering
A survey of fairness interventions in software engineering that organizes prior work into a taxonomy and adds a small empirical analysis of open-source fairness repository maintenance.
Discussion (0). Continue with ORCID to comment.