REVIEW 2 cited by
LLM Attributor: Interactive Visual Attribution for LLM Generation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
While large language models (LLMs) have shown remarkable capability to generate convincing text across diverse domains, concerns around its potential risks have highlighted the importance of understanding the rationale behind text generation. We present LLM Attributor, a Python library that provides interactive visualizations for training data attribution of an LLM's text generation. Our library offers a new way to quickly attribute an LLM's text generation to training data points to inspect model behaviors, enhance its trustworthiness, and compare model-generated text with user-provided text. We describe the visual and interactive design of our tool and highlight usage scenarios for LLaMA2 models fine-tuned with two different datasets: online articles about recent disasters and finance-related question-answer pairs. Thanks to LLM Attributor's broad support for computational notebooks, users can easily integrate it into their workflow to interactively visualize attributions of their models. For easier access and extensibility, we open-source LLM Attributor at https://github.com/poloclub/ LLM-Attribution. The video demo is available at https://youtu.be/mIG2MDQKQxM.
Forward citations
Cited by 2 Pith papers
-
NeMo-Inspector: A Visualization Tool for LLM Generation Analysis
NeMo-Inspector combines interactive inference, side-by-side and multi-generation comparison, custom Python statistics, and LaTeX/Markdown rendering to help developers clean and improve LLM-generated datasets.
-
Document Attribution: Examining Citation Relationships using Large Language Models
A zero-shot textual entailment prompt ('Does the REFERENCE entail the CLAIM?') edges out prior baselines on AttributionBench (73.8 ID, 83.43 OOD with flan-ul2), while per-layer attention probes reduce mostly to trivia...
Discussion (0). Continue with ORCID to comment.