REVIEW 4 cited by
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Recent advancements in Large Language Models (LLMs) have demonstrated exceptional capabilities in complex tasks like machine translation, commonsense reasoning, and language understanding. One of the primary reasons for the adaptability of LLMs in such diverse tasks is their in-context learning (ICL) capability, which allows them to perform well on new tasks by simply using a few task samples in the prompt. Despite their effectiveness in enhancing the performance of LLMs on diverse language and tabular tasks, these methods have not been thoroughly explored for their potential to generate post hoc explanations. In this work, we carry out one of the first explorations to analyze the effectiveness of LLMs in explaining other complex predictive models using ICL. To this end, we propose a novel framework, In-Context Explainers, comprising of three novel approaches that exploit the ICL capabilities of LLMs to explain the predictions made by other predictive models. We conduct extensive analysis with these approaches on real-world tabular and text datasets and demonstrate that LLMs are capable of explaining other predictive models similar to state-of-the-art post hoc explainers, opening up promising avenues for future research into LLM-based post hoc explanations of complex predictive models.
Forward citations
Cited by 4 Pith papers
-
Accurate Ensembles, Fragile Narratives: Multi-Scale Stacking and a Fidelity Audit of LLM-Generated Explanations for Credit Risk
A constrained-prompt LLM credit-risk explainer inverted the sign of three of four supplied drivers in an audited case, while the stacking ensemble's AUC gain over random forest was real but operationally small.
-
CONCORD: Concept-Informed Diffusion for Dataset Distillation
A training-free concept-informed diffusion method, using LLM-retrieved and CLIP-filtered visual descriptions, improves dataset distillation accuracy on ImageNet subsets and ImageNet-1K.
-
How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives
This paper introduces an automated evaluation framework with extraction-based faithfulness metrics, perplexity for assumptions, and embedding-based human similarity, and shows it can reveal LLM sign self-correction on...
-
Explingo: Explaining AI Predictions using Large Language Models
Explingo uses GPT-4o to convert SHAP explanations into natural-language narratives and to grade them on four quality metrics, with a few exemplars improving style but nudging correctness down.
Discussion (0). Continue with ORCID to comment.