Pith. sign in

REVIEW 2 cited by

Language in a Bottle: Language Model Guided Concept Bottlenecks for Interpretable Image Classification

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2211.11158 v2 pith:NKRXPW45 submitted 2022-11-21 cs.CV cs.CL

classification cs.CVcs.CL
keywords labobottlenecksconceptsblacklanguagemodelmodelsgpt-3
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Concept Bottleneck Models (CBM) are inherently interpretable models that factor model decisions into human-readable concepts. They allow people to easily understand why a model is failing, a critical feature for high-stakes applications. CBMs require manually specified concepts and often under-perform their black box counterparts, preventing their broad adoption. We address these shortcomings and are first to show how to construct high-performance CBMs without manual specification of similar accuracy to black box models. Our approach, Language Guided Bottlenecks (LaBo), leverages a language model, GPT-3, to define a large space of possible bottlenecks. Given a problem domain, LaBo uses GPT-3 to produce factual sentences about categories to form candidate concepts. LaBo efficiently searches possible bottlenecks through a novel submodular utility that promotes the selection of discriminative and diverse information. Ultimately, GPT-3's sentential concepts can be aligned to images using CLIP, to form a bottleneck layer. Experiments demonstrate that LaBo is a highly effective prior for concepts important to visual recognition. In the evaluation with 11 diverse datasets, LaBo bottlenecks excel at few-shot classification: they are 11.7% more accurate than black box linear probes at 1 shot and comparable with more data. Overall, LaBo demonstrates that inherently interpretable models can be widely applied at similar, or better, performance than black box approaches.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Does VLM Classification Benefit from LLM Description Semantics?

    cs.CV 2024-12 conditional novelty 6.0 of 10

    LLM-generated descriptions improve VLM classification only when selected to discriminate among ambiguous classes, not when simply ensembled.

  2. Enhancing Patient-Centric Communication: Leveraging LLMs to Simulate Patient Perspectives

    cs.AI 2025-01 conditional novelty 4.0 of 10

    GPT-4 personas aligned with real patient answers 54.97% on average versus 26.7% random, but only when primed with education, and the reported 88% accuracy overstates what was measured.

Pith tools