Pith. sign in

REVIEW 2 cited by

Self-Retrieval: End-to-End Information Retrieval with One Large Language Model

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.00801 v2 pith:54X2HRV7 submitted 2024-02-23 cs.IR cs.AIcs.CL

classification cs.IRcs.AIcs.CL
keywords retrievalllmsself-retrievalsystemsinformationarchitectureend-to-endgeneration
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

The rise of large language models (LLMs) has significantly transformed both the construction and application of information retrieval (IR) systems. However, current interactions between IR systems and LLMs remain limited, with LLMs merely serving as part of components within IR systems, and IR systems being constructed independently of LLMs. This separated architecture restricts knowledge sharing and deep collaboration between them. In this paper, we introduce Self-Retrieval, a novel end-to-end LLM-driven information retrieval architecture. Self-Retrieval unifies all essential IR functions within a single LLM, leveraging the inherent capabilities of LLMs throughout the IR process. Specifically, Self-Retrieval internalizes the retrieval corpus through self-supervised learning, transforms the retrieval process into sequential passage generation, and performs relevance assessment for reranking. Experimental results demonstrate that Self-Retrieval not only outperforms existing retrieval approaches by a significant margin, but also substantially enhances the performance of LLM-driven downstream applications like retrieval-augmented generation.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. QUPID: Quantified Understanding for Enhanced Performance, Insights, and Decisions in Korean Search Engines

    cs.CL 2025-05 conditional novelty 6.0 of 10

    A fine-tuned ensemble of a generative small language model and an embedding model outperformed zero-shot LLMs on Korean search relevance labeling, with reported Cohen's kappa of 0.646 versus 0.387 and 60x lower latency.

  2. DataScout: Automatic Data Fact Retrieval for Statement Augmentation with an LLM-Based Agent

    cs.HC 2025-04 conditional novelty 6.0 of 10

    A new system called DataScout lets data storytellers retrieve supporting and opposing data facts from the World Bank database through an LLM-based agent and a mind-map interface.

Pith tools