Pith. sign in

REVIEW 3 cited by

Generative and Pseudo-Relevant Feedback for Sparse, Dense and Learned Sparse Retrieval

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.07477 v1 pith:B42TDWW6 submitted 2023-05-12 cs.IR

classification cs.IR
keywords retrievalfeedbackquerysparsefirst-passbenefitsdenseexperiments
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Pseudo-relevance feedback (PRF) is a classical approach to address lexical mismatch by enriching the query using first-pass retrieval. Moreover, recent work on generative-relevance feedback (GRF) shows that query expansion models using text generated from large language models can improve sparse retrieval without depending on first-pass retrieval effectiveness. This work extends GRF to dense and learned sparse retrieval paradigms with experiments over six standard document ranking benchmarks. We find that GRF improves over comparable PRF techniques by around 10% on both precision and recall-oriented measures. Nonetheless, query analysis shows that GRF and PRF have contrasting benefits, with GRF providing external context not present in first-pass retrieval, whereas PRF grounds the query to the information contained within the target corpus. Thus, we propose combining generative and pseudo-relevance feedback ranking signals to achieve the benefits of both feedback classes, which significantly increases recall over PRF methods on 95% of experiments.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Upcycling Candidate Tokens of Large Language Models for Query Expansion

    cs.IR 2025-09 conditional novelty 6.0 of 10

    Using unselected top-k candidate tokens from a single LLM decoding pass as extra query terms improves retrieval over standard keyword expansion while using far fewer tokens than document-level methods.

  2. Intended Target Identification for Anomia Patients with Gradient-based Selective Augmentation

    cs.CL 2025-06 conditional novelty 6.0 of 10

    GradSelect uses gradient signals to selectively perturb and expand circumlocution text, improving retrieval of intended target items for anomia patients.

  3. A New Query Expansion Approach via Agent-Mediated Dialogic Inquiry

    cs.IR 2025-02 conditional novelty 4.0 of 10

    AMD uses three LLM agents (Socratic questioning, dialogic answering, reflective feedback) to generate and refine pseudo-answers for query expansion, reporting gains over prior methods on BEIR and TREC benchmarks.

Pith tools