A fixed learned sentence inserted into a frozen LLM's prompt steers its uncertainty so that lowest-entropy answer selection avoids misleading passages, raising mean F1 from 0.5148 to 0.5339 across five QA benchmarks.
LLM-as-judge protocol.Because bothF 1 and EM are lexical, we additionally audit answer cor- rectness with an LLM judge, so that no conclusion rests on string overlap alone
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
LODESTAR: Trustworthy Entropy Is Navigated, Not Merely Measured -- Reinforced Polarizer Keeps a Frozen LLM from Being Confidently Misled by the Wrong Evidence
A fixed learned sentence inserted into a frozen LLM's prompt steers its uncertainty so that lowest-entropy answer selection avoids misleading passages, raising mean F1 from 0.5148 to 0.5339 across five QA benchmarks.