Pith. sign in

Paper Citation Record · LEDGER

Conformal Prediction with Large Language Models for Multi-Choice Question Answering

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2305.18404.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.18404 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:12:03.428043Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:09:45.306878Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4649f54b-f515-424b-b4c2-dbc4914df016 · inbound

Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment cites this paper.

Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 116

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T22:30:44.950008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T22:30:44.520703Z digest=sha256:704da84ff0bc5b01696f7c7f5116033dbde3b503a32876f9b56dbf1365f1cd2f

Observation 43a0f19a-c374-45db-aca2-4fe1715b32cd · inbound

Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning cites this paper.

Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T07:47:42.631390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-23T07:45:50.292586Z digest=sha256:57488c64b3e6f8d5ebf203c7e9e17d783327eff1671fbb8abf168822d9f83007

Observation cf1c964d-8884-498a-b115-0959bd96e9ca · inbound

Retrieval-Augmented Generation with Graphs (GraphRAG) cites this paper.

Retrieval-Augmented Generation with Graphs (GraphRAG) Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 215

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T04:33:39.487193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T04:33:39.076517Z digest=sha256:c5f755df978279c47cdb818150fe1adebf37e62cb18d41bce9f7c4412a03c809

Observation 63c3ab0c-f659-4e20-a109-2fb8b7a2c771 · inbound

PARC: A Quantitative Framework Uncovering the Symmetries within Vision Language Models cites this paper.

PARC: A Quantitative Framework Uncovering the Symmetries within Vision Language Models Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:03.428043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:03.428043Z digest=sha256:900b7d4b80fd59e679e6ab4693974f552044eaa4c4477e469a22b6404c2fc4ef

Observation 211f5f8a-56c2-4c55-b010-b5871c8374a1 · inbound

Large Language Models for Statistical Inference: Context Augmentation with Applications to the Two-Sample Problem and Regression cites this paper.

Large Language Models for Statistical Inference: Context Augmentation with Applications to the Two-Sample Problem and Regression Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:10.479328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:45:10.479328Z digest=sha256:0213d54ae0fac24f7b69361ab6048810150f81f65114f125114f39bc052056eb

Observation f83ed041-8f4c-4561-929b-9ba206e30c06 · inbound

Shapley Uncertainty in Natural Language Generation cites this paper.

Shapley Uncertainty in Natural Language Generation Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T12:55:57.252263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:55:57.252263Z digest=sha256:d3482158cb3404304cf58e84a18cb5ba135a9c605956b7e284b68299b6967e67

Observation 1abc9b59-ae0a-46da-b15a-39dfb676ee6a · inbound

Membership Inference Attacks with False Discovery Rate Control cites this paper.

Membership Inference Attacks with False Discovery Rate Control Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T22:28:06.069255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:28:06.069255Z digest=sha256:36d56b288d9cc8ec295e84fd69475ff32a1660854d762f03c8e6df1c2083ffcb

Observation c4491083-c071-4e44-9c47-ae5e3937875d · inbound

Improving Backward Conformal Prediction via Non-Conformity Score Transformation cites this paper.

Improving Backward Conformal Prediction via Non-Conformity Score Transformation Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T14:20:13.819661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T14:15:26.433846Z digest=sha256:3d7d273ea0263d6a450badba1e5cd2c5b9b432d7b9033461d39151c61c3607ea

Observation e78a6183-3a33-4edc-b38c-33ab9a0c6284 · inbound

Improving Backward Conformal Prediction via Non-Conformity Score Transformation cites this paper.

Improving Backward Conformal Prediction via Non-Conformity Score Transformation Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-03T05:39:30.662335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:39:30.662335Z digest=sha256:8371202d5438c18c61a822241c42caf855da9ec50bb89f5e9e6b359f3840d53c

Observation 9b54431a-a45e-4dbe-b8b8-a585509abdd0 · inbound

Adaptive Stopping for Multi-Turn LLM Reasoning cites this paper.

Adaptive Stopping for Multi-Turn LLM Reasoning Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:23:21.322783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T22:20:57.717657Z digest=sha256:3b7d5c70e67e031e7911b31f6b891df30a69d304d746cc8d0460b05f417be393

Observation 037cd183-8b02-4c20-a9c8-90731c423551 · inbound

Calibrated Confidence Estimation for Tabular Question Answering cites this paper.

Calibrated Confidence Estimation for Tabular Question Answering Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:21:01.436953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T15:00:46.646913Z digest=sha256:4fda75a1a383b7a8d64352800397ee3dcecc625b629842e61f582b9ca40651d7

Observation 9adbf77e-d6fb-43ec-af98-6cb68334d09e · inbound

Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations cites this paper.

Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:44:37.617068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T10:41:21.406201Z digest=sha256:efdb75b8f9f611b0175e67efb7e1cc5447c357ceb76ab200024ff027ee595253

Observation 89f242d3-6f55-4133-94a3-2eb3ebe251e9 · inbound

Continual Calibration: Coverage Can Collapse Before Accuracy in Lifelong LLM Fine-Tuning cites this paper.

Continual Calibration: Coverage Can Collapse Before Accuracy in Lifelong LLM Fine-Tuning Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:41:15.819017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T04:36:42.863594Z digest=sha256:a0d3f99b515f8092e910491ac96fe8b73e3e805885bafc8bf86ac453bcc6747c

Observation 714a6423-68f4-4307-b2df-b51b3f907250 · inbound

VLM Judges Can Rank but Cannot Score: Task-Dependent Uncertainty in Multimodal Evaluation cites this paper.

VLM Judges Can Rank but Cannot Score: Task-Dependent Uncertainty in Multimodal Evaluation Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:31:17.451920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-07T16:46:01.749101Z digest=sha256:97378d8404e0c75859f7ca83f07579fa875fb4ff14f8088d640aff4d59b95813

Observation 33a32adf-0266-4972-83df-a6a9a942245f · inbound

Geometry-Calibrated Conformal Abstention for Language Models cites this paper.

Geometry-Calibrated Conformal Abstention for Language Models Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:31:28.637270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T05:56:03.031192Z digest=sha256:2cc9242dd9affdfb4e7a2688e75ae06a5a7da55a5a27677d7dc48dacb286a548

Observation a009741f-e3f6-4cc0-91d0-357ff45200d9 · inbound

Fair Conformal Classification via Learning Representation-Based Groups cites this paper.

Fair Conformal Classification via Learning Representation-Based Groups Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:02:27.585370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T06:59:58.086633Z digest=sha256:15388df68abbf865b7b7029f1ebf96ced66a3748d4dfc96d8807d899c6cbe1e3

Observation fd12ee8c-6fa5-4ecd-8c13-31a8d2ec74db · inbound

Uncertainty Quantification for LLM-based Code Generation cites this paper.

Uncertainty Quantification for LLM-based Code Generation Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:02:13.084108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T04:01:31.679818Z digest=sha256:b180a47850c3449a405961cf6a08d65a253c6c62c0b9f34e3cd2aca734b61250

Observation c36ef70d-33f0-4110-8c97-1c69cb57d495 · inbound

Reading Calibrated Uncertainty from Language Model Trajectories cites this paper.

Reading Calibrated Uncertainty from Language Model Trajectories Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:45:23.964775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-25T05:44:23.197975Z digest=sha256:5477c73f30fd3eeabba3d43b3b591c68dc6aac1b61054c7e8da761730c1ff3b2

Observation 7abb617f-dcb3-432b-8fc8-40ce0708c670 · inbound

MiRD: Reliable Set-Valued Prediction for Open-Ended Question Answering via Miscoverage Risk Decomposition cites this paper.

MiRD: Reliable Set-Valued Prediction for Open-Ended Question Answering via Miscoverage Risk Decomposition Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T21:33:59.216548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T21:29:34.557764Z digest=sha256:2d52e5672f71cf4562ad63467b558b06e34b76b7ee39ca6f13bd8dd50ab78660

Observation dc50ae31-9d7b-45d5-87a2-755afdd42dc7 · inbound

Entropy Distribution as a Fingerprint for Hallucinations in Generative Models cites this paper.

Entropy Distribution as a Fingerprint for Hallucinations in Generative Models Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:13:26.684025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T12:10:43.626846Z digest=sha256:c31c3133aa5c63d55926007c13a3623b8598b40ca28a0c0a7bc7c673e9700a42

Observation f7e808bb-64df-4d5c-9961-e43204468874 · inbound

Conformal Certification of Reasoning Trace Prefixes cites this paper.

Conformal Certification of Reasoning Trace Prefixes Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:03:12.950380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T06:59:42.656386Z digest=sha256:52beb5f2c8f019582b8d23803564d78bc9adb67c4dbd77f021744e521ddc7fe4

Observation 6475d54f-e6cd-47c8-8eb0-113e256526cf · inbound

Capability Self-Assessment: Teaching LLMs to Know Their Limits cites this paper.

Capability Self-Assessment: Teaching LLMs to Know Their Limits Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:32:44.677621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T22:25:50.579196Z digest=sha256:d3059873897eb1b81ebcfc93528281123b2f359258a48dea58ce8463abf0397f

Observation 9e00818c-01ea-4440-aff0-eecb5df4489c · inbound

Does Compression Preserve Uncertainty? A Unified Benchmark for Quantized and Sparse LLMs via Conformal Prediction cites this paper.

Does Compression Preserve Uncertainty? A Unified Benchmark for Quantized and Sparse LLMs via Conformal Prediction Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:46:18.766333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T15:06:31.471481Z digest=sha256:0b85342b56c35e3f7b915604c253d97fbcd21eac3aa4bb5f76c9d7ef960a749a

Observation 8fc808fc-5950-487e-8d10-39b7274f9ada · inbound

Strategic Decision Support for AI Agents cites this paper.

Strategic Decision Support for AI Agents Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:37:56.422635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T09:57:46.960346Z digest=sha256:aad30ecbd6d627cbbaf2422273690eefac7a75ac5750d9657e9ed0a16ef5431e

Observation 625ba9a7-fb2c-428a-8721-ea17ba229f68 · inbound

The Origins of Stochasticity: Comprehensive Investigations on Uncertainty Quantification for Large Language Models cites this paper.

The Origins of Stochasticity: Comprehensive Investigations on Uncertainty Quantification for Large Language Models Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:09:45.308352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T09:05:08.641544Z digest=sha256:dee86878febde137efb9abb32a28a2516feb03e5fdc6fe9beb4e4cf513f5802f

Observation 524b44ef-f2d2-460c-bff4-3720902e2a7a · inbound

Budgeted Act-or-Defer Multi-Agent LLM Deliberation with Local Reliability Bounds cites this paper.

Budgeted Act-or-Defer Multi-Agent LLM Deliberation with Local Reliability Bounds Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:04:21.958238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T06:56:25.830078Z digest=sha256:37b87991721d0c71b13ed19f630bc794589a1383fa39f2dff4f45980aeac71aa

Observation 5f44c1aa-fbd5-47c6-af45-8604c37df337 · inbound

Cloud-Native Evaluation-as-a-Service: A Microservices Architecture for Scalable AI Monitoring with Conformal Guarantees cites this paper.

Cloud-Native Evaluation-as-a-Service: A Microservices Architecture for Scalable AI Monitoring with Conformal Guarantees Conformal Prediction with Large Language Models for Multi-Choice Question Answering

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-02T08:50:38.484767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:50:38.484767Z digest=sha256:673db488706e20b19a28ca43a3650b4211ee7db0ddfb463eac60bca0978e2d76