Pith. sign in

Paper Citation Record · LEDGER

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control

As of 10 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2508.10022.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10022 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:23:48.288459Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 76451bdc-a42e-4e6e-8aff-97eba88e3b6e · outbound

This paper cites A Gentle Introduction to Conformal Prediction and Distribution-Free Uncertainty Quantification.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control A Gentle Introduction to Conformal Prediction and Distribution-Free Uncertainty Quantification

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:46.709655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:46.709655Z digest=sha256:8f3aefd3df8657a83fde2798ca25939d5315acb7143e0af3369f6e89857ac683

Observation 2ad85750-b998-4bf0-a679-a343480d6dc3 · outbound

This paper cites PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:46.781501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:46.781501Z digest=sha256:51a3db6deef7b7de74d320b667abe4c08b5d50883825e69c66f9a0374259ef83

Observation 451c4316-055f-48e9-81dd-e292512e1eb6 · outbound

This paper cites LL a VA steering: Visual instruction tuning with 500x fewer parameters through modality linear representation-steering.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control LL a VA steering: Visual instruction tuning with 500x fewer parameters through modality linear representation-steering

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:50.030135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T23:23:46.863849Z digest=sha256:59dcde375b3d9ef440ab719f087f20bef10a801c02f1943a67383043eac31354

Observation dd131b02-8b9c-4054-921c-36c772edf56a · outbound

This paper cites CoT-Kinetics: A Theoretical Modeling Assessing LRM Reasoning Process.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control CoT-Kinetics: A Theoretical Modeling Assessing LRM Reasoning Process

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:46.948042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:46.948042Z digest=sha256:279680b817f3713d08906c7437f9b408d20c8aa4d443de91969111cc24526dbc

Observation eafdf066-90b5-4ed9-836f-ff173eb304e4 · outbound

This paper cites Fedbip: Heterogeneous one-shot federated learning with personalized latent diffusion models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Fedbip: Heterogeneous one-shot federated learning with personalized latent diffusion models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.846480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.024290Z digest=sha256:70fdef57cad57f72c320784ab74c843fc0080f4ef00f90fbc989f0ed8a1f2008

Observation 6d7ed467-7a4c-4f14-a85c-be2a120b10e3 · outbound

This paper cites Does machine unlearning truly remove model knowledge? a framework for auditing unlearning in llms.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Does machine unlearning truly remove model knowledge? a framework for auditing unlearning in llms

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.100379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.100379Z digest=sha256:c23ab274a37739b61dfb4c1e144a90106dd6f115ae809440d2b63bbd8f6b6d24

Observation edfb10b0-820c-44f9-b4c9-7dadf750cd4a · outbound

This paper cites Conformal alignment: Knowing when to trust foundation models with guarantees.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Conformal alignment: Knowing when to trust foundation models with guarantees

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.621129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.158934Z digest=sha256:28873913294ad1fbde1d315532e8f25e9cf43391aec35b3c2f0f534414bd5ca9

Observation a58d7397-4577-4cc5-9ca4-b5aab31dee69 · outbound

This paper cites Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.223084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.223084Z digest=sha256:86ee2faffd1eb9f9dc6075aab2a12c6a147321271c3dbf9483ed45e77781190e

Observation c6233eed-af46-4a6d-892b-114be85be1f9 · outbound

This paper cites Backdoor Cleaning without External Guidance in MLLM Fine-tuning.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Backdoor Cleaning without External Guidance in MLLM Fine-tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.312420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.312420Z digest=sha256:7bf625cf84f4b93b3a5001bef8e7a639b2aca9813fe4794ab532a79d6fa9c68a

Observation 3cb58dc8-88a9-4442-b7da-168729b9f851 · outbound

This paper cites Sample then identify: A general framework for risk control and assessment in multimodal large language models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Sample then identify: A general framework for risk control and assessment in multimodal large language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.430114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.445094Z digest=sha256:fcb25b25c0f5f2a428355f45345571d4365ceb70d3716b9dd31e694df6e111a0

Observation 043c2906-a24a-4771-827d-a6552001558b · outbound

This paper cites Conformalized multiple testing after data-dependent selection.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Conformalized multiple testing after data-dependent selection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.235604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.580189Z digest=sha256:a87acec0e3d8777b0c59415940aeee95138bc170b6cc3c4ccec319bb19306f0c

Observation 9a6e9177-439e-4cb4-a35b-84edbef7d3dc · outbound

This paper cites ASCD: Attention-Steerable Contrastive Decoding for Reducing Hallucination in MLLM.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control ASCD: Attention-Steerable Contrastive Decoding for Reducing Hallucination in MLLM

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.662074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.662074Z digest=sha256:1ef657d12e399b39ce19daefabd9e9169e559a972bebcf7832e4f03d4e94fe0f

Observation 5b6655dc-2b0e-4590-9275-709984260001 · outbound

This paper cites Conu: Conformal uncertainty in large language models with correctness coverage guarantees.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Conu: Conformal uncertainty in large language models with correctness coverage guarantees

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.052841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.772855Z digest=sha256:4740c6f71c62f507fe3e8507056ea6ec75f95d50bb4f948e3b485931e69f2118

Observation 48bea371-5e4a-40c3-a8e0-d6ccdac8bdbb · outbound

This paper cites COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.832064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.832064Z digest=sha256:02d8a7cf0cf260637d27cecc1dc4783a5ce27c0ee6baf1e598335088aa093019

Observation bd6def49-0437-46e9-bdda-7b713f81a5af · outbound

This paper cites Word-sequence entropy: Towards uncertainty estimation in free-form medical question answering applications and beyond.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Word-sequence entropy: Towards uncertainty estimation in free-form medical question answering applications and beyond

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:48.927016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.962579Z digest=sha256:f69d69bf074ad682a2ea48f48320a16ba85a0ec67113a5a05dfa3b58900fb250

Observation 17bb5d39-e617-4473-acf5-15e5e2599d61 · outbound

This paper cites SC on U : Selective conformal uncertainty in large language models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control SC on U : Selective conformal uncertainty in large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:48.717705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T23:23:48.072512Z digest=sha256:dcb3a81429bb5d5c4fb0e88366f447ca235e28534e66e0b02d8fbc1327168a02

Observation 67504119-5f72-4e32-95b2-fede8e3a3149 · outbound

This paper cites Benchmarking llms via uncertainty quantification.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Benchmarking llms via uncertainty quantification

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:48.603235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T23:23:48.170468Z digest=sha256:7c9458d9d2c15080da292a9b7cfef2904f958d9674724e061f29e5c23da9a835

Observation 019de6e1-98e3-42a4-a2f2-f2ba73c4ca8f · outbound

This paper cites SPOT! Revisiting Video-Language Models for Event Understanding.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control SPOT! Revisiting Video-Language Models for Event Understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:48.288459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:48.288459Z digest=sha256:f4e021d440e67143a9aff013afd17defa1828f48a1e88f3056a4e1982054c5ad

Pith citing papers

No inbound Pith citation observations are available.