Pith. sign in

Paper Citation Record · LEDGER

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control

As of 14 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2508.10022.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10022 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:23:48.288459Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 76451bdc-a42e-4e6e-8aff-97eba88e3b6e · outbound

This paper cites A Gentle Introduction to Conformal Prediction and Distribution-Free Uncertainty Quantification.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control A Gentle Introduction to Conformal Prediction and Distribution-Free Uncertainty Quantification

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:46.709655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:46.709655Z digest=sha256:5c84e607c989365e505a0de0110fc7c072b308be3656d06f2bc069b4ada9d1be

Observation 2ad85750-b998-4bf0-a679-a343480d6dc3 · outbound

This paper cites PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:46.781501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:46.781501Z digest=sha256:367c5dfe923a7583ae5890f679b1af6fdec4724fbeea269400ebc2e8a7610c25

Observation 451c4316-055f-48e9-81dd-e292512e1eb6 · outbound

This paper cites LL a VA steering: Visual instruction tuning with 500x fewer parameters through modality linear representation-steering.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control LL a VA steering: Visual instruction tuning with 500x fewer parameters through modality linear representation-steering

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:50.030135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T23:23:46.863849Z digest=sha256:38a159549ff810f04ecae3003891af8612bcb74e335a8e920973be1c9c0de212

Observation dd131b02-8b9c-4054-921c-36c772edf56a · outbound

This paper cites CoT-Kinetics: A Theoretical Modeling Assessing LRM Reasoning Process.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control CoT-Kinetics: A Theoretical Modeling Assessing LRM Reasoning Process

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:46.948042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:46.948042Z digest=sha256:688cf943d87be8ed48258e3ff07e03e7bb1752c45c3e95e46f0998b95d7867c0

Observation eafdf066-90b5-4ed9-836f-ff173eb304e4 · outbound

This paper cites Fedbip: Heterogeneous one-shot federated learning with personalized latent diffusion models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Fedbip: Heterogeneous one-shot federated learning with personalized latent diffusion models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.846480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.024290Z digest=sha256:c2b07dfd1c02d526f9d188a1167af7935177f845da1b25fe3a6b1e35eb02c475

Observation 6d7ed467-7a4c-4f14-a85c-be2a120b10e3 · outbound

This paper cites Does machine unlearning truly remove model knowledge? a framework for auditing unlearning in llms.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Does machine unlearning truly remove model knowledge? a framework for auditing unlearning in llms

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.100379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.100379Z digest=sha256:9bac9b9138eab7d05af220776c578ac2cd985ad767eb798c7b2c2605ce66142d

Observation edfb10b0-820c-44f9-b4c9-7dadf750cd4a · outbound

This paper cites Conformal alignment: Knowing when to trust foundation models with guarantees.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Conformal alignment: Knowing when to trust foundation models with guarantees

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.621129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.158934Z digest=sha256:d33254574727870ca5a521edd02d638abdb7e26353289315e1f3ca3260df41ad

Observation a58d7397-4577-4cc5-9ca4-b5aab31dee69 · outbound

This paper cites Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.223084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.223084Z digest=sha256:cb92463f179c732613143c884c50e587b2782755f8177859faca07df6c016590

Observation c6233eed-af46-4a6d-892b-114be85be1f9 · outbound

This paper cites Backdoor Cleaning without External Guidance in MLLM Fine-tuning.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Backdoor Cleaning without External Guidance in MLLM Fine-tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.312420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.312420Z digest=sha256:6d7dee34080eb02838534999861d7c7216138a6e130e28186d657e608bea2378

Observation 3cb58dc8-88a9-4442-b7da-168729b9f851 · outbound

This paper cites Sample then identify: A general framework for risk control and assessment in multimodal large language models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Sample then identify: A general framework for risk control and assessment in multimodal large language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.430114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.445094Z digest=sha256:ee3894661bce697fb7e7d60bc63ef86dfb32290c0e17f5acfd6786be83364fc1

Observation 043c2906-a24a-4771-827d-a6552001558b · outbound

This paper cites Conformalized multiple testing after data-dependent selection.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Conformalized multiple testing after data-dependent selection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.235604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.580189Z digest=sha256:5e313c80c8ccba0606eb3b75dcdccf08e6d47a5bb452666f47fb2e4be42e217a

Observation 9a6e9177-439e-4cb4-a35b-84edbef7d3dc · outbound

This paper cites ASCD: Attention-Steerable Contrastive Decoding for Reducing Hallucination in MLLM.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control ASCD: Attention-Steerable Contrastive Decoding for Reducing Hallucination in MLLM

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.662074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.662074Z digest=sha256:e481cca7a2db08885c4da31944c8b7a3f1b27a32a001adac47adf78c935cafec

Observation 5b6655dc-2b0e-4590-9275-709984260001 · outbound

This paper cites Conu: Conformal uncertainty in large language models with correctness coverage guarantees.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Conu: Conformal uncertainty in large language models with correctness coverage guarantees

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.052841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.772855Z digest=sha256:434974986e45bc8bacdd3ebae8a1d68839d925cd58631cb18a9d9325f4029190

Observation 48bea371-5e4a-40c3-a8e0-d6ccdac8bdbb · outbound

This paper cites COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.832064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.832064Z digest=sha256:39af3473281765d1f2fc09909905a5804a4ba315dd6abd727f98c700134938de

Observation bd6def49-0437-46e9-bdda-7b713f81a5af · outbound

This paper cites Word-sequence entropy: Towards uncertainty estimation in free-form medical question answering applications and beyond.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Word-sequence entropy: Towards uncertainty estimation in free-form medical question answering applications and beyond

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:48.927016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.962579Z digest=sha256:4463d078c286c746d0a7b9a4f77106ee0ac27cc39aab7a069bec9fc26e59a7b5

Observation 17bb5d39-e617-4473-acf5-15e5e2599d61 · outbound

This paper cites SC on U : Selective conformal uncertainty in large language models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control SC on U : Selective conformal uncertainty in large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:48.717705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T23:23:48.072512Z digest=sha256:5b8de287232c71ff473b23229c1f10942b348be552af12cc8469f5ea9415b7a3

Observation 67504119-5f72-4e32-95b2-fede8e3a3149 · outbound

This paper cites Benchmarking llms via uncertainty quantification.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Benchmarking llms via uncertainty quantification

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:48.603235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T23:23:48.170468Z digest=sha256:f9e2b6219c63579785ce76200153315c8e03d3d5ad56de93f74a7c62baa56819

Observation 019de6e1-98e3-42a4-a2f2-f2ba73c4ca8f · outbound

This paper cites SPOT! Revisiting Video-Language Models for Event Understanding.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control SPOT! Revisiting Video-Language Models for Event Understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:48.288459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:48.288459Z digest=sha256:5f58c778d03303ac4c9b8f9523a642d4fada1cabcf218914f483723487254490

Pith citing papers

No inbound Pith citation observations are available.