Pith. sign in

Paper Citation Record · LEDGER

Learning When to Trust via Selective Context Preference Optimization

As of 9 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2608.06377.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06377 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:10:44.120366Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation de5fff26-e8b1-47a2-bcc3-f095138614c9 · outbound

This paper cites SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model.

Learning When to Trust via Selective Context Preference Optimization SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.127049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.127049Z digest=sha256:65fe5436c96f62244125ac1604f19ad154e3c56a6d821899bd6a944fcb84abcd

Observation 6e938e8c-db6b-4dc0-ac0c-5762692f62be · outbound

This paper cites Gemma 3 Technical Report.

Learning When to Trust via Selective Context Preference Optimization Gemma 3 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.503102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.503102Z digest=sha256:a759672bc2f21a990ae8d70a02574baa07a5872b3fd6af0845b7850e7fb1ee79

Observation ebb65bb2-77a6-48d0-b336-fc58a81573cc · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.574558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.574558Z digest=sha256:e3dd8654b5c18cb9fb76d8084f2b5a0dd3f8b68ee41a7f260cfbe09940f96be3

Observation 4d82cb14-4d33-4725-bd02-00f969e27112 · outbound

This paper cites User-Assistant Bias in LLMs.

Learning When to Trust via Selective Context Preference Optimization User-Assistant Bias in LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.689310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.689310Z digest=sha256:755b68570243a273231438105e2f4dd61f30647dd7d5ecf6787365a2e3d62c5d

Observation 89561fec-e23d-4a63-a641-293eab41b16d · outbound

This paper cites Ignore Previous Prompt: Attack Techniques For Language Models.

Learning When to Trust via Selective Context Preference Optimization Ignore Previous Prompt: Attack Techniques For Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.752073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.752073Z digest=sha256:e29584e8d39b3987db91183858eab05474b2953c4e1df10eda688a250b3ebd55

Observation f6d5ff6d-3413-4f9d-87c4-f67d0d645c0f · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:10:44.536920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:10:43.838984Z digest=sha256:dfc89c938e045d8b478859dd168bf5bc36ffcbab512d651fb300c08fd761fe9a

Observation 6757a1d3-49ee-44ec-a07c-bcd9837e1cbe · outbound

This paper cites Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA.

Learning When to Trust via Selective Context Preference Optimization Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:44.016598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:44.016598Z digest=sha256:339d7f36b41695566b68696583364d7553007ffb8be3c9604ed1e75320e11b8f

Observation 11c4f194-8b8c-4dc9-8b93-a3cd0d378853 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Learning When to Trust via Selective Context Preference Optimization Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.266892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.266892Z digest=sha256:3136b0fb09d64cec807c6ee54bd404ba4e9f337ad0686a3b15421bb181ec7bc9

Observation c5bd237e-3b6e-4762-9592-bbeaf23a7bb4 · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:10:44.758253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T04:10:43.423326Z digest=sha256:0d794eef2880d18e8fb98f6335abcaa187c10470864581019b52bacab34cbe23

Observation a2ffcca9-4f3c-4ee8-9a5e-e6bd2b924f5c · outbound

This paper cites Purified OPSD: On-Policy Self-Distillation Without Losing How to Think.

Learning When to Trust via Selective Context Preference Optimization Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.931332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.931332Z digest=sha256:569e171ca83f52184b94949a34beb67caff9d48e62b8c95dc110cde675004bfb

Observation dce39efe-da83-462a-b9b7-a328143600b6 · outbound

This paper cites Phi-4-reasoning Technical Report.

Learning When to Trust via Selective Context Preference Optimization Phi-4-reasoning Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.082824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.082824Z digest=sha256:0281ef7115fe4ef8cd72a3f1a1700de45ba98a457fe00e99874cfde996c3a516

Observation ad3daa9e-4fa4-4c73-a3e0-ae48bf3b87fe · outbound

This paper cites Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models.

Learning When to Trust via Selective Context Preference Optimization Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:44.120366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:44.120366Z digest=sha256:e414df471ecc56d6dd0f791ff1be3163abf96e3b4245d29025eb7a18cf5e6d72

Pith citing papers

No inbound Pith citation observations are available.