Pith. sign in

Paper Citation Record · LEDGER

Learning When to Trust via Selective Context Preference Optimization

As of 10 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2608.06377.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06377 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:10:44.120366Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation de5fff26-e8b1-47a2-bcc3-f095138614c9 · outbound

This paper cites SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model.

Learning When to Trust via Selective Context Preference Optimization SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.127049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.127049Z digest=sha256:8631a1ebe33f70d3a554f82ac3e9a8fd1941962d135b10b1d3013be1e3cd180b

Observation 6e938e8c-db6b-4dc0-ac0c-5762692f62be · outbound

This paper cites Gemma 3 Technical Report.

Learning When to Trust via Selective Context Preference Optimization Gemma 3 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.503102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.503102Z digest=sha256:c95bd62dc0b01c2adb442d40e7cd4d86cb6280463974a496e32063b541e5d16a

Observation ebb65bb2-77a6-48d0-b336-fc58a81573cc · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.574558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.574558Z digest=sha256:7a6a4274e2f26d48bfbba7adec61a92f56ece7f0e939261e04c6e92f1aae9158

Observation 4d82cb14-4d33-4725-bd02-00f969e27112 · outbound

This paper cites User-Assistant Bias in LLMs.

Learning When to Trust via Selective Context Preference Optimization User-Assistant Bias in LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.689310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.689310Z digest=sha256:2e91ead40769a2b404efeea14a5d7a36797d985f59154484a38e8e944dcf2afb

Observation 89561fec-e23d-4a63-a641-293eab41b16d · outbound

This paper cites Ignore Previous Prompt: Attack Techniques For Language Models.

Learning When to Trust via Selective Context Preference Optimization Ignore Previous Prompt: Attack Techniques For Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.752073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.752073Z digest=sha256:a708437c893e28c4f3bb4860eeff0c66c9f41153795ae1281bfa3507943c3936

Observation f6d5ff6d-3413-4f9d-87c4-f67d0d645c0f · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:10:44.536920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:10:43.838984Z digest=sha256:5b34456197b97a59213482301e33f1548281c00e0e564e2f0dd57ddbc3d99b2d

Observation 6757a1d3-49ee-44ec-a07c-bcd9837e1cbe · outbound

This paper cites Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA.

Learning When to Trust via Selective Context Preference Optimization Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:44.016598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:44.016598Z digest=sha256:0b662cdd5206a9c2ab621295cbe78c4e5dd763668332ab06bba5cc78f36aaf1a

Observation 11c4f194-8b8c-4dc9-8b93-a3cd0d378853 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Learning When to Trust via Selective Context Preference Optimization Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.266892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.266892Z digest=sha256:f00b05a03e58de3abeab5e372b3f199663a1a877d6a5afa8d9037c5e0be91321

Observation c5bd237e-3b6e-4762-9592-bbeaf23a7bb4 · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:10:44.758253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:10:43.423326Z digest=sha256:0c8c8e970b46f61765fbaf41f7c1f7684834e606df16defeb394ed9a8de46df9

Observation a2ffcca9-4f3c-4ee8-9a5e-e6bd2b924f5c · outbound

This paper cites Purified OPSD: On-Policy Self-Distillation Without Losing How to Think.

Learning When to Trust via Selective Context Preference Optimization Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.931332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.931332Z digest=sha256:c376dbf5eb2cf440038545f95f74881bc541a49707083eb0f45a3fd2dcb2c2a8

Observation dce39efe-da83-462a-b9b7-a328143600b6 · outbound

This paper cites Phi-4-reasoning Technical Report.

Learning When to Trust via Selective Context Preference Optimization Phi-4-reasoning Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.082824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.082824Z digest=sha256:02233d2933dccc11c8b0058798c907c11e99a2e210a2af361a48d1228db2b81a

Observation ad3daa9e-4fa4-4c73-a3e0-ae48bf3b87fe · outbound

This paper cites Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models.

Learning When to Trust via Selective Context Preference Optimization Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:44.120366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:44.120366Z digest=sha256:483a19cf3a7bfc6265bd3d1da61d25c66578016d0f010722e0b531f48992b4f3

Pith citing papers

No inbound Pith citation observations are available.