Pith. sign in

Paper Citation Record · LEDGER

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints

As of 15 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2508.08466.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.08466 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:33:15.002562Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1c8a340-b8c4-4479-85d6-65ecd3fdf3de · outbound

This paper cites online" 'onlinestring :=.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:13.661566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:13.661566Z digest=sha256:2e106b584d34223e2fb1c517d91c9eaef1d375bf60949a33d44c4e14204f870b

Observation a4d42081-0833-4a1c-b5e2-a4f356c616a9 · outbound

This paper cites write newline.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:13.717146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:13.717146Z digest=sha256:c9bfb31d6fb80b30f73ccda4c709bfc340c5efd79e56a221595c4c8cc4b37339

Observation fec1937c-8fd6-4e10-b526-f01cd2d18c30 · outbound

This paper cites A General Theoretical Paradigm to Understand Learning from Human Preferences.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints A General Theoretical Paradigm to Understand Learning from Human Preferences

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:13.809159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:13.809159Z digest=sha256:61bf5153b40994a947e680a7c8967f91911c01343b22cb1b538c2ef859811493

Observation 7a1e7560-c6e8-4f19-a690-fbd49563bdaf · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Constitutional AI: Harmlessness from AI Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:13.895668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:13.895668Z digest=sha256:3d07121e39c191ada2259ecb93ae1443029f74ab60060ab1bce8557a7e92a29f

Observation d3065ade-670d-4596-94af-13ed7bee8f6d · outbound

This paper cites What is the Role of Small Models in the LLM Era: A Survey.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints What is the Role of Small Models in the LLM Era: A Survey

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:13.933222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:13.933222Z digest=sha256:170b2ef13345f538460f14e7fec929c223cc97189f20b6ff49a8b946d933342d

Observation 9957bac6-a38d-431e-8935-3b7927d24df4 · outbound

This paper cites Anchored Preference Optimization and Contrastive Revisions: Addressing Underspecification in Alignment.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Anchored Preference Optimization and Contrastive Revisions: Addressing Underspecification in Alignment

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:33:15.409840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-05T21:33:14.007798Z digest=sha256:171e00b04357bad74ef24479fa7afdb0b2c7d127d9bc484a1f917e9630934df7

Observation a349f246-6048-441a-af15-2f7c2e04f7bb · outbound

This paper cites Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.093113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.093113Z digest=sha256:dc34c8cfff1f1b88a538f07cbe10867e914f3a1c44007e27aa3c31f4c77abe81

Observation a96aa0da-8eff-4ea9-b056-c1e09d4268bd · outbound

This paper cites Towards Efficient Exact Optimization of Language Model Alignment.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Towards Efficient Exact Optimization of Language Model Alignment

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.147588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.147588Z digest=sha256:43e3fb7e81993fface988253f47aea94898c336efbf59b961f00220cee9bcd0b

Observation 28e58887-3cd0-4674-a756-7ebe4a899221 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Gonzalez, Hao Zhang, and Ion Stoica

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.183442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.183442Z digest=sha256:34d328ebc51844ebcd6d04f6b79309adeaf9af481807a33811b13b781df75428

Observation fc3d78eb-8f11-40c2-bd73-bd94f0ed84bf · outbound

This paper cites Hashimoto.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Hashimoto

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.239825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.239825Z digest=sha256:3a0e38af1efd8c30a3b3c7bd0ccd6f91423d3bce0b692d11c6c21202a995df0e

Observation 92dc04f4-e126-4cce-8711-346323c799db · outbound

This paper cites Distributional Preference Alignment of LLMs via Optimal Transport.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Distributional Preference Alignment of LLMs via Optimal Transport

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.325390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.325390Z digest=sha256:6c0ed444387dd00f5b369b813015667b75c58deb75db0a43be8ee234ecefe77c

Observation 47d1bf11-542b-47d3-8d30-3bb7b62214e2 · outbound

This paper cites an unresolved cited work.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-05T21:33:15.644753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-05T21:33:14.391997Z digest=sha256:92812724e4b56a9bafb9663bac8c01e0055c6d0a594d549fead47832088dac64

Observation 0b776026-6ff1-4919-9074-7634ca861ecf · outbound

This paper cites Training language models to follow instructions with human feedback.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Training language models to follow instructions with human feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.507155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.507155Z digest=sha256:69eca8d7a3a5f5b2dbd9ab5fec5e10fee26f594eaa0d022c1f52cf75824aaee0

Observation 3cc6d8ee-4de4-4718-ae7a-cdf3005b8489 · outbound

This paper cites Disentangling Length from Quality in Direct Preference Optimization.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Disentangling Length from Quality in Direct Preference Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.576084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.576084Z digest=sha256:35271dede038f4fc4bebc41c50c66ebd2c859664b05c3c1af5e82f8d2bef87fd

Observation 2fc2dc83-5dfc-4a65-8c15-2dc86ed03009 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.639533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.639533Z digest=sha256:0ab43eec07b1af7d6469e1a1bfe888cf14c41ce5fbb1957e3c0be10f85e0148b

Observation a9a94c3e-2939-4782-921f-467cd972511c · outbound

This paper cites an unresolved cited work.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.706971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.706971Z digest=sha256:37444b9340efda88384cb3be92a0e507a80cefcb711f37414df1f6ec46f822ed

Observation e3f38a03-98a7-4fcc-8e31-0856c86913a3 · outbound

This paper cites Reinforcement Learning for LLM Post-Training: A Survey.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Reinforcement Learning for LLM Post-Training: A Survey

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.784455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.784455Z digest=sha256:54a5eaa6f03c345c6ff912b8b84c9b1a00bb418096e4c6860c82f6f0f2012205

Observation 5fd91298-c1f0-4ab7-ae4c-bcc8c72fe381 · outbound

This paper cites SLiC-HF: Sequence Likelihood Calibration with Human Feedback.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints SLiC-HF: Sequence Likelihood Calibration with Human Feedback

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.871864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.871864Z digest=sha256:1b96a425ca53a4cccf68b3603a9c85794dee89e9bfc34be149f18c7aa142f6a4

Observation 786d5b9f-dc2f-4472-8ab3-ac935a616e90 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.923297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.923297Z digest=sha256:8f78265f73fa9a312434750d1718062077c6e4efef4fd560ad5b286470c67fd3

Observation 96741aac-5abd-42cd-a100-32adcd6410b9 · outbound

This paper cites LIMA: Less Is More for Alignment.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints LIMA: Less Is More for Alignment

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:15.002562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:15.002562Z digest=sha256:54bec8ac4eaf787199073fb77f71e561ccac8d9e2e49a1c0fb17266b27051560

Pith citing papers

No inbound Pith citation observations are available.