Pith. sign in

Paper Citation Record · LEDGER

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M

As of 15 August 2026, this Paper Citation Record lists 10 of 10 outbound references and 0 inbound Pith citation observations for arXiv:2509.09055.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.09055 v1

Coverage vector

measured 10 of 10 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T19:48:23.200455Z

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

10 of 10 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e05c4997-8370-41ed-bce8-eefa7f93d9e5 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:22.384196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:22.384196Z digest=sha256:6454c9d821c20bbbdf877dd82d9963c8b2523c12dc2daf9041de51d453327b4a

Observation 713eddfd-fe4b-429b-a5e1-ddcdc75da528 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M OPT: Open Pre-trained Transformer Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:22.457324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:22.457324Z digest=sha256:8d10fed186ca6dda3196b87e394ce6fa2b010a11ca64c4962a07d3a3599cdde1

Observation b1d66f49-ccf2-4d08-be9e-eaccbf368904 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:22.569427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:22.569427Z digest=sha256:1ade6d2a3bd529705f2f2baefe6a4f0e976e1ec6cc4c9ba4ec9dc27cb74d5358

Observation 2a73cc95-c9ce-4d0c-a535-f8894b8213b0 · outbound

This paper cites How Far Can Camels Go? Exploring the State of Instruction Tuning on Open Resources.

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M How Far Can Camels Go? Exploring the State of Instruction Tuning on Open Resources

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:22.658957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:22.658957Z digest=sha256:4dabf182e221f914cac577fdcef8b7e80cc4e075fd0fae077db2210c914a48b6

Observation 7740676d-a93a-4607-ba59-16987e2cb33d · outbound

This paper cites Realistic Evaluation of Toxicity in Large Language Models.

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M Realistic Evaluation of Toxicity in Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:22.761053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:22.761053Z digest=sha256:a8ae12f80e59e2ed978f94008c28ac0fb921d4e25d2e773dc0bb6151258957a3

Observation 40651299-5700-4120-a563-dfeaa452eef7 · outbound

This paper cites Crafting Tomorrow's Evaluations: Assessment Design Strategies in the Era of Generative AI.

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M Crafting Tomorrow's Evaluations: Assessment Design Strategies in the Era of Generative AI

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:22.847237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:22.847237Z digest=sha256:46b411bf3ae38f48b996c637d4ee01a4a25c1756a2c3f215a96f438578285282

Observation 91451675-098c-4626-8d54-ab661566828f · outbound

This paper cites Dynabench: Rethinking Benchmarking in NLP.

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M Dynabench: Rethinking Benchmarking in NLP

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:22.922729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:22.922729Z digest=sha256:edcf2a2e03ccd8c5564e00e191b18161256758c3a7c4cdc0b4b754dfed2668da

Observation ced9a9ae-fd1e-4e8f-b6bd-6c8b837f7175 · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M Fine-Tuning Language Models from Human Preferences

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:23.005157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:23.005157Z digest=sha256:a762e53df3e703f38199493c99e07f70394822434f16d63ed856760b7b632248

Observation 4adc5a6b-4463-4caf-831a-8f4ac4bf5faa · outbound

This paper cites Red Teaming Language Models with Language Models.

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M Red Teaming Language Models with Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:23.101317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:23.101317Z digest=sha256:a4b9af38d8926637e1c9b8a1ceb42d5a531d45795a13ddccca503fb895c42160

Observation 98b8aa4c-6825-4719-a95a-f90b96b1c0e0 · outbound

This paper cites Insights into Alignment: Evaluating DPO and its Variants Across Multiple Tasks.

Improving LLM Safety and Helpfulness using SFT and DPO: A Study on OPT-350M Insights into Alignment: Evaluating DPO and its Variants Across Multiple Tasks

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T19:48:23.200455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:48:23.200455Z digest=sha256:c344a562881bd850f76986edadc0f6f639669eb634798e533b1d19cbc1c140bc

Pith citing papers

No inbound Pith citation observations are available.