Pith. sign in

Paper Citation Record · LEDGER

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset

As of 16 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 1 inbound Pith citation observation for arXiv:2411.08243.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.08243 v3

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:52:34.686086Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T12:58:09.227648Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 478f6d08-7e46-4bfa-a518-916494590a2e · outbound

This paper cites Leveraging Large Language Models in Conversational Recommender Systems.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Leveraging Large Language Models in Conversational Recommender Systems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.637138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.637138Z digest=sha256:3428d27519c82790889a2b73bfa66d3786060a50e0986c7600ba52577adc0200

Observation 7f6dae83-1fd2-49f2-ad7d-acbcc3d9751f · outbound

This paper cites BERTopic: Neural topic modeling with a class-based TF-IDF procedure.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset BERTopic: Neural topic modeling with a class-based TF-IDF procedure

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.646442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.646442Z digest=sha256:5ae6f32d8624c149296c4378f4ddc5a11ce58b301dd48abeee9497a5289a2643

Observation 9b0fee32-03aa-4a0f-9155-e1961ec42e11 · outbound

This paper cites The Curious Case of Neural Text Degeneration.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset The Curious Case of Neural Text Degeneration

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.651073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.651073Z digest=sha256:f74e22953a682f7e671a8935ec01604303d5cef73a1df798cfa707e64e99326b

Observation 2d055cf9-5346-4e77-be64-37b7d71acfff · outbound

This paper cites The Alignment Problem from a Deep Learning Perspective.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset The Alignment Problem from a Deep Learning Perspective

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.659958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.659958Z digest=sha256:bafed18afccf5f0fc2674b6659211fef0351692115ab0029164618c59390e1b7

Observation 841dca4e-2f1c-47a2-a249-54ac9aa5b15e · outbound

This paper cites GPT-4 Technical Report.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset GPT-4 Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.664347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.664347Z digest=sha256:2801b8fe00387b0043a0a67bf63f6b229f6a62693cb620fd520b5bff87a0a69f

Observation 4f01ec54-941c-4e56-bd85-34519288d3cb · outbound

This paper cites Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Factually Consistent Summarization via Reinforcement Learning with Textual Entailment Feedback

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.668288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.668288Z digest=sha256:afea7baa3493dded76c7d0c4eab29828332f62a98e987507da524052a9bd0d2b

Observation 08d542fd-f972-452b-b354-6fab2bc2b4d1 · outbound

This paper cites Character-LLM: A Trainable Agent for Role-Playing.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Character-LLM: A Trainable Agent for Role-Playing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.672381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.672381Z digest=sha256:005601b341f19f791fab6de342b42a2cfe30b7d27b055750e984765acc2b1a8c

Observation 67b01a54-9a65-4e89-bea7-243e979e3c72 · outbound

This paper cites Fundamental Limitations of Alignment in Large Language Models.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Fundamental Limitations of Alignment in Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.676905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.676905Z digest=sha256:5f52186bcce23eeef72b8d971ae5645ea0994d8bc32c4ca025f82ea2f8beb29d

Observation 7d3b536a-e6c6-4b8f-a88e-f2be84c5ca21 · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.681096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.681096Z digest=sha256:b90d2baeacf9f69ff67969758ae323d971e1ed5a9be5f44d4746da1e0a1b5b8e

Observation c7deb38d-0573-4119-9cc1-16579962f757 · outbound

This paper cites In total, the dataset contains 22K toxic prompts.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset In total, the dataset contains 22K toxic prompts

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:52:34.850548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T21:52:34.686086Z digest=sha256:254dc287a9a7de349dd1b334861a6648a396d3a08392e44b0b491c939965c488

Observation 02e1a373-9aea-4624-abf1-a42063183e9d · outbound

This paper cites Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.655368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.655368Z digest=sha256:277ef680bbbcafe63cc6307842d61d0d5ccc840c658c851655a08c6bf21a082a

Observation de241576-573b-43f5-962c-f0f635f22ce7 · outbound

This paper cites Safe RLHF: Safe Reinforcement Learning from Human Feedback.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset Safe RLHF: Safe Reinforcement Learning from Human Feedback

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.632566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.632566Z digest=sha256:eefc9f71ed4ce6f51901f944889e656264d742839d8ec22a07e88d276a6839cf

Observation 801ac7b6-4366-48b5-88ff-447686eae7da · outbound

This paper cites CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing.

Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T21:52:34.642407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:52:34.642407Z digest=sha256:31c6bdc2fc7b6eeadbd4a9bdb9038ccabaea06eb83d61a4e5be5a43709e176d2

Pith citing papers

Observation c26fdc0a-8967-4c72-9735-541cb32951c2 · inbound

Discriminatory Compliance: How LLMs Answer Queries from Protected Groups cites this paper.

Discriminatory Compliance: How LLMs Answer Queries from Protected Groups Beyond the Safety Bundle: Auditing the Helpful and Harmless Dataset

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-06-26T12:59:29.240410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-26T12:58:09.227648Z digest=sha256:55cdc66fe2f8d250c54ef8b957dadfc2070b48bb3a4a7736bcd696db077f58ee