Pith. sign in

Paper Citation Record · LEDGER

RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2402.17257.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.17257 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T14:52:27.233425Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-09T05:50:27.096272Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1a9940bf-a6dd-433e-b83f-9670b9018ae3 · inbound

Preference VLM: Leveraging VLMs for Scalable Preference-Based Reinforcement Learning cites this paper.

Preference VLM: Leveraging VLMs for Scalable Preference-Based Reinforcement Learning RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences

Reference 1952

Resolution
unresolved
no resolver link, observed 2026-08-09T14:52:27.233425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:52:27.233425Z digest=sha256:6534d47d61ae4e7d5c9f4e72a0ff3913eedc3d5e0eff293aa2889afb4d62c9b2

Observation d7ec4a76-b6b0-47f6-ae4f-ccec4cdd5be6 · inbound

CLARIFY: Contrastive Preference Reinforcement Learning for Untangling Ambiguous Queries cites this paper.

CLARIFY: Contrastive Preference Reinforcement Learning for Untangling Ambiguous Queries RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:57.502734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:12:57.502734Z digest=sha256:b50785a193f73455d3d7b5f5903cbb4457c88c98173f4a581ce1084d75016de7

Observation 472eb190-1a61-43b5-b310-65446938b326 · inbound

Efficient Preference Poisoning Attack on Offline RLHF cites this paper.

Efficient Preference Poisoning Attack on Offline RLHF RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:50:27.097641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T19:29:25.000361Z digest=sha256:94c08bfa4ac98aa1459b0cf8e68b92ceee0025f67ace0a866c953b95ff6a0a6b

Observation 29e85e52-b7e4-4951-b3ea-3992fe8e8c90 · inbound

Internal Pluralism and the Limits of Pairwise Comparisons cites this paper.

Internal Pluralism and the Limits of Pairwise Comparisons RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences

Reference 174

Resolution
unresolved
no resolver link, observed 2026-07-12T07:49:57.204875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T07:49:57.204875Z digest=sha256:38f2d33c826fee4b6af17ede936b347fa37c99c03312a1ffc5cf4e3dde280d5c

Observation c3e2db5d-f8da-4a84-a301-6e65b215a348 · inbound

Internal Pluralism and the Limits of Pairwise Comparisons cites this paper.

Internal Pluralism and the Limits of Pairwise Comparisons RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences

Reference 271

Resolution
unresolved
no resolver link, observed 2026-08-02T09:03:15.440243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T09:03:15.440243Z digest=sha256:764e7f6cddf6b72491dff99d46da0009ca8b0cde50041f26d4b87fda2fdaf1c9