Pith. sign in

Paper Citation Record · LEDGER

NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2405.01481.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.01481 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:50:25.927351Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bf7613ef-81f8-41ae-b7c2-09a9db61ff12 · inbound

AsyncFlow: An Asynchronous Streaming RL Framework for Efficient LLM Post-Training cites this paper.

AsyncFlow: An Asynchronous Streaming RL Framework for Efficient LLM Post-Training NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T20:50:25.927351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:50:25.927351Z digest=sha256:b947dc5d749b2ff252c7ab6a45856a619a199174f247ff8b87de88c4127308ea

Observation d6f2d8f8-f5b9-4855-9ecc-2958f85c6327 · inbound

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation cites this paper.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:17.701818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:17.701818Z digest=sha256:093f7bdf0f027e4640ab5c81910318ba3ca0d77c42d317f79420c2ca0402f8ce

Observation 2bf3ee51-b65c-4289-b049-862bb2e7007c · inbound

DistFlow: A Fully Distributed RL Framework for Scalable and Efficient LLM Post-Training cites this paper.

DistFlow: A Fully Distributed RL Framework for Scalable and Efficient LLM Post-Training NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T16:19:43.466905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:19:43.466905Z digest=sha256:014808c8376d9ec5cc88d5b332bb1248a39175a984577fadc4e815629446aa81

Observation ebe49702-74e0-46e6-9b87-78bf1d386461 · inbound

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs cites this paper.

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:30:55.184777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T05:29:46.136115Z digest=sha256:f55551346cb12b0df807d9db2b210c6bf1d69fbeddd0bc62337474d106dce50b

Observation 13db1838-6460-4c15-b620-8e6e87b94088 · inbound

Seer: Online Context Learning for Fast Synchronous LLM Reinforcement Learning cites this paper.

Seer: Online Context Learning for Fast Synchronous LLM Reinforcement Learning NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:40:14.517448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T20:38:30.169363Z digest=sha256:9519eb5ad75786e82ee3bc06ed5a141ecd4f132bc9c19d8610214a39c1d947f3

Observation 3de10776-5083-4f30-8968-e8431037f067 · inbound

HetRL: Efficient Reinforcement Learning for LLMs in Heterogeneous Environments cites this paper.

HetRL: Efficient Reinforcement Learning for LLMs in Heterogeneous Environments NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T22:23:36.671858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-16T22:21:26.271796Z digest=sha256:8d9ed3a03bcac5768c6b34b9515cda9e409ef0cc08a0e49369f5288a648f6ee3

Observation 9c070b65-f6a5-4037-ac1c-133cb04b100f · inbound

AIS: Adaptive Importance Sampling for Quantized RL cites this paper.

AIS: Adaptive Importance Sampling for Quantized RL NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:14:52.702508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T03:13:14.384567Z digest=sha256:5a1db924deb5bb05c00033e8627a0f06c86c86d74117daea6b6800e8156212fb

Observation f275f9a5-465a-4df6-90d9-fb8f31b817b3 · inbound

AstraFlow: Dataflow-Oriented Reinforcement Learning for Agentic LLMs cites this paper.

AstraFlow: Dataflow-Oriented Reinforcement Learning for Agentic LLMs NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:13:43.733269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T20:10:32.300423Z digest=sha256:e3eeec4b1d3130e5bff468aea936a7401f28574e3fd5264581501a6982b5b6ad

Observation ce777bb2-6fb7-4c66-90ea-1ba83f659937 · inbound

PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR cites this paper.

PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:29:25.309065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T02:24:48.872065Z digest=sha256:e24b904f38bd973e12d20d4037592c87f83d3c48ac3ed6248bf544524822c9fc

Observation 11cdd8c7-22d3-413d-acc1-725e40234434 · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T01:37:30.462636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:7d607ebf8fdd5d21e197acce8b17415ee07150a877c960ae8d556ed9a781c0eb

Observation 773edd7a-2c84-450c-9864-ff06209b4a50 · inbound

Bidirectional Resource Scheduling for Disaggregated and Asynchronous RL Post-Training cites this paper.

Bidirectional Resource Scheduling for Disaggregated and Asynchronous RL Post-Training NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-13T04:42:35.589143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T04:42:35.589143Z digest=sha256:e16ee8d0e3f7ee161146a3b4330ecef81c9d662502f1fb92fdc7cd20f87cccac