Pith. sign in

Paper Citation Record · LEDGER

Policy-Based Trajectory Clustering in Offline Reinforcement Learning

As of 12 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 2 inbound Pith citation observations for arXiv:2506.09202.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09202 v2

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:01:02.986353Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T10:27:47.896922Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:20:00.517702Z

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c47bc253-0561-4887-8529-9a3809064937 · outbound

This paper cites an unresolved cited work.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:03.300603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:01:02.962537Z digest=sha256:d4bcc33ac4e66462a5dd6a35832199b8c38440615ec71f8af16332a0d2b3cb9f

Observation e1c511b1-b97e-4a24-ac76-3583315dfd99 · outbound

This paper cites an unresolved cited work.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:03.286204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:01:02.966492Z digest=sha256:8348df52923c1d2f0a20c824a09bf4410859b7a60a7b9dcd2158d8af47041311

Observation 379f1434-9361-4d59-8b59-4ccbd6333370 · outbound

This paper cites an unresolved cited work.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:03.271406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:01:02.970261Z digest=sha256:42b6ae3d7f9623d09c0ef64eba6fba203e3990e0ded6653ab0de5a5bed908783

Observation 86f73a97-a64a-4d6b-9036-cd7c669c123f · outbound

This paper cites an unresolved cited work.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:03.256975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:01:02.974127Z digest=sha256:94902fcd7db4637665221e784db8dc7bb650cda4ed9a1dcb8a27ad107fb0aeee

Observation 9c94e70c-2d7b-46da-b108-9da3b240bee5 · outbound

This paper cites Takeball There are four different rule-based policies, i-th policy will pick the i-th ball first.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Takeball There are four different rule-based policies, i-th policy will pick the i-th ball first

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:03.243513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:01:02.977885Z digest=sha256:c7052fe3db3962fb37d9e9cbbae5e2d74e7c9320b9826f7f12df6604b64ca2c6

Observation 2cadc6d3-6a9d-4f48-83b9-e7a07950a58f · outbound

This paper cites an unresolved cited work.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:01:03.229272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:01:02.981668Z digest=sha256:dc9035e690ef6ea3a1ab311c3898540ba074e14204aab23c6400efb63be347c4

Observation adb002fe-3f00-4860-a89e-3e72b48120bc · outbound

This paper cites These policies are corresponding to: No preference, prefer to go up first and prefer to go right first.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning These policies are corresponding to: No preference, prefer to go up first and prefer to go right first

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:01:03.215458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:01:02.986353Z digest=sha256:0883db74bedfcf3b568e1cd275843166f2a5e0642ba185df70eef846ec28e0b8

Observation d36b757d-f311-4590-ad1b-b9c86dbe2772 · outbound

This paper cites Unsupervised Deep Embedding for Clustering Analysis.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Unsupervised Deep Embedding for Clustering Analysis

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:02.957217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:02.957217Z digest=sha256:e6d12c3d4ec392c4246d524f2b3b39c8fc1767de52f83f5614a5e8f97fd07845

Observation 7d8c3380-d34b-49f4-9707-2f6ca7722d5e · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Behavior Regularized Offline Reinforcement Learning

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:02.952757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:02.952757Z digest=sha256:d48a0a560cb80270ee70af7e7990f44ea7ca34e4604938890859337e9fe625ee

Observation 1b92c03a-6ad1-47e1-9b08-ad2ba2ea496f · outbound

This paper cites Contrastive Clustering.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Contrastive Clustering

Reference 2020

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:01:03.186412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:01:02.937922Z digest=sha256:b229b0e760ca6335c98d2676c174a7acb3b81e8f881c288ca6bf679bd90da894

Observation 2e6a45a0-9bee-4608-80df-30b917f176a3 · outbound

This paper cites Deep Reinforcement Learning for Autonomous Driving: A Survey.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Deep Reinforcement Learning for Autonomous Driving: A Survey

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:02.932706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:02.932706Z digest=sha256:0f0605211e5c321a3c3d01819a049d16ea42bd8e4d7a31c66f9b7b050fa6a2e8

Observation 80ea89f9-9ee0-454e-bc69-aece7f06d8c5 · outbound

This paper cites Dataset Clustering for Improved Offline Policy Learning.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning Dataset Clustering for Improved Offline Policy Learning

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:01:03.060042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:01:02.947723Z digest=sha256:8b86def099000e195c4b212c8d8c9534343f0afa0a020cfb4df64cb70cc062e4

Observation 7d86d253-c3c4-42db-ae9c-ef56f3c9b302 · outbound

This paper cites URL http://dx.doi.org/10.1109/TNNLS.2023.

Policy-Based Trajectory Clustering in Offline Reinforcement Learning URL http://dx.doi.org/10.1109/TNNLS.2023

Reference 2388

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T05:01:03.166577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T05:01:02.942412Z digest=sha256:2b15495a4e91bdba1208659eb18b2b3d6459ad077b381bd7cf29ba4deb1e7181

Pith citing papers

Observation 513d2eda-a545-4fcd-b8e7-42e07391a589 · inbound

Implicit Neural Representations of Individual Behavior cites this paper.

Implicit Neural Representations of Individual Behavior Policy-Based Trajectory Clustering in Offline Reinforcement Learning

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:17:48.435994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-27T10:27:47.896922Z digest=sha256:a39d7429625db40aad30d12cd8b1ce03057f4595669a9c3d7860c860e8265df9

Observation 7aa0886d-e7c1-4479-8a86-934233ff7c8d · inbound

ASALT: Adaptive State Alignment for Lateral Transfer in Multi-agent Reinforcement Learning cites this paper.

ASALT: Adaptive State Alignment for Lateral Transfer in Multi-agent Reinforcement Learning Policy-Based Trajectory Clustering in Offline Reinforcement Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:20:00.519286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T23:46:50.183064Z digest=sha256:1ced1e2da6228796cb12745ee533ba451fd0e78ff1f6ee6a92ba584853292abb