Pith. sign in

Paper Citation Record · LEDGER

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning

As of 18 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2608.13026.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.13026 v1

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:21:34.060079Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact0
  • verified fuzzy13
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 869da127-0603-4a4c-ab34-7203f8d9e544 · outbound

This paper cites 2026 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2026 , eprint=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.458290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.676420Z digest=sha256:b5f2e8dca2a360a723da35092a001915a2013cc6e8126f1bb3a931a4696d8e91

Observation a65ec208-8e12-4bb8-b13e-7542844a638f · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.442159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.694079Z digest=sha256:63c073d31a06193c9996c2d464033136c34448686a6c57cf8908956d818c02cf

Observation edf3e6e3-82cb-4d7f-8356-2fbacf0a710c · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.418731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.708722Z digest=sha256:6b63bb3938f499a97816b71a611ba0b95a2e642ae886910f7e7d7e0c07bcba53

Observation 3a46a949-f91e-4bc1-b018-29db7b10cf98 · outbound

This paper cites 2023 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2023 , eprint=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.717518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.717518Z digest=sha256:722994b4f60b9a2faa1d1c4afb7e17bd58a097f7e56a60fb18ecac9b864333d0

Observation 9d75a46c-61e7-41d8-af41-bbd2cc248713 · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.732528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.732528Z digest=sha256:fb179eac2366da8d931508387387e09790df73948db6bc0f3ebb5ae91a08c38a

Observation 8c46405b-e01b-4803-a7fb-eff70999bb0b · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.748196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.748196Z digest=sha256:507694b62000594932198a79b3d81cc49d859433d3e83481078923307857a094

Observation 8b708809-6225-4c7f-8ff2-b9e994a6bd60 · outbound

This paper cites 2024 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2024 , eprint=

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.318982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.758361Z digest=sha256:9a69377abdc4da005919a63116a6cf81f1b00e546f55f19279879cf965003cfa

Observation 682b361c-f657-4973-8262-36bfb41fcbbf · outbound

This paper cites 2026 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2026 , eprint=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.770823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.770823Z digest=sha256:6a95af96b1ca536428ea5503f8a5d0649a1a37d411c3e3f0b91a9e3763ec349f

Observation a5f7eb8f-9713-4f83-b856-48b733bc1ed5 · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.783172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.783172Z digest=sha256:8e918536d2be75e90879b2cc0cf1fdc28bfc474aeb0a46d5f243d69c24ca43a9

Observation 1d69bc32-d0ce-410f-bea8-2561a50cf2eb · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.796752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.796752Z digest=sha256:ae8299bbe0d447eec0d1f496d6b7132c736ec55774ba09cb57a9a1db123dd04f

Observation f7669d75-4a38-44d6-b00f-fa8d48395e8c · outbound

This paper cites 2025 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2025 , eprint=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.802821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.802821Z digest=sha256:27f57dcee6761ec18f10aa147c9cc8e93d67ec0564bb02c183e8e5009e93d05f

Observation afb8e98f-fbc7-4cd6-bd49-cc06cdef7786 · outbound

This paper cites 2023 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2023 , eprint=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.812967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.812967Z digest=sha256:713a44f3c2aaa2854f27ca55e02d15344c7697ab36741103acb37f458829f542

Observation ad7425d2-83b8-4f67-86a0-1c4309d05b77 · outbound

This paper cites 2026 , eprint=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 2026 , eprint=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.147811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.826957Z digest=sha256:072dae340a279a4151ab383d63de3bab449a6d43e89e2793373432f53dbd51b6

Observation 434ab39b-75c8-4f47-98cf-9971e41f8b79 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.843856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.843856Z digest=sha256:fe4cc24164514c0a4ea41eb84bcf360876a4da6a894048955cbf6efd8e9459d8

Observation 899c45df-f990-45da-8988-4ce91fb96a9d · outbound

This paper cites ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.850650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.850650Z digest=sha256:5abc0d5347f65974793ad07b99b08678db33247af1d5e6f8b66722116fa1f623

Observation ddecd92c-4332-4e6e-96a4-82b697484ac0 · outbound

This paper cites SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.858488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.858488Z digest=sha256:30fc13586d2395167e9b6fcbd45feae45ab23777d4eb736722d7860fc4ef3d8d

Observation 7ab6a76f-501a-484c-9f1f-5cd8834cc9ca · outbound

This paper cites arXiv preprint arXiv:2511.09515 , year=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning arXiv preprint arXiv:2511.09515 , year=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.867217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.867217Z digest=sha256:0621dfac3bf9c93d25144d7853349dfdc4d9e6bd920682bb377c03977d5a4aea

Observation ab3ebb63-bbb8-4840-8d38-a33ff3daeb17 · outbound

This paper cites arXiv preprint arXiv:2510.00406 , year=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning arXiv preprint arXiv:2510.00406 , year=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.877681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.877681Z digest=sha256:0a805b2fdb1a29fa10aca9a6e57280ac7ffc067b19fe1b5384c73a9377d3c9f5

Observation ff355e59-c5f4-4dee-a610-8c75c9d1ce71 · outbound

This paper cites arXiv preprint arXiv:2511.00091 , year=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning arXiv preprint arXiv:2511.00091 , year=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.889487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.889487Z digest=sha256:0d15520948cff1a93ce0887b4e28d4e5a162d7babf7ab450d2ae49db57e55a34

Observation 463ffd35-cc7a-4d8a-9694-0c890eb33c4d · outbound

This paper cites RobustVLA: On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning RobustVLA: On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.899884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.899884Z digest=sha256:fde3812cff0335aa98919a56b8369c7ffebe855325c81abc2aa7391f6354bf2d

Observation f9f1cae7-c5cd-4a06-ab10-9b978c160cc7 · outbound

This paper cites Interactive Post-Training for Vision-Language-Action Models.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Interactive Post-Training for Vision-Language-Action Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.909439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.909439Z digest=sha256:3708d9e90dcd2dab2d11e36c5c9e596aecc13f411b5f21b920f9cb7fa92a68a2

Observation 7c113e5e-74d7-4871-9a54-1e942ab25433 · outbound

This paper cites Advances in neural information processing systems , volume=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Advances in neural information processing systems , volume=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.927727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.927727Z digest=sha256:e1fcb4648a2430ef7aef89c04fc2c9498ad17b2ffb4603f669756b61009ea713

Observation a3b986ed-abf6-4da0-bb13-9bb623b6c8ef · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.079596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.947351Z digest=sha256:bc8b617f1fd55af6c42eb41f58da603b7289bf94194193fdc142833314f73753

Observation d60282f9-1343-48c8-9ff0-4e97a6bd045a · outbound

This paper cites Align-RUDDER: Learning From Few Demonstrations by Reward Redistribution.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Align-RUDDER: Learning From Few Demonstrations by Reward Redistribution

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.953684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.953684Z digest=sha256:052a0193288b3c0eafa178e12b21167d62ea845f29b0417ed71e94146af8eb5c

Observation cad20952-1a16-4d52-b11c-c675fbc8d1ff · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:35.053161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:33.964464Z digest=sha256:d1317794493bd8b94b551580497c60a0e4dbc3164916b20de556e8d32a758134

Observation e5a35949-7b26-4bf3-b165-64b8ebb71333 · outbound

This paper cites Reinforcement Learning with Segment Feedback.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Reinforcement Learning with Segment Feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.975739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.975739Z digest=sha256:cccf4edf4bed92053b3a21c40fc3d21fafbc254c1af0207a0427bc4932450994

Observation 79542b24-c7f5-4ec2-b3f4-de899a5c6fe8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:33.989588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:33.989588Z digest=sha256:ecbbdf802f2844ad796f0691d73813c5c0a46b1cfa6544213cca16fbe5f82c85

Observation 0395fe38-777a-41f1-8c98-d990af966289 · outbound

This paper cites SARM: Stage-Aware Reward Modeling for Long Horizon Robot Manipulation.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning SARM: Stage-Aware Reward Modeling for Long Horizon Robot Manipulation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:34.003593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:21:34.003593Z digest=sha256:594cf560343e56be419f4ffabaa2314623476b189640f56e719c3b7bb62f353a

Observation 189c229c-13e8-4f0b-a5f7-f05ecb87fa2e · outbound

This paper cites International Conference on Machine Learning , pages=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning International Conference on Machine Learning , pages=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.971063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.015049Z digest=sha256:9483b1b28555f911726e88532f63433345e6e0f7bfd02fde208c9bd543bdd9c1

Observation 6b6b3e3d-3f48-404e-9d3f-d6cb41d27881 · outbound

This paper cites International Conference on Machine Learning , pages=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning International Conference on Machine Learning , pages=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.930931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.023324Z digest=sha256:d19f582e5fff30b2653b17f90641a1859ae9a5648dd174cde957b222d2958332

Observation 05e21ac1-ce4c-4a78-a8e7-e3e48570d794 · outbound

This paper cites Conference on Robot learning , pages=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Conference on Robot learning , pages=

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.891557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.029391Z digest=sha256:c1932d627ddfbb7e96a7e7e2f8894f11ff3ee79fe96c759f68fe6a16f5e49d9f

Observation 86dd38a3-2008-47c8-ac77-b67ebdc99271 · outbound

This paper cites Conference on Robot Learning , pages=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Conference on Robot Learning , pages=

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.855099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.040918Z digest=sha256:27af2ff1ce3aace4e5b241439971e5f085392ba483520ee67c974315339ff45a

Observation 1b351643-2240-4a64-a884-054ec0cd7cd7 · outbound

This paper cites Conference on Robot Learning , pages=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning Conference on Robot Learning , pages=

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.816203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.048526Z digest=sha256:0b13673f31e0f2f1719dc8b17dafcc890e9cb1754eca1060db23db99e15ee4a5

Observation b4367387-7bd3-45b6-988d-1b16a760ed11 · outbound

This paper cites 8th Annual Conference on Robot Learning , year=.

Temporal GRPO: Beyond Trajectory-Level Credit in Vision-Language-Action Reinforcement Learning 8th Annual Conference on Robot Learning , year=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:21:34.780841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:21:34.060079Z digest=sha256:8626decf08f176beb0e59021cf9c4cf279119898b81721d9e818811192ac4a11

Pith citing papers

No inbound Pith citation observations are available.