Pith. sign in

Paper Citation Record · LEDGER

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards

As of 14 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2501.14513.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.14513 v2

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:10:48.724636Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T16:27:25.150807Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T08:50:59.092753Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved30
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2a6c13df-86bb-4ab5-844b-fa927d35105c · outbound

This paper cites Loquercio, E.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Loquercio, E

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:10:49.353788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.602656Z digest=sha256:d7f990bd561f6e67dbe833d38bfba39a65dafec57b026c80b322f0d8bc08902f

Observation f39976c4-480f-430e-901b-7313d4e1f02f · outbound

This paper cites Loquercio, E.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Loquercio, E

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:10:49.342872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.606745Z digest=sha256:94b8c7b7125fdaa98907485716e714a02c73e9d67072fc9ef1785ab003218ccc

Observation 45562607-49e2-4e8c-84af-fb7895d5b67d · outbound

This paper cites Kaufmann, A.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Kaufmann, A

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:10:49.333036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.609915Z digest=sha256:b84f84507a087854607f7c3757872142159b327327c2c4e9854af7eb1f439bd2

Observation 76dcdfca-35c2-4c1d-9b97-52e3174b8cc8 · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:10:49.321381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.613006Z digest=sha256:d15b3ec0255ca6cdd0ceed36025d4df566c0814242ed1edf3ef4769e4c6f6831

Observation 347fb54a-7ece-4caa-bd65-994b173fdad7 · outbound

This paper cites Back to Newton's Laws: Learning Vision-based Agile Flight via Differentiable Physics.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Back to Newton's Laws: Learning Vision-based Agile Flight via Differentiable Physics

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.616444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.616444Z digest=sha256:bac952d9139e52de2d102f788743caf0d634351b61cb5ee1841d5c06feb05a9e

Observation d4b199b2-798a-4c38-a549-84db8edd5f78 · outbound

This paper cites Wiedemann, V.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Wiedemann, V

Reference 6

Resolution
malformed identifier
no resolver link, observed 2026-08-10T15:10:48.620025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.620025Z digest=sha256:b6b2d350f2d4cffe62ae048dc13e335ff281a39fd0d3ba0abca7381402a8eee1

Observation cff916ca-fe64-46f4-8017-e934760916cf · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:10:49.311453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.623414Z digest=sha256:01ae18e5bee7a9ad65520ff3bfa7dd64218e0921c2b1815b1cd6333dc3c4b4ee

Observation 8f7c4c75-7273-4fe9-a73d-23ee9732b71f · outbound

This paper cites Learning Quadruped Locomotion Using Differentiable Simulation.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Learning Quadruped Locomotion Using Differentiable Simulation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.626404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.626404Z digest=sha256:a83036055092d1993ed743556ddd8dc00c327ce47f8b6ae922bb34276e26ebc2

Observation a3ed2d12-91a8-4c4a-b677-4296a8ba90f9 · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.629534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.629534Z digest=sha256:d2e55bccd844029d8c4c4f42259e21d07ce3152899dbd4a0427e34f280d21dc7

Observation 7add1830-6e13-4a1a-83ea-629e8f04fa1c · outbound

This paper cites Zhang, W.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Zhang, W

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:10:49.301955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.632272Z digest=sha256:020d93466214a143184db833e7f28a75e611d272efd3d11d51ba64c4cd42defc

Observation 400c41c8-6704-48e0-a6bd-249558b1be33 · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:10:49.293561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.635581Z digest=sha256:3fc96cdea0b62def34bf3937e3e8db169cb10aad7b4e2c6eca67b6161f76533d

Observation fd32f32e-a259-4cee-a19a-26e41e0e3e02 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Playing Atari with Deep Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.638379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.638379Z digest=sha256:0e450a16b67ef9f7306c6df5b4195d05a3f1c5f4eb8d4c37d916d608e27a2388

Observation 7541952c-6c51-45bb-bd5c-422ab98394e1 · outbound

This paper cites Continuous control with deep reinforcement learning.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Continuous control with deep reinforcement learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.641907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.641907Z digest=sha256:e34ab2a55b323c2e69c0f8786ac215b931bbc1333ba2d3be4f4097fb5aa9bc7c

Observation 17aeecff-0047-4d6d-8480-85f28b681b0a · outbound

This paper cites Fujimoto, H.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Fujimoto, H

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:10:49.284169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.645239Z digest=sha256:f4ccb9365eb3fa9886dcb1b2f21c75c20f544331177abd1344932f1cbcaa5f92

Observation 4a649408-30d0-4204-af3c-9a67ed14f5ab · outbound

This paper cites Haarnoja, A.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Haarnoja, A

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.648213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.648213Z digest=sha256:457861b49a966153f3e046f1d86cb37c130f3984f3a8a0fcc5397521d49add7a

Observation 7a4a7514-f73b-4c86-aac7-ed59c30b87bb · outbound

This paper cites Trust Region Policy Optimization.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Trust Region Policy Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.651760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.651760Z digest=sha256:daed110be3f04b5b628166990e54352ec4ec9c515669f920c66cbd9d92079753

Observation 19e6577e-580f-4d35-b307-cbe404a74168 · outbound

This paper cites Proximal Policy Optimization Algorithms.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Proximal Policy Optimization Algorithms

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.655628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.655628Z digest=sha256:b68a836ce1b0fa63c59bd2b198c2401ba550a6ba03247eaa3f449cc81dd00c22

Observation 3282080b-e644-4e22-ae24-a23756b66830 · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:10:49.268170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.659274Z digest=sha256:943973f25bad0b5be69d19038089aa27f3b127088b48efce74d903569cac05dd

Observation 33b0d98d-fbac-4889-8b3d-984f6a4777a7 · outbound

This paper cites Deisenroth and C.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Deisenroth and C

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.663178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.663178Z digest=sha256:125433fcac220763d995a71bad1e33a031d031cb0e8384c1be8bb730432d553f

Observation 89d80766-acd2-4fc4-ab74-39af80008f1c · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:10:49.251386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.666518Z digest=sha256:ab3b19dfba8ab3b5f24734062e4399b5effe446f3daea5f659d263d72fe83ead

Observation f128fc34-672c-413b-9455-c0091379df41 · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.669930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.669930Z digest=sha256:1a64f5ad9b116285d7826f67b507678c713309065a244894a794b3f39b60d53c

Observation f80059c9-6d7c-4d3c-90b7-61df54726a2c · outbound

This paper cites Watter, J.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Watter, J

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.673338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.673338Z digest=sha256:ec3fe6efc9ef4767b6d922e0bbc2860fee2c878093c7e770d4e72377e1048748

Observation 44269d92-9f4d-49a8-a6a7-8fa60d2121d6 · outbound

This paper cites Dream to Control: Learning Behaviors by Latent Imagination.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Dream to Control: Learning Behaviors by Latent Imagination

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.676127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.676127Z digest=sha256:a7eeda71b370bab6818ee43ec0f26247ef5c0c48cb7caf06ea62bdba4084cbc1

Observation 16ba7ee5-3adb-4f6a-b810-88682f641455 · outbound

This paper cites DiffTaichi: Differentiable Programming for Physical Simulation.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards DiffTaichi: Differentiable Programming for Physical Simulation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.679345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.679345Z digest=sha256:85364fee15d850ba40a508661fa8273c265286f99c649c0d36ee5eabaa618243

Observation b24d34d9-1fe3-423f-88c8-e060a774792c · outbound

This paper cites Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.682411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.682411Z digest=sha256:6fae8bf2a8f8020ee969f9018a9749a6e3ccf0ba7046c91089f187de04e6ac55

Observation b65633a5-3bf1-4189-95f9-b7995508fb8d · outbound

This paper cites Todorov, T.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Todorov, T

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.685299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.685299Z digest=sha256:a41215bb03e6a357b00639360fcfb2811416d825864216889293c4abbfdc8dbf

Observation d9036fca-56ba-4228-bf6a-1147d597a461 · outbound

This paper cites Heiden, D.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Heiden, D

Reference 27

Resolution
metadata mismatch
raw_fallback, observed 2026-08-10T15:10:48.855662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.688100Z digest=sha256:631d5109f05eaa5ab506a85cca0a46bc81c6e9ba4d2abc6ce091ef39bfd5b2a6

Observation fe272e94-86ba-45a1-8198-9a65ea4d0140 · outbound

This paper cites Dojo: A Differentiable Physics Engine for Robotics.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Dojo: A Differentiable Physics Engine for Robotics

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.690868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.690868Z digest=sha256:17ac2ed5ed22f48b74b5e5c3a63641f57cd2c9ab9f0d04066d1894ff603a96fb

Observation b5ac3004-5f7c-4700-8691-2ef5ae39098e · outbound

This paper cites VisFly: An Efficient and Versatile Simulator for Training Vision-based Flight.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards VisFly: An Efficient and Versatile Simulator for Training Vision-based Flight

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.694231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.694231Z digest=sha256:8bf9506c0360f8b567bf64ccc71e47a994b3d561c8b17f430ef61ea0669e7550

Observation c9b09ff2-12dc-452b-874c-933a5043cb92 · outbound

This paper cites Savva, A.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Savva, A

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:10:49.230830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.697111Z digest=sha256:589414c9e454d505897fd83881bc5052527d82dbb6c34514a3bab74563d81c3a

Observation bd1987eb-84d1-454c-915c-ba001b222dad · outbound

This paper cites Schoenholz and E.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Schoenholz and E

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:10:49.220966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.699836Z digest=sha256:f04a952a534f2b1eade20a495338ceb520f3907fa624052beeb6a3269fa3a3ff

Observation 55011a73-a570-4ff0-a2b4-089610a5fbee · outbound

This paper cites Paszke, S.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Paszke, S

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.702421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.702421Z digest=sha256:df99caaaf71fa0674eed1391a37244fb784bbd8d1d2078d592ccb466b551dd53

Observation 7239ef10-55ae-4a66-8fbe-5d8e30adfe93 · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:10:49.204429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.705121Z digest=sha256:4eadd49703af34fb4da225005ab9aa0ce91af0b014d14bb6ddb997ec68466292

Observation f66c94d8-eb33-4d30-81ea-cca8aa979c40 · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:10:49.194264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.707788Z digest=sha256:576c703623f8b411b525435b65f2f1f19e1768143d43f7c946755c586ac09be2

Observation 7c01b7eb-38dc-4cb7-a07d-ee093e2e55d1 · outbound

This paper cites Accelerated Policy Learning with Parallel Differentiable Simulation.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Accelerated Policy Learning with Parallel Differentiable Simulation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.710302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.710302Z digest=sha256:7a776ea273a3b3a5d2dde2024170b58b258e1fdcf5dd4957b5d9efb2d0f5af9f

Observation 35d460f6-6cb4-440d-b663-31922bd13f68 · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:10:49.184123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.713293Z digest=sha256:c3e06b087139a8d17730be10969e4e724d9fa79854657aee39dfad70817bb869

Observation 914e794d-130e-4aae-8efb-3407f3df2b52 · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:10:49.173553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.716317Z digest=sha256:c52f95485ace1a0320e8c9dfd9ad0b2907a17780b415b2b6639bd98234f117af

Observation 64f414f3-b693-4cc7-8005-2d0709843394 · outbound

This paper cites an unresolved cited work.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.719001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.719001Z digest=sha256:240c1d3c34afec88c9663fcf742506690dc1f022e27dcb2d2a87bd1a748529ef

Observation f144207a-ff4e-4f5d-99db-4c575ce70b95 · outbound

This paper cites Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T15:10:48.721707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:10:48.721707Z digest=sha256:b70a460a81053a2dc6b8c34395fb0f7c34eaad596bbe83c2a1b5af1dfb485f6b

Observation 92b81cd6-9db5-4b6b-acce-458de32401a5 · outbound

This paper cites Raffin, A.

ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards Raffin, A

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:10:49.157346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-10T15:10:48.724636Z digest=sha256:4cacea2305caf5d2094d73cb2fa7af44d1795c67b9cf55dab7ae487b44753af0

Pith citing papers

Observation 9b8e006d-5641-4025-aac8-5e7acac6e7c9 · inbound

Simple but Stable, Fast and Safe: Achieve End-to-end Control by High-Fidelity Differentiable Simulation cites this paper.

Simple but Stable, Fast and Safe: Achieve End-to-end Control by High-Fidelity Differentiable Simulation ABPT: Amended Backpropagation through Time with Partially Differentiable Rewards

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:50:59.094277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T16:27:25.150807Z digest=sha256:a2a74bee1da91ce4e9fb0b5efbb6e09c00e111d59c463d37a78fd11292885c18