Pith. sign in

Paper Citation Record · LEDGER

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning

As of 10 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 2 inbound Pith citation observations for arXiv:2507.12977.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.12977 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:38:36.716242Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:34:38.462448Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T02:38:17.343103Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact1
  • verified fuzzy36
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation faf8fca2-41d8-4109-8203-e470469574af · outbound

This paper cites St-p3: End- to-end vision-based autonomous driving via spatial-temporal feature learning,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning St-p3: End- to-end vision-based autonomous driving via spatial-temporal feature learning,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:47.150044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:30.038782Z digest=sha256:0a70a985390e75f538ed3e455aae72e0a57be069e16415a5696aa76f6d9b561f

Observation c6ab466e-753e-4f9c-a006-8e9635ed060b · outbound

This paper cites Per- ceive, predict, and plan: Safe motion planning through interpretable semantic representations,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Per- ceive, predict, and plan: Safe motion planning through interpretable semantic representations,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:46.831482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:30.116542Z digest=sha256:2bd65314ae4529d50ae1c69833e08abb414868694ddbf5eade4cd01532c594ee

Observation 66c0482e-4272-4b80-8524-1b3f73df66bc · outbound

This paper cites Dsdnet: Deep structured self-driving network,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Dsdnet: Deep structured self-driving network,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:46.581600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:30.242568Z digest=sha256:fdc9bb7d04a0b853d5e6b3d8913a90d6240301983b0b1f212f039b4ef4b51007

Observation 1d7728ca-2930-480e-8e9d-d10de4f472a0 · outbound

This paper cites Multi-modal knowl- edge distillation-based human trajectory forecasting,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Multi-modal knowl- edge distillation-based human trajectory forecasting,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:46.330651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:30.405170Z digest=sha256:2a3bcdad27101d18a0e1b38c67d14493392b9e5c7ae40b20729efdc2fd01b3ee

Observation e5b3705b-37f2-4c92-bb5a-dc368687b641 · outbound

This paper cites Fast-replanning motion control for non-holonomic vehicles with aborting a*,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Fast-replanning motion control for non-holonomic vehicles with aborting a*,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:46.085755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:30.516413Z digest=sha256:2f45d3e88630fd03e9adb2502cb998f1a1969ec040010bd1144872843b026943

Observation 22693db8-4841-4acb-8402-031293c46cf0 · outbound

This paper cites Informed rrt*: Optimal sampling-based path planning focused via direct sampling of an admissible ellipsoidal heuristic,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Informed rrt*: Optimal sampling-based path planning focused via direct sampling of an admissible ellipsoidal heuristic,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:45.826865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:30.672535Z digest=sha256:dc6e17491b21a15e942b39dccb2550f93bc116d2cdc07321db92fc92f270c0ce

Observation 7c348d81-fff9-49f8-842a-7c9e1a1ed09e · outbound

This paper cites Path planning using neural a* search,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Path planning using neural a* search,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:45.632277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:30.823754Z digest=sha256:a8a80f7070b244f73cae3492673d288e50b47eca0ee88bb1919a5e8590508bac

Observation a6c96eca-9795-4be6-a548-d90b8f389437 · outbound

This paper cites Sampling-based algorithms for optimal motion planning,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Sampling-based algorithms for optimal motion planning,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:30.938700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:30.938700Z digest=sha256:dc84e937cb8b8722e358491a55cab7f447f916c542364cd043477a94a78efaff

Observation 96472994-1514-49d9-b45a-517e6eb99ca2 · outbound

This paper cites Deep imitation learning for autonomous driving in generic urban scenarios with enhanced safety,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Deep imitation learning for autonomous driving in generic urban scenarios with enhanced safety,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:45.440891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:31.065718Z digest=sha256:045a58d5d86230164bc31e2e4fd5d59091fa450c566c7f2e6f51bdb706bb20c1

Observation dd7a703f-1053-4687-8ca0-bb91cc1921d0 · outbound

This paper cites Safe reinforcement learning with stability guarantee for motion planning of autonomous vehicles,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Safe reinforcement learning with stability guarantee for motion planning of autonomous vehicles,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:45.213239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:31.176073Z digest=sha256:fdac384bb749725f12bd28cff46c7fc6a83a4f1468f2811382c634c2d94be1f0

Observation 233752d2-440c-4f18-9d56-686eea03e744 · outbound

This paper cites Parting with misconceptions about learning-based vehicle motion planning,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Parting with misconceptions about learning-based vehicle motion planning,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:44.980803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:31.336254Z digest=sha256:e08f11f823cd29413a1c60fd293de56d5c2781e67f6b1a85782a00d947c1078d

Observation 424207c6-bbf6-45dd-9d4e-f28c123a3bde · outbound

This paper cites Differentiable Constrained Imitation Learning for Robot Motion Planning and Control.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Differentiable Constrained Imitation Learning for Robot Motion Planning and Control

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:31.433701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:31.433701Z digest=sha256:d7f4937e66e14e9c7566b009ce6458d1e273f68cf3731c30ff1dce9f1fd5f5dc

Observation 71c2c8ac-a34f-407b-bbaf-910628542620 · outbound

This paper cites Diffusion-es: Gradient-free planning with diffusion for autonomous and instruction-guided driving,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Diffusion-es: Gradient-free planning with diffusion for autonomous and instruction-guided driving,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:44.782389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:31.584764Z digest=sha256:35fb6134c8a16a505864aa858b12eef94ea04a937c326db0eff7255283ae6778

Observation f4136118-a487-43f7-8e4d-7c557a609ca7 · outbound

This paper cites Motiondiffuser: Controllable multi-agent motion prediction using diffusion,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Motiondiffuser: Controllable multi-agent motion prediction using diffusion,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:44.543937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:31.696711Z digest=sha256:31d0cbdfcce79a98ed024d7910c172a0a15d4b29e28a4996197a86730e2c85d4

Observation 56940f92-5fd7-4254-8b75-aed2a20daac9 · outbound

This paper cites Diffusion models beat gans on image synthesis,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Diffusion models beat gans on image synthesis,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:31.813252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:31.813252Z digest=sha256:152c768b930726dc866fcff665492033db2ffc64b373d36032b573be3faab221

Observation ce492428-58aa-4220-8a1b-5977589a9920 · outbound

This paper cites Language-guided traffic simulation via scene-level diffu- sion,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Language-guided traffic simulation via scene-level diffu- sion,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:44.267630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:31.960323Z digest=sha256:3006fb527468c5a358e14d1b76551c67130af8ebd73507c3a5ebbb3fa6002405

Observation d1e7c7ef-8d61-4728-8256-36fc0c2fa831 · outbound

This paper cites Safe-sim: Safety-critical closed-loop traffic simulation with diffusion- controllable adversaries,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Safe-sim: Safety-critical closed-loop traffic simulation with diffusion- controllable adversaries,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:43.946422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:32.094513Z digest=sha256:142fe2c2b3e61a46b11436cf0ea2e7062c4f623dc28c4efad0497a197c7aceff

Observation 40bb854f-01ac-48bd-aa67-351e05df5eba · outbound

This paper cites Fine-Tuning Language Models with Reward Learning on Policy.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Fine-Tuning Language Models with Reward Learning on Policy

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:38:37.143823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:32.272467Z digest=sha256:796cacb613dcaf18cb0ec13c51302d05c9b9615f47c3cc042a5e18fda6d903c9

Observation 1404fe10-6cb0-471d-8729-1f59112944bf · outbound

This paper cites Training diffusion models with reinforcement learning,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Training diffusion models with reinforcement learning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:43.624323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:32.426350Z digest=sha256:d4357a898d3a920e43fb9a91d05257d8bccb6768fc3ec79a7a4578cfdb69d4c9

Observation e3af2e56-e04b-48c8-9eb3-fed313891c33 · outbound

This paper cites Denoising diffusion probabilistic models,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Denoising diffusion probabilistic models,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:32.615647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:32.615647Z digest=sha256:cc0fcffb83ea64b0a904907003a8b4d2fb967ecc855c8ef3c83bf476670a65d6

Observation a0861cf5-8a3e-4acf-b34c-0c0acbb43bd0 · outbound

This paper cites Improved denoising diffusion prob- abilistic models,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Improved denoising diffusion prob- abilistic models,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:43.355616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:32.731481Z digest=sha256:ef7f7ffe98335705f40c54ee23d52e4668df4e4f680cca3421e76a3b03cc055b

Observation 6ad6cba7-d1f8-492e-bf7c-8418c0efb21a · outbound

This paper cites Improving transferability for cross- domain trajectory prediction via neural stochastic differential equa- tion,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Improving transferability for cross- domain trajectory prediction via neural stochastic differential equa- tion,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:43.044648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:32.880110Z digest=sha256:9e1ab5c228c76f9035347d1f63d9b7908c373eb4024c67f16b25ed41cade8707

Observation 92f19643-57e4-4231-b71d-872aeed50c45 · outbound

This paper cites Maximum likelihood training of score-based diffusion models,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Maximum likelihood training of score-based diffusion models,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:42.756492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:33.044271Z digest=sha256:574113c4789459db1851f7b930706018839458e182c4c47ab59f4e4f85706977

Observation 022fdce8-2253-41df-84a9-a17987f66f26 · outbound

This paper cites Data-driven Diffusion Models for Enhancing Safety in Autonomous Vehicle Traffic Simulations.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Data-driven Diffusion Models for Enhancing Safety in Autonomous Vehicle Traffic Simulations

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:33.177376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:33.177376Z digest=sha256:8b72a2e5f66e32cf80e9170065cf80c0862d201d64ed86d336137e09fb3578bc

Observation 5f04bf3a-a7c5-4256-b534-67cd80d675f5 · outbound

This paper cites Planning with Diffusion for Flexible Behavior Synthesis.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Planning with Diffusion for Flexible Behavior Synthesis

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:33.283453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:33.283453Z digest=sha256:9c0ab0ae073bee759c2ca375c417fe504040f564d0c34babb4c7c61d606f80ce

Observation 5a0a298e-080c-43fa-95ac-a15f4603fb61 · outbound

This paper cites Leveraging future relationship reasoning for vehicle trajectory prediction,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Leveraging future relationship reasoning for vehicle trajectory prediction,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:42.507763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:33.400990Z digest=sha256:4c0e77b4bfd982d2465ebb6d65cbe6ca96a34feadb33537227b49e8fd4015368

Observation 4dc59fa4-4fac-47b7-9879-ccd76c0ca9fd · outbound

This paper cites an unresolved cited work.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:33.539006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:33.539006Z digest=sha256:03215a054e6f60948e66a82898ebe5ee3b39f1dc15043ba872b325ff38f25bf0

Observation 697deb55-7215-4d23-aa20-cd32d652cc1d · outbound

This paper cites Human-level control through deep reinforcement learning,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Human-level control through deep reinforcement learning,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:42.194764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:33.686261Z digest=sha256:d0970a5b4c36ecb39d37b0476f90c27e1ba5eb9b8ca37b7960b92acc95ec0e55

Observation bd6870bd-9931-4757-b50a-81e8076a8731 · outbound

This paper cites an unresolved cited work.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:33.811079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:33.811079Z digest=sha256:603bf36eb7965e0f3e0d0182066abe1c805ab1f2ec2f90715d798667a701c951

Observation a9ef963a-46d0-4348-8568-0c152cea7b0d · outbound

This paper cites A markovian decision process,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning A markovian decision process,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:33.928478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:33.928478Z digest=sha256:d047760d95e72fec604e72d85ea9e2b981253fa512cfb7d0c56d7a459e1e12bf

Observation 65ec2c56-ae69-4d0d-b0d5-efac8f5567b4 · outbound

This paper cites Deterministic policy gradient algorithms,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Deterministic policy gradient algorithms,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:41.817679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:34.055208Z digest=sha256:8363b0f896dee7f1369bbc9281c9cdfed29b97814218667cb4c1507d66b43bdb

Observation e1163659-b9b2-4450-8d61-959a4728aaa3 · outbound

This paper cites Comprehensive reactive safety: No need for a trajectory if you have a strategy,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Comprehensive reactive safety: No need for a trajectory if you have a strategy,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:41.510705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:34.275662Z digest=sha256:34fbe8b65acc10eb2b84946b217c7840d68a5215d0f14bcec3c100fd7e5954a3

Observation 43a29ff3-811f-40de-8e6a-cacfe5926601 · outbound

This paper cites Autonomous driving motion planning with constrained iterative lqr,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Autonomous driving motion planning with constrained iterative lqr,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:41.216894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:34.388940Z digest=sha256:49a597c65a196c500a89da56ad5e58b2fe05a6eebe50a14f73ec85ddfb017787

Observation e57640ad-2962-4086-b6e5-013ec4440039 · outbound

This paper cites Leader: Learning attention over driving behaviors for planning under uncertainty,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Leader: Learning attention over driving behaviors for planning under uncertainty,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:40.876894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:34.495658Z digest=sha256:206cdfa7f52ab82bfa18a9313fe1f788191ba7dc4210e91367224de457a586aa

Observation 4e2ae681-4809-4b72-8e39-20da45b38e36 · outbound

This paper cites Kb-tree: Learnable and continuous monte- carlo tree search for autonomous driving planning,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Kb-tree: Learnable and continuous monte- carlo tree search for autonomous driving planning,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:40.618371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:34.660594Z digest=sha256:563b4bba9fc281264d53456ae5c3f69eed35b05190c5a0fa737ffdfc82461ad4

Observation cc8420a6-6c07-47f9-9e84-c42e8b3b83ea · outbound

This paper cites Driving maneuvers prediction based autonomous driving control by deep monte carlo tree search,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Driving maneuvers prediction based autonomous driving control by deep monte carlo tree search,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:40.307644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:34.837308Z digest=sha256:81e1ec702527ce3fac5681fd60b1442e5c760eecd892e757ba25fa4a31788fdf

Observation 5a5bd1f6-1ae4-48ce-8f39-3ec7acbff972 · outbound

This paper cites Crowd-robot interaction: Crowd-aware robot navigation with attention-based deep reinforce- ment learning,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Crowd-robot interaction: Crowd-aware robot navigation with attention-based deep reinforce- ment learning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:40.057558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:35.004659Z digest=sha256:be9be77ec8a01c7fb50656ff2e79ccb1ffa137044a3aaffcd77d2545ef2c498a

Observation 2426f62b-00e3-43c9-bb91-ba36bea1ced6 · outbound

This paper cites Rethinking closed-loop training for autonomous driving,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Rethinking closed-loop training for autonomous driving,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:39.793267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:35.175208Z digest=sha256:ea49a2c8e859f62db173bc6e41a1e8b27dfa8e1a7d93f828658b6ab4e62b30dd

Observation cfa90fd5-bbc6-4a6f-ab0d-a081db8ea412 · outbound

This paper cites UMBRELLA: Uncertainty-Aware Model-Based Offline Reinforcement Learning Leveraging Planning.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning UMBRELLA: Uncertainty-Aware Model-Based Offline Reinforcement Learning Leveraging Planning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:35.279970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:35.279970Z digest=sha256:bc8623066657232b781bd676ececec4bbbeab599a988d64c2604bd5bbe40ac64

Observation 5480b033-6841-4eb0-8a9e-36d72bf2e98d · outbound

This paper cites Model-Based Reinforcement Learning for Atari.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Model-Based Reinforcement Learning for Atari

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:35.402897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:35.402897Z digest=sha256:723e1e18f07b929d6cef6669b2964e6237135cbd037910711d9861fe14590cb2

Observation e4b71ebc-b3b0-4a22-97d6-0684ffe7abb5 · outbound

This paper cites Approximately optimal approximate reinforcement learning,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Approximately optimal approximate reinforcement learning,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:39.465099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:35.558259Z digest=sha256:f763fc9aab8fa2ce5dc8aee8247a710cebcca61328ee0f747dee26814176819a

Observation cfc50aff-d961-484a-8a3a-33a276314092 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T16:38:35.692097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:38:35.692097Z digest=sha256:8c11b23204260346f4924943e06c4a9108236804be8b1aeab7932d39f3297813

Observation 8bee159c-9eb9-454b-ab7e-b295b75b1d86 · outbound

This paper cites You’ll never walk alone: Modeling social behavior for multi-target tracking,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning You’ll never walk alone: Modeling social behavior for multi-target tracking,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:39.188418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:35.873363Z digest=sha256:8e19248569720403b0d4835b4c9ed9b8c89fac8fb83598aa1c0f1f519e3b5d79

Observation b87a31c1-40b2-4398-9c5b-dc162035a206 · outbound

This paper cites Crowds by example,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Crowds by example,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:38.877846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:36.019114Z digest=sha256:b591676b1b8e1a8eb4857bfca30067ce85c337f51d22195a6cedf5bc822fb267

Observation fbb7d610-3bb4-4584-8ab6-f7114cbe6d42 · outbound

This paper cites A game-theoretic framework for joint forecasting and planning,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning A game-theoretic framework for joint forecasting and planning,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:38.596213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:36.142468Z digest=sha256:96d28e40136e7110fe8bebf76cf4945aa066fbb7bf289273a10b02deebfad5b8

Observation a8c5f799-b57f-4978-a72f-d43ba4163ddf · outbound

This paper cites Human trajectory prediction via neural social physics,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Human trajectory prediction via neural social physics,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:38.346033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:36.311109Z digest=sha256:699be9284803ebc4a7758312c36f67f155469ce518e41337df568bf1feddb6c0

Observation 36f629b8-3a35-4601-9aa3-1289bb3193c4 · outbound

This paper cites Dtpp: Differentiable joint conditional prediction and cost evaluation for tree policy planning in autonomous driving,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Dtpp: Differentiable joint conditional prediction and cost evaluation for tree policy planning in autonomous driving,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:38.084375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:36.450342Z digest=sha256:e7d1ddb26b1d3f39d0c9877404dc4f30cdafb63a6839bdebdf01a5a08e6823e4

Observation 18d390bf-2c40-41fc-9771-baac2345d6b5 · outbound

This paper cites Differentiable integrated motion prediction and planning with learnable cost function for autonomous driving,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Differentiable integrated motion prediction and planning with learnable cost function for autonomous driving,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:37.809300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:36.598662Z digest=sha256:0680c1db296db10942adf838279fa8340fdde12a9ed946ca6517f49aa9d96c49

Observation 1bbfc78a-f87f-4e42-a868-01890a07cac1 · outbound

This paper cites Stochastic trajectory prediction via motion indeterminacy diffusion,.

Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning Stochastic trajectory prediction via motion indeterminacy diffusion,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:38:37.526520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T16:38:36.716242Z digest=sha256:a456f8c905a707424554298f8fdedb988a6b5242c306c1b6b9cb6c94efe80e42

Pith citing papers

Observation 9477e58b-7f42-474b-930f-950d54d316e3 · inbound

Generative Active Learning for Long-tail Trajectory Prediction via Controllable Diffusion Model cites this paper.

Generative Active Learning for Long-tail Trajectory Prediction via Controllable Diffusion Model Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T11:34:38.462448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:34:38.462448Z digest=sha256:ad0b5959d6581b1e2b66944775f78696341ad8da420ca92f8f76b6b87e5332a4

Observation 5e59496e-71a2-4290-a56d-2d6428f439fa · inbound

Multimodal embodiment-aware navigation transformer cites this paper.

Multimodal embodiment-aware navigation transformer Non-differentiable Reward Optimization for Diffusion-based Autonomous Motion Planning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-10T02:38:17.344445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T02:37:50.624886Z digest=sha256:ee3dfb8be007774652d257e11a58cded91e0ec44aef4be10fc4106c679f5d429