Pith. sign in

Paper Citation Record · LEDGER

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence

As of 15 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2508.00299.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.00299 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:16:25.195091Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f1709da7-3301-4bd1-9047-90e206426fb2 · outbound

This paper cites nuscenes: A multi- modal dataset for autonomous driving.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence nuscenes: A multi- modal dataset for autonomous driving

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:26.185487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:24.899974Z digest=sha256:39e55e7a0d695f292011aa9bc8d8b327513c4632f0c71c6c0734d5a5f8f528cc

Observation 344670a6-f5c9-433e-b2d2-1274f5d0436c · outbound

This paper cites Realtime multi-person 2d pose estimation using part affinity fields.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Realtime multi-person 2d pose estimation using part affinity fields

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:26.164549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:24.906469Z digest=sha256:7b8d81520df4614dd62267e16281f87dcd880b5eaa8e0a0b506229f54940264c

Observation 2d753344-c695-471f-a7e5-00e479dbfade · outbound

This paper cites MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:24.912356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:24.912356Z digest=sha256:6a31698bedfacc7ec115712c5e5f68a94c17417263582b78347ffcbecf1831e2

Observation 3b5b99bc-8835-4833-955b-e8aa173fa971 · outbound

This paper cites Structure and content-guided video synthesis with diffusion models.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Structure and content-guided video synthesis with diffusion models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:26.137906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:24.919206Z digest=sha256:dfac404c09cc6e3f6ae8f31b09284fb694c2fb74440a9c5c14239a275173f05f

Observation e04cbcc0-5027-43ca-b8dc-f6f4fa088e25 · outbound

This paper cites MagicDrive3D: Controllable 3D Generation for Any-View Rendering in Street Scenes.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence MagicDrive3D: Controllable 3D Generation for Any-View Rendering in Street Scenes

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:24.928194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:24.928194Z digest=sha256:8ac1ec4aad9a34929345a24412a229e2b1f10c7a65b48e9969a005169e7ae1f8

Observation a0b0984a-ad08-4229-97c8-ea0ac9aa733f · outbound

This paper cites Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllability

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:24.935021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:24.935021Z digest=sha256:ebd9b4267c5db65a10120911a728014b7e877c17bcce4e325471751fea15758f

Observation 646d4b11-bdc7-44b7-92ab-4b3199161c35 · outbound

This paper cites 8 Densepose: Dense human pose estimation in the wild.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence 8 Densepose: Dense human pose estimation in the wild

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:26.109398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:24.945679Z digest=sha256:4e437a4cf55b33265409fb478b1b4d71ac5d6ea8a7e09836ce12603f138f7474

Observation 9c1c7d5b-4774-4b90-9fd9-d732087ce681 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:24.951761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:24.951761Z digest=sha256:24859acfdf1e36c10d8dbeb03c25132846acd0b3671a1ef770c19f2b54c27691

Observation 2c446e23-4e59-48f8-9921-2e145ef193f8 · outbound

This paper cites Animate anyone: Consistent and controllable image- to-video synthesis for character animation.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Animate anyone: Consistent and controllable image- to-video synthesis for character animation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:26.086274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:24.961062Z digest=sha256:9b86e254d420b3e065476acaa6bae015ec5c739cb9e7be72a381a12307489825

Observation 0c54ec4b-1903-412c-a5c1-c33ef59082f3 · outbound

This paper cites Composer: creative and controllable im- age synthesis with composable conditions.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Composer: creative and controllable im- age synthesis with composable conditions

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:26.061135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:24.967038Z digest=sha256:98be90844be0f6a0631ccc7b56d5882b2110735ecdfe6218420ca732fa85155f

Observation 7f3ebcc3-659e-4b6d-8079-f3139bf4e40c · outbound

This paper cites Text2performer: Text- driven human video generation.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Text2performer: Text- driven human video generation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:26.026724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:24.973620Z digest=sha256:9fd4c50d023ab1320bf9d98968dbd159ea89121f5656a93c53a75dd762e263c3

Observation c05a73ef-79e9-4afe-b590-977e1d27adf9 · outbound

This paper cites Dreampose: Fashion image-to-video synthesis via stable diffusion.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Dreampose: Fashion image-to-video synthesis via stable diffusion

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:26.000347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:24.980160Z digest=sha256:ae9210041f9e3db9876f04698ac942a020aa5d3d92270063778772a642e80bbd

Observation d265aa3a-7c9f-443e-bc05-d5566f713e7d · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence 3d gaussian splatting for real-time radiance field rendering

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:24.985880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:24.985880Z digest=sha256:8bbe5ef7746076a40ef84e21cec5a47a9ea10e79a02d3b5f7f0e1daacbbfad1f

Observation c68de3f0-fb6f-45ef-a711-eccf83756222 · outbound

This paper cites Text2video-zero: Text- to-image diffusion models are zero-shot video generators.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Text2video-zero: Text- to-image diffusion models are zero-shot video generators

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.962621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:24.996857Z digest=sha256:8dfa109ca875e7a44e1eda76e62a3a46988203fe1b1de444bebca2c98c363ed4

Observation c43b8eca-1d7c-4971-9c64-66c158d3560c · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.003240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.003240Z digest=sha256:a47ec5980da475efad15a29d47c2242348b3a7ca0b70d7a183e38d33fb9e49bb

Observation 15f9c03f-ab2a-4c2c-ae1d-1ac1c14c2c01 · outbound

This paper cites Smpl: a skinned multi- person linear model.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Smpl: a skinned multi- person linear model

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.939959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.009837Z digest=sha256:7937123e7813de04c6f4e6bce77e99419d1a0479c47ec6f3e0078026ad6f5b8e

Observation 4a51c89b-3c3c-4abe-912a-984fc405fb81 · outbound

This paper cites Follow your pose: Pose- guided text-to-video generation using pose-free videos.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Follow your pose: Pose- guided text-to-video generation using pose-free videos

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.015386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.015386Z digest=sha256:c8823138364a4245b2def0eca1f704fa15e809645c4a27db609b2d344a2959e4

Observation 9f3c09a6-a6c0-4a6b-800d-c5d6c0ec6e26 · outbound

This paper cites T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.021308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.021308Z digest=sha256:4b39cbe2bf6e2b934a6cb33209dc64047054ed373912bc685ed81cd856550f80

Observation f916f8f5-d6ec-4e8a-b33d-e1c2f0da9113 · outbound

This paper cites Recondreamer: Crafting world models for driving scene reconstruction via online restora- tion.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Recondreamer: Crafting world models for driving scene reconstruction via online restora- tion

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.867244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.027525Z digest=sha256:7cf03443496658f5134c58b940612b76b799d362b6120e4e203f23488fcd66f1

Observation 33b947be-c410-403d-837c-2cdb30db6f0f · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.034289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.034289Z digest=sha256:5f0a7d86310a67b5d8e981648883fbe1e0de37f1c98f12473bb8fc3558b05c48

Observation a6d1d0ac-9b70-40a4-991e-452020146a0f · outbound

This paper cites Fatezero: Fus- ing attentions for zero-shot text-based video editing.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Fatezero: Fus- ing attentions for zero-shot text-based video editing

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.040307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.040307Z digest=sha256:7ee1207e5a0b9e2d1664fc6a71ab8dbc7446d84888ca2e33f05ebabd93635e17

Observation 3ed40cfc-7d3f-4c3b-b90b-28feca6f9edb · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Learning transferable visual models from natural language supervi- sion

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.046103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.046103Z digest=sha256:e5b43b69f3d67c863052ac6306bb498d41b905a9699a7ea5e39c8b852e5f11b5

Observation 7629ee3c-368d-4faa-9ed2-17d8b213c0bb · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence High-resolution image synthesis with latent diffusion models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.053151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.053151Z digest=sha256:f0438d78ef716d7ee945cfa8954aa1a78d11f09ec66f0da4a3a593239839080b

Observation 293dd31f-36d3-4237-bb29-df108da3dcd8 · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Photorealistic text-to-image diffusion models with deep language understanding

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.060499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.060499Z digest=sha256:7536992382d39abcc1d82afb16a32ca81a3c25022b675b62d818df99b8edf74c

Observation 79e179b4-eba7-4aab-b0c2-e0f7eafe47ac · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.068937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.068937Z digest=sha256:93042e8311d787b6cd93315fbd4960c499784227da90a18efd36535a3ed3df2e

Observation 1c489a80-86e5-4ab7-ac51-649b990bb574 · outbound

This paper cites Object- stitch: Object compositing with diffusion model.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Object- stitch: Object compositing with diffusion model

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.077011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.077011Z digest=sha256:2386b43c72bab908c6af4d7a0374df3a18c77266c8814d47ff419e38682e0a7f

Observation 6529f11c-d287-4fac-953b-5ba1e316d623 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Wan: Open and Advanced Large-Scale Video Generative Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.086406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.086406Z digest=sha256:53d726f461ec0046b4fd0339176c0a733b772b42776f2482c1b7f7cd6f596579

Observation 278b15de-6716-47bc-9cec-64043197d453 · outbound

This paper cites Disco: Disentangled control for realistic human dance generation.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Disco: Disentangled control for realistic human dance generation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.754803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.093477Z digest=sha256:62af5763c7dc5a580451bb9e68b01f5e6e685e003e1afca7089bc83b31f38680

Observation 4a87a7df-35ec-44e2-aade-c41a4a05309e · outbound

This paper cites UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.100015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.100015Z digest=sha256:33219b26dbd4005c0fac36131db20fbe8960a013c01a66c021e26066f77827f7

Observation 7c0d608a-10e5-46ef-bb4f-cd2c987a75c7 · outbound

This paper cites Drivedreamer: Towards real-world- drive world models for autonomous driving.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Drivedreamer: Towards real-world- drive world models for autonomous driving

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.729069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.107949Z digest=sha256:5174cff59a432caa99b231168cb0ca49af016cd692c979f7800848b2034f8f35

Observation ba92fa14-71e3-428c-b3de-32b6294b2deb · outbound

This paper cites Panacea: Panoramic and controllable video generation for autonomous driving.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Panacea: Panoramic and controllable video generation for autonomous driving

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.703052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.115568Z digest=sha256:1c90d051d7352ec7c4aff095d0863d96ab3d72992ab95c11c04132670f1050b1

Observation dc2bce14-12f2-4c6e-92cf-7911afe5a1cc · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.675918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.121320Z digest=sha256:ceb3333fb0363c2e8e116825fac38467e762c4a9681b2db71c3c638502cd1998

Observation 94803cd8-7bf2-4624-9192-3ae2a51a657e · outbound

This paper cites Magicanimate: Temporally consistent human im- age animation using diffusion model.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Magicanimate: Temporally consistent human im- age animation using diffusion model

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.655331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.127100Z digest=sha256:3afeb17cb0ba586873b57e164d71026fe8cd3c78aee0de0b1f20ba39591d22d9

Observation 96b44558-ee8b-4255-a4c8-8a9aadabf8c6 · outbound

This paper cites Effec- tive whole-body pose estimation with two-stages distillation.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Effec- tive whole-body pose estimation with two-stages distillation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.633887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.132995Z digest=sha256:5d544a3470734d45cce0a5baaca1c2dc5194f699225b6d6112272abb060858b8

Observation 5e7b4c7b-6062-441a-a41a-df15c7f93c68 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.143215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.143215Z digest=sha256:2218944b3b45c7928c92b8e29a4749756d58ec7f00dc3d4a6c9d27ad74945c41

Observation 0d1e3167-947e-4909-94d5-33eb166a5415 · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.151569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.151569Z digest=sha256:6527ad0d61aa128f37f3484e9582fd6eaea497bd2ef6a23423df7d34c0523398

Observation cbd7a803-d5ef-44d0-8883-ed06cf4491b6 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Adding conditional control to text-to-image diffusion models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.160399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.160399Z digest=sha256:8f954a96a11f4dcdfc22d60c5fa3850dc9cd51532664b428d9db432913cb5848

Observation b41ac559-04a8-4c33-b39c-e14d02517e70 · outbound

This paper cites MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.167688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.167688Z digest=sha256:37cab4d8dd75012d5a0cefe908d5054a005f0bd6027f79480581ce8232f75dad

Observation 81c170be-a10b-40bd-b804-9404866298e7 · outbound

This paper cites Drivedreamer4d: World models are effective data machines for 4d driving scene rep- resentation.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Drivedreamer4d: World models are effective data machines for 4d driving scene rep- resentation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.602121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.174086Z digest=sha256:776168f90a19f790992a210cf4fea70239b662449fa2bc7b853aae491b89389b

Observation 5e484dbd-cf9a-43d9-a7ef-b36e74bcbdc8 · outbound

This paper cites Drivedreamer-2: Llm-enhanced world models for diverse driving video generation.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Drivedreamer-2: Llm-enhanced world models for diverse driving video generation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.581456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.182522Z digest=sha256:f43f12c31def23270f8bee9e5584b0c14afdffd15dfcdc606d19460a275c2688

Observation 1441e9cf-66fc-4c77-bae3-035f902a7308 · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Open-Sora: Democratizing Efficient Video Production for All

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T10:16:25.189045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:16:25.189045Z digest=sha256:c85491fe4b60eaf4bfd9a6fe77e3570769993dba687491d30810cb878381c520

Observation 302a763e-381b-46db-917f-d4535f5dba29 · outbound

This paper cites Champ: Controllable and consistent human image an- imation with 3d parametric guidance.

Controllable Pedestrian Video Editing for Multi-View Driving Scenarios via Motion Sequence Champ: Controllable and consistent human image an- imation with 3d parametric guidance

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:16:25.559561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T10:16:25.195091Z digest=sha256:5a9859873a15c858b55aaeadf29a33e5d64dbcf1c26c3234a2a9a911a5930400

Pith citing papers

No inbound Pith citation observations are available.