Pith. sign in

Paper Citation Record · LEDGER

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models

As of 23 August 2026, this Paper Citation Record lists 86 of 86 outbound references and 2 inbound Pith citation observations for arXiv:2507.13344.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13344 v1

Coverage vector

measured 86 of 86 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:30:34.764898Z

measured 88 of 88 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T00:22:55.503532Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T00:25:32.931076Z

Reference resolution

86 of 86 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8571a637-e73d-4f7f-855d-d5a9b3b2307b · outbound

This paper cites an unresolved cited work.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.420799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.420799Z digest=sha256:38c21749b1de8044c7bf1b74cf60194d0ff39f793c3dbc0d3a83966b14704c46

Observation b9608fa7-51de-47d9-aa6d-230bc712c26f · outbound

This paper cites Creation of 3d human avatar using kinect.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Creation of 3d human avatar using kinect

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.425120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.425120Z digest=sha256:2fe46dfd7f768dc9e6dc7cc8e5e7f0b0ab84bce61ba306d60aa7ccd3608dad76

Observation 37df16ba-3007-4cda-8f54-942195d81f19 · outbound

This paper cites Tc4d: Trajectory-conditioned text-to-4d generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Tc4d: Trajectory-conditioned text-to-4d generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.429227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.429227Z digest=sha256:c6233ca77b82bec597f338e3a42f21a92ec6446a53694f7b29c08e0eca81882a

Observation f1515e09-b1b1-4f20-827c-c2fa1ab62330 · outbound

This paper cites 4d-fy: Text-to-4d generation using hybrid score dis- tillation sampling.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4d-fy: Text-to-4d generation using hybrid score dis- tillation sampling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.433775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.433775Z digest=sha256:2561b68de7ab0ba38225834049126d75ac1fcb52ac86b302918ca79f34b353e4

Observation 8e6c181b-52a0-45f0-8d1e-2bbfcfc8f81a · outbound

This paper cites Detailed full-body reconstructions of moving peo- ple from monocular rgb-d sequences.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Detailed full-body reconstructions of moving peo- ple from monocular rgb-d sequences

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.437662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.437662Z digest=sha256:8a41808b08dcf9a49181f9468a2b96d6ba2289dc81c0fb167d98e6bd57ea6136

Observation 3e90e633-72a8-449c-a8d7-26581edcb1e1 · outbound

This paper cites Video generation models as world simulators.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Video generation models as world simulators

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.441532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.441532Z digest=sha256:4a74510764c2b9819cc6cb26a012085c94d24db4f9a90ea2d57b5d1231fa0530

Observation 7f9d07fd-e859-4e60-9c03-e96119de7221 · outbound

This paper cites Hexplane: A fast representa- tion for dynamic scenes.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Hexplane: A fast representa- tion for dynamic scenes

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.445677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.445677Z digest=sha256:d35719ca382bcc9e57a82ca2c3ce1ca8f47cecf4cb7ea749966af92aa7632b92

Observation 35acfce3-5e25-4342-a3e4-6988038785cf · outbound

This paper cites pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.742244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.450060Z digest=sha256:ae97f56bfba1f93d52cd3fcd2e71724b1f8c35d923f1bc43e576d5591984635f

Observation 5ed6c430-2746-45c9-8dbe-79c1b92c16ce · outbound

This paper cites MVSplat: Efficient 3D Gaussian Splatting from Sparse Multi-View Images.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models MVSplat: Efficient 3D Gaussian Splatting from Sparse Multi-View Images

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.453665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.453665Z digest=sha256:dcb0ccf64138ffc144b7372123e3e506df85ac2655b4dfd77b2c6232e8b0048c

Observation 41e765d0-2b3f-47a7-a50c-f944d2ee17b8 · outbound

This paper cites DNA-Rendering: A Diverse Neural Actor Repository for High-Fidelity Human-centric Rendering.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models DNA-Rendering: A Diverse Neural Actor Repository for High-Fidelity Human-centric Rendering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.457845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.457845Z digest=sha256:a446cc2c1c9035c81eb6ba6e79d02c029392444710e5cc65e445cab245481b48

Observation 3487a70a-6586-4f65-94e6-11311f655dce · outbound

This paper cites High-quality streamable free-viewpoint video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models High-quality streamable free-viewpoint video

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.463175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.463175Z digest=sha256:ffced4bc7f3a5f2f34484577cd46d1df4f0e748e310bc0b12eb15e72b7af2ef4

Observation 57b4a544-a3ba-4c54-b5c4-de162356275a · outbound

This paper cites Objaverse: A Universe of Annotated 3D Objects.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Objaverse: A Universe of Annotated 3D Objects

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.468225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.468225Z digest=sha256:7b41eba15f466e339e78d4dda45fbc940289e181563affc3ea7391175598802e

Observation 07c93530-0bef-4680-9950-d5a12ea4c3b3 · outbound

This paper cites Objaverse-XL: A Universe of 10M+ 3D Objects.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Objaverse-XL: A Universe of 10M+ 3D Objects

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.473031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.473031Z digest=sha256:b17be5496888fe5c2d3e8d9f5a5a4680232f2ed67f59639dd10cb684d03b5c97

Observation 3c9e237d-1edb-4b47-b239-9cba292a755c · outbound

This paper cites 4d-rotor gaussian splatting: towards efficient novel view synthesis for dynamic scenes.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4d-rotor gaussian splatting: towards efficient novel view synthesis for dynamic scenes

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.478269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.478269Z digest=sha256:6a3522029fa62b0d4543150100e60280e6f536fa856c4b2bfaa6abde16be02e1

Observation 63044d39-968e-4b00-8c94-4cb1d275358f · outbound

This paper cites Fast dynamic radiance fields with time-aware neural vox- els.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Fast dynamic radiance fields with time-aware neural vox- els

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.482394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.482394Z digest=sha256:f57cc885daed9ae8fb920a0c31b6e8dbcf6c7b77289db687fa6174be549bfcde

Observation ebab4608-131e-46c1-b912-45a500bcd16f · outbound

This paper cites K-planes: Explicit radiance fields in space, time, and appearance.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models K-planes: Explicit radiance fields in space, time, and appearance

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.486116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.486116Z digest=sha256:a77bea9dd106fc3c99cf1d140ad45167a9d8f37205c4f01a8bf28a4a374de6b9

Observation 412dcaad-27c0-451c-ae6d-b8eb5798775e · outbound

This paper cites Massively parallel multiview stereopsis by surface normal diffusion.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Massively parallel multiview stereopsis by surface normal diffusion

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.702801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.489747Z digest=sha256:14b8b6188d2e050722ff55622876485aaac6e0f2b6bfc23d5d4cc9873253b719

Observation e5a64fc4-8fd2-41f3-b092-fa094ce8fcc8 · outbound

This paper cites Srinivasan, Jonathan T.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Srinivasan, Jonathan T

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.691475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.493299Z digest=sha256:396ab004f277619d2b7e63ef5e86ddd8076753d63717c242bfc3028a861ab8da

Observation 7398098b-d54f-40be-98ac-d56d3e690634 · outbound

This paper cites Studio production system for dynamic 3d con- tent.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Studio production system for dynamic 3d con- tent

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.680019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.497112Z digest=sha256:416c972c910d8274193be4841ab1e32a7a463c906f0b5496a92236d221f7b418

Observation 4a3206aa-25ad-4a9d-ba32-0d9fee30c9bf · outbound

This paper cites Viewdiff: 3d-consistent image generation with text-to-image models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Viewdiff: 3d-consistent image generation with text-to-image models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.500818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.500818Z digest=sha256:edf158034a8e05833fc9db855521e07ba17fe119a181b1a6789f0b51c3856173

Observation 9ce25f9d-15f4-4aab-a6a4-28439adf213a · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.505019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.505019Z digest=sha256:a1ca8d0a892ed3602fab636748f721f80cd66fb76d795e1663ef0da2a0294340

Observation c42cf2d4-c328-43b1-b609-b938e70a2715 · outbound

This paper cites Gauhuman: Articulated gaus- sian splatting from monocular human videos.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Gauhuman: Articulated gaus- sian splatting from monocular human videos

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.658452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.509099Z digest=sha256:e3d95a6f696bff792a433d368565077bc2920bd66b3a85f7607329a430c2aed1

Observation 82d4f14e-363a-4454-b731-3c1f90dfd63c · outbound

This paper cites Humanrf: High-fidelity neural radiance fields for humans in motion.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Humanrf: High-fidelity neural radiance fields for humans in motion

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.641401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.513178Z digest=sha256:416d60d239cb1109dee2e9c0c8428cac162b7be6c7caa2031658b9c2afa5d415

Observation c329a7ff-e1fd-4c29-91fd-4d5bde5f8707 · outbound

This paper cites Pl ¨ucker coordinates for lines in the space.Prob- lem Solver Techniques for Applied Computer Science, Com- S-477/577 Course Handout, 3, 2020.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Pl ¨ucker coordinates for lines in the space.Prob- lem Solver Techniques for Applied Computer Science, Com- S-477/577 Course Handout, 3, 2020

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.628383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.517297Z digest=sha256:4335d1123e3af00164f416104bfb2d3f670e7ea82a9619c3e567dba1547d9752

Observation c1115a64-0404-4544-b732-75954c1545f3 · outbound

This paper cites Virtualized reality: Constructing virtual worlds from real scenes.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Virtualized reality: Constructing virtual worlds from real scenes

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.616211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.521154Z digest=sha256:97dd51f8b3f2b840cc79e6c31699f265fdd0b015a7cc593a9f965d8aa153cc82

Observation 819f7afa-a1b7-40c0-b24d-b68a3d82104b · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 3d gaussian splatting for real-time radiance field rendering

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.603769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.524599Z digest=sha256:5cfe542644e10758243dbb676d1e3fef726905dc8b6e9a867805cb23f77ca1b7

Observation 6ef0743c-0968-4fe6-b29d-cadef456ac28 · outbound

This paper cites Sapiens: Foundation for Human Vision Models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Sapiens: Foundation for Human Vision Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.528734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.528734Z digest=sha256:043c8e141dfb98dd50a8a511f9a8f953c59f541a248a427054f8efa903f9223d

Observation 00ad95cc-d284-4256-8955-d3ed759e87b9 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Adam: A Method for Stochastic Optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.532489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.532489Z digest=sha256:c4a364f71003d7917c6f5063cf8f5f679f9fb7e1cb7e2ba9c8b7b8d5612c9c94

Observation 6e6b6568-9bab-4190-9aec-bc1f1ead95c3 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.536274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.536274Z digest=sha256:f5934cdfe603bf88c270b2bb0f6d59666716479a692dbf3c1e08d3f0f23dc86b

Observation 770f6a6c-308c-4c7a-973c-45dc1bdeb87a · outbound

This paper cites A theory of shape by space carving.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models A theory of shape by space carving

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.590851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.540369Z digest=sha256:4ec9ec86b88c4de3c8d893758bc7fe54ab9687319578ed7a7c9c7609d91f0807

Observation aa7ce023-2622-481b-b80b-4f6d67963812 · outbound

This paper cites Vivid-zoo: Multi-view video generation with diffusion model.Advances in Neural Information Processing Systems, 37:62189–62222,.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Vivid-zoo: Multi-view video generation with diffusion model.Advances in Neural Information Processing Systems, 37:62189–62222,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.577477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.543908Z digest=sha256:9c82996eabbc6a7ff7622eb1e59478ec66a377e617512646f9c67afd4a58c4cd

Observation 29b26c72-bd8d-49b8-b107-93cf2560d301 · outbound

This paper cites Neural 3d video synthesis from multi-view video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Neural 3d video synthesis from multi-view video

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.547337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.547337Z digest=sha256:07fc18ba22ec173645bb9cfbf6ddd322051df85342ab438a3fa98c62b5b3ae88

Observation 84bb4f8a-9083-4fc6-87f4-9c849252f5c4 · outbound

This paper cites Ani- matable gaussians: Learning pose-dependent gaussian maps for high-fidelity human avatar modeling.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Ani- matable gaussians: Learning pose-dependent gaussian maps for high-fidelity human avatar modeling

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.557675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.551854Z digest=sha256:15dc9429d178fb60844bfb0a9a509c658d09ec92955193da75b7a52b361bd578

Observation cd2ac44b-3338-4bf8-ae12-f587abed1008 · outbound

This paper cites Efficient neural radiance fields for interactive free-viewpoint video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Efficient neural radiance fields for interactive free-viewpoint video

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.545973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.555847Z digest=sha256:307e718ceb7e2230f6f50267d208738070ce405fc274a13a4c4ec9b4e295d337

Observation 7df6b6d9-b38f-4d19-bda8-baf082d481d9 · outbound

This paper cites Real-time high-resolution background matting.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Real-time high-resolution background matting

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.534476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.560525Z digest=sha256:813d53162924b736d1b5388e2768efbb5c642e7abf46df9c0e48b498a6c45d44

Observation 4e1b59c8-bc22-4328-8b9e-4d310fc6d246 · outbound

This paper cites Raft-stereo: Multilevel recurrent field transforms for stereo matching.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Raft-stereo: Multilevel recurrent field transforms for stereo matching

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.564842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.564842Z digest=sha256:aa872c06f7a6aa7707aa826ae1c3584b88126d4c8f09a07e780be018350a06fe

Observation eca86fea-97ca-430e-a778-5bd85a7cc633 · outbound

This paper cites Zero-1-to-3: Zero-shot one image to 3d object, 2023.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Zero-1-to-3: Zero-shot one image to 3d object, 2023

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.514838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.569260Z digest=sha256:05b50dfc8c2d140009b2aaadc4cfb4f12e119c7c8439e69bfa2bb02abe1e455d

Observation ca18d8f5-57fc-4cae-a43b-475f194e0e6e · outbound

This paper cites Mvsgaussian: Fast generalizable gaussian splatting recon- struction from multi-view stereo.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Mvsgaussian: Fast generalizable gaussian splatting recon- struction from multi-view stereo

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.502262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.574009Z digest=sha256:99bb9496c5f41b644bea1998cf7a2a2f1fdeeb65f1bd34e37dd9096fd1df4a98

Observation b85b8ff9-775d-4be5-a5dc-5a943bc29cae · outbound

This paper cites SyncDreamer: Generating Multiview-consistent Images from a Single-view Image.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models SyncDreamer: Generating Multiview-consistent Images from a Single-view Image

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.578276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.578276Z digest=sha256:79a18698ffa1237234e3a57939fea963a9da4e2f32c41ec96b7ee79547f27d40

Observation 6df977f3-65a2-4882-9929-0db30d540cd6 · outbound

This paper cites Smpl: A skinned multi- person linear model.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Smpl: A skinned multi- person linear model

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.582315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.582315Z digest=sha256:14a595f91899edc792174ee82c3e88e792a19592b06aa2b7613c03dbd4c23966

Observation 6df2aee7-a11a-408d-86ed-6d762fd7be18 · outbound

This paper cites DPM-Solver: A Fast ODE Solver for Diffusion Probabilistic Model Sampling in Around 10 Steps.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models DPM-Solver: A Fast ODE Solver for Diffusion Probabilistic Model Sampling in Around 10 Steps

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.586204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.586204Z digest=sha256:2f94169eb43ff5d9febf489b9d1fb26120af7e763261a8c90bff419b3407f657

Observation 3da23ece-d28f-44a0-b728-6337a9acfe48 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view syn- thesis.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Nerf: Representing scenes as neural radiance fields for view syn- thesis

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.590395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.590395Z digest=sha256:2ca24313f822b546be986e7f9a7d34da23c8b6d411a4615e188913a5afa23e65

Observation e8523673-e024-4775-b9a8-779c17c5f045 · outbound

This paper cites Dynamicfusion: Reconstruction and tracking of non-rigid scenes in real-time.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Dynamicfusion: Reconstruction and tracking of non-rigid scenes in real-time

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.475108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.594323Z digest=sha256:9ece87f0ad4bd0c13b9b4353aae023ee8b2a1c47a599b4a80cb6079ccf1b0906

Observation d8eee92d-9d03-4e63-8bf0-ac7a052f6fc8 · outbound

This paper cites Effi- cient4d: Fast dynamic 3d object generation from a single- view video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Effi- cient4d: Fast dynamic 3d object generation from a single- view video

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.598925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.598925Z digest=sha256:32f3c5f90fb94b99cedb18f33c49875b8adfe21340bac8bb99105c2be5947176

Observation 2b4724ea-e076-402b-a305-ebb61c952642 · outbound

This paper cites Barron, Sofien Bouaziz, Dan B Goldman, Steven M.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Barron, Sofien Bouaziz, Dan B Goldman, Steven M

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.602436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.602436Z digest=sha256:fd639859496d227da3204c887d4f6ba82a386f398090a4343418cf551e81c2ae

Observation fbe7732f-c675-4c74-ba8d-535217bf7715 · outbound

This paper cites Barron, Sofien Bouaziz, Dan B Goldman, Ricardo Martin- Brualla, and Steven M.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Barron, Sofien Bouaziz, Dan B Goldman, Ricardo Martin- Brualla, and Steven M

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.456265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.605845Z digest=sha256:a694edfa41a7ba69f98a25f5fc1101f0a8dd043adfb1d4fa81344960428f9ac3

Observation 0d08a2b5-a4f7-48c2-9398-b13573a35f82 · outbound

This paper cites Ani- matable neural radiance fields for modeling dynamic human bodies.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Ani- matable neural radiance fields for modeling dynamic human bodies

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.444141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.609625Z digest=sha256:c182ad6dbd76e878e011acdd17afe4b94bc171101cd4ecf888927bbd5cf088c9

Observation da7146fd-6d79-4243-85e4-db4e97ea0cba · outbound

This paper cites Neural body: Implicit neural representations with structured latent codes for novel view synthesis of dynamic humans.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Neural body: Implicit neural representations with structured latent codes for novel view synthesis of dynamic humans

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.613234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.613234Z digest=sha256:fd42bad2301df09c96fc9211ec53935de815a853252f6bca239d0656bd14ed43

Observation d75891c5-45a7-4007-b204-f625daac911a · outbound

This paper cites DreamFusion: Text-to-3D using 2D Diffusion.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models DreamFusion: Text-to-3D using 2D Diffusion

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.616681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.616681Z digest=sha256:ec1d8c133e65e65afee8483e7fb4b43c3d984bc0531b465e459160c00c935076

Observation 6c6535d4-9fcf-44f8-896c-5510a5e0cb09 · outbound

This paper cites D-NeRF: Neural Radiance Fields for Dynamic Scenes.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models D-NeRF: Neural Radiance Fields for Dynamic Scenes

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.424968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.620537Z digest=sha256:41f3572de80ffa9631e0d2526a0cd9846c565f9779883708c27242e4798ace40

Observation d88181dc-54f9-44c7-ba87-cac7c5bec71f · outbound

This paper cites DreamGaussian4D: Generative 4D Gaussian Splatting.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models DreamGaussian4D: Generative 4D Gaussian Splatting

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.624666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.624666Z digest=sha256:1e3d46b4966b0aaf264ff622b9f9c72fe91ebca75f53216ea727c5a4f4cb5e87

Observation 6fb2c1b8-9158-42b5-873e-d34c29d1f12b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models High-resolution image synthesis with latent diffusion models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.410653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.628519Z digest=sha256:acd854014e5652c29b65883b814e921c6401fb81b0742a1433054f806320fa9f

Observation 7c2900e9-3051-4dd5-8801-8eeac70d29f8 · outbound

This paper cites Structure-from-motion revisited.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Structure-from-motion revisited

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.397435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.632269Z digest=sha256:dfa9d94809469fc4b89412b6d0269b3013c5ed405c8f99cfa47ef0f5d3352804

Observation 85467ce6-6700-4754-a610-6ba810eae861 · outbound

This paper cites Pixelwise view selection for un- structured multi-view stereo.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Pixelwise view selection for un- structured multi-view stereo

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.386295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.636478Z digest=sha256:fb6fbdcff88302618edcaefeabf93d3c14f334b4749cbb34705a33e0c91ffb04

Observation 109f7b72-5c41-49ff-8fed-6b6d6dc0bb7f · outbound

This paper cites Tensor4d: Efficient neural 4d decomposition for high-fidelity dynamic reconstruction and rendering.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Tensor4d: Efficient neural 4d decomposition for high-fidelity dynamic reconstruction and rendering

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.640007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.640007Z digest=sha256:32726d9d8dd42ce2003507193dc466191b332c7c7f9a23519bd64c0cdf2cc392

Observation 36f7a853-d23e-4e47-96e4-e7f94e51d0c4 · outbound

This paper cites Rapid avatar capture and simulation using commodity depth sensors.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Rapid avatar capture and simulation using commodity depth sensors

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.367216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.643880Z digest=sha256:9a6abe3079764437b5843a3ef0f3faabbcc6434cf9070e8c3ca50800a4319961

Observation c8afb796-2adf-4e6a-8e56-566a988e6d4c · outbound

This paper cites MVDream: Multi-view Diffusion for 3D Generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models MVDream: Multi-view Diffusion for 3D Generation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.647825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.647825Z digest=sha256:a18aa877c071191a6da3a7c2825d55672aa2cfda5f347f92befb3f1150e6e46f

Observation d67ffa65-560f-45f4-8755-0c07a3b4cf69 · outbound

This paper cites Text-To-4D Dynamic Scene Generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Text-To-4D Dynamic Scene Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.652426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.652426Z digest=sha256:db511fa8b855fc7eb34b3c894256de0ceca4af6554764d900e5ab0c3501ca030

Observation e81720c4-0c83-4348-a42b-d4935844e092 · outbound

This paper cites Virtual view synthesis of people from multiple view video sequences.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Virtual view synthesis of people from multiple view video sequences

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.355417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.656745Z digest=sha256:91a4152db50670b3af8e440eed496e21b5891e1e1131456441d9b08520b772c3

Observation 5c534f71-c523-48a8-a77a-d6b5d7159a72 · outbound

This paper cites Surface capture for performance-based animation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Surface capture for performance-based animation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.343986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.660850Z digest=sha256:cf2ba3897601054701433fafc58d6bc535ab180b228f968893023825c2ef0da0

Observation fb674737-27b8-4ded-ae19-96bf00061333 · outbound

This paper cites Scanning 3d full human bodies using kinects.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Scanning 3d full human bodies using kinects

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.332853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.664462Z digest=sha256:0a06d93e8e026b9f5d9d78bedf386179d6dfd7fa24abb05c22abcc92bb9611f5

Observation 37ba4bb2-3918-4330-87f1-b3a0e036720c · outbound

This paper cites Diffusers: State-of-the-art diffu- sion models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Diffusers: State-of-the-art diffu- sion models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.668115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.668115Z digest=sha256:db44ca47404f3118903d065e9c7702fde660312ebaf8f4f39cfb6eece6d26c75

Observation 34eab7ac-c5f1-4ba1-99f3-005bee505737 · outbound

This paper cites Fourier plenoctrees for dynamic radiance field ren- dering in real-time.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Fourier plenoctrees for dynamic radiance field ren- dering in real-time

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.312840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.672266Z digest=sha256:4ece7f730b328a4cef80e6fb3cc5325234992d7166daea6dda0fd0a694129a71

Observation adc958c2-78f9-4542-8d6a-645e4139ff0d · outbound

This paper cites ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.675972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.675972Z digest=sha256:5102a7e145cbf388a1e3feb05320399262663b4d30b5511a033434d75a8121ba

Observation 59f7878a-1b95-44ba-a5b8-a5049e69896e · outbound

This paper cites Controlling Space and Time with Diffusion Models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Controlling Space and Time with Diffusion Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.679622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.679622Z digest=sha256:d69e8cadeac46a7927e05cca5e9014a3ef33310c0b538f127ab1f245fc194fd7

Observation bc604099-22a5-4753-8bc1-3224b8a45be5 · outbound

This paper cites Hu- mannerf: Free-viewpoint rendering of moving people from monocular video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Hu- mannerf: Free-viewpoint rendering of moving people from monocular video

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.683825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.683825Z digest=sha256:85895249904353c664d234d61790af20e68f62fe9098eb2eea33ab7499660304

Observation 08c4c103-43fd-4a18-884c-7c2d26f98bb0 · outbound

This paper cites 4d gaussian splatting for real-time dynamic scene render- ing.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4d gaussian splatting for real-time dynamic scene render- ing

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.294926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.688073Z digest=sha256:ca971be007f25935d3a0ff24616799aa089cbfe1a7c6e1084a6bba39905a2635

Observation 791357f3-5e62-4e85-b0e5-dd097bf68ffb · outbound

This paper cites CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.692191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.692191Z digest=sha256:1d9de36da2d46cfe5801823d9db2969e63c0f64c1fbb0f3fc181e8c9950d8e56

Observation 185d1a01-d5b7-4354-a110-2222020e9d07 · outbound

This paper cites SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.696845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.696845Z digest=sha256:d7ccc5679ce4a3553dcd8d065c4c0ed1d62af5ef1e663c45a27ed9a52b36b4a6

Observation 39c07d01-d3e0-4685-b97b-b45c7187bd69 · outbound

This paper cites Easyvolcap: Accelerating neural volumetric video research.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Easyvolcap: Accelerating neural volumetric video research

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.283732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.700951Z digest=sha256:bc534ef06ad58366f51581faf53b7ed7ba230834c7ffd85eca51cae0a9a7eb7a

Observation a70031ba-c695-4a77-9dfc-fa5e4c2d674a · outbound

This paper cites Relightable and animatable neural avatar from sparse-view video.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Relightable and animatable neural avatar from sparse-view video

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.704528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.704528Z digest=sha256:aefd812ac7dd18db477b308c65f969724754dc2813af5bb2fc337683dcb6175b

Observation 3a52ecad-f94e-4650-b2d9-c9b2d96a5b26 · outbound

This paper cites 4k4d: Real-time 4d view synthesis at 4k resolution.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4k4d: Real-time 4d view synthesis at 4k resolution

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.264293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.708737Z digest=sha256:1f8e10bc5368bf714cdfffe40bed95688b90049dac71c6c4e6be5d39c68d4e61

Observation b428c9b2-58f8-413f-9c12-45a569c91d1e · outbound

This paper cites Representing long volumet- ric video with temporal gaussian hierarchy.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Representing long volumet- ric video with temporal gaussian hierarchy

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.252346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.712365Z digest=sha256:1a96887fb41491c3cfe0b4929be76ee6ad30f158e261093756dec8db5a529c5e

Observation eb6042ec-0c17-4911-83d9-1702e3ca5368 · outbound

This paper cites Diffusion2: Dynamic 3d content generation via score composition of orthogonal diffusion models.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Diffusion2: Dynamic 3d content generation via score composition of orthogonal diffusion models

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.240943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.715946Z digest=sha256:1632dd032894a59e33b108bc816ab1939e1873aec073ab00bfb5155487130750

Observation 185457a4-e58a-42e4-8b12-9ab04db6eff9 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.720259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.720259Z digest=sha256:5e5fe4fd8e8da20881dcf649eb0fafcc13a1295096c5ddb5ae48c7bc0b2b4d26

Observation 4ed2c0d3-3c54-404f-bce2-82d27be50d6f · outbound

This paper cites Real- time photorealistic dynamic scene representation and render- ing with 4d gaussian splatting.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Real- time photorealistic dynamic scene representation and render- ing with 4d gaussian splatting

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.226367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.725313Z digest=sha256:beab0d755fc2980ef296f96904516ec4ab9985092879bcc62a8c333c9093004f

Observation 68e6eebc-89f4-4db3-a0b3-ed439169d524 · outbound

This paper cites Mvsnet: Depth inference for unstructured multi-view stereo.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Mvsnet: Depth inference for unstructured multi-view stereo

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.212285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.728762Z digest=sha256:bdf9e046c95dfe3a9ab50a0191478f0bb2857f9268436b69df2494867dfcd106

Observation 3c1f007c-f168-4c06-9255-ed07bb53c4fd · outbound

This paper cites Recurrent mvsnet for high-resolution multi-view stereo depth inference.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Recurrent mvsnet for high-resolution multi-view stereo depth inference

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.200000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.732702Z digest=sha256:105d4d679dc7116f55fb08d7744159dd12da981e32089d68adfb79e9cd5bfb0c

Observation 88bc8735-bdfe-495f-9fb7-1a061878c8eb · outbound

This paper cites 4DGen: Grounded 4D Content Generation with Spatial-temporal Consistency.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4DGen: Grounded 4D Content Generation with Spatial-temporal Consistency

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.736671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.736671Z digest=sha256:bcba72b7397045f0e79157161f35d6a0701ae93b6e8b84c4d8852d7ecb1ad2c0

Observation 3ddff97b-ce40-4da3-a71a-cb32d626f213 · outbound

This paper cites Stag4d: Spatial-temporal anchored generative 4d gaussians.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Stag4d: Spatial-temporal anchored generative 4d gaussians

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.188208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.741102Z digest=sha256:5957685a6f6bd71060c9f2eeeb12a026a128ebb525e8d173688637da9f067192

Observation e2d7be86-4861-458a-85c0-c2dde556a466 · outbound

This paper cites 4diffusion: Multi-view video dif- fusion model for 4d generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models 4diffusion: Multi-view video dif- fusion model for 4d generation

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.175996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.745262Z digest=sha256:d25b38d02e8aba4b0913ff52cc9b239e4ec3f7bfca4b1e62b9d39fd16cb032ff

Observation 61ec5bc2-708c-44ce-b4ce-4c3fa398894a · outbound

This paper cites Cameras as Rays: Pose Estimation via Ray Diffusion.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Cameras as Rays: Pose Estimation via Ray Diffusion

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.749251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.749251Z digest=sha256:a9f0f33e2fe8b6c1b35d9c62f61aa6f992d9df8300e33003fd34f979226c902e

Observation f84d5aed-3a19-4720-abed-d554226d3f24 · outbound

This paper cites Animate124: Animating One Image to 4D Dynamic Scene.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Animate124: Animating One Image to 4D Dynamic Scene

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.753710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.753710Z digest=sha256:a399a0c2ec3d1c77f3e597dcd1d365fcc954e0d1bc8c31f1c6f56f49fa17f757

Observation 13578719-044e-47d8-8176-daf2a4d380f8 · outbound

This paper cites Bilateral refer- ence for high-resolution dichotomous image segmentation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Bilateral refer- ence for high-resolution dichotomous image segmentation

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:34.757492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:30:34.757492Z digest=sha256:99b70460cbc2fa181d4f003d7f6c0b1430426496c05a98edfdea9b51645ee900

Observation 3231362c-a61f-4779-aab8-491d4ca0f60e · outbound

This paper cites Gps- gaussian: Generalizable pixel-wise 3d gaussian splatting for real-time human novel view synthesis.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models Gps- gaussian: Generalizable pixel-wise 3d gaussian splatting for real-time human novel view synthesis

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.156419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.761142Z digest=sha256:bf3140c3af96b9c24e30ca7dd2f2a90fabaf9240f31da17338b5cad8f1229d32

Observation 57c8a991-f2bd-4308-b0ac-7061bb2d9260 · outbound

This paper cites A unified approach for text- and image-guided 4d scene generation.

Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models A unified approach for text- and image-guided 4d scene generation

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:35.140720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-06T16:30:34.764898Z digest=sha256:eb13c3765cac6c2bf0a01e89c249b7757b38c218bd194d81d241603087fb66ed

Pith citing papers

Observation 1cc16329-431c-4af8-8058-054d349c6e10 · inbound

Splatography: Sparse multi-view dynamic Gaussian Splatting for filmmaking challenges cites this paper.

Splatography: Sparse multi-view dynamic Gaussian Splatting for filmmaking challenges Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:25:32.934331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-18T00:22:55.503532Z digest=sha256:37a6002f08f899dbd9ae54f934d43c51035b4cc50136beb65719c9d56c230741

Observation 026f7edc-e9a8-445e-ad1e-15107d66761c · inbound

SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse Cameras cites this paper.

SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse Cameras Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:43:18.026626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-14T23:41:24.423859Z digest=sha256:541c27fd1000a3a7203cf6055b82e240d629ca3306228447e6a67bf9ccca1ead