Pith. sign in

Paper Citation Record · LEDGER

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation

As of 8 August 2026, this Paper Citation Record lists 95 of 95 outbound references and 3 inbound Pith citation observations for arXiv:2506.18839.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.18839 v1

Coverage vector

measured 95 of 95 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:50:23.911081Z

measured 98 of 98 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T06:54:11.233650Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:48:32.946157Z

Reference resolution

95 of 95 outbound references displayed

  • verified exact0
  • verified fuzzy58
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d83a5d1c-7035-4c75-98fb-5dbf7c4804e2 · outbound

This paper cites Controlling space and time with diffusion models.ICLR, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Controlling space and time with diffusion models.ICLR, 2025

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.570058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.570058Z digest=sha256:cbfeb48504e61381aecf0738ab740d38cc587ee1c3a8a8b1aaf4746f7895f977

Observation ae9424f3-dd9c-4fe7-a767-2a0790e7b326 · outbound

This paper cites Barron, and Aleksander Holynski.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Barron, and Aleksander Holynski

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.574568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.574568Z digest=sha256:dbe2d09051bf3594daaf24950bd7208e66c3501c2bdedd04663e60a065157b8d

Observation 96b5f06b-b1a8-4375-8783-e3eb3d5f9232 · outbound

This paper cites Genxd: Generating any 3d and 4d scenes.ICLR, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Genxd: Generating any 3d and 4d scenes.ICLR, 2025

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.579369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.579369Z digest=sha256:76027f2875f50ca1afb94ed83298341becd113864e8109de28179afccf8e9341

Observation 39d6a984-e7fe-4b97-aca5-b04bfa74189c · outbound

This paper cites Dimensionx: Create any 3d and 4d scenes from a single image with controllable video diffusion.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Dimensionx: Create any 3d and 4d scenes from a single image with controllable video diffusion

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.582571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.582571Z digest=sha256:e7b5fbc120c5c91aae87368edcf9847b7a26cffbdc717f879b72eeb5181106ad

Observation 1035cd84-dff6-4d65-985f-ea81a3c8ec38 · outbound

This paper cites Gen3c: 3d-informed world- consistent video generation with precise camera control.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Gen3c: 3d-informed world- consistent video generation with precise camera control

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.586936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.586936Z digest=sha256:ea8e3642ad8c26da2706ba9676e0273d207c35b01a94641406fd588d8c923a3f

Observation 7d883fac-9313-4588-ae5c-c8928f351cbc · outbound

This paper cites Trajectorycrafter: Redirecting camera trajectory for monocular videos via diffusion models, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Trajectorycrafter: Redirecting camera trajectory for monocular videos via diffusion models, 2025

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.591047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.591047Z digest=sha256:4bd8b3a9a9fd8ba359049b2a1c892812fa21365bdd3cd6aef2cfe4ff529d410a

Observation 48fa30c4-61f5-4f87-b06f-8cbd9a59056f · outbound

This paper cites Generative camera dolly: Extreme monocular dynamic novel view synthesis.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Generative camera dolly: Extreme monocular dynamic novel view synthesis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.594988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.594988Z digest=sha256:db168fba83dfaf0068850f09473b9859757a9a8b6cfce7e67c48c45079845e72

Observation 0520b25a-677a-4c3a-8eb2-95b3b04b1eba · outbound

This paper cites Recammaster: Camera-controlled generative rendering from a single video, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Recammaster: Camera-controlled generative rendering from a single video, 2025

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.598368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.598368Z digest=sha256:e660b0bab96b51bf281137d11d98e806046d7a4c2fab57810c6787f66ecde7d8

Observation 8f23826f-04f9-4281-8d06-d10d8fbdeaf1 · outbound

This paper cites Wetzstein.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Wetzstein

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.602196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.602196Z digest=sha256:6b99e5611cf0e8761f1ae6c3dd5aa4c4f1835fe782c1b2987472217068bf68b9

Observation 9d95e326-df56-433c-be0b-66ea76ac393c · outbound

This paper cites Human4dit: Free-view human video generation with 4d diffusion transformer.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Human4dit: Free-view human video generation with 4d diffusion transformer

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.605505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.605505Z digest=sha256:cc0fe8eefe586c9f01cdc0901327b8aca1c0b2cc695b52f1a066cb87c79dc577

Observation e5340efa-6d88-419b-a3e4-abe3a4dbb953 · outbound

This paper cites Vivid-zoo: Multi-view video generation with diffusion model, 2024.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Vivid-zoo: Multi-view video generation with diffusion model, 2024

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.609218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.609218Z digest=sha256:78de95444d54231859cbb0ec1dd9d3b967d0094ae96c8520b3db9868429ee6bc

Observation 853c4b84-080c-4854-ae55-3bbbbddf82c5 · outbound

This paper cites 4diffusion: Multi-view video diffusion model for 4d generation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation 4diffusion: Multi-view video diffusion model for 4d generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.612527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.612527Z digest=sha256:dac7c585b1ba0e25dd391b18039489aa6f0f781a514624792b1094304fde194b

Observation 10c2b3cd-b206-447f-bc61-4188719b4e9b · outbound

This paper cites Sv4d: Dynamic 3d content generation with multi-frame and multi-view consistency.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Sv4d: Dynamic 3d content generation with multi-frame and multi-view consistency

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.615985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.615985Z digest=sha256:00eac128d1f526d73ebe8c0854c5de1646b2f0ec8e69989b0f1ee60d37a84c92

Observation 27eea4e4-1d4c-4762-b76d-2ab56a3513e6 · outbound

This paper cites Syncammaster: Synchronizing multi-camera video generation from diverse viewpoints, 2024.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Syncammaster: Synchronizing multi-camera video generation from diverse viewpoints, 2024

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.619364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.619364Z digest=sha256:7bb8c14d6663a2eaceee5ce0770f00a6461be58ffc198a4c2cec1be3d14d500b

Observation 7f98586b-a49e-4e64-8c5e-18d674f52adc · outbound

This paper cites 4real-video: Learning generalizable photo-realistic 4d video diffusion, 2024.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation 4real-video: Learning generalizable photo-realistic 4d video diffusion, 2024

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.623512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.623512Z digest=sha256:18590c97df4f04c23d77033aea5d823221e0d9724f560edb77eb51c85bfbf714

Observation bcc02f03-ce18-4d57-97c3-6e1f2431e067 · outbound

This paper cites Cogvideox: Text-to-video diffusion models with an expert transformer.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Cogvideox: Text-to-video diffusion models with an expert transformer

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.620017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.627236Z digest=sha256:e9e9d39f1c8a166e0762831e40a331f74ad0026db7f5c7a366e5886135bfb7d1

Observation 45cb1825-5510-46db-8b66-2da8dd573f03 · outbound

This paper cites Wan: Open and advanced large-scale video generative models.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Wan: Open and advanced large-scale video generative models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.609459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.630596Z digest=sha256:11272c56d0e121c55d5745e63b332d8f87c05b1023b46e6cab4b8e8e88448ab6

Observation 7f01170e-046c-4bd0-931a-de3607e2ba3c · outbound

This paper cites Step-video-ti2v technical report: A state-of-the-art text-driven image-to-video generation model, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Step-video-ti2v technical report: A state-of-the-art text-driven image-to-video generation model, 2025

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.599456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.634087Z digest=sha256:835ee7946420d7140bfab4b051f5f48a772a73e3265eb1e1ec8f15128c27b2e8

Observation 71d157dc-b285-4454-bbba-a5de3571d2ac · outbound

This paper cites Sampson, Shikai Li, Simone Parmeggiani, Steve Fine, Tara Fowler, Vladan Petrovic, and Yuming Du.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Sampson, Shikai Li, Simone Parmeggiani, Steve Fine, Tara Fowler, Vladan Petrovic, and Yuming Du

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.590637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.637656Z digest=sha256:8cad4f1cbf475f2a925b2cd1c40977d9dc91468bfa58d08b51a692781a3a7ddb

Observation 05b83f60-3b6e-4775-9db9-780101cc464c · outbound

This paper cites an unresolved cited work.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:50:24.580460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.641024Z digest=sha256:11957d7b1ea5bbb247c75eb0abd11f9e2c5c850253836d9c0d97353d9bbe1898

Observation e62f7e6a-f2b6-4a6d-9817-aa9a3cd1c0fc · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view synthesis.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Nerf: Representing scenes as neural radiance fields for view synthesis

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.572170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.644912Z digest=sha256:534c695242722ca13ae8d9b7aeec664402f0e07188ea0a2c8e3492a490f6a3e4

Observation 4dee6952-4f20-45e5-a9c9-a798317ffa4a · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.ToG, 2023.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation 3d gaussian splatting for real-time radiance field rendering.ToG, 2023

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.562267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.648964Z digest=sha256:14d87767ea2752bafb6036fca27f914adaaf900fdabe3a5a013747e4474b92d4

Observation 1c29e439-43b1-4c70-b636-f1d5cc0ce6d1 · outbound

This paper cites Vggt: Visual geometry grounded transformer.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Vggt: Visual geometry grounded transformer

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.652933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.652933Z digest=sha256:bba73db604dcbc5a982f6e13a25bd1590b0e6b9e2b326f5649d148ef9e72976a

Observation cb049e13-3f70-4880-a03f-bcf89e9f1a5b · outbound

This paper cites Dreamfusion: Text-to-3d using 2d diffusion.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Dreamfusion: Text-to-3d using 2d diffusion

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.657140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.657140Z digest=sha256:6b62d2cc63a3757c7e09c301d26a19ff4ef4d6e1a6313eb0050ff9745a288bf4

Observation 5cb3ad7d-2367-4fad-8a1d-c691c596b3ce · outbound

This paper cites Pro- lificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Pro- lificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.661246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.661246Z digest=sha256:c35d75df40b01f52fe144cf46784402363567e7f92f6312ab55bb897fc6e338b

Observation 938c6441-a563-4c68-9f31-abc77aa45f02 · outbound

This paper cites Score jacobian chaining: Lifting pretrained 2d diffusion models for 3d generation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Score jacobian chaining: Lifting pretrained 2d diffusion models for 3d generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.665412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.665412Z digest=sha256:f36fc401300ec34a9665c05e0d1ad0636721ff3f979613c79b0376701b3e1c38

Observation c1473008-793b-4693-bfaa-e70e428d37a0 · outbound

This paper cites Fantasia3d: Disentangling geometry and appearance for high-quality text-to-3d content creation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Fantasia3d: Disentangling geometry and appearance for high-quality text-to-3d content creation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.669068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.669068Z digest=sha256:4cdfaa1dab52fa126434d4b03bc468d70698547441b0279076477948ba59cf6a

Observation 8783decc-7a9d-49c6-b26d-f6d16f3b8524 · outbound

This paper cites Magic3d: High-resolution text-to-3d content creation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Magic3d: High-resolution text-to-3d content creation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.672950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.672950Z digest=sha256:e68c45664717dceab5bbf18f12754fe21fb7babece95ba69e8030237af1b3574

Observation 40e0bad9-e968-4204-acf2-7bd07d9a100f · outbound

This paper cites Hifa: High-fidelity text-to-3d with advanced diffusion guidance.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Hifa: High-fidelity text-to-3d with advanced diffusion guidance

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.522558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.676972Z digest=sha256:40ab22d2af6c1ff0219d29a4ad8d464dd55066dc60238020df61368a25a86e6c

Observation a5e1badd-d05b-4a3d-bd67-a716b092bc83 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation High-resolution image synthesis with latent diffusion models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.681360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.681360Z digest=sha256:4e14f6cd695eebac3eea66d904ed24e54ec117fedad25be56db2b9cfe181f148

Observation d91a71d9-94a8-4db4-96e8-fc012c271e90 · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.NeurIPS, 2022.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Photorealistic text-to-image diffusion models with deep language understanding.NeurIPS, 2022

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.684804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.684804Z digest=sha256:067f4a7421043944195166d7a37765aa157296eda48b22e25b35de7d08a71e0a

Observation 5a1cb132-5d9b-40da-8347-1ea55b9348f7 · outbound

This paper cites Zero-1-to-3: Zero-shot one image to 3d object.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Zero-1-to-3: Zero-shot one image to 3d object

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.688142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.688142Z digest=sha256:9f762d02114478430561fe11ffd3ef7a71f2930931dd430c1fbf3c8d93571bbc

Observation b2b945ad-7f0c-4395-9d8b-88996b2bf591 · outbound

This paper cites Mvdream: Multi- view diffusion for 3d generation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Mvdream: Multi- view diffusion for 3d generation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.499067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.691396Z digest=sha256:1642be569ad0179c7108700a6ea07ae98e0dfb134fd266d91aa36f131a9da257

Observation 4f031122-fcea-43c3-ac72-a77dc54ae1e1 · outbound

This paper cites 4d-fy: Text-to-4d generation using hybrid score distillation sampling.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation 4d-fy: Text-to-4d generation using hybrid score distillation sampling

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.490096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.694956Z digest=sha256:127487284f52d9af17ebcfd1a8a8187d4fde2bad241fc637575bef031326e118

Observation ca3b555c-09e9-4939-b7bf-5f7af8bc7b4d · outbound

This paper cites Align your gaussians: Text-to-4d with dynamic 3d gaussians and composed diffusion models.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Align your gaussians: Text-to-4d with dynamic 3d gaussians and composed diffusion models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.480993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.698796Z digest=sha256:c1724ea2cdc51a3100f062de14911a7d2ffbabc5e4e5e7d88d0ff1741a5f676e

Observation c015f36a-976b-473b-ab03-17f6bb246565 · outbound

This paper cites Consistent4d: Consistent 360 {\deg}dynamic object generation from monocular video.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Consistent4d: Consistent 360 {\deg}dynamic object generation from monocular video

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.471590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.703148Z digest=sha256:4ddaa751568af72f4fea3995908bd6bff9bd3e30c81ab8c3b266d738ba7d7412

Observation b0ffa42d-cb11-4a4c-b529-75396b367e1d · outbound

This paper cites Dreamgaussian4d: Generative 4d gaussian splatting.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Dreamgaussian4d: Generative 4d gaussian splatting

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.462541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.706973Z digest=sha256:3ff5bfba7d4235e282059ac9c835e3d4c5196582ac012b1ddfd49c8c9052d0c3

Observation 807c601c-3ec9-4e15-86d1-7ce2d53f424e · outbound

This paper cites 4dgen: Grounded 4d content generation with spatial-temporal consistency.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation 4dgen: Grounded 4d content generation with spatial-temporal consistency

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.711037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.711037Z digest=sha256:2e5fa92679879a9eb003b1a329fafc09081cab67ef471761e43d56d059f332e9

Observation f580c5d4-d82d-4260-a7c2-20e592ae8f51 · outbound

This paper cites Ani- mate124: Animating one image to 4d dynamic scene.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Ani- mate124: Animating one image to 4d dynamic scene

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.448324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.714639Z digest=sha256:92c92211abbe4638c1cdffcc432752cae58309614b9b9e50d44da5ce3fcb6c33

Observation c72ccb59-d113-402d-8d20-c8b10e4dd2d7 · outbound

This paper cites Text-to-4d dynamic scene generation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Text-to-4d dynamic scene generation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.437607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.718436Z digest=sha256:22753178ddeb63d1d17a3fddd2516f0290bc4ae17cdd319705f286711ed23b5c

Observation ad48c7f8-51f0-4eb9-88d9-0f4373148984 · outbound

This paper cites 4real: Towards photorealistic 4d scene generation via video diffusion models.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation 4real: Towards photorealistic 4d scene generation via video diffusion models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.427298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.721479Z digest=sha256:72fc318c368c48e7ca079530e04b0bfc8b3509455898db7220e404c5fc77edbe

Observation 844cb7b2-df56-4310-bfc6-97d3e4fc143a · outbound

This paper cites Animatediff: Animate your personalized text-to-image diffusion models without specific tuning.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Animatediff: Animate your personalized text-to-image diffusion models without specific tuning

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.418408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.725224Z digest=sha256:a64c3731ee5fca0405a6456d3409a20ee122afcc89eea350e31cbdbab4b905ac

Observation 2d5b7943-81f4-4b80-afd2-5297ce19e875 · outbound

This paper cites Imagen video: High definition video generation with diffusion models.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Imagen video: High definition video generation with diffusion models

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.409426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.728387Z digest=sha256:61094822b579b0ce92d826c1e1f17afe1be737cbf9d63b8ffe12c11d0d8b38ad

Observation 8e41b232-7088-4e89-946a-caa024cdcf53 · outbound

This paper cites Videocrafter1: Open diffusion models for high-quality video generation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Videocrafter1: Open diffusion models for high-quality video generation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.400272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.732250Z digest=sha256:6fa296de3a4a6f41b1dc5a2f3ecfcb205942117df840490e7dda11a0e2326548

Observation 7f18dc49-9abb-4733-a927-348c0d3e2f7b · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Objaverse: A universe of annotated 3d objects

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.735576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.735576Z digest=sha256:285e0ab45d9125f9b3bf5293434c852a0f82084f9de57297f40adfc4f6f15d6b

Observation 5edbd2aa-2825-4420-96b0-32875371e2bc · outbound

This paper cites Video generation models as world simulators, 2024.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Video generation models as world simulators, 2024

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.738569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.738569Z digest=sha256:fbc8454fa4c8073ad36790b4783f61698e6425c4912f23b6c44b016b701e561d

Observation 3ba617db-3f6d-4817-9149-f8c92381408d · outbound

This paper cites Snap video: Scaled spatiotemporal transformers for text-to-video synthesis.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Snap video: Scaled spatiotemporal transformers for text-to-video synthesis

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.381186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.741499Z digest=sha256:c72453138816d5a00e652ce25a55c00a19595dc61dc59dda7d5c2347c4bf376e

Observation 7f3351a3-900a-4256-8d3e-862b9cfa8b50 · outbound

This paper cites Vd3d: Taming large video diffusion transformers for 3d camera control.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Vd3d: Taming large video diffusion transformers for 3d camera control

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.372220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.744383Z digest=sha256:889d46e6707fc8b463158732bc952a113b39686a8ed1971fd857db88b48ee50c

Observation b365e525-4771-4899-8ee2-eeaffad7bda8 · outbound

This paper cites Motionctrl: A unified and flexible motion controller for video generation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Motionctrl: A unified and flexible motion controller for video generation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.362670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.748477Z digest=sha256:7d605aba99a9d30ab8a7f72bff42f0b01ab012e18a359c88f0084516a771cacd

Observation 671fb44e-c1fe-43e8-9293-ff2ccb613427 · outbound

This paper cites Cameractrl: Enabling camera control for text-to-video generation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Cameractrl: Enabling camera control for text-to-video generation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.354139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.751759Z digest=sha256:2fcd7cdb7d90daffeec35fe7135843d1a96bfb0e810480e6575001c7535ceced

Observation ff44c7b2-3b0a-4bf5-8216-159eb39a8e47 · outbound

This paper cites Direct-a-video: Customized video generation with user-directed camera movement and object motion.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Direct-a-video: Customized video generation with user-directed camera movement and object motion

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.345593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.755463Z digest=sha256:4751fe9c8344f4ac8683165a2ed041d15f161bb5bffdc11d602bfbf625c13c44

Observation ccfd719a-a69d-498b-afd0-1f2a115c2dd1 · outbound

This paper cites Camco: Camera-controllable 3d-consistent image-to-video generation.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Camco: Camera-controllable 3d-consistent image-to-video generation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.335918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.758985Z digest=sha256:1ec93eab80e4b8df3e38e0fbdf80ce8a623b84098a8dfbdb6d905bd61d82fccc

Observation f757fea2-5703-4e1b-a5df-3370b917cc6f · outbound

This paper cites Freevs: Generative view synthesis on free driving trajectory.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Freevs: Generative view synthesis on free driving trajectory

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.327304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.762933Z digest=sha256:ec81cf23dd7d500ec231fbe3861403c55a76392e0ad79d9a82c641c1aa5e17d2

Observation 3bd9309a-2eff-42e0-bf00-27b54ffc2154 · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Frozen in time: A joint video and image encoder for end-to-end retrieval

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.766509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.766509Z digest=sha256:ed31a53e3e135b4d9e4fcbe9ec366e7814856f6a4fbf1f6f4a8b2b97eaefdcac

Observation 6685641e-fb72-4e78-a55c-01d1db995f8f · outbound

This paper cites Stereo magnification: Learning view synthesis using multiplane images.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Stereo magnification: Learning view synthesis using multiplane images

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.313751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.770040Z digest=sha256:e3d6912dc8c2dbeccb13f1cfa59fa4526b625009bbdaeffbe1df286b0dc093e5

Observation 04634c8b-86a5-42af-b5c7-63ecf2845064 · outbound

This paper cites Instantsplat: Sparse-view gaussian splatting in seconds, 2024.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Instantsplat: Sparse-view gaussian splatting in seconds, 2024

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.303502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.774059Z digest=sha256:36a7963e810ad76982c722bee58df0f9b5f5f23d4d4eff55ffaa9fd1beee8c98

Observation e1ca18a6-3eac-415a-be4e-11a7f6c1164a · outbound

This paper cites Putting nerf on a diet: Semantically consistent few-shot view synthesis.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Putting nerf on a diet: Semantically consistent few-shot view synthesis

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.293134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.777563Z digest=sha256:16159c949431523a4c5f5bd038fea93d8d32e2b148a903573dea3087c9020689

Observation 59d5383c-86d3-489a-8354-7b963f800b0e · outbound

This paper cites G3r: Gradient guided generalizable reconstruction.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation G3r: Gradient guided generalizable reconstruction

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.283957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.781352Z digest=sha256:5ad13b46ef2b189cc2eaeed00241eb72fc81f4418a7979465a525751f2421d83

Observation 0d516c04-0831-4067-8eac-c50b173e8dbe · outbound

This paper cites MonoNeRF: Learning a generalizable dynamic radiance field from monocular videos.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation MonoNeRF: Learning a generalizable dynamic radiance field from monocular videos

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.274579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.785199Z digest=sha256:901ac3277e6fb8b6cc4a134dbf6118a6e625b8b72eb49edd66553e4b842b8273

Observation 393a2ed3-5194-4013-8639-538cf5e8d523 · outbound

This paper cites Enhancing neRF akin to enhancing LLMs: Generalizable neRF transformer with mixture-of-view-experts.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Enhancing neRF akin to enhancing LLMs: Generalizable neRF transformer with mixture-of-view-experts

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.265358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.789136Z digest=sha256:78da44dbbebb2b4e41c68496786e7d1f8988c4959b2ab46f8269ad602789a66e

Observation ba2cb804-be68-4957-b800-c9adadf5a64e · outbound

This paper cites Is attention all that neRF needs? InICLR, 2023.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Is attention all that neRF needs? InICLR, 2023

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.256083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.792598Z digest=sha256:b3cfd9458d3ea0b2ac620258623063228ac9b4ae7b4394dc68867ff57a7c150b

Observation e788f295-de1b-4187-a96f-d660d1556101 · outbound

This paper cites Lrm: Large reconstruction model for single image to 3d.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Lrm: Large reconstruction model for single image to 3d

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.245531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.796652Z digest=sha256:dd1de6f96b84817477bd182ec331a4a8452ae93fe9b480c63c399e469149740f

Observation 396fceb1-b23a-4680-83a8-670d5aa223e1 · outbound

This paper cites pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation pixelsplat: 3d gaussian splats from image pairs for scalable generalizable 3d reconstruction

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.800523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.800523Z digest=sha256:794ebb986bd5ec16b9e22c49037026b16a24ee2059350a925452f410a4d6c326

Observation 0891cf3f-1f9d-4bdf-b3a2-b27c98d91e18 · outbound

This paper cites Mvsplat: Efficient 3d gaussian splatting from sparse multi-view images.ECCV, 2024.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Mvsplat: Efficient 3d gaussian splatting from sparse multi-view images.ECCV, 2024

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.230300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.803999Z digest=sha256:ebe8ee302995c0374947edc69f8453eb0be8808ef053f76f1301b8c6ce09cb6f

Observation 6701339f-c233-47ca-9d27-5dc0b2805224 · outbound

This paper cites Mvsplat360: Feed-forward 360 scene synthesis from sparse views.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Mvsplat360: Feed-forward 360 scene synthesis from sparse views

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.807113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.807113Z digest=sha256:1b1a864f96d9dc070715608586a96a735fd66307d96260b5ac5dc8a95057649e

Observation f264ddc9-bcd9-4f8e-996d-24121164bf99 · outbound

This paper cites Gs-lrm: Large reconstruction model for 3d gaussian splatting.ECCV, 2024.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Gs-lrm: Large reconstruction model for 3d gaussian splatting.ECCV, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.215140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.810632Z digest=sha256:9ed87b911946d88ad9f604851a5875c3c26ad4953a8d7d15798b9e5e984f2ec4

Observation d78589ec-b1e4-4d3a-8266-3991496a7271 · outbound

This paper cites Scube: Instant large-scale scene reconstruction using voxsplats.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Scube: Instant large-scale scene reconstruction using voxsplats

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.206565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.814457Z digest=sha256:40856321b7cd36d867c1ab08358e323eb8e64df58160840564518b05865cb004

Observation 9a120bfa-962a-47b6-a71a-e6ef6f4923d0 · outbound

This paper cites Lvsm: A large view synthesis model with minimal 3d inductive bias.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Lvsm: A large view synthesis model with minimal 3d inductive bias

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.196276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.817822Z digest=sha256:495f2c0137c6bdca00cae24cf646b54b872f06810648c185d24c9d73ffa6eb79

Observation f09b4afb-a913-4566-8e56-141ba62157c6 · outbound

This paper cites Rayzer: A self-supervised large view synthesis model.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Rayzer: A self-supervised large view synthesis model

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.188005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.821231Z digest=sha256:b1cbdfc41ecfe713c75b580cb9855e3fd3c3be37bd0336087338e7511cba03b2

Observation 282faeee-b9c7-4633-93ca-9f696fb98b5d · outbound

This paper cites Splatt3r: Zero-shot gaussian splatting from uncalibrated image pairs.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Splatt3r: Zero-shot gaussian splatting from uncalibrated image pairs

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.178642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.824569Z digest=sha256:d129768fc218b00336ca53377c7b108e23d3228dc888744374ec69da42c5a305

Observation 286d7cca-fe03-456d-a60f-38ee63287c63 · outbound

This paper cites Flare: Feed-forward geometry, appearance and camera estimation from uncalibrated sparse views.CVPR, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Flare: Feed-forward geometry, appearance and camera estimation from uncalibrated sparse views.CVPR, 2025

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.168548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.828387Z digest=sha256:82cde8670328220129b29f122981b7ad9ab3ac15b252f03a0e07708c1639abed

Observation 7b44d003-86be-4a92-9832-f3fe78bd9189 · outbound

This paper cites No pose, no problem: Surprisingly simple 3d gaussian splats from sparse unposed images.ICLR, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation No pose, no problem: Surprisingly simple 3d gaussian splats from sparse unposed images.ICLR, 2025

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.158491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.831420Z digest=sha256:bdb8f6d9c5e72071ff1f09333863b1aa88a088ed6e8b63fb4f0a4ad0df4788c6

Observation 292d743f-ca24-4e8f-9d70-bc5f88f55df8 · outbound

This paper cites Pf3plat: Pose-free feed-forward 3d gaussian splatting.ICML, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Pf3plat: Pose-free feed-forward 3d gaussian splatting.ICML, 2025

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.148307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.834350Z digest=sha256:a2f8bb29c5b4a39338ef8e38f74b88add2b848b804b7ebf6b36bf1d9762b7114

Observation 5fc49928-cb37-4c43-8a39-00ed473718e0 · outbound

This paper cites Susskind, and Alexander G.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Susskind, and Alexander G

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.138385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.837357Z digest=sha256:1b1d9cd5c284fa40b2681712a3eb544d7af69e621e7b4d2ee6d17b1465d3fdd0

Observation 7677d66f-752d-4a87-81cf-0a6fb2c3e3e3 · outbound

This paper cites Monst3r: A simple approach for estimating geometry in the presence of motion.ICLR, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Monst3r: A simple approach for estimating geometry in the presence of motion.ICLR, 2025

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.129088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.840169Z digest=sha256:c949a8ad1d6a1b54e5360439961eeb13cbaae47cc04218b5f06885ce9eae9096

Observation dd87c8b4-6514-44e2-8d9a-2586c0a325ef · outbound

This paper cites Dust3r: Geometric 3d vision made easy.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Dust3r: Geometric 3d vision made easy

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.843011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.843011Z digest=sha256:6c3b51b944ad80715960fff3c2f08a74c489ce30c74e652087d3218fb83ca35c

Observation a03f0d00-bdec-4e85-83b8-b91682087dd5 · outbound

This paper cites Grounding image matching in 3d with mast3r, 2024.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Grounding image matching in 3d with mast3r, 2024

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.846627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.846627Z digest=sha256:7a8ea8385a93bc01c6aad1150e817e00e7a6d71c0b18a037035c4b72c6f91727

Observation 139fdb88-0697-4c5c-8ae8-2d22a7e2dd38 · outbound

This paper cites L4gm: Large 4d gaussian reconstruction model.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation L4gm: Large 4d gaussian reconstruction model

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.107515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.849817Z digest=sha256:8562fa3cfc34e0e54f0d841385cad4c3972bf11ac30c202d7cef418077cbb921

Observation d008ac3b-befa-430a-ac3c-e8b1c5ebf8e7 · outbound

This paper cites Feed-forward bullet- time reconstruction of dynamic scenes from monocular videos.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Feed-forward bullet- time reconstruction of dynamic scenes from monocular videos

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.097579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.852934Z digest=sha256:f0e08965ff57ee900d7cc4d2d9783636b0cbee64413622d144f74a38d5e73b85

Observation 66ba1738-44de-49a9-bae0-ee1cb7752108 · outbound

This paper cites Scalable diffusion models with transformers.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Scalable diffusion models with transformers

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.088506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.856515Z digest=sha256:b86868fffe7dfc0fec5252b37c5f1bcf29d0baf594d3ad240df4a6b5932983d4

Observation 616dbc40-57ec-47aa-ae62-60e55a1da6b9 · outbound

This paper cites Cosmos world foundation model platform for physical ai, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Cosmos world foundation model platform for physical ai, 2025

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.079666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.859564Z digest=sha256:9757f33941da313fa3a11d8d1278f0a5174fefd6e4490a274f906a4f47589964

Observation 665d75be-f46c-479a-95b9-6f034c843508 · outbound

This paper cites Sv4d 2.0: Enhancing spatio-temporal consistency in multi-view video diffusion for high-quality 4d gener- ation, 2025.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Sv4d 2.0: Enhancing spatio-temporal consistency in multi-view video diffusion for high-quality 4d gener- ation, 2025

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.069200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.862630Z digest=sha256:fc8b8c2efb554fcdc74f0182be197d6d54adba7f617f131f98896631db67c929

Observation f651e6f6-0bbe-4800-818c-ec068ae5af5b · outbound

This paper cites Flex attention: A programming model for generating optimized attention kernels, 2024.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Flex attention: A programming model for generating optimized attention kernels, 2024

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.866147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.866147Z digest=sha256:abb25d5bad8401734ed0e10111d72e723a1df49736987c52fd85c047bdfa730e

Observation c33a9a83-d002-4aaf-8586-8311feeb9050 · outbound

This paper cites Flow straight and fast: Learning to generate and transfer data with rectified flow.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Flow straight and fast: Learning to generate and transfer data with rectified flow

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.052940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.869966Z digest=sha256:00a6da201ae2995a7b58804375c3de0cd9a070cafcf59540936789f36ab6ad1c

Observation c81b07fc-eb6e-40d9-a5f4-1ae72b828243 · outbound

This paper cites an unresolved cited work.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:50:24.042911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.874407Z digest=sha256:a59e18377a3411470a130033364fac97a326a71baeba45597e288f534318b57c

Observation 9e6d9a1e-b6ce-4997-8df1-62ee68f1a52c · outbound

This paper cites Novel view synthesis of dynamic scenes with globally coherent depths from a monocular camera.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Novel view synthesis of dynamic scenes with globally coherent depths from a monocular camera

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.877689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.877689Z digest=sha256:35567d719098a49d7146faab3ab34da5fa0d5ad23e2d91ee57737e576c813a0e

Observation 7d589cd8-80db-4734-a193-ef5d931e1fc2 · outbound

This paper cites The unreason- able effectiveness of deep features as a perceptual metric.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation The unreason- able effectiveness of deep features as a perceptual metric

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T23:50:23.881503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:50:23.881503Z digest=sha256:b98377711f3032b2aac177d41cc69752543424e65fb15a2954272dd536f93600

Observation c03e6d85-9d73-433a-98ae-1070b2da8744 · outbound

This paper cites Dl3dv-10k: A large-scale scene dataset for deep learning-based 3d vision.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Dl3dv-10k: A large-scale scene dataset for deep learning-based 3d vision

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.018348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.885477Z digest=sha256:8e6582c5afd6a0826a6f0fa4139197e25d828d94ca1f8572df75c0e270589b49

Observation cb319de6-14e3-4966-b3d6-3493797bd69b · outbound

This paper cites Mvimgnet: A large-scale dataset of multi-view images.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Mvimgnet: A large-scale dataset of multi-view images

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:24.006724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.889573Z digest=sha256:ec37241723087cdcc16c064f2ff7e7bf28d4e6c76e7a64bde968a19d8e10104c

Observation 55b52fa2-3894-41fc-84fb-fdc31089851c · outbound

This paper cites Infinite nature: Perpetual view generation of natural scenes from a single image.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Infinite nature: Perpetual view generation of natural scenes from a single image

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:23.995769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.893578Z digest=sha256:d6334a1f000db4f8c3e39a46dd352d59680197db1e5fefbec60cbd556846501d

Observation 1a23ae38-515b-4a11-a353-dfe323cde999 · outbound

This paper cites Vbench++: Comprehensive and versatile benchmark suite for video generative models.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Vbench++: Comprehensive and versatile benchmark suite for video generative models

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:23.984476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.897609Z digest=sha256:b80bd82fade9323da8c611dde50ad0dce0967000d0abddf99de01ad168931d1b

Observation 080817de-0c0b-4f07-836f-00a72ddd011f · outbound

This paper cites Met3r: Measuring multi-view consistency in generated images.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Met3r: Measuring multi-view consistency in generated images

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:23.974625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.901370Z digest=sha256:3e5028fa4845d9aa3c8976e8e3e168314c3086ecd1c2f709af0596bf77518651

Observation 5d41ebdd-d15f-4fb3-aca5-b4c79007aa84 · outbound

This paper cites Tanks and temples: Bench- marking large-scale scene reconstruction.ToG, 2017.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Tanks and temples: Bench- marking large-scale scene reconstruction.ToG, 2017

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:23.965374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.904561Z digest=sha256:ec904b022ec49c76873e96c067ec8c7b0df54df633465b07aec00f0859e1948c

Observation 3eec1823-de78-46f3-895e-64ba1e5b460b · outbound

This paper cites Srinivasan, Rodrigo Ortiz-Cayon, Nima Khademi Kalantari, Ravi Ramamoorthi, Ren Ng, and Abhishek Kar.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Srinivasan, Rodrigo Ortiz-Cayon, Nima Khademi Kalantari, Ravi Ramamoorthi, Ren Ng, and Abhishek Kar

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:23.955740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.907620Z digest=sha256:c5c28d60b0bf0c0b75e52f39ff3950f1d664a2db71a0ef770644e52094ef7eb3

Observation 19ec201f-06d1-43ee-93fa-d61da4095288 · outbound

This paper cites Neural 3d video synthesis from multi-view video.

4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation Neural 3d video synthesis from multi-view video

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:50:23.945498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:50:23.911081Z digest=sha256:5386d325cea0da5d1a2874d333dbdbab2648bdbd9febd74e45b4b4c79a056801

Pith citing papers

Observation 2988db5a-305d-4c5b-b4af-93b4c7720bee · inbound

SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse Cameras cites this paper.

SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse Cameras 4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:43:18.002631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T23:41:24.423859Z digest=sha256:a850e9226fc624bf5d0175267326639ab670cc7a3565771f7fff3ad8c0dc83b5

Observation 3d140238-b633-4b88-8df6-c85b0968aa18 · inbound

Feed-Forward 3D Scene Modeling: A Problem-Driven Perspective cites this paper.

Feed-Forward 3D Scene Modeling: A Problem-Driven Perspective 4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation

Reference 193

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:20:26.111329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T13:13:42.052386Z digest=sha256:8e6a2526d36ced0b82cd82890c8d87592dac0ca877ed667ad7217c4b7fa58a43

Observation f4926fba-b7ec-4a2d-a519-5dad5811ae2c · inbound

Flex4DHuman: Flexible Multi-view Video Diffusion for 4D Human Reconstruction cites this paper.

Flex4DHuman: Flexible Multi-view Video Diffusion for 4D Human Reconstruction 4Real-Video-V2: Fused View-Time Attention and Feedforward Reconstruction for 4D Scene Generation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:48:32.947695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T06:54:11.233650Z digest=sha256:e79a829e1f2df2d45617a40e07ee506b82150f8fe516f30272cf9180acb532ab