Pith. sign in

Paper Citation Record · LEDGER

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

As of 21 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 15 inbound Pith citation observations for arXiv:2506.18903.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.18903 v3

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:44:55.425136Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T05:06:51.863744Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T17:58:47.750659Z

Reference resolution

48 of 48 outbound references displayed

  • verified exact0
  • verified fuzzy26
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fd6c16f4-997a-4808-b95f-a6bca4642e77 · outbound

This paper cites SF3D: Stable Fast 3D Mesh Reconstruction with UV-unwrapping and Illumination Disentanglement.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory SF3D: Stable Fast 3D Mesh Reconstruction with UV-unwrapping and Illumination Disentanglement

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.208668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.208668Z digest=sha256:d5b7aaa53c889d509ae0508bc9c97aa516f7a5150087232c36c021ef30b82c99

Observation add98839-8a72-4a22-9955-ec86ad76498c · outbound

This paper cites Ge- nie: Generative interactive environments.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Ge- nie: Generative interactive environments

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.214283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.214283Z digest=sha256:2d7f752ad6621584b454d5200a801a50a06f6de567e92d493aaa0aa2430f3f15

Observation 77b16276-def5-471b-b7a1-1c9508db77af · outbound

This paper cites Mvsnerf: Fast general- izable radiance field reconstruction from multi-view stereo.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Mvsnerf: Fast general- izable radiance field reconstruction from multi-view stereo

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:56.083602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.219075Z digest=sha256:67faf7894299783a53e76474c04cc06efc291aad69dcac3c3ec5a5046fb184fb

Observation c6664651-6cd3-49d8-9f1f-a86eb188dc0c · outbound

This paper cites Single-Stage Diffusion NeRF: A Unified Approach to 3D Generation and Reconstruction,.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Single-Stage Diffusion NeRF: A Unified Approach to 3D Generation and Reconstruction,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:56.068808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.225074Z digest=sha256:82f28269dfc0bacf86382296b22b2f2e1d56c0220b6ccf9c1ee66d556e189a30

Observation 5f623029-d7ba-4e32-83dc-5c66ddb746a6 · outbound

This paper cites MVSplat: Efficient 3D Gaussian Splatting from Sparse Multi-View Images.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory MVSplat: Efficient 3D Gaussian Splatting from Sparse Multi-View Images

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.234369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.234369Z digest=sha256:539c03ed564af6040d3773164c8ce0fefe1ec1544026db0ac1ceb98afd04cf99

Observation 2443616f-1e82-4804-9408-9f3a9f1f7e00 · outbound

This paper cites Mvsplat360: Feed-forward 360 scene synthesis from sparse views.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Mvsplat360: Feed-forward 360 scene synthesis from sparse views

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.239841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.239841Z digest=sha256:c01ab11feaad6c533c8bfc12e4e6a3720f527067d28061a5feaa8adf61e4e845

Observation 7ce0cb6b-711b-4c85-a335-48872021ae75 · outbound

This paper cites SceneScape: Text-Driven Consistent Scene Generation.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory SceneScape: Text-Driven Consistent Scene Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.244278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.244278Z digest=sha256:524afe1e4556bee6d25013b65086c90d95f894a3353c4894f2bfedce4a2c026a

Observation 29a4e433-366f-4979-9be6-4e0873d68e30 · outbound

This paper cites Srinivasan, Jonathan T.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Srinivasan, Jonathan T

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:56.045240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.248946Z digest=sha256:f06c66dfcff04cbb1ac2f3ed3267e31a1fa39b879bf4a4a2fbeba7fdd74c5ae1

Observation b8b79e16-57b8-43f2-8fa6-63b001bdce6f · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.253743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.253743Z digest=sha256:63643567c39099546befc1baa26251baa9e5e6c694737021bd7f5aff44a909f3

Observation 8b3104ac-bd76-4f68-a3cb-b9466745bf82 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Gans trained by a two time-scale update rule converge to a local nash equilib- rium

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.258500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.258500Z digest=sha256:3bb90fa563324fc52b34035e52b6439ab175dd2ef3077c64413a4980c57d836b

Observation 97a3b91b-bbef-4ea5-9440-dea261f7a1f6 · outbound

This paper cites Text2room: Extracting textured 3d meshes from 2d text-to-image models.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Text2room: Extracting textured 3d meshes from 2d text-to-image models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:56.022683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.263360Z digest=sha256:de8fc5487106ca1b8676d18ea92396c28427d38678a5a54f8681cc7d264da93e

Observation a96b9ace-a67e-419f-9a15-1e9e4497eaf8 · outbound

This paper cites LRM: Large reconstruction model for single image to 3D.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory LRM: Large reconstruction model for single image to 3D

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:56.008118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.267806Z digest=sha256:660a448376aa4dde7617c339f79fa26f809402285f94043ebe95a0eea52cfe57

Observation 12f4eec1-c7d0-4b51-98c8-b948f44b65a8 · outbound

This paper cites Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen- Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen- Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.993610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.272032Z digest=sha256:4b06bcc21cd08b6c8a06ac7f31450f9e4c1b1ef20d002a0cbfcba65d2a6d4d2f

Observation 3fa5edc2-a6f7-4b3e-82f1-f33edd827b13 · outbound

This paper cites Tanks and temples: Benchmarking large-scale scene reconstruction.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Tanks and temples: Benchmarking large-scale scene reconstruction

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.276226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.276226Z digest=sha256:d0a18e77badaa4c869dc256eb222277fb66a76e197dcf67009dcd270491f4765

Observation 76bf3b31-54ff-4263-a706-96c6e567517a · outbound

This paper cites Simple and effective synthesis of in- door 3d scenes.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Simple and effective synthesis of in- door 3d scenes

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.970046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.280486Z digest=sha256:38dc8d7588c654f4129f566dc8a77b213f7ef5d7b6ef9ce096ffe1ddd650a372

Observation 522ef07d-661d-470f-82bd-26074ea1104b · outbound

This paper cites Infinite na- ture: Perpetual view generation of natural scenes from a sin- gle image.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Infinite na- ture: Perpetual view generation of natural scenes from a sin- gle image

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.955697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.284750Z digest=sha256:1e9440b2b20a460735be9e82eca9f8e53738e141fe755d813b7f13f644d7884f

Observation 3f2d1189-5ab2-4e4d-bebc-a563a1538555 · outbound

This paper cites RealFusion: 360 reconstruction of any ob- ject from a single image.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory RealFusion: 360 reconstruction of any ob- ject from a single image

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.941018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.289240Z digest=sha256:13c84fdb5026bb3918ffae48713c1b4cefdc72a64c5de17a8582731563fe9308

Observation cd710309-14a7-4aa2-ada5-63548ab43bc2 · outbound

This paper cites Multidiff: Consistent novel view synthesis from a single image.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Multidiff: Consistent novel view synthesis from a single image

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.926091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.293669Z digest=sha256:bd4b7f3892e64b22aa3fee8e1b7d7c5a3a8ba30de3f434082b6b2e41629eeade

Observation f579dda6-066d-473c-9442-eb29520baa37 · outbound

This paper cites Reg- nerf: Regularizing neural radiance fields for view synthesis from sparse inputs.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Reg- nerf: Regularizing neural radiance fields for view synthesis from sparse inputs

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.910463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.298196Z digest=sha256:2488ece150d106bf8eb1672e55c56236bd28916eb2260504ca9914a466dc3e73

Observation 8ba15dd5-6070-4b62-965b-f38a573eb15e · outbound

This paper cites Genie 2: A large-scale foundation world model, 2024.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Genie 2: A large-scale foundation world model, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.895654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.302459Z digest=sha256:3e3ad74f2127686ae24ad73dd1b306edef56c5d37e71aa037fa593814d650569

Observation 7081efb7-ea7e-45a3-b550-3f656ea89aee · outbound

This paper cites Freeman, and Michael Rubinstein.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Freeman, and Michael Rubinstein

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.881241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.306738Z digest=sha256:81c759e5db6c480e3a7c48053f1252fb0655709a301f3ad3486685253283bbe3

Observation 56c8bafd-2187-4a89-84d7-f839015c4abe · outbound

This paper cites Look outside the room: Synthesizing a consistent long-term 3d scene video from a single image.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Look outside the room: Synthesizing a consistent long-term 3d scene video from a single image

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.866790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.310977Z digest=sha256:501c85e1db07152f9a639c7fd7cb9f575c0d6833787d84a1d847bbe3f1dc6c54

Observation 85bdd673-92de-462f-8f13-a136417e4ef2 · outbound

This paper cites Gen3c: 3d-informed world-consistent video generation with precise camera con- 9 trol.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Gen3c: 3d-informed world-consistent video generation with precise camera con- 9 trol

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.851839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.315241Z digest=sha256:1a70386390e82011a55dafe4d45225b555f94f88ce2076d28a74049704d76e70

Observation 26acf9de-a973-4e3f-add2-e1f96c4e581d · outbound

This paper cites Pixel- synth: Generating a 3d-consistent experience from a single image.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Pixel- synth: Generating a 3d-consistent experience from a single image

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.319542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.319542Z digest=sha256:5e2795a1a0a2196acf0f3719c18a9b2fa96db77f0cee20df491c86fc8b4a78d8

Observation d9994c7b-426a-4ad3-abd8-8ce97e726d8a · outbound

This paper cites Geometry-free view synthesis: Transformers and no 3d pri- ors.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Geometry-free view synthesis: Transformers and no 3d pri- ors

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.828098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.323810Z digest=sha256:3e6d191342869175984f9f7bca2e98feec4dd6a5bdce55151610039848719540

Observation 7b19f49a-8e83-45c9-a6c7-92098039c04d · outbound

This paper cites GenWarp: Single Image to Novel Views with Semantic-Preserving Generative Warping.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory GenWarp: Single Image to Novel Views with Semantic-Preserving Generative Warping

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.328050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.328050Z digest=sha256:cdf10e312219793f20547a9383bfa783f1d012321fe38265ff2ecccf4516b368

Observation b2b7f035-5276-48cb-903c-ab9a51de64be · outbound

This paper cites Zero123++: a Single Image to Consistent Multi-view Diffusion Base Model.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Zero123++: a Single Image to Consistent Multi-view Diffusion Base Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.332942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.332942Z digest=sha256:a92d5a37df23589208c37907ae8390a0102a934689aa034315280bfcb3be8af1

Observation c12bb08e-27c3-4249-9c70-b57a01ec86bd · outbound

This paper cites Flash3d: Feed-forward gener- alisable 3d scene reconstruction from a single image.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Flash3d: Feed-forward gener- alisable 3d scene reconstruction from a single image

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.813889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.337573Z digest=sha256:c21bbe4effd3315f5fdd51bfeb04c044445fd277bb80ad4dcd748e8732f3d3a5

Observation 37e85b76-892e-4633-8138-ca5008a893f8 · outbound

This paper cites Splatter Image: Ultra-fast single-view 3D recon- struction.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Splatter Image: Ultra-fast single-view 3D recon- struction

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.342114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.342114Z digest=sha256:317dcf7d7bb9b6e3f56b98cd7053bb8cd219fa136f80dc16e7e61484889dd3aa

Observation 97969906-7731-41f5-8542-8db6c3b70c0a · outbound

This paper cites Sparf: Neural radiance fields from sparse and noisy poses.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Sparf: Neural radiance fields from sparse and noisy poses

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.346293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.346293Z digest=sha256:e9126f077e62f88c22e72b306639fbaba212349b47833080d14f8172a0e3b69f

Observation d4e791ea-bea5-4ef8-a0de-2a256589acbc · outbound

This paper cites Consistent view synthe- sis with pose-guided diffusion models.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Consistent view synthe- sis with pose-guided diffusion models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.781560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.350562Z digest=sha256:26fcdaee13e0c41a55ea916b367a0d306dd016bf83a7c840eaeca8549f4a6533

Observation 60d1c29b-b81b-4af0-833a-b650fa015388 · outbound

This paper cites Efros, and Angjoo Kanazawa.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Efros, and Angjoo Kanazawa

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.767513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.354695Z digest=sha256:276f883e7f877c78f62bfd3d3402d0b24576781d6a329d4f94c1c1ec39b0d10f

Observation 10407b3e-656a-4128-be13-5b7d8ef6f9d5 · outbound

This paper cites Dust3r: Geometric 3d vi- sion made easy.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Dust3r: Geometric 3d vi- sion made easy

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.753361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.358921Z digest=sha256:beefb144991a1278348a0dd2912747d91656ade7f28bf0b5cb5ae0c902770416

Observation 8ac958c1-8384-498a-ba75-d1c06aab8d01 · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Image quality assessment: from error visibility to structural similarity

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.363135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.363135Z digest=sha256:5cb99978f3e5da10f8a3ec512ca90f58d719095ec46ac5c23b05231af3c8a01d

Observation a85e073b-92cf-436e-afea-3db49c5284a1 · outbound

This paper cites Motionctrl: A unified and flexible motion controller for video generation.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Motionctrl: A unified and flexible motion controller for video generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.729382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.367690Z digest=sha256:379b5e9babfd68433c25593fe5d60aac0bb4cedfaf491b81be0e1e484d62e890

Observation aa654a19-ab04-4d5d-a84f-f0ef55ba838a · outbound

This paper cites SynSin: End-to-end view synthesis from a single image.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory SynSin: End-to-end view synthesis from a single image

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.714501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.372120Z digest=sha256:6571442acfa79fda2fc526cae50edb4e12615bb851c87d6dade9fc394be80a8f

Observation 72ca79eb-9775-48ee-b752-03bf820cfcff · outbound

This paper cites Reconfusion: 3d reconstruction with diffusion priors.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Reconfusion: 3d reconstruction with diffusion priors

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.700711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.376410Z digest=sha256:098d3d9b277892bf6c2fb7b94b397b3a596a738d78e2062f3f30ee0072a56b26

Observation f47c0f9c-ed6a-49cd-98cb-76a2fadaccbe · outbound

This paper cites Structured 3D Latents for Scalable and Versatile 3D Generation.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Structured 3D Latents for Scalable and Versatile 3D Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.380754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.380754Z digest=sha256:8d941551c5a761f73198b716c20866eec43a324318993752c55c3213f7c0d7ca

Observation ae01d324-f964-412f-9043-cc83edc457b2 · outbound

This paper cites Worldmem: Long- term consistent world simulation with memory, 2025.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Worldmem: Long- term consistent world simulation with memory, 2025

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.686515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.385553Z digest=sha256:57636c839047f9b5c2fe0655f8956af8b4e48978c93d70cf06fb0d3144461037

Observation 2fdfa670-b2b8-4fab-a882-45a16422e152 · outbound

This paper cites WonderJourney: Going from Anywhere to Everywhere.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory WonderJourney: Going from Anywhere to Everywhere

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.390198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.390198Z digest=sha256:87c4f3657a287e5079fc98c7e8b4dadb7bd227f2143ecd6f05dbd6a3d50dd38c

Observation e52e608a-0d15-4768-994e-7b4ffa8a1af0 · outbound

This paper cites WonderWorld: Interactive 3D Scene Generation from a Single Image.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory WonderWorld: Interactive 3D Scene Generation from a Single Image

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.395021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.395021Z digest=sha256:ec55a70a0ca2d3ba82a4103b87a8a549046c5c44819115ac8c28da009dcdac3b

Observation 0b32d7fc-e62f-44a1-a168-d39ca32cb2d2 · outbound

This paper cites Long-term photometric consistent novel view synthesis with diffusion models.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Long-term photometric consistent novel view synthesis with diffusion models

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.672684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.400815Z digest=sha256:e97d275d6504e942f2baee3c64f7278e2c3a6a31dbbbfc21ffc60cc4df823457

Observation dcf4d948-3bc3-4372-8a02-2ecc49237467 · outbound

This paper cites ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.405963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.405963Z digest=sha256:5bc2288bcf3b0636c67b4f0cdb39cab71ee66fbd027f01e32640d1983e9fbb39

Observation 8882c8b9-204b-46dc-b8de-0a52f82ec8a1 · outbound

This paper cites Stargen: A spatiotemporal autoregression framework with video dif- fusion model for scalable and controllable scene generation,.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Stargen: A spatiotemporal autoregression framework with video dif- fusion model for scalable and controllable scene generation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:44:55.658212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T18:44:55.410547Z digest=sha256:0bd0d7fba54fc4629b60969064599292722f5cd86cd132504e9a1b5ca1dfbc13

Observation 28f4a810-5771-4100-9514-b2432af03caf · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory The unreasonable effectiveness of deep features as a perceptual metric

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.415264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.415264Z digest=sha256:4f41ebcbc8050c151ce99bf8780b55544a4120aa77631ed0413aa8807138444b

Observation b7987984-6ecf-468e-b6fd-77a64e697e88 · outbound

This paper cites Stable Virtual Camera: Generative View Synthesis with Diffusion Models.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Stable Virtual Camera: Generative View Synthesis with Diffusion Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.419913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.419913Z digest=sha256:576e36212ec0ec6c746db729ee63433520e78b20a8cec6add72e6ebbbee4b871

Observation 8eb2ca01-6884-4ff9-bbae-f56586a1fbcf · outbound

This paper cites Stereo Magnification: Learning View Synthesis using Multiplane Images.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Stereo Magnification: Learning View Synthesis using Multiplane Images

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.425136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.425136Z digest=sha256:517c97487fa7b99632b2ce269606ff718636554c5df627bd93b929610c8ebe66

Observation c584eebf-9972-40e9-bf18-66558bbd58a7 · outbound

This paper cites Single-Stage Diffusion NeRF: A Unified Approach to 3D Generation and Reconstruction.

VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory Single-Stage Diffusion NeRF: A Unified Approach to 3D Generation and Reconstruction

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:55.229391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:55.229391Z digest=sha256:da0c17163579f60f8718fd8868c00e836df008b9746f4e23905306bcea7b3706

Pith citing papers

Observation e0e09699-5170-49d9-b20d-3498f807d1b4 · inbound

CustomX: Unified Character, Action, and Scene Customization in Video World Models cites this paper.

CustomX: Unified Character, Action, and Scene Customization in Video World Models VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T15:28:54.859365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:28:54.859365Z digest=sha256:8d5fb2fc57d59fac0144b5707d8e1142b44f211af9ee5d4446c14d33dcaac604

Observation 9e0238c0-1202-4efe-ab8c-69888bb3f6cf · inbound

UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models cites this paper.

UCM: Unified Modeling of Camera Control and Memory with Time-aware Positional Encoding Warping for World Models VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T20:36:11.425146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:36:11.425146Z digest=sha256:6a553477ca732525672e188cf18ef3c7ebd1f2a37f31e81ec7e426b05f8e03b4

Observation 62eb21f5-d460-4442-ada8-5f9fa0bdc3ae · inbound

Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding cites this paper.

Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T17:54:05.655438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:54:05.655438Z digest=sha256:8f8927b5837a65d13130fb707e9e0a96bbd27cbe0d38877af3136ef90682c9cd

Observation cf8a707d-f829-4682-9576-03af98789a77 · inbound

I3DM: Implicit 3D-aware Memory Retrieval and Injection for Consistent Video Scene Generation cites this paper.

I3DM: Implicit 3D-aware Memory Retrieval and Injection for Consistent Video Scene Generation VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T02:30:19.482363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:30:19.482363Z digest=sha256:b931a0ba45cf2c40b44d1a43c91dbcffb49767301f095c1d21e1c919a1c8fdd2

Observation d4919605-74b9-4e4c-abb6-af4493587453 · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 287

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.422882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:1870fc6bbdc33c76e8cb35ef58bade4eea87adeae953c9642e71d1ed843c1454

Observation 539fda55-bd04-4f4a-a8c2-d659a81ca089 · inbound

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds cites this paper.

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:45:27.241056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T13:45:24.961208Z digest=sha256:86f57ff5ee4fc1f0a6dd3e22f3db4eed1ea6325400ce33de0faa3daec8f9b459

Observation 48857495-eb75-4a88-803c-68b0c6c75231 · inbound

WorldKV: Efficient World Memory with World Retrieval and Compression cites this paper.

WorldKV: Efficient World Memory with World Retrieval and Compression VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:21:10.125264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T06:20:18.637998Z digest=sha256:f67e009be0ee4d0a33536918bee5d783210aa6a381c177b1b9819b793d4e715c

Observation 1f5e6537-9dde-4498-955a-9bad94af393a · inbound

E$^3$C: Video Generation with 3D Environmental Memory and Ego-Exo Human Pose Control cites this paper.

E$^3$C: Video Generation with 3D Environmental Memory and Ego-Exo Human Pose Control VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T22:34:02.493682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T22:24:38.389294Z digest=sha256:c895fe5af2da0aaec4b171b4dea649a72d8d32dc479300ba9632afc61218fdbe

Observation d6e19b2a-9e74-4d4b-a3e8-ec79cef350f3 · inbound

Robust Dreamer: Deviation-Aware Latent Gaussian Memory for Action-Controlled AR Video Generation cites this paper.

Robust Dreamer: Deviation-Aware Latent Gaussian Memory for Action-Controlled AR Video Generation VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:02:46.964385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T22:56:35.530309Z digest=sha256:59cced69f0be667a774dc605ff02cd0e85de2876de0646a63eae0273b7f40438

Observation d6ad3910-ebaf-4067-bdfe-de77ebedf16d · inbound

DecMem: Towards Minute-Long Consistent World Generation with Decoupled Memory cites this paper.

DecMem: Towards Minute-Long Consistent World Generation with Decoupled Memory VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:42:46.841506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T22:38:10.473453Z digest=sha256:8b3f9e34ec7c2e4e6ef23ae37ac84febfb1e50edfbfe2dc033e2b066858f2e0a

Observation d23b27c4-f2da-4dae-8c53-36a2b1dd05e3 · inbound

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data cites this paper.

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:56:20.197617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T14:52:30.406683Z digest=sha256:5a98446180098849c073937516a456599ff8310e5d28470c0ccc5dc17d5547a6

Observation 9abdcd1c-0cdf-44fa-acaa-293c453449d1 · inbound

Latent Spatial Memory for Video World Models cites this paper.

Latent Spatial Memory for Video World Models VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:07:30.579093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T16:47:42.761342Z digest=sha256:fd1866b39794c3aba98a94c52a2dc08de4d46a4bec1031f9e41d07437c62ea12

Observation 80c55eca-3c03-4ad1-932c-dcd2be0be63a · inbound

PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory cites this paper.

PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:58:47.752156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T03:23:58.455778Z digest=sha256:81a023f282ed57be9fb29b37d01a4e3ff7f8d17a17270fa6859e702d03d212b0

Observation f7f4d308-7a1d-48e0-b60e-b026f81abe37 · inbound

MemLearner: Learning to Query Context memory for Video World Models cites this paper.

MemLearner: Learning to Query Context memory for Video World Models VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.888623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-01T05:30:56.140465Z digest=sha256:6620d8162372250e1817502e1e6e22d87d362a863c26104f379e206556f290c4

Observation de891703-bf83-4994-a4e0-ab76812214f2 · inbound

Addressable Memory for Video World Models cites this paper.

Addressable Memory for Video World Models VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.863744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.863744Z digest=sha256:078fdc465fbc4f550a80d0791820cd12c94d182d919ee7e9c1a053fcb1842b86