Pith. sign in

Paper Citation Record · LEDGER

Addressable Memory for Video World Models

As of 10 August 2026, this Paper Citation Record lists 100 of 120 outbound references and 0 inbound Pith citation observations for arXiv:2608.07408.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07408 v1

Coverage vector

measured 100 of 120 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T05:06:52.022607Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 120 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved97
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 91dddf21-8cd7-473a-a03a-66e3ae65cbcf · outbound

This paper cites Cosmos 3: Omnimodal World Models for Physical AI.

Addressable Memory for Video World Models Cosmos 3: Omnimodal World Models for Physical AI

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.699806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.699806Z digest=sha256:0322157d4855cf285e0ca52643c456f0ac4dfb4f502aaa5dd61ea4f22ff85662

Observation 7528a57e-f4fd-4f91-afb6-ee3666f0aa7e · outbound

This paper cites Round and Round We Go! What makes Rotary Positional Encodings useful?.

Addressable Memory for Video World Models Round and Round We Go! What makes Rotary Positional Encodings useful?

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.704237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.704237Z digest=sha256:3bbb0c58fe29e04b04dafa571f9fc485faa3e6e99a4b36162e5bf679426fa5f3

Observation a8f3e74f-95e1-41ce-9902-adf2a33ca7dc · outbound

This paper cites NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation.

Addressable Memory for Video World Models NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:06:56.404991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:06:51.707884Z digest=sha256:9a1f563fbdc3cc0599aab805c894a5bff2f38ab1b02217f00f0915ed85777f1a

Observation 5d95ce7c-d24c-4c40-9d2a-26901cf9f142 · outbound

This paper cites Longformer: The Long-Document Transformer.

Addressable Memory for Video World Models Longformer: The Long-Document Transformer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.711233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.711233Z digest=sha256:e626e70155153d94643bdcf1e2242962007b783d92bf5cca2e149ad49f454b62

Observation c47d67ba-feee-4af2-ab43-c0b46352b11f · outbound

This paper cites Variance Reduction for Expectations with Diffusion Teachers.

Addressable Memory for Video World Models Variance Reduction for Expectations with Diffusion Teachers

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:06:56.381256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:06:51.714636Z digest=sha256:1b7ac78f67722c256d3c51976a92e4c2657ff7c618a94aea712c9af85c3d8e5f

Observation d29b49ed-1b39-413e-a48e-73f7cd861f3c · outbound

This paper cites Token merging: Your ViT but faster.

Addressable Memory for Video World Models Token merging: Your ViT but faster

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.718099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.718099Z digest=sha256:a3d1f9838117b0486b09ce3a10630de2e27f204ae52bd6d406baf58862470652

Observation c4c3c9d2-ba07-46f1-9448-e4fe74264f76 · outbound

This paper cites Recurrentmemorytransformer.

Addressable Memory for Video World Models Recurrentmemorytransformer

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.721638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.721638Z digest=sha256:fa2365d1c8d1b2ee459b2f49f5e2f8d8ea88e24c3223f0c67d38dd66a34b7199

Observation fe7ebef4-6ad7-4617-bf71-6a922bc30e3e · outbound

This paper cites Mixture of contexts for long video generation.

Addressable Memory for Video World Models Mixture of contexts for long video generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.724686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.724686Z digest=sha256:e5c028c4c9e1eb6bb17370e2607bd84e66456edd06b7c4b28ebbe364f2c1778d

Observation 8f8f6f4e-e63a-4705-8c2c-04581997312a · outbound

This paper cites PyramidKV: Dynamic KV cache compression based on pyramidal information funneling.

Addressable Memory for Video World Models PyramidKV: Dynamic KV cache compression based on pyramidal information funneling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.727509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.727509Z digest=sha256:8c20899f872904e40b834303fafd3b2ccace0d51d5ce4dbea066abb3c1491d8f

Observation 66121e07-549e-442d-9b59-46cfcd0d5f47 · outbound

This paper cites Diffusion forcing: Next-token prediction meets full-sequence diffusion.

Addressable Memory for Video World Models Diffusion forcing: Next-token prediction meets full-sequence diffusion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.730698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.730698Z digest=sha256:411a28bcfbef4963c7818b3b55d7c849be88fee3eb5c6de041c8381bd1355f0d

Observation 72569696-3196-41d8-a983-122bbe4ae8d2 · outbound

This paper cites Past- and future-informed kv cache policy with salience estimation in autoregressive video diffusion.arXiv preprint arXiv:2601.21896, 2026.

Addressable Memory for Video World Models Past- and future-informed kv cache policy with salience estimation in autoregressive video diffusion.arXiv preprint arXiv:2601.21896, 2026

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.733972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.733972Z digest=sha256:cd8c3b02629d997cc3439cbe86882fbf35bf41412b17e6f8a974369b9b8e7875

Observation 1ffda83b-fb20-41b7-af03-d9bafca2f1e2 · outbound

This paper cites Grounded Forcing: Bridging Time-Independent Semantics and Proximal Dynamics in Autoregressive Video Synthesis.

Addressable Memory for Video World Models Grounded Forcing: Bridging Time-Independent Semantics and Proximal Dynamics in Autoregressive Video Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.737173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.737173Z digest=sha256:6e652dbc91c66b7602786826bec0e60ddb9a27b41f07bb9b919f3bee6ecc99df

Observation fc86cc5f-0460-4489-9718-5e58d3fe3b02 · outbound

This paper cites Extending Context Window of Large Language Models via Positional Interpolation.

Addressable Memory for Video World Models Extending Context Window of Large Language Models via Positional Interpolation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.740643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.740643Z digest=sha256:27ee11e5d57ec7f1fac7c6d1e4844590151c4da0e09cd9d8d684c7f1e77568dd

Observation 64a0907e-d5ff-458c-bdeb-997c74c5484a · outbound

This paper cites Context forcing: Consistent autoregressive video generation with long context.arXiv preprint arXiv:2602.06028, 2026.

Addressable Memory for Video World Models Context forcing: Consistent autoregressive video generation with long context.arXiv preprint arXiv:2602.06028, 2026

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.744146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.744146Z digest=sha256:675020dc487a426d6364d7582eb1b650c68860809276b63403ba28bed2b8bc4d

Observation 6baf39b6-e159-4933-80e4-39e35ece8d5b · outbound

This paper cites VRAG: Learning World Models for Interactive Video Generation.

Addressable Memory for Video World Models VRAG: Learning World Models for Interactive Video Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.747413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.747413Z digest=sha256:6257c582a8acf06511ea441c3b49d63521086ea0884cfcf9b1278dde84e8ed5f

Observation 634e47f8-392f-448a-aff0-da0d87426578 · outbound

This paper cites FINCH: Prompt-guided key-value cache compression for large language models.

Addressable Memory for Video World Models FINCH: Prompt-guided key-value cache compression for large language models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.751167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.751167Z digest=sha256:3d32578fd7ca5b2f75eb6c6389de33a001d794bd2bbab8b4574d3270e27baaa1

Observation cbd6ee8d-35dc-491d-8390-ac9135b57c4f · outbound

This paper cites Lol: Longer than longer, scaling video generation to hour.arXiv preprint arXiv:2601.16914, 2026.

Addressable Memory for Video World Models Lol: Longer than longer, scaling video generation to hour.arXiv preprint arXiv:2601.16914, 2026

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.754420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.754420Z digest=sha256:bedfb56176b738d992e053db671b26e4e635c128e843e57e171075a8d6cc8056

Observation 19f274db-1514-4a23-8140-2e0eacc5720d · outbound

This paper cites Self-Forcing++: Towards Minute-Scale High-Quality Video Generation.

Addressable Memory for Video World Models Self-Forcing++: Towards Minute-Scale High-Quality Video Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.757833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.757833Z digest=sha256:cd50c6f1ac11acee0f802d2fdef4ac89b2fc383736071063a62fc8889311e96c

Observation 40a1ac60-fc60-45cb-8c83-9d9ada4bd0cf · outbound

This paper cites Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context.

Addressable Memory for Video World Models Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.761506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.761506Z digest=sha256:2694908516f6eccac68511bb2dd52612d67914624a1159f1f266453c9487374e

Observation ea1c6f51-d05b-4b31-96ce-bb17f324f423 · outbound

This paper cites Oasis: A universe in a transformer.https://oasis-model.github.io, 2024.

Addressable Memory for Video World Models Oasis: A universe in a transformer.https://oasis-model.github.io, 2024

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.765198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.765198Z digest=sha256:e28280f637bbf2ad83af74e0fb56035de55949e628db351d0c8654728a73f60b

Observation f437397a-bae7-430f-8b66-5650f628291b · outbound

This paper cites WorldScore: A unified evaluation benchmark for world generation.

Addressable Memory for Video World Models WorldScore: A unified evaluation benchmark for world generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.768499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.768499Z digest=sha256:617062c1a50b1a036273834bd7ba86a540f370d52b6fa7bd4cd3778dbd57fdf1

Observation f019ce2e-f743-4eb5-89d2-ca3c65e8642a · outbound

This paper cites A unified framework for approximating and clustering data.

Addressable Memory for Video World Models A unified framework for approximating and clustering data

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.771607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.771607Z digest=sha256:6d50a55a67eb23b48eb3130c098e5ec8b70a212e5a1c831a867f64c67a6666b6

Observation 03483878-b4e0-499b-81bd-85e31a3f7192 · outbound

This paper cites Memcam: Memory- augmented camera control for consistent video generation.arXiv preprint arXiv:2603.26193, 2026.

Addressable Memory for Video World Models Memcam: Memory- augmented camera control for consistent video generation.arXiv preprint arXiv:2603.26193, 2026

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.775432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.775432Z digest=sha256:3bc731e1ca140750b762f0da79ee9e03c69d4f13f30666df2c4001e39ca6608e

Observation e3736da0-d508-4182-964d-a81e8e3f08e5 · outbound

This paper cites Contextual Position Encoding: Learning to Count What's Important.

Addressable Memory for Video World Models Contextual Position Encoding: Learning to Count What's Important

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.778667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.778667Z digest=sha256:87f5b021373ec248c812f048ea2ecd1bfa28ccb46188aa6dcada98e06099d990

Observation a6f6d7e9-d04d-43b0-b90f-1d8a30699903 · outbound

This paper cites Genie 3: A new frontier for world models.

Addressable Memory for Video World Models Genie 3: A new frontier for world models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.781948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.781948Z digest=sha256:81bf48e6160cc926a0e9b9a5cd90f2e3b7de1fd4784074b09b1a50f4f5e44cca

Observation 45dba7e6-fa47-460a-947d-89942018d259 · outbound

This paper cites Mamba: Linear-time sequence modeling with selective state spaces.

Addressable Memory for Video World Models Mamba: Linear-time sequence modeling with selective state spaces

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.784638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.784638Z digest=sha256:45daa2408f354f68dd91f90e7170631d957492ef0b9610dcc07f6c3e84b07539

Observation 5b919a9d-fe83-4fc2-966b-325295ad91bb · outbound

This paper cites Long-Context Autoregressive Video Modeling with Next-Frame Prediction.

Addressable Memory for Video World Models Long-Context Autoregressive Video Modeling with Next-Frame Prediction

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.787420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.787420Z digest=sha256:8a02dd898aead57c894643c1fe2ce880d9dce8d6808843f3ba742e3e82a1571a

Observation 6e8384ef-1551-4dec-8358-f5370072c761 · outbound

This paper cites Memorize When Needed: Decoupled Memory Control for Spatially Consistent Long-Horizon Video Generation.

Addressable Memory for Video World Models Memorize When Needed: Decoupled Memory Control for Spatially Consistent Long-Horizon Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.790280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.790280Z digest=sha256:208a10ace5ccfc5f3d6ffbea217afe3bccb1afa3c33ca8c7d0e031e902631fe2

Observation d25dd89a-6b4d-4c51-b4aa-3497b5ac48b2 · outbound

This paper cites Recurrent world models facilitate policy evolution.

Addressable Memory for Video World Models Recurrent world models facilitate policy evolution

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.793059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.793059Z digest=sha256:90453494132fd037bf3f5ce91b4bb402fabfd4bd44e33f6192e17d984584b0b9

Observation 52bc8748-c20a-4621-be85-6e16b6584354 · outbound

This paper cites Mastering diverse control tasks through world models.Nature, 640:647–653, 2025.

Addressable Memory for Video World Models Mastering diverse control tasks through world models.Nature, 640:647–653, 2025

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.795696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.795696Z digest=sha256:a30a107d5c1146ae9b98d39bae227a4ba7ac422b55b01541d7c2fd585c45257a

Observation d3ed0f70-18c3-4768-826a-8056568bb32c · outbound

This paper cites A$^2$ATS: Retrieval-Based KV Cache Reduction via Windowed Rotary Position Embedding and Query-Aware Vector Quantization.

Addressable Memory for Video World Models A$^2$ATS: Retrieval-Based KV Cache Reduction via Windowed Rotary Position Embedding and Query-Aware Vector Quantization

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:06:55.948862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:06:51.798489Z digest=sha256:0623f91799c503bfdc690affa9c0dcf3d1b249be0b7fd14fcd36f1287c0161e1

Observation 613f3692-f94d-4567-af19-01e552b5d47d · outbound

This paper cites Matrix-game 2.0: An open-source real-time and streaming interactive world model.

Addressable Memory for Video World Models Matrix-game 2.0: An open-source real-time and streaming interactive world model

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.801289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.801289Z digest=sha256:fd45f1698423f86fc7024b5b5587f32cce336be4350f9477e840086f786f3bb4

Observation 1ecc7a16-0e7d-4c3a-91e9-62594a23f9c8 · outbound

This paper cites StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text.

Addressable Memory for Video World Models StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.804256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.804256Z digest=sha256:61cfaca6ddb040f9699837dbacb10e6d69a5b08e258be8345e9e92126251a6b6

Observation 22778dbf-972b-4f61-a063-b0dcb288ae08 · outbound

This paper cites Rotary Position Embedding for Vision Transformer.

Addressable Memory for Video World Models Rotary Position Embedding for Vision Transformer

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.807153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.807153Z digest=sha256:881d7ce35bf6eb8318c09c09401b2ba71477499cde650833a2c8e0b07ccd4332

Observation d5056227-bf25-4709-829e-8880f2b2cb0b · outbound

This paper cites RELIC: Interactive video world model with long-horizon memory.arXiv preprint arXiv:2512.04040, 2025.

Addressable Memory for Video World Models RELIC: Interactive video world model with long-horizon memory.arXiv preprint arXiv:2512.04040, 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.810004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.810004Z digest=sha256:86960547d13bb30f98de68e62145be5e4ef1950ba9f4b5735d0bcae5c87555e6

Observation bfa75425-95e9-4cc8-95a2-268522f777dc · outbound

This paper cites Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization.

Addressable Memory for Video World Models Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.812677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.812677Z digest=sha256:e87dd9f1b806e4ea26bcf6f0582fa3b19a21f3a25ccedfb9c2ac9867e3458f91

Observation 578fb85c-1927-4c6b-b32e-f9ebd635c0e0 · outbound

This paper cites Vid2World: Crafting video diffusion models to interactive world models.

Addressable Memory for Video World Models Vid2World: Crafting video diffusion models to interactive world models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.815893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.815893Z digest=sha256:008a7a041af895a9b0f94646ebc35e77fb73b7f398878229231963a6cf4398a1

Observation a7b5482e-334b-40b0-9d81-117133e1dfaa · outbound

This paper cites Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion.

Addressable Memory for Video World Models Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.818857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.818857Z digest=sha256:c608d6cff9a5bc7512aed7aba10c5f7e036cefc8c26683e3fc73f6924f82d22f

Observation c161ec1f-f0c5-4703-939c-b9dcdc7b853a · outbound

This paper cites Block-Recurrent Transformers.

Addressable Memory for Video World Models Block-Recurrent Transformers

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.822520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.822520Z digest=sha256:58dc6b5b139b7e43597fc48f65428d75d7263705092a9189f3d21b5ac12e1ea2

Observation ee220334-8634-4eca-a680-40f0eee0b7e0 · outbound

This paper cites Transformers are RNNs: Fast autore- gressive transformers with linear attention.

Addressable Memory for Video World Models Transformers are RNNs: Fast autore- gressive transformers with linear attention

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.825902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.825902Z digest=sha256:91dfb6a4965f1cc59dd4fa8c09ff1f8b070b32bc8384cf16a7e6a5e2cafe07aa

Observation 8f2611ed-07eb-4757-83a6-b167dacfe868 · outbound

This paper cites The Impact of Positional Encoding on Length Generalization in Transformers.

Addressable Memory for Video World Models The Impact of Positional Encoding on Length Generalization in Transformers

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.829142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.829142Z digest=sha256:274c824068a87f1c5979cc37be8093d74b39d428b4fd1c09c65bbfb14fb6816d

Observation 1962a171-d7bb-4c8a-979f-14fbaf6a2575 · outbound

This paper cites Jay Kuo, and Peter A.

Addressable Memory for Video World Models Jay Kuo, and Peter A

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.832659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.832659Z digest=sha256:c2272f1f3d5ebe859dc3fc3c563b02e3e39684dd79b7640275157aa31d8fd267

Observation b421cdb8-9622-4ef8-b681-bbcdd47326a1 · outbound

This paper cites Robust nonnegative matrix factorization using l21-norm.

Addressable Memory for Video World Models Robust nonnegative matrix factorization using l21-norm

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.835834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.835834Z digest=sha256:5a335b6a411a34f1b25231b477f29a26881be4c8effdf4c7fedb43f6f4c511ac

Observation 88b6d623-057b-4204-9a35-ba8d372a4def · outbound

This paper cites Lee and H.

Addressable Memory for Video World Models Lee and H

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.839051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.839051Z digest=sha256:63f24aa67a1b15e9db4816d8721d223fa609d8e7575b87afcebd380a7dce3a52

Observation 383659e2-0643-4ac9-824d-9b509bfdda14 · outbound

This paper cites Understanding and Improving Length Generalization in Hierarchical Sparse Attention Models.

Addressable Memory for Video World Models Understanding and Improving Length Generalization in Hierarchical Sparse Attention Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.842394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.842394Z digest=sha256:c2d235515e42954827316a81a0184014490dccf9b4011c990647daaa92dae6c7

Observation 905facbf-630e-4981-9740-9d27a14b048c · outbound

This paper cites Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion.

Addressable Memory for Video World Models Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.845827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.845827Z digest=sha256:45635259e47801517327670edf1a918640e049dceeca4dcd2d62886223cd305c

Observation 6723b123-bdad-46d4-bdbb-9b27cc723b67 · outbound

This paper cites Train short, inference long: Training-free horizon extension for autoregressive video generation.arXiv preprint arXiv:2602.14027, 2026.

Addressable Memory for Video World Models Train short, inference long: Training-free horizon extension for autoregressive video generation.arXiv preprint arXiv:2602.14027, 2026

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.849683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.849683Z digest=sha256:9543d8de12685ca80b78af4d9733fe2d2d788b5b68c79518b7930c5739282dbe

Observation 7e76c29e-770c-49e0-a061-6c12c4a6f8b7 · outbound

This paper cites PackCache: A training-free acceleration method for unified autoregressive video generation via compact KV-cache.arXiv preprint arXiv:2601.04359, 2026.

Addressable Memory for Video World Models PackCache: A training-free acceleration method for unified autoregressive video generation via compact KV-cache.arXiv preprint arXiv:2601.04359, 2026

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.853092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.853092Z digest=sha256:9fc180b2fc21adb00c213f8ef9604ab5a55b2900f926a42934023d8e50a0024a

Observation 02f61ce5-176a-42d0-b586-cf9e55dc44fb · outbound

This paper cites Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation.

Addressable Memory for Video World Models Long-Horizon Streaming Video Generation via Hybrid Attention with Decoupled Distillation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.856531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.856531Z digest=sha256:04846d74f9aab7d8c0160a0fd8635dd2046dde5586148d4f4d77b89cbd9172e9

Observation bb0d9e24-d699-45a6-ac54-f152584b9165 · outbound

This paper cites Cameras as relative positional encoding.

Addressable Memory for Video World Models Cameras as relative positional encoding

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.860251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.860251Z digest=sha256:db3f13ab01944a6641ae2038ad34890a6b0b9bdc6c660e7b902372dbdfca543d

Observation de891703-bf83-4994-a4e0-ab76812214f2 · outbound

This paper cites VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory.

Addressable Memory for Video World Models VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.863744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.863744Z digest=sha256:75285f0d12e2cd114d307e73d7190df29ef21ebfe39becb229326e92cad1c3f5

Observation 32182234-1248-4cdc-bd14-87e95f0bce03 · outbound

This paper cites Stable video infinity: Infinite-length video generation with error recycling.arXiv preprint arXiv:2510.09212, 2025.

Addressable Memory for Video World Models Stable video infinity: Infinite-length video generation with error recycling.arXiv preprint arXiv:2510.09212, 2025

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.866780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.866780Z digest=sha256:722e57c285ab0c9c01bbbda325f1cb085d8896432a88dc262c4f73b59aec0e3e

Observation 94be4b49-5b3b-4c47-9e14-b5e75d728040 · outbound

This paper cites SnapKV: LLM Knows What You are Looking for Before Generation.

Addressable Memory for Video World Models SnapKV: LLM Knows What You are Looking for Before Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.869780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.869780Z digest=sha256:b2b32ef1e57244268616a1ef81c031520f04ad6125d31cd97ab76c96ba51edea

Observation 562e98e3-2e84-4069-9556-a10cd7445178 · outbound

This paper cites LoopNav: Benchmarking Spatial Consistency in World Models.

Addressable Memory for Video World Models LoopNav: Benchmarking Spatial Consistency in World Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.873150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.873150Z digest=sha256:31834271b5ff307ff08f0eaa2baf246bfa81069511b95f32fabe3c779f7b6507

Observation 9329cf75-9c5e-4d92-8c27-a520c49578a8 · outbound

This paper cites Rolling Forcing: Autoregressive Long Video Diffusion in Real Time.

Addressable Memory for Video World Models Rolling Forcing: Autoregressive Long Video Diffusion in Real Time

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.876309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.876309Z digest=sha256:1e317c2d3d6e51d0afa57f7e355d95b9ffc19a9e0346faeac0088b5e6b5ab838

Observation 5e2835f8-be93-4c5a-acbf-5a5e80310f76 · outbound

This paper cites KIVI: A Tuning-Free Asymmetric 2bit Quantization for KV Cache.

Addressable Memory for Video World Models KIVI: A Tuning-Free Asymmetric 2bit Quantization for KV Cache

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.879558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.879558Z digest=sha256:b22d297e89ec57c86363faf82e57db5fdad2d93ead1d149f45c32e1722d44cd4

Observation f941b651-7c6a-4d9e-be46-daeada6be491 · outbound

This paper cites JacNet: Learning functions with structured Jacobians.

Addressable Memory for Video World Models JacNet: Learning functions with structured Jacobians

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.882731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.882731Z digest=sha256:f3f72b6c05c109a12558daaa962f61dc4d92d352dbbc9bcf3d205085560f3cb4

Observation 552639f3-182c-4464-b22a-91d52fc77809 · outbound

This paper cites Task Selection for AutoML System Evaluation.

Addressable Memory for Video World Models Task Selection for AutoML System Evaluation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.885612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.885612Z digest=sha256:b428496f18d294d06c5a9fa2266b9f994be18107de14a738fe67e9c54d4ecc3c

Observation 9525ee8d-21fe-4950-9617-1cc7aa81e211 · outbound

This paper cites ATT3D: Amortized text-to-3D object synthesis.

Addressable Memory for Video World Models ATT3D: Amortized text-to-3D object synthesis

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.888849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.888849Z digest=sha256:dd6718e3241b65c5adb989be5e76a3e00348afb30d78b9422ed963bf684d650a

Observation 710479b7-ebba-4f2e-a7f6-96d07dceb05c · outbound

This paper cites Flow caching for autoregressive video generation.arXiv preprint arXiv:2602.10825, 2026.

Addressable Memory for Video World Models Flow caching for autoregressive video generation.arXiv preprint arXiv:2602.10825, 2026

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.891520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.891520Z digest=sha256:16bc0e695b1203c889580d7b33ef234d02429a514299180143061810facd6bc2

Observation db742686-b706-49a0-b81c-697c9cdfe559 · outbound

This paper cites TriAttention: Efficient Long Reasoning with Trigonometric KV Compression.

Addressable Memory for Video World Models TriAttention: Efficient Long Reasoning with Trigonometric KV Compression

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.894378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.894378Z digest=sha256:3e58887e4a6a4e13d052d17821d669e4f56539040562915ce2281cd9ba505939

Observation 5a383fa8-5c2d-4d8f-aa09-a6fae5d36853 · outbound

This paper cites PackForcing: Short video training suffices for long video sampling and long context inference.arXiv preprint arXiv:2603.25730, 2026.

Addressable Memory for Video World Models PackForcing: Short video training suffices for long video sampling and long context inference.arXiv preprint arXiv:2603.25730, 2026

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.897344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.897344Z digest=sha256:dbff79d60d18b89aefffcf8607dea050fc2c9b6361f946ee69364f7354c218a5

Observation 4d4e19e2-0b8a-485f-8a01-00bd839a6c22 · outbound

This paper cites Landmark Attention: Random-Access Infinite Context Length for Transformers.

Addressable Memory for Video World Models Landmark Attention: Random-Access Infinite Context Length for Transformers

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.900992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.900992Z digest=sha256:1feb3de2d7123a38e8d0ba2c4bf1e9f396b3d82188291a059932d61f75f04bea

Observation a788e2bf-48d1-495e-8bb0-e2821e928d0a · outbound

This paper cites Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention.

Addressable Memory for Video World Models Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.904460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.904460Z digest=sha256:cc76812cbdfa5eb354f69e2fde7d0d6fdef602348aac92f1ac62ed3cd094f002

Observation f7d0633f-ef88-46a1-8be5-ae07b9b18d9f · outbound

This paper cites KVPress: A compression library for transformer KV caches.https://github.com/NVIDIA/kvpress, 2024.

Addressable Memory for Video World Models KVPress: A compression library for transformer KV caches.https://github.com/NVIDIA/kvpress, 2024

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.908045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.908045Z digest=sha256:b91cdc5204ab4e03cc9f443a60e57e7b2af0faed0bc3d92e281e6777fa091a86

Observation 866ad531-9c4e-4db8-952d-3b7609f609b1 · outbound

This paper cites WorldPack: Dynamic Frame Compression for Long-context Video World Modeling.

Addressable Memory for Video World Models WorldPack: Dynamic Frame Compression for Long-context Video World Modeling

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.911350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.911350Z digest=sha256:d4f3b52a87201ed821216c0b49b297473fa8d30c2fd2ca16f2e771161014fd18

Observation e412048b-faec-4d7f-921a-bd699312e06b · outbound

This paper cites YaRN: Efficient Context Window Extension of Large Language Models.

Addressable Memory for Video World Models YaRN: Efficient Context Window Extension of Large Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.914934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.914934Z digest=sha256:891eba57fc7c969dcc96bd157432a598162e2fc3937279b426fcb0ea9db43753

Observation 56cb717b-9da6-449f-b17c-be8f860716db · outbound

This paper cites Long-context state-space video world models.

Addressable Memory for Video World Models Long-context state-space video world models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.918609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.918609Z digest=sha256:1c2d575a9a047ea0aa5acca54317cb259c09cc15422e8ce028c543974d063f16

Observation 79e8cde5-5f8c-425a-ac7c-ea83f84bc309 · outbound

This paper cites Smith, and Mike Lewis.

Addressable Memory for Video World Models Smith, and Mike Lewis

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.921685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.921685Z digest=sha256:8691ad91974054eca742b906849ac8b1a664ebd3b48fc294bea1b52d7a0a4386

Observation 1443c73a-8bf6-4d74-96cb-d73da9750e97 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Addressable Memory for Video World Models Learning transferable visual models from natural language supervision

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.924869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.924869Z digest=sha256:0a695124fdc8241e11bcc3bfe49d4dfd965c95c987bb01cb3aaec500bd16fd16

Observation 930ccfba-f3f9-4bc8-a89a-01a40f5b13ea · outbound

This paper cites Rae, Anna Potapenko, Siddhant M.

Addressable Memory for Video World Models Rae, Anna Potapenko, Siddhant M

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.928398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.928398Z digest=sha256:6440f2c76475e1af25e87a4e7dd5e7643cbde7d1a316da2695aa0b05660fb8ed

Observation 0aefec3a-6ed3-4e7c-b7a2-edc059d454c3 · outbound

This paper cites Score Distillation Sampling for Audio: Source Separation, Synthesis, and Beyond.

Addressable Memory for Video World Models Score Distillation Sampling for Audio: Source Separation, Synthesis, and Beyond

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.931568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.931568Z digest=sha256:701a4079e822648fd7893b18b5c40ea189c77e0d71da69b94b1bbc6a7be243bc

Observation 53088a46-4a55-42aa-bd86-9a3623203d8b · outbound

This paper cites Fast autoregressive video diffusion and world models with temporal cache compression and sparse attention.arXiv preprint arXiv:2602.01801, 2026.

Addressable Memory for Video World Models Fast autoregressive video diffusion and world models with temporal cache compression and sparse attention.arXiv preprint arXiv:2602.01801, 2026

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.935154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.935154Z digest=sha256:12bd602d5bd823560cde8cfba0a3349ea4f53d2fbf13f140cf1f395afdb1cbe7

Observation 97763b1a-405f-4127-8696-1a6dd8e498d4 · outbound

This paper cites LongRoPE2: Near-lossless LLM context window scaling.

Addressable Memory for Video World Models LongRoPE2: Near-lossless LLM context window scaling

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.938252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.938252Z digest=sha256:fd1ac6476032ebc3047069a4065d52d52bdab03b700534e7d7a24ced8b0b4a67

Observation 219c96eb-d2b2-44d7-ad62-c753a18d971a · outbound

This paper cites History-Guided Video Diffusion.

Addressable Memory for Video World Models History-Guided Video Diffusion

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.941856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.941856Z digest=sha256:4e5dfbf6c4367c2aa2bdadcd3cd79b2885cfb232d1c30ca906179acec45d866b

Observation 8722fe6f-b7dc-4e9d-9be1-807e6a546a3b · outbound

This paper cites Multi-student Diffusion Distillation for Better One-step Generators.

Addressable Memory for Video World Models Multi-student Diffusion Distillation for Better One-step Generators

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.945440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.945440Z digest=sha256:9e60ca51ed81ee36fb9eda3a250759bc6d0e242ededea1c0172d40071c9a6c2f

Observation bb5f8c81-180b-4a75-aa0e-09ae3ecb1078 · outbound

This paper cites Composition of memory experts for diffusion world models.

Addressable Memory for Video World Models Composition of memory experts for diffusion world models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.948895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.948895Z digest=sha256:86b325de35847b7d84f83788c8299b6e2d0e3196c753e3efd4fe13e9cd535f3e

Observation 22efbd61-f2aa-4b2c-b95f-ab977bbe44dc · outbound

This paper cites RoFormer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063, 2024.

Addressable Memory for Video World Models RoFormer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063, 2024

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.951728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.951728Z digest=sha256:257410c4a9165576ad9f37089cfe2506efb08eaaad9aeb78383f4ca5056effa9

Observation 2ced6be5-c8ec-446c-b8be-33feb282cf57 · outbound

This paper cites WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling.

Addressable Memory for Video World Models WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.954581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.954581Z digest=sha256:6f6c4f91b4973e7584820a033af5527b8ae13642b3651e7d56462a90ccf0da14

Observation dcb9b390-15dd-4de5-96cd-15fd81384c7e · outbound

This paper cites Learning to (learn at test time): RNNs with expressive hidden states.

Addressable Memory for Video World Models Learning to (learn at test time): RNNs with expressive hidden states

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.957468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.957468Z digest=sha256:3e6d3e5aa096c29b47018eff79b98d7f9b4e25740ce9da718173cc34d48cb6c7

Observation 0e61fcb2-60a9-41fe-b00d-cac9563f47a0 · outbound

This paper cites A Length-Extrapolatable Transformer.

Addressable Memory for Video World Models A Length-Extrapolatable Transformer

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.960146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.960146Z digest=sha256:eaac09b0a87ad457c85874690837ee0e2ad2742c4c7c7453df7e894504b08d50

Observation 924e2455-efe2-417e-a089-a8c83e870d69 · outbound

This paper cites Advancing Open-source World Models.

Addressable Memory for Video World Models Advancing Open-source World Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.963230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.963230Z digest=sha256:bfd3abc8fcb6572f3661565506566113a2404db31dc2ccc45e83a3eef896b482

Observation df4eb73f-d4b9-4344-aaf1-80db9dbf98d8 · outbound

This paper cites KeepKV: Achieving periodic lossless KV cache compression for efficient LLM inference.

Addressable Memory for Video World Models KeepKV: Achieving periodic lossless KV cache compression for efficient LLM inference

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.966146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.966146Z digest=sha256:187b10355274f4ca3b532bb56a4f75c2640fb495d19d4e4a89a39be412c6bf21

Observation a5e1beab-7c38-4a85-8e41-2c2aa5a3315e · outbound

This paper cites Diffusion models are real-time game engines.

Addressable Memory for Video World Models Diffusion models are real-time game engines

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.968987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.968987Z digest=sha256:28b54737e82bf1b232f1b048e9bbf1c26251b641fb9b08702339d6c0647aa8fa

Observation c6859384-b193-442f-9602-689ba0412c20 · outbound

This paper cites an unresolved cited work.

Addressable Memory for Video World Models Unresolved cited work

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.971985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.971985Z digest=sha256:6163b95f4e7a44a39a099ad375969dda9aebecb5a230a8c4bd14ae2930419c73

Observation 878363cd-acd2-4c73-8f18-0e636e4ee5fd · outbound

This paper cites Fast transformers with clustered attention.

Addressable Memory for Video World Models Fast transformers with clustered attention

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.974974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.974974Z digest=sha256:c8d2d8b83591a5848d4f50b82b31adbb55f2d1852826a67835fe93cfd4e07e5e

Observation 140f24d8-7ecf-4516-941f-a07656c5dd05 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Addressable Memory for Video World Models Wan: Open and Advanced Large-Scale Video Generative Models

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.977937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.977937Z digest=sha256:a8ed3b31c5ec4d495a84bf8506159d4b21f567d654c592c6221ae541c7d975d9

Observation a1dd92ba-2805-4c71-b625-635ae12b24d6 · outbound

This paper cites When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training.

Addressable Memory for Video World Models When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.980988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.980988Z digest=sha256:8e8f3b667712b0356781548a90a6ae6af4fa2e8b11669014b551cdb42edf9c5a

Observation 50d18257-3043-4ec0-a6c0-2fd075c40d31 · outbound

This paper cites Linformer: Self-Attention with Linear Complexity.

Addressable Memory for Video World Models Linformer: Self-Attention with Linear Complexity

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.983926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.983926Z digest=sha256:ee4e04cf010112760dd1d894c2b28e0db53776cd212427e507b646b1ed54d85e

Observation 44e33ebc-3b39-4e69-bb25-391ef2aca2a3 · outbound

This paper cites LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models.

Addressable Memory for Video World Models LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.987484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.987484Z digest=sha256:76deffb46c2017936cb8706b1ae48e48bb2e1f39127003e3de731d2cc1978d34

Observation 81426352-910e-48fa-85ad-b2c179e619bd · outbound

This paper cites Bovik, Hamid R.

Addressable Memory for Video World Models Bovik, Hamid R

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.991027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.991027Z digest=sha256:f5ea79a1313de17612e4a372652fe2f5e72e6bcf542856407dd13c5711ef7347

Observation c11f2c74-cb6f-4071-8458-ab4a9998adef · outbound

This paper cites Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory.

Addressable Memory for Video World Models Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.994602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.994602Z digest=sha256:5c35cd78a1f530163f7e3b0c566507b8775566abceb9063400e6cd01b8a3e4bb

Observation bf6b92ab-aa87-4a59-8ff8-96e0a87c7b1b · outbound

This paper cites VideoRoPE: What Makes for Good Video Rotary Position Embedding?.

Addressable Memory for Video World Models VideoRoPE: What Makes for Good Video Rotary Position Embedding?

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:51.998314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:51.998314Z digest=sha256:b9d229d6ba5590fb991cfed85d2766699d09bc942f734a7f23acfc9f3c4c701b

Observation 45019931-55ad-42dd-9581-576a360adc4c · outbound

This paper cites Video World Models with Long-term Spatial Memory.

Addressable Memory for Video World Models Video World Models with Long-term Spatial Memory

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:52.002085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:52.002085Z digest=sha256:6bb42a95adb4b02485c06d3737aea86de556fc1719a4f69e55b55a5865cb8e15

Observation 2a0e4fff-a6e0-4dda-822e-60fcaf2c95d5 · outbound

This paper cites Corgi: Cached memory guided video generation.

Addressable Memory for Video World Models Corgi: Cached memory guided video generation

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:52.005790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:52.005790Z digest=sha256:144f7cc61219f6f2530c18892112a6f20708fee32dc49f9cd4ee86b115faa8d4

Observation 7563504e-80c6-4111-a504-757d267ce84f · outbound

This paper cites Motion attribution for video generation.

Addressable Memory for Video World Models Motion attribution for video generation

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:52.009046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:52.009046Z digest=sha256:c092c3fbcd15c72b07d157ac023a849681a88ecf016fb7e0e532166ff80f5bc3

Observation 9dff48ed-dcf2-486a-93ee-30cedd49be66 · outbound

This paper cites Rabe, DeLesley Hutchins, and Christian Szegedy.

Addressable Memory for Video World Models Rabe, DeLesley Hutchins, and Christian Szegedy

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:52.012470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:52.012470Z digest=sha256:8b32e1d9ee86159834885930ef4ec18d898bb798748a822601fe9327f3644109

Observation fc0d617b-5107-4687-b585-3c4dbbe1e106 · outbound

This paper cites Efficient streaming language models with attention sinks.

Addressable Memory for Video World Models Efficient streaming language models with attention sinks

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:52.015881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:52.015881Z digest=sha256:f4a2003f17af8c8dca73888c46c9f832410ef96bbc2baede23b45125f993196b

Observation a7be016b-15e7-4e9c-a3dd-2d7e75e1965e · outbound

This paper cites DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads.

Addressable Memory for Video World Models DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:52.019067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:52.019067Z digest=sha256:0aad4a25bd43c684f739f25d8e276d7c292bab1557ffad7bf8495000ee6bbe1e

Observation b9d48c43-e88a-44e5-9b91-b8019f21f757 · outbound

This paper cites WorldMem: Long-term consistent world simulation with memory.

Addressable Memory for Video World Models WorldMem: Long-term consistent world simulation with memory

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-10T05:06:52.022607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:06:52.022607Z digest=sha256:0a49e885b81530a062039bb407f7a895cdabd30f99dde7b3208492e81bdfd376

Pith citing papers

No inbound Pith citation observations are available.