Pith. sign in

Paper Citation Record · LEDGER

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity

As of 12 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 6 inbound Pith citation observations for arXiv:2412.09856.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.09856 v2

Coverage vector

measured 76 of 76 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T16:43:08.574966Z

measured 82 of 82 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:03:20.460990Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T07:27:08.875128Z

Reference resolution

76 of 76 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 67be81be-11f0-4b4e-b2e0-c5f26f9a2a65 · outbound

This paper cites Video generation models as world simu- lators.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Video generation models as world simu- lators

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.974969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.180513Z digest=sha256:0f67bc41517dfff96dc203fbad534423271740a23a4f2112dd776fcd78f3d91e

Observation 4546eec7-2c32-4ebc-8dde-fd71c16bff2a · outbound

This paper cites VideoCrafter1: Open Diffusion Models for High-Quality Video Generation.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.187386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.187386Z digest=sha256:743cabe8151ff14ebe1a0b10bb6375a493be2aaceeb75652ef731b25b3d6a08d

Observation df97383c-259c-4653-9fba-4ce95c81c15d · outbound

This paper cites VideoCrafter2: Overcoming data limitations for high-quality video diffu- sion models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity VideoCrafter2: Overcoming data limitations for high-quality video diffu- sion models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.959193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.192920Z digest=sha256:f43dd8097f8ba425a36430db7c0b463c90eb2102e046737843909bbd4fed3e09

Observation 8f8f8991-fcd0-4764-9516-13c1a0eae939 · outbound

This paper cites DiffEdit: Diffusion-based semantic image editing with mask guidance.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity DiffEdit: Diffusion-based semantic image editing with mask guidance

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.198731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.198731Z digest=sha256:7b8241aebe2d33b489f1cba7d657fb05cdbbf280cf44fbab9ebb91f16a2efe86

Observation 3fa3e2dc-c0a3-451d-ac0c-e511652471a0 · outbound

This paper cites Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.204201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.204201Z digest=sha256:3ffc2dc555aea59f1c1afbb17a9c2cfc7b54aac2fafc49ad6ac7e03c55d0d8bb

Observation 1a6c9d8d-052b-4f2b-82a6-6806e2bd1ac3 · outbound

This paper cites Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.210442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.210442Z digest=sha256:f8350b6b4296bbecc53082c890b0b9dd883bc9ac7d6bf08f0ce4b7e9c7d730d3

Observation 434e6069-d101-43fc-a7de-3ca178f17454 · outbound

This paper cites FlashAttention: Fast and memory-efficient exact attention with IO-awareness.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity FlashAttention: Fast and memory-efficient exact attention with IO-awareness

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.941900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.216087Z digest=sha256:d15f5fd9a29e488d2662c2c319aeafbe727a36a69091f9fb050c1fe76c532d6b

Observation f707be84-4ea5-4a18-8d5c-5e57d64d9060 · outbound

This paper cites Alabdul- mohsin, et al.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Alabdul- mohsin, et al

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.926042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.221900Z digest=sha256:74a4f41154cde5c93262b69e999b420e2d56ed667e4cf6d7b45be3df0b978aee

Observation 98cbeb64-83b5-48aa-82dc-8039ff9a7c47 · outbound

This paper cites The Llama 3 Herd of Models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.228074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.228074Z digest=sha256:3365bab60e1e82dff5b38855fb2262223416d9ed351d135f2b9377719252030b

Observation f72e4609-bef2-4b9f-a8f0-2769afa8be51 · outbound

This paper cites Matten: Video Generation with Mamba-Attention.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Matten: Video Generation with Mamba-Attention

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.233267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.233267Z digest=sha256:1a1fa0d471e1741dd43c5fad538e9496b5651dfd79cc622cea6152bb939078fc

Observation e1c7108d-86e9-429d-a487-cdda0b005945 · outbound

This paper cites Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.238185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.238185Z digest=sha256:a4bdb15c85793db8d4e2d2e6fde4275951aef3573890ebb9110f7323a4791ffa

Observation a39fe9b4-09d7-466f-aa97-d98a4220b74b · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.245160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.245160Z digest=sha256:d2800406d7cf739ead103f568dfde46d7434ae8638468767fbdc798607575a6a

Observation 87fb6a4b-1dda-4eb5-b2e8-90a140483ca2 · outbound

This paper cites HiPPO: Recurrent memory with optimal polyno- mial projections.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity HiPPO: Recurrent memory with optimal polyno- mial projections

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.910237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.250778Z digest=sha256:9ebbd7a410553d5994eb4aa1a0cb5f1ab64623759e5a604c7cd989cdacb99920

Observation 1f31fdff-1860-4e54-b596-3d9f364712c7 · outbound

This paper cites Efficiently Modeling Long Sequences with Structured State Spaces.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Efficiently Modeling Long Sequences with Structured State Spaces

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.255538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.255538Z digest=sha256:4e4ebebd0ea9118cb5f44367abdeb27325c199e3039889a657028b093012b8f6

Observation 8016844c-8074-43e6-ac6c-97465382bbf7 · outbound

This paper cites MambaAD: Exploring State Space Models for Multi-class Unsupervised Anomaly Detection.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity MambaAD: Exploring State Space Models for Multi-class Unsupervised Anomaly Detection

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.260634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.260634Z digest=sha256:7f36693825ecbcb94de087742ad07e3569bf2348e5945812d52cacc2f2bcb31c

Observation 69c39173-1b98-43c4-bfcd-3fc74f86e966 · outbound

This paper cites Denoising diffu- sion probabilistic models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Denoising diffu- sion probabilistic models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.265594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.265594Z digest=sha256:f78cc3f234c819559878ec5e399f9c25aed557978515e0556c158bb98b40073d

Observation 73af2afb-1f10-42fb-abe2-0fb5db852272 · outbound

This paper cites Imagen Video: High Definition Video Generation with Diffusion Models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Imagen Video: High Definition Video Generation with Diffusion Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.270287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.270287Z digest=sha256:30b9271a81c3c641a7876373aa18b83b99e6434e1be5ca81562b7d840c24d7d3

Observation 5584c0cd-a842-4134-932b-2bfe21d80aea · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.276151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.276151Z digest=sha256:18b14f2a037e62f592f012f784cd5b6d2a504311df2da9c0487228b0d8e06944

Observation b87f8b7d-304b-4486-a080-f9a33cad1989 · outbound

This paper cites ZigMa: A DiT-style Zigzag Mamba Diffusion Model.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity ZigMa: A DiT-style Zigzag Mamba Diffusion Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.281382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.281382Z digest=sha256:8cd8f6b822840d0700d3fb1b2a34e197e058f3ae3fd99702e65b6ac39cc4ffc7

Observation 93a6088d-2350-4891-84b1-d660d273be6d · outbound

This paper cites VBench: Com- prehensive benchmark suite for video generative models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity VBench: Com- prehensive benchmark suite for video generative models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.286787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.286787Z digest=sha256:08ee16ea45397cc3298a6b5ba1ae153f2acaeac220a3cd45a8ed88a4a24f8de5

Observation ebc306c6-34c7-4484-8efc-8ab750767b79 · outbound

This paper cites MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.291243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.291243Z digest=sha256:56cd2ac64eb050b82e3f448cc704c7ce72042c2350cf40a82fb9b9b174a005d5

Observation 304a5e6a-d94e-4091-ac81-b411f442c56b · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Imagic: Text-based real image editing with diffusion models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.295497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.295497Z digest=sha256:72a096cb3e23a9663f395d3d5be73f904b9b3e5fdcea125bb3cae4122dc80a82

Observation c8adbf2d-ef64-428e-be91-d8aabb0f3530 · outbound

This paper cites BK-SDM: A Lightweight, Fast, and Cheap Version of Stable Diffusion.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity BK-SDM: A Lightweight, Fast, and Cheap Version of Stable Diffusion

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.300020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.300020Z digest=sha256:3848e0e29aad80eaa91ffff6aa78eeb43dd4a34c9c8ce90005bc59851d024028

Observation 4338dae4-dca8-43f4-9558-30ae7b5529c1 · outbound

This paper cites Kling AI: Next-generation AI creative studio.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Kling AI: Next-generation AI creative studio

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.861129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.305666Z digest=sha256:5486a3d50d183f5abc529e7b427ab512ded9de4d51c050f53cb1b20f881a37a8

Observation 096d883a-35b0-4f2e-ae7c-ebe3d2b8644d · outbound

This paper cites VideoPoet: A Large Language Model for Zero-Shot Video Generation.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity VideoPoet: A Large Language Model for Zero-Shot Video Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.311136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.311136Z digest=sha256:debe711cb9e4a0fa65b22bf71a145c68fbf5eb241a7147c6254ff7112f257a5e

Observation 2d3acca8-e53b-4e2b-8c1e-65fb9a39e256 · outbound

This paper cites Pika labs.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Pika labs

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.843977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.317114Z digest=sha256:d7704792e9f5d2e0af01631f7759a675ba9d2f256d8de5025eaf965ff85daacc

Observation dced3c44-002a-4e03-9d1e-b8d38aaf19b1 · outbound

This paper cites xFormers: A modular and hackable trans- former modelling library.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity xFormers: A modular and hackable trans- former modelling library

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.826627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.322165Z digest=sha256:233822e3e0afe2bf8f6c783446c4c93897a4b5e291e1825aa05f8c5b5dbcd6b8

Observation 49a8df5e-1cf7-4675-a3cd-d7887bad7f70 · outbound

This paper cites T2V-Turbo: Breaking the Quality Bottleneck of Video Consistency Model with Mixed Reward Feedback.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity T2V-Turbo: Breaking the Quality Bottleneck of Video Consistency Model with Mixed Reward Feedback

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.327092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.327092Z digest=sha256:a0ab41565429ead3b05fb95bda21af635ed346a8a9fc2f5af91f23ba944577b2

Observation e5c55d33-5deb-4072-9510-1050b5d2165c · outbound

This paper cites T2V- Turbo-v2: Enhancing video generation model post-training through data, reward, and conditional guidance design.arXiv preprint arXiv:2410.05677, 2024.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity T2V- Turbo-v2: Enhancing video generation model post-training through data, reward, and conditional guidance design.arXiv preprint arXiv:2410.05677, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.332179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.332179Z digest=sha256:0ab57aa66573c27c3c8042720b9deb23db13d2be6e5f47d6681fd298c5e7be7e

Observation 9d4794f3-5e65-4028-a7b2-3285f26d365e · outbound

This paper cites Flow Matching for Generative Modeling.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Flow Matching for Generative Modeling

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.337755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.337755Z digest=sha256:15a73330e9e43bdfe4f7dd7b2ed79d8cbf31a3d4143364b174528529161cb1c0

Observation 6c9fc96a-32fe-40db-b94b-5173b6635299 · outbound

This paper cites Swin Transformer: Hierarchical vision transformer using shifted windows.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Swin Transformer: Hierarchical vision transformer using shifted windows

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.811957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.342684Z digest=sha256:2727d400bcc053388faa76ac178ebad4f2d8f7d292e74bbea4e108132800bee5

Observation aba8da3d-3592-41ea-862f-9e30436b6564 · outbound

This paper cites VDT: General-purpose Video Diffusion Transformers via Mask Modeling.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity VDT: General-purpose Video Diffusion Transformers via Mask Modeling

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.347341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.347341Z digest=sha256:1d5a42f65c10234339dfaef2919d7a462fb6fee199b3c8453779ad152e52cf69

Observation 52824075-8794-4904-a44b-b57723da92ab · outbound

This paper cites Dream machine.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Dream machine

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.796124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.351882Z digest=sha256:b00c9ffc3e123678b6253d2ff4657da4dafdf3069a4cc657660ab33f54973a9d

Observation 93f206e9-c541-4573-9808-e797e531f9dc · outbound

This paper cites Diffusion probabilistic models for 3D point cloud generation.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Diffusion probabilistic models for 3D point cloud generation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.780129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.356402Z digest=sha256:4ffd9cab2e42cfc5e247fdcc9d33eb3e9ee05c60b13115b0287e5211f83fd093

Observation 45ec4db2-4e09-4354-9913-ae343f0d2ee1 · outbound

This paper cites Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.362974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.362974Z digest=sha256:bf73c9a9916b67f3c93aba636b0760571fc3fa58dcd8300d22fef96ea77200fc

Observation 9f108b34-3684-43b0-9e5b-1acf437d6dd6 · outbound

This paper cites On distillation of guided diffusion models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity On distillation of guided diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.368317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.368317Z digest=sha256:ba68bc4b84b8b24cdeebc037770a6d8f2c948d1f509a7367c233e420ae8d1610

Observation 3271f270-fff0-492f-893e-ee838cc80e68 · outbound

This paper cites Scaling Diffusion Mamba with Bidirectional SSMs for Efficient Image and Video Generation.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Scaling Diffusion Mamba with Bidirectional SSMs for Efficient Image and Video Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.374170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.374170Z digest=sha256:64f6972ea4a6ba18079f5789e9ecee6dbe200eafc08b644fd1546ae3ba74fca5

Observation 1d66c585-fee5-45db-bbc1-d996606ef9bb · outbound

This paper cites Transframer: Arbitrary Frame Prediction with Generative Models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Transframer: Arbitrary Frame Prediction with Generative Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.379341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.379341Z digest=sha256:055b420b706f2b57bbcf459a27fb2c55b225533d00da0f663d91c7162c16b2cd

Observation 1fd1fe16-4527-444e-9c70-efd6d6929ca6 · outbound

This paper cites Scalable diffusion models with transformers.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Scalable diffusion models with transformers

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.384349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.384349Z digest=sha256:b70126e56770fbe10e7379768af2c4babd769eec0098a0f196290f92cee3acf1

Observation 93b99cae-553e-4607-b021-cf266fc24959 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.389315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.389315Z digest=sha256:a489b68c0bc86ce154427ce3e6ba3e5bdc34d6a450d0d2582ddae1005c2aae34

Observation 9a67080f-8ff0-4101-b294-704c6b458a16 · outbound

This paper cites Movie Gen: A Cast of Media Foundation Models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Movie Gen: A Cast of Media Foundation Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.395562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.395562Z digest=sha256:80fd63a9b63ced09f2a1eedfd95664b37541e6c79f6092293548c83d8abfb40b

Observation 94240f2a-ecc1-4110-b093-b3a8473e482f · outbound

This paper cites RawFilm: 8k cinematic royalty-free stock footage.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity RawFilm: 8k cinematic royalty-free stock footage

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.739126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.401196Z digest=sha256:d72ea0ae8adb24e7c5163aeb482608cdbb5c7d010650ca560d5447d2c6661556

Observation 394572c4-2cf3-4c0e-b9c2-57989059a4f5 · outbound

This paper cites Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language Modeling.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language Modeling

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.406817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.406817Z digest=sha256:b4daa5b631d1cdd9e3513c3b48d699aecce4b32a2d79020cf55b2c9ea2919c1a

Observation a9140747-2199-4b26-803c-511bd120043f · outbound

This paper cites MambaCSR: Dual-Interleaved Scanning for Compressed Image Super-Resolution With SSMs.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity MambaCSR: Dual-Interleaved Scanning for Compressed Image Super-Resolution With SSMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.411772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.411772Z digest=sha256:2996acce65626c614d2a58c3b1d0daf44faa43f494fc8613f9ff1ee2115c152b

Observation 2761b435-eb27-4d91-a71d-b347209851f5 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity High-resolution image synthesis with latent diffusion models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.417340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.417340Z digest=sha256:a1c0c06429c46ed503a84bc2a65fb1a9a50b40654464685f1076613e00a69473

Observation 6c18eb9f-9ee1-45aa-bc60-eea27b3554e3 · outbound

This paper cites Introducing Gen-3 alpha.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Introducing Gen-3 alpha

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.713492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.421498Z digest=sha256:2d7ca29484d8f9d437875e8b32b63307a0b6735b7de97ee375ac36c253610819

Observation 79c12c70-83fa-4901-885e-a794f0a3b58a · outbound

This paper cites Denton, Kamyar Ghasemipour, Raphael Gontijo Lopes, Burcu Karagol Ayan, Tim Salimans, et al.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Denton, Kamyar Ghasemipour, Raphael Gontijo Lopes, Burcu Karagol Ayan, Tim Salimans, et al

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.697332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.426432Z digest=sha256:f49ef11da985edad7d123394fa3a1c9c1f13c15fccd92a3330c9b29a41ef7ca1

Observation 3a31a639-25df-401b-bf69-137bad6fdfad · outbound

This paper cites GLU Variants Improve Transformer.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity GLU Variants Improve Transformer

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.430805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.430805Z digest=sha256:6b4a56a67fd05971e1e2bed2fe287c0512543c0ec80d60b3760d6ee2d7c685ce

Observation aa30885d-27f4-41ed-8eb5-11bb5c6c12ad · outbound

This paper cites Emu Edit: Precise image editing via recognition and gen- eration tasks.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Emu Edit: Precise image editing via recognition and gen- eration tasks

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.680306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.435207Z digest=sha256:76504815f1d0f24903de78937fd6a8a8281131426f0e4e93b16e0b8a39cdd436

Observation ce1f816d-4e84-412c-ba00-2b1ab475e770 · outbound

This paper cites NormFormer: Improved Transformer Pretraining with Extra Normalization.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity NormFormer: Improved Transformer Pretraining with Extra Normalization

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.440239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.440239Z digest=sha256:5c5e1af0bc0d72d76ca6de55501f95333bd1fb072fe0dc3658692ca7ff3b8a6b

Observation 64de6b84-dbc5-4660-b28d-ec1d84c83d1e · outbound

This paper cites Shutterstock: Stock photos, royalty-free images, graphics, vectors, videos, and music.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Shutterstock: Stock photos, royalty-free images, graphics, vectors, videos, and music

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.664288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.445562Z digest=sha256:cbad595b62d9526e60f4c0fca08d6558c89dc002d226b733828abc132f121bf5

Observation cfdc7c31-2fd1-45d8-9e8f-c630d5920afc · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.450278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.450278Z digest=sha256:36c6a845992a4ec9e1963f0c585334f32dda08f30c32862a109309bcaec6c40b

Observation 5b0c444a-68c3-4be6-b776-e0ef90e36fed · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Deep unsupervised learning using nonequilibrium thermodynamics

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.647984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.455265Z digest=sha256:409b9337a90c28255b711885b7c434f2c00a3e4c9fef11785776059364c188bb

Observation 66110414-57a4-40c8-851e-b00797916aa4 · outbound

This paper cites Consistency Models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Consistency Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.461195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.461195Z digest=sha256:fc06cc3efbd58d9d33057f0ad60dfb01d7b7d17a8a3bc7401755ff2dff2ce7ff

Observation 9b74f8bf-1982-4a81-ab30-b09496863bc0 · outbound

This paper cites UL2: Unifying Language Learning Paradigms.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity UL2: Unifying Language Learning Paradigms

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.467829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.467829Z digest=sha256:0bf1b976e3a4073a6baeba3f4c649ad748f5e6f0c62b859ef9f12746c628f883

Observation 4ccbca24-6cc3-449e-9468-ba4b5f26a920 · outbound

This paper cites an unresolved cited work.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-11T16:43:09.630832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.473388Z digest=sha256:716eef36aa5b1de8a77fc96a5f6b058fb1ba403be5d0e0987652261a8377b3e1

Observation 6edfc395-e805-48a7-9f56-300d4160948b · outbound

This paper cites DiM: Diffusion Mamba for Efficient High-Resolution Image Synthesis.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity DiM: Diffusion Mamba for Efficient High-Resolution Image Synthesis

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.478362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.478362Z digest=sha256:9a30741d1977852bc96557b672948889736bcfc98839fc79a288ba5a3592b140

Observation 93cf7088-e7da-40bf-9ff8-692f012afd68 · outbound

This paper cites LION: Latent point dif- fusion models for 3D shape generation.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity LION: Latent point dif- fusion models for 3D shape generation

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.613422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.484763Z digest=sha256:67b566dc0bbc8f2d0caf677da20255778ba5565120ee03373123b67fd52cb933

Observation ff7a6544-0c04-43ad-a68e-435c6c98d6f4 · outbound

This paper cites Phenaki: Variable length video generation from open domain textual descriptions.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Phenaki: Variable length video generation from open domain textual descriptions

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.489615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.489615Z digest=sha256:e8109822fca5aff7645db19ef964d0ae97059756f1c14243e599355595770c5a

Observation 5e2ed4d1-b2d8-4131-ad92-1a0a21c701cc · outbound

This paper cites An Empirical Study of Mamba-based Language Models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity An Empirical Study of Mamba-based Language Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.494872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.494872Z digest=sha256:2e353070c5885826dacf5bc559cbaee2cbb0508d2807502eba917d3bb08f35b0

Observation 50e6d2a5-ff15-4a22-9ce9-9b479f9317e5 · outbound

This paper cites Jha, and Yuchen Liu.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Jha, and Yuchen Liu

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.587749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.499943Z digest=sha256:90e249eaff62cba9be76b5fea19b18f3f7929c6aba1aae9c555bcbd88b68dc24

Observation 1f44977b-4c10-4806-8cf6-d2f2b9a797bd · outbound

This paper cites ModelScope Text-to-Video Technical Report.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity ModelScope Text-to-Video Technical Report

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.504673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.504673Z digest=sha256:0cf5d7d24248238ce46e6d9e5576cc002c6035e89e16c4f7090c4824260d73ba

Observation b87730bf-0482-421d-9086-c9aec0aed188 · outbound

This paper cites VideoLCM: Video Latent Consistency Model.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity VideoLCM: Video Latent Consistency Model

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.509845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.509845Z digest=sha256:21e2f67c42e8735703384d79479a9e8c19588373e62234799c8d2b76869b1eec

Observation fc670c36-7bbe-4b30-b704-235d62799298 · outbound

This paper cites LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.514912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.514912Z digest=sha256:71836449be9ddc1422a479fd28f3fa2951b31391c2734142ce00874e7604469b

Observation 13f00a1f-ca3e-4c96-9bb6-99320135d3d4 · outbound

This paper cites Loong: Generating Minute-level Long Videos with Autoregressive Language Models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Loong: Generating Minute-level Long Videos with Autoregressive Language Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.519752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.519752Z digest=sha256:910ac0c299197b79235e1cdcc1e27a1bc26fabaf39d9e033be646b6a8f69fdef

Observation 986da5f9-b3bf-4441-a306-b32c54140b37 · outbound

This paper cites Progressive Autoregressive Video Diffusion Models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Progressive Autoregressive Video Diffusion Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.525261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.525261Z digest=sha256:d0d15dc567147dcdaa40560fdf369c6dc5c0181c6ac1e2acb8ee6ea9b567d409

Observation 30536006-04cb-4964-9e3c-614def0e3e23 · outbound

This paper cites Demystifying CLIP Data.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Demystifying CLIP Data

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.531267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.531267Z digest=sha256:bf72444329b0cb93fd1b96a0410f530b9c281ab6b6e9b6587c72af634c1243e2

Observation c4ec7859-8615-4d98-a3d8-131b31574633 · outbound

This paper cites ByT5: Towards a token-free future with pre-trained byte- to-byte models.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity ByT5: Towards a token-free future with pre-trained byte- to-byte models

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.573235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.536061Z digest=sha256:a8d2e95ffa0c7d1daadc9c589731c34d78ab601682f6c98884968a2decc824db

Observation b5b75df8-c35b-4007-8684-bcf332f3247a · outbound

This paper cites Dif- fusion models without attention.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Dif- fusion models without attention

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.558055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.541049Z digest=sha256:8758fd4e9726fd0c10b6a3f410b73de7a52a4bdc42e2abcacd30d7db85e4f618

Observation a3c68c58-0f51-48e8-9131-ed880c767a6d · outbound

This paper cites VideoGPT: Video Generation using VQ-VAE and Transformers.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity VideoGPT: Video Generation using VQ-VAE and Transformers

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.545220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.545220Z digest=sha256:6f2719386ed69ebd63e3e75cff76b2e038a8deab4d5900fb8e72af67cc7b7f76

Observation bfa55274-aaff-4222-8c16-acdcc4db6dd9 · outbound

This paper cites Paint by example: Exemplar-based image editing with diffusion mod- els.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Paint by example: Exemplar-based image editing with diffusion mod- els

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.550153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.550153Z digest=sha256:d42c94c05003ee1210bf019ddd70e83cd128eda7bcfb511eec9195415d4684ef

Observation f16bd07c-6c7e-4ddc-89cb-47d5a430be8b · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.554893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.554893Z digest=sha256:180d4b98f0f31051b26fa859967d09f2b5b7cc0b5a51bcc259d32458299a9b99

Observation bbb8c895-b74a-43b8-a0dd-8d11a7d08921 · outbound

This paper cites Hauptmann, Ming- Hsuan Yang, Yuan Hao, Irfan Essa, et al.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Hauptmann, Ming- Hsuan Yang, Yuan Hao, Irfan Essa, et al

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.531284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.559795Z digest=sha256:7119dc814a2553869c9d5688b08ab19178b9acf690a2f972525f41ba73ab97a4

Observation e90ea079-7759-4fd6-9fc4-a3250ced6a98 · outbound

This paper cites Root mean square layer nor- malization.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Root mean square layer nor- malization

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.515407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.565377Z digest=sha256:97b3ad7b3d0c4dc447a34306aa2c153b444875bb9e552b7d3d78127502a411eb

Observation d4003444-edac-41d9-9900-71d939834724 · outbound

This paper cites Show-1: Marrying pixel and latent diffusion models for text-to-video generation.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Show-1: Marrying pixel and latent diffusion models for text-to-video generation

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.499522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.570224Z digest=sha256:1def16c023d116b32216a15f65e3377a26c62f1d301d2f6f60d533a7ff7d4295

Observation 5bb90295-a349-43af-bedc-e0295fb25444 · outbound

This paper cites Open-Sora: Democratizing efficient video production for all.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity Open-Sora: Democratizing efficient video production for all

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:43:09.483317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T16:43:08.574966Z digest=sha256:a92ba4f7c1b56785556159cea090008205e83cfa75773c62e43483244328f2ba

Pith citing papers

Observation 0b237622-2fa2-45f5-8386-99bd91f8cf11 · inbound

Long-Context State-Space Video World Models cites this paper.

Long-Context State-Space Video World Models LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T14:03:20.460990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:03:20.460990Z digest=sha256:9fbb06af11f79c6bf38be57397a8b8acec7c29b932ef8747d7bc07467854e650

Observation 0c076b86-9971-4220-859b-86de36be2480 · inbound

Video World Models with Long-term Spatial Memory cites this paper.

Video World Models with Long-term Spatial Memory LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:41.279115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:41.279115Z digest=sha256:6dfda80d379f0b360c61b42b85a4ef18d7195a98b6fc287b398da92e468e6a5d

Observation de71e09f-eb27-4bd3-bf72-0227ef3f457e · inbound

Exploring Diffusion Transformer Designs via Grafting cites this paper.

Exploring Diffusion Transformer Designs via Grafting LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:04.156409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:29:04.156409Z digest=sha256:6e5d797c7645ee54d813afe988eff6532243cf77d620c82c101c24084addb441

Observation a8fe1376-17c3-4580-a606-0c671997fc7a · inbound

M4V: Multimodal Mamba for Efficient Text-to-Video Generation cites this paper.

M4V: Multimodal Mamba for Efficient Text-to-Video Generation LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:21:19.044277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:21:19.044277Z digest=sha256:a754cf25bdc3da83fa180c35d449031c967db1ba332854250f0993d6464c23ec

Observation d2d1e529-6b84-4fb0-92c2-28dede2602f8 · inbound

GenHSI: Controllable Generation of Human-Scene Interaction Videos cites this paper.

GenHSI: Controllable Generation of Human-Scene Interaction Videos LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:27:08.877210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T07:25:41.219753Z digest=sha256:59ccaafe092d893d7ab13ef7d352619fa9fc8b52855f914d726752b479963037

Observation 17f7f2af-7ba3-415a-84f0-f6c164e11758 · inbound

VMoBA: Mixture-of-Block Attention for Video Diffusion Models cites this paper.

VMoBA: Mixture-of-Block Attention for Video Diffusion Models LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:04.914075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:04.914075Z digest=sha256:e0de79834259f09990bae9600ab9cbbf656c9bacfc66a063a1f8e07e16b53b0b