Pith. sign in

Paper Citation Record · LEDGER

FitVid: Overfitting in Pixel-Level Video Prediction

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2106.13195.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2106.13195 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:21:29.978421Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T11:38:04.583669Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6c18a44a-d8d0-43f6-ace2-7fb804c3be46 · inbound

Video Diffusion Models cites this paper.

Video Diffusion Models FitVid: Overfitting in Pixel-Level Video Prediction

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-13T14:38:27.960088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T14:38:27.919104Z digest=sha256:60ce665b3bba9807c68ef2e5ff751f1646c6e44e8bd5ca0d095d2a72951e9d84

Observation f482da99-14b8-4e1d-b1e9-5ca8a5b3665e · inbound

Imagen Video: High Definition Video Generation with Diffusion Models cites this paper.

Imagen Video: High Definition Video Generation with Diffusion Models FitVid: Overfitting in Pixel-Level Video Prediction

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:31:08.192650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T03:31:08.155347Z digest=sha256:1837d146c7bae45c0db724425aa4654024adcf4afd1649e9403d43b971f0a9ee

Observation aa2f687b-557d-4f0c-aa7b-9ba40bc185f9 · inbound

Phenaki: Variable Length Video Generation From Open Domain Textual Description cites this paper.

Phenaki: Variable Length Video Generation From Open Domain Textual Description FitVid: Overfitting in Pixel-Level Video Prediction

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:43:34.050645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T01:43:34.024375Z digest=sha256:6ffa936f65ce802767c268935ae5e62f8eb91854ab59224d2f64c47bdbbb13ff

Observation abce743b-f5bf-4dc1-b0f0-75f3068540f9 · inbound

MagicVideo: Efficient Video Generation With Latent Diffusion Models cites this paper.

MagicVideo: Efficient Video Generation With Latent Diffusion Models FitVid: Overfitting in Pixel-Level Video Prediction

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:47:57.426709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T18:47:57.322437Z digest=sha256:6b095412d76c4196ea2421a9a8b7ae295063311ae71e6abf9e11fc81aca2219c

Observation 69b97509-1433-41ef-a792-cb9db08b0e5a · inbound

Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation cites this paper.

Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation FitVid: Overfitting in Pixel-Level Video Prediction

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:06:44.737606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-13T20:06:44.480769Z digest=sha256:19ca58ee72920a6c2e87454b0e7bb927aa6f3e2d5c6b691c945412914dde1a91

Observation 3562646d-d0dc-4a2d-ad79-b43e8dc7ea2e · inbound

Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models cites this paper.

Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models FitVid: Overfitting in Pixel-Level Video Prediction

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:54:59.044933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T05:54:58.940428Z digest=sha256:a15ae021de2ba83352167634b458b683d4d45d3d508f458273b34e0ee4867141

Observation b7e0497a-368c-4635-ba60-af7f4c7e5131 · inbound

Cell as Point: One-Stage Framework for Efficient Cell Tracking cites this paper.

Cell as Point: One-Stage Framework for Efficient Cell Tracking FitVid: Overfitting in Pixel-Level Video Prediction

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:01.341460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:53:01.341460Z digest=sha256:b8a806c1ac36f566641a9e1f498b8dfd9427792ecf59c8aa320d4292292d288f

Observation 5ec36a95-9b70-4d7c-9087-5e50c676d990 · inbound

Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction cites this paper.

Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction FitVid: Overfitting in Pixel-Level Video Prediction

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T21:15:07.339610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:15:07.339610Z digest=sha256:0aae9ba693dc2f8dcfbd40aa51acde245cc1161397636122fc6ed6c9ab5e35b4

Observation 63bb87c6-876d-4857-82a6-4f2ff2273f44 · inbound

Efficient Continuous Video Flow Model for Video Prediction cites this paper.

Efficient Continuous Video Flow Model for Video Prediction FitVid: Overfitting in Pixel-Level Video Prediction

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:18.752749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:18.752749Z digest=sha256:6bf7b775a5cc9ba7b1bd0735ad3101791f131e1768a71fd07b0cb959eb76d557

Observation fe3c55c9-84b8-4521-ade6-8839b8b8fb9f · inbound

Long-Context Autoregressive Video Modeling with Next-Frame Prediction cites this paper.

Long-Context Autoregressive Video Modeling with Next-Frame Prediction FitVid: Overfitting in Pixel-Level Video Prediction

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:05:17.391798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T23:05:17.201790Z digest=sha256:15fecf25939bbce41fd9667ca747884e23fa438c6c5c91db68776a81c48ec20b

Observation c7aa28f1-f9b9-4140-a3fa-8389ad11989d · inbound

FlowDreamer: A RGB-D World Model with Flow-based Motion Representations for Robot Manipulation cites this paper.

FlowDreamer: A RGB-D World Model with Flow-based Motion Representations for Robot Manipulation FitVid: Overfitting in Pixel-Level Video Prediction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T21:21:29.978421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:21:29.978421Z digest=sha256:0fd55fdb3c290865ddc722f5713d0f6ec31b27d8be7502ff5b56f951373ae5be

Observation 25ada730-dc2c-4e25-add8-77e9bf2b3b62 · inbound

Inferring Dynamic Physical Properties from Video Foundation Models cites this paper.

Inferring Dynamic Physical Properties from Video Foundation Models FitVid: Overfitting in Pixel-Level Video Prediction

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:11:14.091739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T10:08:11.191706Z digest=sha256:759b840a04c40d21b6b080934edf7259d116d9da48a9f5afeedb97b08bb89345

Observation a2b64de2-c769-4c68-809f-fd99638de2cb · inbound

Geometry-Aware Single-Image 4D Synthesis via Dense Trajectory Generation cites this paper.

Geometry-Aware Single-Image 4D Synthesis via Dense Trajectory Generation FitVid: Overfitting in Pixel-Level Video Prediction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T18:30:01.052327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:30:01.052327Z digest=sha256:b010f6988ee88e48f9740f1673c3ba09190e6b42383f7d144d2a243efd129b20

Observation 796d7e21-43de-40bb-9d2e-c90dd14f03bd · inbound

Stochastic Lifting for Generating Trajectories of Stochastic Physical Systems cites this paper.

Stochastic Lifting for Generating Trajectories of Stochastic Physical Systems FitVid: Overfitting in Pixel-Level Video Prediction

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:53:15.996121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T08:48:25.141614Z digest=sha256:1963c17f2785f13be1aab93465e84ca55fa427574f3eeebaeacf1e7ba63ba0d6

Observation 04a51190-0d50-4a96-97ca-a72f17930270 · inbound

Bridge-WA: Predicting Where and How the World Changes for Robotic Action cites this paper.

Bridge-WA: Predicting Where and How the World Changes for Robotic Action FitVid: Overfitting in Pixel-Level Video Prediction

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-03T11:38:04.585207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-03T11:28:50.286894Z digest=sha256:2865d145bcd8c42603f427ca1c21cea0941adc764dfe0eb50e53d0aa4fcce335

Observation e47c69d9-50e8-4966-9a5c-3e5cf2c231d8 · inbound

Quo Vadis, World Modeling? cites this paper.

Quo Vadis, World Modeling? FitVid: Overfitting in Pixel-Level Video Prediction

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:13:50.539483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:13:50.539483Z digest=sha256:41215aa6966372af658eb00916bf3ce8f2e4b24cdd27463387c5a1885d2c0131

Observation e63d17cc-ab61-4758-b102-ec7bdf1bde34 · inbound

Overcoming Statistical Bias in Action-Controllable World Models cites this paper.

Overcoming Statistical Bias in Action-Controllable World Models FitVid: Overfitting in Pixel-Level Video Prediction

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T19:49:02.076976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:49:02.076976Z digest=sha256:ef4cbc9044bcf4c4bb239ebe390540628bcd8a6d6cc18419a7b8612f7d56ffe2