Pith. sign in

Paper Citation Record · LEDGER

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:2501.12375.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.12375 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:42:05.975670Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T20:00:07.459591Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 56e8e8b7-e70e-44c4-b103-8c88dfbc9e82 · inbound

DissolveStereo: Coarse Depth Injection for Zero-Shot Stereo Video Generation cites this paper.

DissolveStereo: Coarse Depth Injection for Zero-Shot Stereo Video Generation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:18:13.976065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T17:17:48.961672Z digest=sha256:7a68c11c416ad0737ce3ecb1b952e7c315770156952094e3388496ebf3ed1064

Observation 8a153380-aa41-472a-b42f-d93b78d6ce3f · inbound

Seeing World Dynamics in a Nutshell cites this paper.

Seeing World Dynamics in a Nutshell Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T04:42:05.975670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:42:05.975670Z digest=sha256:d80ce5040359979e9b08071c565bad9d8b4db4fbc999b7300a2af6d74ea10936

Observation 677d1a65-46d6-4a29-8c90-a851b74eb0e9 · inbound

UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation cites this paper.

UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:26:57.495990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:26:57.495990Z digest=sha256:c2f7b39e525f6df3785753ceb1b8722499f466fd94a72a989903fcae9c58f141

Observation 02a6ffb5-61ac-4584-91e1-58d98f7acf98 · inbound

E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models cites this paper.

E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:36:06.197375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:36:06.197375Z digest=sha256:c94728bfa727f5e96df8d8d6e5cc6d38242cf40d909ef82d7aa915d7c5e6dcc0

Observation 53500216-ec5f-460f-b62c-3f9bbeec206c · inbound

IllumiCraft: Unified Geometry and Illumination Diffusion for Controllable Video Generation cites this paper.

IllumiCraft: Unified Geometry and Illumination Diffusion for Controllable Video Generation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:43.647637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:43.647637Z digest=sha256:08a993e119baaef32d53e882d19b568fca0549062b7c89e599a6eacbe08acde0

Observation 118f8165-5b8a-471a-a0ce-a0ed023223e8 · inbound

RaCalNet: Radar Calibration Network for Sparse-Supervised Metric Depth Estimation cites this paper.

RaCalNet: Radar Calibration Network for Sparse-Supervised Metric Depth Estimation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T23:58:41.807576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:58:41.807576Z digest=sha256:4b5b4b4f92662f0d4c2f058008bd9eaadc8ae022833ebfcf6af0ececbce7d743

Observation ddf523f2-5879-4fe0-9116-f32619e7454c · inbound

RoboScape: Physics-informed Embodied World Model cites this paper.

RoboScape: Physics-informed Embodied World Model Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:56:09.428813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:56:09.428813Z digest=sha256:ce60ec278ff5108ce14327d24f8da4329ddb1de6481f95c01c8740f70a00dbc6

Observation ecf4f2db-7073-4213-beb6-daefed166cac · inbound

MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion cites this paper.

MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:06.939745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:06.939745Z digest=sha256:58cdafd601bd8fb19f5e9936864363d26a478c7b5920e92aa0e912c5d2afa42a

Observation 98217f03-2f15-440f-9424-261dcd1fb3f8 · inbound

RCG: Safety-Critical Scenario Generation for Robust Autonomous Driving via Real-World Crash Grounding cites this paper.

RCG: Safety-Critical Scenario Generation for Robust Autonomous Driving via Real-World Crash Grounding Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T17:34:05.903479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:34:05.903479Z digest=sha256:62200c00e278d6741dc11fd13634dec2856c9fc4391cb0f51d33b3e97933019e

Observation c3c5ea58-a021-4c05-b163-8065c0f45e55 · inbound

SpatialTrackerV2: 3D Point Tracking Made Easy cites this paper.

SpatialTrackerV2: 3D Point Tracking Made Easy Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:49:48.565573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:49:48.565573Z digest=sha256:e36c9bae6fc1f526a570bb4ac7c783f1eba0140456a8245f0e7033a4ee5150f3

Observation 96348a5e-db49-4c7b-86cc-e3b688b5be64 · inbound

Reconstructing 4D Spatial Intelligence: A Survey cites this paper.

Reconstructing 4D Spatial Intelligence: A Survey Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 135

Resolution
unresolved
no resolver link, observed 2026-08-06T13:02:29.086195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:02:29.086195Z digest=sha256:a0c5a97a5fef9d298b8fee5191b04d9d45e5a3b316ff13fa9f69134fa4f9a029

Observation 3a193b00-fc7b-44d1-a0ab-53127283cb3e · inbound

ViPE: Video Pose Engine for 3D Geometric Perception cites this paper.

ViPE: Video Pose Engine for 3D Geometric Perception Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.727344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:45e92282cee9b6cd51d828002389b4f40c5c3ec7f8bb1a347822c28bdff7c8de

Observation c11040f0-350d-4e24-a23f-8f6aa3498ec0 · inbound

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation cites this paper.

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T13:46:44.930591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:46:44.930591Z digest=sha256:ee8db8b3185ccb4037b650f03831d0fe30b494d8b1d53e26af804704160e169e

Observation 96c1b55d-adfa-44e5-9ef0-2d72c5c1f332 · inbound

Feedback Matters: Augmenting Autonomous Dissection with Visual and Topological Feedback cites this paper.

Feedback Matters: Augmenting Autonomous Dissection with Visual and Topological Feedback Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T11:35:58.419641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:35:58.419641Z digest=sha256:9607d18318fbce0810a82d9878af237b05d8df3da56b554b0e0fba2ad50a1bd9

Observation 6be966d9-301c-4180-901d-c6d8cd0236ce · inbound

CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction cites this paper.

CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:38:37.596051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T22:38:01.008280Z digest=sha256:63f40c329d4f587f4e6b077b1177f249b00fa1bb41161c12634e906a745873a8

Observation e0f3cdf7-c86a-49ef-a33e-0492f7d05222 · inbound

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation cites this paper.

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T22:13:53.383917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:13:53.383917Z digest=sha256:43fda1296741991b47f3853c648f60a13ce982ef5310871bbf147a1dac4bcf22

Observation f87dda88-df7c-41f9-ba6e-74ead34e61c1 · inbound

HOIGS: Human-Object Interaction Gaussian Splatting cites this paper.

HOIGS: Human-Object Interaction Gaussian Splatting Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T17:28:02.544902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T17:24:50.043908Z digest=sha256:df0022ac20db5bc275dd9df441358fcbd0e44ea5247028c34bc80a5135370c86

Observation 2bb3e8db-2aea-477e-b6f7-934d607c3940 · inbound

From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data cites this paper.

From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:03:01.206201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T17:02:18.358675Z digest=sha256:b0b75a1f700010457724d5f218c83e462dfa53f31c24ab822642466fe8202aed

Observation fbd445bf-0125-4f54-bbcb-db9b20b08ad8 · inbound

LuMon: A Comprehensive Benchmark and Development Suite with Novel Datasets for Lunar Monocular Depth Estimation cites this paper.

LuMon: A Comprehensive Benchmark and Development Suite with Novel Datasets for Lunar Monocular Depth Estimation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:06:00.772158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:18:51.726719Z digest=sha256:3aeb59e6726672ad16bdbf3f748777ab65015062b020b182ce4250e7250628ed

Observation 2c3a4796-b1a6-4246-94e3-9497a21452a6 · inbound

Controllable Video Object Insertion via Multiview Priors cites this paper.

Controllable Video Object Insertion via Multiview Priors Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:20.659728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T11:44:17.033051Z digest=sha256:8947716a8ed95ee67e45486d1fa0418a07748060032de43cf1432aaf6c847969

Observation 3d744e5f-9ae0-4b6c-936c-3826d068ad69 · inbound

GenMatter: Perceiving Physical Objects with Generative Matter Models cites this paper.

GenMatter: Perceiving Physical Objects with Generative Matter Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:01:18.256786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T12:53:18.504365Z digest=sha256:556d9b9869cd97f5a6ba0b5f7aa6ca784924a27fcfb17e5d9cf11e416af5de71

Observation a9a13e12-5f04-48e4-b699-7817f3f75bbd · inbound

GenMatter: Perceiving Physical Objects with Generative Matter Models cites this paper.

GenMatter: Perceiving Physical Objects with Generative Matter Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:00:07.462600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-04T19:55:05.116930Z digest=sha256:319a820f23795091eada8d7c8ac6301487b80bd002191cb9f9edd8342053471f

Observation 86fb3299-e315-4310-b09a-74670ef9662b · inbound

WildLIFT: Lifting monocular drone video to 3D for species-agnostic wildlife monitoring cites this paper.

WildLIFT: Lifting monocular drone video to 3D for species-agnostic wildlife monitoring Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:46:40.192466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T04:26:53.225501Z digest=sha256:23ed0938d600a609a85bd5592d45ab3f68831b54eb6f953fba8607b1acf576bf

Observation 05596cf5-ac2a-47a5-9472-5b6e82108ab7 · inbound

UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors cites this paper.

UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:26:07.902348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T20:05:21.723724Z digest=sha256:49722cf823e8c320dd474d6275201d72c5d42e64328e1568c494b35272c616b7

Observation dd2f7c12-0bcf-428f-afbb-cffd18c90535 · inbound

UfM*: Uncertainty from Motion* for DNN Depth Estimation Using Gaussians cites this paper.

UfM*: Uncertainty from Motion* for DNN Depth Estimation Using Gaussians Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:16:39.146720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T05:16:28.339361Z digest=sha256:9b734d9e67f8bbf71560e17fdab8fe028e4cdc27a04027e6f07e8f3dd2215eaf

Observation faf378b5-508b-4f61-bdbc-ef01898e89df · inbound

Stabilizing Streaming Video Geometry via Dynamic Feature Normalization cites this paper.

Stabilizing Streaming Video Geometry via Dynamic Feature Normalization Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:24:02.104002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T23:15:20.254402Z digest=sha256:2dcb5153f9179d59fe6adb7a8c433949d359ea493f680e9b6b341887e28a8869

Observation 0bbdd005-672f-4bdc-a482-a51255ba26fe · inbound

Neural Voxel Dynamics: Learning Implicit 3D Physics via Volumetric Feature Advection cites this paper.

Neural Voxel Dynamics: Learning Implicit 3D Physics via Volumetric Feature Advection Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:49:58.020111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T01:17:06.587807Z digest=sha256:52883b3a4f21936ef98f71b74725eb0ac66612fc76e282ff9b008e0f65cfee78

Observation 1c980236-bde6-4537-a8d9-05d5838c4609 · inbound

PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation cites this paper.

PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T20:03:56.956882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T04:34:38.286863Z digest=sha256:ea2d62bf7e034a870b2a41efcfc1e5c023b6482e3a345611fda3c74665db6f85

Observation 7b749165-ee2a-424f-81b7-5157dd18f758 · inbound

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras cites this paper.

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.239660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:16:27.239660Z digest=sha256:e4bb4a663ab8df79054e39895a7f3a5308bcc4bd499df762b8997884f648a512

Observation a1ac38ed-ec1a-4c70-8835-a89107fa843d · inbound

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models cites this paper.

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:39:26.592889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:39:26.592889Z digest=sha256:c56716ce778d46deb2d81d83bd4bfc844affaa34c4e811bfa7114c8c0fcc016c