Pith. sign in

Paper Citation Record · LEDGER

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

As of 10 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 30 inbound Pith citation observations for arXiv:2501.12375.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.12375 v3

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T17:16:42.754439Z

measured 87 of 87 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:42:05.975670Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T20:00:07.459591Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact1
  • verified fuzzy36
  • unresolved17
  • parse uncertain2
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d3d3d1b4-fcd8-4d86-abbe-517b01c981b6 · outbound

This paper cites Adabins: Depth estimation using adaptive bins.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Adabins: Depth estimation using adaptive bins

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.431315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.528682Z digest=sha256:6c7241e2972ed87d04e0740b777194b36ce344fe12db5fa377196bc23386ede2

Observation 8d288226-58ad-4f9c-8d53-577ef3170211 · outbound

This paper cites ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.533465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.533465Z digest=sha256:f4650c32100ff3f18cb7c4a41dcb0f2eabea2556eaf5f4790b5d4f340d2f5185

Observation a07b5346-78f5-486d-a793-05906da5d1af · outbound

This paper cites MiDaS v3.1 -- A Model Zoo for Robust Monocular Relative Depth Estimation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos MiDaS v3.1 -- A Model Zoo for Robust Monocular Relative Depth Estimation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.537994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.537994Z digest=sha256:b98b17b8d30ec0840a4b56d3b379b872e2e1897e1cd7ee716f91c208cd44e708

Observation e730a4f8-f9af-4623-aeb4-99bef793b4eb · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.543959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.543959Z digest=sha256:3674a1572e4a9315516f8877451ab6105ba11b722ac12b5463e877845855ed7c

Observation 73ed30e3-30b8-434d-a4e4-1087fe3afc66 · outbound

This paper cites Butler, Jonas Wulff, Garrett B.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Butler, Jonas Wulff, Garrett B

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.417215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.548569Z digest=sha256:74f4810335a19787fbcdadac62190e8c64fb013e4e5675970eef76f1dd637172

Observation 891a2586-fdd2-48cb-ad07-672423e6cdab · outbound

This paper cites Virtual KITTI 2.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Virtual KITTI 2

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.552597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.552597Z digest=sha256:df35ec519965160ded3c9a1d7f033d77696e3c861a4ad1e80f45b66b41422614

Observation d476ff43-e3cd-4ef4-aa26-3cdb8ea26f61 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.405049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.557215Z digest=sha256:5249386fb57efa9e04c2cc275c37d226c4e595cfce41c8e862e94d3ef98afcf6

Observation 6a473dbf-e0b6-4efe-bf8e-75638e0c58b4 · outbound

This paper cites Towards real-time monocular depth estimation for robotics: A survey.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Towards real-time monocular depth estimation for robotics: A survey

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.393099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.560882Z digest=sha256:318d116e933907ece2516576f21f69d50b1a60ce744cc7a8d00f437e2eaef351

Observation 4f33b360-56e9-4ce0-96f0-d9e1aa312a6f · outbound

This paper cites Depth map prediction from a single image using a multi-scale deep network.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Depth map prediction from a single image using a multi-scale deep network

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.380549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.564628Z digest=sha256:13abec4ad92d47d6ccd563cd28754487cb0a24b402f635c20c7ed719e1700fb1

Observation 1e17322c-76c3-47dd-8823-2465205f2787 · outbound

This paper cites Deep ordinal regression net- work for monocular depth estimation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Deep ordinal regression net- work for monocular depth estimation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.369581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.568349Z digest=sha256:de38f50dd6c0620e11c6cab68f037fe3d19ad3917525bb3ac7e8fe168b372aa6

Observation 8f8194b0-3357-4d6b-abc9-4fde44f0ec5b · outbound

This paper cites Vision meets robotics: The kitti dataset.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Vision meets robotics: The kitti dataset

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.571805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.571805Z digest=sha256:b310ce8eb3fa63e07322f529637c1e8316d820940938259d94c1bfae17f666ad

Observation c365942c-0ca0-4132-a698-4f0abc04218b · outbound

This paper cites Fast depth densifi- cation for occlusion-aware augmented reality.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Fast depth densifi- cation for occlusion-aware augmented reality

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.340237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.580025Z digest=sha256:3b56bf8ef6126ce3980e36dad8d9af11ccf8f409e28e7e4f222991bc7e4c020f

Observation 827ce4a9-1f2d-4dfa-bf51-70073d971803 · outbound

This paper cites DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.584184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.584184Z digest=sha256:3a5e7568a95fee0046bf9630d42031283e4523b5b38777b0aa6e03ae321e941a

Observation c6ba5fdb-3929-4236-8a1c-214e17adfa5a · outbound

This paper cites WildAvatar: Learning In-the-wild 3D Avatars from the Web.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos WildAvatar: Learning In-the-wild 3D Avatars from the Web

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-10T17:16:42.866440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.588339Z digest=sha256:df7f4175c4a4f680cb1344161904a0b0b274e2bac8bb8adaab4a515f07e2ad34

Observation 4c6dca98-977c-4dc7-9270-4da3bf9c82af · outbound

This paper cites Dai, AndreaF.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Dai, AndreaF

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.328742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.592309Z digest=sha256:16d486edd35f60aa490864d406182715a934ecf7d1b6eb53d601dd9182d3af7c

Observation 3aa68f3a-9252-48c8-9d8c-1807ff7d90fd · outbound

This paper cites Match- stereo-videos: Bidirectional alignment for consistent dynamic stereo matching.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Match- stereo-videos: Bidirectional alignment for consistent dynamic stereo matching

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.316535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.596729Z digest=sha256:69b1ef474193ec7ca3e83d2028faa451ab296b57115ebf7ec100495a192b4a0a

Observation 7fe7c9e9-5cb5-4a60-8432-b17e8ac0fa42 · outbound

This paper cites Repurpos- ing diffusion-based image generators for monocular depth estimation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Repurpos- ing diffusion-based image generators for monocular depth estimation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.305974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.600349Z digest=sha256:52b749c2c6f536aa4ff046d70a6276050b26698aea4495026e307ef60879c8ac

Observation 03b5a6b3-e8c5-4e8e-8534-9615c61ca32f · outbound

This paper cites Ro- bust consistent video depth estimation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Ro- bust consistent video depth estimation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.294969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.604220Z digest=sha256:ed3d810b79a5df39f6221aba88af77ed927ef29e058ad1532fb3b85945056de5

Observation 0b858364-bc81-4f36-9518-9bd20e8c8871 · outbound

This paper cites Towards practical consistent video depth estimation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Towards practical consistent video depth estimation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.283784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.607910Z digest=sha256:73d96258ca8d96d09ad3ab5e6baf624c324a62cf9d9a54bd45818b80feb39c15

Observation d9211d6b-5f7d-4b9a-a5cc-5f5aa989d081 · outbound

This paper cites Decoupled Weight Decay Regularization.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Decoupled Weight Decay Regularization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.611804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.611804Z digest=sha256:11f35cd12d89ab5347de2f24686e3bab19dcbc5e21c12762068a92987167641e

Observation 92127fdc-d022-42f0-85bd-7b7c419d1d17 · outbound

This paper cites Consistent video depth estimation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Consistent video depth estimation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.271325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.616052Z digest=sha256:011cdcbc6e94eca0fc2731af4a3791e0b1ca49782904ca81d697fd7523d314c4

Observation 99ac958b-3d38-4b8f-a84d-1dc402ad2747 · outbound

This paper cites Indoor segmentation and support inference from rgbd images.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Indoor segmentation and support inference from rgbd images

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.259975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.619788Z digest=sha256:b649a853c927d60566fe931898a21b2ae4b11ec802224e177cd876ac734e9205

Observation a445e850-e6b3-439b-830a-6c6fd9266774 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos DINOv2: Learning Robust Visual Features without Supervision

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.623301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.623301Z digest=sha256:479bad02f8b19094725a0c225fb3e366fc91446fc61e31949f450270ffaca910

Observation 71449da6-e297-4ea9-b218-5cfbc1b390fe · outbound

This paper cites Palazzolo, J.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Palazzolo, J

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.248793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.627338Z digest=sha256:475549ea6452f6da8fb7ef38bbb4d9b4ed6fe2368439b39fba50d8c93ec84694

Observation 9893152f-55c3-4f4a-90c4-bb01def1116d · outbound

This paper cites ControlNeXt: Powerful and Efficient Control for Image and Video Generation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos ControlNeXt: Powerful and Efficient Control for Image and Video Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.630740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.630740Z digest=sha256:1b630b44cc5d5c0dfefbd19c70bcc2d9ade6c5b560225ba7a9f93a5e24e78ddb

Observation 0bebba61-e9f4-4297-916e-78d524e0a0c5 · outbound

This paper cites Perazzi, J.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Perazzi, J

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.237391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.634724Z digest=sha256:f7dae83652b28525ed697143ca09a6f9e22b379849f9823a6fbcf22379a94324

Observation ed9888d5-aff9-495e-b72a-ce8745389117 · outbound

This paper cites Unidepth: Universal monocular metric depth estimation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Unidepth: Universal monocular metric depth estimation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.225757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.638617Z digest=sha256:524ea3830da063f8fa8b2198e4d004aa8f9509918520cf050c9a152a4d20c26f

Observation 66d3dc7d-2554-41b4-92f2-882546c95624 · outbound

This paper cites Vi- sion transformers for dense prediction.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Vi- sion transformers for dense prediction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.212892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.642980Z digest=sha256:23e3fce75132f3392bbb899f41f3821f95ec3a09ad051dcfc146dafe8983dbd3

Observation e83c7004-539f-4097-8aed-94aaf976a0bf · outbound

This paper cites Simplere- con: 3d reconstruction without 3d convolutions.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Simplere- con: 3d reconstruction without 3d convolutions

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.199762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.646505Z digest=sha256:abdb71f4f535db3a593b62f635c7083a62d8ab2dfe637ce4e0355ad055884a37

Observation 5068126c-0dff-4e04-8978-2160652500da · outbound

This paper cites Structure- from-motion revisited.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Structure- from-motion revisited

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.187334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.650052Z digest=sha256:07a140532efcead3fc850b1a2930dc2a4105db0c4e7a037b304de0919d62cee0

Observation 1c11288d-f483-4055-a12a-d43df2fafc61 · outbound

This paper cites Schonberger, Silvano Galliani, Torsten Sattler, Konrad Schindler, Marc Pollefeys, and An- dreas Geiger.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Schonberger, Silvano Galliani, Torsten Sattler, Konrad Schindler, Marc Pollefeys, and An- dreas Geiger

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.175296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.654140Z digest=sha256:a5f9cf86e1b3c86d77a1adbc8ffc4a36a77a1ce29b5ff97329f859ee9a745b46

Observation a557e620-21f8-40ef-aefb-e9015d500253 · outbound

This paper cites Learning Temporally Consistent Video Depth from Video Diffusion Priors.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.657944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.657944Z digest=sha256:3b4e51d9175152634d668e7d9abac13d0da4fa5f90efdae45619ed15f9392860

Observation b259e54b-b159-4b3d-8ff8-976e515c0c4f · outbound

This paper cites Raft: Recurrent all-pairs field transforms for optical flow.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Raft: Recurrent all-pairs field transforms for optical flow

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.162861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.661927Z digest=sha256:8a4ad48eeebf05c1aa1ca12b1a9be462eddaefb9332f4075982f53f5f00ad11e

Observation 30e7535f-9480-4030-9cc6-6b8d1ea83696 · outbound

This paper cites Attention is all you need.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Attention is all you need

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.665719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.665719Z digest=sha256:7e9a59fde329acee2ded9935b7f45a633ea44966957a70523d562aff2a2678f5

Observation b52cba32-3c53-4b8b-99ea-a7988c2c7c56 · outbound

This paper cites Irs: A large naturalistic indoor robotics stereo dataset to train deep models for disparity and surface normal estimation, 2021.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Irs: A large naturalistic indoor robotics stereo dataset to train deep models for disparity and surface normal estimation, 2021

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.143347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.669835Z digest=sha256:1aa0f4342c01b8ab2604822f1eea0beeff5f2abcd07285b6c933d1029083b0f2

Observation 0a6184a8-c7d4-4a1f-a762-f9185fb6fb56 · outbound

This paper cites Moge: Unlocking accurate monocular geometry estimation for open-domain images with optimal training supervision, 2024.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Moge: Unlocking accurate monocular geometry estimation for open-domain images with optimal training supervision, 2024

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.130365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.673488Z digest=sha256:a5a1946d71575df3ed1e963efe85d5cc0a1e29be84bbe621825b4a772a828ae8

Observation b73f032f-3097-4878-938a-e9d93a66b986 · outbound

This paper cites Tartanair: A dataset to push the limits of visual slam.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Tartanair: A dataset to push the limits of visual slam

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.116978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.677479Z digest=sha256:34de966958ccd451d6b8fca3025bbef76b051371ba7c920132df3f8341a8a1ce

Observation 98c4bc60-96d9-4570-9c12-cbc0c56bfbe5 · outbound

This paper cites Less is more: Consistent video depth estimation with masked frames modeling.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Less is more: Consistent video depth estimation with masked frames modeling

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.104835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.681304Z digest=sha256:12e9fedb5902a78b3d1d9bf126decec2b6996510ae63ee5a7e83992e0754dda1

Observation 5cffe7e7-6ace-4998-8f3e-afdda3707646 · outbound

This paper cites Neural video depth stabilizer.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Neural video depth stabilizer

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.685170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.685170Z digest=sha256:be3448ded8b19ae04bf16d8227fda56497e15741d3de13f4d71f43567197fd19

Observation 5df45105-0661-4cbe-8120-8c071d0a5eb4 · outbound

This paper cites Gmflow: Learning optical flow via global matching.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Gmflow: Learning optical flow via global matching

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.693355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.693355Z digest=sha256:79438673943cce3e4be6fdaf1d7b7b98b2b3c11b1f6b56e4ca1d17fc042f0d88

Observation 968b365f-2842-49ff-9244-bda754e2dc54 · outbound

This paper cites Depth Any Video with Scalable Synthetic Data.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Depth Any Video with Scalable Synthetic Data

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.697279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.697279Z digest=sha256:2ef93e9de5a834bd1d48db12f7e8259f47c6dbef383b7974981ea4c86213bf2e

Observation fc3685f3-a462-4440-aa17-3c573adc9a22 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Depth anything: Unleashing the power of large-scale unlabeled data

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.066741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.701620Z digest=sha256:66a78d779751a6cafe0759e37b094096c0ee1a4e8b87f89f408d8a87d4e567d7

Observation 0fb9ef0c-ca73-4a68-987d-0722d00d47fb · outbound

This paper cites Depth Anything V2.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Depth Anything V2

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.705513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.705513Z digest=sha256:65766121b30232a877ce12c3c5c2802007c8f5b12a16fb1d444276cff4f22fa8

Observation 9dfc5840-6544-47cc-a7d3-cf34b5b6537e · outbound

This paper cites Mamo: Leveraging memory and attention for monocular video depth estimation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Mamo: Leveraging memory and attention for monocular video depth estimation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.055642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.709708Z digest=sha256:04b38c52b8ca13162d57bcb1706d834bbfafa4b9dca8ac1e4bba41e050aae05b

Observation e32d5c73-5566-48e4-a34a-90b46ff9dfc9 · outbound

This paper cites Metric3d: Towards zero-shot metric 3d prediction from a single image.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Metric3d: Towards zero-shot metric 3d prediction from a single image

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.044566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.713324Z digest=sha256:49ee1a6b6e1de497cc690a3163e20d2836643cb4fd8ae75b2975c38b2d92d3c7

Observation 9b851ce3-14fd-4e44-96e3-b484dc6d2096 · outbound

This paper cites Neural window fully-connected crfs for monocu- lar depth estimation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Neural window fully-connected crfs for monocu- lar depth estimation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.033010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.717269Z digest=sha256:e3742d5ff17eb3c357f99901f2ac89b14c2e6133d24e3df03d905c73b545deb0

Observation 7c8ead4a-2144-429a-8e7a-8c6963282a0c · outbound

This paper cites ControlVideo: Training-free Controllable Text-to-Video Generation.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.721276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.721276Z digest=sha256:83b751f1803715c0eebce20ca2bb4a411440503f0b05e3a23a906eef450e1503

Observation 6bcbb9a6-5c9b-403a-84eb-3030e2267c36 · outbound

This paper cites Consistent depth of moving objects in video.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Consistent depth of moving objects in video

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.021156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.725219Z digest=sha256:79f388e5179740643cedbfbe3adb1c88e1e1d6f8646622b85d04b43742c14a2b

Observation 19c1d010-2365-4e36-92dd-651209660d49 · outbound

This paper cites Pointodyssey: A large-scale synthetic dataset for long-term point tracking.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Pointodyssey: A large-scale synthetic dataset for long-term point tracking

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:43.009362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.728998Z digest=sha256:29d30133ef23734535f103ffbb3e04b723529535004404888e0747039674aac3

Observation 35859513-8ef7-4af7-bb33-cdcf21510635 · outbound

This paper cites In-the-wild image results.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos In-the-wild image results

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:42.997096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.732819Z digest=sha256:35806fa3d78b0b55f445b1475c8e11209444c87db8f77d7eb621dec44462a26e

Observation 87a903d2-8396-49a7-8889-81829b9655fc · outbound

This paper cites As shown in Tab.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos As shown in Tab

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:42.983233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.736378Z digest=sha256:97d86ba3bbfca2b4caf3cd456f485ed99b4767eaadd528c853bbd3f4cb75b329

Observation 9c37abb0-cd46-4d41-aa80-28fecef0ecc6 · outbound

This paper cites We believe that with more data, the model’s performance can be further improved, and the backbone network can be unlocked for fine-tuning.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos We believe that with more data, the model’s performance can be further improved, and the backbone network can be unlocked for fine-tuning

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:42.970075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.740811Z digest=sha256:ae2fc26ccce5ed43a766c3829d567d6ccc8ec2d7b542a56c3edb617396b44baa

Observation 9b306b64-d356-4f30-9f4b-13db9fef4c96 · outbound

This paper cites Among the four temporal layers, two are inserted after the Reassemble layers at the two smallest resolutions, and the other two are inserted before the last two Fusion layers.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Among the four temporal layers, two are inserted after the Reassemble layers at the two smallest resolutions, and the other two are inserted before the last two Fusion layers

Reference 55

Resolution
malformed identifier
raw_fallback, observed 2026-08-10T17:16:42.958653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.745823Z digest=sha256:437a212144484aa15783bc2df8a1d5d8ad43d4f1c1ad97759b5ee74e91b5f192

Observation b1256bde-70cd-411d-b8fa-196442820c1d · outbound

This paper cites We use a total of five datasets for video depth evaluation: KITTI [ 11], Scannet [ 7], Bonn [24], NYUv2 [ 22], and Sintel [ 5].

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos We use a total of five datasets for video depth evaluation: KITTI [ 11], Scannet [ 7], Bonn [24], NYUv2 [ 22], and Sintel [ 5]

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T17:16:42.946830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.750503Z digest=sha256:bbcd202b791cd242e59955328c159fddc9ee9fd41bddb9ae9baede91ad75939d

Observation c5b24286-6d96-426f-8ea3-b00c4e6add9c · outbound

This paper cites an unresolved cited work.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-10T17:16:42.933293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.754439Z digest=sha256:a8be355b3c4cbb49e8db656b1d204d3a35bb62a3a7131c54d0b595c0c5da0882

Observation 34300afb-759f-4307-b9b5-213e605a2f4d · outbound

This paper cites an unresolved cited work.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Unresolved cited work

Reference 2013

Resolution
parse uncertain
raw_fallback, observed 2026-08-10T17:16:43.351728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.575871Z digest=sha256:c4f0fcf75a1ba69c5dca1122797a8aed68acdbbcdbdd688c986a7bede50052c5

Observation 808c5845-7ecd-41e2-b8cf-d1e977f7524e · outbound

This paper cites an unresolved cited work.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Unresolved cited work

Reference 2023

Resolution
parse uncertain
raw_fallback, observed 2026-08-10T17:16:43.085125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T17:16:42.689159Z digest=sha256:b97c584022ae4de161d4065360f0ef9c2429a12b27d295cc6cb43abfc765df88

Pith citing papers

Observation 56e8e8b7-e70e-44c4-b103-8c88dfbc9e82 · inbound

DissolveStereo: Coarse Depth Injection for Zero-Shot Stereo Video Generation cites this paper.

DissolveStereo: Coarse Depth Injection for Zero-Shot Stereo Video Generation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:18:13.976065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T17:17:48.961672Z digest=sha256:4c1166641d5b8e47a9d2323b97af53d307fb8034134f7ed91362a6925e37697b

Observation 8a153380-aa41-472a-b42f-d93b78d6ce3f · inbound

Seeing World Dynamics in a Nutshell cites this paper.

Seeing World Dynamics in a Nutshell Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T04:42:05.975670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:42:05.975670Z digest=sha256:0e163fc60708367e9d5d6c9f60f3846a91fcbbcc3f4103ad6992c61070d024dc

Observation 677d1a65-46d6-4a29-8c90-a851b74eb0e9 · inbound

UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation cites this paper.

UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:26:57.495990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:26:57.495990Z digest=sha256:1cf6e5eaa4f9d1a60b774ac5205a120b483719a76c7fea5326cdeeb8ffe3053c

Observation 02a6ffb5-61ac-4584-91e1-58d98f7acf98 · inbound

E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models cites this paper.

E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:36:06.197375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:36:06.197375Z digest=sha256:400520478a9950ccef1f945de348efebfa748cb0913d05243ebc42be49010abc

Observation 53500216-ec5f-460f-b62c-3f9bbeec206c · inbound

IllumiCraft: Unified Geometry and Illumination Diffusion for Controllable Video Generation cites this paper.

IllumiCraft: Unified Geometry and Illumination Diffusion for Controllable Video Generation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:43.647637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:43.647637Z digest=sha256:fd645d1777b7594e64e8262f5b67204b06a45895b5d0971e76a50f35416cfe61

Observation 118f8165-5b8a-471a-a0ce-a0ed023223e8 · inbound

RaCalNet: Radar Calibration Network for Sparse-Supervised Metric Depth Estimation cites this paper.

RaCalNet: Radar Calibration Network for Sparse-Supervised Metric Depth Estimation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T23:58:41.807576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:58:41.807576Z digest=sha256:c58cd32747b8304f0c8433231decdea1fd11afac485e650dbc0f3cad0abad5f1

Observation ddf523f2-5879-4fe0-9116-f32619e7454c · inbound

RoboScape: Physics-informed Embodied World Model cites this paper.

RoboScape: Physics-informed Embodied World Model Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:56:09.428813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:56:09.428813Z digest=sha256:12f86062d179fc956dc4e22b8d21489e141edd30f1f4fa86bc7ac08cf5faa8fd

Observation ecf4f2db-7073-4213-beb6-daefed166cac · inbound

MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion cites this paper.

MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:06.939745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:06.939745Z digest=sha256:1e012cb4dc85c696d53dd4a65268ce96889b174919e3d8fdc1207f941dad6be7

Observation 98217f03-2f15-440f-9424-261dcd1fb3f8 · inbound

RCG: Safety-Critical Scenario Generation for Robust Autonomous Driving via Real-World Crash Grounding cites this paper.

RCG: Safety-Critical Scenario Generation for Robust Autonomous Driving via Real-World Crash Grounding Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T17:34:05.903479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:34:05.903479Z digest=sha256:b028beca9bfc7aaa6f0f76047a104f2ed96b12382fa101195cab923c422e6a9a

Observation c3c5ea58-a021-4c05-b163-8065c0f45e55 · inbound

SpatialTrackerV2: 3D Point Tracking Made Easy cites this paper.

SpatialTrackerV2: 3D Point Tracking Made Easy Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:49:48.565573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:49:48.565573Z digest=sha256:3d1e2158b89eb1c34eca5a3bb4f1a6b5fc7660259df562ea30e4202fe4d028fd

Observation 96348a5e-db49-4c7b-86cc-e3b688b5be64 · inbound

Reconstructing 4D Spatial Intelligence: A Survey cites this paper.

Reconstructing 4D Spatial Intelligence: A Survey Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 135

Resolution
unresolved
no resolver link, observed 2026-08-06T13:02:29.086195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:02:29.086195Z digest=sha256:7f140f6639ecc7b3b9357a659fa7b895d63d983e0fd6f69b16204c134070cc71

Observation 3a193b00-fc7b-44d1-a0ab-53127283cb3e · inbound

ViPE: Video Pose Engine for 3D Geometric Perception cites this paper.

ViPE: Video Pose Engine for 3D Geometric Perception Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.727344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:9bfaaca3beaecb5b2959d43bc45899b1cb6eb56dea285c0e5a93b33b56d9e87c

Observation c11040f0-350d-4e24-a23f-8f6aa3498ec0 · inbound

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation cites this paper.

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T13:46:44.930591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:46:44.930591Z digest=sha256:73b0cbb47a2c4603ed1f7932aa2fa70535ee7267b9c405e0c4e43a8ba71da2d9

Observation 96c1b55d-adfa-44e5-9ef0-2d72c5c1f332 · inbound

Feedback Matters: Augmenting Autonomous Dissection with Visual and Topological Feedback cites this paper.

Feedback Matters: Augmenting Autonomous Dissection with Visual and Topological Feedback Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T11:35:58.419641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:35:58.419641Z digest=sha256:5cc5d558b6d7e3a99812334945c8c3886780f16da4c2d63cc13225411fd776d1

Observation 6be966d9-301c-4180-901d-c6d8cd0236ce · inbound

CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction cites this paper.

CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:38:37.596051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T22:38:01.008280Z digest=sha256:1e4ff7d58ec840b85643c622d1db4f04ed7c96037876b491c5e16b601878007c

Observation e0f3cdf7-c86a-49ef-a33e-0492f7d05222 · inbound

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation cites this paper.

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T22:13:53.383917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:13:53.383917Z digest=sha256:6f846a8dc35d9effba5ac0a50224aaa865d1c4b413b6fbbf946fb4e8cd75276e

Observation f87dda88-df7c-41f9-ba6e-74ead34e61c1 · inbound

HOIGS: Human-Object Interaction Gaussian Splatting cites this paper.

HOIGS: Human-Object Interaction Gaussian Splatting Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T17:28:02.544902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:50.043908Z digest=sha256:b6938e6a4320a79fe066c52c4488769fe6703e4fc2072e66293543c41f6d2988

Observation 2bb3e8db-2aea-477e-b6f7-934d607c3940 · inbound

From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data cites this paper.

From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:03:01.206201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T17:02:18.358675Z digest=sha256:7ffeee1a15991817f6c24f54cb62dc052f72d9cc5e493470eed86d18c3231b69

Observation fbd445bf-0125-4f54-bbcb-db9b20b08ad8 · inbound

LuMon: A Comprehensive Benchmark and Development Suite with Novel Datasets for Lunar Monocular Depth Estimation cites this paper.

LuMon: A Comprehensive Benchmark and Development Suite with Novel Datasets for Lunar Monocular Depth Estimation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:06:00.772158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:18:51.726719Z digest=sha256:ae3054bbdeaea3abb8fa574082c060d4199306d99c8e933474b897f2b783ee6e

Observation 2c3a4796-b1a6-4246-94e3-9497a21452a6 · inbound

Controllable Video Object Insertion via Multiview Priors cites this paper.

Controllable Video Object Insertion via Multiview Priors Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:20.659728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T11:44:17.033051Z digest=sha256:97062f2f994fd7d54e8d78d0a3c35caf8d0f463e71336afd6ced05f178b7afc6

Observation 3d744e5f-9ae0-4b6c-936c-3826d068ad69 · inbound

GenMatter: Perceiving Physical Objects with Generative Matter Models cites this paper.

GenMatter: Perceiving Physical Objects with Generative Matter Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:01:18.256786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T12:53:18.504365Z digest=sha256:cbd30701a21c0ce1ea32d313e66742f44c48e05cd08728590015348d22af78b4

Observation a9a13e12-5f04-48e4-b699-7817f3f75bbd · inbound

GenMatter: Perceiving Physical Objects with Generative Matter Models cites this paper.

GenMatter: Perceiving Physical Objects with Generative Matter Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:00:07.462600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-04T19:55:05.116930Z digest=sha256:fcceee0f968446e0eb6bd691d6f6f1af441df8051c5eac7ef5ec9c1a121484ee

Observation 86fb3299-e315-4310-b09a-74670ef9662b · inbound

WildLIFT: Lifting monocular drone video to 3D for species-agnostic wildlife monitoring cites this paper.

WildLIFT: Lifting monocular drone video to 3D for species-agnostic wildlife monitoring Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:46:40.192466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T04:26:53.225501Z digest=sha256:155fcd3997a50cb542fdf9b8205ac610e8c247578605897f1d1166aaa014f911

Observation 05596cf5-ac2a-47a5-9472-5b6e82108ab7 · inbound

UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors cites this paper.

UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:26:07.902348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-09T20:05:21.723724Z digest=sha256:dec20919a4b0317cbf9a510628deff9a77a80daceede9f8f671d4d57379206bf

Observation dd2f7c12-0bcf-428f-afbb-cffd18c90535 · inbound

UfM*: Uncertainty from Motion* for DNN Depth Estimation Using Gaussians cites this paper.

UfM*: Uncertainty from Motion* for DNN Depth Estimation Using Gaussians Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:16:39.146720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T05:16:28.339361Z digest=sha256:b605b710076dcc9ab93884bb9570c56b560ade115aa7c24c58e115dbebec79cc

Observation faf378b5-508b-4f61-bdbc-ef01898e89df · inbound

Stabilizing Streaming Video Geometry via Dynamic Feature Normalization cites this paper.

Stabilizing Streaming Video Geometry via Dynamic Feature Normalization Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:24:02.104002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T23:15:20.254402Z digest=sha256:b4c867f2df350b8abbb9d18aa6e009de8cc37a77c5d3cb115afcd1d2196d9e11

Observation 0bbdd005-672f-4bdc-a482-a51255ba26fe · inbound

Neural Voxel Dynamics: Learning Implicit 3D Physics via Volumetric Feature Advection cites this paper.

Neural Voxel Dynamics: Learning Implicit 3D Physics via Volumetric Feature Advection Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:49:58.020111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T01:17:06.587807Z digest=sha256:7a91e23aec144bbb5dd1567e33154834dc8d5f43092a365532b7e389c6b02509

Observation 1c980236-bde6-4537-a8d9-05d5838c4609 · inbound

PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation cites this paper.

PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T20:03:56.956882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T04:34:38.286863Z digest=sha256:89e4557e0f08942e90a3cde50346f1981aede2d76c626c0009913a7499f8f7e7

Observation 7b749165-ee2a-424f-81b7-5157dd18f758 · inbound

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras cites this paper.

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.239660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:16:27.239660Z digest=sha256:29fd4686baa90951c77c8aa1fce51f4a772e2527b066f9b6ae58871e9cfbe91e

Observation a1ac38ed-ec1a-4c70-8835-a89107fa843d · inbound

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models cites this paper.

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:39:26.592889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:39:26.592889Z digest=sha256:00d44b14062658bd9617eac6047bfa92bcdf146f19d587028362ec7fe5ad27bf