Pith. sign in

Paper Citation Record · LEDGER

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors

As of 19 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 1 inbound Pith citation observation for arXiv:2411.17249.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17249 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:26:49.610580Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T00:56:18.867382Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

64 of 64 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 187910a2-3575-461a-9c28-bc9b1ac1baef · outbound

This paper cites Stable diffusion version 2.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Stable diffusion version 2

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.611236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.326039Z digest=sha256:088e2d8a8ea0358696cf47dee0d8a90e122c45ab4140a743f3016323cd0a6ea4

Observation 59d925b9-d6a3-43cc-b7ef-504d0728e725 · outbound

This paper cites Rethinking induc- tive biases for surface normal estimation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Rethinking induc- tive biases for surface normal estimation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.595455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.331132Z digest=sha256:fb44b365af37f8a98c538a7198d5ee7cd10eac9aef1121d4c2292ce2c9a6258f

Observation e4575005-7e21-49cc-9ca6-4366e1e6a895 · outbound

This paper cites Es- timating and exploiting the aleatoric uncertainty in surface normal estimation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Es- timating and exploiting the aleatoric uncertainty in surface normal estimation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.335953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.335953Z digest=sha256:5b2f7c427a73c139cea1aa27c7484533283c746de6bd9b5aa1c7bcc4d7999b65

Observation 262d1cd3-b732-4e78-8a39-4f313c02ed5d · outbound

This paper cites Marr revisited: 2d-3d alignment via surface normal prediction.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Marr revisited: 2d-3d alignment via surface normal prediction

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.570136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.340963Z digest=sha256:cff423a458c5ccf631c0e4e82db634ed42c2a78ec17e6748101b6651cfe8c7eb

Observation 2518f65c-8653-4fde-acf1-c5c1fb3e8f37 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.345431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.345431Z digest=sha256:6080547026794cf69a555d3354a05fe0e93a5a9c228cf6b1741ea3424ebe3f37

Observation 7e247130-d8f9-43b0-9ce8-11808caffa17 · outbound

This paper cites Video generation models as world simulators.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Video generation models as world simulators

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.350473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.350473Z digest=sha256:836ca03ca0a2cd5da511008cfff4d65f7bf4083226f8067267f8762cf363bff2

Observation 65ececb4-b366-4fa1-bc0b-fbd7eea0694e · outbound

This paper cites A naturalistic open source movie for opti- cal flow evaluation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors A naturalistic open source movie for opti- cal flow evaluation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.354841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.354841Z digest=sha256:8fb23b6a90b68195d83d8b16299ecfdc59b86cefc01b219d27f2d0739f4eda13

Observation 245df49e-2c83-4604-98f2-8341d16205c1 · outbound

This paper cites A computational approach to edge detection.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors A computational approach to edge detection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.359265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.359265Z digest=sha256:5f22913cd4ca799fc0e9c286a93777ae20a798ebaa72e61618a6a103e11c579b

Observation d2c3d993-fed7-4cfe-a952-3673610f9586 · outbound

This paper cites Learning structure affinity for video depth estima- tion.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Learning structure affinity for video depth estima- tion

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.525276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.364198Z digest=sha256:4dd5b1f18dc781c69d39721f2d91df792f4126150401f1881eb4d70d909515e1

Observation 56c2cd09-9da3-4dc2-b004-a6773999b77c · outbound

This paper cites Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffusion,.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffusion,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.499768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.373053Z digest=sha256:df9448eb1acd973abe1a69de6bda6ef7b3d034f71897ac6a40d7ef127ae1763e

Observation 01b72b87-6db6-4e18-b6fc-9db078e79c6c · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.378047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.378047Z digest=sha256:1470ceed58f747b51ab8fb8724638ac784a68584a37397a56916728420b64707

Observation 97050a54-893b-4653-b268-dfea066a0245 · outbound

This paper cites Surface normal estimation of tilted images via spatial rectifier.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Surface normal estimation of tilted images via spatial rectifier

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.474955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.382110Z digest=sha256:51fc000dde314e87bb2dd7ad444921b5701c407cf6faae0a9cd806e34582f5bd

Observation 658f4f7f-8352-4d5b-8e2e-38a96f3f628e · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors An image is worth 16x16 words: Transformers for image recognition at scale

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.386304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.386304Z digest=sha256:c16153ca37e221d861374ae3ce2a4c3e14836ecb48e9192795d40b8f4179e5c0

Observation f2f4f897-f0be-4af4-9219-7787f9111105 · outbound

This paper cites Omnidata: A scalable pipeline for making multi- task mid-level vision datasets from 3d scans.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Omnidata: A scalable pipeline for making multi- task mid-level vision datasets from 3d scans

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.390520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.390520Z digest=sha256:b619922d3d0266ad490fecde254f95de5a75eceab9c5c63e6307355ebecf0a48

Observation 6ec9e73d-34da-4b78-aaa4-2128f5f59dc9 · outbound

This paper cites Predicting depth, surface nor- mals and semantic labels with a common multi-scale con- volutional architecture.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Predicting depth, surface nor- mals and semantic labels with a common multi-scale con- volutional architecture

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.394856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.394856Z digest=sha256:29801c11a90199c2fc6bbe592c0fecc52fce6788a85cda025b2a912468a3c3de

Observation 2c9d8aa0-67eb-462d-8b0e-dd08457fff4e · outbound

This paper cites Gpt-3: Its nature, scope, limits, and consequences.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Gpt-3: Its nature, scope, limits, and consequences

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.399804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.399804Z digest=sha256:e7144f9eb01899389ed733b9581048cbd11daaae28bfb2b8bc44881d145ef63c

Observation 032f69e2-a540-41a1-8fa4-26a649dd0315 · outbound

This paper cites Unfolding an indoor origami world.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Unfolding an indoor origami world

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.419189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.404192Z digest=sha256:68669e7a2252b26d4f9ad7a44856ed7d7047ff8fa4eec29e2138a5c607ba8ab8

Observation 82fae043-fe50-4ef7-99ee-7871cdf01673 · outbound

This paper cites Deep ordinal regression net- work for monocular depth estimation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Deep ordinal regression net- work for monocular depth estimation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.408572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.408572Z digest=sha256:cf1490d1989151347c55eeb5d21ffc9b82ed4aa8a3455e5dfe9d4731e772df35

Observation 2ae88873-2835-4af3-abb9-d76d0f0c1ef9 · outbound

This paper cites Geowiz- ard: Unleashing the diffusion priors for 3d geometry esti- mation from a single image.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Geowiz- ard: Unleashing the diffusion priors for 3d geometry esti- mation from a single image

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.394471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.413308Z digest=sha256:7f9af7dfe45bfe37a5026659fdf0873a41ef5020d01373dbc40f732ece652eae

Observation 10fd91f8-9759-4813-ada6-a8b09b7ad15b · outbound

This paper cites Fine-tuning image-conditional diffusion models is easier than you think.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Fine-tuning image-conditional diffusion models is easier than you think

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.417483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.417483Z digest=sha256:76215ae7108279db0727f1552797e13ed91e417456b68c001ecb04c7d4135a07

Observation 1abc2e7b-9865-47f3-b0df-773249ae66a6 · outbound

This paper cites Vision meets robotics: The kitti dataset.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Vision meets robotics: The kitti dataset

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.421736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.421736Z digest=sha256:8a8f354c8026bba15e23077533e33c2b44f7d664904ca0e5548205fbb6507406

Observation f616ba72-e35a-4819-8b28-52d37ebe15e0 · outbound

This paper cites Unsupervised monocular depth estimation with left- right consistency.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Unsupervised monocular depth estimation with left- right consistency

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.425894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.425894Z digest=sha256:87a6cbd6aa5d86a16711bf1dd8dc5231b73e54b1f92d78334a3350cf967095f5

Observation e8dbd863-3354-471b-b356-c0b7b7935331 · outbound

This paper cites Sparsectrl: Adding sparse controls to text-to-video diffusion models.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Sparsectrl: Adding sparse controls to text-to-video diffusion models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.360134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.430555Z digest=sha256:d44bac3869998fe03a04d2dc65362af6d505bba7df56708cc65b4fe8256ce6f8

Observation 975036b5-c9bd-48f2-964e-05b94accd2ff · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.434878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.434878Z digest=sha256:23f27145fe67e7695bd88fe2b5037d11b2f1e2f7feaefef117982181d3596a32

Observation 9ac0534f-b42b-44f2-aa42-633ffb811955 · outbound

This paper cites Cameractrl: Enabling camera control for text-to-video generation, 2024.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Cameractrl: Enabling camera control for text-to-video generation, 2024

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.439392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.439392Z digest=sha256:4afbe37e176d411f49823960913e0b54027fab96ec57f4c0f3317bc474b3df66

Observation 9b02d17b-0e95-4da9-aa2b-5900149ae4f9 · outbound

This paper cites Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.443740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.443740Z digest=sha256:54089ebe97d0d767c21775bec4c5151f0a01553582b80ceb3d5b40669e35fe00

Observation 7ac5a2f9-31bf-4f7f-91c8-ba9acc66daf9 · outbound

This paper cites Au- tomatic photo pop-up.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Au- tomatic photo pop-up

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.335238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.448544Z digest=sha256:9038e29395909164dcd82192b67bcf83ae50c3068111ab666d3f2c94724f9476

Observation cf3d1765-4faa-4c7b-8986-b284eb72a750 · outbound

This paper cites Recov- ering surface layout from an image.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Recov- ering surface layout from an image

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.320375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.453035Z digest=sha256:1a70203c682cb0d68873abbaf86cf0409a448fd53355acb20aaa63a773980eaf

Observation 1df4caab-b875-4548-954e-4105eb5519bc · outbound

This paper cites Animate anyone: Consistent and controllable image-to-video synthesis for character animation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Animate anyone: Consistent and controllable image-to-video synthesis for character animation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.457920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.457920Z digest=sha256:18793fb860ca74be8f4d999986dc4e371e6c406025b8df70d8bd6c0660c56c96

Observation 40020932-d4fa-4969-82a0-695cf3e8a991 · outbound

This paper cites DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.462374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.462374Z digest=sha256:65e619308c9064f63eedcf28bb69a3699c6ec16ddfd6d69966c5229925787975

Observation 2568f948-f0d4-4a91-adc3-4ff448e40c6d · outbound

This paper cites 3d common corruptions and data augmentation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors 3d common corruptions and data augmentation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.294598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.467369Z digest=sha256:4a9f01b11649eff92f537504a4e364aa9df96108093b40e1415322f5bd2a3b27

Observation 555eb39d-6469-49f0-9625-4f2497ded9d8 · outbound

This paper cites Repurpos- ing diffusion-based image generators for monocular depth estimation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Repurpos- ing diffusion-based image generators for monocular depth estimation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.279786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.472397Z digest=sha256:712a076a8cb8a97d69aa18e0fd27a28deb7fc5538ef3c010cf7020f6f42808a8

Observation a67698da-6d59-4469-b48f-c56f66fa77f2 · outbound

This paper cites Wetzstein.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Wetzstein

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.263587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.476689Z digest=sha256:95c462b7030183afc8b07e79a66db0454e05d5716024619318f86bd05eb9b5f8

Observation 787e8a8c-88ec-48e2-85cc-c6f8f7caf426 · outbound

This paper cites Sift flow: Dense correspondence across different scenes.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Sift flow: Dense correspondence across different scenes

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.248950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.480809Z digest=sha256:a52459ecbeccf17b8b966f7a8b5c91b16178b781d65b9ea237d9f60574158786

Observation f251f2a2-e6d2-4adc-9891-fab44589eaad · outbound

This paper cites Decoupled weight decay regularization, 2019.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Decoupled weight decay regularization, 2019

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.234128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.485369Z digest=sha256:4f73987c40950023080d8443bac349835b251280402b6d41e6ea8aa67967de6e

Observation 88326bde-0c7e-4fe4-86e7-a20d4c0c19fd · outbound

This paper cites Refusion: 3d reconstruction in dynamic environments for rgb-d cameras exploiting resid- uals.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Refusion: 3d reconstruction in dynamic environments for rgb-d cameras exploiting resid- uals

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.489604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.489604Z digest=sha256:a268edcc97095667e7a12316dd84ee2a541323155baa0ab8dd92736d9526a4fd

Observation a2df41bb-4bb0-4aa4-94fe-464ef2fc8c1d · outbound

This paper cites Automatic differentiation in pytorch.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Automatic differentiation in pytorch

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.208525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.493902Z digest=sha256:ebf495f59b7a7bcd0e56395e3150b4723f119d1b7fa1cd5b3d6bf5f4bf9f1420

Observation dc4db023-3044-4f21-98b7-1917724f6012 · outbound

This paper cites State of the art on diffusion models for visual computing.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors State of the art on diffusion models for visual computing

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.192683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.498450Z digest=sha256:30bbb0bf1111fbea4618a664c378a40dd335df86460243038508fc4a6eb2f714

Observation 41b6bb34-8ea3-4e49-aadf-4b030892ad42 · outbound

This paper cites Movie Gen: A Cast of Media Foundation Models.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Movie Gen: A Cast of Media Foundation Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.502751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.502751Z digest=sha256:82dfaa645016ba0f50ad5443721c9ed079e28120e1eeae8973881d0338660019

Observation e9c72c4b-9ab3-4974-a4db-0507721a899a · outbound

This paper cites Geonet: Geometric neural network for joint depth and surface normal estimation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Geonet: Geometric neural network for joint depth and surface normal estimation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.508084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.508084Z digest=sha256:bde2bbbf6a2f9cea4332bbcf328b967d75407afab51c3fc03fb7d87378370340

Observation a4a05eea-5326-40b2-871a-9ad9aa4a2b96 · outbound

This paper cites Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.512476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.512476Z digest=sha256:1e91b1425f1b685f359f4a9449c4d8cc03e76d05cdea3076c1f8d719b46b0ff7

Observation cc6973c1-3e45-4488-8bf9-8313d8c3d00e · outbound

This paper cites Vi- sion transformers for dense prediction.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Vi- sion transformers for dense prediction

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.516819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.516819Z digest=sha256:4121a99288dc74e9819a285b0127222e3f118dae97c4c76d7a12caf66f87495f

Observation 1a86e612-d6b3-4773-82be-6f67fc50394b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors High-resolution image synthesis with latent diffusion models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.521458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.521458Z digest=sha256:309b6e5d5a075b85f6d302a07ff31fffc236a12069311ed6dfbb1120399e41c0

Observation 447cf863-7a9f-4cf0-8274-1461c587a32e · outbound

This paper cites Make3d: Learning 3d scene structure from a single still image.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Make3d: Learning 3d scene structure from a single still image

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.525708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.525708Z digest=sha256:a82b588dafe1251bd641944be3fc5994f7f734d9e3efdec94844adee93d60770

Observation 55069968-05a1-4281-9857-e5a3511f756e · outbound

This paper cites Learning Temporally Consistent Video Depth from Video Diffusion Priors.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.531022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.531022Z digest=sha256:64f5e4048b745936a46392fba76613326cbaaa4cb1b0811c093baad359dcbb2a

Observation 20332be9-bc3b-422c-8fc7-30c45a9ad1cb · outbound

This paper cites Human4dit: 360-degree human video gen- eration with 4d diffusion transformer.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Human4dit: 360-degree human video gen- eration with 4d diffusion transformer

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.535752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.535752Z digest=sha256:8b5767069625f0ec4fca357941296fee794d23cc5a7c950e98354cb69232b8d6

Observation a0b2460d-c2dd-4561-868d-92c11a7716e5 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.540029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.540029Z digest=sha256:a6a238b5ad3892c8df06cb9dc292d3ce3c2fe77c022e02914c4d5281070fa3fe

Observation 0a6d8dc2-3177-467d-bc8f-d2e7e7f35f97 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors LLaMA: Open and Efficient Foundation Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.544620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.544620Z digest=sha256:58a62dbf83c5a7b582f4dffc365c94d09ea3f677fc0313257392d8b7fbad6b4f

Observation c8729ca6-1a4e-4c09-95f8-6d771d0c42de · outbound

This paper cites Diffusers: State-of-the-art diffu- sion models.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Diffusers: State-of-the-art diffu- sion models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.549250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.549250Z digest=sha256:f0ab568a73e11568ee52ecb187ed8a29d6fc515c973b20380213cde0ef067a1e

Observation 10f8e96c-0a6e-4de1-bf79-f984be9415ae · outbound

This paper cites Vplnet: Deep single view normal estimation with vanishing points and lines.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Vplnet: Deep single view normal estimation with vanishing points and lines

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.105611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.553546Z digest=sha256:cc62db8fc6d1078b445fa3666c33f20ac40c1fd15f680cde52c6032078cbc077

Observation 11754409-b13d-45f4-9751-d988e5fbbe34 · outbound

This paper cites De- signing deep networks for surface normal estimation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors De- signing deep networks for surface normal estimation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.089789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.558243Z digest=sha256:a5d692ef9636b16b6116317d86eac7da2a38c9224747e6499505b935496fb0ee

Observation b2cbdd1c-6395-4ba8-a6b6-f9f690b4a448 · outbound

This paper cites Less is more: Consistent video depth estimation with masked frames modeling.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Less is more: Consistent video depth estimation with masked frames modeling

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.074061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.562494Z digest=sha256:b8f78a39757e58eebeb8cd152d8bc01aa3779dc52e292b3d7fbc952a51e781fc

Observation a85f3e65-fcab-43ec-bdca-3f32631c5da8 · outbound

This paper cites Neural video depth stabilizer.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Neural video depth stabilizer

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.566738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.566738Z digest=sha256:209105f23d90f7f263150aa02247480156aab8778ac540851775859f92e688d1

Observation b0acf96b-5ada-4329-8bae-d08283b2dda1 · outbound

This paper cites CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.571053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.571053Z digest=sha256:2792a0d04c5cdd4d2cce7acbd21ea7cb4d3617c8d0df67b8fb840a9ca6674d08

Observation d05fd471-d3b3-4e86-bdd8-70c6a9869666 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Depth anything: Unleashing the power of large-scale unlabeled data

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.049053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.575647Z digest=sha256:be5880ad4c111c6d703754c36397c984a2d07816e296bb9e7e5cbe11aabd8861

Observation 117cd4ed-22cd-4fbd-b815-4464841292b2 · outbound

This paper cites Depth Anything V2.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Depth Anything V2

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.579883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.579883Z digest=sha256:2ac99404c1d593a3423c669830283f858bfa3d696e0c34ad8855bbbb85cdd6ba

Observation 8a1668ba-09a1-4b4b-8c4b-55c34ccf21cb · outbound

This paper cites En- forcing geometric constraints of virtual normal for depth pre- diction.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors En- forcing geometric constraints of virtual normal for depth pre- diction

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.034622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.584342Z digest=sha256:b88b3561524a428153a6fa7086b146ca28fb547465a5f0e340c3859fe775b2e5

Observation d5764505-c86a-48da-9dd9-939a82876222 · outbound

This paper cites Rgb ↔x: Image decomposition and synthesis using material-and lighting-aware diffusion models.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Rgb ↔x: Image decomposition and synthesis using material-and lighting-aware diffusion models

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.018555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.588513Z digest=sha256:6d350af85f6f415e0db5ee85f9bce34965137add5bcf0d10fda8f3f0128fa154

Observation 2acbd3f7-9950-4622-82d2-f444f9d6f573 · outbound

This paper cites Hierarchical normalization for robust monocular depth estimation.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Hierarchical normalization for robust monocular depth estimation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:50.002051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.593053Z digest=sha256:5ab169914f37e7b60f1566f25c0e7c7396858d5423387f3de0f06a85d39fb309

Observation 0c1eebdc-368c-4d59-b42d-321006368869 · outbound

This paper cites Arf: Artistic radiance fields.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Arf: Artistic radiance fields

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:49.986336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.597204Z digest=sha256:85c60eca4163a96754c0b2dad2cf61529fd25763a34ab2d10a4e35ca0d3afd66

Observation 2a277178-70c8-4b7b-9787-2d804b36647f · outbound

This paper cites an unresolved cited work.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-12T12:26:49.970703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.601496Z digest=sha256:cd4f05433be71ffb1ad9a172a423225dcdb4358c07fefb8e3cebea8cb29e3f8a

Observation 1d80dda8-fece-4ef0-aa11-14c1ef3d8e8e · outbound

This paper cites We utilize the official implementations of Depth Anything V2 [56] and Marigold-E2E-FT [20], adapting temporal blocks from the UnetMotion architecture in the Diffusers [49] library.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors We utilize the official implementations of Depth Anything V2 [56] and Marigold-E2E-FT [20], adapting temporal blocks from the UnetMotion architecture in the Diffusers [49] library

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:26:49.955597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.605761Z digest=sha256:ffc35f58c466db355bac5d79a90a16f20794018ab7d91a10b997c6793b1416a9

Observation 3c3225c1-1da2-49aa-916e-77646c08e3d5 · outbound

This paper cites an unresolved cited work.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-12T12:26:49.938514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T12:26:49.610580Z digest=sha256:f8b505335d6a95d706c0bcb5fc4bb1741072bd5774581c7475ec7f622139fb96

Observation 55c7ddf7-4846-4016-ab8f-eb0e4350a923 · outbound

This paper cites an unresolved cited work.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Unresolved cited work

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.368602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.368602Z digest=sha256:9d48154ca7826b67791e6619aed33e173e0f791792a95fe2247eab446460c8ad

Pith citing papers

Observation c46a1f6e-6f13-4b8f-9a6c-98450bc7130f · inbound

Video Generation Models are General-Purpose Vision Learners cites this paper.

Video Generation Models are General-Purpose Vision Learners Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-13T00:56:18.867382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T00:56:18.867382Z digest=sha256:5b68c6d202cc1810384ef33cf9093506ee74f5562346f4346e99aa8b3de12dec