Pith. sign in

Paper Citation Record · LEDGER

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion

As of 16 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2505.23085.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23085 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:59:28.175667Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-09T21:59:43.755956Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T14:21:06.813117Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy51
  • unresolved25
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 49941d42-b991-47f8-8fde-a70a8c25d5a9 · outbound

This paper cites Photorealistic monocular 3d reconstruction of humans wear- ing clothing.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Photorealistic monocular 3d reconstruction of humans wear- ing clothing

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:39.780982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:20.930873Z digest=sha256:4f6d7efd6ffd5fc5bacc20f3be992ebd5d4a7bd5c513b44929471dd8b3b1cd1b

Observation 367974e2-3c0b-494b-9592-b44c88e3edad · outbound

This paper cites Lumiere: A Space-Time Diffusion Model for Video Generation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Lumiere: A Space-Time Diffusion Model for Video Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.994210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.994210Z digest=sha256:e4e1e57d2589a0c6a15335acd89fdecb9bd426a699b794174b94cad4b49faa91

Observation 34b33b62-3e13-48f7-a25c-febeff7487d5 · outbound

This paper cites ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.107495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.107495Z digest=sha256:48227ca77b8f1f536c7bb31972d68f6f1d5c0acb2286ce730033896653ab3144

Observation 5eeb489c-94f3-4068-90fd-ef64ce9dc5bb · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.225920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.225920Z digest=sha256:69b0382df22c8017647e294b7086639f61e880bdd25fb3bdd9e813f096568505

Observation 761a963f-c4dd-4e56-8b3a-b1be18517d16 · outbound

This paper cites Depth Pro: Sharp Monocular Metric Depth in Less Than a Second.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Depth Pro: Sharp Monocular Metric Depth in Less Than a Second

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.343777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.343777Z digest=sha256:71f766a4e6c37eb3113f01225ca3bb13094986abf4b8e6772b40355257a5d8db

Observation 3f355d39-2a11-42f3-af29-3aed37b395c5 · outbound

This paper cites Video generation models as world simulators,.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Video generation models as world simulators,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.482823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.482823Z digest=sha256:411c1356fc5cb142824de15bf50facc0bda07c5457889b09da17dcd0d2300e6b

Observation 472d4005-bd5f-4931-80a7-7d08cb53fa49 · outbound

This paper cites High accuracy optical flow estimation based on a theory for warping.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion High accuracy optical flow estimation based on a theory for warping

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:39.611170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:21.585722Z digest=sha256:790620e5ba4ab5604bcfd0517873c2b91a699cd98cbbbef40b61f28431c9f891

Observation e60237ce-cecf-49a6-9e0a-a7594c0d0c74 · outbound

This paper cites Stable- video: Text-driven consistency-aware diffusion video edit- ing.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Stable- video: Text-driven consistency-aware diffusion video edit- ing

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:39.442026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:21.742575Z digest=sha256:09aad287d61b63ebab7fa2e73db18f3e440dcb53ae1ec9717e202aa598757556

Observation b1b1a472-c114-448c-b728-2b231a1625c6 · outbound

This paper cites VideoCrafter1: Open Diffusion Models for High-Quality Video Generation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.840777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.840777Z digest=sha256:138e687cf4d5489609ed062d5c7945ed02756d35b1eaa22c9e4dfd01ab6aff27

Observation f5894e7a-d201-4265-923d-2cb0cc62fd58 · outbound

This paper cites Videocrafter2: Overcoming data limitations for high-quality video diffusion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Videocrafter2: Overcoming data limitations for high-quality video diffusion models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:39.235242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:21.956046Z digest=sha256:dc6faae20a4c7d24f997720ce99969f2a8fa40497a9c88c6a300df2467cfcdff

Observation 4671f298-d307-4434-97d0-d02a079fa059 · outbound

This paper cites Self-supervised learning with geometric constraints in monocular video: Connecting flow, depth, and camera.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Self-supervised learning with geometric constraints in monocular video: Connecting flow, depth, and camera

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:39.104114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:22.077547Z digest=sha256:7b077cd3a081367d6e08bacd9ee9abb0eba9df9c8c5f9e7b25f74df1b7adfa68

Observation 29448567-24b0-4ad9-8531-348f540a34cc · outbound

This paper cites Cogview2: Faster and better text-to-image generation via hi- erarchical transformers.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Cogview2: Faster and better text-to-image generation via hi- erarchical transformers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.826291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:22.208238Z digest=sha256:1ff57602fd671d3b33928945a3197831e1aa8cf623e4cfc30bb5fb4b77d52944

Observation a65020e1-9b5f-4763-be62-1b7958887b6b · outbound

This paper cites Depth map prediction from a single image using a multi-scale deep net- work.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Depth map prediction from a single image using a multi-scale deep net- work

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.623717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:22.296951Z digest=sha256:dd7bc2db38ab85be971fde2861fa578275e2f49b453d10f58dc0b5493e3e485e

Observation 330baaba-8581-4169-b767-c241eb71aedc · outbound

This paper cites Structure and content-guided video synthesis with diffusion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Structure and content-guided video synthesis with diffusion models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.389610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:22.365317Z digest=sha256:fd3e86eab5d82ca6d1e98ac227498ca098995e9bd53e680e2feaa930696153bc

Observation 098ae50c-bc3d-4f47-a039-bf1d6b073d83 · outbound

This paper cites GeoWiz- ard: Unleashing the diffusion priors for 3d geometry estima- tion from a single image.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion GeoWiz- ard: Unleashing the diffusion priors for 3d geometry estima- tion from a single image

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.252183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:22.458937Z digest=sha256:c1c080f1c1973696fbd946f9550ab0a3268212fd07fb8047884651703d9175f3

Observation a57da66e-926d-49ab-91ad-d91136b814be · outbound

This paper cites Humans in 4D: Re- constructing and tracking humans with transformers.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Humans in 4D: Re- constructing and tracking humans with transformers

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.131541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:22.597289Z digest=sha256:2e3ed66e938bd0647a81cb74472b6a3869ca8a46f79dee384cd20157c6326725

Observation a3564360-2e18-4929-ad3f-7e35594a9d46 · outbound

This paper cites High-fidelity 3d human digitization from single 2k resolution images.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion High-fidelity 3d human digitization from single 2k resolution images

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:38.024814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:22.692204Z digest=sha256:97507d50daeb509f707575a408ce0a7d5d575a0f5427f5e70cdebf8476957db2

Observation 180f7715-6b33-47ed-a81d-1f40606278b3 · outbound

This paper cites Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Lotus: Diffusion-based Visual Foundation Model for High-quality Dense Prediction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.779637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.779637Z digest=sha256:25afee54640bc6eeb7f4f8f7c3b938a3c4fa6338ed659bc0b86d67927e4d0216

Observation 862928a1-904f-4c57-acad-c5d91e36df0f · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.901882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.901882Z digest=sha256:05147ca71cb382e2314fbed3f92fd2de477b37dda9087b0fe5d6cbd59cfd401b

Observation 38c55ea5-a2ac-40b3-b5b4-e22544322618 · outbound

This paper cites Denoising diffu- sion probabilistic models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Denoising diffu- sion probabilistic models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:37.901721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:23.018296Z digest=sha256:f76fe2a909fe7d4c6e93177483073ac1d49e6b91208332ec030fd56b86784454

Observation ff12a275-d4bb-475c-b9d0-8c0273694b09 · outbound

This paper cites Video Diffusion Models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Video Diffusion Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:23.141059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:23.141059Z digest=sha256:35a5f0c3935981480251c37601a24dc798831cb7ecd81c80ddd6550721097245

Observation 1d0655c6-1b64-4688-905f-55c08ea4372e · outbound

This paper cites Animate anyone: Consistent and controllable image-to-video synthesis for character animation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Animate anyone: Consistent and controllable image-to-video synthesis for character animation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:37.784968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:23.223405Z digest=sha256:099e1ffbbf09f75b6c18f819bdd091f776964f4ee05a4d3ee76a8b32aa111ba9

Observation 4fe6de2f-4a3e-4eff-b9fd-a9a91d1f95da · outbound

This paper cites Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:23.363527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:23.363527Z digest=sha256:315fddd96a02d69cd84a0b3e735735fdb93cf48bbc9f3c832e6201b8688aa94b

Observation c9d314d8-7044-45ca-a973-8831b2525b77 · outbound

This paper cites DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:23.490758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:23.490758Z digest=sha256:607881f825f3229a9b78d32105da1ea1cabf2bb9acbf22cefc92747bd07e6723

Observation 5f48e3d9-f51f-44ff-a2c0-eb692f198ae3 · outbound

This paper cites ARCH: Animatable reconstruction of clothed humans.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion ARCH: Animatable reconstruction of clothed humans

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:37.448315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:23.710031Z digest=sha256:36c4cad2558fb0980587155a5f3e3baff69c96205df9259c5029e4899f1e76a0

Observation 3ce18d9f-1d23-4551-b295-da670be10b2d · outbound

This paper cites Flowformer: A transformer architecture for optical flow.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Flowformer: A transformer architecture for optical flow

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:37.215736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:23.844526Z digest=sha256:4f8c2476a2b7b1624ed924aca9b98e37dc3c53afce622b78f952dedb2685a14e

Observation 5e3d4528-7f7b-4b23-a769-4f5f35e85eb1 · outbound

This paper cites HumanRF: High-fidelity neural radiance fields for humans in motion.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion HumanRF: High-fidelity neural radiance fields for humans in motion

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:37.098110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:23.941445Z digest=sha256:b85af4e943642db50d8d0a61b5193506cebc37d737bbb42d7b0220f78d46f201

Observation 5edd3c3b-c62a-4836-bcb2-b0bc54bc82f6 · outbound

This paper cites Jafarian and H.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Jafarian and H

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:36.950238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:24.038240Z digest=sha256:b51a9792a7c9902c76432cdf5396cddd0b3ef879d46e896482d37acd428d395d

Observation 65ad9760-e6e4-4e9b-a895-13947f567f5e · outbound

This paper cites Learning high fidelity depths of dressed humans by watching social media dance videos.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Learning high fidelity depths of dressed humans by watching social media dance videos

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:36.671335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:24.158543Z digest=sha256:d91c8ee59de7e521b98f390aa690151293b4ee1adf8cb5d9d6565641e21cca4a

Observation 7ef05ac4-86e6-461b-9ca2-c88b95a586d6 · outbound

This paper cites Repurpos- ing diffusion-based image generators for monocular depth estimation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Repurpos- ing diffusion-based image generators for monocular depth estimation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:36.532733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:24.272479Z digest=sha256:dae3dea0c29a8e10e3b9ac8258f6b20139e24b8b95f850a036bece901f6b02af

Observation 1b188c7c-7d3d-441e-a9c9-9d1942134c17 · outbound

This paper cites Sapiens: Foundation for human vision mod- els.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Sapiens: Foundation for human vision mod- els

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:36.304940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:24.420233Z digest=sha256:94538a9149fa6aacf42228945df6cd08fc4467b2248ce9af04b2ff48894278fb

Observation ddab9752-2070-4210-beda-baa514076afb · outbound

This paper cites Auto-Encoding Variational Bayes.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Auto-Encoding Variational Bayes

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:24.519563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:24.519563Z digest=sha256:1eb74e6c5270923454046635ea4b4beabb8e559a11de9cedacab10a33bf45cdb

Observation 0eacca16-ca2e-49ed-9b1a-b7d76a46f35d · outbound

This paper cites Adam: A Method for Stochastic Optimization.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Adam: A Method for Stochastic Optimization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:24.610161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:24.610161Z digest=sha256:3403efc2217ca59d55192b140384196400329b1596fd9199db0137a24ad80816

Observation c02b8bbd-ff08-46f6-8dd8-0878a7f494a1 · outbound

This paper cites Robust consistent video depth estimation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Robust consistent video depth estimation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:36.168095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:24.706712Z digest=sha256:7dc36878c790e8a65b3a1c9cff0a8b9f323d6f742227b39d0215bd29411d9b1a

Observation 9156ee3e-aa92-4401-ac86-7428000a56ec · outbound

This paper cites Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion Model.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:24.778818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:24.778818Z digest=sha256:2ce195c634c8c05d24d255b7365ce83240fefb43cba1de1b5f3feddae4b2c582

Observation aca88010-beee-4e50-8f11-1441357d59a2 · outbound

This paper cites an unresolved cited work.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:35.989939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:24.879018Z digest=sha256:680c6171a1c82a56242e4cfda06d9720e6daa0f93ff03f2c15aa095fb71b004d

Observation 83973606-f072-4174-b1ad-cadb94cd9cb4 · outbound

This paper cites Consistent video depth estimation.ACM Transactions on Graphics (TOG), 39(4), 2020.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Consistent video depth estimation.ACM Transactions on Graphics (TOG), 39(4), 2020

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:35.783342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:24.977142Z digest=sha256:e6db45a2eb7cb9d3b7293546b2235b8e332244df70eeceb88ed144a553555d96

Observation cbc13d4e-e859-4ed7-9445-24b28b22e216 · outbound

This paper cites Jewett, Simon Ven- shtain, Christopher Heilman, Yueh-Tung Chen, Sidi Fu, Mo- hamed Ezzeldin A.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Jewett, Simon Ven- shtain, Christopher Heilman, Yueh-Tung Chen, Sidi Fu, Mo- hamed Ezzeldin A

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:35.523612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.067043Z digest=sha256:ae37f97bdeccfec833f8184e8710080c6a863ae5f9b12fa39bc1a980dc0ec9ce

Observation 867a13c8-b40c-4a9a-9d4a-e4b41e60cd48 · outbound

This paper cites an unresolved cited work.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:35.350523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.114459Z digest=sha256:b5c491f0e1cb1e1f7812f4f00a7dd3e8e4310307ead3e4fe67b126a590944879

Observation d4af9a8c-a7b2-47d6-97dc-f46fcd957eee · outbound

This paper cites UniDepth: Universal monocular metric depth estimation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion UniDepth: Universal monocular metric depth estimation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:35.140737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.183400Z digest=sha256:a3494f6e40dad4a296c905c403143e3b5726f42c118686ff6a146d2808ce0305

Observation 2987710d-1973-4166-aa8a-dea10b3e31b0 · outbound

This paper cites SDXL: Improving latent diffusion models for high-resolution image synthesis.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion SDXL: Improving latent diffusion models for high-resolution image synthesis

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:34.925641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.231414Z digest=sha256:01783fea8638b09d1d45d80180f530b42a3c947085dd2e3932dd6cb438e8dc8b

Observation b8fc48b4-11c9-4495-9c15-5f78d9a96f20 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:25.277658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:25.277658Z digest=sha256:0b22483058e99dbba3693cbc18390b94a7613221945ea1ce83e081981e565144

Observation acec3087-4f57-44b1-b37e-effc5550f785 · outbound

This paper cites Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:34.651003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.340557Z digest=sha256:7556d18a80c239e7d22b19a6e8258d405d85f647308de50202026c0b3a34c24c

Observation 0ca20f74-3fb0-48dd-81a7-26e57ddd8f05 · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion High-resolution image syn- thesis with latent diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:34.524732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.393763Z digest=sha256:9be5ca9fdbb49fc0f4acd8bf8200d9c202af1b621988448509fe5025095eb06c

Observation 4f78e827-66ea-4b05-8b91-f2a5c145f163 · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:34.315839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.463051Z digest=sha256:e0bdc89bba9eecf2ca7aa6f0f9bee33de924e3fcc37745f745b1e178e25904ee

Observation f860c84d-9f30-4a5c-aa7d-1be88a0bd5ae · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Photorealistic text-to-image diffusion models with deep language understanding

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:34.127227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.530959Z digest=sha256:37007a0dc6ac509ab4f81dfd8f8a30efd5d7531b246bc72b95d4da579bb0903b

Observation 862d2a01-f574-4b15-a755-2bb648e0e223 · outbound

This paper cites Pifu: Pixel-aligned implicit function for high-resolution clothed human digitiza- tion.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Pifu: Pixel-aligned implicit function for high-resolution clothed human digitiza- tion

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:33.956629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.621719Z digest=sha256:f283d99d306ffbfe1bf4473a950aa0558042c969b8bedea4839c08ce5c089c5c

Observation a0017d01-24ec-48c7-a2be-393dbf3feb68 · outbound

This paper cites Pifuhd: Multi-level pixel-aligned implicit function for high-resolution 3d human digitization.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Pifuhd: Multi-level pixel-aligned implicit function for high-resolution 3d human digitization

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:33.776720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.680799Z digest=sha256:94c6fa10b550cb4b4855bafb282a4323aaab26296088392d4a6d45f4db06d8b7

Observation 88961aa5-be61-49d3-ac5a-ce28397d65a6 · outbound

This paper cites Progressive distillation for fast sampling of diffusion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Progressive distillation for fast sampling of diffusion models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:33.632255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.747364Z digest=sha256:4a2f0f0dbf81ce63c2c0bfeb5f6101c9be581c20116548841ac21b9a24a23945

Observation 98bcd4a8-55b4-47ab-b3eb-8e6b35804c24 · outbound

This paper cites X-Avatar: Ex- pressive human avatars.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion X-Avatar: Ex- pressive human avatars

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:33.395997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.834387Z digest=sha256:0844a7392200642dfe1635fce4c079b6e2d22c25013dbd5fcbcf2eb26eb0a7ad

Observation 1d8443e4-d83a-4c5a-85bd-205c0c06c2cc · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Deep unsupervised learning using nonequilibrium thermodynamics

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:33.116872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.919767Z digest=sha256:de5899fa0a9d5d687e2999146fe8264cf1ff2d46d9385ed8ca46fb3cfeb12766

Observation ff0dd391-6b33-425b-a5b4-453ced72beaf · outbound

This paper cites Diffusers: State-of-the-art diffu- sion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Diffusers: State-of-the-art diffu- sion models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:32.825490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:25.970093Z digest=sha256:5b4d50b8da081444fb710d9386ba0eb8288f1cf658817b363627aa233a5a5683

Observation 961a342c-79f9-4bc0-a26a-3397d5b7b854 · outbound

This paper cites Modelscope text-to-video technical report.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Modelscope text-to-video technical report

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:32.568013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.032794Z digest=sha256:75e14a20a6447f8d83c9561aec9e769a8a213dcf9ef7ccaf77b65b4f232c97a3

Observation 58349eb9-8492-4cc6-9e65-5e30617d8781 · outbound

This paper cites 4D-DRESS: A 4d dataset of real-world human clothing with semantic annotations.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion 4D-DRESS: A 4d dataset of real-world human clothing with semantic annotations

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:32.308039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.071652Z digest=sha256:420cd9109389de389f6b33a3d34e9da4a4c3ed239ed069de974c13a3732bc07c

Observation 1fa9faa1-24e3-409a-933b-90eb67cd7ed2 · outbound

This paper cites MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:26.160150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:26.160150Z digest=sha256:913bd140f1831b072135bffbe0040447ef008b024c4c23b4a9d6f8641e6bc182

Observation be46c9aa-c0a1-438e-86d5-da82c6780477 · outbound

This paper cites Videocomposer: Compositional video synthesis with motion controllability.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Videocomposer: Compositional video synthesis with motion controllability

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:32.016949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.205797Z digest=sha256:22fabe770a2a0217fa6010f5966a878b4218360afd59e969eaa9768682418acd

Observation dfeaf70c-1b4a-4219-b331-b41c37b6b0db · outbound

This paper cites Less is more: Consistent video depth estimation with masked frames modeling.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Less is more: Consistent video depth estimation with masked frames modeling

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:31.737807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.280443Z digest=sha256:713aa454a74dae3b07a2caeb1f3ac33116a9dc56a3d8c8f8c81efc026cd27755

Observation edd21a6f-b0df-4956-b331-04c9a679840d · outbound

This paper cites TRAM: Global trajectory and motion of 3d humans from in- the-wild videos.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion TRAM: Global trajectory and motion of 3d humans from in- the-wild videos

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:31.453658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.351944Z digest=sha256:c1532953d01c9b5e3892d63eedba4c20f107113422888f258c2bdea27e99f647

Observation 0d9b665b-ae44-48b1-9c76-2fcb8cd12e02 · outbound

This paper cites ICON: Implicit clothed humans obtained from nor- mals.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion ICON: Implicit clothed humans obtained from nor- mals

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:31.252652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.425490Z digest=sha256:9c6ada3994fbcda94a5f577e249a018d825707d0fdc5c902b3aaa96c7bad2b71

Observation dce71b8a-2f4a-43a4-9806-2ef4bf7cf110 · outbound

This paper cites an unresolved cited work.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:31.012696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.491077Z digest=sha256:4c4589bfc9f4b9b41974386d2ba514b22558c22dfd2e8340de300bfa9b0d35e7

Observation fbfb37ba-2077-416a-9b13-f6c4e8fad643 · outbound

This paper cites Depth Any Video with Scalable Synthetic Data.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Depth Any Video with Scalable Synthetic Data

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:26.572903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:26.572903Z digest=sha256:529e3f5cb0c5efb20d852fed48b2677ff7d915dddf3281f856edc87a6b77de16

Observation 907ed995-7673-44d8-b6c4-7662a0a9e360 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Depth anything: Unleashing the power of large-scale unlabeled data

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:30.804175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.634045Z digest=sha256:8d95dc028a8f87f55b671913048f863cbd419c357f83bad82f7ce76750b57eb5

Observation 2fda88d2-6e6f-482a-ad69-38d26ce581be · outbound

This paper cites Depth Anything V2.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Depth Anything V2

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:26.713658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:26.713658Z digest=sha256:20e2072dfcc677d9edc4d40853a0f8c4049c8f20315e8d4d2214f57b18c35de1

Observation d8e321ed-141b-4bb9-b7bb-2a68a195f911 · outbound

This paper cites Metric3d: Towards zero-shot metric 3d prediction from a single image.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Metric3d: Towards zero-shot metric 3d prediction from a single image

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:30.592483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.790452Z digest=sha256:021c55c34b84b5b009b54444338dccf12e8a8be4681352ff962c489a0a696de2

Observation 5e4588df-5554-48bc-9615-3cc5b313e01b · outbound

This paper cites Function4d: Real-time human vol- umetric capture from very sparse consumer rgbd sensors.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Function4d: Real-time human vol- umetric capture from very sparse consumer rgbd sensors

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:30.362228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.845197Z digest=sha256:1d73d8fc191f2ccb8df3361edf852150dcb7a0353e13ecc0901c7643ef681302

Observation 3cd973db-6364-460c-b8e1-6482f89fcd0b · outbound

This paper cites GLAMR: Global occlusion-aware human mesh recovery with dynamic cameras.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion GLAMR: Global occlusion-aware human mesh recovery with dynamic cameras

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:30.093061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:26.919244Z digest=sha256:537208c33ecb232e254467551cf657493bc3ea96e27548a097f97c29e8d4d99a

Observation d0ff54db-8de9-4f3a-984c-bd65f7edc42b · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Adding conditional control to text-to-image diffusion models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:26.972501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:26.972501Z digest=sha256:9a40cf0a697f2e75ccd41e52ff923dca0b2b68eac7f8b481d212eb3257024b57

Observation 5f650424-84f2-4430-97a0-a5a51ea0005b · outbound

This paper cites Adding conditional control to text-to-image diffusion models.ICCV,.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Adding conditional control to text-to-image diffusion models.ICCV,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:27.029876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:27.029876Z digest=sha256:f1903c62233b96fb23d68bdad7487c90e2e5a761186a13ea8236c21c1bd7e990

Observation 5697c2bc-bf4a-4209-8275-bec6e5e5dd87 · outbound

This paper cites IC-light github page, 2024.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion IC-light github page, 2024

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:29.954643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:27.085835Z digest=sha256:a620eb91f996dc2412b243247597ce33af6b686e4b204342f72978c2c1e75ff9

Observation c4cc89b7-8c1b-4a7f-a543-4af6c43f4f6a · outbound

This paper cites I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:27.197607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:27.197607Z digest=sha256:d51682e3382e4f2801537be76c85a8dbcfd3f12d4ee645dee73ad6a3e8951fa9

Observation 551c7ffe-eca2-4e39-9245-198e351bb707 · outbound

This paper cites Consistent depth of moving objects in video.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Consistent depth of moving objects in video

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:29.688301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:27.290642Z digest=sha256:52c1c9d83f8470cc0547b0c91f41dde09489be54fa349244c202e63cee2d78e4

Observation f2299730-b323-4bea-bfea-67b1ec70959b · outbound

This paper cites SIFU: Side- view conditioned implicit function for real-world usable clothed human reconstruction.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion SIFU: Side- view conditioned implicit function for real-world usable clothed human reconstruction

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:29.466161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:27.465300Z digest=sha256:a31172797974e8a65a983b1e6e8914a7b56b6dc70bed9605509a27987f0a89f1

Observation 3f8c97b7-cf18-42a5-91c6-8db32c16e8c8 · outbound

This paper cites Bilateral refer- ence for high-resolution dichotomous image segmentation.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Bilateral refer- ence for high-resolution dichotomous image segmentation

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:29.220649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:27.622635Z digest=sha256:cfeee16c07398a98c63ea0443ff65afc70b49f6bed04b99994c436f254d4974d

Observation 26fc7723-ac14-4cb9-9337-5a7c06e93f79 · outbound

This paper cites PaMIR: Parametric model-conditioned implicit representa- tion for image-based human reconstruction.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion PaMIR: Parametric model-conditioned implicit representa- tion for image-based human reconstruction

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:29.038110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:27.745112Z digest=sha256:f0a008d3ec139d12aa587ae428dfa9edba7d34e32c94401fb6972ad21956557c

Observation 4182cb87-e610-47f8-9822-d9ce8182df6b · outbound

This paper cites MagicVideo: Efficient Video Generation With Latent Diffusion Models.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion MagicVideo: Efficient Video Generation With Latent Diffusion Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:27.891963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:27.891963Z digest=sha256:4e88be69a2a9c7157035862ef745e9b284983695814240069060978652f7f5d7

Observation 838da50b-0235-417d-b38b-897d9c48bdd3 · outbound

This paper cites “1q ř ktPK,dtPDt ››dt´ dgt t ››2 RMSEplogq: b 1řpKt““1q ř ktPK,dtPDt ››log dt´ log dgt t ››2 δă thr: 1řpKt““1q ř ktPK,dtPDt Kt.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion “1q ř ktPK,dtPDt ››dt´ dgt t ››2 RMSEplogq: b 1řpKt““1q ř ktPK,dtPDt ››log dt´ log dgt t ››2 δă thr: 1řpKt““1q ř ktPK,dtPDt Kt

Reference 77

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:59:28.872014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:28.006204Z digest=sha256:a998a107258f76001b81e3ba6333d1c01cd36672b426401ebbb0b4ae6fd38872

Observation 1cc54783-db91-4d93-9ed1-e99a46b31234 · outbound

This paper cites 24 Figure S18.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion 24 Figure S18

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:28.563642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:28.175667Z digest=sha256:f2ab6a85d477ea29ea5ffb9c0499404c309fed1eaacf7e514f7dc298e2674b4e

Observation 3644b666-5d63-48a1-be02-28e9bf3e2bb4 · outbound

This paper cites an unresolved cited work.

GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion Unresolved cited work

Reference 2024

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T12:59:37.639984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T12:59:23.588927Z digest=sha256:d166020f7a0cc8bf5a7a7ef804c8460e1796bec74838a0dcfc6d1de4779d338b

Pith citing papers

Observation e42001aa-f210-455b-b801-30b8bb9a1ca1 · inbound

Sapiens2 cites this paper.

Sapiens2 GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:21:06.816106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-09T21:59:43.755956Z digest=sha256:c5a08c11634e4c25a39525e82904ca7e6085b5fc2d423f308a7911d8842b05b0