Pith. sign in

Paper Citation Record · LEDGER

Rethinking Encoder-Decoder Flow Through Shared Structures

As of 12 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2501.14535.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.14535 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:06:59.593165Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ab59d56d-c1e1-4879-8da0-ab204d237580 · outbound

This paper cites Vision trans- formers for dense prediction,.

Rethinking Encoder-Decoder Flow Through Shared Structures Vision trans- formers for dense prediction,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:07:00.018182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.463606Z digest=sha256:2996fb6816acb6460f6debb122c75490a440bd0644877091957e50b23d71b997

Observation 0042a627-bb2c-434b-992c-d92c6ee4683c · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Rethinking Encoder-Decoder Flow Through Shared Structures DINOv2: Learning Robust Visual Features without Supervision

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T15:06:59.468124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:06:59.468124Z digest=sha256:3377478ba34b656113bf2d29303e04748aeb0e95df48929fda1a4945347af61f

Observation f93e330a-4bd0-4eab-8d99-207733666ea2 · outbound

This paper cites RefineNet: Multi-path refinement networks for high-resolution semantic segmenta- tion,.

Rethinking Encoder-Decoder Flow Through Shared Structures RefineNet: Multi-path refinement networks for high-resolution semantic segmenta- tion,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.998882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.472325Z digest=sha256:178cd1d4624875266c18c26197fb89658922967476a8f139919c81e90aec56f7

Observation 060f2921-613e-41ea-b129-aac012b674ab · outbound

This paper cites Eva: Exploring the limits of masked visual representation learning at scale,.

Rethinking Encoder-Decoder Flow Through Shared Structures Eva: Exploring the limits of masked visual representation learning at scale,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.987597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.477122Z digest=sha256:c3d00ca54f61831bf0753755b077e26541ac9c4829d65766c97e05bc72f84450

Observation bbfad913-ac07-4424-9235-10b6153a80e4 · outbound

This paper cites Depth anything: Unleashing the power of large- scale unlabeled data,.

Rethinking Encoder-Decoder Flow Through Shared Structures Depth anything: Unleashing the power of large- scale unlabeled data,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.976797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.481829Z digest=sha256:4ff5463b0235211df0a3c83546bcc7e366e781a3d5b41d8b0113230fe2b23ae3

Observation 64afed73-17f2-4291-a194-3bd92e96a43c · outbound

This paper cites Repurposing diffusion-based image generators for monocular depth estimation,.

Rethinking Encoder-Decoder Flow Through Shared Structures Repurposing diffusion-based image generators for monocular depth estimation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.966993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.486243Z digest=sha256:f4d8137fa8b0552571e3d47b4e68ad6dd55c476dd04da03daeae2563e0af18d2

Observation 876ec342-8c97-4bbf-95ed-6b3fa1c90ba8 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Rethinking Encoder-Decoder Flow Through Shared Structures An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T15:06:59.491556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:06:59.491556Z digest=sha256:5fc46ff9ad4ce6b867277e7094563fad40d1937d8c46f6f41b60cfcebd88b460

Observation 3d4aab51-528c-4e60-91ac-5185b1e4ef12 · outbound

This paper cites MiDaS v3.1 -- A Model Zoo for Robust Monocular Relative Depth Estimation.

Rethinking Encoder-Decoder Flow Through Shared Structures MiDaS v3.1 -- A Model Zoo for Robust Monocular Relative Depth Estimation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T15:06:59.496317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:06:59.496317Z digest=sha256:b6221a22609f074dbc8ba419904940659fdb11799180c28b37a269be6bb2bea8

Observation 0e896ab9-8f0d-4412-b394-816cbfe07e82 · outbound

This paper cites ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth.

Rethinking Encoder-Decoder Flow Through Shared Structures ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T15:06:59.500958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:06:59.500958Z digest=sha256:76a62e247af02f998fc5a9631c5a07574cb0fabc073ca6fe4d2aea85fdfe6908

Observation 01cbb2bb-fb1a-4a58-9fff-0290ed21ff4c · outbound

This paper cites DepthFM: Fast Monocular Depth Estimation with Flow Matching.

Rethinking Encoder-Decoder Flow Through Shared Structures DepthFM: Fast Monocular Depth Estimation with Flow Matching

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T15:06:59.504757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:06:59.504757Z digest=sha256:ce33e94a956c63cedce4ba9af334f10c33d2f40bddb3803768979f4364edb43c

Observation c401638b-fbc7-4514-9d88-854889d52861 · outbound

This paper cites Bins- former: Revisiting adaptive bins for monocular depth estimation,.

Rethinking Encoder-Decoder Flow Through Shared Structures Bins- former: Revisiting adaptive bins for monocular depth estimation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.957239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.508892Z digest=sha256:fd84598ae99403b06415c32139af52321514852bd2fcd200242dd382bfdf7b7c

Observation 313e1eb8-e4cb-4400-985d-7215f594657b · outbound

This paper cites Ad- abins: Depth estimation using adaptive bins,.

Rethinking Encoder-Decoder Flow Through Shared Structures Ad- abins: Depth estimation using adaptive bins,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.947053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.515730Z digest=sha256:1220e3c1211eef14b041be5fab2681a604ec2bcfabb70eed4052e0f49948fab3

Observation 792b4371-b577-453f-b0ca-a62aee7388aa · outbound

This paper cites Shvit: Single-head vision transformer with memory efficient macro design,.

Rethinking Encoder-Decoder Flow Through Shared Structures Shvit: Single-head vision transformer with memory efficient macro design,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.935246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.518878Z digest=sha256:44e3423152835b51fa1b24aa3299e6858ce2a9ef05f0e9f6264b41efb2b2f3f7

Observation 7ed27f85-5c3a-4919-9096-5e8949479743 · outbound

This paper cites Fast vision transformers with hilo attention,.

Rethinking Encoder-Decoder Flow Through Shared Structures Fast vision transformers with hilo attention,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.923495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.522915Z digest=sha256:6908e833306efad79e831fa47c5825c7a6ccd9bdb4bf5111a9a2ea4e665b39d4

Observation d460affb-467a-4c40-8cb4-2ee8c556360e · outbound

This paper cites Swin transformer: Hier- archical vision transformer using shifted windows,.

Rethinking Encoder-Decoder Flow Through Shared Structures Swin transformer: Hier- archical vision transformer using shifted windows,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.911904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.526043Z digest=sha256:66086cfb892f35f787dd2e3975526c8e4c04491c084788da9b5fef0c55c4442f

Observation bb2f93d3-02d7-4c9f-a569-65a5434e0051 · outbound

This paper cites Swin-unet: Unet-like pure transformer for medical image segmentation,.

Rethinking Encoder-Decoder Flow Through Shared Structures Swin-unet: Unet-like pure transformer for medical image segmentation,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.899885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.530238Z digest=sha256:8c1cf066ad341a1d77d65c9bac0436523d2b8f211aa5662a8d30770708ae3fba

Observation 5e54d188-d6be-4535-a486-8631216ae583 · outbound

This paper cites Efficientformer: Vision trans- formers at mobilenet speed,.

Rethinking Encoder-Decoder Flow Through Shared Structures Efficientformer: Vision trans- formers at mobilenet speed,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.886239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.534916Z digest=sha256:48a8dc37c8842efb2f61b7765808c43ef5dbf62ce1ed1717413f8f6cd50da87a

Observation d45a0931-ebc7-485a-8c96-4a0c31fa138a · outbound

This paper cites Swiftformer: Efficient additive attention for transformer-based real-time mobile vision applications,.

Rethinking Encoder-Decoder Flow Through Shared Structures Swiftformer: Efficient additive attention for transformer-based real-time mobile vision applications,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.876699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.539526Z digest=sha256:102138b76ba99aca041e4755256f3e8e17ea7f3151fd5ddf391f8d0b18fa31fa

Observation b167f16a-a848-4c81-819e-8733562d28f7 · outbound

This paper cites Fastvit: A fast hybrid vision transformer using struc- tural reparameterization,.

Rethinking Encoder-Decoder Flow Through Shared Structures Fastvit: A fast hybrid vision transformer using struc- tural reparameterization,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.865084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.543625Z digest=sha256:e75ed6859a1d080d4af7d33927fbfcb561f0fcbe0344fb5362f1f2ce9f0e3162

Observation 269cce47-0bd5-417b-b29f-c81370b1a5a1 · outbound

This paper cites Repvit: Revisiting mobile cnn from vit perspective,.

Rethinking Encoder-Decoder Flow Through Shared Structures Repvit: Revisiting mobile cnn from vit perspective,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.853908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.547471Z digest=sha256:f7a491c6bba02b7b6f9636dc5053c9668290dca9114814cbdd4ab98e99235187

Observation 6a1f778a-f50b-4054-b837-341c6ca73cf6 · outbound

This paper cites Segformer: Simple and efficient design for se- mantic segmentation with transformers,.

Rethinking Encoder-Decoder Flow Through Shared Structures Segformer: Simple and efficient design for se- mantic segmentation with transformers,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.842777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.551479Z digest=sha256:021fd6be495f8733b7b446cd872b13a299ec3c249a0d9c89758afc734f5c2dd6

Observation fba264c2-d66b-4ac4-9ca6-3b0eecb59c35 · outbound

This paper cites Segmenter: Transformer for semantic segmentation,.

Rethinking Encoder-Decoder Flow Through Shared Structures Segmenter: Transformer for semantic segmentation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.830495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.555872Z digest=sha256:c14952c105d0e6171ac49998415016f0b2a17fc9bd785065e4f45311d2b2eeb0

Observation a46b2914-2327-485a-a791-6cce23382a2f · outbound

This paper cites More than encoder: Introducing transformer decoder to upsample,.

Rethinking Encoder-Decoder Flow Through Shared Structures More than encoder: Introducing transformer decoder to upsample,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.816265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.560947Z digest=sha256:7b983d73bc4b9fbd6344e8803aa52deb1940ffc0e80bf3088f2fe85b221de5e4

Observation 14cdf7e4-2780-454a-a415-fefcdba6bcf5 · outbound

This paper cites Denoising Vision Transformers.

Rethinking Encoder-Decoder Flow Through Shared Structures Denoising Vision Transformers

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T15:06:59.564819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:06:59.564819Z digest=sha256:cfbd4379d993e0ac757886513df1c7b83185e8b2f65a0cf97527b1eb2d5397ea

Observation 434e7862-d873-42d9-9f9f-8d3582019d0c · outbound

This paper cites Learning to upsample by learning to sample,.

Rethinking Encoder-Decoder Flow Through Shared Structures Learning to upsample by learning to sample,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.803185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.568972Z digest=sha256:bdf21f63efe13c14c6d807a2b59246bdcf2957be4b934fef7813aefe575bcba9

Observation d5b6f143-37b5-4c7b-8a89-43bb81a5d63c · outbound

This paper cites Imagenet-21k pretraining for the masses,.

Rethinking Encoder-Decoder Flow Through Shared Structures Imagenet-21k pretraining for the masses,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.791769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.572097Z digest=sha256:6611684dd6e5eac46ff90bb076c6b0397705b0baf5dfe45056db770165e51a54

Observation a6163c73-4605-44a6-8945-63adcf640615 · outbound

This paper cites Indoor segmentation and support inference from rgbd images,.

Rethinking Encoder-Decoder Flow Through Shared Structures Indoor segmentation and support inference from rgbd images,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.781025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.575338Z digest=sha256:49f5a7d9a14f3461a798d1d278223eea6f00cca5012d7c6b96775e8353077abe

Observation ee294dc1-6015-427e-9e52-bfb42f18b10f · outbound

This paper cites Learning the depths of moving people by watching frozen people,.

Rethinking Encoder-Decoder Flow Through Shared Structures Learning the depths of moving people by watching frozen people,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.768155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.578339Z digest=sha256:d7475d3ee5936cacf8edb446b0dfb751a5200a7c77202f76c6439c75e6308270

Observation b969f40a-ddf2-4f81-a8bd-9eb4f46e60e7 · outbound

This paper cites Irs: A large naturalistic indoor robotics stereo dataset to train deep models for disparity and surface normal estimation,.

Rethinking Encoder-Decoder Flow Through Shared Structures Irs: A large naturalistic indoor robotics stereo dataset to train deep models for disparity and surface normal estimation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.755471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.581949Z digest=sha256:e8b2d37997b7116cdeaa7073ec4106c58bb6e4928e3e5bc7af1e06e90a49e9f7

Observation 21cc9aba-e41b-4218-80f4-d3a6f9df9fa7 · outbound

This paper cites Hypersim: A photorealistic synthetic dataset for holistic indoor scene understanding,.

Rethinking Encoder-Decoder Flow Through Shared Structures Hypersim: A photorealistic synthetic dataset for holistic indoor scene understanding,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.742261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.585759Z digest=sha256:52131f4467ed589e3efcaa2bc4742ab1b81d83a78932b8218e2b861ff888327f

Observation 5b102dc6-554f-414d-97d7-97c3a8ac3b55 · outbound

This paper cites an unresolved cited work.

Rethinking Encoder-Decoder Flow Through Shared Structures Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-10T15:06:59.725072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.589399Z digest=sha256:ba27d41ca63c66d864ab0022bac9ecdc02c546574d6fb0869bc197dc1fe8cd47

Observation 3093b120-ccb8-41f2-8c36-5282fe4a2fd8 · outbound

This paper cites The open images dataset v4: Unified image classification, object detection, and visual relationship detection at scale,.

Rethinking Encoder-Decoder Flow Through Shared Structures The open images dataset v4: Unified image classification, object detection, and visual relationship detection at scale,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T15:06:59.710981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:06:59.593165Z digest=sha256:b5c49cc5c6cdc9e6784a6a167084369d9fd2f0d1d24137e09f7999d8e218826f

Pith citing papers

No inbound Pith citation observations are available.