Pith. sign in

Paper Citation Record · LEDGER

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching

As of 10 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 1 inbound Pith citation observation for arXiv:2507.10318.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10318 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:39:29.580035Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T14:24:49.968980Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact2
  • verified fuzzy48
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4bc141d3-7675-4a91-93f2-d87a7f3465e5 · outbound

This paper cites Burst: A benchmark for unifying object recognition, segmentation and tracking in video.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Burst: A benchmark for unifying object recognition, segmentation and tracking in video

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:37.856565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.108553Z digest=sha256:ed9c58ab0ce0d63e93599666d1285328da2f7cc58320d79175969ebdf10ad248

Observation 756c8d63-ff4b-467d-9dd6-15e914fccb81 · outbound

This paper cites Hpatches: A benchmark and evaluation of handcrafted and learned local descriptors.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Hpatches: A benchmark and evaluation of handcrafted and learned local descriptors

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:37.846283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.201361Z digest=sha256:b6a01dfb98fd8665aa27744813d305c309852e21713304d55ac2b6225913982d

Observation 24de8596-906f-4ce2-bd45-d38b533b1ad8 · outbound

This paper cites an unresolved cited work.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:39:37.778915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.298170Z digest=sha256:254d5daef16e411bcd2e48a09e10c068e3fe4ab5035c06b17a31594a4a38c265

Observation a2ca3393-b810-4c80-97cd-4e163f48e6bc · outbound

This paper cites Surf: Speeded up robust features.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Surf: Speeded up robust features

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:25.377933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:25.377933Z digest=sha256:84671d847754c638c2647c4e3a9938aeed92055b40e4c64883db8cab4ec640be

Observation 904d3fe5-893a-4e2d-bcfd-ce89a4d86e40 · outbound

This paper cites Speeded-up robust features (surf).

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Speeded-up robust features (surf)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:37.567856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.474100Z digest=sha256:15f2ec03deab8beb102111f4f3cb4d80cdf85a07b0412ca4982d50d18807506e

Observation 84983732-8d76-4fd7-a9e9-a9d50bf1c739 · outbound

This paper cites VOLoc: Visual Place Recognition by Querying Compressed Lidar Map.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching VOLoc: Visual Place Recognition by Querying Compressed Lidar Map

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:39:29.984495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.536456Z digest=sha256:6e7ae8b327e5758e9ef5cc347fbdd381b7b6b19f5dcbf04f1fe49a025bf039a7

Observation dc0dafdf-0a80-4042-9b63-3a67dac07e34 · outbound

This paper cites Prism: Pro- gressive dependency maximization for scale-invariant image matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Prism: Pro- gressive dependency maximization for scale-invariant image matching

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:37.392675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.625802Z digest=sha256:2d815b088aa68629bdb8b50a09017b177b3d51cbb96419c233c6f4dd90f20197

Observation d33e7e9d-8a53-48c3-ab78-6d6a09fea765 · outbound

This paper cites Improving transformer-based image matching by cascaded capturing spatially informative keypoints.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Improving transformer-based image matching by cascaded capturing spatially informative keypoints

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:37.154073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.728919Z digest=sha256:60ddcd65c5133665c2c1bfaaf9fde6e59a0e0fe4162a96b50a3596ae3c74493c

Observation 35df9448-77ba-495d-8ec8-2a9c4003d9dc · outbound

This paper cites Aspanformer: Detector-free image matching with adaptive span transformer.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Aspanformer: Detector-free image matching with adaptive span transformer

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.959512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.775270Z digest=sha256:c3c65ab522ba9b4f84a0b994a1d3e10ca402698c4934df5b91841e075fefb250

Observation e43532a9-5957-4d6b-b6bf-ca6b5f0704d0 · outbound

This paper cites Ecomatcher: Efficient clustering oriented matcher for detector-free image matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Ecomatcher: Efficient clustering oriented matcher for detector-free image matching

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.754168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.817823Z digest=sha256:cf53873d525ee641005cc646c72db79ddb8e4cca3381296fe63d88388116ad5e

Observation a9318b76-92f6-4235-9996-27774f6251cc · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.560632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.855956Z digest=sha256:befebb1ea6c72639f7f4d772e12b3c38b3b233f552a9c7b4e0aaf727b704d376

Observation 070a27cd-51d9-4b48-b33c-3588af7b8a5c · outbound

This paper cites Superpoint: Self-supervised interest point detection and description.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Superpoint: Self-supervised interest point detection and description

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.378666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.919632Z digest=sha256:dbd6e8ae14eff1e0be7c05b1c66783c949094977e616cd71932baea63fb09e6d

Observation 1c8f08ac-794d-40a3-8521-6addefadc609 · outbound

This paper cites Diffusion models beat gans on image synthesis.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Diffusion models beat gans on image synthesis

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.200418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:25.983654Z digest=sha256:ec2e92e3df0ec2dee4b929e53324a5cd050c969f5ca2f578d14818d91cde405f

Observation da261533-5483-4852-8b58-cb367799727c · outbound

This paper cites Dkm: Dense kernelized feature matching for geometry estimation.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Dkm: Dense kernelized feature matching for geometry estimation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.103809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:26.075536Z digest=sha256:d70224e414b14513cf8cd2cbe33b4949a4024fb11822f7cc534f725f6c260abe

Observation a12f004c-20c9-4518-b327-c2b995336d2b · outbound

This paper cites Roma: Robust dense fea- ture matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Roma: Robust dense fea- ture matching

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.972971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:26.145757Z digest=sha256:ac8c9eb7ad52af5a04cbb3d87257bb74d406279efae6f0332f2941b5bc2d5af5

Observation bbe42dae-55cc-4b36-9f7c-861494824c09 · outbound

This paper cites Top- icfm: Robust and interpretable topic-assisted feature match- ing.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Top- icfm: Robust and interpretable topic-assisted feature match- ing

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.824253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:26.220843Z digest=sha256:d86d982921d389a49c4d908f100d5aa91b32399d46c20dbbec56c3f1201be59c

Observation 4548f61f-6638-4d1e-b0b5-b6cc1fe6d4f7 · outbound

This paper cites Deep residual learning for image recognition.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Deep residual learning for image recognition

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:26.290930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:26.290930Z digest=sha256:d70054aecd889756b605ea88a0b241a86319099028161e50317aa68a1ba21154

Observation d3577e46-a4a3-4ed2-9e62-011b3aa12c3c · outbound

This paper cites Masked autoencoders are scalable vision learners.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Masked autoencoders are scalable vision learners

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.675513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:26.392578Z digest=sha256:d0a8c8d2380c3c4fde8f7d82592f350dd004e9f170061c1b882e33d6fdba291a

Observation e8c594b2-c443-471c-a17e-5569b80f6f26 · outbound

This paper cites Denoising dif- fusion probabilistic models.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Denoising dif- fusion probabilistic models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.517902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:26.469629Z digest=sha256:cbfb64583e5a75b3f43bcfd6918a36085b91482c97ef4a4fa34c873d186621d3

Observation 4c876dab-15d7-424d-a0be-6d4397234437 · outbound

This paper cites Dynamicid: Zero-shot multi-id image personalization with flexible facial editabil- ity.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Dynamicid: Zero-shot multi-id image personalization with flexible facial editabil- ity

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.300228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:26.535857Z digest=sha256:1e9451e842856aff83a13015a0756febd3bc84d17c3d2017f8458166e0a3ebb6

Observation c7ec0d92-1951-4c36-a0cf-b39b8651e222 · outbound

This paper cites Omniglue: Generalizable feature match- ing with foundation model guidance.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Omniglue: Generalizable feature match- ing with foundation model guidance

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.162338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:26.616999Z digest=sha256:892db62ddbc10eda2250aad8e71bc5a702572fa014eb7a5abd083807af4428b6

Observation 55c6ccaf-9893-4b15-a619-1b6ce19ca8e3 · outbound

This paper cites Segment any- thing.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Segment any- thing

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:26.711760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:26.711760Z digest=sha256:364a69f382a94ec9226eee727d5846882bdf738a8b599b00686420db471dd4eb

Observation a95c3a97-891d-4892-8be2-85a0a65f7c1a · outbound

This paper cites Sd4match: Learning to prompt stable diffu- sion model for semantic matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Sd4match: Learning to prompt stable diffu- sion model for semantic matching

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.017486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:26.781562Z digest=sha256:c77ca2aee970bce7efb8a92e9df801e2315df40d1ede1c7d0e5d58564d9ff99c

Observation 03a73a54-fa32-40d8-9682-852623082ad7 · outbound

This paper cites Megadepth: Learning single- view depth prediction from internet photos.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Megadepth: Learning single- view depth prediction from internet photos

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.858580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:26.854437Z digest=sha256:f62f4896c012adb9429defab27c7c91d2b9bdae077f660a6733a59201b0e6922

Observation b140ce15-c14e-40e1-8239-8b8992a360dc · outbound

This paper cites Recrecnet: Rectangling rectified wide-angle images by thin-plate spline model and dof-based curriculum learn- ing.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Recrecnet: Rectangling rectified wide-angle images by thin-plate spline model and dof-based curriculum learn- ing

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.697273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:26.915593Z digest=sha256:bb5e07209e61c2e55f10294de1684eb7b8ac5068d50afea6a57e04be8f968d73

Observation cbb360a1-5166-40d3-a3a8-e95e79844a2e · outbound

This paper cites Feature pyra- mid networks for object detection.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Feature pyra- mid networks for object detection

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.562423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.008107Z digest=sha256:a266ceba30e4eef4fec662d5eb1f3370c4c18e2303c7fed41e8bca46dd16bf09

Observation 0e3b00dc-2f28-4bae-a4df-b1f79a9861ef · outbound

This paper cites Lightglue: Local feature matching at light speed.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Lightglue: Local feature matching at light speed

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.421039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.078922Z digest=sha256:875de56e6acf296427b35b7dd7c4133804b387d8944de2f6627ac38a479af7d8

Observation a27d2dc7-f8b6-49cf-8f23-14d8b8dfff31 · outbound

This paper cites Semantic- aware representation learning for homography estimation.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Semantic- aware representation learning for homography estimation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.293249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.136387Z digest=sha256:09dcbfaac1da29c26453d5f0f0ac8d6d8c9ed9a6a6ca5120aacf3e7d4597f96c

Observation 67a59f89-9e45-488f-9897-935d63b327c4 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Swin transformer: Hierarchical vision transformer using shifted windows

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:27.196330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:27.196330Z digest=sha256:1aeb882b52c3df7a0836d7d4e0c723f20e163f412086d78e9891085e1275b4d8

Observation 94568ca6-1309-498e-97d4-f7b6be224f49 · outbound

This paper cites Distinctive image features from scale- invariant keypoints.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Distinctive image features from scale- invariant keypoints

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.172494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.266560Z digest=sha256:cde222cfde6b23bda90e4c50873f734aba1242a61f483ea552abac569b13dd2f

Observation ff7ace8f-0a0d-4e7d-925c-e4841cde794e · outbound

This paper cites Raising the ceiling: Conflict- free local feature matching with dynamic view switching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Raising the ceiling: Conflict- free local feature matching with dynamic view switching

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.031882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.308887Z digest=sha256:3bbe1efb190609188fa0f60e5d911ba0d0c451e6ff5d7705a9907711f2240b6d

Observation 0a3cc0d3-66f8-4460-b93e-cca54871611e · outbound

This paper cites Diffusion Model for Dense Matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Diffusion Model for Dense Matching

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:27.364440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:27.364440Z digest=sha256:6113c5221108400330d837a744ec49081795439fccea947e348d1bda2a7f28a9

Observation 366924d7-07d4-4d99-b556-10c53e212838 · outbound

This paper cites Unsupervised deep image stitching: Reconstructing stitched features to images.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Unsupervised deep image stitching: Reconstructing stitched features to images

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:33.884826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.422565Z digest=sha256:96e022022c688e02a2f5b5d6e6c461d5aa8d6dd051f1ecb0dec15b9ad43c2839

Observation 6e3d36ea-47ca-45b8-ab3f-75e0e8c32329 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching DINOv2: Learning Robust Visual Features without Supervision

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:27.473301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:27.473301Z digest=sha256:0ee3160090edd471820cbf782b0919c1e3ffb32c924456db009c8f3faec1f30f

Observation fd2881af-efa5-4df9-9ad6-c19daaa7db92 · outbound

This paper cites Enhancing deformable local features by jointly learning to detect and describe keypoints.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Enhancing deformable local features by jointly learning to detect and describe keypoints

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:33.747990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.541307Z digest=sha256:cfb4681a0036bed7753cede7d95449e67d7c7b727815de0ad1cbb79115572489

Observation 5a3f4967-3b91-4f78-9039-865fce183440 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Learning transferable visual models from natural language supervi- sion

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:33.625802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.599471Z digest=sha256:c8293c78ebdfb7bbd3cd3cc2dc89e645f44cbbb0f98d8ed3aa65fbab86218619

Observation 70bc9d2e-73e1-4040-8a00-03d86913dc32 · outbound

This paper cites R2d2: Reliable and repeatable detec- tor and descriptor.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching R2d2: Reliable and repeatable detec- tor and descriptor

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:33.457903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.655373Z digest=sha256:14ca91eab02e41e45dfffadd45f9fdd78fd258772f93733b1d1b52305c941082

Observation b21463f9-122c-4b08-bc1a-cb102224dff0 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching High-resolution image synthesis with latent diffusion models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:33.226013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.718222Z digest=sha256:fe06db8542edbeabe7bba5fe5a9030f54b1e7891f5177f1bf45b2649c9575265

Observation 2348e819-b850-451f-aec9-b22fe4533446 · outbound

This paper cites Orb: An efficient alternative to sift or surf.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Orb: An efficient alternative to sift or surf

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:27.778044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:27.778044Z digest=sha256:4b4dc21816539fc4348a48bd8707507ae89d9618a65347baf1ad2a0962cc09a2

Observation 4a3840b2-ea6d-49a1-a160-26f54d845d5e · outbound

This paper cites Where’s waldo: Diffusion features for person- alized segmentation and retrieval.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Where’s waldo: Diffusion features for person- alized segmentation and retrieval

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.994498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.842497Z digest=sha256:00b380723ef9532a6a084f1e7db7009b165b84d11dacf796b4ef58c5793fec6b

Observation d522f11c-1717-4367-bed7-b0e5294addd8 · outbound

This paper cites From coarse to fine: Robust hierarchical localization at large scale.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching From coarse to fine: Robust hierarchical localization at large scale

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.663381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.888299Z digest=sha256:d196af5df7cc3ac87b7b5233a39395e697a8801f415c224b44ef7827bf1bbd17

Observation 00c9f4eb-a961-45e8-ae7f-d6025a30d812 · outbound

This paper cites Superglue: Learning feature matching with graph neural networks.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Superglue: Learning feature matching with graph neural networks

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.425716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:27.953760Z digest=sha256:a88bb21c5ad7cdce1fe7fa38ecf72cb81851f84377739d8323590bb69a27dfcb

Observation cba10385-a76f-4133-b112-10efd16b501d · outbound

This paper cites Back to the feature: Learning robust camera localization from pixels to pose.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Back to the feature: Learning robust camera localization from pixels to pose

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.291176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.004243Z digest=sha256:d1d0e235a0c8364e48794d399b3e57f2ee2b771358f50b594e7e36f8db555409

Observation 2d36e1a3-6cb1-4a23-96c6-3b81badb7119 · outbound

This paper cites Benchmarking 6dof outdoor visual localization in changing conditions.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Benchmarking 6dof outdoor visual localization in changing conditions

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.147368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.063872Z digest=sha256:b5f42f451b8309371a708e48a21850baf3057baa3f02d99cde9f88aafd6bbe04

Observation fdf1efb3-359f-48af-8a60-1ab49c542121 · outbound

This paper cites Structure- from-motion revisited.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Structure- from-motion revisited

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.002162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.163328Z digest=sha256:4cf392eed0108d5e2f5e469b71265dc797e860866dc045a984099f820bab59f1

Observation df2232f0-de18-4a57-9fc8-bda30024a113 · outbound

This paper cites Pixelwise view selection for unstructured multi-view stereo.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Pixelwise view selection for unstructured multi-view stereo

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:28.275287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:28.275287Z digest=sha256:cec6e544558c7066db2e1e5d52c714d9d4821a11a7c0e16f0fc6ca5b68a226fc

Observation 427d96aa-3386-4763-91d0-d57e5adc142b · outbound

This paper cites Laion-5b: An open large-scale dataset for training 10 next generation image-text models.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Laion-5b: An open large-scale dataset for training 10 next generation image-text models

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.869337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.359469Z digest=sha256:448aabaff47eefbb9d6684448d6fc0063c1b3130bb9eb57244ebe3601b8fe880

Observation 74b1138f-c65a-4a4d-bf43-563ca1732b1a · outbound

This paper cites Denoising Diffusion Implicit Models.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Denoising Diffusion Implicit Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:28.447431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:28.447431Z digest=sha256:658f99dc1286d6bfd974a779db810683be8c788383a05fc470eb8e370fd8affa

Observation 0767a589-9a9d-4861-8944-d19cbe56baec · outbound

This paper cites Loftr: Detector-free local feature matching with transformers.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Loftr: Detector-free local feature matching with transformers

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.711218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.520920Z digest=sha256:044e4e232bdde78088fe173c2c62abc6435c52bfee47d75de397983d9f61ac6b

Observation dc90f8ae-589d-4c75-8125-0eea407c6c5b · outbound

This paper cites Inloc: Indoor visual localization with dense matching and view synthesis.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Inloc: Indoor visual localization with dense matching and view synthesis

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.577976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.614403Z digest=sha256:80f7b6d68b2bf74bf74d8f2c86e2c8bf465a775523f14610f02da9d417589dac

Observation f85b2797-b92d-41d2-a5ad-ddb97e4c6582 · outbound

This paper cites Emergent correspondence from image diffusion.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Emergent correspondence from image diffusion

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.429567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.704467Z digest=sha256:281b949fb5f1dfd90496b7e041d9b688b2ea2cb6eb233b5f8f1f7ba724e4f967

Observation 0d1b7fd6-57a2-4cb9-b2ef-591173c50acf · outbound

This paper cites QuadTree Attention for Vision Transformers.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching QuadTree Attention for Vision Transformers

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:28.721662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:28.721662Z digest=sha256:a48449a08ff8b770979300f0dcf97d456abe4c4363194114ca9f832ddecc4b24

Observation 61b4e027-6585-49a9-9d8a-0b944f0982ea · outbound

This paper cites HomoMatcher: Dense Feature Matching Results with Semi-Dense Efficiency by Homography Estimation.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching HomoMatcher: Dense Feature Matching Results with Semi-Dense Efficiency by Homography Estimation

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:39:29.795682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.725362Z digest=sha256:dbb7d0818d16a18bf5ac40bf4b5b201d257ac869effb781bfe60b876a72540c4

Observation 462b7df8-e90a-4399-b2c3-2587db946e2c · outbound

This paper cites Efficient loftr: Semi-dense local feature matching with sparse-like speed.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Efficient loftr: Semi-dense local feature matching with sparse-like speed

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.294865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.728936Z digest=sha256:1fdfa62a221bcf4f9867dca3b297434c233f5f931a7a4548ec92081f404f86c3

Observation f616e622-d0f0-4e29-bd54-aa0f2a3de402 · outbound

This paper cites Croco: Self-supervised pre-training for 3d vision tasks by cross-view completion.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Croco: Self-supervised pre-training for 3d vision tasks by cross-view completion

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.149646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.829196Z digest=sha256:cd506237175a003302dcbc90c6cdaba0c94f809ed010b1b417ffed6b6631f7eb

Observation 09b711a5-6c11-4a40-80b8-44f20bb91b39 · outbound

This paper cites Croco v2: Improved cross-view completion pre- training for stereo matching and optical flow.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Croco v2: Improved cross-view completion pre- training for stereo matching and optical flow

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:28.904750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:28.904750Z digest=sha256:e43eb5c77190c54b4f878c7b707a8a27d4e0247f1dddeb8ac61d8859bdc1eb23

Observation bf421cb8-49cc-49ee-a1c1-76e69ab1ef8b · outbound

This paper cites Open-vocabulary panop- tic segmentation with text-to-image diffusion models.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Open-vocabulary panop- tic segmentation with text-to-image diffusion models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.981161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:28.981804Z digest=sha256:9b5d4a41c7442f5efa12be4f7f15765cc77a44f06f8739875ecf7adc5dc63b09

Observation 26a49359-f5fb-41d4-9b9e-09785c3cabae · outbound

This paper cites Telling left from right: Identifying geometry-aware semantic corre- spondence.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Telling left from right: Identifying geometry-aware semantic corre- spondence

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.843729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:29.111538Z digest=sha256:d0333d71dd6e8c219e2196793afbc66608721cc35ef78ba4dfed3b27097d4799

Observation 4e6a71ee-c64b-4a4b-b7ff-298e4763f738 · outbound

This paper cites A tale of two features: Stable diffusion complements dino for zero-shot semantic correspondence.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching A tale of two features: Stable diffusion complements dino for zero-shot semantic correspondence

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.731910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:29.252854Z digest=sha256:70b850dce6bbbcbc40e99fe2aa0a139fee92d1f95f61025c87f79373254111a5

Observation e38d7acc-1889-43f7-8ca8-e58f981098e6 · outbound

This paper cites Diffglue: Diffusion-aided image feature matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Diffglue: Diffusion-aided image feature matching

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.584616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:29.347751Z digest=sha256:f2abc245b16308d7df35f281eab1b12db4c90fa33252a2a3fbd69c4cf3e298ae

Observation 0105457e-345d-44ef-abc4-6841eb6c8711 · outbound

This paper cites Mesa: Matching everything by segmenting anything.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Mesa: Matching everything by segmenting anything

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.407579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:29.391843Z digest=sha256:f8cf406dea56fdc2b0ade93a46ba2d2d68beada206ed1c8b1d72211dc491736f

Observation 9e353c44-b44c-4873-b468-0e978bdc14db · outbound

This paper cites Unleashing text-to-image diffu- sion models for visual perception.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Unleashing text-to-image diffu- sion models for visual perception

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:29.514537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:29.514537Z digest=sha256:6d9da20d2d41926952d05c83a1b1243d725db262a10974fb883c4edfa84d8eeb

Observation 99ebaf2a-ed60-458b-8219-a47bfbe0fe54 · outbound

This paper cites Pmatch: Paired masked image modeling for dense geometric matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Pmatch: Paired masked image modeling for dense geometric matching

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.196322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:39:29.580035Z digest=sha256:597ac1229ef1cccd477b4987c5f87f8278af849d2e11f0034aa244effe3dab76

Pith citing papers

Observation 5470f779-d45d-474b-a5cd-7888123a85c0 · inbound

Probing and Leveraging Video Diffusion Transformer Features for Robust Point Tracking cites this paper.

Probing and Leveraging Video Diffusion Transformer Features for Robust Point Tracking Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T14:24:49.968980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:24:49.968980Z digest=sha256:3485cf3b16123c66099e3dfb59c529006ffb3d7afa8b25f0e00a8e328d144599