Pith. sign in

Paper Citation Record · LEDGER

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching

As of 7 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 1 inbound Pith citation observation for arXiv:2507.10318.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10318 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:39:29.580035Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T14:24:49.968980Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact2
  • verified fuzzy48
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4bc141d3-7675-4a91-93f2-d87a7f3465e5 · outbound

This paper cites Burst: A benchmark for unifying object recognition, segmentation and tracking in video.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Burst: A benchmark for unifying object recognition, segmentation and tracking in video

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:37.856565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.108553Z digest=sha256:c57f2b661fc3aa0d2450c1c3cc5ee24bc236e9a03518fc63dce4ef2a2c0bd725

Observation 756c8d63-ff4b-467d-9dd6-15e914fccb81 · outbound

This paper cites Hpatches: A benchmark and evaluation of handcrafted and learned local descriptors.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Hpatches: A benchmark and evaluation of handcrafted and learned local descriptors

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:37.846283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.201361Z digest=sha256:ea363729004e205fc59d0d55fc167776efcbfc9c621803bf6b67bd6a1797e8f2

Observation 24de8596-906f-4ce2-bd45-d38b533b1ad8 · outbound

This paper cites an unresolved cited work.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:39:37.778915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.298170Z digest=sha256:8e2f001d3adbd2e89719f93dd4563ea1b0df73718f26dff168491ecf47573271

Observation a2ca3393-b810-4c80-97cd-4e163f48e6bc · outbound

This paper cites Surf: Speeded up robust features.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Surf: Speeded up robust features

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:25.377933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:25.377933Z digest=sha256:35d312851113c0824f45622cc9eb297c86d760abba767419d31a75a3678f63bf

Observation 904d3fe5-893a-4e2d-bcfd-ce89a4d86e40 · outbound

This paper cites Speeded-up robust features (surf).

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Speeded-up robust features (surf)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:37.567856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.474100Z digest=sha256:e0baadc1a036653c481559ed8e56b1f5d7962420f2fadb5f0f771a68571ef7b2

Observation 84983732-8d76-4fd7-a9e9-a9d50bf1c739 · outbound

This paper cites VOLoc: Visual Place Recognition by Querying Compressed Lidar Map.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching VOLoc: Visual Place Recognition by Querying Compressed Lidar Map

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:39:29.984495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.536456Z digest=sha256:5d9be2be0f2b8ad43046c8ed29b21f8c731196631b61be4e6a0bb199dfcd1c80

Observation dc0dafdf-0a80-4042-9b63-3a67dac07e34 · outbound

This paper cites Prism: Pro- gressive dependency maximization for scale-invariant image matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Prism: Pro- gressive dependency maximization for scale-invariant image matching

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:37.392675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.625802Z digest=sha256:154efe603449592ade301475cf7c9433ef97dfd183c0fe04db4c5153a7c964ee

Observation d33e7e9d-8a53-48c3-ab78-6d6a09fea765 · outbound

This paper cites Improving transformer-based image matching by cascaded capturing spatially informative keypoints.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Improving transformer-based image matching by cascaded capturing spatially informative keypoints

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:37.154073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.728919Z digest=sha256:87946a3938c1347bce438207b10123926397489a4fbfee67434c241a27169082

Observation 35df9448-77ba-495d-8ec8-2a9c4003d9dc · outbound

This paper cites Aspanformer: Detector-free image matching with adaptive span transformer.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Aspanformer: Detector-free image matching with adaptive span transformer

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.959512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.775270Z digest=sha256:ba16f5622fe03bd2d6fc3079490b002de3920b2855dce96757c82b11d95001c4

Observation e43532a9-5957-4d6b-b6bf-ca6b5f0704d0 · outbound

This paper cites Ecomatcher: Efficient clustering oriented matcher for detector-free image matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Ecomatcher: Efficient clustering oriented matcher for detector-free image matching

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.754168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.817823Z digest=sha256:e0522bda4e5d9c268394ffe3c9b4f6623a9a9dfc15e066f997581d9aa0e8a5fc

Observation a9318b76-92f6-4235-9996-27774f6251cc · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.560632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.855956Z digest=sha256:b27bd629f32d6dd4538e772f628e7cb2b73d0822d8b5d20841220a4f21e2e20a

Observation 070a27cd-51d9-4b48-b33c-3588af7b8a5c · outbound

This paper cites Superpoint: Self-supervised interest point detection and description.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Superpoint: Self-supervised interest point detection and description

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.378666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.919632Z digest=sha256:855e5b88b543b8e5e6d16adaaaf9e447664480c92594dfd774671d57612be6a0

Observation 1c8f08ac-794d-40a3-8521-6addefadc609 · outbound

This paper cites Diffusion models beat gans on image synthesis.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Diffusion models beat gans on image synthesis

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.200418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:25.983654Z digest=sha256:9f8de635cee2796a4931cd04c05326cea70115b2bd710937a1d214f9925a787e

Observation da261533-5483-4852-8b58-cb367799727c · outbound

This paper cites Dkm: Dense kernelized feature matching for geometry estimation.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Dkm: Dense kernelized feature matching for geometry estimation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:36.103809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:26.075536Z digest=sha256:3c0ee88059e5801ad90bfc9ab3d9a90e1e685a29e6b3a634b6287ca3fc803185

Observation a12f004c-20c9-4518-b327-c2b995336d2b · outbound

This paper cites Roma: Robust dense fea- ture matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Roma: Robust dense fea- ture matching

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.972971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:26.145757Z digest=sha256:f6920e224b1ecc4ccbe39fab6ec2c6d912cae68b99082c2049368e44359b237a

Observation bbe42dae-55cc-4b36-9f7c-861494824c09 · outbound

This paper cites Top- icfm: Robust and interpretable topic-assisted feature match- ing.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Top- icfm: Robust and interpretable topic-assisted feature match- ing

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.824253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:26.220843Z digest=sha256:0cb23d5b9749ec7a0c09f82dc633f3dd13de62f5dde5f40a99d516678fd12918

Observation 4548f61f-6638-4d1e-b0b5-b6cc1fe6d4f7 · outbound

This paper cites Deep residual learning for image recognition.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Deep residual learning for image recognition

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:26.290930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:26.290930Z digest=sha256:4a29f22fc0c58b7307f284a088c7a2de667356a047ec7a523238839156d4d11e

Observation d3577e46-a4a3-4ed2-9e62-011b3aa12c3c · outbound

This paper cites Masked autoencoders are scalable vision learners.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Masked autoencoders are scalable vision learners

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.675513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:26.392578Z digest=sha256:5c3b300faf06e43c8f3166714183b675c35f7e9792c8351f23f8c280a224e2d8

Observation e8c594b2-c443-471c-a17e-5569b80f6f26 · outbound

This paper cites Denoising dif- fusion probabilistic models.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Denoising dif- fusion probabilistic models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.517902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:26.469629Z digest=sha256:6830063bdf5ac363af61ed91f2687195efd914ca24bc1669c2576e6a003aaf6d

Observation 4c876dab-15d7-424d-a0be-6d4397234437 · outbound

This paper cites Dynamicid: Zero-shot multi-id image personalization with flexible facial editabil- ity.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Dynamicid: Zero-shot multi-id image personalization with flexible facial editabil- ity

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.300228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:26.535857Z digest=sha256:2a02c96df53811b4eb205dac931fb07813edc1e58cd072a3bc929c089a9d6ad4

Observation c7ec0d92-1951-4c36-a0cf-b39b8651e222 · outbound

This paper cites Omniglue: Generalizable feature match- ing with foundation model guidance.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Omniglue: Generalizable feature match- ing with foundation model guidance

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.162338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:26.616999Z digest=sha256:dd4380cc64dcef9f7d38c3cad6eafcaaf8f1408df52a6d731e2a4291555bb4f1

Observation 55c6ccaf-9893-4b15-a619-1b6ce19ca8e3 · outbound

This paper cites Segment any- thing.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Segment any- thing

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:26.711760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:26.711760Z digest=sha256:7c96bd76e7b183087ecebbddb53e78fe0b8120e3a44a8bfb3f25431a28b1bdb8

Observation a95c3a97-891d-4892-8be2-85a0a65f7c1a · outbound

This paper cites Sd4match: Learning to prompt stable diffu- sion model for semantic matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Sd4match: Learning to prompt stable diffu- sion model for semantic matching

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:35.017486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:26.781562Z digest=sha256:1b9fd7611ee59e8cf92de0d22428053855830437b9ca0fd79b92e71aacd6864b

Observation 03a73a54-fa32-40d8-9682-852623082ad7 · outbound

This paper cites Megadepth: Learning single- view depth prediction from internet photos.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Megadepth: Learning single- view depth prediction from internet photos

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.858580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:26.854437Z digest=sha256:2353b4dbb8eac51db2bd409a0b85a406683bdcec24331cb533575183e77f80cf

Observation b140ce15-c14e-40e1-8239-8b8992a360dc · outbound

This paper cites Recrecnet: Rectangling rectified wide-angle images by thin-plate spline model and dof-based curriculum learn- ing.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Recrecnet: Rectangling rectified wide-angle images by thin-plate spline model and dof-based curriculum learn- ing

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.697273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:26.915593Z digest=sha256:ef665a8e920257e6402e51e111aad2e38947fdba2b988120f21a1c30addc1a7f

Observation cbb360a1-5166-40d3-a3a8-e95e79844a2e · outbound

This paper cites Feature pyra- mid networks for object detection.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Feature pyra- mid networks for object detection

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.562423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.008107Z digest=sha256:d320f447fee37375aafcae9194e1d9fa10a1cf9efd28e35e1456605f6fbe4e11

Observation 0e3b00dc-2f28-4bae-a4df-b1f79a9861ef · outbound

This paper cites Lightglue: Local feature matching at light speed.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Lightglue: Local feature matching at light speed

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.421039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.078922Z digest=sha256:2484575cd7504692aed1e7fa6efab69987f9a4f1d9a7cd68726934180bc9b08f

Observation a27d2dc7-f8b6-49cf-8f23-14d8b8dfff31 · outbound

This paper cites Semantic- aware representation learning for homography estimation.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Semantic- aware representation learning for homography estimation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.293249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.136387Z digest=sha256:2e00240ae4a471cc638b5a660f8799f7c98f96eaad4b7dd0fb7d1862a8cd8e8b

Observation 67a59f89-9e45-488f-9897-935d63b327c4 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Swin transformer: Hierarchical vision transformer using shifted windows

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:27.196330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:27.196330Z digest=sha256:9169ea894bc7e76683785c45a778deee231947c06163d9df2129d93e5d4e9e87

Observation 94568ca6-1309-498e-97d4-f7b6be224f49 · outbound

This paper cites Distinctive image features from scale- invariant keypoints.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Distinctive image features from scale- invariant keypoints

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.172494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.266560Z digest=sha256:10874a59e8bb39d758fd56aedd6f9982ce24a908339a24c384c7172753129550

Observation ff7ace8f-0a0d-4e7d-925c-e4841cde794e · outbound

This paper cites Raising the ceiling: Conflict- free local feature matching with dynamic view switching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Raising the ceiling: Conflict- free local feature matching with dynamic view switching

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:34.031882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.308887Z digest=sha256:eb9c7f2128197e92cfc8dabb4da7f61defe57fc75a7307151f08c9c1f5183a4b

Observation 0a3cc0d3-66f8-4460-b93e-cca54871611e · outbound

This paper cites Diffusion Model for Dense Matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Diffusion Model for Dense Matching

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:27.364440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:27.364440Z digest=sha256:57475b329cbd51600224262bea2ab281acfe95bce3e7a3f1d0e954dca5ae9b08

Observation 366924d7-07d4-4d99-b556-10c53e212838 · outbound

This paper cites Unsupervised deep image stitching: Reconstructing stitched features to images.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Unsupervised deep image stitching: Reconstructing stitched features to images

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:33.884826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.422565Z digest=sha256:23ca7e7ec7821121db8d98cf927ea04a93ec7218aab586147b42049e4ce6c6fc

Observation 6e3d36ea-47ca-45b8-ab3f-75e0e8c32329 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching DINOv2: Learning Robust Visual Features without Supervision

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:27.473301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:27.473301Z digest=sha256:b383889f302066eb94b7c6af8437130274942a201a7a3007af2b345f8c4aeade

Observation fd2881af-efa5-4df9-9ad6-c19daaa7db92 · outbound

This paper cites Enhancing deformable local features by jointly learning to detect and describe keypoints.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Enhancing deformable local features by jointly learning to detect and describe keypoints

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:33.747990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.541307Z digest=sha256:efcbb47ad94fcfd65a403d40e146f9809b5b9737015d78cc7b096f754cecf8d9

Observation 5a3f4967-3b91-4f78-9039-865fce183440 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Learning transferable visual models from natural language supervi- sion

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:33.625802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.599471Z digest=sha256:796d7db23cbb6d8bbc1e94caaa8cf5034aba9c6c23abbb25cfd0280c3ca6529d

Observation 70bc9d2e-73e1-4040-8a00-03d86913dc32 · outbound

This paper cites R2d2: Reliable and repeatable detec- tor and descriptor.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching R2d2: Reliable and repeatable detec- tor and descriptor

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:33.457903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.655373Z digest=sha256:292eeea554439b4e21309a8c29e9b086cf8e8527f5eb6bf1a3db96d9c927b825

Observation b21463f9-122c-4b08-bc1a-cb102224dff0 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching High-resolution image synthesis with latent diffusion models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:33.226013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.718222Z digest=sha256:220cdb55bed59d0788b4f1d4d0894aaffb62a8968e7351d4b97bd4f554a75ea1

Observation 2348e819-b850-451f-aec9-b22fe4533446 · outbound

This paper cites Orb: An efficient alternative to sift or surf.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Orb: An efficient alternative to sift or surf

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:27.778044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:27.778044Z digest=sha256:36c85945c57de8ae9f7e1032213d9ccdc1fa2de05963a2387b988d164a990a5f

Observation 4a3840b2-ea6d-49a1-a160-26f54d845d5e · outbound

This paper cites Where’s waldo: Diffusion features for person- alized segmentation and retrieval.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Where’s waldo: Diffusion features for person- alized segmentation and retrieval

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.994498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.842497Z digest=sha256:6268ef02bd40bb716f7e6df2af7644f854a7888fe33b18de94db30b2b8fa6c5a

Observation d522f11c-1717-4367-bed7-b0e5294addd8 · outbound

This paper cites From coarse to fine: Robust hierarchical localization at large scale.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching From coarse to fine: Robust hierarchical localization at large scale

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.663381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.888299Z digest=sha256:72ea47c6d70257d116cbf3eb545fa97966252f6c7d34c146bf1d5ecab9c75504

Observation 00c9f4eb-a961-45e8-ae7f-d6025a30d812 · outbound

This paper cites Superglue: Learning feature matching with graph neural networks.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Superglue: Learning feature matching with graph neural networks

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.425716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:27.953760Z digest=sha256:50e970cd9a6351e2765a5826641c5633a5b9138e3158cfce093138ebbda851df

Observation cba10385-a76f-4133-b112-10efd16b501d · outbound

This paper cites Back to the feature: Learning robust camera localization from pixels to pose.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Back to the feature: Learning robust camera localization from pixels to pose

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.291176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.004243Z digest=sha256:27458a53ee98ed4a566d6eef7ba75afc3ddb76283352dc5faa14b61686ec0515

Observation 2d36e1a3-6cb1-4a23-96c6-3b81badb7119 · outbound

This paper cites Benchmarking 6dof outdoor visual localization in changing conditions.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Benchmarking 6dof outdoor visual localization in changing conditions

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.147368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.063872Z digest=sha256:ba40bedf3782031cc406e2e73243a4c7c83731fe3cfaedd06411a8149d69ddfb

Observation fdf1efb3-359f-48af-8a60-1ab49c542121 · outbound

This paper cites Structure- from-motion revisited.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Structure- from-motion revisited

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:32.002162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.163328Z digest=sha256:617b6189a12efef978d533b7ba92b10a230a3de0c79cd3d2ccc6b300014204ec

Observation df2232f0-de18-4a57-9fc8-bda30024a113 · outbound

This paper cites Pixelwise view selection for unstructured multi-view stereo.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Pixelwise view selection for unstructured multi-view stereo

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:28.275287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:28.275287Z digest=sha256:152b20d1ff6aceb18e55c1d7f33612080b62d0d95592fc8441e773fab789dc82

Observation 427d96aa-3386-4763-91d0-d57e5adc142b · outbound

This paper cites Laion-5b: An open large-scale dataset for training 10 next generation image-text models.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Laion-5b: An open large-scale dataset for training 10 next generation image-text models

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.869337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.359469Z digest=sha256:23cb4141d1a0016cf7a5158c7762f4cc8e511db3282e59b37de78f214975be0b

Observation 74b1138f-c65a-4a4d-bf43-563ca1732b1a · outbound

This paper cites Denoising Diffusion Implicit Models.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Denoising Diffusion Implicit Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:28.447431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:28.447431Z digest=sha256:d30f8119a5831f749a83484502e4c3ab91afbc656ba31d38625db14b55fdc263

Observation 0767a589-9a9d-4861-8944-d19cbe56baec · outbound

This paper cites Loftr: Detector-free local feature matching with transformers.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Loftr: Detector-free local feature matching with transformers

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.711218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.520920Z digest=sha256:59a794fb0ebe3bf2112f6c63ebc030b0875a769631d84c41cd37f9b98b9e2886

Observation dc90f8ae-589d-4c75-8125-0eea407c6c5b · outbound

This paper cites Inloc: Indoor visual localization with dense matching and view synthesis.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Inloc: Indoor visual localization with dense matching and view synthesis

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.577976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.614403Z digest=sha256:0e5f91463f8501d9fc60fb2c85b6d9300a696203c276b0242404e02134e008c0

Observation f85b2797-b92d-41d2-a5ad-ddb97e4c6582 · outbound

This paper cites Emergent correspondence from image diffusion.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Emergent correspondence from image diffusion

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.429567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.704467Z digest=sha256:a3293c5868388a2a570e1a40f6295817636b66849721b4b98d1c596af003e95a

Observation 0d1b7fd6-57a2-4cb9-b2ef-591173c50acf · outbound

This paper cites QuadTree Attention for Vision Transformers.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching QuadTree Attention for Vision Transformers

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:28.721662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:28.721662Z digest=sha256:c80b4974d27cec96e13b68014c5dbefdfb3d6d848fdf1587633accf18fe42b1e

Observation 61b4e027-6585-49a9-9d8a-0b944f0982ea · outbound

This paper cites HomoMatcher: Dense Feature Matching Results with Semi-Dense Efficiency by Homography Estimation.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching HomoMatcher: Dense Feature Matching Results with Semi-Dense Efficiency by Homography Estimation

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:39:29.795682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.725362Z digest=sha256:3cf7a408b1763cf1047a4daebf4e5c8c80d785f67ecab826fe68d572ce79e7e3

Observation 462b7df8-e90a-4399-b2c3-2587db946e2c · outbound

This paper cites Efficient loftr: Semi-dense local feature matching with sparse-like speed.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Efficient loftr: Semi-dense local feature matching with sparse-like speed

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.294865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.728936Z digest=sha256:b5ef73e8abebe331cd67fe6f59aa918814efc8b9a35adc2f7318bd9b5c282cb8

Observation f616e622-d0f0-4e29-bd54-aa0f2a3de402 · outbound

This paper cites Croco: Self-supervised pre-training for 3d vision tasks by cross-view completion.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Croco: Self-supervised pre-training for 3d vision tasks by cross-view completion

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:31.149646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.829196Z digest=sha256:df3f35ad223289d0811e2b2809e5417cf7724473bb3a04366de1af2e811d27dc

Observation 09b711a5-6c11-4a40-80b8-44f20bb91b39 · outbound

This paper cites Croco v2: Improved cross-view completion pre- training for stereo matching and optical flow.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Croco v2: Improved cross-view completion pre- training for stereo matching and optical flow

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:28.904750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:28.904750Z digest=sha256:db819fce47ab089c49d568fe33c1173785bde4b31aec5b5cc4d118b00bc419a3

Observation bf421cb8-49cc-49ee-a1c1-76e69ab1ef8b · outbound

This paper cites Open-vocabulary panop- tic segmentation with text-to-image diffusion models.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Open-vocabulary panop- tic segmentation with text-to-image diffusion models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.981161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:28.981804Z digest=sha256:5de31f2e39a896f223967eca3263cd22dddde3e7f653f9918f9ee9e9f73f433b

Observation 26a49359-f5fb-41d4-9b9e-09785c3cabae · outbound

This paper cites Telling left from right: Identifying geometry-aware semantic corre- spondence.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Telling left from right: Identifying geometry-aware semantic corre- spondence

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.843729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:29.111538Z digest=sha256:313b2615d89fdfe70da1ebed824dd6245156df7793187fdfb5819afb7333fe5e

Observation 4e6a71ee-c64b-4a4b-b7ff-298e4763f738 · outbound

This paper cites A tale of two features: Stable diffusion complements dino for zero-shot semantic correspondence.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching A tale of two features: Stable diffusion complements dino for zero-shot semantic correspondence

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.731910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:29.252854Z digest=sha256:8eb64578829875803dbfd743ec53fe658e9e1bb5f6edfbfc7ac491078693f2fd

Observation e38d7acc-1889-43f7-8ca8-e58f981098e6 · outbound

This paper cites Diffglue: Diffusion-aided image feature matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Diffglue: Diffusion-aided image feature matching

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.584616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:29.347751Z digest=sha256:8080acaa99b1986fbdedd9711235e118a84fe9339aa6e3aed3a905ee0a1b3a26

Observation 0105457e-345d-44ef-abc4-6841eb6c8711 · outbound

This paper cites Mesa: Matching everything by segmenting anything.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Mesa: Matching everything by segmenting anything

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.407579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:29.391843Z digest=sha256:7dd6582b5aa581016a53f2616a51ec6ea0197335af40829a1fc8e71300bbed50

Observation 9e353c44-b44c-4873-b468-0e978bdc14db · outbound

This paper cites Unleashing text-to-image diffu- sion models for visual perception.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Unleashing text-to-image diffu- sion models for visual perception

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:29.514537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:29.514537Z digest=sha256:2c8f8c2e4e6c157fd0d02deb9bb0bff52d4155ae3321666d764005f1c7433ab0

Observation 99ebaf2a-ed60-458b-8219-a47bfbe0fe54 · outbound

This paper cites Pmatch: Paired masked image modeling for dense geometric matching.

Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching Pmatch: Paired masked image modeling for dense geometric matching

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:39:30.196322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:39:29.580035Z digest=sha256:06bfd309b92a3b3306543c47b552d8653a5fcc10e966c5058c20706d2b72df37

Pith citing papers

Observation 5470f779-d45d-474b-a5cd-7888123a85c0 · inbound

Probing and Leveraging Video Diffusion Transformer Features for Robust Point Tracking cites this paper.

Probing and Leveraging Video Diffusion Transformer Features for Robust Point Tracking Mind the Gap: Aligning Vision Foundation Models to Image Feature Matching

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T14:24:49.968980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:24:49.968980Z digest=sha256:fb02e5c779b13a38e0a75a6599584228fd95939ed8120d484b091294b69f46fb