Pith. sign in

Paper Citation Record · LEDGER

DINOv3

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 100 inbound Pith citation observations for arXiv:2508.10104.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10104 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 100 of 909 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:10:05.205162Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

4
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8fa472b4-1096-44f6-aebc-468bf59c8337 · inbound

Seeing SDG 6 from space: local-scale monitoring of piped water and sewage system access across Africa using satellite imagery and self-supervised learning cites this paper.

Seeing SDG 6 from space: local-scale monitoring of piped water and sewage system access across Africa using satellite imagery and self-supervised learning DINOv3

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T17:18:14.596527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-23T17:16:11.599623Z digest=sha256:64881a4328f1281a1e9b786c77a2623d286522eff0abb0980092d29e0aa7476f

Observation 55877c7a-c5d2-4b5e-8450-75b782632df0 · inbound

3D Foundation Model for Generalizable Disease Detection in Head Computed Tomography cites this paper.

3D Foundation Model for Generalizable Disease Detection in Head Computed Tomography DINOv3

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-23T03:27:27.046641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-23T03:26:46.665351Z digest=sha256:dbf8d6a7eb285897dfd169d6ddd4185e55d918ccd9285a1c7c3ebb70cc865b95

Observation b7ff6680-af79-413a-8fff-13fa87622959 · inbound

Deepfake Detection that Generalizes Across Benchmarks cites this paper.

Deepfake Detection that Generalizes Across Benchmarks DINOv3

Reference 46

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T23:51:55.062617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T23:47:27.729587Z digest=sha256:292316e6182c745a235d2294e5c076a6edd74b40e1921c51d56fbfe03053713a

Observation 0b8aa0d4-6938-44dc-a674-241ebe62c9cd · inbound

MobQA: A Benchmark Dataset for Semantic Understanding of Human Mobility Data through Question Answering cites this paper.

MobQA: A Benchmark Dataset for Semantic Understanding of Human Mobility Data through Question Answering DINOv3

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-05T20:08:56.016306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:08:56.016306Z digest=sha256:562a260d8a7602e644b2155934772080f7072d98fd949a8b6b00c0af6ee9efbd

Observation 09b92571-f072-4d47-828c-9f186b734b1f · inbound

VFM-Guided Semi-Supervised Detection Transformer under Source-Free Constraints for Remote Sensing Object Detection cites this paper.

VFM-Guided Semi-Supervised Detection Transformer under Source-Free Constraints for Remote Sensing Object Detection DINOv3

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-05T20:10:05.205162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:10:05.205162Z digest=sha256:29f766adc80b4710490d6217d2ff4cb8aa4a70d66e72c927009702b149499815

Observation cc534ef3-b65a-4e55-81d3-93c0862876d2 · inbound

Annotation-Free Open-Vocabulary Segmentation for Remote-Sensing Images cites this paper.

Annotation-Free Open-Vocabulary Segmentation for Remote-Sensing Images DINOv3

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T16:41:41.779734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:41:41.779734Z digest=sha256:28451f42664069cd5a24ee836a3a386d324d3a6fa472edc868623db8746ceaf3

Observation 18397311-5588-4a30-bede-3eb00d3eb645 · inbound

Dino U-Net: Exploiting High-Fidelity Dense Features from Foundation Models for Medical Image Segmentation cites this paper.

Dino U-Net: Exploiting High-Fidelity Dense Features from Foundation Models for Medical Image Segmentation DINOv3

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-18T20:22:50.696311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T20:22:41.555806Z digest=sha256:e9834d3bfd0a9239a202f661c092e1a948e8e9409295e2f111c1959838162cdb

Observation fbd06ad7-5de4-40a3-960a-1aad392da16e · inbound

SegDINO: An Efficient Design for Medical and Natural Image Segmentation with DINO-V3 cites this paper.

SegDINO: An Efficient Design for Medical and Natural Image Segmentation with DINO-V3 DINOv3

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T13:14:15.332805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:14:15.332805Z digest=sha256:208a6809a0ee517d481cf5b46ce91dbb98bfc598bc16c8ebcedf04b562da5aba

Observation f55aa540-f0de-4be5-ba1e-6c35a8f6240b · inbound

M3Ret: Unleashing Zero-shot Multimodal Medical Image Retrieval via Self-Supervision cites this paper.

M3Ret: Unleashing Zero-shot Multimodal Medical Image Retrieval via Self-Supervision DINOv3

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T12:45:29.498792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:45:29.498792Z digest=sha256:1cd74e50c5cff5abf719399d7e49ffadcdbaaa8621e0dd75b454159016c94132

Observation db84ad64-d230-43e9-a50f-07d7f5184446 · inbound

AI-Driven Marine Robotics: Emerging Trends in Underwater Perception and Ecosystem Monitoring cites this paper.

AI-Driven Marine Robotics: Emerging Trends in Underwater Perception and Ecosystem Monitoring DINOv3

Reference 57

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T20:31:50.462746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T20:30:33.239934Z digest=sha256:e915b6d967a63a6fa32f64cdbf8a827afa3fb7a7516d6a2f40bcf96505de1056

Observation ed74ebb5-da7a-4143-8ffc-b80935ddd53b · inbound

FastVGGT: Training-Free Acceleration of Visual Geometry Transformer cites this paper.

FastVGGT: Training-Free Acceleration of Visual Geometry Transformer DINOv3

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T23:36:06.987511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T23:36:06.870759Z digest=sha256:3af3509c5819ba878ee0a63cfe869aa686d5163a21bfcd9b6ac3a045ba9375ef

Observation 1ecad42c-2b37-4bec-b2f3-0cae7d3c80f5 · inbound

Generalist versus Specialist Vision Foundation Models for Ocular Disease and Oculomics cites this paper.

Generalist versus Specialist Vision Foundation Models for Ocular Disease and Oculomics DINOv3

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:57:48.267016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:57:48.267016Z digest=sha256:988eff5e57791bb1939d428597a14c32db6c2e7522d8c7b102cc7933f20f24a6

Observation 8112021b-dd7f-4a21-901a-fd2e6726d6e9 · inbound

UniVerse-1: Unified Audio-Video Generation via Stitching of Experts cites this paper.

UniVerse-1: Unified Audio-Video Generation via Stitching of Experts DINOv3

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T00:05:47.846026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:05:47.846026Z digest=sha256:108a65485389806d19059f2a355eb3ceffb115045b495de770c8afb181a977d1

Observation 086f93d2-2f9b-429d-8f4c-6b0307f3dac7 · inbound

DIET-CP: Lightweight and Data Efficient Self Supervised Continued Pretraining cites this paper.

DIET-CP: Lightweight and Data Efficient Self Supervised Continued Pretraining DINOv3

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T11:31:41.917100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:31:41.917100Z digest=sha256:5df9b08a5d1cd56924fc85fba6977886652b830929255e03e9b3024b5eef99fb

Observation 0dee32a7-b8c2-4df6-95e2-86fe16e57298 · inbound

PeftCD: Leveraging Vision Foundation Models with Parameter-Efficient Fine-Tuning for Remote Sensing Change Detection cites this paper.

PeftCD: Leveraging Vision Foundation Models with Parameter-Efficient Fine-Tuning for Remote Sensing Change Detection DINOv3

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T18:55:35.641588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:55:35.641588Z digest=sha256:5fc7272a0ae4f18a464ff7d4d0a7cec3a385ece24c66a77d62f6897f8f3b9b1b

Observation 3055d09b-c43c-4f66-9681-9b489e28bd83 · inbound

Mars Traversability Prediction: A Multi-modal Self-supervised Approach for Costmap Generation cites this paper.

Mars Traversability Prediction: A Multi-modal Self-supervised Approach for Costmap Generation DINOv3

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T17:10:04.387734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:10:04.387734Z digest=sha256:024d5d2d7815643061e21ef7db401e645e91a696d8eeea4dd4f06a8400464d3e

Observation 2569de7a-fb28-405e-b8b1-964876bc1b24 · inbound

Advancing Metallic Surface Defect Detection via Anomaly-Guided Pretraining on a Large Industrial Dataset cites this paper.

Advancing Metallic Surface Defect Detection via Anomaly-Guided Pretraining on a Large Industrial Dataset DINOv3

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T15:43:09.085047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:43:09.085047Z digest=sha256:64a346bc71c36292bc54b29a769991a725e06fba57e322627321a147322b9365

Observation 1231fcf8-febf-4fd7-9b86-f3e0b4dc96c0 · inbound

PartCo: Part-Level Correspondence Priors Enhance Category Discovery cites this paper.

PartCo: Part-Level Correspondence Priors Enhance Category Discovery DINOv3

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-05-22T13:31:36.026328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-22T13:31:12.201742Z digest=sha256:3da65047b463e0c916436651f880e664e0973f946a855fcd3f4b83ca0d3f4b67

Observation fb226236-97c9-4ff6-9a21-10254b929a27 · inbound

CardioBench: Do Echocardiography Foundation Models Generalize Beyond the Lab? cites this paper.

CardioBench: Do Echocardiography Foundation Models Generalize Beyond the Lab? DINOv3

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T21:44:22.749484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T21:41:55.451141Z digest=sha256:04f62262091cc56d71031f4512838d02d664ba5ca5558bd7a20ad9e8440ec5a2

Observation 6604e520-b58a-4334-8838-4dcff8208d4f · inbound

Activation Quantization of Vision Encoders Needs Prefixing Registers cites this paper.

Activation Quantization of Vision Encoders Needs Prefixing Registers DINOv3

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T11:31:29.827005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:31:29.827005Z digest=sha256:47e59988b132d39b2d03119617bd184534a5cd7fb45c5a07dcf95c1c6a6aa150

Observation 8b5ea262-818a-40cd-bc93-250db9f6ea2d · inbound

Resolution scaling governs DINOv3 transfer performance in chest radiograph classification cites this paper.

Resolution scaling governs DINOv3 transfer performance in chest radiograph classification DINOv3

Reference 23

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T09:06:09.131842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T09:05:29.448447Z digest=sha256:d4172c77e6917473a27ee42fbfc62549fee92a0ba93f0440ba5956f2ce98b26a

Observation 693385d1-de7c-4db3-b244-c16a6e954a9f · inbound

SpaceVista: All-Scale Visual Spatial Reasoning from mm to km cites this paper.

SpaceVista: All-Scale Visual Spatial Reasoning from mm to km DINOv3

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-04T10:37:39.470308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:37:39.470308Z digest=sha256:d4781c4a7592febbf0ad3302c496ef1f6e5dd8dab14bb1e08574fdbf62ab4dc7

Observation 6048d556-e9be-4fe7-9a63-09ce43de4dec · inbound

DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving cites this paper.

DriveVLA-W0: World Models Amplify Data Scaling Law in Autonomous Driving DINOv3

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T06:48:01.091554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T06:48:00.943591Z digest=sha256:95344a7a333c4f6ddf07de06305614f526ad3a81d573ac4a63a575639a6e7c57

Observation b2138c60-baf3-4c57-ad32-e78f4d503726 · inbound

Memory-SAM: Human-Prompt-Free Tongue Segmentation via Retrieval-to-Prompt cites this paper.

Memory-SAM: Human-Prompt-Free Tongue Segmentation via Retrieval-to-Prompt DINOv3

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T06:05:57.473592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T06:03:45.684267Z digest=sha256:5d214290b244833de67b8f2ae1074e72de1e3000f93a0c23cd5493f64f02a9ef

Observation c9ef338c-3cf7-4b3e-bd4b-80cf25fa87eb · inbound

Memory-SAM: Human-Prompt-Free Tongue Segmentation via Retrieval-to-Prompt cites this paper.

Memory-SAM: Human-Prompt-Free Tongue Segmentation via Retrieval-to-Prompt DINOv3

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T09:21:21.558253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:21:21.558253Z digest=sha256:7bcc046b8b025d31cf2f91f850eb4519aff58ac481f7fa0836f052979b0a65df

Observation cfb4958b-b9fd-4ceb-b9ec-f7dbf8162639 · inbound

Elastic ViTs from Pretrained Models without Retraining cites this paper.

Elastic ViTs from Pretrained Models without Retraining DINOv3

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T09:03:09.336349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:03:09.336349Z digest=sha256:b84e0e6995b677727d431878d85b1b7c84cb01eaf43c2dbb17816270b5af90d7

Observation 030bf170-f131-478c-a790-ffc5894cb964 · inbound

Not Every Time and Frequency Need to Be Forgotten in Diffusion Unlearning cites this paper.

Not Every Time and Frequency Need to Be Forgotten in Diffusion Unlearning DINOv3

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:38.166432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:12:38.166432Z digest=sha256:72622653843815d85178c74897ee34707a43cc5681f93e83b2b9681d86ac6f67

Observation c184aa1d-fd04-474f-b800-6dfcd6d2aa0d · inbound

VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models cites this paper.

VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models DINOv3

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-18T05:22:23.995050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T05:22:05.125849Z digest=sha256:b3100273524112b3604f09a078a0db2d2f44e12e67ccc87264b50191e2136499

Observation 2165d00d-4f98-4d24-9e37-4aaf47718d28 · inbound

S3OD: Towards Generalizable Salient Object Detection with Synthetic Data cites this paper.

S3OD: Towards Generalizable Salient Object Detection with Synthetic Data DINOv3

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T08:20:01.265096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:20:01.265096Z digest=sha256:98b5001dbdb97239af69a2c3fee90740b0bbe4473170a6964fecfc8d1056b0be

Observation 56b3683f-8dcc-44a3-a193-e76ccbb82679 · inbound

A geometric and deep learning reproducible pipeline for monitoring floating anthropogenic debris in urban rivers using in situ cameras cites this paper.

A geometric and deep learning reproducible pipeline for monitoring floating anthropogenic debris in urban rivers using in situ cameras DINOv3

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-04T07:54:10.509113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:54:10.509113Z digest=sha256:fd91b7d542ffc626613d269d896d2e83e127f52efb3ec340cf8ad7fe4e2bcc37

Observation 8b71d498-3164-4406-b33d-d690ff747651 · inbound

Densemarks: Learning Canonical Embeddings for Human Heads Images via Point Tracks cites this paper.

Densemarks: Learning Canonical Embeddings for Human Heads Images via Point Tracks DINOv3

Reference 18

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T00:55:35.242807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T00:53:24.323059Z digest=sha256:27403ad036214c4e651b6e067f60524f6e8f09406641d0e7c6ea0c0a9be33b8d

Observation ee5451e6-3aa8-4ce2-a2c4-90d347b1c3b7 · inbound

Towards Cellular-Scale Interpretability in Pathology Foundation Models for Biomarker Assessment cites this paper.

Towards Cellular-Scale Interpretability in Pathology Foundation Models for Biomarker Assessment DINOv3

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T23:34:15.815600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:34:15.815600Z digest=sha256:513c8e9bc504c5c009bb7db03eef6e8c8a18bcd7f6ad35aaa27a6738d5d38202

Observation db4ce52f-d9db-4c44-9deb-47fc22d613b8 · inbound

UniADC: A Unified Framework for Anomaly Detection and Classification cites this paper.

UniADC: A Unified Framework for Anomaly Detection and Classification DINOv3

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T23:17:32.936917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:17:32.936917Z digest=sha256:9ae796e33a302262187beaa35c48c0286c2d018a4e097e32e3294fa70b8df0bb

Observation 5b27e588-645e-4af4-a313-ec1edec4893a · inbound

LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics cites this paper.

LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics DINOv3

Reference 39

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T07:23:00.016578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-16T07:22:59.854042Z digest=sha256:a8e5786b21e2a74e160851cd994fb7b89734d375f8ca02868150ea003bbff688

Observation 64af748a-f9e2-4b95-bfba-45e178038f68 · inbound

{\Phi}eat: Physically Grounded Material Feature Representation cites this paper.

{\Phi}eat: Physically Grounded Material Feature Representation DINOv3

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T22:16:16.783061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:16:16.783061Z digest=sha256:6c9f4e892e9cc61ead7d73c9122ed50c804b5460ebda83ac942f1fd241863689

Observation 6ab52e01-b898-4d54-a9fd-d1bc6072cc10 · inbound

The Persistence of Cultural Memory: Investigating Multimodal Iconicity in Diffusion Models cites this paper.

The Persistence of Cultural Memory: Investigating Multimodal Iconicity in Diffusion Models DINOv3

Reference 31

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T21:55:20.311565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T21:52:26.944212Z digest=sha256:1c8ec866465d76dd062ed1122905b79b6e4aceadda7ac453f118fc2dbbf19e2d

Observation e9be355a-964d-433a-ab75-2019a2ff43a8 · inbound

Sat2RealCity: Geometry-Aware and Appearance-Controllable 3D Urban Generation from Satellite Imagery cites this paper.

Sat2RealCity: Geometry-Aware and Appearance-Controllable 3D Urban Generation from Satellite Imagery DINOv3

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T22:13:46.241078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:13:46.241078Z digest=sha256:e642059a8b47ec21cdccd65d9c19f15cd0d8f6046440b9fd7f10f6bc804cd271

Observation 7b8f6809-21fa-41a1-bcb1-b3e48cf9e0a7 · inbound

Pixels or Positions? Benchmarking Modalities in Group Activity Recognition cites this paper.

Pixels or Positions? Benchmarking Modalities in Group Activity Recognition DINOv3

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T21:50:19.718324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T21:45:20.428547Z digest=sha256:d9c2d9708a0a62ac38d4250f28369d6bfba74e9f65e8b1e03c14e89ff80c5e0a

Observation d2e9cd46-da43-4cc1-adc9-5786cf7c55c3 · inbound

RoMa v2: Harder Better Faster Denser Feature Matching cites this paper.

RoMa v2: Harder Better Faster Denser Feature Matching DINOv3

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T21:24:49.186860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:24:49.186860Z digest=sha256:380255a1053ac8592055baae834656bcb29fdbfae578d641dda8c22905ba22db

Observation 403e5516-7a4b-44d8-97c0-d86d50677a92 · inbound

Align & Invert: Solving Inverse Problems with Diffusion and Flow-based Models via Representation Alignment cites this paper.

Align & Invert: Solving Inverse Problems with Diffusion and Flow-based Models via Representation Alignment DINOv3

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-03T21:07:14.817049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:07:14.817049Z digest=sha256:f22188bab238768bd62bcbf203aafad2afd2b13e949663efed63c293fee864db

Observation 237f50d5-c6a6-4ceb-a631-dfc883569ba5 · inbound

QueryOcc: Query-based Self-Supervision for 3D Semantic Occupancy cites this paper.

QueryOcc: Query-based Self-Supervision for 3D Semantic Occupancy DINOv3

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T21:03:43.448478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:03:43.448478Z digest=sha256:f40e3bee9a333a14150aa342cfef3cdeed8afb3571ee88721734ee62894d9629

Observation 1a2304ef-5e90-4ab0-86d5-1e419724144f · inbound

Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer cites this paper.

Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer DINOv3

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-03T20:30:35.520090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:30:35.520090Z digest=sha256:afc3cfaae08870bc90255cc0134291e7d84f7cc13fef6edd612475576aa002ab

Observation d2e00fd4-96a9-4e24-a333-5c2990bffc3c · inbound

SparseWorld-TC: Trajectory-Conditioned Sparse Occupancy World Model cites this paper.

SparseWorld-TC: Trajectory-Conditioned Sparse Occupancy World Model DINOv3

Reference 33

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T05:29:04.967759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T05:26:34.859975Z digest=sha256:e2dd82b08875bfe0799550c3c7859409e4edbfa76d2b3b1770fcc246f9da1d8e

Observation dfdd92c0-131f-498b-9d25-7aa40c51088a · inbound

Contrastive Heliophysical Image Pretraining for Solar Dynamics Observatory Records cites this paper.

Contrastive Heliophysical Image Pretraining for Solar Dynamics Observatory Records DINOv3

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-17T04:19:00.395837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T04:17:57.951707Z digest=sha256:4e3227cfde5ebe6edf00606fe1c57e39d8a2e9297bb55722d1bbdcc7508420b4

Observation 31621c73-e55f-4265-9a93-de2c6b4c2045 · inbound

Goal-Driven Reward by Video Diffusion Models for Reinforcement Learning cites this paper.

Goal-Driven Reward by Video Diffusion Models for Reinforcement Learning DINOv3

Reference 25

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T02:53:54.876698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T02:52:22.919967Z digest=sha256:128469656dd63f869544ee7098a0dddabf0814324b996db5df74e55160cd3410

Observation 02b1417e-253f-4910-87a0-a6d4ce7991e8 · inbound

ShelfGaussian: Shelf-Supervised Open-Vocabulary Gaussian-based 3D Scene Understanding cites this paper.

ShelfGaussian: Shelf-Supervised Open-Vocabulary Gaussian-based 3D Scene Understanding DINOv3

Reference 62

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T03:28:58.119925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T03:24:32.577420Z digest=sha256:fcf74dd5a93ef0f45da743d7604eb407eab161f2552ece2135c64b2240ecfc8b

Observation 0081e8c6-f88e-4cde-910e-c815499703f2 · inbound

C3G: Learning Compact 3D Representations with 2K Gaussians cites this paper.

C3G: Learning Compact 3D Representations with 2K Gaussians DINOv3

Reference 60

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T02:08:51.295886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T02:06:54.730272Z digest=sha256:cdb0dcc6d1c25b481320bbf25627f1ec7ec45a48b3efb3ec5c49d510fae88e99

Observation f43ce3e8-0197-42ce-bb20-98599db06f26 · inbound

From Orbit to Ground: Generative City Photogrammetry from Extreme Off-Nadir Satellite Images cites this paper.

From Orbit to Ground: Generative City Photogrammetry from Extreme Off-Nadir Satellite Images DINOv3

Reference 53

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T00:58:46.324965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T00:57:35.745943Z digest=sha256:a7871a0f3a8ff6cd184b885cee2d2b8617be25320167cb4fed3185481233d179

Observation 6b6c1676-a0be-4151-816d-08a85c14284c · inbound

Multi-view Pyramid Transformer: Look Coarser to See Broader cites this paper.

Multi-view Pyramid Transformer: Look Coarser to See Broader DINOv3

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T17:52:11.789138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:52:11.789138Z digest=sha256:c7b5a28a1faff91ac3a5c6f47f559d4a6c003fd9dd3e0a292245a4256f58bcf6

Observation b85b3468-59e9-442f-be55-da40c0a36691 · inbound

Is Generation Required for Data-Efficient Perception? cites this paper.

Is Generation Required for Data-Efficient Perception? DINOv3

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-03T17:39:46.607876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:39:46.607876Z digest=sha256:af2b6096958a24532de7f19930b83f00d37edd3ef015d87c5c8278101dae4e14

Observation 492f5b9c-66dc-44fa-b53d-5ff0feaa3cfb · inbound

ConceptPose: Training-Free Zero-Shot Object Pose Estimation using Concept Vectors cites this paper.

ConceptPose: Training-Free Zero-Shot Object Pose Estimation using Concept Vectors DINOv3

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-16T23:43:42.037967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T23:43:07.754702Z digest=sha256:33a5d8c1380ac7121ad3e43f318f0763744a148feb191705d2776d48a94bbc3d

Observation 7e05af4e-9ee9-40f3-9e1c-d74cd159069c · inbound

Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization cites this paper.

Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization DINOv3

Reference 59

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T22:53:38.244109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T22:52:41.592714Z digest=sha256:129159efa98ad5e161d47a763466a0c567d506e9de0955c003b44057811d5e0d

Observation 160684e0-1ec8-49b0-955b-bd703b05026a · inbound

SoccerMaster: A Vision Foundation Model for Soccer Understanding cites this paper.

SoccerMaster: A Vision Foundation Model for Soccer Understanding DINOv3

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T02:27:20.912025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T23:08:08.584111Z digest=sha256:32d22e37ccc5e11f3934a2abdb4d47c4b40a90f69923094c78f7666612e0c997

Observation a647ab63-dde3-4f63-8de1-691c7f324188 · inbound

Revisiting 2D Foundation Models for Scalable 3D Medical Image Classification cites this paper.

Revisiting 2D Foundation Models for Scalable 3D Medical Image Classification DINOv3

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T16:36:36.378043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:36:36.378043Z digest=sha256:558efe43a8794930792a91a2e684cc13a5f812dd1f88a34528df0ce7b792e5c5

Observation da686c8d-b660-4472-81d4-5c950e054d7b · inbound

JoVA: Unified Multimodal Learning for Joint Video-Audio Generation and Editing cites this paper.

JoVA: Unified Multimodal Learning for Joint Video-Audio Generation and Editing DINOv3

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T16:25:58.357377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:25:58.357377Z digest=sha256:0f464423271f5c7c05e553ff7ae4bc9562e291abfb686d5e6f4c9da7ae70a173

Observation 07180ba0-d5ae-4314-b973-d4ccf495a314 · inbound

Native and Compact Structured Latents for 3D Generation cites this paper.

Native and Compact Structured Latents for 3D Generation DINOv3

Reference 54

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T05:20:42.429791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T05:20:42.300247Z digest=sha256:f36e03b29645c93b74b8b837c01b5fa1056d1bd45c01cde163dc8724aeb95d93

Observation 3297648f-e5dd-450b-b824-745e08929433 · inbound

MoonSeg3R: Monocular Online Zero-Shot Segment Anything in 3D with Reconstructive Foundation Priors cites this paper.

MoonSeg3R: Monocular Online Zero-Shot Segment Anything in 3D with Reconstructive Foundation Priors DINOv3

Reference 40

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T21:38:34.028822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T21:36:51.329792Z digest=sha256:1de19874cbd26261d7ce283fc15b97b1a15de19b97b26bb21aae302e05c41270

Observation 9aa87c0c-c621-4ee5-8f36-1e3f2c300beb · inbound

DVGT: Driving Visual Geometry Transformer cites this paper.

DVGT: Driving Visual Geometry Transformer DINOv3

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T15:27:19.710548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:27:19.710548Z digest=sha256:ae97bf5ad974eaf2be1643aeae1f32ab5930a64feac2993e631ce8d3f942d5c3

Observation 7a516cf8-d4a3-4bf8-93d9-7264b21aa0a2 · inbound

MMLANDMARKS: a Cross-View Instance-Level Benchmark for Geo-Spatial Understanding cites this paper.

MMLANDMARKS: a Cross-View Instance-Level Benchmark for Geo-Spatial Understanding DINOv3

Reference 74

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T20:51:15.275262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T20:49:11.961079Z digest=sha256:d3738960016f4f6d719c528fe111fa5066177cadfac7be647f15f803e05d042d

Observation a8c005e7-5a87-47b4-808c-86c8990ac397 · inbound

Chorus: Multi-Teacher Pretraining for Holistic 3D Gaussian Scene Encoding cites this paper.

Chorus: Multi-Teacher Pretraining for Holistic 3D Gaussian Scene Encoding DINOv3

Reference 47

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T20:38:24.853538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T20:34:16.666956Z digest=sha256:205a7e222dcd1ad8651d152135760d96e1ad4f26c324b278b316ca04bdc4a985

Observation 84491694-4222-4b63-a73e-8e77fa4a39fc · inbound

SigLino: Efficient Multi-Teacher Distillation for Agglomerative Vision Foundation Models cites this paper.

SigLino: Efficient Multi-Teacher Distillation for Agglomerative Vision Foundation Models DINOv3

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T20:23:23.747485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T20:21:47.430865Z digest=sha256:3590af0c639575b1c8a3570a13e4886e38cd9a8ecb2c57662dd22978a4eae8d4

Observation 409e1d96-c49c-4808-b728-e9457e084840 · inbound

AlignPose: Generalizable 6D Pose Estimation via Multi-view Feature-metric Alignment cites this paper.

AlignPose: Generalizable 6D Pose Estimation via Multi-view Feature-metric Alignment DINOv3

Reference 47

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T12:34:52.503487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T12:33:14.120076Z digest=sha256:0c33e29f45eaeb9664fca74d2301ec7ca5e67345bad3070d9fa55054a9c4e77c

Observation 49177057-4adf-4cc4-ad48-879945916bd6 · inbound

Probing and Leveraging Video Diffusion Transformer Features for Robust Point Tracking cites this paper.

Probing and Leveraging Video Diffusion Transformer Features for Robust Point Tracking DINOv3

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T14:24:50.001185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:24:50.001185Z digest=sha256:73a4d819ec0e2dbbd6eaf9e92e9b59e0b397c75eed6f8bb1e32b93481b30d53d

Observation 1cd0e8c4-7e8c-4876-8e01-6ce1ccf72642 · inbound

EraseLoRA: MLLM-Driven Foreground Exclusion and Background Subtype Aggregation for Dataset-Free Object Removal cites this paper.

EraseLoRA: MLLM-Driven Foreground Exclusion and Background Subtype Aggregation for Dataset-Free Object Removal DINOv3

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T14:08:22.945466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T14:08:22.945466Z digest=sha256:f9f9a3b7e5395383ff906b9f1cb8749b1ed787fc8dfd8569a7bb2f473666296a

Observation c1a1dc19-950c-4e04-ac01-2589a76aac09 · inbound

FinPercep-RM: A Fine-grained Reward Model and Co-evolutionary Curriculum for RL-based Real-world Super-Resolution cites this paper.

FinPercep-RM: A Fine-grained Reward Model and Co-evolutionary Curriculum for RL-based Real-world Super-Resolution DINOv3

Reference 23

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T19:01:11.689919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T19:00:03.367921Z digest=sha256:af7ad8d476d5167925ffd6321350f537cb69a46dd4c9229995bfbeb8b8ac8f4a

Observation 4c0a3303-c3cc-41c3-8201-68831b42c261 · inbound

What Drives Success in Physical Planning with Joint-Embedding Predictive World Models? cites this paper.

What Drives Success in Physical Planning with Joint-Embedding Predictive World Models? DINOv3

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-05-21T15:34:15.010899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-21T15:33:24.616338Z digest=sha256:6fa3bee3752588ed0eea0af8d433c47774efc4a4f5b6bc0fcbbf7bcacd2d6d06

Observation bca09ef9-e511-4ebc-adb0-e311176a51a4 · inbound

Towards Any-Quality Image Segmentation via Generative and Adaptive Latent Space Enhancement cites this paper.

Towards Any-Quality Image Segmentation via Generative and Adaptive Latent Space Enhancement DINOv3

Reference 98

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T18:13:13.153258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T18:11:47.141366Z digest=sha256:9202422152efcf6de2d5904020e9f710786785f8f295962f029b23015e8de8f9

Observation bef07bcb-2def-414a-991c-018c6fa28a35 · inbound

PatchAlign3D: Local Feature Alignment for Dense 3D Shape Understanding cites this paper.

PatchAlign3D: Local Feature Alignment for Dense 3D Shape Understanding DINOv3

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T06:36:47.448156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:36:47.448156Z digest=sha256:60ca3d8d67a912459e57fe0a91d60406e673095f0f59df4a0caa9b7a212de7fa

Observation c29bb0d1-0435-47b3-a0bd-9a9c5a67934e · inbound

DiT-JSCC: Rethinking Deep JSCC with Diffusion Transformers and Semantic Representations cites this paper.

DiT-JSCC: Rethinking Deep JSCC with Diffusion Transformers and Semantic Representations DINOv3

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T12:26:40.101386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:26:40.101386Z digest=sha256:c1aa6e9c7dc703718ea5139cea4871d2f726ac2a589a7bea68320937daf00659

Observation d152726d-620b-4897-b735-fbf8247f8bc2 · inbound

Cross-Scale Pretraining: Enhancing Self-Supervised Learning for Low-Resolution Satellite Imagery for Semantic Segmentation cites this paper.

Cross-Scale Pretraining: Enhancing Self-Supervised Learning for Low-Resolution Satellite Imagery for Semantic Segmentation DINOv3

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-16T13:17:54.986097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T13:16:29.137187Z digest=sha256:3b786fee792f77b4426fe4960ad55bb28f3e7f57e4fec734df1d0afcbe580522

Observation 01cfa052-4f17-4c60-9e3f-9f72f4bf3d65 · inbound

Mirai: Autoregressive Visual Generation Needs Foresight cites this paper.

Mirai: Autoregressive Visual Generation Needs Foresight DINOv3

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:17:52.082557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T12:16:16.461488Z digest=sha256:da0de9384f939ecfdb407ccbe714553515a142d2854c097228e3f24eaad5e7b4

Observation d7febf5e-ce5a-46b8-bda4-57f503a47a35 · inbound

RayRoPE: Projective Ray Positional Encoding for Multi-view Attention cites this paper.

RayRoPE: Projective Ray Positional Encoding for Multi-view Attention DINOv3

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T09:02:49.202458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:02:49.202458Z digest=sha256:f1abd44320688d59e5e77f98e4bd74a878730cad0f2f5c4eabdc979d44a30964

Observation 826a8c7d-4e84-4393-8a75-a7a3228fc8b9 · inbound

UPLiFT: Efficient Pixel-Dense Feature Upsampling with Local Attenders cites this paper.

UPLiFT: Efficient Pixel-Dense Feature Upsampling with Local Attenders DINOv3

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T08:13:44.241700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:13:44.241700Z digest=sha256:dba1df8945743e2b65c49f5745bd86536d1354e1c531b4bda590c7197401ca45

Observation fe2d11e2-aa9d-4042-bf3a-2cc5aa9a0d6e · inbound

The pretraining domain outweighs the training objective in setting the privacy-utility trade-off of differentially private medical image analysis cites this paper.

The pretraining domain outweighs the training objective in setting the privacy-utility trade-off of differentially private medical image analysis DINOv3

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T07:37:46.236062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:37:46.236062Z digest=sha256:4133e7cf8e7534a1819d09f20dcc1d1ea69b6d662a64b13c243e5338f2e7c45b

Observation ea0ce8ba-2d8a-47f4-b0d1-9cbb093fca9f · inbound

AnomalyVFM -- Transforming Vision Foundation Models into Zero-Shot Anomaly Detectors cites this paper.

AnomalyVFM -- Transforming Vision Foundation Models into Zero-Shot Anomaly Detectors DINOv3

Reference 62

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T10:47:45.459816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T10:47:17.480722Z digest=sha256:acd4b4d96cc46d7cabc4482855f70e3bb54e63250f2c7107a015cb3ae3186f46

Observation 51e8d30c-4c30-4189-9c3d-0d338512ed80 · inbound

MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources cites this paper.

MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources DINOv3

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-03T06:50:24.779427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:50:24.779427Z digest=sha256:6ab8cf2d3d9dcc1ef15a604b425aeeb05820f28a048733fa2885600edbed7b3c

Observation 57cce296-036f-4d54-8312-31ae09027a35 · inbound

Relighting as a Probe of Visual Priors via Augmented Latent Intrinsics cites this paper.

Relighting as a Probe of Visual Priors via Augmented Latent Intrinsics DINOv3

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T05:44:12.034823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:44:12.034823Z digest=sha256:1db3be37b7d9cac30b2e4e8cc2e01dce5e4742af3ee0eba05d1f5c6f00acd84a

Observation 61c92240-ff23-43c6-ac9a-b2eb2a022e32 · inbound

Simplicity Prevails: The Emergence of Generalizable AIGI Detection in Visual Foundation Models cites this paper.

Simplicity Prevails: The Emergence of Generalizable AIGI Detection in Visual Foundation Models DINOv3

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:40:46.248899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T08:40:09.385451Z digest=sha256:7ada694eef6c8f6120cb9c870a3a47174ce3039861be621c480693db33ae6fdb

Observation af13348a-839e-4a08-a6d9-5ad0701e0312 · inbound

Toxicity Assessment in Preclinical Histopathology via Class-Aware Mahalanobis Distance for Known and Novel Anomalies cites this paper.

Toxicity Assessment in Preclinical Histopathology via Class-Aware Mahalanobis Distance for Known and Novel Anomalies DINOv3

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T05:32:09.513188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:32:09.513188Z digest=sha256:6c75c19a126d1b64f60e978335d6ec3d942d747c4c200fc0cfd38f7015a830d0

Observation b2edd37d-8966-4e1c-a74b-1d6c9f4af436 · inbound

Diversity-Preserved Distribution Matching Distillation for Fast Visual Synthesis cites this paper.

Diversity-Preserved Distribution Matching Distillation for Fast Visual Synthesis DINOv3

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T14:40:14.532646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T14:37:09.487669Z digest=sha256:a64c2bb82938a676fa503e3fd28c0f2a1c58a0bef5b64d60ce123fc6aa053b07

Observation 1021f588-4279-4938-b861-5e2df567ecaf · inbound

SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos cites this paper.

SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos DINOv3

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T07:17:30.314610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T07:16:29.588452Z digest=sha256:a22cb36a0c477a2e8d8bec3c811bc292e7274bf4b1c0c6108e8ef28790d85cc4

Observation 45d74fc1-e692-4531-9bf6-9e0d09e40b28 · inbound

Self-Supervised Learning with a Multi-Task Latent Space Objective cites this paper.

Self-Supervised Learning with a Multi-Task Latent Space Objective DINOv3

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T04:08:18.513267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:08:18.513267Z digest=sha256:33fbfc6ef904614c7a0e66fd8d23c88195232a979a60a40ad3e9c251a334c121

Observation e9cc85df-1af1-4f93-b7ca-c43f1144e5e5 · inbound

PANC: Prior-Aware Normalized Cut via Anchor-Augmented Token Graphs cites this paper.

PANC: Prior-Aware Normalized Cut via Anchor-Augmented Token Graphs DINOv3

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-16T06:42:27.814490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T06:41:17.449437Z digest=sha256:523fe2180774bf864237052a75253276e961513e3b693fd25ad4bde380ab573f

Observation 64f21cdd-2499-4ad0-aecf-62c8f7b2d042 · inbound

Brep2Shape: Boundary and Shape Representation Alignment via Self-Supervised Transformers cites this paper.

Brep2Shape: Boundary and Shape Representation Alignment via Self-Supervised Transformers DINOv3

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-03T03:42:05.400340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:42:05.400340Z digest=sha256:375597c0fef009b958a02181394199ee2804d5618282dcb0678cce3cdf9b338d

Observation 2b5282fd-cedd-431e-a6bc-58381023636a · inbound

Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation cites this paper.

Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation DINOv3

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T00:02:14.849577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:02:14.849577Z digest=sha256:9c26cc8b507d74875d8019b6a0ceaf0902597f43d8f86c3c5d8c12dd4af0e071

Observation e2167caf-095e-4f7b-bd71-915bfd384348 · inbound

FAIL: Flow Matching Adversarial Imitation Learning for Image Generation cites this paper.

FAIL: Flow Matching Adversarial Imitation Learning for Image Generation DINOv3

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T23:59:17.124942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:59:17.124942Z digest=sha256:7462e9eb6b0a7dddac390efb798e8292ed5d3e7b74b29739af72e8ada0cc822b

Observation 54dc58ce-6d6a-4f11-a59f-46bc602cabb2 · inbound

LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion cites this paper.

LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion DINOv3

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T23:57:44.675360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:57:44.675360Z digest=sha256:a7d1ee8a5a5ae0fa1e7fd08229c9302e58531a9d9d1e80fd75a3209bacc4c374

Observation a4cc25f9-992f-4ba6-9599-89d16fb03e88 · inbound

Xray-Visual Models: Scaling Vision models on Industry Scale Data cites this paper.

Xray-Visual Models: Scaling Vision models on Industry Scale Data DINOv3

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T22:25:59.691930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:25:59.691930Z digest=sha256:6ce17126b148b14881dc5e6ca7f2d91d7dd432fe16e6bec93632f6f4948538fd

Observation f006d4b0-1ead-4599-8cc4-0ed51e3e01b5 · inbound

Learning to Localize Reference Trajectories in Image-Space for Visual Navigation cites this paper.

Learning to Localize Reference Trajectories in Image-Space for Visual Navigation DINOv3

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T21:56:35.956566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:56:35.956566Z digest=sha256:2a474dbf11092410faf095e374b641ad595cef83927dcc9d1b6fb0819ff4e0d2

Observation 32d63a07-3782-4dd6-ae28-e3cc24b39b90 · inbound

SegMoTE: Token-Level Mixture of Experts for Medical Image Segmentation cites this paper.

SegMoTE: Token-Level Mixture of Experts for Medical Image Segmentation DINOv3

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T21:47:27.543396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:47:27.543396Z digest=sha256:1a258cbfd4f5efca7e39ccd7bfe191d22068463147541c407b81f0e209d03a8d

Observation 201fdea5-2628-4538-856e-6ee6f001249e · inbound

Brewing Stronger Features: Dual-Teacher Distillation for Multispectral Earth Observation cites this paper.

Brewing Stronger Features: Dual-Teacher Distillation for Multispectral Earth Observation DINOv3

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T21:36:06.425621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:36:06.425621Z digest=sha256:997b8d2c24fa091cdb6acff2d9e4edf69ec78d0a53d2bbe1e7ed21a698b03653

Observation 17498c73-ff66-4375-83db-7ec4081a9124 · inbound

MultiModalPFN: Extending Prior-Data Fitted Networks for Multimodal Tabular Learning cites this paper.

MultiModalPFN: Extending Prior-Data Fitted Networks for Multimodal Tabular Learning DINOv3

Reference 51

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T20:11:34.272487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:11:22.146641Z digest=sha256:244bde5e62532a3ea3458c49125d074263b19120f71ba14967ab30c4b1464938

Observation 41aeb9b8-c3fb-4927-bdf1-1369a9e06810 · inbound

SubspaceAD: Training-Free Few-Shot Anomaly Detection via Subspace Modeling cites this paper.

SubspaceAD: Training-Free Few-Shot Anomaly Detection via Subspace Modeling DINOv3

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T18:50:16.775488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T18:48:08.212050Z digest=sha256:9bd918918f8f1b355e2ad331041f76bcaa95a9297110d6a7e5650dece7f168b7

Observation 323c7d98-0c37-4159-8ac3-8caf63d2c8a6 · inbound

PRIMA: Pre-training with Risk-integrated Image-Metadata Alignment for Medical Diagnosis via LLM cites this paper.

PRIMA: Pre-training with Risk-integrated Image-Metadata Alignment for Medical Diagnosis via LLM DINOv3

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T20:27:16.942403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:27:16.942403Z digest=sha256:284dc3b8724305ef41cef9229aa98cec288c6fe5567290e7fb706ecb2e6f4ab9

Observation d9d02977-36d3-4638-94af-0306b148792e · inbound

SOTAlign: Semi-Supervised Alignment of Unimodal Vision and Language Models via Optimal Transport cites this paper.

SOTAlign: Semi-Supervised Alignment of Unimodal Vision and Language Models via Optimal Transport DINOv3

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T20:27:15.665278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:27:15.665278Z digest=sha256:3d64b5003e9ffafc094a374a2e548cc0464d62fdd7ffec4bc4ca07ad75ecfb2d

Observation 36919d1c-4681-4f4f-b0e9-fb11f5a256b1 · inbound

Multimodal Optimal Transport for Training-free Temporal Segmentation in Surgical Robotics cites this paper.

Multimodal Optimal Transport for Training-free Temporal Segmentation in Surgical Robotics DINOv3

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T11:40:03.376835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T11:38:01.845488Z digest=sha256:c0f778d35296d70c8f97cd999343662299372941ec828cb183f8cd9af5505872

Observation b75d8483-1bf4-4d40-8b04-0505d750cb14 · inbound

Differential privacy representation geometry for medical image analysis cites this paper.

Differential privacy representation geometry for medical image analysis DINOv3

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T18:06:25.239050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T18:05:40.153547Z digest=sha256:5ada9ef8092e772bc5a227a4338f576b67b2ee3d4bd0fb7f1281216222cb5779

Observation 54ec6ce8-f183-4260-afb1-ff50dddd6bb6 · inbound

Compact Task-Aligned Imitation Learning for Laboratory Automation cites this paper.

Compact Task-Aligned Imitation Learning for Laboratory Automation DINOv3

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T19:45:29.468990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T19:45:29.468990Z digest=sha256:74bdc43c179c2e332fbdef75a67dbece613b610ce34b6de8b7f6a657822f25b3

Observation 376a34b1-e4fb-41d4-9556-10c6378ee4bd · inbound

LUMOS: Latent Universal Medical Priors for Segmentation cites this paper.

LUMOS: Latent Universal Medical Priors for Segmentation DINOv3

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T19:46:36.682107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:46:36.682107Z digest=sha256:420543438f6013f0a1a57981f6f31ae23c6e85949530ca26a1cb77b528d573d8

Observation 3a22770f-d492-4842-9138-4bcf95c148a2 · inbound

RA-Det: Towards Universal Detection of AI-Generated Images via Robustness Asymmetry cites this paper.

RA-Det: Towards Universal Detection of AI-Generated Images via Robustness Asymmetry DINOv3

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T19:39:17.016668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:39:17.016668Z digest=sha256:22cbf597a0316c0cf8a81cb0f77cfe1a09abba7a26c0e29d09cc1f123c109639