Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-09T04:17:19.878360Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 100 inbound Pith citation observations for arXiv:2304.07193.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-09T04:17:19.878360Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T21:31:08.235510Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-11T00:07:42.299741Z
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 55b016cc-551c-47ad-818b-92042c006700 · outbound
DINOv2: Learning Robust Visual Features without Supervision Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c0eead4b-6bd2-4b7d-a913-e85abd8f0764 · outbound
DINOv2: Learning Robust Visual Features without Supervision MultiGrain: a unified image embedding for classes and instances
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation af299ae4-93fd-4081-9dad-3e486c880ac6 · outbound
DINOv2: Learning Robust Visual Features without Supervision Are we done with ImageNet?
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e9595bf0-923e-469b-ac36-f611f6ad4762 · outbound
DINOv2: Learning Robust Visual Features without Supervision On the Opportunities and Risks of Foundation Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 57ca71f7-d270-4429-9d0d-72c6731cf7dd · outbound
DINOv2: Learning Robust Visual Features without Supervision Symbolic Discovery of Optimization Algorithms
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 22f17709-8612-49f1-88e7-495ae501f91e · outbound
DINOv2: Learning Robust Visual Features without Supervision An empirical study of training self-supervised vision transformers
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2e1f5db9-d418-43e4-9e18-357bf242629b · outbound
DINOv2: Learning Robust Visual Features without Supervision PaLM: Scaling Language Modeling with Pathways
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 37267cd6-9275-444c-baf4-9082af89a338 · outbound
DINOv2: Learning Robust Visual Features without Supervision A Simple Recipe for Competitive Low-compute Self supervised Vision Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9b809046-68ed-42d5-914a-67591100a026 · outbound
DINOv2: Learning Robust Visual Features without Supervision Are Large-scale Datasets Necessary for Self-Supervised Pre-training?
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9b24cbce-dfda-43fa-9423-a5e6a1a5e296 · outbound
DINOv2: Learning Robust Visual Features without Supervision Eva: Exploring the limits of masked visual representation learning at scale
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9976cf06-4437-4342-ba15-fce8aad4bc35 · outbound
DINOv2: Learning Robust Visual Features without Supervision Self-supervised Pretraining of Visual Features in the Wild
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8ba28ccd-9876-4e0d-9bab-498307162344 · outbound
DINOv2: Learning Robust Visual Features without Supervision Vision Models Are More Robust And Fair When Pretrained On Uncurated Images Without Supervision
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 59a753fb-2ab0-4bd2-b0b0-4b73ef2635e2 · outbound
DINOv2: Learning Robust Visual Features without Supervision The many faces of robustness: A critical analysis of out-of-distribution generalization
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6da0582a-384f-406c-b65b-f9bed07efe79 · outbound
DINOv2: Learning Robust Visual Features without Supervision Training Compute-Optimal Large Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1d52adaa-7998-4af3-a14b-c171b5807a2e · outbound
DINOv2: Learning Robust Visual Features without Supervision The Kinetics Human Action Video Dataset
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 171fa669-b057-4b32-98a4-f60e80f231fc · outbound
DINOv2: Learning Robust Visual Features without Supervision BinsFormer: Revisiting Adaptive Bins for Monocular Depth Estimation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 62e0ff4e-9338-4509-bf01-155d3599c779 · outbound
DINOv2: Learning Robust Visual Features without Supervision Polarized Self-Attention: Towards High-quality Pixel-wise Regression
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 124182fc-b3af-4ada-bf24-26a8e07d4439 · outbound
DINOv2: Learning Robust Visual Features without Supervision Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 73cc2f39-7ce9-48e1-b823-1660fe4dd218 · outbound
DINOv2: Learning Robust Visual Features without Supervision Carbon Emissions and Large Neural Network Training
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 013e819f-8aa5-41c6-a2be-3dbd69246de3 · outbound
DINOv2: Learning Robust Visual Features without Supervision Learning to Generate Reviews and Discovering Sentiment
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation db2f8fc9-39e0-4edf-8cea-71d977674c59 · outbound
DINOv2: Learning Robust Visual Features without Supervision Imagenet large scale visual recognition challenge.IJCV
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fc68bd92-588a-4e1b-b843-40313b096556 · outbound
DINOv2: Learning Robust Visual Features without Supervision GLU Variants Improve Transformer
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation db7e4380-5da3-4887-8f2c-f7a3599c91c2 · outbound
DINOv2: Learning Robust Visual Features without Supervision UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6ed76909-00b3-40a8-8b28-c91d7db6b4b7 · outbound
DINOv2: Learning Robust Visual Features without Supervision LLaMA: Open and Efficient Foundation Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fc3561b8-0b4e-412c-afce-e9676ed9742c · outbound
DINOv2: Learning Robust Visual Features without Supervision Bench- marking representation learning for natural world image collections
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4efad419-ca9d-4f8a-817e-a438decb6382 · outbound
DINOv2: Learning Robust Visual Features without Supervision Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d5ad5530-cbb2-4705-9211-1bfe7ba1c95d · outbound
DINOv2: Learning Robust Visual Features without Supervision Masked Autoencoders that Listen
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2b6c4006-bbb4-477a-9f9b-89b82d51594e · outbound
DINOv2: Learning Robust Visual Features without Supervision Billion-scale semi-supervised learning for image classification
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5368219e-fe1e-45ec-9f52-1946bd270291 · outbound
DINOv2: Learning Robust Visual Features without Supervision Mugs: A Multi-Granular Self-Supervised Learning Framework
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 51f78af0-e787-4f12-8270-1277c70523bd · outbound
DINOv2: Learning Robust Visual Features without Supervision Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7af36918-2a23-4a77-8ca4-e411a2ca391e · outbound
DINOv2: Learning Robust Visual Features without Supervision We apply the KoLeo regularizer with a weight of 0.1 between the class tokens of the first global crop, for all samples within a GPU without cross-communication for this step
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9d2a2454-8422-42a0-8e64-f0e9ac780900 · outbound
DINOv2: Learning Robust Visual Features without Supervision Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f91d300c-1f50-42d5-9752-77788ab4c79c · outbound
DINOv2: Learning Robust Visual Features without Supervision EMA update for the teacher
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation da1bec99-2754-4351-b94f-13602ac2a4fd · outbound
DINOv2: Learning Robust Visual Features without Supervision (Everingham et al
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8f6af0b6-6c74-4f2d-9268-cf84126ab147 · outbound
DINOv2: Learning Robust Visual Features without Supervision (Van Horn et al
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8b529756-49db-4c09-8226-a3bd04f896f6 · outbound
DINOv2: Learning Robust Visual Features without Supervision (Van Horn et al
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e912152b-8116-4a04-a0a0-f19fd0189464 · outbound
DINOv2: Learning Robust Visual Features without Supervision (Everingham et al
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 577d983d-58a0-4a35-a2ee-2e27ae4f9a2f · inbound
RoMa: Robust Dense Feature Matching DINOv2: Learning Robust Visual Features without Supervision
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2e7b95bf-ef3d-4ca8-b2ba-ab0734c7791a · inbound
A Survey on Multimodal Large Language Models DINOv2: Learning Robust Visual Features without Supervision
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c758da9f-4ae8-4c2a-829c-5088e4d69c04 · inbound
Project Aria: A New Tool for Egocentric Multi-Modal AI Research DINOv2: Learning Robust Visual Features without Supervision
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a1b0f07f-ab34-41b3-9bb8-417fb5194bdc · inbound
Vision Transformers Need Registers DINOv2: Learning Robust Visual Features without Supervision
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c18b963d-d5e6-4add-aa51-6e8965ce4b82 · inbound
Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution DINOv2: Learning Robust Visual Features without Supervision
Reference 249
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a28bd355-7217-4e35-879d-d1ba5c41ef21 · inbound
Causal Unsupervised Semantic Segmentation DINOv2: Learning Robust Visual Features without Supervision
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c4a0bc69-d527-426e-bb7a-b1c3b5d43e85 · inbound
InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks DINOv2: Learning Robust Visual Features without Supervision
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1d8194b2-b8a3-4c50-922f-23b075ee3a32 · inbound
Data-Centric Foundation Models in Computational Healthcare: A Survey DINOv2: Learning Robust Visual Features without Supervision
Reference 215
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ba32c141-fd1b-4623-b3e4-291510747b5d · inbound
Massive Activations in Large Language Models DINOv2: Learning Robust Visual Features without Supervision
Reference 140
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3cb14679-d2ff-4fa7-8471-f157b95c1c26 · inbound
TempCompass: Do Video LLMs Really Understand Videos? DINOv2: Learning Robust Visual Features without Supervision
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6827c88e-c9f4-4974-9d28-a097dea10b28 · inbound
MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training DINOv2: Learning Robust Visual Features without Supervision
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e49771d3-488e-4d21-b591-bf848bf5d706 · inbound
Revisiting Feature Prediction for Learning Visual Representations from Video DINOv2: Learning Robust Visual Features without Supervision
Reference 263
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 80e3b255-1a0f-4d1d-be93-5dab2b44f614 · inbound
Leveraging Medical Foundation Model Features in Graph Neural Network-Based Retrieval of Breast Histopathology Images DINOv2: Learning Robust Visual Features without Supervision
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2e95bcef-086c-4e18-9468-a39b72a2bae7 · inbound
The Platonic Representation Hypothesis DINOv2: Learning Robust Visual Features without Supervision
Reference 197
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9d98c612-6b4b-46ef-80d9-e0a50e3b47bb · inbound
OpenVLA: An Open-Source Vision-Language-Action Model DINOv2: Learning Robust Visual Features without Supervision
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bcc57d77-533f-44e5-ba13-1e1ea51e6e6d · inbound
LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models DINOv2: Learning Robust Visual Features without Supervision
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 507ab40c-7d9b-4c79-bb9b-06b34cea6ec6 · inbound
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation DINOv2: Learning Robust Visual Features without Supervision
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cc5d14ed-fe89-45d6-b204-2febe308a956 · inbound
MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark DINOv2: Learning Robust Visual Features without Supervision
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e13f6784-27fc-4f31-a098-5fce1a425133 · inbound
Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models DINOv2: Learning Robust Visual Features without Supervision
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 61c9ea69-b5ee-4590-bad6-dba33e6bffd7 · inbound
TOAST: Transformer Optimization using Adaptive and Simple Transformations DINOv2: Learning Robust Visual Features without Supervision
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a07877d8-63bb-4be8-9e92-280e74695b8a · inbound
UNCOM: Zero-shot Context-Aware Command Understanding for Tabletop Scenarios DINOv2: Learning Robust Visual Features without Supervision
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2493597c-661c-4c95-aa04-690c68546a8b · inbound
LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding DINOv2: Learning Robust Visual Features without Supervision
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 954f3313-166b-48fd-9fad-01a0ebc22257 · inbound
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning DINOv2: Learning Robust Visual Features without Supervision
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 939fee17-02ae-477d-a9f4-6f81433fc6f7 · inbound
DissolveStereo: Coarse Depth Injection for Zero-Shot Stereo Video Generation DINOv2: Learning Robust Visual Features without Supervision
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3c98600f-acdc-4a4f-854f-28d7deed86b4 · inbound
Large Language Model-Brained GUI Agents: A Survey DINOv2: Learning Robust Visual Features without Supervision
Reference 229
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a3cb008c-22b0-44a4-af34-5c67146fb598 · inbound
Multimodal Contextualized Support for Enhancing Video Retrieval System DINOv2: Learning Robust Visual Features without Supervision
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3171fb7b-80e1-4b08-81e8-5d4a6fa242a5 · inbound
TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies DINOv2: Learning Robust Visual Features without Supervision
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation da509556-579c-4bb6-a1d7-df292ecc4451 · inbound
DepthMaster: Taming Diffusion Models for Monocular Depth Estimation DINOv2: Learning Robust Visual Features without Supervision
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 90e48ea4-efbe-412f-9243-777fcc6fd9c2 · inbound
Benchmarking Vision Foundation Models for Input Monitoring in Autonomous Driving DINOv2: Learning Robust Visual Features without Supervision
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f217d42f-e05c-4275-a0ac-8787f3afc269 · inbound
Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps DINOv2: Learning Robust Visual Features without Supervision
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4ae3679b-685b-40b1-8724-444754e2495c · inbound
Personalization Toolkit: Training Free Personalization of Large Vision Language Models DINOv2: Learning Robust Visual Features without Supervision
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ac2ce2e8-9d79-460e-baff-044e19feca99 · inbound
TripoSG: High-Fidelity 3D Shape Synthesis using Large-Scale Rectified Flow Models DINOv2: Learning Robust Visual Features without Supervision
Reference 134
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 228ae164-ff72-4721-b5c7-b8646e294103 · inbound
Geometry-aided Vision-based Localization of Future Mars Helicopters in Challenging Illumination Conditions DINOv2: Learning Robust Visual Features without Supervision
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7d4a3c5e-9deb-412a-b49d-140949da65fd · inbound
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success DINOv2: Learning Robust Visual Features without Supervision
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9de94152-88ce-4e88-9fce-ae6004b8eb52 · inbound
UniDepthV2: Universal Monocular Metric Depth Estimation Made Simpler DINOv2: Learning Robust Visual Features without Supervision
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 434c806a-a7d3-4c2e-905c-20df8b78e3c6 · inbound
Primus: Enforcing Attention Usage for 3D Medical Image Segmentation DINOv2: Learning Robust Visual Features without Supervision
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 41ba1e14-2956-46f7-b53f-3424dd92297e · inbound
Adaptive Camera Sensor for Vision Models DINOv2: Learning Robust Visual Features without Supervision
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c48a6835-bca2-433b-a9e2-86a5c833d0a4 · inbound
Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model DINOv2: Learning Robust Visual Features without Supervision
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 547e41d2-bed1-405e-ad5d-21b1c3ab55fc · inbound
HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model DINOv2: Learning Robust Visual Features without Supervision
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b53f411a-70d2-43a5-b099-6a64821ed13e · inbound
GAIR: Location-Aware Self-Supervised Contrastive Pre-Training with Geo-Aligned Implicit Representations DINOv2: Learning Robust Visual Features without Supervision
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2de09c84-0f0f-4d7d-9a02-a7ad63efd874 · inbound
Toward Generalizable Forgery Detection and Reasoning DINOv2: Learning Robust Visual Features without Supervision
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 44fcb28a-8b3a-482d-a361-35daef47c51a · inbound
Seedream 3.0 Technical Report DINOv2: Learning Robust Visual Features without Supervision
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 395041c0-217e-4114-8a32-235e63c4abf7 · inbound
Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive Evaluation DINOv2: Learning Robust Visual Features without Supervision
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 64e7dc58-8ace-4e75-a10d-0716daf9c294 · inbound
FreeGraftor: Training-Free Cross-Image Feature Grafting for Subject-Driven Text-to-Image Generation DINOv2: Learning Robust Visual Features without Supervision
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6170fdc2-8c73-4355-9335-c2b64d8b021f · inbound
NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks DINOv2: Learning Robust Visual Features without Supervision
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9252b305-6182-4cad-bb40-52ad42356e16 · inbound
In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer DINOv2: Learning Robust Visual Features without Supervision
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e348e3f9-83b2-42d7-85a6-d4f4880692db · inbound
GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data DINOv2: Learning Robust Visual Features without Supervision
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c1be779b-1e02-44fb-9007-29bec9901230 · inbound
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation DINOv2: Learning Robust Visual Features without Supervision
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6f2c3e87-a354-4abf-bb94-4c044ab34a4e · inbound
Seed1.5-VL Technical Report DINOv2: Learning Robust Visual Features without Supervision
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c060c368-9ac5-48ff-9cb7-a943d60c43bc · inbound
VGGT-SLAM: Dense RGB SLAM Optimized on the SL(4) Manifold DINOv2: Learning Robust Visual Features without Supervision
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dd52bf2f-0608-4efe-abea-051583d32cbe · inbound
MAGI-1: Autoregressive Video Generation at Scale DINOv2: Learning Robust Visual Features without Supervision
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7556c271-7c8a-4958-8cc7-d183c8d72026 · inbound
Policy Contrastive Decoding for Robotic Foundation Models DINOv2: Learning Robust Visual Features without Supervision
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 34967fdb-6d6f-45bc-9e83-b086bbcb7a24 · inbound
FractalMamba++: Scaling Vision Mamba Across Resolutions via Hilbert Fractal Geometry DINOv2: Learning Robust Visual Features without Supervision
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d16b66a9-182f-40ab-bfcd-13469f92fcde · inbound
Interactive Post-Training for Vision-Language-Action Models DINOv2: Learning Robust Visual Features without Supervision
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a65220d4-64fe-4246-b377-f6a8418344d8 · inbound
Slot-MLLM: Object-Centric Visual Tokenization for Multimodal LLM DINOv2: Learning Robust Visual Features without Supervision
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d50dc313-3588-4bc8-ba8c-80acc68977f3 · inbound
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning DINOv2: Learning Robust Visual Features without Supervision
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3a85a4d7-8d77-4fbb-99b8-a7a6480fdf5f · inbound
ImgEdit: A Unified Image Editing Dataset and Benchmark DINOv2: Learning Robust Visual Features without Supervision
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation df892426-61d1-41f0-a5ea-88a4d230b471 · inbound
Geometry-Editable and Appearance-Preserving Object Compositon DINOv2: Learning Robust Visual Features without Supervision
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e56d951b-5ba2-40db-8de8-f45ba64782f0 · inbound
A European Multi-Center Breast Cancer MRI Dataset DINOv2: Learning Robust Visual Features without Supervision
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6ba41af7-6b24-486f-93a4-085ee0cbc71a · inbound
AuralSAM2: Enabling SAM2 Hear Through Pyramid Audio-Visual Feature Prompting DINOv2: Learning Robust Visual Features without Supervision
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 276b12ab-a255-4faf-b8fa-e263f013080b · inbound
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering DINOv2: Learning Robust Visual Features without Supervision
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 396c38b5-9580-44d3-b0d8-77bc02935d8e · inbound
UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation DINOv2: Learning Robust Visual Features without Supervision
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7664a232-90a5-4b13-a6f7-17530da8d901 · inbound
Adapting Vision-Language Foundation Model for Next Generation Medical Ultrasound Image Analysis DINOv2: Learning Robust Visual Features without Supervision
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2bda31ff-adde-4223-ae84-e3c5287a7ff7 · inbound
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models DINOv2: Learning Robust Visual Features without Supervision
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 433bc99e-583f-4e76-9b53-8f9b6dc37c0c · inbound
The Less You Depend, The More You Learn: Synthesizing Novel Views from Sparse, Unposed Images with Minimal 3D Knowledge DINOv2: Learning Robust Visual Features without Supervision
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 57365b87-d7e6-4990-aca8-c82a929eeb0a · inbound
V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning DINOv2: Learning Robust Visual Features without Supervision
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 72851fcb-8c27-4421-a253-0c85d09b6649 · inbound
CLIP the Landscape: Automated Tagging of Crowdsourced Landscape Images DINOv2: Learning Robust Visual Features without Supervision
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation efb1dbd7-41fa-4565-9923-076a759c2bbb · inbound
Hunyuan3D 2.1: From Images to High-Fidelity 3D Assets with Production-Ready PBR Material DINOv2: Learning Robust Visual Features without Supervision
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 78d241c5-bccb-4740-99ee-1162d93747af · inbound
OmniGen2: Towards Instruction-Aligned Multimodal Generation DINOv2: Learning Robust Visual Features without Supervision
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c4e4a262-85db-4ee2-8300-f1c4f750e3ca · inbound
GenHSI: Controllable Generation of Human-Scene Interaction Videos DINOv2: Learning Robust Visual Features without Supervision
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a30cfea2-24dd-4de9-b4c3-379cc83011a6 · inbound
MoGe-2: Accurate Monocular Geometry with Metric Scale and Sharp Details DINOv2: Learning Robust Visual Features without Supervision
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d8de1717-616a-4f19-986b-1fd583b7780c · inbound
Perception-Aware Policy Optimization for Multimodal Reasoning DINOv2: Learning Robust Visual Features without Supervision
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6c6d406f-35f3-480c-946f-bd2dde97eb95 · inbound
SCOOTER: A Human Evaluation Framework for Unrestricted Adversarial Examples DINOv2: Learning Robust Visual Features without Supervision
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bfd871df-9dab-4102-98ff-45d0d4354cc0 · inbound
Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling DINOv2: Learning Robust Visual Features without Supervision
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a4350520-e548-4337-82dc-7a2d069ce964 · inbound
Navigating the Challenges of AI-Generated Image Detection in the Wild: What Truly Matters? DINOv2: Learning Robust Visual Features without Supervision
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ddd65d56-4e3d-437a-af53-33287ad36be9 · inbound
Streaming 4D Visual Geometry Transformer DINOv2: Learning Robust Visual Features without Supervision
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 66191822-d84f-4aad-baa4-d435340c8d80 · inbound
$\pi^3$: Permutation-Equivariant Visual Geometry Learning DINOv2: Learning Robust Visual Features without Supervision
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3ba6a66d-6633-4545-bcdc-788539450d06 · inbound
Frozen Forecasting: A Unified Evaluation DINOv2: Learning Robust Visual Features without Supervision
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 42ee7a51-17f5-44c2-a573-e9ac9fc8f509 · inbound
ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning DINOv2: Learning Robust Visual Features without Supervision
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fa37519e-3991-459a-b01e-5b462ab3d50d · inbound
AI in Agriculture: A Survey of Deep Learning Techniques for Crops, Fisheries and Livestock DINOv2: Learning Robust Visual Features without Supervision
Reference 151
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 92f37ce5-ae18-4595-b74b-214ac051d98c · inbound
Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation DINOv2: Learning Robust Visual Features without Supervision
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 80a9f126-b12e-4b48-a13a-770070cccaf7 · inbound
IntrinsicWeather: Controllable Weather Editing in Intrinsic Space DINOv2: Learning Robust Visual Features without Supervision
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3853af0a-42cc-4146-b354-2c67a8ed117d · inbound
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation DINOv2: Learning Robust Visual Features without Supervision
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 370fa0db-f45b-47ec-a1bf-8106d6fd89f3 · inbound
Dino U-Net: Exploiting High-Fidelity Dense Features from Foundation Models for Medical Image Segmentation DINOv2: Learning Robust Visual Features without Supervision
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3382f296-6469-4f09-a436-c37ae4d106a7 · inbound
FastVGGT: Training-Free Acceleration of Visual Geometry Transformer DINOv2: Learning Robust Visual Features without Supervision
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 16c39280-a7d5-4a04-8cf1-a25cab76ea5b · inbound
LUIVITON: Learned Universal Interoperable VIrtual Try-ON DINOv2: Learning Robust Visual Features without Supervision
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 151f8100-3a5c-43fd-a11f-d3ff3ad44984 · inbound
Block-Sparse Global Attention for Efficient Multi-View Geometry Transformers DINOv2: Learning Robust Visual Features without Supervision
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 48747c6e-fdcd-4b0e-80cd-33aed443f89f · inbound
One View, Many Worlds: Single-Image to 3D Object Meets Generative Domain Randomization for One-Shot 6D Pose Estimation DINOv2: Learning Robust Visual Features without Supervision
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18868781-093f-4806-b8ec-b558adfd1a18 · inbound
Domain Knowledge is Power: Leveraging Physiological Priors for Self Supervised Representation Learning in Electrocardiography DINOv2: Learning Robust Visual Features without Supervision
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b690062-15d8-4053-9c37-476f70e7730b · inbound
Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities DINOv2: Learning Robust Visual Features without Supervision
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e81d0d67-8d36-43ac-acb7-b7fd53065b58 · inbound
ViewSparsifier: Killing Redundancy in Multi-View Plant Phenotyping DINOv2: Learning Robust Visual Features without Supervision
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c06163f8-3aff-44a6-b1ed-7e820b97f396 · inbound
RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation DINOv2: Learning Robust Visual Features without Supervision
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30cfddf2-47f1-4eb0-9d19-1375bfe987fb · inbound
Mind Meets Space: Rethinking Agentic Spatial Intelligence from a Neuroscience-inspired Perspective DINOv2: Learning Robust Visual Features without Supervision
Reference 150
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf5f5831-113e-4ebc-be0c-58225de6a9f9 · inbound
Learning Object-Centric Representations in SAR Images with Multi-Level Feature Fusion DINOv2: Learning Robust Visual Features without Supervision
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0fcf758-83ae-4563-850a-aaa7e5619bee · inbound
Image Recognition with Vision and Language Embeddings of VLMs DINOv2: Learning Robust Visual Features without Supervision
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 074847cd-7d7e-4632-90ab-04f6897547da · inbound
Unsupervised Integrated-Circuit Defect Segmentation via Image-Intrinsic Normality DINOv2: Learning Robust Visual Features without Supervision
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7a1aacb-ddf0-4850-9ced-6162638e0d3f · inbound
MimicDroid: In-Context Learning for Humanoid Robot Manipulation from Human Play Videos DINOv2: Learning Robust Visual Features without Supervision
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3708538-b30e-4440-a28d-55cdbdf4754c · inbound
Self-supervised Learning Of Visual Pose Estimation Without Pose Labels By Classifying LED States DINOv2: Learning Robust Visual Features without Supervision
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ba4bd1b-cc4e-4b7b-ab2e-342bd5ebda52 · inbound
ImMimic: Cross-Domain Imitation from Human Videos via Mapping and Interpolation DINOv2: Learning Robust Visual Features without Supervision
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a283243-ebdd-43af-8725-57af9b3aa62b · inbound
RAPTOR: A Foundation Policy for Quadrotor Control DINOv2: Learning Robust Visual Features without Supervision
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.