Pith. sign in

Paper Citation Record · LEDGER

Object-level Self-Distillation for Vision Pretraining

As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.05409.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05409 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:52:54.973978Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact1
  • verified fuzzy20
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0c4c77bd-525f-4379-b8ac-ad7ba49c8b22 · outbound

This paper cites Deep ViT Features as Dense Visual Descriptors.

Object-level Self-Distillation for Vision Pretraining Deep ViT Features as Dense Visual Descriptors

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.657235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.657235Z digest=sha256:ebfa6c98dc8a40e9018376ea4b5134d642974c0cdf5088b187ddfc5d19f74ed5

Observation 8f4f1e59-bdc1-417a-94f1-0755f3bcd612 · outbound

This paper cites Self-supervised learning from images with a joint-embedding predictive architecture.

Object-level Self-Distillation for Vision Pretraining Self-supervised learning from images with a joint-embedding predictive architecture

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.662402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.662402Z digest=sha256:182139b813f1ef3f0b17b18cc59cc51a6d757b722df1317cfff054a1788dc5a5

Observation 39f8aa84-736b-4159-b3c9-ac111e486596 · outbound

This paper cites Towards in-context scene understanding.

Object-level Self-Distillation for Vision Pretraining Towards in-context scene understanding

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.613282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.666837Z digest=sha256:5b27469a4bcd9035c9416eff714f7d45314ca66f7de2fff480c21b8756f60a66

Observation d601e02e-6555-4c2b-bed4-86e731552f7b · outbound

This paper cites BEiT: BERT Pre-Training of Image Transformers.

Object-level Self-Distillation for Vision Pretraining BEiT: BERT Pre-Training of Image Transformers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.797778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.797778Z digest=sha256:30a008ef6b0640d4ffb12a829e8347eccf51eb43f4f524f890caddc6d43f72fa

Observation d4e66b74-3c82-41a3-91cd-9a3d2daea8e2 · outbound

This paper cites Are we done with ImageNet?.

Object-level Self-Distillation for Vision Pretraining Are we done with ImageNet?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.802249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.802249Z digest=sha256:de1d62804c1f439b12a88b61b9d121fb75094726a77356a9c95bf5f5feba55cb

Observation fd8d6619-be2a-41c8-80f6-6679c54ededb · outbound

This paper cites MONet: Unsupervised Scene Decomposition and Representation.

Object-level Self-Distillation for Vision Pretraining MONet: Unsupervised Scene Decomposition and Representation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.806304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.806304Z digest=sha256:636c8c8353334c897dd73ff409bd9365f97b240073cb080fd045707c91501213

Observation 3ab378cd-c2ec-4af3-aa9d-9ce881e0176d · outbound

This paper cites End-to-end object detection with transformers.

Object-level Self-Distillation for Vision Pretraining End-to-end object detection with transformers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.811166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.811166Z digest=sha256:2c375e0f3b98755e86c23245fa10d7905a3560c6da9ff7af8d624401c28ce692

Observation 2bc2c9a5-7fa3-49c1-a992-c00bf875414b · outbound

This paper cites Emerging properties in self-supervised vision transformers.

Object-level Self-Distillation for Vision Pretraining Emerging properties in self-supervised vision transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.815084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.815084Z digest=sha256:153ac7538767e53e5a805ac6ea3518beffe7d922576b5e995217de408b7c746a

Observation 9c82dc22-3163-4437-b965-6312223d0ff1 · outbound

This paper cites A simple framework for contrastive learning of visual representations.

Object-level Self-Distillation for Vision Pretraining A simple framework for contrastive learning of visual representations

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.818818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.818818Z digest=sha256:5aa54004350c6a8e92c5784582c1cc4a27e95052b78419485a3d363fcac1f1ac

Observation 6299418c-e640-408f-9168-6fa5a6d2a5e3 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Object-level Self-Distillation for Vision Pretraining Imagenet: A large-scale hierarchical image database

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.822453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.822453Z digest=sha256:0547aeff38ae158e28170e2b4b6ca3a2c03e265c2914e091a2dca2154140fe29

Observation 67f2652a-8366-43d9-8cec-7b9faaf0147f · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Object-level Self-Distillation for Vision Pretraining BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.826195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.826195Z digest=sha256:b3cf0dbcb27dedae1cf83242e9cbf0875bd5a97d7c54258d2767173e11f0df57

Observation 85f2dc7d-ebbe-4e57-8242-e868f51b76f3 · outbound

This paper cites On the transfer of object-centric representation learning.

Object-level Self-Distillation for Vision Pretraining On the transfer of object-centric representation learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.559232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.830333Z digest=sha256:961c9119117d47ece34ecbd7d0a60da256c8bcaf491a1c88b68a8289ed98563e

Observation 263ed725-ad70-4fa0-bbb4-5400f5cb24b8 · outbound

This paper cites Attention over learned object embeddings enables complex visual reasoning.

Object-level Self-Distillation for Vision Pretraining Attention over learned object embeddings enables complex visual reasoning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.545518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.833978Z digest=sha256:0c883711f70c5dfba0a3777d98e34726e7a2616ae51efa16408748973273a265

Observation 54bf1c76-00f2-4e1e-9baa-e2726dd3ca84 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Object-level Self-Distillation for Vision Pretraining An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.837670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.837670Z digest=sha256:17f3402553b1a933101796bea575128e5e78276225414706c443649f30f0ec35

Observation d8b15774-5d64-403f-80cc-e47d18cce68e · outbound

This paper cites The pascal visual object classes challenge: A retrospective.

Object-level Self-Distillation for Vision Pretraining The pascal visual object classes challenge: A retrospective

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.841715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.841715Z digest=sha256:cb53c3185d691d47ca63bdcbc2a738a8621ab1bbe604c22e616a64960d58481e

Observation 2662183d-8910-4f6d-b7c4-603567a3e4f7 · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning.

Object-level Self-Distillation for Vision Pretraining Bootstrap your own latent-a new approach to self-supervised learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.845527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.845527Z digest=sha256:06138df653d521836de1ba77faeca03a0d308bc300a3e396f39ae46e58ac4dfc

Observation 72ff2598-229a-4858-af51-980ec1c33472 · outbound

This paper cites Unsupervised Semantic Segmentation by Distilling Feature Correspondences.

Object-level Self-Distillation for Vision Pretraining Unsupervised Semantic Segmentation by Distilling Feature Correspondences

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.849121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.849121Z digest=sha256:9ba3cf3a27748d5e6aa0875bccdba570406eaab3e0fe4d011783b82b888c16ed

Observation af5ffdda-990a-4b83-baf3-f15dbf0de28e · outbound

This paper cites Masked autoencoders are scalable vision learners.

Object-level Self-Distillation for Vision Pretraining Masked autoencoders are scalable vision learners

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.513013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.853099Z digest=sha256:7adc70fa16cd0975669c2de2a8ff6de08085d2448aad00977f544673635c9730

Observation 1fb37551-0c9e-4e79-9c16-9b27a6ff328c · outbound

This paper cites Efficient visual pretraining with contrastive detection.

Object-level Self-Distillation for Vision Pretraining Efficient visual pretraining with contrastive detection

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.499495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.856862Z digest=sha256:199a7712e72dd3916e5d762ff1611780ae6486633ee47fcb1123c86321f0b25b

Observation 808d2e15-e936-4218-8433-32d75d576558 · outbound

This paper cites Object discovery and representation networks.

Object-level Self-Distillation for Vision Pretraining Object discovery and representation networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.486121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.860732Z digest=sha256:bb3e678df96ea68eb99f675acca4a9b99ad45f45c0876d28b05e11357f98fe7b

Observation 218af056-abeb-4cd7-aa74-1ac8f2c8b639 · outbound

This paper cites Segment anything.

Object-level Self-Distillation for Vision Pretraining Segment anything

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.864448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.864448Z digest=sha256:0e5d4432ec25fd7bbf44d8890ea2e10570154558019a96196a025213b531beb2

Observation bfbcf340-dcd5-43b8-92d1-87691d22e3bd · outbound

This paper cites CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping.

Object-level Self-Distillation for Vision Pretraining CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:52:55.106571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.868199Z digest=sha256:047d2aa4daae400b7682e0166eb2ed3af115c16c93ec6af0cd6e6b4b9ded83ce

Observation 32f05046-d05f-442a-b0ca-db80917fbbd6 · outbound

This paper cites Microsoft coco: Common objects in context.

Object-level Self-Distillation for Vision Pretraining Microsoft coco: Common objects in context

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.872053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.872053Z digest=sha256:839033b04ef4e3bdf04ee568d9aeec201bbf653ec986390fec63d66cc8a3ae30

Observation dc097102-f3a6-4cac-bc94-2a7590b8a648 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

Object-level Self-Distillation for Vision Pretraining Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.453723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.875853Z digest=sha256:941618afa22d977e2565a5a8202e5bfea33352500cc7d60281aa84242405849e

Observation 77737830-6a7d-4300-a8ba-e13d6cac174d · outbound

This paper cites Object-centric learning with slot attention.

Object-level Self-Distillation for Vision Pretraining Object-centric learning with slot attention

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.879463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.879463Z digest=sha256:870271061881aedcf87acf558758a7f632160456eea9f06d54f325036bcef1d5

Observation 3090c0ee-6db9-4f8e-9af0-b3132d803359 · outbound

This paper cites Class-agnostic object detection with multi-modal transformer.

Object-level Self-Distillation for Vision Pretraining Class-agnostic object detection with multi-modal transformer

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.430040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.883202Z digest=sha256:66c32d7f905c6ff6e6d4e63ae6ffaee1f4e4eb09651ebae592a95f9b62a2d80a

Observation 4b8de29f-a972-4d86-b301-0374c377f368 · outbound

This paper cites Exploring the Effectiveness of Object-Centric Representations in Visual Question Answering: Comparative Insights with Foundation Models.

Object-level Self-Distillation for Vision Pretraining Exploring the Effectiveness of Object-Centric Representations in Visual Question Answering: Comparative Insights with Foundation Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.887043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.887043Z digest=sha256:bf9da71df286147b2909cf80fced032a7352432b66317c8be75c45cd32177fab

Observation d30f897e-0ba8-4e5d-9208-a8839eadecd8 · outbound

This paper cites Unsupervised learning of dense visual representations.

Object-level Self-Distillation for Vision Pretraining Unsupervised learning of dense visual representations

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.416858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.891247Z digest=sha256:fc25351030a08f26fcc1f51884bd70351f403129894be22d4bf85225eb20aaf7

Observation 969274a8-d206-4e24-b87b-dc8e53d88812 · outbound

This paper cites Neural congealing: Aligning images to a joint semantic atlas.

Object-level Self-Distillation for Vision Pretraining Neural congealing: Aligning images to a joint semantic atlas

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.403336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.894863Z digest=sha256:f3bd5421b35e339991bf8105a00a14e894d0656462bd1b08f95efad2d7917a9c

Observation 861c41db-1346-4146-bde7-de33800adf3e · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Object-level Self-Distillation for Vision Pretraining DINOv2: Learning Robust Visual Features without Supervision

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.898506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.898506Z digest=sha256:5c2fe8160a3cf692dcdac27567dd295798d3b8212b9d1a2a337429e062f78b13

Observation 027c8501-67b8-4c3c-ad2b-f0a534165b98 · outbound

This paper cites Improving language understanding by generative pre-training.(2018), 2018.

Object-level Self-Distillation for Vision Pretraining Improving language understanding by generative pre-training.(2018), 2018

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.389838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.902139Z digest=sha256:3a2ecf42eaffc224ef17f0c389ee3c5bfed4c0491951456fcbc8212a74541ca4

Observation 3b797630-d64c-402c-a2ca-98a1076a9a89 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Object-level Self-Distillation for Vision Pretraining Learning transferable visual models from natural language supervision

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.906427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.906427Z digest=sha256:b7138fe8bb1f496945d85ecd73e92651f0b88b30ed80b9f148bc79e54fd45956

Observation d89c2bf2-2718-4400-95ba-7d5a8618fc91 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Object-level Self-Distillation for Vision Pretraining SAM 2: Segment Anything in Images and Videos

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.910228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.910228Z digest=sha256:b9abcbc50acb1ce9d702904a0bdd80341331c9f63ffe4ce8504ef236cf027133

Observation 4cf3507e-0984-47a2-b9a9-dc5fd72da7f1 · outbound

This paper cites Do imagenet classifiers generalize to imagenet? In International conference on machine learning, pages 5389--5400.

Object-level Self-Distillation for Vision Pretraining Do imagenet classifiers generalize to imagenet? In International conference on machine learning, pages 5389--5400

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.914054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.914054Z digest=sha256:c5b6754269b6541746d56359425ed5d36542c3e584a727721c8fe1a2825f0659

Observation 7619d09c-7cee-4089-9e7a-f6e24c628ba5 · outbound

This paper cites You only look once: Unified, real-time object detection.

Object-level Self-Distillation for Vision Pretraining You only look once: Unified, real-time object detection

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.918254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.918254Z digest=sha256:d48fe1e54e45dfe09a9a8714e0413e5f68acb5172a0f8219a90cc9234d5dec90

Observation baf18923-4db5-43d0-af5b-28b6817b5366 · outbound

This paper cites Are We Done with Object-Centric Learning?.

Object-level Self-Distillation for Vision Pretraining Are We Done with Object-Centric Learning?

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.922097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.922097Z digest=sha256:2292146f62f9eec568107f124ec0a7dc37575e5d5f94faed6d7b4a70df7b8f12

Observation 1b30982e-6436-4d6c-927c-eb499cc00399 · outbound

This paper cites Bridging the Gap to Real-World Object-Centric Learning.

Object-level Self-Distillation for Vision Pretraining Bridging the Gap to Real-World Object-Centric Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.925946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.925946Z digest=sha256:9453775264fe5957338b88431440b1f56041a7ffa65a96a7b2d8f4d7c878c882

Observation 99228d01-3c03-42b2-b57d-d507dc5750ff · outbound

This paper cites Evaluating machine accuracy on imagenet.

Object-level Self-Distillation for Vision Pretraining Evaluating machine accuracy on imagenet

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.349341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.930160Z digest=sha256:2bd4a01e092a101f790f0967bb4b9438b0a86d68b90bcad186cc7f1bf3c69c15

Observation c37011a0-5ce3-4742-8d9c-072e910e3825 · outbound

This paper cites Croc: Cross-view online clustering for dense visual representation learning.

Object-level Self-Distillation for Vision Pretraining Croc: Cross-view online clustering for dense visual representation learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.336102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.934008Z digest=sha256:ca3a256a28ee3bbc5ec21cb0bdf27f6ecf1a997fc67b47905cf81950a3aa842a

Observation e3432fdd-81b0-4249-ae5b-d1a5451d2af9 · outbound

This paper cites Convnets and imagenet beyond accuracy: Understanding mistakes and uncovering biases.

Object-level Self-Distillation for Vision Pretraining Convnets and imagenet beyond accuracy: Understanding mistakes and uncovering biases

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.322766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.937717Z digest=sha256:74f998a1955338293cde5ad07435b3aab5e3cc908bce1fbe668daca2ce5a4de5

Observation 72da9c3d-4061-4759-9807-9c8d44cef805 · outbound

This paper cites Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results.

Object-level Self-Distillation for Vision Pretraining Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.941851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.941851Z digest=sha256:dadc32446bb596e0c6972a52325e9eba9cafe901c7b2f5490efc782b97b7a9ed

Observation 767d379c-3c0b-4cf9-a642-97cef43c6653 · outbound

This paper cites From imagenet to image classification: Contextualizing progress on benchmarks.

Object-level Self-Distillation for Vision Pretraining From imagenet to image classification: Contextualizing progress on benchmarks

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.299105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.945813Z digest=sha256:d38be4b324d4520a451c0dabd0e237a9378759a8472b9e04438b32caaea4606b

Observation 00c2909d-3221-4241-827c-b060f92fc604 · outbound

This paper cites Splicing vit features for semantic appearance transfer.

Object-level Self-Distillation for Vision Pretraining Splicing vit features for semantic appearance transfer

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.284466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.950041Z digest=sha256:45a8d07c15d8adcb32db52814e1f67dfac49c49e99cf83889ef1d247c2ef1c09

Observation 06994615-4d3a-46b6-b97e-381d17e14ad5 · outbound

This paper cites Dense contrastive learning for self-supervised visual pre-training.

Object-level Self-Distillation for Vision Pretraining Dense contrastive learning for self-supervised visual pre-training

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.271415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.953887Z digest=sha256:ef6ce687b6b049ff852fc3969e939ec1c4e1696dd9d95ad125541d58324992be

Observation 5b50c22f-d140-43ab-89bc-aa9cd0b8cb87 · outbound

This paper cites Self-supervised visual representation learning with semantic grouping.

Object-level Self-Distillation for Vision Pretraining Self-supervised visual representation learning with semantic grouping

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.258331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.957725Z digest=sha256:7ed017f49ebd0660dc83fc71e03115a8b8c6afaa74f8e8a278acb46b76a85ed6

Observation 2b237aeb-0368-412d-b3c6-afe5ac14ecd9 · outbound

This paper cites Unsupervised object-level representation learning from scene images.

Object-level Self-Distillation for Vision Pretraining Unsupervised object-level representation learning from scene images

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.244587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.961699Z digest=sha256:9689f1c707ef9c380b28c6ad095850b89c4c6df3253c6f361c429fc578fc9046

Observation b57b8659-4670-4432-ae26-c2d605449ecf · outbound

This paper cites Re-labeling imagenet: from single to multi-labels, from global to localized labels.

Object-level Self-Distillation for Vision Pretraining Re-labeling imagenet: from single to multi-labels, from global to localized labels

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.230722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.965391Z digest=sha256:f463acf341871795c10d87bdb4c950e52d88e34568b7403c9a9aa407fae8a274

Observation 562eab54-3c33-493d-8da8-fa3795e212ab · outbound

This paper cites Scene parsing through ade20k dataset.

Object-level Self-Distillation for Vision Pretraining Scene parsing through ade20k dataset

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.969995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.969995Z digest=sha256:56b1339f08ff9a3fc27c6fd395eb6e72350209aeb130daa87534466467ca2e55

Observation 18ca66a9-4989-4c00-b86e-31483b866a09 · outbound

This paper cites iBOT: Image BERT Pre-Training with Online Tokenizer.

Object-level Self-Distillation for Vision Pretraining iBOT: Image BERT Pre-Training with Online Tokenizer

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.973978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.973978Z digest=sha256:565f2c8402ed653c23e9a6cf36ea2f819ea09686cd4ae8dff3fe6515995e033e

Pith citing papers

No inbound Pith citation observations are available.