Pith. sign in

Paper Citation Record · LEDGER

Object-level Self-Distillation for Vision Pretraining

As of 8 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.05409.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05409 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:52:54.973978Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact1
  • verified fuzzy20
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0c4c77bd-525f-4379-b8ac-ad7ba49c8b22 · outbound

This paper cites Deep ViT Features as Dense Visual Descriptors.

Object-level Self-Distillation for Vision Pretraining Deep ViT Features as Dense Visual Descriptors

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.657235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.657235Z digest=sha256:ebfa6c98dc8a40e9018376ea4b5134d642974c0cdf5088b187ddfc5d19f74ed5

Observation 8f4f1e59-bdc1-417a-94f1-0755f3bcd612 · outbound

This paper cites Self-supervised learning from images with a joint-embedding predictive architecture.

Object-level Self-Distillation for Vision Pretraining Self-supervised learning from images with a joint-embedding predictive architecture

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.662402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.662402Z digest=sha256:182139b813f1ef3f0b17b18cc59cc51a6d757b722df1317cfff054a1788dc5a5

Observation 39f8aa84-736b-4159-b3c9-ac111e486596 · outbound

This paper cites Towards in-context scene understanding.

Object-level Self-Distillation for Vision Pretraining Towards in-context scene understanding

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.613282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.666837Z digest=sha256:2e619042adf508222dee8dbc431b6ff36bef68bde3401a1c9e2ba029619b17e1

Observation d601e02e-6555-4c2b-bed4-86e731552f7b · outbound

This paper cites BEiT: BERT Pre-Training of Image Transformers.

Object-level Self-Distillation for Vision Pretraining BEiT: BERT Pre-Training of Image Transformers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.797778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.797778Z digest=sha256:30a008ef6b0640d4ffb12a829e8347eccf51eb43f4f524f890caddc6d43f72fa

Observation d4e66b74-3c82-41a3-91cd-9a3d2daea8e2 · outbound

This paper cites Are we done with ImageNet?.

Object-level Self-Distillation for Vision Pretraining Are we done with ImageNet?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.802249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.802249Z digest=sha256:de1d62804c1f439b12a88b61b9d121fb75094726a77356a9c95bf5f5feba55cb

Observation fd8d6619-be2a-41c8-80f6-6679c54ededb · outbound

This paper cites MONet: Unsupervised Scene Decomposition and Representation.

Object-level Self-Distillation for Vision Pretraining MONet: Unsupervised Scene Decomposition and Representation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.806304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.806304Z digest=sha256:636c8c8353334c897dd73ff409bd9365f97b240073cb080fd045707c91501213

Observation 3ab378cd-c2ec-4af3-aa9d-9ce881e0176d · outbound

This paper cites End-to-end object detection with transformers.

Object-level Self-Distillation for Vision Pretraining End-to-end object detection with transformers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.811166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.811166Z digest=sha256:2c375e0f3b98755e86c23245fa10d7905a3560c6da9ff7af8d624401c28ce692

Observation 2bc2c9a5-7fa3-49c1-a992-c00bf875414b · outbound

This paper cites Emerging properties in self-supervised vision transformers.

Object-level Self-Distillation for Vision Pretraining Emerging properties in self-supervised vision transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.815084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.815084Z digest=sha256:153ac7538767e53e5a805ac6ea3518beffe7d922576b5e995217de408b7c746a

Observation 9c82dc22-3163-4437-b965-6312223d0ff1 · outbound

This paper cites A simple framework for contrastive learning of visual representations.

Object-level Self-Distillation for Vision Pretraining A simple framework for contrastive learning of visual representations

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.818818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.818818Z digest=sha256:5aa54004350c6a8e92c5784582c1cc4a27e95052b78419485a3d363fcac1f1ac

Observation 6299418c-e640-408f-9168-6fa5a6d2a5e3 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Object-level Self-Distillation for Vision Pretraining Imagenet: A large-scale hierarchical image database

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.822453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.822453Z digest=sha256:0547aeff38ae158e28170e2b4b6ca3a2c03e265c2914e091a2dca2154140fe29

Observation 67f2652a-8366-43d9-8cec-7b9faaf0147f · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Object-level Self-Distillation for Vision Pretraining BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.826195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.826195Z digest=sha256:b3cf0dbcb27dedae1cf83242e9cbf0875bd5a97d7c54258d2767173e11f0df57

Observation 85f2dc7d-ebbe-4e57-8242-e868f51b76f3 · outbound

This paper cites On the transfer of object-centric representation learning.

Object-level Self-Distillation for Vision Pretraining On the transfer of object-centric representation learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.559232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.830333Z digest=sha256:41d311db3e21932da90ff626859d30bc8798c241d81b97804e856c2ce85e2817

Observation 263ed725-ad70-4fa0-bbb4-5400f5cb24b8 · outbound

This paper cites Attention over learned object embeddings enables complex visual reasoning.

Object-level Self-Distillation for Vision Pretraining Attention over learned object embeddings enables complex visual reasoning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.545518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.833978Z digest=sha256:74e3d81b9638bb8266e85dfb48a1563ec3655644b0e1afd06dec8680e201c3cb

Observation 54bf1c76-00f2-4e1e-9baa-e2726dd3ca84 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Object-level Self-Distillation for Vision Pretraining An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.837670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.837670Z digest=sha256:17f3402553b1a933101796bea575128e5e78276225414706c443649f30f0ec35

Observation d8b15774-5d64-403f-80cc-e47d18cce68e · outbound

This paper cites The pascal visual object classes challenge: A retrospective.

Object-level Self-Distillation for Vision Pretraining The pascal visual object classes challenge: A retrospective

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.841715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.841715Z digest=sha256:cb53c3185d691d47ca63bdcbc2a738a8621ab1bbe604c22e616a64960d58481e

Observation 2662183d-8910-4f6d-b7c4-603567a3e4f7 · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning.

Object-level Self-Distillation for Vision Pretraining Bootstrap your own latent-a new approach to self-supervised learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.845527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.845527Z digest=sha256:06138df653d521836de1ba77faeca03a0d308bc300a3e396f39ae46e58ac4dfc

Observation 72ff2598-229a-4858-af51-980ec1c33472 · outbound

This paper cites Unsupervised Semantic Segmentation by Distilling Feature Correspondences.

Object-level Self-Distillation for Vision Pretraining Unsupervised Semantic Segmentation by Distilling Feature Correspondences

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.849121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.849121Z digest=sha256:9ba3cf3a27748d5e6aa0875bccdba570406eaab3e0fe4d011783b82b888c16ed

Observation af5ffdda-990a-4b83-baf3-f15dbf0de28e · outbound

This paper cites Masked autoencoders are scalable vision learners.

Object-level Self-Distillation for Vision Pretraining Masked autoencoders are scalable vision learners

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.513013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.853099Z digest=sha256:edf0495827dae0246e92e9e4da801b7ec36fec4130368b618a50b3ff5c8cd577

Observation 1fb37551-0c9e-4e79-9c16-9b27a6ff328c · outbound

This paper cites Efficient visual pretraining with contrastive detection.

Object-level Self-Distillation for Vision Pretraining Efficient visual pretraining with contrastive detection

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.499495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.856862Z digest=sha256:f70ede7821259166047e82587757569796d3c4f951f817ab98928f285471364e

Observation 808d2e15-e936-4218-8433-32d75d576558 · outbound

This paper cites Object discovery and representation networks.

Object-level Self-Distillation for Vision Pretraining Object discovery and representation networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.486121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.860732Z digest=sha256:768bd9732361ec1ff29d8863d93b4fa300bb90210c9105a29d6bc00871e6a7e6

Observation 218af056-abeb-4cd7-aa74-1ac8f2c8b639 · outbound

This paper cites Segment anything.

Object-level Self-Distillation for Vision Pretraining Segment anything

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.864448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.864448Z digest=sha256:0e5d4432ec25fd7bbf44d8890ea2e10570154558019a96196a025213b531beb2

Observation bfbcf340-dcd5-43b8-92d1-87691d22e3bd · outbound

This paper cites CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping.

Object-level Self-Distillation for Vision Pretraining CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:52:55.106571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.868199Z digest=sha256:16921070f108fe6957bd9a1c316df2c27c4f8c8369f096256dba23e644c0ce6e

Observation 32f05046-d05f-442a-b0ca-db80917fbbd6 · outbound

This paper cites Microsoft coco: Common objects in context.

Object-level Self-Distillation for Vision Pretraining Microsoft coco: Common objects in context

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.872053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.872053Z digest=sha256:839033b04ef4e3bdf04ee568d9aeec201bbf653ec986390fec63d66cc8a3ae30

Observation dc097102-f3a6-4cac-bc94-2a7590b8a648 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

Object-level Self-Distillation for Vision Pretraining Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.453723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.875853Z digest=sha256:c549c998513fe1cb6c8b5c5f35e02244f29c289fc63e9153ca7ab21abe7f71a6

Observation 77737830-6a7d-4300-a8ba-e13d6cac174d · outbound

This paper cites Object-centric learning with slot attention.

Object-level Self-Distillation for Vision Pretraining Object-centric learning with slot attention

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.879463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.879463Z digest=sha256:870271061881aedcf87acf558758a7f632160456eea9f06d54f325036bcef1d5

Observation 3090c0ee-6db9-4f8e-9af0-b3132d803359 · outbound

This paper cites Class-agnostic object detection with multi-modal transformer.

Object-level Self-Distillation for Vision Pretraining Class-agnostic object detection with multi-modal transformer

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.430040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.883202Z digest=sha256:5ace17760256e81b000ddbcdd928f10d5bc3f3e3866af20b5043d40b6d849dd7

Observation 4b8de29f-a972-4d86-b301-0374c377f368 · outbound

This paper cites Exploring the Effectiveness of Object-Centric Representations in Visual Question Answering: Comparative Insights with Foundation Models.

Object-level Self-Distillation for Vision Pretraining Exploring the Effectiveness of Object-Centric Representations in Visual Question Answering: Comparative Insights with Foundation Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.887043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.887043Z digest=sha256:bf9da71df286147b2909cf80fced032a7352432b66317c8be75c45cd32177fab

Observation d30f897e-0ba8-4e5d-9208-a8839eadecd8 · outbound

This paper cites Unsupervised learning of dense visual representations.

Object-level Self-Distillation for Vision Pretraining Unsupervised learning of dense visual representations

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.416858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.891247Z digest=sha256:eff7f21a8aee8e68769a35bcc3a16e3416cdfd865ac3923e7fa7857f25e05e00

Observation 969274a8-d206-4e24-b87b-dc8e53d88812 · outbound

This paper cites Neural congealing: Aligning images to a joint semantic atlas.

Object-level Self-Distillation for Vision Pretraining Neural congealing: Aligning images to a joint semantic atlas

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.403336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.894863Z digest=sha256:a31a2e7320da2f55d5a079fae72e0f50a6b8a0a3cb772c135eadb29d52bd3ad8

Observation 861c41db-1346-4146-bde7-de33800adf3e · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Object-level Self-Distillation for Vision Pretraining DINOv2: Learning Robust Visual Features without Supervision

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.898506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.898506Z digest=sha256:5c2fe8160a3cf692dcdac27567dd295798d3b8212b9d1a2a337429e062f78b13

Observation 027c8501-67b8-4c3c-ad2b-f0a534165b98 · outbound

This paper cites Improving language understanding by generative pre-training.(2018), 2018.

Object-level Self-Distillation for Vision Pretraining Improving language understanding by generative pre-training.(2018), 2018

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.389838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.902139Z digest=sha256:659ac3e2711eec3cb5a48b741b7ed4860826379f2af7cc54a518af1f58e95d00

Observation 3b797630-d64c-402c-a2ca-98a1076a9a89 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Object-level Self-Distillation for Vision Pretraining Learning transferable visual models from natural language supervision

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.906427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.906427Z digest=sha256:b7138fe8bb1f496945d85ecd73e92651f0b88b30ed80b9f148bc79e54fd45956

Observation d89c2bf2-2718-4400-95ba-7d5a8618fc91 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Object-level Self-Distillation for Vision Pretraining SAM 2: Segment Anything in Images and Videos

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.910228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.910228Z digest=sha256:b9abcbc50acb1ce9d702904a0bdd80341331c9f63ffe4ce8504ef236cf027133

Observation 4cf3507e-0984-47a2-b9a9-dc5fd72da7f1 · outbound

This paper cites Do imagenet classifiers generalize to imagenet? In International conference on machine learning, pages 5389--5400.

Object-level Self-Distillation for Vision Pretraining Do imagenet classifiers generalize to imagenet? In International conference on machine learning, pages 5389--5400

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.914054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.914054Z digest=sha256:c5b6754269b6541746d56359425ed5d36542c3e584a727721c8fe1a2825f0659

Observation 7619d09c-7cee-4089-9e7a-f6e24c628ba5 · outbound

This paper cites You only look once: Unified, real-time object detection.

Object-level Self-Distillation for Vision Pretraining You only look once: Unified, real-time object detection

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.918254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.918254Z digest=sha256:d48fe1e54e45dfe09a9a8714e0413e5f68acb5172a0f8219a90cc9234d5dec90

Observation baf18923-4db5-43d0-af5b-28b6817b5366 · outbound

This paper cites Are We Done with Object-Centric Learning?.

Object-level Self-Distillation for Vision Pretraining Are We Done with Object-Centric Learning?

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.922097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.922097Z digest=sha256:2292146f62f9eec568107f124ec0a7dc37575e5d5f94faed6d7b4a70df7b8f12

Observation 1b30982e-6436-4d6c-927c-eb499cc00399 · outbound

This paper cites Bridging the Gap to Real-World Object-Centric Learning.

Object-level Self-Distillation for Vision Pretraining Bridging the Gap to Real-World Object-Centric Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.925946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.925946Z digest=sha256:9453775264fe5957338b88431440b1f56041a7ffa65a96a7b2d8f4d7c878c882

Observation 99228d01-3c03-42b2-b57d-d507dc5750ff · outbound

This paper cites Evaluating machine accuracy on imagenet.

Object-level Self-Distillation for Vision Pretraining Evaluating machine accuracy on imagenet

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.349341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.930160Z digest=sha256:ec64a81a350017a42820d60e6970930c3fc2f33a179429599223170fdb3a27fe

Observation c37011a0-5ce3-4742-8d9c-072e910e3825 · outbound

This paper cites Croc: Cross-view online clustering for dense visual representation learning.

Object-level Self-Distillation for Vision Pretraining Croc: Cross-view online clustering for dense visual representation learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.336102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.934008Z digest=sha256:11914624cb6b0d4300bcce729e74260f773998e8529a405106acf7a254818bbc

Observation e3432fdd-81b0-4249-ae5b-d1a5451d2af9 · outbound

This paper cites Convnets and imagenet beyond accuracy: Understanding mistakes and uncovering biases.

Object-level Self-Distillation for Vision Pretraining Convnets and imagenet beyond accuracy: Understanding mistakes and uncovering biases

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.322766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.937717Z digest=sha256:6bc4e557e51af6fcaac34b49bf08f9581db2542e21a89090dd708495abb36510

Observation 72da9c3d-4061-4759-9807-9c8d44cef805 · outbound

This paper cites Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results.

Object-level Self-Distillation for Vision Pretraining Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.941851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.941851Z digest=sha256:dadc32446bb596e0c6972a52325e9eba9cafe901c7b2f5490efc782b97b7a9ed

Observation 767d379c-3c0b-4cf9-a642-97cef43c6653 · outbound

This paper cites From imagenet to image classification: Contextualizing progress on benchmarks.

Object-level Self-Distillation for Vision Pretraining From imagenet to image classification: Contextualizing progress on benchmarks

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.299105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.945813Z digest=sha256:318d26bec9c3fb995bc1f6a04fe6a4acaadfa93e253d74a476a8a7dadd9a14fc

Observation 00c2909d-3221-4241-827c-b060f92fc604 · outbound

This paper cites Splicing vit features for semantic appearance transfer.

Object-level Self-Distillation for Vision Pretraining Splicing vit features for semantic appearance transfer

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.284466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.950041Z digest=sha256:945894d54bf6f01268bbe641dba3447f730ca11b84f407fe498d7af757346758

Observation 06994615-4d3a-46b6-b97e-381d17e14ad5 · outbound

This paper cites Dense contrastive learning for self-supervised visual pre-training.

Object-level Self-Distillation for Vision Pretraining Dense contrastive learning for self-supervised visual pre-training

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.271415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.953887Z digest=sha256:ebc19d043537d97f105302d2e9d5bfca68c99bc91817052725fc0327eb544a87

Observation 5b50c22f-d140-43ab-89bc-aa9cd0b8cb87 · outbound

This paper cites Self-supervised visual representation learning with semantic grouping.

Object-level Self-Distillation for Vision Pretraining Self-supervised visual representation learning with semantic grouping

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.258331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.957725Z digest=sha256:2ce47cb54dd5324ff179d117dabf65b1aafe6d899207754e9143fc0981334368

Observation 2b237aeb-0368-412d-b3c6-afe5ac14ecd9 · outbound

This paper cites Unsupervised object-level representation learning from scene images.

Object-level Self-Distillation for Vision Pretraining Unsupervised object-level representation learning from scene images

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.244587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.961699Z digest=sha256:83d386fed1340310dbcf164a9bafa5f299a0f8cf4292fc46f9b1f4a6b733f7f0

Observation b57b8659-4670-4432-ae26-c2d605449ecf · outbound

This paper cites Re-labeling imagenet: from single to multi-labels, from global to localized labels.

Object-level Self-Distillation for Vision Pretraining Re-labeling imagenet: from single to multi-labels, from global to localized labels

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:52:55.230722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:52:54.965391Z digest=sha256:9e4672de3a5eeb1061013d5ea32680f84d4c5ea7cdd0fa15933c12bffe79fa78

Observation 562eab54-3c33-493d-8da8-fa3795e212ab · outbound

This paper cites Scene parsing through ade20k dataset.

Object-level Self-Distillation for Vision Pretraining Scene parsing through ade20k dataset

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.969995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.969995Z digest=sha256:56b1339f08ff9a3fc27c6fd395eb6e72350209aeb130daa87534466467ca2e55

Observation 18ca66a9-4989-4c00-b86e-31483b866a09 · outbound

This paper cites iBOT: Image BERT Pre-Training with Online Tokenizer.

Object-level Self-Distillation for Vision Pretraining iBOT: Image BERT Pre-Training with Online Tokenizer

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T10:52:54.973978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:52:54.973978Z digest=sha256:565f2c8402ed653c23e9a6cf36ea2f819ea09686cd4ae8dff3fe6515995e033e

Pith citing papers

No inbound Pith citation observations are available.