Pith. sign in

Paper Citation Record · LEDGER

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance

As of 13 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2412.03871.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.03871 v2

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:06:00.353383Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

75 of 75 outbound references displayed

  • verified exact1
  • verified fuzzy30
  • unresolved44
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 91946815-836d-4d45-a606-ecccd9ecf1eb · outbound

This paper cites Contrastive learning of medical visual representations from paired images and text,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Contrastive learning of medical visual representations from paired images and text,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.090553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.090553Z digest=sha256:0887d203f0776ebdc0592434d61282545675bc409bf27541520bac5a10b11562

Observation 0aea3cd7-3cae-4250-8527-1152206ae3ee · outbound

This paper cites Learning transferable visual models from natural language supervision,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Learning transferable visual models from natural language supervision,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.095675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.095675Z digest=sha256:70e081363bb3ba7d353a12260cf512755580545ee234f85405e88072fd563fb6

Observation bdbee442-6410-4f1b-8cf0-00cf44c82418 · outbound

This paper cites Vlmo: Unified vision-language pre-training with mixture-of-modality-experts,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Vlmo: Unified vision-language pre-training with mixture-of-modality-experts,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:01.049391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.100573Z digest=sha256:2cf49b78f0b7fff3389bf3f943b6e57c7bf6f48238c5be13f9daef9ad4da8eca

Observation ad94a0b6-cecf-40c4-8f46-45737d4c9243 · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Scaling up visual and vision-language representation learning with noisy text supervision,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.104843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.104843Z digest=sha256:777a0a6220ec7168538ab056ffe1ea808bb124755fc155ff23003d03140edd57

Observation ca36181f-c706-41ec-8116-a6c34923517b · outbound

This paper cites ImageBERT: Cross-modal Pre-training with Large-scale Weak-supervised Image-Text Data.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance ImageBERT: Cross-modal Pre-training with Large-scale Weak-supervised Image-Text Data

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.109020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.109020Z digest=sha256:280b7f0a19800b36d1b0eb531f93a03860cb8407287035fdae12cc75ffe76ab9

Observation e7c8122b-57ed-4cea-843a-0ee1e67cef29 · outbound

This paper cites Reproducible scaling laws for contrastive language-image learning,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Reproducible scaling laws for contrastive language-image learning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:01.029833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.113671Z digest=sha256:a979c753e3ba520ecac01854edcf1f0a9f11ffb1391f55ce7df9bd207fd4d1e1

Observation e5845aab-7b1b-41e8-a85d-afd68ca91055 · outbound

This paper cites Scaling language- image pre-training via masking,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Scaling language- image pre-training via masking,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.117928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.117928Z digest=sha256:51bec2682b4b45a646782f8dc171ab955822e3405ce369f75d72160caff5144d

Observation 7ca70f1a-34ae-4d63-b0f3-1b88000e0dfe · outbound

This paper cites An inverse scaling law for clip training,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance An inverse scaling law for clip training,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:01.010717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.121503Z digest=sha256:d18cdb73587f7e82f0223d111bc6c7479fc4535ca35214f42c6f49e25fe5ca5d

Observation 8e6cf8b0-f751-40e8-b4c1-fb9357e61d26 · outbound

This paper cites Supervision exists everywhere: A data efficient contrastive language-image pre-training paradigm,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Supervision exists everywhere: A data efficient contrastive language-image pre-training paradigm,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.998963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.125049Z digest=sha256:efc2e996294067651935d713fb41c64ab5b640fff9aa79965ce47137860fcb41

Observation 6b4bbdfb-7b62-4e36-b559-e21730856a97 · outbound

This paper cites Too large; data reduction for vision-language pre-training,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Too large; data reduction for vision-language pre-training,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.988030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.130194Z digest=sha256:ac608ff444ee159dff16c2660bd3b9d52ad92b1ca476d50339a7946248b83a59

Observation 3e8eba59-80c7-4892-879a-b7070098a7dc · outbound

This paper cites Slip: Self-supervision meets language-image pre-training,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Slip: Self-supervision meets language-image pre-training,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.977198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.134764Z digest=sha256:23b9ed6d9b922c8a8b5544301b19e7d4a7e23b796504bfa774ba6defc2e06bfa

Observation e8482758-cd0d-4f74-9b3f-87a6816d2436 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.138881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.138881Z digest=sha256:9caa5af1e384d2c224c98be261f7fa64dd32137a9455d1c5544f1550451ad9ef

Observation 88ca24f7-999f-4835-b6c1-ac0fa8ddc1f8 · outbound

This paper cites Microsoft coco: Common objects in context,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Microsoft coco: Common objects in context,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.142872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.142872Z digest=sha256:3d48d17d9503d6b558398e66aeb5cbc8c1469f844a17776e684113b5bed13d42

Observation 20bf84cc-3960-46ff-b846-cb466ae9e1dc · outbound

This paper cites Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image cap- tioning,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image cap- tioning,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.147060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.147060Z digest=sha256:0d7b5f837a5cbcaaba52b6a81accb4eab8804093ae2252f85a247de33f935f4e

Observation 89f84f21-96d0-47ff-9ca2-a5aa5a72519c · outbound

This paper cites Image as a foreign language: Beit pretraining for vision and vision-language tasks,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Image as a foreign language: Beit pretraining for vision and vision-language tasks,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.941175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.150607Z digest=sha256:e52d106ffae4d919f2c1a2e51ec968fcdc20c463e30a5ee69d8881c1e785d3f1

Observation 236d816a-d82f-4d93-9908-254da5e0ef32 · outbound

This paper cites The crucial role of data collection in research: Techniques, challenges, and best practices,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance The crucial role of data collection in research: Techniques, challenges, and best practices,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.924510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.154213Z digest=sha256:507c12eeb35a472046d9fb56ed3393fc682ebc406561e4237c057105e9f3f84a

Observation 8288412d-19ea-4192-9e22-fd875dd932ea · outbound

This paper cites Tinyclip: Clip distillation via affinity mimicking and weight inheritance,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Tinyclip: Clip distillation via affinity mimicking and weight inheritance,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.911260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.159558Z digest=sha256:f241aeae05bdebfd4bd0bde762466132ea701e6e9c2889d7f90e05671dfd2928

Observation 512d4a61-f85b-45fc-934b-32f5880b9690 · outbound

This paper cites Clip-kd: An empirical study of clip model distillation,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Clip-kd: An empirical study of clip model distillation,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.163378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.163378Z digest=sha256:7b12554382b64a30295dd6603e5ca550939935d068521665ccee1bb922a20be9

Observation 4d795658-3725-4daa-b2ad-9e4431d2f4cd · outbound

This paper cites Mobileclip: Fast image-text models through multi-modal reinforced training,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Mobileclip: Fast image-text models through multi-modal reinforced training,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.167013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.167013Z digest=sha256:c1f46323d72c9eb2a4525f743fa329a31cdf871ef0a89eb7c8d8505f99f9869a

Observation abced0f7-245b-4d67-89f8-3de1d07393e5 · outbound

This paper cites ComKD-CLIP: Comprehensive Knowledge Distillation for Contrastive Language-Image Pre-traning Model.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance ComKD-CLIP: Comprehensive Knowledge Distillation for Contrastive Language-Image Pre-traning Model

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-11T22:06:00.470348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.170818Z digest=sha256:2ae91976274afa14e3743f3e65ac29ed1168b029cac5f4c21cd39338c99bd516

Observation 11b0426b-1949-4225-a587-212e3adc196a · outbound

This paper cites CLIP-CID: Efficient CLIP Distillation via Cluster-Instance Discrimination.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance CLIP-CID: Efficient CLIP Distillation via Cluster-Instance Discrimination

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.174773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.174773Z digest=sha256:063dd7bc79b107750e1ed52f2bc36beba2585f3d1f665632d84e9e5cdd4f5ec1

Observation 98bb1143-be61-4147-842b-fc4dcc342547 · outbound

This paper cites Module-wise adaptive distillation for multimodality foun- dation models,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Module-wise adaptive distillation for multimodality foun- dation models,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.880132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.179242Z digest=sha256:e6c3597112f2a3a2025eb0202e0ee9eefaa4d17ea92b9aea4a9d17e031bbf4a4

Observation 64a7479b-8a93-41b1-b089-879819650dfe · outbound

This paper cites Self-supervised co-training for video representation learning,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Self-supervised co-training for video representation learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.867734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.182919Z digest=sha256:c21ab204f06085dab888b194269c0a11be443ee2f5a5fa6f8238502c57693437

Observation a58589fe-c8c8-4090-a7e6-35ce51f73929 · outbound

This paper cites Improving generalization via scalable neighborhood component analysis,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Improving generalization via scalable neighborhood component analysis,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.854101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.186788Z digest=sha256:3fe81baa6344f2509d2ddf3871a3e236295cd4ae7d5f71e23384218d10459304

Observation 94da781f-6637-4cea-8bb5-9e987f1f8841 · outbound

This paper cites With a little help from my friends: Nearest-neighbor contrastive learning of visual representations,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance With a little help from my friends: Nearest-neighbor contrastive learning of visual representations,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.190524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.190524Z digest=sha256:3682c55f1ba74dbe42736eb45d36cf50d0daf41c426218dcb72255b6f6bc9973

Observation 1190b91c-0203-492d-ad08-ee158f2b05a1 · outbound

This paper cites Promoting semantic connectivity: Dual nearest neighbors contrastive learning for unsupervised domain generalization,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Promoting semantic connectivity: Dual nearest neighbors contrastive learning for unsupervised domain generalization,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.833909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.193769Z digest=sha256:f0c065188da72cc455a96032fea1077c78c7bca220bf091933996c00d276a2e2

Observation 026e746f-b4b5-41f5-aeef-8fa9ecbd9c5b · outbound

This paper cites Imagenet: A large-scale hierarchical image database,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Imagenet: A large-scale hierarchical image database,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.197171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.197171Z digest=sha256:6417606153ed59bbf996a7920f4bf786b02ac267c26fbd2dd2f52e5e1c61552c

Observation c06e4be7-ac28-48c5-8350-0831648a41c9 · outbound

This paper cites From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.200851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.200851Z digest=sha256:c7a7141660f1dd56daf47d47f477375543e4fcfe0d59921d4ffa83cba79dc80b

Observation 4a203864-f0f4-446b-8b9f-f7c74ca00559 · outbound

This paper cites Vision-language models for vision tasks: A survey,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Vision-language models for vision tasks: A survey,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.204377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.204377Z digest=sha256:ae6d5eeb028a87326464593bd561ab4158984a2ff8cce20df5606f64844466e7

Observation d1db178b-18c5-4aab-af74-5a9d5e858327 · outbound

This paper cites A Survey of Vision-Language Pre-Trained Models.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance A Survey of Vision-Language Pre-Trained Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.207800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.207800Z digest=sha256:3ecb3968fe467f04ed31e643628b037192ee9dc989de0c3ee9f0de9b991fcce4

Observation f8bb8222-28f2-4be0-a75b-cea192dc845b · outbound

This paper cites Vlp: A survey on vision-language pre-training,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Vlp: A survey on vision-language pre-training,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.211525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.211525Z digest=sha256:71846f7b4b5073dee2a79e8e4d0ca60d7bce8e977ed483714b8009c2807000ab

Observation 90cd535b-2824-4fe1-ad29-2685a4b6ed02 · outbound

This paper cites Self- supervised learning of visual features through embedding images into text topic spaces,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Self- supervised learning of visual features through embedding images into text topic spaces,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.784451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.214468Z digest=sha256:6a9f7bd58a8566b1d6d055b4427f59548176c8a8b1aa75e8b8448fd2d3e8e8ce

Observation ac14d9e5-3188-4772-9bff-50f865e083a9 · outbound

This paper cites Beyond instance-level image retrieval: Lever- aging captions to learn a global visual representation for semantic retrieval,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Beyond instance-level image retrieval: Lever- aging captions to learn a global visual representation for semantic retrieval,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.771795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.217478Z digest=sha256:b0260c633c4d8963cd1c3317e0d37a9be578700bba4cd11a6fbb2f4f55630793

Observation 2b30d2c5-aa1d-4108-b940-11841b3b8923 · outbound

This paper cites Learning visual n-grams from web data,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Learning visual n-grams from web data,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.759054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.220250Z digest=sha256:4de244a077d414a2ef4ae1ee9ad53e6279515d0e24726a25ea668e95a98b806c

Observation 0b009b12-0821-4b51-aea5-1e57736515ff · outbound

This paper cites Virtex: Learning visual representations from textual annotations,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Virtex: Learning visual representations from textual annotations,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.223160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.223160Z digest=sha256:43929c3987c550f178072615be431dd1c1dc0e51e7598053efc72eb2920cee54

Observation 7f330f13-f532-4857-a9cc-48442ff3b788 · outbound

This paper cites Learning visual representa- tions with caption annotations,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Learning visual representa- tions with caption annotations,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.226288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.226288Z digest=sha256:5f4875d5ab62153e2b16bf21291070d5669faec89d7f93f2f2ec5335824b907e

Observation 925d989d-082b-43e3-b996-21cc8e9798e9 · outbound

This paper cites Combined scaling for zero-shot transfer learning,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Combined scaling for zero-shot transfer learning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.724957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.229123Z digest=sha256:ff21f5de619ebacd759726a80d79d8d67042ae31eeb68b4577633593107263d4

Observation 156132af-75b2-4825-9262-f41f3e8c6ad6 · outbound

This paper cites SimVLM: Simple visual language model pretraining with weak supervision,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance SimVLM: Simple visual language model pretraining with weak supervision,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.714698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.232081Z digest=sha256:e453c1db6806f38610c4f0ca76be625451d10b3fbb43a904950c60e8c43a7a6e

Observation b52c001c-90ff-4127-9a22-83093896cc27 · outbound

This paper cites Florence: A New Foundation Model for Computer Vision.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Florence: A New Foundation Model for Computer Vision

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.234692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.234692Z digest=sha256:cebc9316fcd3879ece97315ac4954401f5e93c6be4d511216f2fd888aef359fd

Observation 063d9e71-4045-4488-9ea2-588180f339e7 · outbound

This paper cites Lit: Zero-shot transfer with locked-image text tuning,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Lit: Zero-shot transfer with locked-image text tuning,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.704730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.238602Z digest=sha256:335e4f488d360863e80f2d35871b2cd5fc998fcd49335b76c26474787af1e29e

Observation a5d540bb-1243-4d9e-bab9-45b2237b970a · outbound

This paper cites Compressing visual-linguistic model via knowledge distillation,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Compressing visual-linguistic model via knowledge distillation,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.693660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.241819Z digest=sha256:969076cbf30894d96a064974149edd08ec3663873919f75308a752d302ff93f2

Observation 3ca00690-42f1-4dff-b713-28007b1779d7 · outbound

This paper cites Distilling large vision-language model with out-of-distribution generalizability,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Distilling large vision-language model with out-of-distribution generalizability,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.683490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.245036Z digest=sha256:3f5229c0814a2bc5b9df90240b822b0d9f1a1725e9dc9ab9c518006ad6857849

Observation d168bf67-5684-46a2-901a-3d40d0ec970d · outbound

This paper cites Multimodal Adaptive Distillation for Leveraging Unimodal Encoders for Vision-Language Tasks.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Multimodal Adaptive Distillation for Leveraging Unimodal Encoders for Vision-Language Tasks

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.248566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.248566Z digest=sha256:a55d67d94fd8e2b5e79d4b37d55e66f6b3085e894c65cdc457ca0e0afc9d5225

Observation 96d6f81d-acd0-499d-a518-27fd20d71bc7 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Representation Learning with Contrastive Predictive Coding

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.252477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.252477Z digest=sha256:001be2dde08ecb0163accd85e139b9077ebad948031996e926977edeebf22121

Observation 38f597db-d92e-4f1d-8958-428760d598e7 · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Pytorch: An imperative style, high-performance deep learning library,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.256071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.256071Z digest=sha256:d20f1b537a8f5b29df265f76d725515a37cd182ec99434073a2def3062fa12eb

Observation f4209112-e2f7-4097-bd4b-46c69d80807a · outbound

This paper cites Pytorch image models,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Pytorch image models,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.259290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.259290Z digest=sha256:b3535221141aa3b88eb398671dd6c2c59118aaf306a8b99bb6f4c3ef454e0f93

Observation 3a61defe-dde8-4b7c-ba72-7e3c30db0a81 · outbound

This paper cites MobileBERT: a compact task-agnostic BERT for resource-limited devices,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance MobileBERT: a compact task-agnostic BERT for resource-limited devices,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.661894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.262663Z digest=sha256:0055b181aafd9a781d1439cd85b2a744a25ebd0eb3e4a3796e1565bfa8579afc

Observation 6e400ef1-71af-4e03-a2d6-8db114ad7506 · outbound

This paper cites A convnet for the 2020s,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance A convnet for the 2020s,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.266138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.266138Z digest=sha256:ce14b74e71c905668063532849cea19751cf8e1e3b7543d44a3c678be65d2c63

Observation 1fb58ed0-7f95-4fb1-af0a-cd1c801b0c5e · outbound

This paper cites MobileNetV4 -- Universal Models for the Mobile Ecosystem.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance MobileNetV4 -- Universal Models for the Mobile Ecosystem

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.269384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.269384Z digest=sha256:1fccc47faf0107e588a9fff0ecd8ca789aea161634e526100347d589a850d625

Observation b26d086a-b239-4fbb-a27d-2f4f475085da · outbound

This paper cites Big transfer (bit): General visual representation learning,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Big transfer (bit): General visual representation learning,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.645346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.273076Z digest=sha256:7e1d12598e5ddfb1ec517f6a2fcec0eedba6d81998fd481e55e8045e52252253

Observation 01984534-8d08-4e14-a204-b54f5b76e920 · outbound

This paper cites Identity mappings in deep residual networks,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Identity mappings in deep residual networks,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.276429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.276429Z digest=sha256:3b9840066836ffdd372a29d8e6779f15aa98eb69395749eac2c05468d92c5773

Observation 53c711ba-8155-49c4-9ca8-549a181a1cf7 · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Bert: Pre-training of deep bidirectional transformers for language understanding,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.279965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.279965Z digest=sha256:fa13c5023179eef2436e9ded7b4fef915ab623a8d0c152a2e1655871a59d00ed

Observation 9528bc2e-a8cb-414b-99a2-83d8f1c54636 · outbound

This paper cites Decoupled weight decay regularization,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Decoupled weight decay regularization,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.283293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.283293Z digest=sha256:e9f85455bc574d51b541e3a467023e5044ce9a091fa00f2b212278021730fc4a

Observation 037f5969-2b5b-4763-bf50-9e497fe88522 · outbound

This paper cites Algorithm 799: revolve: an implementa- tion of checkpointing for the reverse or adjoint mode of computational differentiation,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Algorithm 799: revolve: an implementa- tion of checkpointing for the reverse or adjoint mode of computational differentiation,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.619079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.286716Z digest=sha256:5ce35b92edde903073107cc10b67114f6d940cd576dc2c83f2c8f946a2e4d1b6

Observation 49120462-847a-42d0-8b3d-9e0d85a1b312 · outbound

This paper cites Training Deep Nets with Sublinear Memory Cost.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Training Deep Nets with Sublinear Memory Cost

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.289929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.289929Z digest=sha256:78024f5ac582434df1069e46207b2275151f0f7b5f8346fb8453d1334fd64e7a

Observation dd24f8c8-fb05-4cb6-b22d-7cc28335a625 · outbound

This paper cites Mixed precision training,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Mixed precision training,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.608386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.293603Z digest=sha256:c1905dc307cd50d6689698202a7836fa34e8309209736c12dfd05f56a85cfcae

Observation cc95b177-27c6-41dc-b72e-b3cf8d6e0854 · outbound

This paper cites An analysis of single-layer networks in unsupervised feature learning,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance An analysis of single-layer networks in unsupervised feature learning,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.296954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.296954Z digest=sha256:526967d43e0e8f8f91a7b454181339307f034e7a91bd594d34cf728868512d66

Observation c7bee03e-4e57-4228-983c-ec90e8a1a218 · outbound

This paper cites Learning multiple layers of features from tiny images,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Learning multiple layers of features from tiny images,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.591862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.300463Z digest=sha256:195125802e018a6f15faf468dd33a40fc71d8f4c25616f8eecf8030c83345165

Observation fd03d16d-6366-4ec8-aa8f-172ff09cee95 · outbound

This paper cites Human action recognition by learning bases of action attributes and 14 parts,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Human action recognition by learning bases of action attributes and 14 parts,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.581109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.303985Z digest=sha256:8ecbd4540771264a05ff4278c4cceef13272bc506376690eaaf2dca75e708ad5

Observation a80240a3-fbb7-4a3f-93cf-be9ac967d8a6 · outbound

This paper cites Do imagenet clas- sifiers generalize to imagenet?.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Do imagenet clas- sifiers generalize to imagenet?

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.571263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.307645Z digest=sha256:d5b7a4106c86d600a86ae12776d8f2cf54f947719a813473a8a25c5d6304fba1

Observation aefd8b45-d5df-49b4-a8e9-d8c3b99f703a · outbound

This paper cites The many faces of robustness: A critical analysis of out-of-distribution generalization,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance The many faces of robustness: A critical analysis of out-of-distribution generalization,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.310953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.310953Z digest=sha256:0fcc1fa5ec36056915ab540dd8e87439d6c3b6e5309950e96869cc45891b8f66

Observation 539b9788-7af3-40cf-b5e9-16ba2ea8a00c · outbound

This paper cites Natural adversarial examples,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Natural adversarial examples,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.314256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.314256Z digest=sha256:64fed08d0c836385c9231b0cea45655da2af8e0a0b9805f188cbcedb2c298f80

Observation 3e797959-6474-4019-ab3c-9e05e8030a02 · outbound

This paper cites Learning robust global representations by penalizing local predictive power,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Learning robust global representations by penalizing local predictive power,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.549967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.317590Z digest=sha256:92eaf6b259fbb35f889ae510d83b1e22951591967befdb5c5cfac4d1b1a25dc6

Observation f2ac2cc0-e8e2-44a5-81a3-e69ae3f039d5 · outbound

This paper cites Cats and dogs,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Cats and dogs,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.320891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.320891Z digest=sha256:305bda8bef25a650a0fe0edca4674c655189f54bbd612b7eee438edf8cbe49aa

Observation a1449af8-2a7c-4dec-87dd-f55bcb7359df · outbound

This paper cites Learning generative visual models from few training examples: An incremental bayesian approach tested on 101 object categories,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Learning generative visual models from few training examples: An incremental bayesian approach tested on 101 object categories,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.323621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.323621Z digest=sha256:c23f195373215b89a398c504af1e771cea2377fa349b42b1d19141fd41bc2e89

Observation 29563316-e331-4aa3-80d8-a946108cc0d1 · outbound

This paper cites Automated flower classification over a large number of classes,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Automated flower classification over a large number of classes,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.326785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.326785Z digest=sha256:7208cf25db3a4b98823eb6ff1974dce3e08de2857391fb9b947727a4851ffa95

Observation 098783a8-9ee8-4094-be52-00cd110e5e21 · outbound

This paper cites Fine-Grained Visual Classification of Aircraft.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Fine-Grained Visual Classification of Aircraft

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.329530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.329530Z digest=sha256:aa06f289ccbe40039402095f3e4348d34b7527183137d41db610b29fbd0793fd

Observation 2ba57d51-3a69-41aa-8b41-9e61daa0cb44 · outbound

This paper cites Food-101–mining discriminative components with random forests,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Food-101–mining discriminative components with random forests,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.333125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.333125Z digest=sha256:f32d557b03e5e3e57174707848d3f00489241793370378335207e8ca2d59779a

Observation c08a64fc-c9f0-4911-a6e4-b56e013a7d3f · outbound

This paper cites Describing textures in the wild,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Describing textures in the wild,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.335701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.335701Z digest=sha256:5edce460a4331bcfa723197c065f17a89748940a4765237d5f8256a3d6241bfd

Observation 52fe8b59-0b0e-4a61-a699-08119d0c2d7e · outbound

This paper cites Sun database: Large-scale scene recognition from abbey to zoo,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Sun database: Large-scale scene recognition from abbey to zoo,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.338608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.338608Z digest=sha256:bf991043fb843f970826ed2508985ee7219b0b7b101d967c5b10ae651561e525

Observation 4cc371d4-12f0-41b8-b9c9-5bb66dc9ef3f · outbound

This paper cites Collecting a large-scale dataset of fine-grained cars,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Collecting a large-scale dataset of fine-grained cars,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.341399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.341399Z digest=sha256:075f99fd680b3e4381228f5d29a65b9b8d9a4c636e2fa27bf9a650aa0c70213f

Observation de58c476-3a39-4ef7-a0a4-50997fcf47cc · outbound

This paper cites 3d object representations for fine-grained categorization,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance 3d object representations for fine-grained categorization,

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.344080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.344080Z digest=sha256:27d7c133f1bf81450683da1e69d8109c3e638337d5609d2fa50a1b40620f7b44

Observation 51ec20a6-c923-4b42-b048-a8bdd9703bae · outbound

This paper cites Adam: A Method for Stochastic Optimization.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Adam: A Method for Stochastic Optimization

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.347067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.347067Z digest=sha256:b503eb10fc3e41f5f3c6c83add22e57ddd3112157e38a5c6e2a77134ce01795e

Observation 470a26e8-36c3-46ed-a93a-cf747da79d27 · outbound

This paper cites Knowledge distillation: A good teacher is patient and consistent,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance Knowledge distillation: A good teacher is patient and consistent,

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T22:06:00.350128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:06:00.350128Z digest=sha256:712523180bf41069ef401593e3758315f3151d97f26eda68d25e9669bc120944

Observation c0a7247d-e83d-4c79-ab06-b97a3876c7fa · outbound

This paper cites The efficiency misnomer,.

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance The efficiency misnomer,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:06:00.490177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:06:00.353383Z digest=sha256:5df909be68a588bed45e909015317d620b95761653fa3d734b8d568c665e6fac

Pith citing papers

No inbound Pith citation observations are available.