Pith. sign in

Paper Citation Record · LEDGER

On the rankability of visual embeddings

As of 19 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 0 inbound Pith citation observations for arXiv:2507.03683.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.03683 v1

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:11:51.717772Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

73 of 73 outbound references displayed

  • verified exact11
  • verified fuzzy29
  • unresolved29
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 72fcb7d1-1878-4a5b-987a-7a8956bbd1b7 · outbound

This paper cites Understanding intermediate layers using linear classifier probes.

On the rankability of visual embeddings Understanding intermediate layers using linear classifier probes

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:48.623426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:48.623426Z digest=sha256:2571c941ab3374b9770b84a46f8669aa38e6ee2e07121592a889880b3d1a7750

Observation 72f8616e-862e-4547-92c6-4f8ae4d82f13 · outbound

This paper cites Conditioned and composed image retrieval combining and partially fine-tuning CLIP-based features.

On the rankability of visual embeddings Conditioned and composed image retrieval combining and partially fine-tuning CLIP-based features

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.679728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:48.662843Z digest=sha256:d66f5e8526e6ee875cc8471283fbdb12416e11f3ba310d364432b6a8eeb3d7cc

Observation 671d080d-4f91-48ef-b91b-bc3ff30dd0f1 · outbound

This paper cites Not Only Text: Exploring Compositionality of Visual Representations in Vision-Language Models.

On the rankability of visual embeddings Not Only Text: Exploring Compositionality of Visual Representations in Vision-Language Models

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:11:53.140579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:48.742002Z digest=sha256:6790b4b7758f1b7c30963f3087d765eb33e9d6c7fcdb1b26e9caf9fa7d42f781

Observation 82b92998-93e2-45a5-9a65-d480c670884b · outbound

This paper cites A Simple Framework for Contrastive Learning of Visual Representations.

On the rankability of visual embeddings A Simple Framework for Contrastive Learning of Visual Representations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.663979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:48.803202Z digest=sha256:801b57a3d4df0df1f8e8de6254f8c9f2f03a2b8b3e58a0856ebe1078451ff7a5

Observation 511cf3cc-d07f-4c12-a0ae-b1d11993093b · outbound

This paper cites Deep Learning for Instance Retrieval: A Survey.

On the rankability of visual embeddings Deep Learning for Instance Retrieval: A Survey

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.647220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:48.894792Z digest=sha256:6b8a496f884a44d44aed3dd6ed490c4ba56747ed5d5a7e01b839a1fa1a1a11ef

Observation 0208f5c9-189e-4893-a53b-1e5093e7b4a4 · outbound

This paper cites Deep learning for instance retrieval: A survey.

On the rankability of visual embeddings Deep learning for instance retrieval: A survey

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.630656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:48.981497Z digest=sha256:be270383e4a6ab9ab61a46b56d93908653ee67b9f070ffbea6c351db45a122c3

Observation af97d169-f7c6-4cc0-8ed5-efa72c98a39c · outbound

This paper cites Composition Loss for Counting, Density Map Estimation and Localization in Dense Crowds.

On the rankability of visual embeddings Composition Loss for Counting, Density Map Estimation and Localization in Dense Crowds

Reference 7

Resolution
malformed identifier
no resolver link, observed 2026-08-06T20:11:49.072090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:49.072090Z digest=sha256:cbca4ba98c7d1d68c21306ef59b95e177403550c7febd6af275d720858869ed2

Observation 94a871ac-e6e3-41c2-b3b3-06ddd5a78bde · outbound

This paper cites Hyperbolic Image-Text Representations.

On the rankability of visual embeddings Hyperbolic Image-Text Representations

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:11:53.112878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:49.128084Z digest=sha256:8f8c71c83ce9836894154b3264248b0f2ba13e9218a15cd7b6e8d2e1b0249b49

Observation 2d825a8c-c3ea-4a01-a78d-877669bf5ce3 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

On the rankability of visual embeddings An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:49.229769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:49.229769Z digest=sha256:bba8feefd46f03957f56fb8e25916fa6067a83e4a9a4e10bc1b30ceeee711234

Observation b6cc826c-934f-489b-94e1-f44b9422ddc1 · outbound

This paper cites Teach CLIP to Develop a Number Sense for Ordinal Regression.

On the rankability of visual embeddings Teach CLIP to Develop a Number Sense for Ordinal Regression

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:11:52.059438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:49.329152Z digest=sha256:bd46bfcddaad1114035ca3b0c7b361c8d4f0b071348c1702dde02c9ce3b3e9d8

Observation 8883acf4-b68b-4343-903e-8706086d6021 · outbound

This paper cites Age and Gender Estimation of Unfiltered Faces.

On the rankability of visual embeddings Age and Gender Estimation of Unfiltered Faces

Reference 11

Resolution
verified exact
raw_fallback, observed 2026-08-06T20:11:53.088420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:49.450137Z digest=sha256:56af4d45289da0d66e9ca7fb79839c104e1019bdf6956ac4d9b30c167334b525

Observation 9cfaf071-0d3c-4582-83b4-b5dd8d280749 · outbound

This paper cites It's Not a Modality Gap: Characterizing and Addressing the Contrastive Gap.

On the rankability of visual embeddings It's Not a Modality Gap: Characterizing and Addressing the Contrastive Gap

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:49.531103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:49.531103Z digest=sha256:568d383f07a10e44b96973b95d608d4446769b0902b36bf4b433687cb2a4f5e4

Observation 9dd7610a-3d37-4134-a268-e1343c54c0f3 · outbound

This paper cites Heterogeneous face attribute estimation: A deep multi-task learning approach.

On the rankability of visual embeddings Heterogeneous face attribute estimation: A deep multi-task learning approach

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.613502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:49.721697Z digest=sha256:1ccc5c10675d77caf89d5238e9ce4fb1ecd1be68fb8e678f08d7e20da4179800

Observation 477da1c2-50ff-4f16-8b1d-05c3201567d5 · outbound

This paper cites Deep Residual Learning for Image Recognition.

On the rankability of visual embeddings Deep Residual Learning for Image Recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:49.805256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:49.805256Z digest=sha256:95a67a1f001354fbc4d9ab786e4cc4e28d85b199be8b8a4e66be01a8ae8da4fe

Observation 307ea798-7ea6-4281-aa6e-2432bac0103f · outbound

This paper cites Multilayer feedforward networks are universal approximators.

On the rankability of visual embeddings Multilayer feedforward networks are universal approximators

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.597532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:49.928316Z digest=sha256:55689d97d37adbd27eacaea9c17baebb00f45197ee77a2d20de5616dabc7c0fd

Observation bb2b5fda-2abe-47f1-b5e9-2c4c8f7635f0 · outbound

This paper cites KonIQ-10k: An Ecologically Valid Database for Deep Learning of Blind Image Quality Assessment.

On the rankability of visual embeddings KonIQ-10k: An Ecologically Valid Database for Deep Learning of Blind Image Quality Assessment

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.582027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:50.055806Z digest=sha256:d269970978dce017c9ceb9cfba6d59633838c27ca5c2a2ed98a40cb031852e3a

Observation 97fb22fc-96d4-4760-9b18-e002309f4483 · outbound

This paper cites Lp++: A surprisingly strong linear probe for few-shot clip.

On the rankability of visual embeddings Lp++: A surprisingly strong linear probe for few-shot clip

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.566871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:50.309455Z digest=sha256:6af9c6247d7e95bf768b414f030bef79981250479bd902ee815aa238fb5b43aa

Observation 9d9125c5-7393-4fdd-90fd-9213f230accf · outbound

This paper cites The Platonic Representation Hypothesis.

On the rankability of visual embeddings The Platonic Representation Hypothesis

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:50.427501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:50.427501Z digest=sha256:e942eb59ee612a7313a0a55f01e182426877a37391b5e23502e035acc819ea43

Observation 1e02910c-73b7-4adf-a889-8aa29d70c8f0 · outbound

This paper cites CLIP-Count: Towards Text-Guided Zero- Shot Object Counting.

On the rankability of visual embeddings CLIP-Count: Towards Text-Guided Zero- Shot Object Counting

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:50.544893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:50.544893Z digest=sha256:6d5ab4a9becbef76e70da1c79a4cf2eefff18ab6010f7d4994953b1ff543e636

Observation 22adfc6d-f764-4110-8eea-9146e82b38b7 · outbound

This paper cites Hyperbolic Image Embeddings.

On the rankability of visual embeddings Hyperbolic Image Embeddings

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.549000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:50.625607Z digest=sha256:8812607e9c51a4ce17618c569886c22175690225dadb1dcacc7d28987860ac8d

Observation 86d92a71-4aea-4f83-ad09-41ba0cd792a0 · outbound

This paper cites MLLM-CompBench: A Comparative Reasoning Benchmark for Multimodal LLMs.

On the rankability of visual embeddings MLLM-CompBench: A Comparative Reasoning Benchmark for Multimodal LLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:50.693133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:50.693133Z digest=sha256:0a728d558b378bf9de6b52535044a85c287fb0e7fd5dd8913ca57d496b0114ed

Observation 4c0d978c-48c2-461d-9110-e11b77b6153b · outbound

This paper cites Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav).

On the rankability of visual embeddings Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav)

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.529295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:50.813651Z digest=sha256:507ab6ce27b0a0c3a3fc130c7837008a290f30d4a175867bd61be542d3150d9e

Observation d74f666a-fe1f-4221-be11-5cb58b0a6fd4 · outbound

This paper cites CLIP Behaves like a Bag-of-Words Model Cross-modally but Not Uni-modally.

On the rankability of visual embeddings CLIP Behaves like a Bag-of-Words Model Cross-modally but Not Uni-modally

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:50.960886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:50.960886Z digest=sha256:3730222c44a9c8b2b9931039cc2ce26e1fb5844f451041d0190ac8f6ea299ffc

Observation de4d0571-cc07-4331-a0c2-034a6b53f7b1 · outbound

This paper cites CLIP Behaves like a Bag-of-Words Model Cross-modally but not Uni-modally.

On the rankability of visual embeddings CLIP Behaves like a Bag-of-Words Model Cross-modally but not Uni-modally

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.002032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.002032Z digest=sha256:7373c87ab5e1ba196de6aa8e53efab14ee503d5df0d2c6014540c906b04e1b3c

Observation d160e3ef-edc5-4886-b0e0-5c3704ab4ed2 · outbound

This paper cites Beyond a Pre-Trained Object Detector: Cross-Modal Textual and Visual Context for Image Captioning.

On the rankability of visual embeddings Beyond a Pre-Trained Object Detector: Cross-Modal Textual and Visual Context for Image Captioning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.512741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.033881Z digest=sha256:a9c3de40fdfe61de45d82c5d030f0271115f647d3864c035e121320817d0dcc5

Observation c5ada6c4-84e0-41ad-b510-1f10d8b09f98 · outbound

This paper cites MiVOLO: Multi-input Transformer for Age and Gender Estimation.

On the rankability of visual embeddings MiVOLO: Multi-input Transformer for Age and Gender Estimation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.039704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.039704Z digest=sha256:b18ebf50e1b8fd5c814a91a4087b2cada2641c38ef5c9fd0a3f005aee106771f

Observation 9e393176-8412-4a5a-816e-ea283721eaac · outbound

This paper cites The Double-Ellipsoid Geometry of CLIP.

On the rankability of visual embeddings The Double-Ellipsoid Geometry of CLIP

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.045027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.045027Z digest=sha256:9ad84b772ef0f1a6892c61d4e7d6e888bba0be4c97a0e114257d710b558bf43a

Observation 4d502446-6440-47de-b3a7-37f7735ac5f9 · outbound

This paper cites Does CLIP Bind Concepts? Probing Compositionality in Large Image Models.

On the rankability of visual embeddings Does CLIP Bind Concepts? Probing Compositionality in Large Image Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.071811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.071811Z digest=sha256:3f45caf34f37ee601325afc8145e4feba3908c8cbf850cc5d214a3be2116f1be

Observation 2624fd97-d6cf-4533-b4f8-02e1e764d3dd · outbound

This paper cites Align before fuse: Vision and language representation learning with mo- mentum distillation.

On the rankability of visual embeddings Align before fuse: Vision and language representation learning with mo- mentum distillation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.494930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.171974Z digest=sha256:02092e275337910c333174113a9b0b6bba94e46595b7f56daafba478a8a69c38

Observation bf413c30-217d-4123-95e4-6fe4efb58ffe · outbound

This paper cites BLIP: Bootstrapping Language-Image Pre-training for Unified Vision- Language Understanding and Generation.

On the rankability of visual embeddings BLIP: Bootstrapping Language-Image Pre-training for Unified Vision- Language Understanding and Generation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.478419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.232772Z digest=sha256:bc22f109f4fd8e007ead04f8e26027747051bb004aae70b2533b627b96a156ea

Observation f366b8f4-65fd-4e5b-8ae5-52c09621ca3b · outbound

This paper cites OrdinalCLIP: Learning Rank Prompts for Language-Guided Ordinal Regression.

On the rankability of visual embeddings OrdinalCLIP: Learning Rank Prompts for Language-Guided Ordinal Regression

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:11:51.956760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.302801Z digest=sha256:3d8b6b072e9f7f55695a9518d94d1389fa8dc0adbcd996520e807a5dc37d234e

Observation 00efac1b-a70d-4467-b8f4-0ba5c81b0c6d · outbound

This paper cites CrowdCLIP: Unsupervised Crowd Counting via Vision-Language Model.

On the rankability of visual embeddings CrowdCLIP: Unsupervised Crowd Counting via Vision-Language Model

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:11:51.933505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.386439Z digest=sha256:4a7e6d1ea6d06e75885e2f61f2f05723be6356202adf565f02a15f343431fcaf

Observation ffb1b424-5803-447c-983e-274b6ec42009 · outbound

This paper cites Beyond comparing image pairs: Setwise active learning for relative attributes.

On the rankability of visual embeddings Beyond comparing image pairs: Setwise active learning for relative attributes

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.462867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.467629Z digest=sha256:2ab5f323748002b02763472306fdd77231cb2957db628fccc3a6909762bbf323

Observation 9cdd71ea-a5ce-4033-93fd-6d02518fd554 · outbound

This paper cites CLIP-EBC: CLIP Can Count Accurately through Enhanced Blockwise Classification.

On the rankability of visual embeddings CLIP-EBC: CLIP Can Count Accurately through Enhanced Blockwise Classification

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.500131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.500131Z digest=sha256:9b6b8ea6a9cfc0d9399a716ec00f5abd21bb452d3ac22226bc56d13cc62e63f6

Observation ca8ccce3-3a66-418b-96d1-4e4cd7f48a02 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

On the rankability of visual embeddings ClipCap: CLIP Prefix for Image Captioning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.511161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.511161Z digest=sha256:7ef7afd2d058f987c2e18e58e232fb4321bd064523b50c36fcb8780098bccf5d

Observation ead6e43e-4d76-4484-8b92-a25d7b2ded35 · outbound

This paper cites A V A: A Large-Scale Database for Aesthetic Visual Analysis.

On the rankability of visual embeddings A V A: A Large-Scale Database for Aesthetic Visual Analysis

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.516129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.516129Z digest=sha256:4fdf144b232f70feb0590fa204037caa804e405655d72a3ed3944de3fb9c8170

Observation 55c49fff-8449-4320-8028-83eba01abd3f · outbound

This paper cites A metric learning reality check.

On the rankability of visual embeddings A metric learning reality check

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.441046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.520788Z digest=sha256:1ef2a5e1e8a567e0886c871f2a654a9b1069ec8f6e28446983f44eb1e12502d8

Observation f8ee04c5-2aa8-46e1-be7d-9572b7fa1707 · outbound

This paper cites Parts of Speech-Grounded Subspaces in Vision-Language Models.

On the rankability of visual embeddings Parts of Speech-Grounded Subspaces in Vision-Language Models

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:11:52.533367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.526319Z digest=sha256:6468cf2ba1942a60c03f0243ba133c21635417c5e9f56280e4aac74808c26f02

Observation 18e7cfd6-4087-4ef9-beec-b45a1a211e6a · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

On the rankability of visual embeddings DINOv2: Learning Robust Visual Features without Supervision

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.531605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.531605Z digest=sha256:7b8f0292a7e41ec26052c6b60308b4d870001b07a5393623fc8d0b7b64c42a9b

Observation bd3a1569-da58-45a3-826f-5ab8b6110eaf · outbound

This paper cites Teaching CLIP to Count to Ten.

On the rankability of visual embeddings Teaching CLIP to Count to Ten

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.537121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.537121Z digest=sha256:0adaa30ac47ef1a9e318fd7647241271c83d1fa551ea1ea2eb68aa5cb0fb1315

Observation 75fa21f4-3934-4458-8b3d-b3fd6aaebe20 · outbound

This paper cites Dating Historical Color Images.

On the rankability of visual embeddings Dating Historical Color Images

Reference 42

Resolution
malformed identifier
no resolver link, observed 2026-08-06T20:11:51.542600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.542600Z digest=sha256:25d55439eec321389a29148f5adc7c96892952306ca8bb4a163864419110d1a8

Observation 5e8fa665-8c00-4a0b-b22f-f10365a35b98 · outbound

This paper cites Relative Attributes.

On the rankability of visual embeddings Relative Attributes

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.548487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.548487Z digest=sha256:7ac019494eb1d2d0711bf7507cef18eafd450e2412206b182bd9921f4bdd8dc8

Observation 0babb719-b950-468c-a839-5b56f2279032 · outbound

This paper cites HYDEN: Hyperbolic Density Representations for Medical Images and Re- ports.

On the rankability of visual embeddings HYDEN: Hyperbolic Density Representations for Medical Images and Re- ports

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.422955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.553683Z digest=sha256:53b2cb08776c0c1ca633bee1c2d2023bbc426df8e763d92f9ff08299144da3ae

Observation a0e88aef-b4aa-4aff-90ec-4a7203a7dd5e · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

On the rankability of visual embeddings Learning Transferable Visual Models From Natural Language Supervision

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.559058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.559058Z digest=sha256:631c42aa7ce1b9284756c056e9b8a8a92acd487958de02d29546a9d2fb04a5f6

Observation 025ec07b-d7e3-459d-99a9-9abf9a1a44e9 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Super- vision.

On the rankability of visual embeddings Learning Transferable Visual Models From Natural Language Super- vision

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.404122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.564088Z digest=sha256:f187dbec2212075d43ca4822089e67520ec06bdc88063b69e2439eeb925dc49c

Observation e2d4c7c3-cb2a-4eed-bf32-218603569642 · outbound

This paper cites Steering Llama 2 via Contrastive Activation Addition.

On the rankability of visual embeddings Steering Llama 2 via Contrastive Activation Addition

Reference 47

Resolution
malformed identifier
no resolver link, observed 2026-08-06T20:11:51.569573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.569573Z digest=sha256:7bff404226d989935895f2a526eae9b288ce86739f2e37447b605c8c8ea14730

Observation fe3b8e31-c769-4e73-ae32-dd0e949665a4 · outbound

This paper cites Finetuning CLIP to Reason about Pairwise Differences.

On the rankability of visual embeddings Finetuning CLIP to Reason about Pairwise Differences

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.575686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.575686Z digest=sha256:45b2ea0dd73e0813080a7ff25af5d7d44648cda9f2163a7c840858b49f595f60

Observation 3fd03dbe-f488-4d9f-be8d-8c43e02a1b47 · outbound

This paper cites Improving Image Encoders for General-Purpose Nearest Neighbor Search and Classification.

On the rankability of visual embeddings Improving Image Encoders for General-Purpose Nearest Neighbor Search and Classification

Reference 49

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T20:11:52.391746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.583085Z digest=sha256:b6fc7eedbbb083a5537bc808ce5eac118536ea007331b133ffaceccd2f72f6ea

Observation bae4703a-499c-46ee-8035-7a7ab5c85cd7 · outbound

This paper cites Linear Representations of Sentiment in Large Language Models.

On the rankability of visual embeddings Linear Representations of Sentiment in Large Language Models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.386329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.588973Z digest=sha256:cc3e050ad33a0a73f697e82c364ce623c616acfd947c02da5579531f69976a04

Observation 3238f178-81da-4011-99c2-d6fe4be79ac2 · outbound

This paper cites Linear Spaces of Meanings: Compositional Structures in Vision- Language Models.

On the rankability of visual embeddings Linear Spaces of Meanings: Compositional Structures in Vision- Language Models

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.368075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.601038Z digest=sha256:03ce2935b4d2904467a92e913ef078ccc1187e7683743fdf5b417bf53f97689c

Observation 1b0cdd12-01f9-4d10-9361-19f08ebb1ee4 · outbound

This paper cites Deep image prior.

On the rankability of visual embeddings Deep image prior

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.348792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.607467Z digest=sha256:6bf5528ef585b81312e1a164560577adf89d99795fcc129e7634fcfe0ffd9f40

Observation e4fe7ec6-2185-47a2-af26-0d7a5a3c787e · outbound

This paper cites BEYOND DECODABILITY: LIN- EAR FEATURE SPACES ENABLE VISUAL COMPOSITIONAL GENERALIZATION.

On the rankability of visual embeddings BEYOND DECODABILITY: LIN- EAR FEATURE SPACES ENABLE VISUAL COMPOSITIONAL GENERALIZATION

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.330263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.612865Z digest=sha256:00608026bbab6b08c15c0c0f389aad4abbaa5b3b5a12c0d4d6f4a7f9380db733

Observation 4230b46f-7018-4a08-8e07-12086fe5af61 · outbound

This paper cites Intermediate Layer Classifiers for OOD Generalization.

On the rankability of visual embeddings Intermediate Layer Classifiers for OOD Generalization

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.308628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.617981Z digest=sha256:a0bad3d920f3fd914a1396ab652a89ea94c2b5a8f506023e1cb3da4e950f7333

Observation 9247635d-3ab4-46b7-847f-bf53ff49548e · outbound

This paper cites Order-Embeddings of Images and Language.

On the rankability of visual embeddings Order-Embeddings of Images and Language

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.623271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.623271Z digest=sha256:5e5b99159454850d38d5b8eb6abd3a60e257fad570c31597c35a84aae6e6fee3

Observation 6b5e8000-421a-459c-a408-8b76f589aed6 · outbound

This paper cites Exploring CLIP for Assessing the Look and Feel of Images.

On the rankability of visual embeddings Exploring CLIP for Assessing the Look and Feel of Images

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.628732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.628732Z digest=sha256:6c1c9704579ee6961dd3a496da0e9a4938d3e00690be16339860e646d801313c

Observation 936ac420-362e-4fbf-8e94-e91a8469f375 · outbound

This paper cites Learning-to-Rank Meets Language: Boosting Language-Driven Ordering Alignment for Ordinal Classification.

On the rankability of visual embeddings Learning-to-Rank Meets Language: Boosting Language-Driven Ordering Alignment for Ordinal Classification

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.290182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.634171Z digest=sha256:01039dacafb5cdeab30144efb202814a45e29553e913eb26e3181b631ab95938

Observation e4732314-04f8-4044-ba8f-e92acfb33c58 · outbound

This paper cites Learning-to-rank meets language: Boosting language-driven ordering align- ment for ordinal classification.

On the rankability of visual embeddings Learning-to-rank meets language: Boosting language-driven ordering align- ment for ordinal classification

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.271360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.640208Z digest=sha256:956c816d3dc437f0a48a1dd34106b624726bd36e85134c9eeca06b464611c613

Observation be34882b-7800-4c43-bc56-e814d8e5511c · outbound

This paper cites Understanding Contrastive Representation Learning through Alignment and Uniformity on the Hypersphere.

On the rankability of visual embeddings Understanding Contrastive Representation Learning through Alignment and Uniformity on the Hypersphere

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.646034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.646034Z digest=sha256:a316c2740384b009e079ae3284acc6a7d2bc11e0de1c04d0e801e1f63f041b77

Observation 6f1bb15f-0122-4804-9671-9633a12c7e06 · outbound

This paper cites Disentangled representation learning.

On the rankability of visual embeddings Disentangled representation learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.252755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.651508Z digest=sha256:69dd59202f3d80f30469f2e2e55b10cc88c277de0198787b06d1be6d077a2067

Observation f0ff0cf9-3c50-4858-9037-a99cb6c6d5d9 · outbound

This paper cites ConvNeXt V2: Co-designing and Scaling ConvNets with Masked Autoencoders.

On the rankability of visual embeddings ConvNeXt V2: Co-designing and Scaling ConvNets with Masked Autoencoders

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.656002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.656002Z digest=sha256:c38fc2480a491c69803e2a6b7216d18ed8fa1b980bc6c038cbee53f673544cad

Observation c36c12cd-79c4-42e7-a276-151ea491f4ee · outbound

This paper cites CLIP Brings Better Features to Visual Aesthetics Learners.

On the rankability of visual embeddings CLIP Brings Better Features to Visual Aesthetics Learners

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:11:52.250347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.661258Z digest=sha256:a54031d2418ebd65c2e0afd876edde9c31303419aa60fc31a141992d15b3f9e4

Observation 2dd22aa8-820d-4c84-b04a-02be321730dd · outbound

This paper cites FILIP: Fine-grained Interactive Language-Image Pre-Training.

On the rankability of visual embeddings FILIP: Fine-grained Interactive Language-Image Pre-Training

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.667082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.667082Z digest=sha256:c5307b7f6230a3a507424fac078cc20a65dba21d1e40ab32dca41453791b659d

Observation 01c94fe3-87ce-4cef-8ba2-51249241ef8a · outbound

This paper cites Just noticeable differences in visual attributes.

On the rankability of visual embeddings Just noticeable differences in visual attributes

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.233523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.673658Z digest=sha256:0cbb94a18dfdb3e10b0ec2b9accfda3b58bf2e1736b5ef32800f8d8b72752fcb

Observation 82ad9c1e-42a9-416c-b056-06aac0007fb0 · outbound

This paper cites CoCa: Contrastive Captioners are Image-Text Foundation Models.

On the rankability of visual embeddings CoCa: Contrastive Captioners are Image-Text Foundation Models

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.216137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.679214Z digest=sha256:c036535d107fa55d38123761927a80028a0210fb827e4e1c2bbde17fe356eb2d

Observation db9bfc04-6741-4d31-9d31-7c52aab8b687 · outbound

This paper cites RANKING-AWARE ADAPTER FOR TEXT-DRIVEN IMAGE OR- DERING WITH CLIP.

On the rankability of visual embeddings RANKING-AWARE ADAPTER FOR TEXT-DRIVEN IMAGE OR- DERING WITH CLIP

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.200041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.684176Z digest=sha256:977bf79435b73f8a8968d112a1f0b38b6485d11d3fd67008388694f7d0188873

Observation 3f14a96a-9b7b-416c-97c0-65bfe82f4afc · outbound

This paper cites Ranking-aware adapter for text-driven image ordering with CLIP.

On the rankability of visual embeddings Ranking-aware adapter for text-driven image ordering with CLIP

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:11:51.801956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.689502Z digest=sha256:1bf5ee3a50c547f8f373c0013495c6679989cb47644594ff9f1e4f3d71bc4872

Observation 3ed8fddf-ba6d-48cb-8c12-f3952ea698e1 · outbound

This paper cites When and why vision-language models behave like bags-of-words, and what to do about it?.

On the rankability of visual embeddings When and why vision-language models behave like bags-of-words, and what to do about it?

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.695389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.695389Z digest=sha256:c7eaade33de9797890a2d4da39f386670fff604c30772eb3e98171bf0dd6642a

Observation 68d75df2-e218-4e43-b214-cd4e2f6db13d · outbound

This paper cites Single-Image Crowd Counting via Multi-Column Convolutional Neural Network.

On the rankability of visual embeddings Single-Image Crowd Counting via Multi-Column Convolutional Neural Network

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.701459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.701459Z digest=sha256:7448ae51e7e7603ab06430f8c8ea5defb10567479f6b8dbcb8932c8e7ee135b4

Observation ef6fcadf-a217-4378-ae55-e1e091615786 · outbound

This paper cites an unresolved cited work.

On the rankability of visual embeddings Unresolved cited work

Reference 70

Resolution
verified exact
raw_fallback, observed 2026-08-06T20:11:52.187722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.706634Z digest=sha256:0d266320fed0988e80292fef1794fe50deddecdd8b7885067e58179baecb7e43

Observation e330120c-f2ee-478c-809d-c67b5537b483 · outbound

This paper cites Learning Ordinal Relationships for Mid-Level Vision.

On the rankability of visual embeddings Learning Ordinal Relationships for Mid-Level Vision

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:11:53.183574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.717772Z digest=sha256:fbe29b92ea1296003d448eea6e78a019f75b131dddd767a3f67b6675d7dfbe74

Observation 73c93712-e3f5-4205-97ab-3869d8d60f4d · outbound

This paper cites Age Progression/Regression by Conditional Adversarial Autoencoder.

On the rankability of visual embeddings Age Progression/Regression by Conditional Adversarial Autoencoder

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:11:51.767240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T20:11:51.712243Z digest=sha256:76984d62432af07a98160be7de727e47b829956cd3a7cb0b2fa88dd6c1de7e16

Observation ec1d0391-a4a0-4c85-84d5-c2f4ae582f49 · outbound

This paper cites Linear Representations of Sentiment in Large Language Models.

On the rankability of visual embeddings Linear Representations of Sentiment in Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:51.594734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:51.594734Z digest=sha256:87595273a008e6983bc4ff9142a83d82c4da984c3b26fc4cf700a5824bc68a69

Observation 7d03c2f3-6c39-41dc-89c0-7120d1ade950 · outbound

This paper cites DOI: 10.1109/TIP.2020.2967829.

On the rankability of visual embeddings DOI: 10.1109/TIP.2020.2967829

Reference 4056

Resolution
unresolved
no resolver link, observed 2026-08-06T20:11:50.176194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:11:50.176194Z digest=sha256:1cb5eb1d77e2b9b2dc5569f6d4ac2059e636b9ea5d13ce676dda9355b924939c

Pith citing papers

No inbound Pith citation observations are available.