Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:13:47.973601Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2507.03542.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:13:47.973601Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 514e91e1-0e31-4f17-ae39-97a34219086b · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Flamingo: a Visual Language Model for Few-Shot Learning,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b2230a53-3e9d-4ba8-86b6-b639a7ca850b · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Vision-language models do not understand negation, 2025
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d092a9c8-e782-4bf8-9063-2ac28c24e2ce · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Interpreting CLIP with Sparse Linear Concept Embeddings (SpLiCE)
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 320ef00d-9715-47d0-be98-97bb7e4d2b81 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Evolving interpretable visual classifiers with large language models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8c94f42e-3714-46a5-b39f-d5ff3fe20e6c · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor The platonic representation hypothesis, 2024
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4d74d4bf-483a-4ca9-9d77-c693da85ec94 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Scaling up visual and vision-language representa- tion learning with noisy text supervision
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d82a0fa9-6153-4417-9d18-8c851062505e · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor What's "up" with vision-language models? Investigating their struggle with spatial reasoning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12f6097e-142c-47d1-b06c-549592b1a386 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Large language models struggle to learn long-tail knowledge, 2023
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e4a5d569-e6ae-401d-93b7-87eec66b9952 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Cifar- 100 (canadian institute for advanced research)
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ae75167e-f446-46e0-a689-3ca8eba802ac · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Align before fuse: Vision and language representation learn- ing with momentum distillation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8910940d-96a0-4bda-be08-75a09dfa55d2 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Improved baselines with visual instruction tuning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cffcfc55-cc6d-4671-93b8-a982536abc0a · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Visual instruction tuning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e51f2db3-70d4-4c48-a252-12925c062df8 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Visual classification via description from large language models, 2022
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c50dfedd-ed59-455a-b089-498faed52d00 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Visual Classification via Description from Large Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 56026ccf-c83c-4f76-8e27-db45f747e6e7 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Slip: Self-supervision meets language-image pre- training
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5a8d458-3ca0-4235-a84f-6e78a3ee3794 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Dinov2: Learning robust visual features with- out supervision, 2024
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58fbfcf4-8fab-4580-9678-2f8346bafff2 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Learning Transferable Visual Models From Natural Language Supervision
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c9fdb4a-06e6-4ba2-8ebd-64d9158315fe · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Learning transferable visual models from natural language supervision
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f61589dd-d120-4037-aa20-442ae9ede13b · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Sophia Koepke, Oriol Vinyals, Cordelia Schmid, and Zeynep Akata
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a567c6e1-e12f-4ac8-880e-9a6f36b24095 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Sun, and Swarat Chaudhuri
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation faf6d0b9-3eef-4c57-b727-a935747683a1 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Love, Christo- pher J
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 335682cf-102e-4107-8b7d-25d1d37a5bdc · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Understanding the emergence of multimodal representation alignment, 2025
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 98ee8c2b-cd79-4be2-852d-65a645b5068b · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Eyes Wide Shut? Exploring the Vi- sual Shortcomings of Multimodal LLMs
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6835cae2-5dcb-4d53-9c6f-1e658f69a0ca · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Building a bird recognition app and large scale dataset with citizen scientists: The fine print in fine-grained dataset collection
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f318744b-75a4-477c-9b73-e1fdda7efc10 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 044a7e41-72bb-41ee-9e8f-f053afa823bf · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Learning Concise and Descriptive Attributes for Visual Recognition
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 40433328-06c2-4aa4-8d6f-d9e4f40a3787 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Learning concise and descriptive attributes for visual recognition
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2511dd34-242d-489f-a07b-a6243ef53d17 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Language in a bottle: Language model guided concept bottlenecks for in- terpretable image classification, 2023
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d193fc5f-09a5-488c-9559-38c6f818b9a8 · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Filip: Fine-grained interactive language-image pre-training
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 19274119-75c2-4d67-9c60-8d015a7e368e · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor laysan albatross, which is a
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0bdd19ee-a710-47ff-b531-5ce9f118ab5d · outbound
Beyond Accuracy: Metrics that Uncover What Makes a 'Good' Visual Descriptor Flamingo: a Visual Language Model for Few-Shot Learning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.