Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-21T23:56:31.695833Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2507.01201.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-21T23:56:31.695833Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:04:28.680647Z
A source-named dated measurement, never combined with another source.
Source: cited_works
54 of 54 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1884349b-28d1-4e9c-829f-c1d82e12a68b · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models The platonic representation hypothesis
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation aa671778-35d3-4358-9db2-183f61e2ab58 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Republic (De Republica)
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6661b766-67bb-4632-a074-af8229b7c29b · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Revisiting model stitching to compare neural representations
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 09b798a0-7858-4b71-8c3a-43fe2e59ee1a · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Wit: Wikipedia-based image text dataset for multimodal multilingual machine learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d4e11f50-3246-4ef3-82a2-cb983f9693dd · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Kornblith, M
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 015c6b74-2f46-437e-9b8c-acc4242d9629 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Raghu, J
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 70185435-131b-4f8f-a325-1e7ef61c47b5 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Insights on representational similarity in neural networks with canonical correlation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 93f621e8-18c8-48b0-8454-70f5a71b96ab · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Similarity of neural network models: A survey of functional and representational measures
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ba5ce699-0dc4-4b4d-b7ba-458d5856a30f · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Visiolinguistic attention learning for multimodal coreference resolution
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6e63313e-af84-41c8-9eb5-d6c52f445a02 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Language Is Not All You Need: Aligning Perception with Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1865d481-b8c5-4bca-9027-86f7531156af · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Understanding image representations by measuring their equivariance and equivalence
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f40a460e-0f9c-431a-97db-0ce13c1ea7e8 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Gemini: a family of highly capable multimodal models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation df6e5033-126a-4172-8c9d-358e45fab2c6 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Gpt-4 with vision
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 112e250e-94a7-4ead-933b-02bf2b1f2866 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Llama 3 model card
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e20b53b3-ebef-41b9-a8ab-7d15cc8ceac3 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Learning transferable visual models from natural language supervision
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f6ff8060-6d3c-428b-802c-74dad1b4042f · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Scaling up visual and vision-language representation learning with noisy text supervision
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation afc26b20-4288-45b1-8e11-2a7407fd45de · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bf14bd92-1a09-484b-9876-8684d686a519 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models BLIP-2: bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b93e8f73-4c1d-47c1-a945-13b128d1e4e0 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models A Conditional Singular Value Decomposition
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1c74bd89-e1dc-4034-aaa0-6ce08acc7071 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Sugar- crepe: Fixing hackable benchmarks for vision-language compositionality
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f7a57e19-6d70-4e95-8107-6c5407cd5af9 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fe11eed0-059d-4a1f-93e9-ba6787197171 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models 2 olmo 2 furious
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f7c615b2-89e6-4841-85e3-0c88df16f2d3 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d26a58e9-2d1b-485b-851b-fa1354904948 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Deep residual learning for image recognition
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b726d433-1636-42af-9c0b-24d6905c3980 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models What regularized auto-encoders learn from the data- generating distribution
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2cc34b5a-d44e-4b71-800a-6c07812c231a · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Regularized linear autoen- coders recover the principal components, eventually
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 64ef0c76-b5ff-4d18-84ac-14ef94b5fc65 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0e128f5a-2cb4-4f89-9477-af6dfb3db34b · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Glu variants improve transformer
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9abf31b7-e272-497c-a2f2-5f3959c7716a · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Gomez, Lukasz Kaiser, and Illia Polosukhin
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0819e5dd-5dc1-43ea-a242-dfe4e7fe1d6c · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models When and why vision-language models behave like bags-of-words, and what to do about it? In International Conference on Learning Representations
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation de1aacee-62e3-49ee-8b5d-87c326c62888 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Chen, Daniel Y
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation fceee09c-7378-4188-a9c9-e8bef8b1ac20 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Fu, Mayee F
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ba228811-c1fa-4b6a-932d-fa5760ff1ebb · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Supervised Contrastive Learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f96e9c78-2966-4dc0-bca5-9e5efebe71d4 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Winoground: Probing vision and language models for visio-linguistic compositionality
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 92f2ec17-e16a-4464-8f8a-645a437f9f24 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Masked Autoencoders Are Scalable Vision Learners
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 413bd0b0-6595-46b7-9e31-c7a78cbd8126 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Unsupervised learning of visual features by contrasting cluster assignments
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 393e72a8-c88d-4821-8836-b4c8d89ea711 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Swin transformer: Hierarchical vision transformer using shifted windows
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0cd0f8e7-f8aa-4739-a41b-5bd57d5f91b6 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models An image is worth 16x16 words: Transformers for image recognition at scale
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e22e1674-a40c-42f2-9bee-fab1408a1c54 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Decoupled weight decay regularization
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 47fc6a20-e9fb-4fae-8649-781de989cfe7 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Openclip, July 2021
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a3e116a6-18b4-4366-b852-a283934dfa30 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models LAION-5b: An open large-scale dataset for training next generation image-text models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 21c867b0-b8ff-47f5-aa2e-8909202fb47c · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Sigmoid Loss for Language Image Pre-Training
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 77c04473-9739-4271-8420-707b515d46a9 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Understanding dimensional collapse in contrastive self-supervised learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 22686a86-f3fc-4633-af21-e226d54fceec · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Curriculum learning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 14356c54-5c0c-4cbe-91cd-ca6d331ed493 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav)
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 980043ab-76ef-4d94-b9bb-65a0343a36f2 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Network dissection: Quantifying interpretability of deep visual representations
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b4847d06-45cd-4dd8-ad15-d03dbfb3174e · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Foundation models for time series analysis: A tutorial and survey
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8ad6a9e1-77f1-4219-841b-e2732da0fe8e · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Totem: Tokenized time series embeddings for general time series analysis
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 71d08308-1006-449f-8cc3-f930f33cc915 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models MOMENT: A Family of Open Time-series Foundation Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2890e573-f06a-4de3-af5b-525cf6b937c7 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models A decoder-only foundation model for time-series forecasting
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0bdff118-588d-4801-9d1b-14a9ec146888 · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Relations between two sets of variates
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b4689f2c-96b2-4014-85c9-959dc96c994c · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Reproducing kernel hilbert space, mercer’s theorem, eigenfunctions, nyström method, and use of kernels in machine learning: Tutorial and survey
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 847a45d3-0880-48a9-909b-72182a59025e · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 32af4824-70e7-4fc8-bee6-8a3980c8392e · outbound
Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models High-dimensional canonical correlation analysis
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2e5ce6c1-74e4-43b8-8da9-6df64d02a772 · inbound
Multi-Way Representation Alignment Escaping Plato's Cave: JAM for Aligning Independently Trained Vision and Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.