Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T12:21:22.544303Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 1 inbound Pith citation observation for arXiv:2507.21871.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T12:21:22.544303Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T16:32:28.422930Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T16:32:28.671394Z
19 of 19 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9bdf7fb3-8a6f-4f18-8ba3-ce6f87e63edf · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities Emerging evidence suggests that human brain representations in both vision and language are well predicted by semantic feature spaces obtained from large language models (LLMs)
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 808fd061-f75c-4eae-8e7f-de22674eed3a · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities (A) Cross-validated non-negative least squares regression was used to model the brain RDMs at every searchlight location using the behavioural RDMs derived from our MA tasks
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5bbd874c-07bd-406d-97ab-218d9a5d3fd2 · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 09891198-0646-410c-acc0-115c75458ddb · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities linguistic modality
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2a6fb438-672d-4b6a-b27c-01a0d165e43e · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities (A) Participants completed the MA task either on 100 natural scene images (visual modality left) or 100 sentence captions 8 describing the images (linguistic modality right)
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6cb16c37-3857-4953-adea-9511c086b9a4 · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities Scaling Laws for Transfer
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 698df9c9-edc8-4cb8-b09b-0509435d986b · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities VoLTA: Vision-Language Transformer with Weakly-Supervised Local-Feature Alignment
Reference 128
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 678bc90a-5ea0-4743-acde-4b632dd97752 · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities A., Schmitz, T
Reference 134
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 24dcb9c7-e65c-45d3-be8e-6ace6e4c0c1b · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities A., Kiani, R., Bodurka, J., Esteky, H., Tanaka, K., & Bandettini, P
Reference 245
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0f02e926-5b21-49b1-aaf4-c65258533fd1 · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities visual modality
Reference 2012
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c8bfdb37-929c-4e63-94b7-f280d91f9d21 · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities We show that a similar relational structure emerges for both linguistic and visual inputs
Reference 2013
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5545cf45-c302-4c49-ac71-85dcceb97a61 · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities The significance of correlations was tested using one-sided t-test across participants and corrected for multiple comparisons at FDR p < 0.05
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b50d41b6-b50e-4c2d-bff9-78862206375c · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities Unresolved cited work
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 044c86dc-3609-493a-a98d-061e1f87a0a4 · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities It may be that the visual system translates sensory inputs into modality-agnostic representations that reflect stable, relational patterns observed in the real-world environment
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c1b187da-df2f-4334-a4a4-b53ba3204fdc · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities The sentence captions were collected from five human annotators as part of the Microsoft Common Objects in Context database (Lin et al., 2014)
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8746e166-da76-43e8-851d-94d81dc74e1e · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities Unresolved cited work
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation be82ce1a-0acd-48f6-9705-f40062c5cb0f · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities Unresolved cited work
Reference 4081
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 233de5ba-3820-4e13-b61c-6c4c11479083 · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks
Reference 6241
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2112a0ed-47c6-4cc6-acf6-2f903dea97d9 · outbound
Representations in vision and language converge in a shared, multidimensional space of perceived similarities Visual representations in the human brain are aligned with large language models
Reference 9383
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e30cad62-9bbc-4c50-a7ce-b698b07ceb1e · inbound
Disentangling the Factors of Convergence between Brains and Computer Vision Models Representations in vision and language converge in a shared, multidimensional space of perceived similarities
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.