Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T15:00:24.297044Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2606.02168.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T15:00:24.297044Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9f695afd-c87c-47b1-9790-c4df90d71020 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Vqa: Visual question answering,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4457f8f5-fbf4-4ffa-9dba-0e57cc100178 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c1bc5aa-0459-450a-8f87-2c13003ec651 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Image as a foreign language: Beit pretraining for vision and vision-language tasks,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1ee886b-4cb1-456b-a3a8-0f302ee759f9 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Connectionism and cognitive archi- tecture: A critical analysis,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65222092-fd15-40b3-b9b4-9701a8da8b92 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Systematic Generalization: What Is Required and Can It Be Learned?
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0b607505-1732-4b83-8733-8c91f2a66987 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA CLOSURE: Assessing Systematic Generalization of CLEVR Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b87c9b17-dd30-4251-9545-19d00601e586 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA A benchmark for systematic generalization in grounded language under- standing,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fba93d66-587b-4ac7-8b74-f3c11a71e192 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA How modular should neural module networks be for systematic generalization?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c7b0d6e-87c1-44d3-bb8d-6be96d409160 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Meta module network for compositional visual reasoning,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e449ac8-373b-4afc-9fa4-e73261789a87 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Transformer module networks for systematic generalization in visual question answering,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f084baa-277c-4598-b2a3-30da671d3a4c · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Detection-based intermediate supervision for visual question answering,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c74a796-66f4-498b-ac7a-6f43f7c3d3ac · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Neural module networks,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0bc193f-b735-4ab7-90fd-2ebbef8c892d · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Linguistically routing capsule network for out-of-distribution visual question answering,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5b9e607-d1ae-4257-a2bd-0af933afc54f · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Neural- symbolic vqa: Disentangling reasoning from vision and language under- standing,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01199c1f-7d17-4ae4-b218-63bd8eb946f5 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Multimodal graph networks for composi- tional generalization in visual question answering,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c76e7700-179c-495c-8ed8-f954ff50c167 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Compositional generalization in neuro- symbolic visual question answering,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8765eb5-b42d-491a-9752-cb6f5dfccc1c · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Mdetr-modulated detection for end-to-end multi-modal understanding,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e75fec6e-a4a2-425b-bcfe-8bf105aa7729 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Exploring the effect of primitives for compositional generalization in vision-and-language,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebd91367-320c-487d-a84d-184ccc174617 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Towards a Definition of Disentangled Representations
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b7b3ef2f-a364-4b1b-bfd4-67041aa338ff · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Causal inference in statistics: An overview,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 258d9ccf-cc4c-44de-b1b1-18147eb74029 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Group equivariant convolutional networks,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a41f3597-08c7-4efc-8e0a-7316e03724d7 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Mutan: Multi- modal tucker fusion for visual question answering,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb009872-6290-49f3-8e1a-e7ddef126de4 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Joint embedding vqa model based on dynamic word vector,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01a0301e-1ca1-4e54-9a71-b08348506c9e · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Where to look: Focus regions for visual question answering,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca384247-0ec2-4078-9424-34498047a0bb · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d47d5ad8-33a3-4fe5-8afe-92d825fb68f7 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Visual question answering with dense inter-and intra-modality interactions,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cc4150e-e47b-4b05-99e9-0b3c88ffdd9c · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Attention is all you need,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dd99fb5-5bb4-45b4-9e2a-353802844ec5 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA LXMERT: Learning Cross-Modality Encoder Representations from Transformers
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 80c0c08e-dfbc-43c6-b4ca-4ff972963b9f · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Boosting generic visual-linguistic representation with dynamic contexts,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 335dc2dc-8591-4188-bb68-ed35c7241019 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA MM-LLMs: Recent Advances in MultiModal Large Language Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6e5f3483-d568-407b-ba10-ce9cfc691609 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Exploring compositional generalization of large language models,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a192a7dd-0f1a-4b78-b1b1-3f3de31dfc57 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Learning, reasoning, and compositional gen- eralisation in multimodal language models,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b00e832f-6212-4d6e-9ca0-fd58ecdc653f · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA An empirical study of gpt-3 for few-shot knowledge-based vqa,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d935252-7e24-42eb-8709-040f78080aae · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Learning to supervise knowledge retrieval over a tree structure for visual question answering,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63888886-6f4c-493d-9d62-95357b596296 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Towards causal vqa: Revealing and reducing spurious correlations by invariant and covariant semantic editing,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5aafe7d4-7a87-42d1-94a2-89fbb2739e45 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Film: Visual reasoning with a general conditioning layer,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da6f634c-78ca-46de-b677-f820a9c6c336 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Transparency by design: Closing the gap between performance and interpretability in visual reasoning,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53ac0129-d8e0-4012-aae4-6eb949a2eef0 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Lexsym: Compositionality as lexical symmetry,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc064c12-6695-4991-9407-44541884bcef · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86af0957-9e41-46cb-a8e0-c02afd9b4751 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Disentangled representation learning,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53d9498d-4e46-4b55-991e-7c966ce6c30b · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Robustly disentangled causal mechanisms: Validating deep representations for interventional robustness,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61ccfafb-9bd2-4675-9ce4-a8b312d830b3 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA On causally disentangled representations,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96a10e15-d4bd-4f1f-bbfc-153104aee8a8 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Derf: Decomposed radiance fields
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8341495a-bedf-4f62-b050-9cf983060fe9 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA A meta-transfer objective for learning to disentangle causal mechanisms,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ca05e7a-8be0-41c9-bead-4de496e65a48 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Disen- tangled generative causal representation learning,
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30fb5c3b-0d91-4263-84d5-8caf94e14969 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Equivariant flows: Exact likeli- hood generative learning for symmetric densities,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56c8af20-d85e-4e30-8c9c-b5b6c7797425 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Group equivariant capsule networks,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4302c82-8bac-44fc-82b1-60650abac4f1 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Learning generalized transformation equivariant representations via autoencoding transforma- tions,
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16cd8704-c1fd-4a72-8d9a-a44a1b2d5715 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Equivariant and invariant grounding for video question answering,
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa544d7f-30fb-4361-b3f8-1fb9263dc572 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Vilt: Vision-and-language transformer without convolution or region supervision,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9881197d-0f97-43aa-b7d6-4aeea8839f8f · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Invariant Risk Minimization
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 637be5b8-5187-4b19-8d72-d63f0f60b91d · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Bottom-up and top-down attention for image captioning and visual question answering,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7d0607d-2664-4545-848c-4946f2a69d3d · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Auto-Encoding Variational Bayes
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c2c2093a-fb46-4478-8693-e3d95e68f4d4 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Grad-cam: Visual explanations from deep networks via gradient-based localization,
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddaeb4a4-9767-4c97-aae1-6ce792e38b79 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Clevr: A diagnostic dataset for compositional language and elementary visual reasoning,
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b193d3b6-e1ac-4ae8-885a-a4a28387502f · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Making the v in vqa matter: Elevating the role of image understanding in visual question answering,
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afa23648-02e5-42d1-b46c-f31b83374864 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Gqa: A new dataset for real-world visual reasoning and compositional question answering,
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33fd2797-4764-4720-9ee9-d266fa796a1b · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA In: Fleet, D., Pajdla, T., Schiele, B., Tuytelaars, T
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 41daf745-af95-48c6-a5bd-5de2f9c9c2e8 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9c4e75a6-2522-421d-b4b1-ddddd8ee115b · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Faster r-cnn: Towards real-time object detection with region proposal networks,
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 241e80d4-6cf4-4e75-9b52-4583a6893089 · outbound
Disentanglement-Based Equivariant Learning for Compositional VQA Unresolved cited work
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.