Pith. sign in

Paper Citation Record · LEDGER

Disentanglement-Based Equivariant Learning for Compositional VQA

As of 6 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2606.02168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.02168 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T15:00:24.297044Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact10
  • verified fuzzy0
  • unresolved50
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9f695afd-c87c-47b1-9790-c4df90d71020 · outbound

This paper cites Vqa: Visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Vqa: Visual question answering,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:56cf206414a9f5673ed20d81556de039a1d82fb47bcc5999d3310bcece6b0e58

Observation 4457f8f5-fbf4-4ffa-9dba-0e57cc100178 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Disentanglement-Based Equivariant Learning for Compositional VQA Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:2f413adfe1d1a4877886f3480461e2333d2b94f07ef51219663ad1ad87c1be0c

Observation 3c1bc5aa-0459-450a-8f87-2c13003ec651 · outbound

This paper cites Image as a foreign language: Beit pretraining for vision and vision-language tasks,.

Disentanglement-Based Equivariant Learning for Compositional VQA Image as a foreign language: Beit pretraining for vision and vision-language tasks,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:5c9b383c1003da95f8dc6cb5e21c2f1465d20164933c849246bc4b78599cc99b

Observation c1ee886b-4cb1-456b-a3a8-0f302ee759f9 · outbound

This paper cites Connectionism and cognitive archi- tecture: A critical analysis,.

Disentanglement-Based Equivariant Learning for Compositional VQA Connectionism and cognitive archi- tecture: A critical analysis,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:c055fd96be4978320fab25eef41e2a1fe68463b95aff0a4b4772f8ef720890da

Observation 65222092-fd15-40b3-b9b4-9701a8da8b92 · outbound

This paper cites Systematic Generalization: What Is Required and Can It Be Learned?.

Disentanglement-Based Equivariant Learning for Compositional VQA Systematic Generalization: What Is Required and Can It Be Learned?

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:46:20.198556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:a3f0941be690714f119c5093264f8ce7c436ba46f9599e557f4ccde75f12f7b8

Observation 0b607505-1732-4b83-8733-8c91f2a66987 · outbound

This paper cites CLOSURE: Assessing Systematic Generalization of CLEVR Models.

Disentanglement-Based Equivariant Learning for Compositional VQA CLOSURE: Assessing Systematic Generalization of CLEVR Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:46:20.206113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:3ca0ae0a07e8d1083c84a33102b2895590c7e5b6aac9344619304ef149a5abe8

Observation b87c9b17-dd30-4251-9545-19d00601e586 · outbound

This paper cites A benchmark for systematic generalization in grounded language under- standing,.

Disentanglement-Based Equivariant Learning for Compositional VQA A benchmark for systematic generalization in grounded language under- standing,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:398443ba5dd88db25bf4985a8d98fc4b283344ebf30505dc59f60506a13d9f3d

Observation fba93d66-587b-4ac7-8b74-f3c11a71e192 · outbound

This paper cites How modular should neural module networks be for systematic generalization?.

Disentanglement-Based Equivariant Learning for Compositional VQA How modular should neural module networks be for systematic generalization?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:64eeae893c062e595297e4f3f8c9ed9f323022d23a26ef45734e4b0ecfa88c22

Observation 1c7b0d6e-87c1-44d3-bb8d-6be96d409160 · outbound

This paper cites Meta module network for compositional visual reasoning,.

Disentanglement-Based Equivariant Learning for Compositional VQA Meta module network for compositional visual reasoning,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:640df6bc4b19858dabdc6255a0eff7900b8ba278725e26a916e1354e809d519e

Observation 6e449ac8-373b-4afc-9fa4-e73261789a87 · outbound

This paper cites Transformer module networks for systematic generalization in visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Transformer module networks for systematic generalization in visual question answering,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:a4f34d131759b8282d6f57922070f186701ae24062164b651cc8fbe90ab3d966

Observation 1f084baa-277c-4598-b2a3-30da671d3a4c · outbound

This paper cites Detection-based intermediate supervision for visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Detection-based intermediate supervision for visual question answering,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:d6bec20ca19e32016560e632bd9dc8fb594d75566370b604a569039de3fa9992

Observation 0c74a796-66f4-498b-ac7a-6f43f7c3d3ac · outbound

This paper cites Neural module networks,.

Disentanglement-Based Equivariant Learning for Compositional VQA Neural module networks,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:61ab5df82f6a3f263eb8e017fcd78c8b3dc93d2211d3b67ccec63c265357b51a

Observation f0bc193f-b735-4ab7-90fd-2ebbef8c892d · outbound

This paper cites Linguistically routing capsule network for out-of-distribution visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Linguistically routing capsule network for out-of-distribution visual question answering,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:7f51f6b04fe070f935e8e0bdd45b5b511307c4995709bfbb8b281f01e9c5503f

Observation c5b9e607-d1ae-4257-a2bd-0af933afc54f · outbound

This paper cites Neural- symbolic vqa: Disentangling reasoning from vision and language under- standing,.

Disentanglement-Based Equivariant Learning for Compositional VQA Neural- symbolic vqa: Disentangling reasoning from vision and language under- standing,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:dbac7efa1f8cdec10d5f578a3eaa9dc0c002da5563447b8d157ec61b6ccda831

Observation 01199c1f-7d17-4ae4-b218-63bd8eb946f5 · outbound

This paper cites Multimodal graph networks for composi- tional generalization in visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Multimodal graph networks for composi- tional generalization in visual question answering,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:571a4da6a444463041979c5dd43f7f4e7ae75c0c86dd078a02ad22a68e89e5e5

Observation c76e7700-179c-495c-8ed8-f954ff50c167 · outbound

This paper cites Compositional generalization in neuro- symbolic visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Compositional generalization in neuro- symbolic visual question answering,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:92259917f16b81306d378bf238bbd5fc33123b03460df8fba98a9379ece76460

Observation e8765eb5-b42d-491a-9752-cb6f5dfccc1c · outbound

This paper cites Mdetr-modulated detection for end-to-end multi-modal understanding,.

Disentanglement-Based Equivariant Learning for Compositional VQA Mdetr-modulated detection for end-to-end multi-modal understanding,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:265907579637699c0661427d7253453707bde6e15baae59332ea762e1e92d5c3

Observation e75fec6e-a4a2-425b-bcfe-8bf105aa7729 · outbound

This paper cites Exploring the effect of primitives for compositional generalization in vision-and-language,.

Disentanglement-Based Equivariant Learning for Compositional VQA Exploring the effect of primitives for compositional generalization in vision-and-language,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:477e2a8688303a69965f2a47dd3df61d94de32c87d182bf2bd10961d243c3a64

Observation ebd91367-320c-487d-a84d-184ccc174617 · outbound

This paper cites Towards a Definition of Disentangled Representations.

Disentanglement-Based Equivariant Learning for Compositional VQA Towards a Definition of Disentangled Representations

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:46:20.190915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:fe0e3c30899f64a6daf59ee4b6a06aeb943cce175485788233f700765ff394a2

Observation b7b3ef2f-a364-4b1b-bfd4-67041aa338ff · outbound

This paper cites Causal inference in statistics: An overview,.

Disentanglement-Based Equivariant Learning for Compositional VQA Causal inference in statistics: An overview,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:52c74f6a3f2d0ad1d99e94013cf940bcd478306ec2ebd60c35627bd70038dec2

Observation 258d9ccf-cc4c-44de-b1b1-18147eb74029 · outbound

This paper cites Group equivariant convolutional networks,.

Disentanglement-Based Equivariant Learning for Compositional VQA Group equivariant convolutional networks,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:37eb98422022b7ca2f59c5bdc25aa0b77655b8f9480080838066aef97be42a4c

Observation a41f3597-08c7-4efc-8e0a-7316e03724d7 · outbound

This paper cites Mutan: Multi- modal tucker fusion for visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Mutan: Multi- modal tucker fusion for visual question answering,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:036c3cd85ce0bbe1ffd2afffdaedbb4136a8cf7bb0259c5f4a4c63c60948ab5a

Observation eb009872-6290-49f3-8e1a-e7ddef126de4 · outbound

This paper cites Joint embedding vqa model based on dynamic word vector,.

Disentanglement-Based Equivariant Learning for Compositional VQA Joint embedding vqa model based on dynamic word vector,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:169f53cd0855bafa0fe29779dcdadac14038c3e768f8b1f7afc0738058c09869

Observation 01a0301e-1ca1-4e54-9a71-b08348506c9e · outbound

This paper cites Where to look: Focus regions for visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Where to look: Focus regions for visual question answering,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:84d113551b3d069b1e6240db0dda17306ef982a846efefa1d493e43dd6c6dcd4

Observation ca384247-0ec2-4078-9424-34498047a0bb · outbound

This paper cites The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision.

Disentanglement-Based Equivariant Learning for Compositional VQA The Neuro-Symbolic Concept Learner: Interpreting Scenes, Words, and Sentences From Natural Supervision

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:46:20.194421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:86d27b343d06caa777af0996eb9b716d279dcf4e44c09774a7edc45c50b71bd3

Observation d47d5ad8-33a3-4fe5-8afe-92d825fb68f7 · outbound

This paper cites Visual question answering with dense inter-and intra-modality interactions,.

Disentanglement-Based Equivariant Learning for Compositional VQA Visual question answering with dense inter-and intra-modality interactions,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:f2a22afc1c9a60792e3ff0fe371fbbc40de96b184baac50cf73fa45b9e3d8f12

Observation 3cc4150e-e47b-4b05-99e9-0b3c88ffdd9c · outbound

This paper cites Attention is all you need,.

Disentanglement-Based Equivariant Learning for Compositional VQA Attention is all you need,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:bfecb5accb5c57c3765c7a6da257f970e4eb0515e7c3f5020e85248e565f36e8

Observation 6dd99fb5-5bb4-45b4-9e2a-353802844ec5 · outbound

This paper cites LXMERT: Learning Cross-Modality Encoder Representations from Transformers.

Disentanglement-Based Equivariant Learning for Compositional VQA LXMERT: Learning Cross-Modality Encoder Representations from Transformers

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:46:20.180376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:b904cbf9e095acb79af6edd7007008b73653ece4920d43375d4477f5d1f5eae8

Observation 80c0c08e-dfbc-43c6-b4ca-4ff972963b9f · outbound

This paper cites Boosting generic visual-linguistic representation with dynamic contexts,.

Disentanglement-Based Equivariant Learning for Compositional VQA Boosting generic visual-linguistic representation with dynamic contexts,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:d570658da6c94cab87bae2197915c26d4fd91d2f0b4abbe9f621fa363dea48aa

Observation 335dc2dc-8591-4188-bb68-ed35c7241019 · outbound

This paper cites MM-LLMs: Recent Advances in MultiModal Large Language Models.

Disentanglement-Based Equivariant Learning for Compositional VQA MM-LLMs: Recent Advances in MultiModal Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:46:20.183807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:725f0472d1b6c85ae1bdfc02dff5e626c3296fbd8f69f626809b3c6e5d052dc1

Observation 6e5f3483-d568-407b-ba10-ce9cfc691609 · outbound

This paper cites Exploring compositional generalization of large language models,.

Disentanglement-Based Equivariant Learning for Compositional VQA Exploring compositional generalization of large language models,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:b82b74a9143361b2232d9e770b80c49773a3a790d7e93b14c5941504f660fdbe

Observation a192a7dd-0f1a-4b78-b1b1-3f3de31dfc57 · outbound

This paper cites Learning, reasoning, and compositional gen- eralisation in multimodal language models,.

Disentanglement-Based Equivariant Learning for Compositional VQA Learning, reasoning, and compositional gen- eralisation in multimodal language models,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:28b7e383ea5052bfac063302afa2bae00f1a3d59e4abee2a0a536e07723fd6dc

Observation b00e832f-6212-4d6e-9ca0-fd58ecdc653f · outbound

This paper cites An empirical study of gpt-3 for few-shot knowledge-based vqa,.

Disentanglement-Based Equivariant Learning for Compositional VQA An empirical study of gpt-3 for few-shot knowledge-based vqa,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:dfd92d2b475fc91f6301adaf6481bdf44f7501c960fa4956bdf9c37f6577b335

Observation 9d935252-7e24-42eb-8709-040f78080aae · outbound

This paper cites Learning to supervise knowledge retrieval over a tree structure for visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Learning to supervise knowledge retrieval over a tree structure for visual question answering,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:07a07b300b6ed44da008cd99d914ede97c43df6156bc3decf08cf47426abc626

Observation 63888886-6f4c-493d-9d62-95357b596296 · outbound

This paper cites Towards causal vqa: Revealing and reducing spurious correlations by invariant and covariant semantic editing,.

Disentanglement-Based Equivariant Learning for Compositional VQA Towards causal vqa: Revealing and reducing spurious correlations by invariant and covariant semantic editing,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:0a38b6d438c04f75e04b8b482892be398333cbb6be34a61c7b7a5216f396054d

Observation 5aafe7d4-7a87-42d1-94a2-89fbb2739e45 · outbound

This paper cites Film: Visual reasoning with a general conditioning layer,.

Disentanglement-Based Equivariant Learning for Compositional VQA Film: Visual reasoning with a general conditioning layer,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:d13a03b8e43a4e975c92c2a45f83dc39997d994214a19c9c5180b8733bf6adea

Observation da6f634c-78ca-46de-b677-f820a9c6c336 · outbound

This paper cites Transparency by design: Closing the gap between performance and interpretability in visual reasoning,.

Disentanglement-Based Equivariant Learning for Compositional VQA Transparency by design: Closing the gap between performance and interpretability in visual reasoning,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:cf56063b41f979d074602bb84d26b5cf5525e1ceb4b28e87fefe4db4826e7c31

Observation 53ac0129-d8e0-4012-aae4-6eb949a2eef0 · outbound

This paper cites Lexsym: Compositionality as lexical symmetry,.

Disentanglement-Based Equivariant Learning for Compositional VQA Lexsym: Compositionality as lexical symmetry,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:ba721d5ccc8d2f05e103d926445272bf623075655cb80ea93c53bfc63249166b

Observation dc064c12-6695-4991-9407-44541884bcef · outbound

This paper cites an unresolved cited work.

Disentanglement-Based Equivariant Learning for Compositional VQA Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:abc93ec246d6ac9ca6569c165159e3d5d091b28bbb0054136a1ba7eac881c157

Observation 86af0957-9e41-46cb-a8e0-c02afd9b4751 · outbound

This paper cites Disentangled representation learning,.

Disentanglement-Based Equivariant Learning for Compositional VQA Disentangled representation learning,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:c21df7af803cda6830de1202b9ab5aa6156394bf5d2b18b38843812bbdf4c263

Observation 53d9498d-4e46-4b55-991e-7c966ce6c30b · outbound

This paper cites Robustly disentangled causal mechanisms: Validating deep representations for interventional robustness,.

Disentanglement-Based Equivariant Learning for Compositional VQA Robustly disentangled causal mechanisms: Validating deep representations for interventional robustness,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:8b402cf90ca22be579c23d11c4fdc74c5220c0ca350f493a6746d7677bfb6f51

Observation 61ccfafb-9bd2-4675-9ce4-a8b312d830b3 · outbound

This paper cites On causally disentangled representations,.

Disentanglement-Based Equivariant Learning for Compositional VQA On causally disentangled representations,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:190d8f19f7b54c26c0f662a1b63caac3572ba801f62605cb6dd01b58cad3b716

Observation 96a10e15-d4bd-4f1f-bbfc-153104aee8a8 · outbound

This paper cites Derf: Decomposed radiance fields.

Disentanglement-Based Equivariant Learning for Compositional VQA Derf: Decomposed radiance fields

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T15:02:18.383830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:9584c1e98a3a17d4ed1185556bb50e5ef417940583c4cfadbc8de0070ccb6645

Observation 8341495a-bedf-4f62-b050-9cf983060fe9 · outbound

This paper cites A meta-transfer objective for learning to disentangle causal mechanisms,.

Disentanglement-Based Equivariant Learning for Compositional VQA A meta-transfer objective for learning to disentangle causal mechanisms,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:f8d7d25ba86aa7449de04683ef923da8be148d622a3c2c31e99816c6a3d336ac

Observation 0ca05e7a-8be0-41c9-bead-4de496e65a48 · outbound

This paper cites Disen- tangled generative causal representation learning,.

Disentanglement-Based Equivariant Learning for Compositional VQA Disen- tangled generative causal representation learning,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:34b7efcd1bdb12e6be208db5c628791e70e0b5a16713795c580c141c3bc9b1e3

Observation 30fb5c3b-0d91-4263-84d5-8caf94e14969 · outbound

This paper cites Equivariant flows: Exact likeli- hood generative learning for symmetric densities,.

Disentanglement-Based Equivariant Learning for Compositional VQA Equivariant flows: Exact likeli- hood generative learning for symmetric densities,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:2d281f43405c065905a4571b0a6fa6211d57a8e20edf589bda6852a03c77acdf

Observation 56c8af20-d85e-4e30-8c9c-b5b6c7797425 · outbound

This paper cites Group equivariant capsule networks,.

Disentanglement-Based Equivariant Learning for Compositional VQA Group equivariant capsule networks,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:94240dbbf0234b7240e26ffd110107cadb5f7022960cabcc2927e54ccfe01760

Observation d4302c82-8bac-44fc-82b1-60650abac4f1 · outbound

This paper cites Learning generalized transformation equivariant representations via autoencoding transforma- tions,.

Disentanglement-Based Equivariant Learning for Compositional VQA Learning generalized transformation equivariant representations via autoencoding transforma- tions,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:e8f058b1b71aa615cc7d93d7f9ace5f4f364bf0fd3051f46ab16d49acd50e348

Observation 16cd8704-c1fd-4a72-8d9a-a44a1b2d5715 · outbound

This paper cites Equivariant and invariant grounding for video question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Equivariant and invariant grounding for video question answering,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:9d40041f144322d94360e11a39df3299e57319ba8b428fcd347a0d77326bd5e0

Observation fa544d7f-30fb-4361-b3f8-1fb9263dc572 · outbound

This paper cites Vilt: Vision-and-language transformer without convolution or region supervision,.

Disentanglement-Based Equivariant Learning for Compositional VQA Vilt: Vision-and-language transformer without convolution or region supervision,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:76c80b8c4c213a5f885efe8cb749e36460510c8c4acf9e872ed68de4b32171b6

Observation 9881197d-0f97-43aa-b7d6-4aeea8839f8f · outbound

This paper cites Invariant Risk Minimization.

Disentanglement-Based Equivariant Learning for Compositional VQA Invariant Risk Minimization

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:46:20.186924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:a187ee07c91a8391763bb878358362d0aef44862cf0502a71d297de8b88a1860

Observation 637be5b8-5187-4b19-8d72-d63f0f60b91d · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Bottom-up and top-down attention for image captioning and visual question answering,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:fc6c0c98a5a8f6fed78a3efed44cae9d35ee38801efd7de11745fbfe102e4895

Observation a7d0607d-2664-4545-848c-4946f2a69d3d · outbound

This paper cites Auto-Encoding Variational Bayes.

Disentanglement-Based Equivariant Learning for Compositional VQA Auto-Encoding Variational Bayes

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:46:20.201937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:cc7d3f606ec723682cccc518c424f5867ae32b9bf1ecb4e2eb935ae26f9c6508

Observation c2c2093a-fb46-4478-8693-e3d95e68f4d4 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization,.

Disentanglement-Based Equivariant Learning for Compositional VQA Grad-cam: Visual explanations from deep networks via gradient-based localization,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:8821e388e3747717d444a78c3fdf07d306aa902b80761da97d2105e50d03b9fb

Observation ddaeb4a4-9767-4c97-aae1-6ce792e38b79 · outbound

This paper cites Clevr: A diagnostic dataset for compositional language and elementary visual reasoning,.

Disentanglement-Based Equivariant Learning for Compositional VQA Clevr: A diagnostic dataset for compositional language and elementary visual reasoning,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:a5475a026ecd6d211553c8b3f2043880249b2e9b15f95e511eb849eb4cefdfe3

Observation b193d3b6-e1ac-4ae8-885a-a4a28387502f · outbound

This paper cites Making the v in vqa matter: Elevating the role of image understanding in visual question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Making the v in vqa matter: Elevating the role of image understanding in visual question answering,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:3fa535fe1c492bf83f57422a76836e06b25d34c22b3fa5abc2dd6030b65c5d98

Observation afa23648-02e5-42d1-b46c-f31b83374864 · outbound

This paper cites Gqa: A new dataset for real-world visual reasoning and compositional question answering,.

Disentanglement-Based Equivariant Learning for Compositional VQA Gqa: A new dataset for real-world visual reasoning and compositional question answering,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:7e8e3e99c3eec7742234e9e708af0d2807a97e888c012728881eb462a611a533

Observation 33fd2797-4764-4720-9ee9-d266fa796a1b · outbound

This paper cites In: Fleet, D., Pajdla, T., Schiele, B., Tuytelaars, T.

Disentanglement-Based Equivariant Learning for Compositional VQA In: Fleet, D., Pajdla, T., Schiele, B., Tuytelaars, T

Reference 58

Resolution
verified exact
doi, observed 2026-06-28T15:02:18.386214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:34e2722615119acfa040936da5a4db5b2b37f693af1200eaf512858f25dba7d3

Observation 41daf745-af95-48c6-a5bd-5de2f9c9c2e8 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Disentanglement-Based Equivariant Learning for Compositional VQA BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:46:20.176301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:30ed24bcddefb4041ef0567f5941e8dd9370f8461d79cfce26989fe7a2860592

Observation 9c4e75a6-2522-421d-b4b1-ddddd8ee115b · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks,.

Disentanglement-Based Equivariant Learning for Compositional VQA Faster r-cnn: Towards real-time object detection with region proposal networks,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:2575afdd8b5353d45d0b565285e287ddc7fba99094bdf56905f8fcb782d73770

Observation 241e80d4-6cf4-4e75-9b52-4583a6893089 · outbound

This paper cites an unresolved cited work.

Disentanglement-Based Equivariant Learning for Compositional VQA Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-06-28T15:00:24.297044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T15:00:24.297044Z digest=sha256:4a4d2d6015af68c148938cc1c25c1d4cbfb088425f4d6cc7505e167175d0c3ea

Pith citing papers

No inbound Pith citation observations are available.