Pith. sign in

Paper Citation Record · LEDGER

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering

As of 8 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2505.19455.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19455 v2

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:18:45.684227Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T11:29:29.185783Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact1
  • verified fuzzy41
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e6ae922a-bab8-4215-b0b6-9dfe69934b24 · outbound

This paper cites VLC-BERT: Visual question answering with contextualized commonsense knowledge.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering VLC-BERT: Visual question answering with contextualized commonsense knowledge

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:58.338712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:39.472572Z digest=sha256:8340a553a311b1fe5a1615219f957b44a18f5ea555b3e90b89dddee507a5cada

Observation ac9a5d60-9325-4594-9b14-5daab540db17 · outbound

This paper cites Align before fuse: Vision and language representation learning with momen- tum distillation.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Align before fuse: Vision and language representation learning with momen- tum distillation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:58.045224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:39.629437Z digest=sha256:6bda2394d37381cb305576806d247a8e593271a5170f276f105eb400a508fe5c

Observation 6fe75be5-1eb4-48a5-84e7-29fca135b16d · outbound

This paper cites Vqacl: A novel visual question answering continual learning setting.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Vqacl: A novel visual question answering continual learning setting

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:57.760674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:39.809754Z digest=sha256:0a7dcb83235cbe66e3f318736dfdd341ec5189e0cdf19ee7c2ad1133ee442ead

Observation 36d15d62-f440-4a59-bac7-880688c382d0 · outbound

This paper cites Enhancing Continual Learning in Visual Question Answering with Modality-Aware Feature Distillation.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Enhancing Continual Learning in Visual Question Answering with Modality-Aware Feature Distillation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:39.958714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:39.958714Z digest=sha256:6f04fbc61fcc30c483b93626ce41a32ee8d411150525dbd797e0c582806896cd

Observation 9d597ad1-f91c-424b-8710-be8908819f64 · outbound

This paper cites Decouple before interact: Multi-modal prompt learning for continual visual question answering.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Decouple before interact: Multi-modal prompt learning for continual visual question answering

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:57.445456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:40.068963Z digest=sha256:28993105562898b0ed4c0077ac0c263ac004a9f2f8658b252abd9c28af1f6db2

Observation cd25b4d6-6e6c-471a-b13d-58b5c22cbb64 · outbound

This paper cites RE- VIVE: Regional visual representation matters in knowledge-based visual question answering.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering RE- VIVE: Regional visual representation matters in knowledge-based visual question answering

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:58.667094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:40.219348Z digest=sha256:bdb7e37edc72566ceb9ed9dd766c1d3404861f82dd388558b14d34639fa464a7

Observation 8c5b0a7f-e4cd-4b7c-999a-0345ea4cb14d · outbound

This paper cites Symbolic replay: Scene graph as prompt for continual learning on VQA task.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Symbolic replay: Scene graph as prompt for continual learning on VQA task

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:57.179424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:40.380500Z digest=sha256:1443e5df7334e07bb0be2cd1bd59da4d1306f42d103ed977cf17467d6b4db410

Observation 9913f3ed-6ddc-4afe-82ca-58238da31f9c · outbound

This paper cites CluMo: Cluster-based Modality Fusion Prompt for Continual Learning in Visual Question Answering.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering CluMo: Cluster-based Modality Fusion Prompt for Continual Learning in Visual Question Answering

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:18:46.155924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:40.495264Z digest=sha256:15cb2a6f4be9b57c9b358672bfc30bc03f4e77e1fd8052e11714e2132de7f180

Observation 3f726f8c-57e2-4b33-8db8-3ed788348b82 · outbound

This paper cites DualPrompt: Complementary prompting for rehearsal-free continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering DualPrompt: Complementary prompting for rehearsal-free continual learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:56.940890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:40.652713Z digest=sha256:dc2ce37859ac51d915e53a9e19f34af7d5a10b2af9081cf3ca49b761932a826a

Observation 43e4801c-2209-4de4-8d8f-b00eaea56e1f · outbound

This paper cites Learning to prompt for continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Learning to prompt for continual learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:56.599667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:40.777953Z digest=sha256:79ccadf44cfeb5d318a128eef3dc5712d4ea2dffde30d8608e71a0724032f08b

Observation 6f9b46ce-0b9f-4b0c-89ae-6c2668842733 · outbound

This paper cites CODA-Prompt: Contin- ual decomposed attention-based prompting for rehearsal-free continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering CODA-Prompt: Contin- ual decomposed attention-based prompting for rehearsal-free continual learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:56.343115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:40.943821Z digest=sha256:1be8b01fe44904866471ff48f9f21f515493eb2bb554a7975d9ff3609cd4a120

Observation 22dffaf8-fe08-4f55-aee0-fb0b0bd2dddf · outbound

This paper cites Semantic residual prompts for continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Semantic residual prompts for continual learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:56.052264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:41.055708Z digest=sha256:a26cc1523235cbbc4e944062a3c527b0d5333111b00cb38f36a5e09ba6071f0e

Observation 6e32edd3-7d65-4cff-867b-0aad9e513f3d · outbound

This paper cites Multi-domain multi- task rehearsal for lifelong learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Multi-domain multi- task rehearsal for lifelong learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:55.799111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:41.235906Z digest=sha256:621a57a25aac7450e970349606069254e46079f39f8bf3902f78494504938c67

Observation ab5d9594-1fa3-4350-9a12-5db1c8d4b6bc · outbound

This paper cites Exploring example influence in continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Exploring example influence in continual learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:55.291288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:41.384164Z digest=sha256:8440f3dd77835b4b9f8c2be276e1bf55d19a38fbe42fd0b11ea401c8fda45eb0

Observation 9b251139-b4cc-4378-9ea7-1bfc4beb2aac · outbound

This paper cites Measuring asymmetric gradient discrepancy in parallel continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Measuring asymmetric gradient discrepancy in parallel continual learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:55.030671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:41.574282Z digest=sha256:39da721c2af1e229ba71e2b9a065f514f4981965fedc90d1e25c4ecfbb02f305

Observation 5db16f86-a349-4ba3-83fa-eb82eecf4be6 · outbound

This paper cites MAPLE: Multi-modal prompt learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering MAPLE: Multi-modal prompt learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:54.732552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:41.733658Z digest=sha256:b1e85fbdbc44a1ecf22a2d0f44b722b482f044da1c0186ce08b02cb162c65155

Observation 43ad371d-77c6-4d4f-a749-2b9d63cc585f · outbound

This paper cites Overcoming language priors in visual question answering with adversarial regularization.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Overcoming language priors in visual question answering with adversarial regularization

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:54.412802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:41.933587Z digest=sha256:6968864cefc1446c36dcf5e19e39cfda3060b62bb948a85052c3742ab1d9673f

Observation 237b70e3-dc83-488d-a892-9b2c00d311a7 · outbound

This paper cites Making the v in VQA matter: Elevating the role of image understanding in visual question answering.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Making the v in VQA matter: Elevating the role of image understanding in visual question answering

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:54.060826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:42.083469Z digest=sha256:183d017b48a76889e523fdda4c391543db5b32f0275752b74ab0199fa4622ff4

Observation 70209ef8-75c9-493f-a996-d6ae11ef9323 · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Lawrence Zitnick, and Devi Parikh

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:53.736161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:42.187884Z digest=sha256:9c9dbcf31d56d582fb1f99b6316f39f68cf444c404d75aab876b2ba14e1f78d1

Observation c80d7d0b-7d58-4398-8070-651ddba2b3c0 · outbound

This paper cites A lifelong learning perspective for mobile robot control.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering A lifelong learning perspective for mobile robot control

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:53.444496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:42.329149Z digest=sha256:7b5c8fe13dc8f36cde457066793f29235de96fa5e61e972e38b195332ff436b5

Observation 5c70ba20-7bea-45a9-8d9e-2e75add3a03d · outbound

This paper cites Pre- venting zero-shot transfer degradation in continual learning of vision-language models.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Pre- venting zero-shot transfer degradation in continual learning of vision-language models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:53.172915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:42.506071Z digest=sha256:70b22525b1e321112e1ea87c28368fea19666270e14cc8483fe1face297bd794

Observation 21d07dd1-800e-4786-85b1-bb83db898b79 · outbound

This paper cites Bakker, Nicu Sebe, and Michael S.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Bakker, Nicu Sebe, and Michael S

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:52.778162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:42.625308Z digest=sha256:c50128a4614970ea2ff9008a885a2bc0748d14829b678c99a1cd5f2b5f38aec7

Observation 6cfdf7fe-11d6-4319-99c8-b6162bfe8e5a · outbound

This paper cites Boosting continual learning of vision-language models via mixture-of-experts adapters.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Boosting continual learning of vision-language models via mixture-of-experts adapters

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:52.553431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:42.809398Z digest=sha256:edc86408f4bdfd9d5a65e777d842c6cb9e079870c23f0142ef7bc07f2f8b65cf

Observation 3e396b1e-cf57-454d-9641-ba791695c299 · outbound

This paper cites Learn to grow: A continual structure learning framework for overcoming catastrophic forgetting.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Learn to grow: A continual structure learning framework for overcoming catastrophic forgetting

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:52.265906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:42.913039Z digest=sha256:42d05f3462376fb01aec6338aae2713ba67187accdd22b53b4ee937c239b8a76

Observation 631255ab-0ad6-4160-b3c4-f0f9169afc88 · outbound

This paper cites Ex- perience replay for continual learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Ex- perience replay for continual learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:51.993430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:43.051597Z digest=sha256:24240f909c8097c2f549a1886c44792b9d1867806bc785323ead708d0084217d

Observation 6f3b90d0-62f3-4fae-8bd0-f18f9d6b6bbc · outbound

This paper cites Dark experience for general continual learning: A strong, simple baseline.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Dark experience for general continual learning: A strong, simple baseline

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:51.689768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:43.197659Z digest=sha256:eb1f7fb09258f1a58726c5efb7f70061d5b75d7c539817b8d0fd462bd6ec17d8

Observation 803b3c9d-a998-4371-8ec5-bca40dc484e1 · outbound

This paper cites A continual learning survey: Defying forgetting in classification tasks.IEEE Transactions on Pattern Analysis and Machine Intelligence, pages 3366–3385, 2021.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering A continual learning survey: Defying forgetting in classification tasks.IEEE Transactions on Pattern Analysis and Machine Intelligence, pages 3366–3385, 2021

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:51.397181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:43.327033Z digest=sha256:6a6df11f7ad93b1f5a42f52a969875745c91dc8c7fdf2b321899e117d3a50c17

Observation c174679f-b047-4d86-921b-77da94b15948 · outbound

This paper cites Yu, and Irwin King.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Yu, and Irwin King

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:43.480038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:43.480038Z digest=sha256:de6a4deb6498714f22cb999359fa928d4ee0151f7683d35cd74baa350fa12753

Observation 743388c0-d14e-45d7-ae35-51a1893b90d8 · outbound

This paper cites Balanced multimodal learning via on-the-fly gradient modulation.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Balanced multimodal learning via on-the-fly gradient modulation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:51.120558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:43.549152Z digest=sha256:f400f595e7c40bad924323168c068a1bf8fcead971a00686e4ecb3c529704d96

Observation cdb1a1a3-d0bf-42d0-a702-8bc1cae5276f · outbound

This paper cites Pre-trained models: Past, present and future.AI Open, pages 225–250, 2021.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Pre-trained models: Past, present and future.AI Open, pages 225–250, 2021

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:50.865018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:43.646275Z digest=sha256:a91d46c16273c827b70b14c8fb5c3daae13ffd508d9973b2dd8d0ecec5305d6b

Observation 86c661f7-5eef-4423-b75d-78c4f288c359 · outbound

This paper cites an unresolved cited work.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:50.550263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:43.718731Z digest=sha256:ce6c25f1e0158a374285a3e48d2101af0f6b1352405d6b57f5baaa064fecbe2e

Observation e3966b13-3e4a-4d4c-850c-bd9290508bb3 · outbound

This paper cites LXMERT: Learning Cross-Modality Encoder Representations from Transformers.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering LXMERT: Learning Cross-Modality Encoder Representations from Transformers

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:43.801881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:43.801881Z digest=sha256:bfb297e5949e821403a6f819b3b5175004114a45046571a5079d92f9ed266697

Observation 955dbc63-ca53-4d0c-aa9a-67732d5d63cb · outbound

This paper cites DyTox: Trans- formers for continual learning with dynamic token expansion.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering DyTox: Trans- formers for continual learning with dynamic token expansion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:50.197997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:43.925124Z digest=sha256:312badb6a702ce79c29b668052a7cb3311512d27303cc6866d316f637a5a5228

Observation 4bff90a1-d510-4dca-94ec-7801bc613390 · outbound

This paper cites Learning transferable visual models from natural language supervision.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Learning transferable visual models from natural language supervision

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:49.892421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:44.029603Z digest=sha256:80187e8dac7255f24319870d0910835f641f0820ed014073e5c9d3fd47dc581a

Observation 62044c85-dfe2-4f54-b81a-d161c3060f30 · outbound

This paper cites Understanding driving risks via prompt learning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Understanding driving risks via prompt learning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:49.565975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:44.129858Z digest=sha256:eb024cab8c0100d3f36cd9390d1759af4d14c926fb53ce8a4d36afa205e5ba99

Observation 8bd7b5b4-bdea-432a-861b-01db032a8db2 · outbound

This paper cites Attention bottlenecks for multimodal fusion.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Attention bottlenecks for multimodal fusion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:49.297569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:44.250330Z digest=sha256:a90c76d58971f9dc093865319552b6bbf5c1b6817e609b4e458dff818699a6c6

Observation 70997690-a4ce-45f6-9772-630cf5f68273 · outbound

This paper cites Difnet: Boosting visual information flow for image captioning.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Difnet: Boosting visual information flow for image captioning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:48.980507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:44.369136Z digest=sha256:2e560015215f94d15e8e292d3fd4e263feaaa317d66a00877561fb8f22d6d799

Observation d494e25d-0efa-4917-8cee-caa28cd4a5f1 · outbound

This paper cites Aligning visual regions and textual concepts for semantic-grounded image representations.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Aligning visual regions and textual concepts for semantic-grounded image representations

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:48.663665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:44.478499Z digest=sha256:84c51c8c069be1c215b010373d1eec26946ff4b57083178c2a9cffaad5c6eb4b

Observation 7bf90daf-328b-4d20-9e9a-b12e51906647 · outbound

This paper cites Masked autoencoders are scalable vision learners.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Masked autoencoders are scalable vision learners

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:48.441359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:44.622402Z digest=sha256:87106d9063e01fdf9ccec4d52280aa5f64750800937aa8a6bb9034e11d2fafab

Observation cb88f53b-56bb-45d3-a44b-db000d993c3a · outbound

This paper cites A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering A Survey on Masked Autoencoder for Self-supervised Learning in Vision and Beyond

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:44.745026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:44.745026Z digest=sha256:a9c32858b248b8975764dfc7b6d05d78d8e2be40828d340ecfc884eb22bef94d

Observation ccaa9d9f-cdee-4e79-bc59-91990f4f9e7f · outbound

This paper cites Can we gain more from orthogonality regularizations in training deep networks? InAdvances in Neural Information Processing Systems (NeurIPS), 2018.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Can we gain more from orthogonality regularizations in training deep networks? InAdvances in Neural Information Processing Systems (NeurIPS), 2018

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:48.119044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:44.854051Z digest=sha256:47bc014b070321bdda6e33e2320f0d962692aecd6c30fda9323fddcc2b24fb54

Observation 96122ca3-70b4-4fb7-9c68-0b641f3927a0 · outbound

This paper cites NExT-QA: Next phase of question answering to explaining temporal actions.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering NExT-QA: Next phase of question answering to explaining temporal actions

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:47.810672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:44.946111Z digest=sha256:81221d8cd57443f77e49efee7be08839da0fbc760eb33f12aa805c495d8a09c5

Observation 72d88bc0-f0aa-417a-91b8-037b2f7322c4 · outbound

This paper cites A simple weight decay can improve generalization.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering A simple weight decay can improve generalization

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:47.551554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:45.088368Z digest=sha256:be101eaf6013cbff778fcc493460e5bc2faa67757b7f24159fb00895145e3ea0

Observation 2e79ac51-df0b-40c6-91ca-cc1e558b8d17 · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Bottom-up and top-down attention for image captioning and visual question answering

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:47.282811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:45.264873Z digest=sha256:24a68cb65140395148e09d0b339d040019e5905df5c01a0e392c455af6397f79

Observation 0aa209b1-8d7c-4312-9310-a85912551606 · outbound

This paper cites Visualizing data using t-SNE.Journal of Machine Learning Research, pages 2579–2605, 2008.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Visualizing data using t-SNE.Journal of Machine Learning Research, pages 2579–2605, 2008

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:46.956737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:45.420900Z digest=sha256:89e2aa0d0ac4eeca4414bedded6b4f9fbd6b99aceacb7c591302a155c8bc3016

Observation 3eb9922d-d7b6-4c2f-9173-1f5e80ce4b86 · outbound

This paper cites Scaling instruction-finetuned language models.Journal of Machine Learning Research, 2024.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Scaling instruction-finetuned language models.Journal of Machine Learning Research, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:46.672087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:45.539148Z digest=sha256:efbfa25e7ea92d51ab29896dec7724a6b9e8307c6811d21c324a1efede3859d6

Observation f0ed1da2-ed8a-4d9b-a2b2-41744b569d0b · outbound

This paper cites Plus”, “Mean Pooling.

MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering Plus”, “Mean Pooling

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:46.411761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:18:45.684227Z digest=sha256:fec9e1fc455c2b55b47a9824ee7eff582b0b24aee51fd39b6b8217d74a3fca30

Pith citing papers

Observation 2d7753f2-fd46-4f6c-a6d5-04b682a174bb · inbound

Group Preference Collapse in Personalized Multimodal Large Language Models cites this paper.

Group Preference Collapse in Personalized Multimodal Large Language Models MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T11:29:29.185783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:29:29.185783Z digest=sha256:dfdfbd1c3a704516daa407d4bb12c3da1c8efdfc73ee58146f6f93be0a4ced02