Pith. sign in

Paper Citation Record · LEDGER

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models

As of 8 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 0 inbound Pith citation observations for arXiv:2508.00260.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.00260 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:21:21.500180Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

55 of 55 outbound references displayed

  • verified exact1
  • verified fuzzy39
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a2cd95b2-bd47-4f87-863c-e5546ad593d3 · outbound

This paper cites GPT-4 Technical Report.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.310768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.310768Z digest=sha256:02fdef488e87c9e6974c8a87072e7ae1f403eca897beb2a2fcfaffb8507663a3

Observation 8c6088e5-da7c-4619-bf4c-78bc5c5ff8be · outbound

This paper cites Expert gate: Lifelong learning with a network of experts.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Expert gate: Lifelong learning with a network of experts

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:22.067100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.315076Z digest=sha256:b63f390abe9e856d927c79944ef828aeaa9b5979c473ea297fc7534dab86732c

Observation 99ba7021-4be1-4c8b-ab97-3ef594dff23e · outbound

This paper cites Memory aware synapses: Learning what (not) to forget.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Memory aware synapses: Learning what (not) to forget

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:22.057628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.318768Z digest=sha256:d3330df40fbb88e013eb5c8fcee2b0eea931e13b8db78fbd72a307e658e4e877

Observation f62ad9ad-b7c2-4400-94b4-e6169e020360 · outbound

This paper cites Generative multi-modal models are good class incremental learners.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Generative multi-modal models are good class incremental learners

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:22.047700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.322689Z digest=sha256:dd4533af77d6f7941f0ca25e7a585c47ff983eaa813cda4f858c8d99f768839f

Observation 8e8e1114-3b47-4343-a4fc-008ddfec2dd0 · outbound

This paper cites Honeybee: Locality-enhanced projector for multimodal llm.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Honeybee: Locality-enhanced projector for multimodal llm

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:22.037476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.326048Z digest=sha256:98075cf13a83812dcab546ee950f806dfee42b9c03bbb67a3608801768ab4462

Observation 570d82c6-cebe-42ef-a1e6-21476761a094 · outbound

This paper cites Riemannian walk for incremen- tal learning: Understanding forgetting and intransigence.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Riemannian walk for incremen- tal learning: Understanding forgetting and intransigence

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:22.027727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.329716Z digest=sha256:c34018c2215ff6017a5a7d75f24d6c01a52ae8f9bce454845fda04e60a0ac840

Observation 30788f19-cb92-480b-8fbd-a5ac6691d02f · outbound

This paper cites CoIN: A benchmark of continual instruction tuning for multimodel large language models.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models CoIN: A benchmark of continual instruction tuning for multimodel large language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:22.017546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.333291Z digest=sha256:2387c856a0975cbdc1e1a9c7da0192c9444e28332921a931d191ecbfbab5ede7

Observation dd7b02c6-1581-4797-8a2c-1351ec226a56 · outbound

This paper cites Lifelong language pretraining with distribution-specialized experts.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Lifelong language pretraining with distribution-specialized experts

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:22.007441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.336761Z digest=sha256:fa52eb9af0a396b85b1836f6bba1de8d6a7245d84d789dbe6fe6827c33fe41c4

Observation 81a25914-8eaa-4039-a0dc-17fafc13e101 · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality, march 2023.URL https://lmsys.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality, march 2023.URL https://lmsys

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.997478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.340156Z digest=sha256:3afdcfd73d5889787ba423c1bb804c9b6215f3114e81b893d5ca5e0165d41d5e

Observation 945b7b39-0945-4280-a61a-bdaa7a41fb4d · outbound

This paper cites V ocabulary-free image classification.Advances in Neural Information Processing Systems, 36:30662–30680, 2023.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models V ocabulary-free image classification.Advances in Neural Information Processing Systems, 36:30662–30680, 2023

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.987564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.343385Z digest=sha256:33a339211f36ff9f3ae09f6761a7de491d01c08da4ba626fa2611c9ede80efc3

Observation 2ab47047-23b1-4284-bff6-c519341f81e7 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.347451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.347451Z digest=sha256:7a87ee6658079cb5586b91012cdf4a75c9fafb9c23cfce163991b10868e6f5ca

Observation 06a9d7c0-b14e-4747-883b-cdd02f613800 · outbound

This paper cites RATT: Recurrent attention to transient tasks for continual image captioning.Advances in Neural Information Processing Systems, 33:16736–16748,.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models RATT: Recurrent attention to transient tasks for continual image captioning.Advances in Neural Information Processing Systems, 33:16736–16748,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.976443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.351302Z digest=sha256:79b787a947c31bd71e0f8c2ab1e403e59d7c479d5e6e96e633be1e9d4d8ff32b

Observation a6ea476f-f29f-493b-a247-ed1bf7a12bc6 · outbound

This paper cites DyTox: Transformers for continual learning with dynamic token expansion.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models DyTox: Transformers for continual learning with dynamic token expansion

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.965024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.355162Z digest=sha256:2e071e1548d8306986c5419914457ec4ffcd57e6c5968881262c88b31c0953eb

Observation 45fca5ca-69f9-48be-8db1-23abe7e24f42 · outbound

This paper cites EV A: Exploring the limits of masked visual representa- tion learning at scale.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models EV A: Exploring the limits of masked visual representa- tion learning at scale

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.954607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.359574Z digest=sha256:8b175f07572b06d0625cb419a8537b99a5a928888ab80ecb2f6b968ccb5498f2

Observation b22df089-9ead-47e1-a943-6e6380b1d531 · outbound

This paper cites Beyond prompt learning: Continual adapter for efficient rehearsal-free continual learning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Beyond prompt learning: Continual adapter for efficient rehearsal-free continual learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.944646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.363466Z digest=sha256:71cddaab9d68058750f2acb1b91296c26b12f1b7a5e37bb59337cd65e2d934d9

Observation 02372aff-844f-4e71-8370-72ec0917cdeb · outbound

This paper cites Continual Instruction Tuning for Large Multimodal Models.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Continual Instruction Tuning for Large Multimodal Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.366736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.366736Z digest=sha256:379b00db7e4513e237485149b84e904a5dc294258c0d19c3199ab038fe6953bb

Observation 0bee0757-8a2a-4ee7-a506-1be9dca4f475 · outbound

This paper cites The many faces of robust- ness: A critical analysis of out-of-distribution generalization.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models The many faces of robust- ness: A critical analysis of out-of-distribution generalization

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.934456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.370202Z digest=sha256:59fcadb7dfaca0bead5c1fec88c7d0f27675a7de3cc85266664dddb5a7a81064

Observation b20eeacf-6977-4ed0-ad91-0c2f538a8ce3 · outbound

This paper cites CLIPScore: A reference-free evaluation metric for image captioning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models CLIPScore: A reference-free evaluation metric for image captioning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.924153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.373331Z digest=sha256:10aab9290bb4a5c22592d38f11c664bb8c9d7606e9dbef03f94dcfb76ed91fd1

Observation 3bdd2268-3957-48f4-88f7-8181ab064744 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models LoRA: Low-Rank Adaptation of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.376708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.376708Z digest=sha256:bb65d1c9d65c3f23628d68236d3c27b682effa88565efbd6bbf395c458886b9a

Observation 5326a5b6-36e1-4a81-92ee-3c77860b2220 · outbound

This paper cites Adaptive mixtures of local experts.Neu- ral computation, 3(1):79–87, 1991.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Adaptive mixtures of local experts.Neu- ral computation, 3(1):79–87, 1991

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.380456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.380456Z digest=sha256:43b4534577f943514e667676cf37f01821dca77c21f1035b961b480cf9ea0786

Observation 8a380d33-a13b-4e35-94f8-41dbb652a36c · outbound

This paper cites Vi- sual prompt tuning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Vi- sual prompt tuning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.908229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.384416Z digest=sha256:0818f0805b3e151bd5ea811a299468cc16450c1b624eeca4eee9b42134b79783

Observation 4046f619-d310-4c43-916e-7e154fa85702 · outbound

This paper cites Helpful or harmful: Inter- task association in continual learning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Helpful or harmful: Inter- task association in continual learning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.898082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.387520Z digest=sha256:8600f9cd056aeca7df7b94a2567e64e7984cd9947409c618971f979bf89e27a4

Observation c2627989-9973-45ca-8420-7fcadd3372a3 · outbound

This paper cites Growing a brain with sparsity-inducing genera- tion for continual learning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Growing a brain with sparsity-inducing genera- tion for continual learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.888147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.391038Z digest=sha256:fef8149d2bb84671a7f52f4b9f081ea64554468d3cdfcf8b53aaca07a4bca79c

Observation 263528cf-4529-4fb3-92b3-b84b4b146685 · outbound

This paper cites Deep visual-semantic align- ments for generating image descriptions.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Deep visual-semantic align- ments for generating image descriptions

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.877582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.394628Z digest=sha256:83bfb595ff3b127c20b8f8754a3fb697749fa9c21b18546de6748e46afab3e52

Observation 6ad4d4fa-f5d2-4bd1-a32d-e55d25b24242 · outbound

This paper cites Overcoming catastrophic forgetting in neu- ral networks.Proceedings of the National Academy of Sci- ences of the United States of America, 114(13):3521–3526,.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Overcoming catastrophic forgetting in neu- ral networks.Proceedings of the National Academy of Sci- ences of the United States of America, 114(13):3521–3526,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.867258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.397882Z digest=sha256:adff93fcbaf64be757d164fcba78ceed476a1f6f13d27b16e1914afff2e58f0a

Observation 2d94a6d3-33dc-493a-b215-ac027f494ba4 · outbound

This paper cites Quantifying the Carbon Emissions of Machine Learning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Quantifying the Carbon Emissions of Machine Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.401198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.401198Z digest=sha256:7f53a0da55496038e37c2af4f400eead70586ffeb345f09618803145b0d771d8

Observation e97382b7-74b5-40ff-9acb-2aa6f612a842 · outbound

This paper cites GShard: Scaling giant models with conditional computation and automatic sharding.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models GShard: Scaling giant models with conditional computation and automatic sharding

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.857066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.404660Z digest=sha256:3a64b0e466db09b53c156f33ea48713e5c77012d02f28ae001b9e7a7f857fad7

Observation e1280ef7-e5aa-4e2e-afc9-d1a6581f1682 · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.408225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.408225Z digest=sha256:5d05a5e3615e9ccbc75e9d3e99cc9884d9d09d9a12af3e72bfe78406d88cddfb

Observation ba8b8395-1e75-4d1a-b2e2-4241ffd20f49 · outbound

This paper cites BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.845884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.411788Z digest=sha256:a6bd6955b0019c34b047865e526c9534e7787b2b42bfad6fd1dbef291f26e71a

Observation 4f8163e8-6e42-49a3-ae1f-3b14c7fd1d8b · outbound

This paper cites Learning without forgetting.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Learning without forgetting

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.835870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.414960Z digest=sha256:ceaad567def0e0e445f1052894eaa5bdf786cd639e25923852e02c2f84719f50

Observation 3ea7bd6f-3ff7-43e2-baac-83a22dff776b · outbound

This paper cites InfLoRA: Interference-free low-rank adaptation for continual learning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models InfLoRA: Interference-free low-rank adaptation for continual learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.825621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.418162Z digest=sha256:39b6ee6639bc11b2f282c8777198fd476d1fe42101bd4be7cdae12710c0ba66e

Observation 444189ce-f650-444a-8d50-5e51921f735a · outbound

This paper cites Improved baselines with visual instruction tuning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Improved baselines with visual instruction tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.421578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.421578Z digest=sha256:a7775631a38a25437bdf1732c2dbba07cc79892f1a73e92787c60ee78e39e63d

Observation 1e4fc8c2-bcee-4400-86e1-ce01a0297c89 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36, 2024.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Visual instruction tuning.Advances in neural information processing systems, 36, 2024

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.424576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.424576Z digest=sha256:44560a32ed4e79486691c490b55ce815ed07ed7b5060f9086fd7a3a43ed67d8a

Observation 2d6433d5-c150-47ed-b400-6d06c5ff7eab · outbound

This paper cites Adaptive aggregation networks for class-incremental learning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Adaptive aggregation networks for class-incremental learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.802726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.427663Z digest=sha256:bec8c6f1e1af8e3b238a82d0a99d8cef2afebdca64b44abf31d25d9991d5318a

Observation 11518c7f-6194-43b9-b38f-a2c4eae7dfa6 · outbound

This paper cites Unified-IO 2: Scaling autoregressive multimodal models with vision language audio and action.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Unified-IO 2: Scaling autoregressive multimodal models with vision language audio and action

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.791017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.430584Z digest=sha256:04d687e01d6a5cddb07b919801b8f8425de82d04ee10d02d724799be52e135e2

Observation 61fee644-8cda-467b-9d08-33eb3dd9dfbb · outbound

This paper cites Not all ex- perts are equal: Efficient expert pruning and skipping for mixture-of-experts large language models.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Not all ex- perts are equal: Efficient expert pruning and skipping for mixture-of-experts large language models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.779647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.433965Z digest=sha256:0703fb65e2de457acdcb8c6b864c41005f094e2762e4ac31dd45a8431ad7d3eb

Observation cf2acef6-ec07-4ca3-a677-6d4a9ef772d6 · outbound

This paper cites Catastrophic inter- ference in connectionist networks: The sequential learning problem.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Catastrophic inter- ference in connectionist networks: The sequential learning problem

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.768403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.437007Z digest=sha256:f54eed2ddb9506622c52da8f410e55112a8bc527eaa790fceafd9db24f547033

Observation 512af70e-e130-4b8b-a7f1-7c739dff4727 · outbound

This paper cites Multimodal contrastive learn- ing with limoe: the language-image mixture of experts.Ad- vances in Neural Information Processing Systems, 35:9564– 9576, 2022.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Multimodal contrastive learn- ing with limoe: the language-image mixture of experts.Ad- vances in Neural Information Processing Systems, 35:9564– 9576, 2022

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.758311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.440162Z digest=sha256:42d81da2314a32e766faad2189ca32d76e9f0ed6f9ff8667f26cff2bd5d16350

Observation c68e684d-ac33-46b2-aa90-aa2bb7345d87 · outbound

This paper cites Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models.Interna- tional Journal of Computer Vision, 123:74–93, 2017.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models.Interna- tional Journal of Computer Vision, 123:74–93, 2017

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.748323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.443261Z digest=sha256:c16b045b27218577095802789cc4596153fe250e83b615935c2fa2daaf2c4b8a

Observation fa114b31-03a8-4d3c-bf60-47870e7f9285 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Learning transferable visual models from natural language supervi- sion

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.446302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.446302Z digest=sha256:5a53da631e390eade1ad2bd83e0684d02b3d6d9379d72b9b55004c788d8505f6

Observation 7dc32692-beca-432e-8e4b-17fde4ca2554 · outbound

This paper cites iCaRL: Incremental clas- sifier and representation learning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models iCaRL: Incremental clas- sifier and representation learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.730959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.449652Z digest=sha256:40f37707985eab5d5a6884740a10e7cfc65b6fa8c46b437fc5f7b355d4ced3a9

Observation b83644da-00c9-4940-8e7b-e4593c362396 · outbound

This paper cites Exploring models and data for image question answering.Advances in neural information processing systems, 28, 2015.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Exploring models and data for image question answering.Advances in neural information processing systems, 28, 2015

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.720656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.452679Z digest=sha256:2e08a64331c0981680ca93e809a54974cca9548379501c385d7dfa92e3bdf95e

Observation aa84c749-4020-40f3-b455-7c2192d7ba52 · outbound

This paper cites Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.455817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.455817Z digest=sha256:69a51adc8b0f41a5def05110033791b6073fbf9e7223077dc44dcd8a9d7e7d38

Observation 90f5229b-c3f4-48f0-8f8f-30b2dbec25f2 · outbound

This paper cites CODA-prompt: Con- tinual decomposed attention-based prompting for rehearsal- free continual learning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models CODA-prompt: Con- tinual decomposed attention-based prompting for rehearsal- free continual learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.710769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.459384Z digest=sha256:3376edf0c8bd5c44c921a2eaa727a72283404e8c70c197b627f8afef280d3dc3

Observation b5ad4abf-dce0-4b27-9427-29e05c0955ae · outbound

This paper cites Constrained contrastive distribution learning for unsupervised anomaly detection and localisation in med- ical images.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Constrained contrastive distribution learning for unsupervised anomaly detection and localisation in med- ical images

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.700593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.463488Z digest=sha256:975297bc6db660f1fa549366e115c79a934c15ea7d28b12426397e20a2ab6ddc

Observation 1655feca-8306-4c3c-b6c9-47273b294f71 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.467333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.467333Z digest=sha256:22db8ff12cc7b45c62724d7c786ce3d72d7ca0e7cac7b215ca2140df39d25d2d

Observation 3da1cda1-5e60-4ce3-85fe-caed156eca0f · outbound

This paper cites DualPrompt: Complementary prompting for rehearsal-free continual learning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models DualPrompt: Complementary prompting for rehearsal-free continual learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.690408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.470608Z digest=sha256:791b81cadf35766675c60fbd561c86c1e8f6b84ff484e9e46f367004780056d5

Observation e34c4a4b-b655-4c20-8fb9-5dba204c4ec0 · outbound

This paper cites Learning to prompt for con- tinual learning.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Learning to prompt for con- tinual learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.474330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.474330Z digest=sha256:2e5de860479a7c6bd2c9f4e59b9459ae0129d66dff2018e5f713a88f84566961

Observation 85a0656a-f59d-4f0a-96d5-9010d31056c4 · outbound

This paper cites mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.478032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.478032Z digest=sha256:b16c3c0eadac8fa936f4317d118c48eb5f66352fcb4d14cee751d26a65261c42

Observation 0e2c69be-6c75-40ee-a6d0-ba572582c434 · outbound

This paper cites Boosting continual learning of vision-language models via mixture-of-experts adapters.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Boosting continual learning of vision-language models via mixture-of-experts adapters

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.674647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.481597Z digest=sha256:ef4883290382877ad58f705ddb215fc260b64e8f634f4f3a75122bd9a596fc00

Observation 15c32d4d-2942-4925-a2f2-e79a44d2b491 · outbound

This paper cites Contin- ual learning through synaptic intelligence.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Contin- ual learning through synaptic intelligence

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.664041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.485352Z digest=sha256:9a1698c1c84b3af18440eda9454c8ede31376ce034a55851c650f65704028bf5

Observation 34554ce1-1655-4f89-9000-5f55e38d0500 · outbound

This paper cites Investigating the catastrophic for- getting in multimodal large language models.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Investigating the catastrophic for- getting in multimodal large language models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.653638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.489745Z digest=sha256:cac7389d3ac4d28ec01f928dee320b5d802bbaf458958d411db2ddf96a1bafdb

Observation 128e48f9-fff9-4e9f-b946-457d78234640 · outbound

This paper cites Prompt-Aware Adapter: Towards Learning Adaptive Visual Tokens for Multimodal Large Language Models.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Prompt-Aware Adapter: Towards Learning Adaptive Visual Tokens for Multimodal Large Language Models

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:21:21.546029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.493130Z digest=sha256:cf6fd4078a40e0f6493980a924c9b8d09159579bc2395a3629b0e05df674595f

Observation 7893dd5b-5c4d-451e-a26d-18c1465b73f8 · outbound

This paper cites Preventing zero-shot transfer degradation in continual learning of vision-language models.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models Preventing zero-shot transfer degradation in continual learning of vision-language models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:21:21.642002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:21:21.496493Z digest=sha256:b029a19be2252dbad52e4a936d7bad52a18927095db1f29cb86b1122d7e0e0bd

Observation 904ad81f-2db6-4e47-bd45-2d68bd2ff969 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Instruction-Grounded Visual Projectors for Continual Learning of Generative Vision-Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T10:21:21.500180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:21:21.500180Z digest=sha256:1b1c3b636710765d081bf60cc2d92b4d7d4e4d695190c48b05aa6463b67d6deb

Pith citing papers

No inbound Pith citation observations are available.