Pith. sign in

Paper Citation Record · LEDGER

Continual Learning in Vision-Language Models via Aligned Model Merging

As of 10 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 6 inbound Pith citation observations for arXiv:2506.03189.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03189 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:13:44.072901Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T16:19:25.059081Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:38:28.618294Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact2
  • verified fuzzy7
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4ef75ae6-ac9c-4b8f-ba59-f93575baf2f6 · outbound

This paper cites an unresolved cited work.

Continual Learning in Vision-Language Models via Aligned Model Merging Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:46.722563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:42.024266Z digest=sha256:545e2da19b260af906f0f2de533fbabdd6edd1097a7093e5eae25d003d8e7cb2

Observation d49cf55c-dda8-4fe0-b900-f4b74541a716 · outbound

This paper cites Fusing finetuned models for better pretraining.

Continual Learning in Vision-Language Models via Aligned Model Merging Fusing finetuned models for better pretraining

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:42.124288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:13:42.124288Z digest=sha256:2b321cc114620910061a881a2b0a782a65b0062ee909dd5b00dde70e945a2a54

Observation afb968f0-77ff-4ff8-bb26-af2e3d0b0261 · outbound

This paper cites How to Merge Your Multimodal Models Over Time?.

Continual Learning in Vision-Language Models via Aligned Model Merging How to Merge Your Multimodal Models Over Time?

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:13:44.658600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:42.442908Z digest=sha256:7469b4c494bc3045dbe2a88d0fff2d24f18682800d0219d8da12f7996342662e

Observation 4101a6c8-5453-44f8-b066-ae04db996fda · outbound

This paper cites Kembhavi, M.

Continual Learning in Vision-Language Models via Aligned Model Merging Kembhavi, M

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:46.291282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:42.827482Z digest=sha256:35a0a19eebaddfe955b60f9747ec53af2bfec6aed65878c322b44baedac46bae

Observation 0d3ca982-d6ef-4874-9269-ad3aebe56ed6 · outbound

This paper cites Learning Attentional Mixture of LoRAs for Language Model Continual Learning.

Continual Learning in Vision-Language Models via Aligned Model Merging Learning Attentional Mixture of LoRAs for Language Model Continual Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:43.137652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:13:43.137652Z digest=sha256:0c027bc1d391914974b6a3a542057144351a3dd188d71e1b91f2b465f7f4b11f

Observation e4380326-7d8a-488d-8a6c-fefd41b3b9b7 · outbound

This paper cites LoRA Soups: Merging LoRAs for Practical Skill Composition Tasks.

Continual Learning in Vision-Language Models via Aligned Model Merging LoRA Soups: Merging LoRAs for Practical Skill Composition Tasks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:43.254006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:13:43.254006Z digest=sha256:c34ad84ccbc988e2f2e9376499b0446d7f9e9d6c2dd3ff3724553b0691740be7

Observation b5f8f3f7-5271-491b-9174-f91cfe0675f9 · outbound

This paper cites The Non-Local Model Merging Problem: Permutation Symmetries and Variance Collapse.

Continual Learning in Vision-Language Models via Aligned Model Merging The Non-Local Model Merging Problem: Permutation Symmetries and Variance Collapse

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:13:44.334300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:43.501858Z digest=sha256:b13c656f0aa1674e12a3a012939e4f703b1dc6a327f55c25fec4d1b9be5ae8cb

Observation f463f304-8578-4c88-a309-10c4972c1fda · outbound

This paper cites an unresolved cited work.

Continual Learning in Vision-Language Models via Aligned Model Merging Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:45.687330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:43.622560Z digest=sha256:cbafb615953cf2c18645b52b195982eba454adc1d0f508991b4741c66b52b427

Observation f8a668b0-9d71-4fde-8766-903060a4af11 · outbound

This paper cites Parameter-Efficient Fine-Tuning Methods for Pretrained Language Models: A Critical Review and Assessment.

Continual Learning in Vision-Language Models via Aligned Model Merging Parameter-Efficient Fine-Tuning Methods for Pretrained Language Models: A Critical Review and Assessment

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:43.689031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:13:43.689031Z digest=sha256:f88d59c27593eae434f655d376abc14c7aec4f1f9c3aa2350ae43934a0fbcc41

Observation dd4a6e48-87b3-483c-b0b7-775d4fc9568c · outbound

This paper cites We used the best hyperparameters reported by Beyer et al.

Continual Learning in Vision-Language Models via Aligned Model Merging We used the best hyperparameters reported by Beyer et al

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:45.503630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:43.780570Z digest=sha256:a796b3f093509a9a0d1c969ecc77c5a53e178659b1eb8ccbf78ab86741ee4a07

Observation 6fba70f4-e967-47c3-b680-9a3390768025 · outbound

This paper cites Baselines We used a LoRA rank of 32 for all baselines.

Continual Learning in Vision-Language Models via Aligned Model Merging Baselines We used a LoRA rank of 32 for all baselines

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:45.331014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:43.865608Z digest=sha256:2c9954bcda64ea49c2bc83b27df72d8e089bd23024ceb26816fe5c48d86fba02

Observation 28542678-6e95-4e94-80a8-4911dd39e937 · outbound

This paper cites At each time step, we report the average forgetting on seen tasks so far (lower is better).

Continual Learning in Vision-Language Models via Aligned Model Merging At each time step, we report the average forgetting on seen tasks so far (lower is better)

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:45.161725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:43.949613Z digest=sha256:0b633b85107712a2723a7bf008f51032f9f6fc0fb9c53239b574baa3240dce19

Observation 9994f912-b970-4cd3-bdd1-93d17180029a · outbound

This paper cites In contrast, merging demonstrates greater robustness to these task variations.

Continual Learning in Vision-Language Models via Aligned Model Merging In contrast, merging demonstrates greater robustness to these task variations

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:44.927081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:44.072901Z digest=sha256:7c612ccf576d626252bd1b76fa051fb8c5f3a706d315de6d0185bde489edd3e9

Observation 5aa130d6-3a08-4964-a631-0b732ee1575a · outbound

This paper cites Soup to go: mitigating forgetting during continual learning with model averaging.

Continual Learning in Vision-Language Models via Aligned Model Merging Soup to go: mitigating forgetting during continual learning with model averaging

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:42.925177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:13:42.925177Z digest=sha256:140114231a487742e9c467834ce748c2dc4d6f3319c384c2d04b274af2ff27ea

Observation 6fe3712f-fcfb-429f-a207-fee8c2ba982f · outbound

This paper cites Don’tforget, thereismorethanforgetting: new metrics for continual learning.

Continual Learning in Vision-Language Models via Aligned Model Merging Don’tforget, thereismorethanforgetting: new metrics for continual learning

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:46.507830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:42.346481Z digest=sha256:75d61097c8107a3b4bd1dd540a8049d12d0b69e164f99ccc4bc848daf878fc34

Observation 9175bb53-8694-447a-87aa-e80a26b1b2d9 · outbound

This paper cites Editing Models with Task Arithmetic.

Continual Learning in Vision-Language Models via Aligned Model Merging Editing Models with Task Arithmetic

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:42.567852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:13:42.567852Z digest=sha256:5a4f31a9dd32825cc16e8c746f498c7b40a139e53d9e5206223fa26e50364369

Observation 18f32e32-86e2-459f-97d7-219fd7d64e2f · outbound

This paper cites On Tiny Episodic Memories in Continual Learning.

Continual Learning in Vision-Language Models via Aligned Model Merging On Tiny Episodic Memories in Continual Learning

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:41.939706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:13:41.939706Z digest=sha256:e2dd18798deab9c307dd711fbc30525644f05452c7a4541e54445b098aaf690a

Observation 79d3fe60-aae9-4339-8372-9a6011244763 · outbound

This paper cites icarl: Incrementalclassifierandrepresentation learning.

Continual Learning in Vision-Language Models via Aligned Model Merging icarl: Incrementalclassifierandrepresentation learning

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:13:45.874169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:43.381228Z digest=sha256:8a61b6623e7781a233ac4cb611f307ed64720d097ac3cd691597e061184735db

Observation 5d3df987-0317-48bb-b9f8-4349e98f1c8e · outbound

This paper cites Averaging Weights Leads to Wider Optima and Better Generalization.

Continual Learning in Vision-Language Models via Aligned Model Merging Averaging Weights Leads to Wider Optima and Better Generalization

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:42.702115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:13:42.702115Z digest=sha256:5a6b532e876cc7b891f1a3c32c94564aa1fd798baf3347e109a21a1629c44353

Observation f80782f5-82fe-415d-a6af-19b27362809e · outbound

This paper cites an unresolved cited work.

Continual Learning in Vision-Language Models via Aligned Model Merging Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:46.071657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:43.035247Z digest=sha256:1d3b255a5b82b4522b1dbd3dbe1ea6cfcac49d57eff9c4467fe54ad8855a6364

Observation 0e4d076c-1cde-4c4f-9839-c86e4fd7d445 · outbound

This paper cites an unresolved cited work.

Continual Learning in Vision-Language Models via Aligned Model Merging Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:13:46.905472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T12:13:41.874808Z digest=sha256:16bf38e282d2df93a7c25a716d9c168e4ff9d6b6c613e6b4d1240e8a754db45c

Observation e5093490-477b-4a7f-b7f4-8003fd429957 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Continual Learning in Vision-Language Models via Aligned Model Merging BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:42.219343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:13:42.219343Z digest=sha256:ace02c80f2e3cce211c1fc5fd23853aca472cfe902ce80e08ee42dd748291bbe

Pith citing papers

Observation 9f4fa37b-6fc2-4e6c-ac28-0de2d49d3eb6 · inbound

Task Alignment: A Simple Proxy for Practical Model Merging Across Diverse Vision Tasks cites this paper.

Task Alignment: A Simple Proxy for Practical Model Merging Across Diverse Vision Tasks Continual Learning in Vision-Language Models via Aligned Model Merging

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:26:02.204702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T14:56:04.775995Z digest=sha256:73dda2147bdb5d1dd7758b8d93eea14276a34e4aa8724932644afcb6762e4b3f

Observation a43963e3-ac4d-4175-a0c6-f3ca4611686e · inbound

Task Alignment: A Simple Proxy for Practical Model Merging Across Diverse Vision Tasks cites this paper.

Task Alignment: A Simple Proxy for Practical Model Merging Across Diverse Vision Tasks Continual Learning in Vision-Language Models via Aligned Model Merging

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T16:19:25.059081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:19:25.059081Z digest=sha256:69143b50a12ffad1e330f6b7b01820ca96a7510b1cfceeeb112c05c6d48a8693

Observation f44fb004-3313-4fb6-a032-aa6675021ae4 · inbound

ORBIT: Preserving Foundational Language Capabilities in GenRetrieval via Origin-Regulated Merging cites this paper.

ORBIT: Preserving Foundational Language Capabilities in GenRetrieval via Origin-Regulated Merging Continual Learning in Vision-Language Models via Aligned Model Merging

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:12:18.078464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T05:07:59.947540Z digest=sha256:c90c86ee767a048bc2e4df9230c41c33a71ee2cb4d01043d1d564547c5e3d0d3

Observation 3699b2d5-7e38-4fc3-8200-0f92c437a7cb · inbound

Amnesia: A Stealthy Replay Attack on Continual Learning Dreams cites this paper.

Amnesia: A Stealthy Replay Attack on Continual Learning Dreams Continual Learning in Vision-Language Models via Aligned Model Merging

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T12:18:06.672725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T09:02:44.548488Z digest=sha256:a8d18d918e86a04da24161d5b0cef6a1037c9105a8d99a4c3a045b0193cc018e

Observation b76f2e21-89fb-4c48-af4c-2fce9a9d651b · inbound

Unlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction Tuning cites this paper.

Unlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction Tuning Continual Learning in Vision-Language Models via Aligned Model Merging

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:38:28.619687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T14:35:17.004680Z digest=sha256:83b633c41ad3668ead52b3fe5d6384dc1c8ef7a27dae8310e884d498262dc02a

Observation 58294f10-753f-46e2-895f-e3ea0d455fe2 · inbound

Towards Human-like Physical Intelligence: Lifelong Vision-Language-Action Learning for Robotic Manipulation cites this paper.

Towards Human-like Physical Intelligence: Lifelong Vision-Language-Action Learning for Robotic Manipulation Continual Learning in Vision-Language Models via Aligned Model Merging

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T00:57:43.881234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:57:43.881234Z digest=sha256:2ecde6fc3cc4738f35a3c93ad47fd3de68a43a22c3361710046d888f8a07d04b