Pith. sign in

Paper Citation Record · LEDGER

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation

As of 11 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2604.16943.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.16943 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T07:12:16.623769Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact10
  • verified fuzzy9
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eaa92d63-5fb4-48e9-b6ff-2cb55a721c12 · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Instructblip: Towards general-purpose vision-language models with instruction tuning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T14:20:14.945477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:7dd9c6e8f79ddedddecd67d8cdb6c4c3974f0afcc90b809a028398336a174b11

Observation 573b71ee-ea35-48bb-b188-0c73dde5c942 · outbound

This paper cites Multimodal large language models: A survey.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Multimodal large language models: A survey

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T14:20:14.937021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:81a8b5a3e9b1355b2d19f64aee16b81034471aec25eaf7ead04a7f76715f970a

Observation b6e6431c-2e2f-4f34-96c3-f18df788a4f5 · outbound

This paper cites MM-LLMs: Recent Advances in MultiModal Large Language Models.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation MM-LLMs: Recent Advances in MultiModal Large Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.427461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:afc05f0b7ed0655510d2aa9e412f2da5f68bb174d9356dbb5668e0bc9d16fa7e

Observation 661aa85c-4f25-4c2d-99f3-15deae39d0c1 · outbound

This paper cites Exploring Better Text Image Translation with Multimodal Codebook.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Exploring Better Text Image Translation with Multimodal Codebook

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.446018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:23b3d03acc870e21c2b547974f9c5439630b222252cb8342ab3c37690ca6c67b

Observation 1b734d92-fc0f-44c2-8006-7c00ed20208e · outbound

This paper cites Automatic detection and translation of text from natural scenes.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Automatic detection and translation of text from natural scenes

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T14:20:14.940192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:b6190ee12d0e7cb9635c64c11a457489ca10e9df998ae9d1e751e631d79c4bbb

Observation 34170d40-a9b7-4438-9860-d6099cb9d404 · outbound

This paper cites Translatotron-V(ison): An end-to-end model for in-image machine translation.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Translatotron-V(ison): An end-to-end model for in-image machine translation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T14:20:14.948906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:f0551d2b00b22d7dd56abe7ce3e01fea732f9463b6ebec825c561f89d0a71511

Observation 055d1fbc-3956-47dd-b87c-4c24aef14ad0 · outbound

This paper cites Towards End-to-End In-Image Neural Machine Translation.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Towards End-to-End In-Image Neural Machine Translation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.436049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:fe819cd002e8167684f6b27f9cb9628c865b00fed646cd26d5331ee314274357

Observation df6614d4-c203-44fe-81d3-ebd20c87eef5 · outbound

This paper cites Image translation network.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Image translation network

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T14:20:14.952037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:c91e509d84f4f378d887b4c0d62f684f02be7fa3a4eb26205936d6e84e1fdbe0

Observation 0907ac0e-8eed-457c-8f03-42856b00a036 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:48:40.999279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:52a6f4252ed6ccd4da667e3fdac08daed6b1a4f7806675e14008c6e206ec4b7f

Observation f13d4557-bb29-4696-8f52-8c33e522407a · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T21:07:32.698795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:fe9cf117b18035d24a5bffcfa17527cd47537661a9c0ba7fec8919f54bfcec0b

Observation 64cdb28d-b052-4988-9eeb-bff080f4ade8 · outbound

This paper cites Qwen2.5-VL Technical Report.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Qwen2.5-VL Technical Report

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-10T09:23:37.439386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:2a8758c48860478af67e8a131b8a50765cabcf659dabc5c7c60595bc45b7bef9

Observation de19b0ad-7ace-4d58-9469-e0103037757e · outbound

This paper cites Finetuned Language Models Are Zero-Shot Learners.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Finetuned Language Models Are Zero-Shot Learners

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T21:14:13.786069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:cfa90c6ed4cbfc1af132911fa4d9a95c4ae8059d6935afde03013b5fc28aa889

Observation 0cea1666-a462-4ebe-81ca-c49d979914e4 · outbound

This paper cites An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation An Empirical Study of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:26:42.058060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:89deae5f442c6f2dcaf492aae18c06b0e991a1c45ced7d2a63cfd5c51b261cf0

Observation 2bc504f4-aa6b-49e2-bb2d-e946d3ca3ad4 · outbound

This paper cites E^2VPT: An Effective and Efficient Approach for Visual Prompt Tuning.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation E^2VPT: An Effective and Efficient Approach for Visual Prompt Tuning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.452774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:7eedf2b657ec1adb3f9d74ff0661049a7340b575ac96d38856c17b92901e65c3

Observation 7e350dba-13cc-41db-8583-2748973f4443 · outbound

This paper cites M 2PT: Multimodal prompt tuning for zero-shot instruction learning.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation M 2PT: Multimodal prompt tuning for zero-shot instruction learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T14:20:14.925339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:83652dc1ddf2b21a57abbd0cf3d9b0edabe5a808645ed6a52b986197c4760cdd

Observation 047ce859-858b-4f24-978f-92140fa43540 · outbound

This paper cites Pruning Convolutional Neural Networks for Resource Efficient Inference.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Pruning Convolutional Neural Networks for Resource Efficient Inference

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.449194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:58ced03d42a17c01bab22976b21ae148fd1a029a8c9894eda12cab8e18651946

Observation 7d5e5d7c-8527-48b9-8da0-983cebe4a5ca · outbound

This paper cites Overview of the iwslt 2017 evaluation campaign.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Overview of the iwslt 2017 evaluation campaign

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T14:20:14.933623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:1b77d31ff26d37b73dbf0361b35cd50f4001e46408ffdab4bd89d541a5e69ddc

Observation da7f256d-2039-4d4f-90b2-48f46148ff10 · outbound

This paper cites PP-OCRv3: More Attempts for the Improvement of Ultra Lightweight OCR System.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation PP-OCRv3: More Attempts for the Improvement of Ultra Lightweight OCR System

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.456021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:412e5338d2c569243e83c5ae76375191d97447a14b6318d9a3398c80097da9fa

Observation 6f52da4d-e593-4746-85ed-9dec2de6af48 · outbound

This paper cites UMTIT: Unifying recognition, translation, and generation for multimodal text image translation.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation UMTIT: Unifying recognition, translation, and generation for multimodal text image translation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T14:20:14.921984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:0174859abca3723ad4009add00ba4c05e880876b73939e5f1d470a6cbeb8a75e

Observation 835c528e-93a7-4ed8-bfb1-14f51ae5d7ed · outbound

This paper cites Document image machine translation with dynamic multi-pre-trained models assembling.

MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translation Document image machine translation with dynamic multi-pre-trained models assembling

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T14:20:14.928521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:12:16.623769Z digest=sha256:1fba460c08e2334126e9a2a3186c4fc4d9346b78e81013cda751af022e62a9aa

Pith citing papers

No inbound Pith citation observations are available.