Pith. sign in

Paper Citation Record · LEDGER

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation

As of 9 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2607.02593.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.02593 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T09:26:29.033271Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 58d73325-0b48-468c-a3a1-ed1744c875e0 · outbound

This paper cites In: The Twelfth International Conference on Learning Representations (2024) 2.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The Twelfth International Conference on Learning Representations (2024) 2

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:25cd65e74bbab8f62b89d0a78d1553a9f3a4c4e24f1b93ef3ec42412636575d7

Observation c6e331ac-e63a-44d1-b919-665aff10ba00 · outbound

This paper cites BD-KD: Balancing the Divergences for Online Knowledge Distillation.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation BD-KD: Balancing the Divergences for Online Knowledge Distillation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:754497eb9f7133906771347ef4492acf4a1fe2512aefd752b2a1490a2287d139

Observation 60a72bd3-e090-45fa-b59f-87aa7f2f9b3f · outbound

This paper cites Jang et al.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Jang et al

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:a55b0e0992ddfa18ec67db71d9eb71acb83c8d192c93fc01b099a9a01f3500be

Observation 82aba099-b0aa-412e-85c8-ec512d18b4db · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF International Conference on Computer Vision

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:6a0bfc611dc0f44bfcbb01612e3b99e3f9cb07b3faf795c84eb51d8576f3b3d1

Observation db8b735c-98a8-4f32-8e4c-abb5783e836c · outbound

This paper cites an unresolved cited work.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:a6b5781539ae0405c63a4b01514a29ae499e424cf3f8d83581139bde538cb6bc

Observation f61cacce-e175-42bc-9506-ea3b54b0b06f · outbound

This paper cites arXiv preprint arXiv:2602.09483 (2026) 3.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation arXiv preprint arXiv:2602.09483 (2026) 3

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:1e3713c51a63da22bd25208571cb3bc3f780ae6350b45e353a8d09170bf95462

Observation 0152ab51-d6bf-4cc9-921e-681cdb785fa3 · outbound

This paper cites arXiv preprint arXiv:2503.01773 (2025) 8, 13.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation arXiv preprint arXiv:2503.01773 (2025) 8, 13

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:2cc8e6fa55ada01e3460f2fbb5d2ca5468f53c60cfdf98b29cdb4b88762f4db9

Observation 4a35c71e-bfa1-401f-804e-d3d5f088be0b · outbound

This paper cites Scientific Reports (2026) 8.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Scientific Reports (2026) 8

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:0976c09fe564058be14a02ed06cade0380f9f743126c047961005420387de007

Observation 453c6014-6e0a-4acd-9318-f564a1eb8d1a · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:b2b5c65975a6e5bc18fc5ed568dc65da34ef3a04df48707cade063c6e5b559a6

Observation 141f3f52-237c-405f-8774-5853f2eef546 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:7545d58ee8258167cb4d4904d3ffc50f7823c213c26fb2388b483c58972b9e69

Observation b4bc6b5b-10f0-440c-94c4-e23bc1ff3330 · outbound

This paper cites In: The Thirty-ninth Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track (2025) 11, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The Thirty-ninth Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track (2025) 11, 22

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:e5f31e829e4212997f05fb6acf3cbeb61fc3e6aec43526231fe486329f3776c1

Observation 1d63e10b-ab3a-47e9-b094-99d9d7bfa121 · outbound

This paper cites In: The Twelfth International Conference on Learning Represen- tations (2024) 2.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The Twelfth International Conference on Learning Represen- tations (2024) 2

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:052299f772b2f502ac868b691c7ce3faa2fa94a211098e9c5e84508ae14fc48b

Observation 915b37af-adf1-4796-8326-5c3e21fc226c · outbound

This paper cites Image and Vision Computing146, 105020 (2024) 3.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Image and Vision Computing146, 105020 (2024) 3

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:dbf0a34d40aec9f23993d73aaf0f167e2ad54d9f0898c9227b872a69e80cd650

Observation 0838025f-b421-4635-be6a-d8590ace867e · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:1a9f90dc594fee3ce3ff56b74daf439247fc97efab2b475688f83f660b616406

Observation 70a16c5b-0303-4681-8cab-f092fc0be3b3 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Gaussian Error Linear Units (GELUs)

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:4e6ef592a69ce2383a577530c94c76ac09d28b50e5b20764fe5100189b77e867

Observation 3515f1e9-96d7-437f-ac0b-0134f858d751 · outbound

This paper cites an unresolved cited work.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:b764aaa884404ba2b1dc300ab7f72edec4a56de48e208e11d581748a418d8ec1

Observation 223bb2d2-99b1-4d15-8244-451d2133af1a · outbound

This paper cites Advances in neural in- formation processing systems36, 31096–31116 (2023) 11, 21, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in neural in- formation processing systems36, 31096–31116 (2023) 11, 21, 22

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:1d0d09d95323d581d90b270ba98911ba92f517314f7f45022e10b3212485cac0

Observation 54168d2f-c6fd-47c4-912d-eca6238cd6be · outbound

This paper cites In: Proceedings of the IEEE/CVF confer- ence on computer vision and pattern recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF confer- ence on computer vision and pattern recognition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:c4845e2a01ae76bf242cc2df559d3ecd97dcbf8a4d985b645343e491e1264db9

Observation 28d6eaa1-e126-4e83-a239-d45845f2624d · outbound

This paper cites In: The Fourteenth International Con- ference on Learning Representations (2026) 2, 3, 5, 6, 8, 10, 13, 22, 24 Token-level Response-visual Attention Guidance for MLLMs KD 17.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The Fourteenth International Con- ference on Learning Representations (2026) 2, 3, 5, 6, 8, 10, 13, 22, 24 Token-level Response-visual Attention Guidance for MLLMs KD 17

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:522a7dddb16f1cab74125424db2bfd72c7265a5d0502d3eddd1af69cf794fdab

Observation 21bb65d7-d824-4af4-a27d-5686d78c322f · outbound

This paper cites In: Forty-first International Conference on Machine Learning (2024) 2.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Forty-first International Conference on Machine Learning (2024) 2

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:c4f7a4c1a925859a84b918c8a5fb54c09886edd5a44f95980531ed3863f636b2

Observation 8426b147-8cb7-4310-8a1f-ede94743c189 · outbound

This paper cites In: The 2023 Conference on Empirical Methods in Natural Language Processing 11, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The 2023 Conference on Empirical Methods in Natural Language Processing 11, 22

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:a843d3ffd82bd597599a16a21aa2422e8614aec4107bee30a354a9621589bfd2

Observation 1450bcb7-6075-472d-8c90-58c686232123 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:ce009549bc50c2bd56e9e704497aff990305b9cff66bf41d74b4164de9835e15

Observation b8c90f58-ad53-4163-8029-22730a179f4d · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:e3a0d4e7e0a4d78bf4031ac1ece0c5bb25bda89c86419bf0db22247051972804

Observation 42f02d38-cd49-406a-8b55-fce3df53bccd · outbound

This paper cites Advances in neural information processing systems36, 34892–34916 (2023) 1, 4.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in neural information processing systems36, 34892–34916 (2023) 1, 4

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:f4866cc9d151e1d18c0dc2ee798374f0162279c912e3a9517a2c5113f00472cc

Observation fa12e0d8-e266-449f-856a-c23fe769f5b9 · outbound

This paper cites an unresolved cited work.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:7e130ed0997a81c04e1cb7c70b99c1590bc621af58344e9537959d9fa45d6fff

Observation cba19a9c-0cba-421b-973b-1b615c90b766 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:809d1570e862b773cebe8719774cc03027ff54bad566616e2fb1fb13e7601339

Observation c1ac0dc3-37fc-4cc8-b3ef-34022f314e5a · outbound

This paper cites Advances in Neural Information Processing Systems 35, 2507–2521 (2022) 11, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in Neural Information Processing Systems 35, 2507–2521 (2022) 11, 22

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:78690a818a3220bb1cc22840eaab652e99aaad180ee96ef5f79b84a235bdf137

Observation 9fe202d8-92f4-4fd6-ba03-04c0882657b1 · outbound

This paper cites Advances in Neural Information Processing Systems37, 101880–101904 (2024) 11, 21, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in Neural Information Processing Systems37, 101880–101904 (2024) 11, 21, 22

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:e2a60fd6c122d09675efbd9d4b18707ac1dbaf3a17310abec5fc5518418f100f

Observation be5d2b2b-517a-4aea-a6d5-e8d9fbb4bf35 · outbound

This paper cites In: International Confer- ence on Learning Representations.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: International Confer- ence on Learning Representations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:d617bd855006c4de8e5773e979851c369927f7558750599a397c52017f38ff69

Observation 33870b8d-7d11-40c6-82e4-2e68cfff7963 · outbound

This paper cites In: Pro- ceedings of the 61st Annual Meeting of the Association for Computational Linguis- tics (Volume 1: Long Papers).

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Pro- ceedings of the 61st Annual Meeting of the Association for Computational Linguis- tics (Volume 1: Long Papers)

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:4c8b29efc0ded63679fbfa0d2256db49e99216aef1e554b4d6f934b8bf88b4f1

Observation 38213dc0-8e6a-49a5-b0ea-f3d261464658 · outbound

This paper cites In: The Thirteenth International Conference on Learning Representations (2025) 3, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The Thirteenth International Conference on Learning Representations (2025) 3, 22

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:d9289c9035139f53f31a5e2b9997453a9b62a0450579c5da083e9068e8ba99c7

Observation f4fc5747-ab7a-4d80-accb-27c52f540972 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:3cb237c53440564192bbcef75eed3c1ecce5b84d8f114997b1a526c9749d4071

Observation 06929c93-75f7-4822-99fe-d15ad45c097b · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pat- tern Recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pat- tern Recognition

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:85e4040aac3620adbfb79ec52f0091f2f728564ade41637cf773fe466066ca2d

Observation c2651038-4c34-4a58-9ee7-4e3cdf7ed84f · outbound

This paper cites Advances in neural information pro- cessing systems30(2017) 4 18 J.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in neural information pro- cessing systems30(2017) 4 18 J

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:07b513d06daddfccb66b61019b48beed6e111812b7f6b0330c54ab8b65717ff0

Observation cf240ae3-4bad-4a29-af40-a31a80e633b2 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:66858e125f149065f73c7509b29ecedfc7c5b29c1e1dde1f1327958051b475ad

Observation 5369a142-9f7d-4250-9e29-d4bd22171cc1 · outbound

This paper cites Advances in Neural Information Processing Systems37, 114553–114573 (2024) 1.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in Neural Information Processing Systems37, 114553–114573 (2024) 1

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:419a45c6ed13fc2d94c01bac947fc225a1d6e25a18c3a0bfabac3eff9c948b12

Observation e33c5704-2d33-4d22-a808-9bfd0a7d1bb1 · outbound

This paper cites In: Pro- ceedings of the 31st International Conference on Computational Linguistics.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Pro- ceedings of the 31st International Conference on Computational Linguistics

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:6858197a26332cb9b8238955e8f10928db6a17d61af87d026103390f2b1184c7

Observation 1c498e38-1e41-4d7f-b5b7-1153890db0ff · outbound

This paper cites LLAVADI: What Matters For Multimodal Large Language Models Distillation.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation LLAVADI: What Matters For Multimodal Large Language Models Distillation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:4d74f53059c81e2cc41c9792ed3116350463f3a1ba0482c4c5d638b274fb5fef

Observation 42a707c8-37ef-4f14-8173-f3f425cb0be4 · outbound

This paper cites Scientific Reports13(1), 18369 (2023) 4.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Scientific Reports13(1), 18369 (2023) 4

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:e3e7cc9e1d3ad544027a96a07caade11040402feed22263b358771f304f49c5a

Observation f91d8a65-7cab-498b-91cf-fc3c479db2df · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:d5adf178df22fc9617633fc934889810dbef4bd3e82413f1e4e0781a32873b9f

Observation 20b604cc-e141-42c3-8c60-34634e86c90e · outbound

This paper cites In: Forty-third International Conference on Machine Learning (2026) 2, 31.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Forty-third International Conference on Machine Learning (2026) 2, 31

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:b86c8bce519bc940c5dffae95e71060b4a15f34801ac2c87b279e49944dd0370

Observation f8a2189d-94d9-4530-b86c-ea12f3d4a513 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:511218d32a15ee644f490f36d440d7c6be94363ec28136b825d7dbb45885c509

Observation a8cf291d-6305-4e88-99a5-7d6c1ebe08c9 · outbound

This paper cites In: Interna- tional Conference on Learning Representations (2017) 3.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Interna- tional Conference on Learning Representations (2017) 3

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:a9cc140441d599f87c428777588ad250401d0443f94fc23f1178322f25bf187e

Observation 98a3b12d-5662-4771-ad2e-4ff8ae2f8c94 · outbound

This paper cites In: Proceedings of the IEEE/CVF international conference on computer vision.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF international conference on computer vision

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:f266e62b2e2d772f055f61bc8d9b8e77f76e4ff5d86471d7d5afe8e4338cf74a

Observation 777a44ec-d2ee-4fd5-a2e6-e6e42772a763 · outbound

This paper cites an airport.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation an airport

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:fc89e01d60ed2a8acf80c9c234961f2e093b985d160ad32e5afe72d3a417e2ff

Pith citing papers

No inbound Pith citation observations are available.