Pith. sign in

Paper Citation Record · LEDGER

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation

As of 10 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2607.02593.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.02593 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T09:26:29.033271Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 58d73325-0b48-468c-a3a1-ed1744c875e0 · outbound

This paper cites In: The Twelfth International Conference on Learning Representations (2024) 2.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The Twelfth International Conference on Learning Representations (2024) 2

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:f740b9d00dec0e3e4598bd771696e4ab28b9b78ef2708455b956ddf18d070d5e

Observation c6e331ac-e63a-44d1-b919-665aff10ba00 · outbound

This paper cites BD-KD: Balancing the Divergences for Online Knowledge Distillation.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation BD-KD: Balancing the Divergences for Online Knowledge Distillation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:f4d16a0854232097c66dad20e8e7034709be341a73da425f0758f2d9fe689b3d

Observation 60a72bd3-e090-45fa-b59f-87aa7f2f9b3f · outbound

This paper cites Jang et al.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Jang et al

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:fa9f51a6a9a39f5983cae10ef33a64f84decd79244a0d1ee4bc4c2d95ef65111

Observation 82aba099-b0aa-412e-85c8-ec512d18b4db · outbound

This paper cites In: Proceedings of the IEEE/CVF International Conference on Computer Vision.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF International Conference on Computer Vision

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:26462baeddab33bce3e3799250555ef7bf12014d2798759d237d98465491e0d6

Observation db8b735c-98a8-4f32-8e4c-abb5783e836c · outbound

This paper cites an unresolved cited work.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:01c34cf242a9575652489731535b6dff3a153c609d2849197d28b4abaf483388

Observation f61cacce-e175-42bc-9506-ea3b54b0b06f · outbound

This paper cites arXiv preprint arXiv:2602.09483 (2026) 3.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation arXiv preprint arXiv:2602.09483 (2026) 3

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:ddee5364651e29fd267fe8e72e1cce77ddddff4ec817029786c6b113c994b7de

Observation 0152ab51-d6bf-4cc9-921e-681cdb785fa3 · outbound

This paper cites arXiv preprint arXiv:2503.01773 (2025) 8, 13.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation arXiv preprint arXiv:2503.01773 (2025) 8, 13

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:81a472a992821d3cdf59cf23de4841897f667b6d9c1de0a6a01d58052f87a3ea

Observation 4a35c71e-bfa1-401f-804e-d3d5f088be0b · outbound

This paper cites Scientific Reports (2026) 8.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Scientific Reports (2026) 8

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:6a3d6517f9ce33f1562aa16fa7f66b2fcec813b4352c6f26435048ec8a53f60c

Observation 453c6014-6e0a-4acd-9318-f564a1eb8d1a · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:0badaa563ed26dae0d44d50a7733c6915f7068be06610db5242c8abad6a2a73e

Observation 141f3f52-237c-405f-8774-5853f2eef546 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:7370baa64756d8482199af65fab1b58e91715d3b905f799e1efc30cf75d8b429

Observation b4bc6b5b-10f0-440c-94c4-e23bc1ff3330 · outbound

This paper cites In: The Thirty-ninth Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track (2025) 11, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The Thirty-ninth Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track (2025) 11, 22

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:d0d389a09857f85a9c847f20cb3ef284d715339f9352a71ddf72d3bd48261307

Observation 1d63e10b-ab3a-47e9-b094-99d9d7bfa121 · outbound

This paper cites In: The Twelfth International Conference on Learning Represen- tations (2024) 2.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The Twelfth International Conference on Learning Represen- tations (2024) 2

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:0b0072d597e423ccc1ba24b54d5b5536db50fd2b7319ca2008b9e59ad741fa6c

Observation 915b37af-adf1-4796-8326-5c3e21fc226c · outbound

This paper cites Image and Vision Computing146, 105020 (2024) 3.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Image and Vision Computing146, 105020 (2024) 3

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:62961b10cf7b29e93a7bf45ee3470fc203eccc3606f9766a3098a573d0b6e9d7

Observation 0838025f-b421-4635-be6a-d8590ace867e · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:dde642ac7262156eaa65ed2066d256246a4c282fd1500c6c16bd640888314048

Observation 70a16c5b-0303-4681-8cab-f092fc0be3b3 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Gaussian Error Linear Units (GELUs)

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:c999e88070c4a7f4b6e998a401bb13bbd89eaed2bb747c9375c8b985be6708cb

Observation 3515f1e9-96d7-437f-ac0b-0134f858d751 · outbound

This paper cites an unresolved cited work.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:4ba289459fe7f323bd054320abebc468b1b57150ee9237fbe80540f46f869106

Observation 223bb2d2-99b1-4d15-8244-451d2133af1a · outbound

This paper cites Advances in neural in- formation processing systems36, 31096–31116 (2023) 11, 21, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in neural in- formation processing systems36, 31096–31116 (2023) 11, 21, 22

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:88758e0a3fa2fd1b0cdf634c5285af2b8b21591e99933ab8a594ab0e2d2fa208

Observation 54168d2f-c6fd-47c4-912d-eca6238cd6be · outbound

This paper cites In: Proceedings of the IEEE/CVF confer- ence on computer vision and pattern recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF confer- ence on computer vision and pattern recognition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:93262b30a2464e2b1627be98ff27f9e165232436b35e5d564f9ddedb12e21665

Observation 28d6eaa1-e126-4e83-a239-d45845f2624d · outbound

This paper cites In: The Fourteenth International Con- ference on Learning Representations (2026) 2, 3, 5, 6, 8, 10, 13, 22, 24 Token-level Response-visual Attention Guidance for MLLMs KD 17.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The Fourteenth International Con- ference on Learning Representations (2026) 2, 3, 5, 6, 8, 10, 13, 22, 24 Token-level Response-visual Attention Guidance for MLLMs KD 17

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:e4792041a71aa612f0ae6922f9953ecd973476fdfe57f7c9bef4d8d1fa8f1a9d

Observation 21bb65d7-d824-4af4-a27d-5686d78c322f · outbound

This paper cites In: Forty-first International Conference on Machine Learning (2024) 2.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Forty-first International Conference on Machine Learning (2024) 2

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:02c8ceee5555854790c2a945e9d663fc32e25c12ba9a031704474da65bdb5010

Observation 8426b147-8cb7-4310-8a1f-ede94743c189 · outbound

This paper cites In: The 2023 Conference on Empirical Methods in Natural Language Processing 11, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The 2023 Conference on Empirical Methods in Natural Language Processing 11, 22

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:464ef457d316e697100d4f10033afb5b66d386c9e6598f393f74d8b1fac8df3a

Observation 1450bcb7-6075-472d-8c90-58c686232123 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:d5205551f1bd221308127c6ef22a45cbc591d6a9a76a7e22b1990247c0ad9b41

Observation b8c90f58-ad53-4163-8029-22730a179f4d · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:94a84198f3494928e5a0c6a463056fc85763a07da71a802107b62bed959bd2e3

Observation 42f02d38-cd49-406a-8b55-fce3df53bccd · outbound

This paper cites Advances in neural information processing systems36, 34892–34916 (2023) 1, 4.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in neural information processing systems36, 34892–34916 (2023) 1, 4

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:c55dc33b9642d03423f6ac39edd12f88bcedba0e2a199a82dc2d654b628dc4e8

Observation fa12e0d8-e266-449f-856a-c23fe769f5b9 · outbound

This paper cites an unresolved cited work.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:d472554565882689132163b04a1ea6a3398bda944536bbc89edda1165e2623d2

Observation cba19a9c-0cba-421b-973b-1b615c90b766 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:7f71efa00f2ab3a9a0a4845a75525046a9dd0d079f88718effb6ce3c295c19df

Observation c1ac0dc3-37fc-4cc8-b3ef-34022f314e5a · outbound

This paper cites Advances in Neural Information Processing Systems 35, 2507–2521 (2022) 11, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in Neural Information Processing Systems 35, 2507–2521 (2022) 11, 22

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:43501861767c6336247deb624453c464761dccb5cf4310c4720fb4bd60d5f1a2

Observation 9fe202d8-92f4-4fd6-ba03-04c0882657b1 · outbound

This paper cites Advances in Neural Information Processing Systems37, 101880–101904 (2024) 11, 21, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in Neural Information Processing Systems37, 101880–101904 (2024) 11, 21, 22

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:3abed393c3dc05b2dd092b453174240a7bd75b0093e5e491e8f49e13c8f0e084

Observation be5d2b2b-517a-4aea-a6d5-e8d9fbb4bf35 · outbound

This paper cites In: International Confer- ence on Learning Representations.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: International Confer- ence on Learning Representations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:0c9565d41d5e8b87ab0ec2c8adacc90f66eef6970eb314ec355f7911d5b433bf

Observation 33870b8d-7d11-40c6-82e4-2e68cfff7963 · outbound

This paper cites In: Pro- ceedings of the 61st Annual Meeting of the Association for Computational Linguis- tics (Volume 1: Long Papers).

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Pro- ceedings of the 61st Annual Meeting of the Association for Computational Linguis- tics (Volume 1: Long Papers)

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:22dfe5227fe443a5c233da67b30976ade79afa23a1be1e769c8285927f9986d3

Observation 38213dc0-8e6a-49a5-b0ea-f3d261464658 · outbound

This paper cites In: The Thirteenth International Conference on Learning Representations (2025) 3, 22.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: The Thirteenth International Conference on Learning Representations (2025) 3, 22

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:ae4d6d176b48a946de4aec513523de01160683abee4b14aa1f1a75ef357e923b

Observation f4fc5747-ab7a-4d80-accb-27c52f540972 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:cb2336cab8f53026a53df5185aebe1e6f7f9b5d966feb10ba94f68565be58529

Observation 06929c93-75f7-4822-99fe-d15ad45c097b · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pat- tern Recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pat- tern Recognition

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:7b22cde93520bc579ad6dcbd3f9e3c7f4441d8035ecdac1dedc7e3959fb216e3

Observation c2651038-4c34-4a58-9ee7-4e3cdf7ed84f · outbound

This paper cites Advances in neural information pro- cessing systems30(2017) 4 18 J.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in neural information pro- cessing systems30(2017) 4 18 J

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:c87b2538b8c52e9a3f088385a011dc6ae6889682fc1e89e49bd9980c5c1edfa7

Observation cf240ae3-4bad-4a29-af40-a31a80e633b2 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:7ba2363cc0cdd8b1f3592fca123e3458990d7682c76e4186113d9163ae6a327b

Observation 5369a142-9f7d-4250-9e29-d4bd22171cc1 · outbound

This paper cites Advances in Neural Information Processing Systems37, 114553–114573 (2024) 1.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Advances in Neural Information Processing Systems37, 114553–114573 (2024) 1

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:5427a2cc708d6af3b8a8f7366e327e95aa5cbd252bdf81bd65525fa0a638c66e

Observation e33c5704-2d33-4d22-a808-9bfd0a7d1bb1 · outbound

This paper cites In: Pro- ceedings of the 31st International Conference on Computational Linguistics.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Pro- ceedings of the 31st International Conference on Computational Linguistics

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:29eb8dfa7f73b2c6ecf8ab220cdf8ca0f988f39184b9cdbea1cfaedea2fddd8f

Observation 1c498e38-1e41-4d7f-b5b7-1153890db0ff · outbound

This paper cites LLAVADI: What Matters For Multimodal Large Language Models Distillation.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation LLAVADI: What Matters For Multimodal Large Language Models Distillation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:9e834acb606bd76bc8d535a351f5b9546513d1964a941493cf75ca0d8050555d

Observation 42a707c8-37ef-4f14-8173-f3f425cb0be4 · outbound

This paper cites Scientific Reports13(1), 18369 (2023) 4.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation Scientific Reports13(1), 18369 (2023) 4

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:8a0e6b923d8905df673849be5feedfce6f2d0d989a82415d5fd82bd12781c2c6

Observation f91d8a65-7cab-498b-91cf-fc3c479db2df · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:b1e937aa1d55af000dfe206a40f4ed1469c19245b9db319af0efb3ee0c7f5dba

Observation 20b604cc-e141-42c3-8c60-34634e86c90e · outbound

This paper cites In: Forty-third International Conference on Machine Learning (2026) 2, 31.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Forty-third International Conference on Machine Learning (2026) 2, 31

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:0081000b2666104de754a65261115bf19901a3ef84466708fe18f64399f93e82

Observation f8a2189d-94d9-4530-b86c-ea12f3d4a513 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:f5bb51022147854f4d6193ee16e8dca836591df5910bff481aa14a7830d9bfcf

Observation a8cf291d-6305-4e88-99a5-7d6c1ebe08c9 · outbound

This paper cites In: Interna- tional Conference on Learning Representations (2017) 3.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Interna- tional Conference on Learning Representations (2017) 3

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:e1a0f523a6811e11d651419086517b1a391fa59e7ab8fa1558b47b5f240a514a

Observation 98a3b12d-5662-4771-ad2e-4ff8ae2f8c94 · outbound

This paper cites In: Proceedings of the IEEE/CVF international conference on computer vision.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation In: Proceedings of the IEEE/CVF international conference on computer vision

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:13517f98b5939f8f017c1894f458892785cdbce60309ed61fcd0a2146e8dc0d1

Observation 777a44ec-d2ee-4fd5-a2e6-e6e42772a763 · outbound

This paper cites an airport.

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation an airport

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-12T09:26:29.033271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:26:29.033271Z digest=sha256:fedf80814533fdd11e25d17acadce076cf2d18e8c39c0c07f16def5339e79c7c

Pith citing papers

No inbound Pith citation observations are available.