Pith. sign in

Paper Citation Record · LEDGER

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning

As of 16 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 0 inbound Pith citation observations for arXiv:2608.10513.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10513 v1

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:24:49.761179Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

72 of 72 outbound references displayed

  • verified exact2
  • verified fuzzy14
  • unresolved56
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6025fc5e-cb13-4b9e-a1f7-bcc2fbfedd4b · outbound

This paper cites Communication, Simulation, and Intelligent Agents: Implications of Personal Intelligent Machines for Medical Education.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Communication, Simulation, and Intelligent Agents: Implications of Personal Intelligent Machines for Medical Education

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.286087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.286087Z digest=sha256:2026b501099affcc9309bd4f6fd179bf7b2a3f577370e3cc2d79d25859824a1b

Observation aafe544d-f17a-4557-a3b7-65bcaa860b07 · outbound

This paper cites Classification Problem Solving.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Classification Problem Solving

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.291857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.291857Z digest=sha256:df2eb1ebfb8fb165b57baec3ba231e42b58d369ad9fbab1936693c564d150f52

Observation 03f39d12-37c9-475e-8996-782a28b02c46 · outbound

This paper cites , title =.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning , title =

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.297126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.297126Z digest=sha256:ba2fa10751139ab1dae76ae95ae16002d415c7f7f3aaa1c45888474db102567a

Observation 9b684ccd-2da1-4a2e-bd08-f0b86098ad39 · outbound

This paper cites New Ways to Make Microcircuits Smaller---Duplicate Entry.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning New Ways to Make Microcircuits Smaller---Duplicate Entry

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.302190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.302190Z digest=sha256:bafe2badc007d3c036d6e5b0d507486c082c7a8fc90c76cbaeacac79138f0b97

Observation 6e7c5bc3-0353-4ce4-a8bb-05eedb4e6db7 · outbound

This paper cites Clancey and Glenn Rennels , abstract =.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Clancey and Glenn Rennels , abstract =

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.307325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.307325Z digest=sha256:7b3bb68941361a299a07a0d7b744324ebd28b7e44209d74bd5fb98fb18f1e554

Observation b370780d-21ed-4d8e-922a-391605fab0d1 · outbound

This paper cites and Rennels, Glenn R.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning and Rennels, Glenn R

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.312726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.312726Z digest=sha256:92fddc0bd019e0bab2ee35dbf819cfa28564dc6e1cef247fab8b892ea313ad0e

Observation 7c2213ad-8c12-4d7b-a616-e70bc73431ad · outbound

This paper cites Poligon: A System for Parallel Problem Solving.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Poligon: A System for Parallel Problem Solving

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.318979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.318979Z digest=sha256:6a6e2277762975fec5b256e3149955ece631f87960f67fa02a1e3a0c019f8c81

Observation 1af58472-120b-4390-a2f6-3e45ba91d558 · outbound

This paper cites Transfer of Rule-Based Expertise through a Tutorial Dialogue.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Transfer of Rule-Based Expertise through a Tutorial Dialogue

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.324337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.324337Z digest=sha256:7203b2d6c2d0f20f110d459d972331b6a94b015ec486384c4ed52a22e995422f

Observation 30e37e8f-9630-4387-a5dc-7562ba936e85 · outbound

This paper cites The Engineering of Qualitative Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning The Engineering of Qualitative Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.330017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.330017Z digest=sha256:3cfc0639b627e205f5ec432bcc60f2fd9acd26fa6a47657154f6ae88615469f8

Observation 8dd50b5f-f6b2-47c0-b120-2f58e7e202f2 · outbound

This paper cites 2023 , eprint=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning 2023 , eprint=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.334945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.334945Z digest=sha256:088e440e530782435f1e7bcbb33c9a2ec5a7a590b62c4ca0e02eacde20b7de5b

Observation 8412b579-2ad8-4563-9fc2-71154bb189a0 · outbound

This paper cites Pluto: The 'Other' Red Planet.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Pluto: The 'Other' Red Planet

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.340125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.340125Z digest=sha256:f177646ef115e3eead940ec92d230c5c2a3927f9a15b92ef8f557ee465afdc03

Observation 228938b6-472c-4002-bb00-eaa25e47f0d7 · outbound

This paper cites European Conference on Computer Vision , pages=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning European Conference on Computer Vision , pages=

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.161453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.366719Z digest=sha256:1a50d643fdb9be979eae3cf5a0a73665a0c70b276fef284dac477d4b95a1bd3f

Observation 08e7d26c-4bf3-49e6-8c3b-0f8a7c03056b · outbound

This paper cites Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.144189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.387989Z digest=sha256:489ac341ca09d61ed4efec7247ce4b3fda28bc1d79709bce2ce6f738d5e980f8

Observation 83099ffa-5463-4801-8505-c206be410151 · outbound

This paper cites Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.127311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.392957Z digest=sha256:b0f1f5a31f48260f2b921c914e9e4919a446df8f577a32fc44dd4ff3a6701855

Observation 3dd945eb-064e-4277-833e-943d8cc322aa · outbound

This paper cites International Conference on Learning Representations , year=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning International Conference on Learning Representations , year=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.110729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.419372Z digest=sha256:b6f59a60f3da7e857129fdd300d5dda0d9111c4f79a908b24833147a7272b132

Observation c0c9e85f-a084-4bbb-96e2-1570013db0a5 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:51.093297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.429522Z digest=sha256:e8f5ff90e8de65276e4a5d0dada34d4f3255d2ec527bca3dee00fe69b9ae8538

Observation 4e7208e8-6538-4df6-b419-8ca1cf767783 · outbound

This paper cites 2024 , organization=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning 2024 , organization=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.076144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.434452Z digest=sha256:509b74499fc4da142d4a587d8d494c973077b5d2e08758357de7aa523ed3f665

Observation 3b98b798-fc97-4541-bc10-1eec8dbbd7f8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.058351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.440055Z digest=sha256:102ebeece6dc90ef3a4d51bf24b07415eb340b8e28849806271296bf8e7ae939

Observation af1397e5-5deb-4e3a-8ee1-3a02f280d7c2 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:51.038568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.445335Z digest=sha256:3abb9528b948956456a7b7a22d78faded629c1ebdbb28ef6fdb44fdb1743aa83

Observation bd1c348e-53c4-4dfd-ac82-5f65f525fdf2 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:51.019929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.455750Z digest=sha256:fce7a93799aa1abe574f5bf861e161c5b4b9f428746f020ee9462f0a0bc98ebc

Observation 7c801524-ce3b-49e5-af20-dae48fee7937 · outbound

This paper cites Immune: Improving Safety Against Jailbreaks in Multi-modal.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Immune: Improving Safety Against Jailbreaks in Multi-modal

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:51.003543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.460698Z digest=sha256:eea981650535ce43b5f906a3a4ac773be2c1bbd7d991ce9564d4b39d1630691c

Observation 217032c7-0fc6-4961-983f-9dfe298829f1 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.985880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.477037Z digest=sha256:691df6dfb015a5be654aee338d22e6619e9b94910d8f78609fc4bc0deda85652

Observation ec4efce1-7f79-4e07-9aa5-0179b1da6d55 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.968545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.481819Z digest=sha256:c9ba37437c31d942c37f3f553bc6080b6e646d0921a9e443ab27d89244f78df5

Observation a63deb8f-bea4-4fa6-b41b-7cc2d9cdeff8 · outbound

This paper cites Computer Vision -- ECCV 2024 , year=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Computer Vision -- ECCV 2024 , year=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.952194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.486867Z digest=sha256:314e8c73771467bd361a54383a471f0721bdb58243343b12dffe999573115fd9

Observation 953d98a7-bbe0-4ebd-b7b5-21f38ce009f0 · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.935088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.491449Z digest=sha256:8f3a1a202434d7ea618d6dfa3ab2f5edc4c49825e19220dae765114d742ad43e

Observation 2c142d4a-a0a0-47df-913c-31da69c8f934 · outbound

This paper cites 2025 , eprint=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning 2025 , eprint=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.918592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.496220Z digest=sha256:3e28fdb275496c86379b9f7b57645d60a2df65a4fd3948cbdf4326808454a2f5

Observation 034ce379-cce3-4d28-9fb0-a351d1e9e384 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.901331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.506246Z digest=sha256:27adc93c2930c143f9878636e8b94e7b8f7cf15b6718245be32b12ce299d7e3d

Observation 4d562cb7-31a3-43b9-a9dc-e10ec02ef7d5 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Advances in Neural Information Processing Systems , volume=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.516080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.516080Z digest=sha256:4a7f6e62d23178bf5fe8fd1ddd696de2c0bdaed6cd049929b0dffc1ae457a91b

Observation 3d8d9528-f3c7-4256-862a-3c5a8ccba9fa · outbound

This paper cites 2025 , eprint=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning 2025 , eprint=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.546118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.546118Z digest=sha256:3b2db4684c9eb3e7bba200bc5fcd06fb645e24b1d0b5ac309dbd125ff67b6b43

Observation 42b41b8a-ddef-45ae-a798-53896366a1fd · outbound

This paper cites 2025 , howpublished=.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning 2025 , howpublished=

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.863889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.551124Z digest=sha256:d5ba9dcd89559b779c9dadcd79a4ba80dada45e410a6b3d18731d66e7f436211

Observation b99c6442-d704-4675-9462-6145ec9c58e7 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.848497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.561077Z digest=sha256:91c5127d50cdfd86479310b1dfc1d8c8ff5768469d1a658f523387783616b00b

Observation 7af71618-73f6-4217-a0d6-926fb2b4d874 · outbound

This paper cites S.; Dong, Y.; Roy-Chowdhury, A.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning S.; Dong, Y.; Roy-Chowdhury, A

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.566112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.566112Z digest=sha256:cee51870d88dfdac513b0f1ef142ffa462838794cc4125e98066a4d93332d388

Observation fcac46c4-dee3-4c64-9f6c-298a3dc99766 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.833135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.570974Z digest=sha256:a442af84e28cf8bd51f32fe194e2ee48b465250d93ed36be69b7ecb2037bfb6b

Observation ead0cd05-33b7-4a54-9afc-9b2b4d1ec861 · outbound

This paper cites Are We on the Right Way for Evaluating Large Vision-Language Models?.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Are We on the Right Way for Evaluating Large Vision-Language Models?

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.575768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.575768Z digest=sha256:cb00d5851f946cb0795555673551f62ceafc2cff7d327472d01abf74be489cc7

Observation 59813561-4a5a-4fbc-88bb-df1b70ebd447 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.580381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.580381Z digest=sha256:3e39027d297e570d2cde0ccdeca7e93b2521272a8789d97bae702d6d58ba2e6e

Observation 22e5aa0d-b540-4990-9f7d-b19fcaec8339 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.585173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.585173Z digest=sha256:dab974f9ed37dc7731a13662c92d0197e497d0e4f9c729ed3f68b46012647858

Observation 5948ede7-3092-4529-985d-e8cfa1c54525 · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Gemini Robotics: Bringing AI into the Physical World

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.589605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.589605Z digest=sha256:e77015cc80d7f5904ab179ebb7c773afbb2d1ffa41bc4bbdc1648b58a783f188

Observation 31627a61-ebd4-4a9b-ab3a-ae7869c2a350 · outbound

This paper cites S.; Chakraborty, S.; Singh, V.; Guan, T.; Wang, M.; Velasquez, A.; Beirami, A.; Huang, F.; Manocha, D.; and Bedi, A.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning S.; Chakraborty, S.; Singh, V.; Guan, T.; Wang, M.; Velasquez, A.; Beirami, A.; Huang, F.; Manocha, D.; and Bedi, A

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.817152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.594546Z digest=sha256:fde405b266a8add5297d5a8a6f1e360b300d90fb99f71036ee192fc399c4673a

Observation 28abac41-ddee-48b4-b94f-af4761b95fc6 · outbound

This paper cites FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.599621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.599621Z digest=sha256:076245bac6394d943e05e9e9dd3bebbcbe666d32ea6fbb6783540650fb2762a1

Observation 25918594-f841-43f1-8862-2212b3e99ad6 · outbound

This paper cites T.; and Zhang, Y.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning T.; and Zhang, Y

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.800740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.604537Z digest=sha256:bb4e4c47f58e8465d77a48b48909e575e7037e7d6f05a02ebb509b2a0958005d

Observation 4dd94b19-98cd-4cea-987c-4013b0862739 · outbound

This paper cites Tit-for-Tat: Safeguarding Large Vision-Language Models Against Jailbreak Attacks via Adversarial Defense.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Tit-for-Tat: Safeguarding Large Vision-Language Models Against Jailbreak Attacks via Adversarial Defense

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-08-15T14:24:50.323213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.609578Z digest=sha256:85caaa127182ec92a3ff365dd8527a0da403910a9b0fde7ff62fbbbf3d9dfacf

Observation 5ac197ff-34c5-4570-82fe-84b6764a67ae · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.784835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.614405Z digest=sha256:0c7f8f610fbeeee6e59599ff5124c21798dde3f8fe52c6a56f2bb6941bd1ee47

Observation 2dc6e744-3892-41f5-815d-f8b9a5e4ff39 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.769285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.619496Z digest=sha256:bc35cd7d6fa34c72401738936ea96da453c58560dc7f459888460c289d55a9cd

Observation b9bf0e8c-3aaa-4677-8afe-b2c4d7fd2db9 · outbound

This paper cites How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-08-15T14:24:50.599930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.624109Z digest=sha256:2172a980282e1865b745e409aefd44cc167019b27cd976f082ae122f093608ad

Observation b330890a-6d8b-4c62-a945-a091af957280 · outbound

This paper cites Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.628739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.628739Z digest=sha256:aa749598e416d099ef9e8d92e2da996ffc8cb90ef0420c59df163271e6d8366e

Observation 8c6746d9-388b-4d77-992a-e3c437860635 · outbound

This paper cites GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.633399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.633399Z digest=sha256:4afb97ce72aa98e6df9d94ac797c09933a4deea248758198c867503fcf35508c

Observation a1c14e02-0921-4471-83ab-86e644c7fb38 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.753489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.638422Z digest=sha256:ceb153dfda1f60c4d4503b53b76a01dec0fa016254a9ac27323c4822a2c42a6b

Observation 269c60e0-8b68-425e-9851-c51d0534576f · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.736673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.643490Z digest=sha256:b58258347acc224cf014625d021579ca41074c293f3f8603e2de5d7e10c235a8

Observation 0566dc98-f71a-4d25-95ae-b9bf57f6730c · outbound

This paper cites UniGuard: Towards Universal Safety Guardrails for Jailbreak Attacks on Multimodal Large Language Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning UniGuard: Towards Universal Safety Guardrails for Jailbreak Attacks on Multimodal Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.648328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.648328Z digest=sha256:4f7596d278e66cdd608c7379750ce3bd2214f83ebf7ca32472598b9a49faab50

Observation 490d3cb1-abf5-456a-8af5-94c4a05e0634 · outbound

This paper cites Learning To See But Forgetting To Follow: Visual Instruction Tuning Makes LLMs More Prone To Jailbreak Attacks.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Learning To See But Forgetting To Follow: Visual Instruction Tuning Makes LLMs More Prone To Jailbreak Attacks

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.653520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.653520Z digest=sha256:2d7941db8f8a86233eefe0bb0bfe46e0fb200e256c7132a8ead14a0ff4ea5e52

Observation f1bdd4e5-71c7-40b5-a5a0-d9479f72aa14 · outbound

This paper cites MLLM-Protector: Ensuring MLLM's Safety without Hurting Performance.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning MLLM-Protector: Ensuring MLLM's Safety without Hurting Performance

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.658164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.658164Z digest=sha256:8bea9308f13c4cee5dba016a5f5bc90e6dbf6738dcd6e7b220808d5a704fcbad

Observation b468061a-8828-426a-a25a-39e7fb2a1094 · outbound

This paper cites Image Textualization: An Automatic Framework for Creating Accurate and Detailed Image Descriptions.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Image Textualization: An Automatic Framework for Creating Accurate and Detailed Image Descriptions

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.663318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.663318Z digest=sha256:ffb0687a28495cb49d4919ad2681feeb29a318d0662b66a7a1a7d0459a7d872d

Observation ff2c9902-8178-4ac7-b812-54158f74d295 · outbound

This paper cites Visual Adversarial Examples Jailbreak Aligned Large Language Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.667991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.667991Z digest=sha256:69c0b28302968dccdfd8b9f6558ad90c3d7cfeedbc67fccdec510fd79320e0df

Observation 80bcc7fb-bb9c-4d73-a9a3-fce47d6851a4 · outbound

This paper cites D.; and Finn, C.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning D.; and Finn, C

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:24:50.719088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.672709Z digest=sha256:72142fdab9333c39710de7fe18e066f79672c9f520b4b5b0953ee882af4e2bb5

Observation a084824b-141e-4c09-b436-0b33272b7c61 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.678351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.678351Z digest=sha256:f9fedd7320e509120d2624464efa43180ef2dc422791fa18f6ad88f0a0ee75ab

Observation ee26d16d-fc61-46ae-9fbc-4cce92410aa5 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.702794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.683390Z digest=sha256:66180d5e015021aa1e2af0fb1bb57782187f59f221423a12bdf6562d3aebddf1

Observation f619761b-f2a6-471e-88ec-11824f613c4f · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.688116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.688116Z digest=sha256:3102a8480d121c184ce1e6c008f34637adc9ad7686023da466defa381a985f92

Observation fc5301ea-d58a-41b1-b04c-c916f780f957 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.685891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.692606Z digest=sha256:9720f90ecc6fa15d7a8c3fa5bd92960372698963fb528eb3921e2284823f76c4

Observation 35a00d1f-c63b-41cb-a266-798754d1fa60 · outbound

This paper cites Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.697447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.697447Z digest=sha256:76920264a54e9e7f5213b6e3b01455a4d030bf2b6ea862f5949b2dff2a7b086f

Observation 319c5e91-143d-4896-964a-00b3ea1e1ccb · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.669292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.702249Z digest=sha256:90bfde16aec7f5075fc8efa47fd895b83c1b2a3df21f97fe53cde3d4734be9f9

Observation 60a371c4-1909-450e-ae9c-02d9c15c21ed · outbound

This paper cites InferAligner: Inference-Time Alignment for Harmlessness through Cross-Model Guidance.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning InferAligner: Inference-Time Alignment for Harmlessness through Cross-Model Guidance

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.707221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.707221Z digest=sha256:378b78ce7ff20b8ad86bac6d0c2047ac6c1be0c29f64919cd841ac440058b456

Observation b252a0b5-1046-4f9f-8096-c4836462f329 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-15T14:24:50.652790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T14:24:49.712315Z digest=sha256:4a5124cc69821dea0009a4381d5e7caafc85f1ef0441963b3dfda1fd427e7fe8

Observation 47cbe7ae-0820-4309-b7a4-a69d2c725db3 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.717212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.717212Z digest=sha256:08551aa9a946af2c5b5c488818e131fde8d83db7cb2b13fc929fc2001f402262

Observation 2c287b8c-e440-414c-8279-c9fa5299dea6 · outbound

This paper cites A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.722202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.722202Z digest=sha256:a77733e1b0c6af608ab83e77af28ab5f9893eebf67d4df13e9bd088319245b34

Observation d9b1a76d-8748-4a6d-ad38-cde6c6a62f53 · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.727045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.727045Z digest=sha256:4dc97a1327f1f399d17e167b4631a1899d518ed69d9b75bcb237b9c4e1ba56f6

Observation 0124c92d-8536-4a13-87e3-6d0eaf31fbb5 · outbound

This paper cites SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Model

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.731853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.731853Z digest=sha256:e258ac8ac095bbb1ccc09b523cf87fa777755f976256356c0b144fb467c16d1a

Observation bff65539-c313-49db-a9ae-bb5b9155bc60 · outbound

This paper cites an unresolved cited work.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Unresolved cited work

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.736633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.736633Z digest=sha256:c5c0f1dd4fe8ad922d9ef839ed064b5ef0691ece3a7cbb7037aa60b065d803bd

Observation c10d224c-4731-4766-8304-f2d4adcf8da4 · outbound

This paper cites MM-RLHF: The Next Step Forward in Multimodal LLM Alignment.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning MM-RLHF: The Next Step Forward in Multimodal LLM Alignment

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.741420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.741420Z digest=sha256:1586d7e893b49b21b11b8a8176a9bf91edd26b6f9ac0c95d69709d8442fa225e

Observation 96f4e123-ea8d-4458-b644-8b37f42e79f0 · outbound

This paper cites Multimodal Situational Safety.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Multimodal Situational Safety

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.746172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.746172Z digest=sha256:9e5f3fff4074b70716d9ab171795ad27b1887e2e15ad2368ad05eded938d13a0

Observation c33304f9-7846-4db1-818e-036cdd66476b · outbound

This paper cites Safety Fine-Tuning at (Almost) No Cost: A Baseline for Vision Large Language Models.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Safety Fine-Tuning at (Almost) No Cost: A Baseline for Vision Large Language Models

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.751283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.751283Z digest=sha256:7e58c6d247d2dc0f1fd73cf801c24e05be7727dca8e86daa143aa966990dc082

Observation 0d7bae29-0166-43e4-84d7-a3053f47ca88 · outbound

This paper cites Understanding and Rectifying Safety Perception Distortion in VLMs.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Understanding and Rectifying Safety Perception Distortion in VLMs

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.755989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.755989Z digest=sha256:32a9d3d3d6b93cadec75b52e54f6e291af7dbc4d402af3c2e03a340ff06366b8

Observation fcbbe27f-914f-4b57-9fd5-cf34162e1e2e · outbound

This paper cites Image-to-Text Logic Jailbreak: Your Imagination can Help You Do Anything.

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning Image-to-Text Logic Jailbreak: Your Imagination can Help You Do Anything

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-15T14:24:49.761179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:24:49.761179Z digest=sha256:6895b7525eccb3710df125207306c6ad08ba3e11d9fbdd2e8101639626a56ab4

Pith citing papers

No inbound Pith citation observations are available.