Pith. sign in

Paper Citation Record · LEDGER

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment

As of 14 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2411.11543.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11543 v4

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:29:33.334887Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy32
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 32d9ab13-412e-44e9-91da-95889faf2b66 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.127861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.127861Z digest=sha256:1697cd357692530c98c6af69d5dc7130dcc7b143721ed82b0ed0a4eb9c873d68

Observation a0429e9e-e10e-4dc9-b726-c6df87f5c87d · outbound

This paper cites Image hijacks: Adversarial images can control generative models at runtime, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Image hijacks: Adversarial images can control generative models at runtime, 2023

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.949694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.133504Z digest=sha256:c3396137e6123af4453b111d1a47e09b55b443fa413d282940c75c84ef7083d1

Observation a6b154bb-a044-4e52-9bdb-6e121b68aaf8 · outbound

This paper cites Introducing our multimodal models, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Introducing our multimodal models, 2023

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.937976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.137541Z digest=sha256:f88b0eea4265be34a369383ca308d725d911afe42f74923b8595a55fe83b9e7a

Observation fc67f042-bb33-46d2-ba9f-4abd0af7faf9 · outbound

This paper cites Image safeguarding: Reasoning with conditional vision language model and obfuscating unsafe content counterfactually, 2024.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Image safeguarding: Reasoning with conditional vision language model and obfuscating unsafe content counterfactually, 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.925201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.141955Z digest=sha256:c7987f4bc271307ff100dc82913698e0158a4c7e5c38b0fc1121df1a0713f3ad

Observation 23467f4c-1b7a-44f7-b46d-b0f62c61e11d · outbound

This paper cites Honeybee: Locality-enhanced projector for multimodal llm.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Honeybee: Locality-enhanced projector for multimodal llm

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.913698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.146610Z digest=sha256:01e5d310979731e8683c1ac8dac85c182cd7b9f8ba73b3dbfa38ba62b1b09263

Observation 72440e1a-9935-4c87-8042-b190e1378edd · outbound

This paper cites Antifakeprompt: Prompt-tuned vision-language models are fake image detectors, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Antifakeprompt: Prompt-tuned vision-language models are fake image detectors, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.901712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.150340Z digest=sha256:44426b6897b2f7446eef36d38af5bcec36afde03a366d366c9a5ceae205d03ef

Observation 384f5638-c27c-4d99-804d-2d910caefcd1 · outbound

This paper cites ShareGPT4V: Improving Large Multi-Modal Models with Better Captions.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.154314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.154314Z digest=sha256:7506115f92d49dd82e56e5bf1efa03d386ea54a3994d6090d0b154ae521632f9

Observation 516ed2bc-7c04-4040-93ae-a7438301a33a · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.159094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.159094Z digest=sha256:22c78aa2dfedb443224e47559984e79cf7028f4e1ac95c3a284d170b47ca927a

Observation 4db07976-ba95-4c85-bf39-62f6d1e6375e · outbound

This paper cites Cogview: Mastering text-to-image generation via transformers.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Cogview: Mastering text-to-image generation via transformers

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.889327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.163265Z digest=sha256:aae5c27c7ff2be62512ac8916030d1c823ea98f5fd34137707929cd208a3de5c

Observation 8c68a258-628a-404e-9980-cc0ac6a8e8af · outbound

This paper cites InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.168107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.168107Z digest=sha256:ac1701266ec60ff0236bcdc061fa5eb1a4640245eff5d0ebe01af1ccbbdbaa96

Observation 82475f71-b3da-4f34-8dd4-0b3969378aae · outbound

This paper cites Glm: General language model pretraining with autoregressive blank infilling.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Glm: General language model pretraining with autoregressive blank infilling

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.876353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.173070Z digest=sha256:5ea23e58c1ee73c97264ff92aba9652e4d758f046c1489522d5f3ac3f2687c33

Observation 9fbfbe23-38ba-4fc4-b04b-bd5d62b480c1 · outbound

This paper cites Mme: A comprehen- sive evaluation benchmark for multimodal large language models, 2024.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Mme: A comprehen- sive evaluation benchmark for multimodal large language models, 2024

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.864604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.176842Z digest=sha256:2858ab3c11f6377847770fdb078b8a85b5eaf9f9d29ec6283d6be489d2248ac5

Observation 9f88201e-ac5f-4c0a-9473-c8ed226d9a55 · outbound

This paper cites Inducing high energy-latency of large vision-language models with verbose images, 2024.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Inducing high energy-latency of large vision-language models with verbose images, 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.853154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.179990Z digest=sha256:54eea2f01f62b2405af56dbc417b81da3171f5f5f8235354b2cfd41a3f1c19c0

Observation 54304045-3415-40dd-a318-4255ee4feecb · outbound

This paper cites Fig- step: Jailbreaking large vision-language models via typo- graphic visual prompts, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Fig- step: Jailbreaking large vision-language models via typo- graphic visual prompts, 2023

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.840982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.183160Z digest=sha256:21c6a2a321a2ed1901de1f2c281e41d61ca7934aab4f9e43a6fe0dcb6ac9e0c5

Observation aea55f20-49c2-4dca-bc3c-7819b1cb748c · outbound

This paper cites OneLLM: One Framework to Align All Modalities with Language.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment OneLLM: One Framework to Align All Modalities with Language

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.186515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.186515Z digest=sha256:ab139c45280b647d4f7a90bba1a55b609942bc2231a84ba5e4c2b9cdfe331eaf

Observation 4158e1b0-d3f4-484f-a553-a10d97d131ab · outbound

This paper cites Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen- Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen- Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.828033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.191274Z digest=sha256:90a07dbd6730b5a4f7d7b48c79fab4d5640052b23958a758204b35bcc0d6e024

Observation e8917f62-199c-4d5e-a97b-0054c61feb2a · outbound

This paper cites Nsfw data scraper.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Nsfw data scraper

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.815309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.195499Z digest=sha256:a381f992f5e6a27fd46fe40fd1412d59cae537ccb128cadde305d623419756dc

Observation b48f14e9-1907-4d64-80d8-83f050a9d852 · outbound

This paper cites Concept bottleneck models.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Concept bottleneck models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.803049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.200222Z digest=sha256:5201d2f229535482cfc68f2c77bb1d40f6a304ac1594e1721fe9446fd6bb388f

Observation 36d6e535-6381-4184-9ed1-bbba99defa90 · outbound

This paper cites A hierarchical approach for generating descriptive image paragraphs, 2017.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment A hierarchical approach for generating descriptive image paragraphs, 2017

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.790177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.204955Z digest=sha256:b2cf183cc4a4a9c9f79ee3a585e4b3e7882d152d86b22a174c017aa3e64b413d

Observation d651c935-b420-4915-a41c-e90ce28adeed · outbound

This paper cites SEED-Bench-2: Benchmarking Multimodal Large Language Models.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment SEED-Bench-2: Benchmarking Multimodal Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.209278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.209278Z digest=sha256:da3165a8f62dd6c084a6f7b40e8129c25a7471083b3c6ff9a1a0f4bc89a2010d

Observation 7da9cca8-247f-47a5-807c-394543fc7752 · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.214041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.214041Z digest=sha256:87c52b355a42e3e05a884c76b33a1457e7e485e4d62db4cdfa12436ce0a9b521

Observation 1af47637-f9f6-44a6-b7f7-cbc4b68ff7dd · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision- language understanding and generation.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Blip: Bootstrapping language-image pre-training for unified vision- language understanding and generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.218280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.218280Z digest=sha256:9c8ffe0ed500cec51be3d82f78ad5c6c1c79c04eef42f44321c83af98a3cddbe

Observation 6f1abdc0-db2c-47be-aa7d-21d262855cdd · outbound

This paper cites Blip- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Blip- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.223174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.223174Z digest=sha256:c5ee836c65e09dffc282239c241af0ee725d8a89706b7a472e24f0b98c1db22e

Observation ddd21dca-9a6a-47fb-9811-88397de0f38f · outbound

This paper cites Red teaming visual language models, 2024.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Red teaming visual language models, 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.759080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.227195Z digest=sha256:a861f8bd93daa1ecbd31c5fd7419730cd4181471222287873a91c71627a6f3d3

Observation 2b5a6728-d3b1-4891-b394-d0747968a253 · outbound

This paper cites Vl-trojan: Mul- timodal instruction backdoor attacks against autoregressive visual language models, 2024.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Vl-trojan: Mul- timodal instruction backdoor attacks against autoregressive visual language models, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.746772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.231094Z digest=sha256:46c48f2b624efc078d9e9af33da5ebfd2e006662618f3b72ec998f4774bde23e

Observation cb154169-5e40-4037-86cf-df5d4bb7bc15 · outbound

This paper cites Microsoft coco: Common objects in context.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Microsoft coco: Common objects in context

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.236310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.236310Z digest=sha256:ae4905bf061eed36bf50dbfac332d2bd2fb7b2caaefaa764fe2dea5e2cb79ed6

Observation d2f33554-7af2-4bb3-9c76-b3567be80bfe · outbound

This paper cites Improved baselines with visual instruction tuning, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Improved baselines with visual instruction tuning, 2023

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.726907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.240512Z digest=sha256:e4e5f4ef6d41dbed3b82e8dcc5fd069494f4bfb6939f1419813afac7b9e06bdf

Observation e7d6b7e4-f1e5-4632-ad78-443b1bf4867a · outbound

This paper cites Visual instruction tuning, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Visual instruction tuning, 2023

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.713400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.244894Z digest=sha256:a1493c702e98866c9f57dd53dc3e94dee4f8dc812e2a01667505c22ba35b79d4

Observation 07a51b75-5fc9-4e81-97af-1374d39169f6 · outbound

This paper cites Llava-next: Improved reason- ing, ocr, and world knowledge, 2024.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Llava-next: Improved reason- ing, ocr, and world knowledge, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.249485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.249485Z digest=sha256:754dcfa7b5c0be9dbd21b2ff471e22f2ad08915af1bd53ce9b15eb074776fa9d

Observation 7eeea764-5569-42c9-96b9-9964801e3ba9 · outbound

This paper cites Visual instruction tuning.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Visual instruction tuning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.253243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.253243Z digest=sha256:d2c82127edca85472b4fc2b443ceba01ca270a71c0ff3f0483f254760920413a

Observation 4a749d1e-df2a-4729-8ad2-43143ac3e85e · outbound

This paper cites Mm-safetybench: A benchmark for safety evaluation of multimodal large language models, 2024.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Mm-safetybench: A benchmark for safety evaluation of multimodal large language models, 2024

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.686040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.256901Z digest=sha256:9fcabf8b9cb49264a62f66bfcc63c58f2a4442602b5a52cb58f34bcc885919bf

Observation 21b551e1-6c26-4d88-b361-fa71c5e1742b · outbound

This paper cites Mmbench: Is your multi-modal model an all-around player?, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Mmbench: Is your multi-modal model an all-around player?, 2023

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.672265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.260390Z digest=sha256:3dc181a842c0de7fe699509f20a12f1eafab318344f402dbb283c156bf497af7

Observation 745a8f39-1d37-4cc8-a75d-73379872db57 · outbound

This paper cites Deep learning face attributes in the wild, 2015.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Deep learning face attributes in the wild, 2015

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.658621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.264015Z digest=sha256:bcc433fed559fe982c64c00c47825390484affdc16c4c9e9718dc522714ba08a

Observation 18ad0113-9428-4918-bb47-c6671d5f068a · outbound

This paper cites Interpretability Beyond Classification Output: Semantic Bottleneck Networks.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Interpretability Beyond Classification Output: Semantic Bottleneck Networks

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.267678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.267678Z digest=sha256:9d7509214e62d463932431d77cff0211d65727df1c661074f51fb4c670b6f8a3

Observation 83cd5ba9-4f87-48a2-8902-e80d08d2aa05 · outbound

This paper cites Stable bias: Evaluating societal representa- tions in diffusion models.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Stable bias: Evaluating societal representa- tions in diffusion models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.646418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.271764Z digest=sha256:b9779c58f100445d3b018f5bcf049cc0fa54f4c23efa0b7d060582105594ebda

Observation 7b50ee66-87f1-4bc2-918a-81570f5edf0c · outbound

This paper cites Gpt-4 technical report, 2024.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Gpt-4 technical report, 2024

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.275572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.275572Z digest=sha256:f7a77c1057a6282824ad75a595f55c6f737b250545b09656ac1b7c20615c5f21

Observation f7d1738d-f720-400f-9c35-38ef544e26a0 · outbound

This paper cites llama3-vision-alpha.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment llama3-vision-alpha

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.626757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.279385Z digest=sha256:f5ec2873880d4a5ee164680fadbc744ccbf94c7d2b2db5df65432fa0abac625d

Observation 85d538d1-28a7-42c8-99d5-2825c96c3f4f · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Learning transferable visual models from natural language supervision, 2021

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.282948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.282948Z digest=sha256:fcefb85e8426a66038eb8802b5bbf63638961828241dc1a3cd7f55820344ba8b

Observation 1184df5d-0577-4fd7-a415-3f016dd08efd · outbound

This paper cites XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.286362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.286362Z digest=sha256:a0b24b806ec156c67d4e09c1159069a7ba8746dedd027b67ac80498a9f976ed2

Observation 3989e4ee-c4bd-4600-9925-a68e3db72797 · outbound

This paper cites How many unicorns are in this image? a safety evaluation benchmark for vision llms, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment How many unicorns are in this image? a safety evaluation benchmark for vision llms, 2023

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.608038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.290064Z digest=sha256:fe4861f3254dde151dc2643e9ed7ba046b4915e4a9783c879e68eda05ac7bb91

Observation ac426aa2-3471-4d53-8eab-6f9773db2503 · outbound

This paper cites Towards understanding and detecting cyberbullying in real-world images.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Towards understanding and detecting cyberbullying in real-world images

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.596340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.293833Z digest=sha256:611d354a832836ecdcfa0c9829f488fa48338aab9e3a40d4ade39e5173950314

Observation c669f67b-2509-4b8a-b24d-16cda2f2e17a · outbound

This paper cites Knowledge mining with scene text for fine-grained recognition, 2022.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Knowledge mining with scene text for fine-grained recognition, 2022

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.582483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.297669Z digest=sha256:1683336ba0db9d5358ff99702b2df123d65ff98189c6422d30bebed95c608017

Observation e68a4896-fd60-4c66-bc0d-4fb006ef4ea3 · outbound

This paper cites Adversarial prompt tuning for vision-language models, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Adversarial prompt tuning for vision-language models, 2023

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.566740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.306703Z digest=sha256:48684d894eb9c3bb6c9c83666a9c625d41d3cd295820b8e501c5d2bf6344137e

Observation 9699e6de-3b32-43c8-89e5-9f4e5394b79a · outbound

This paper cites A mutation-based method for multi-modal jailbreaking attack detection, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment A mutation-based method for multi-modal jailbreaking attack detection, 2023

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.550558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.312676Z digest=sha256:22a9312567bbc671277a50ac4e99a8abfe926a787e86a50eeb9147f1b249a18f

Observation 12984441-25fb-4e15-86b5-f7b0491619dd · outbound

This paper cites Meta-Transformer: A Unified Framework for Multimodal Learning.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Meta-Transformer: A Unified Framework for Multimodal Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T18:29:33.316995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:29:33.316995Z digest=sha256:f0ffd8a14f1cf4f48be29db496bf98b88a012e20354e442bde1c5e003fcddb1c

Observation eaf18c04-9c75-459b-9f2f-c8cb34a113b7 · outbound

This paper cites Privacyalert: A dataset for image privacy prediction.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Privacyalert: A dataset for image privacy prediction

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.536365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.321251Z digest=sha256:fa0a99b5b20f534b6412b014b7c30fe4bbcd4afade90e9ddd560a482b46e7878

Observation bddfeaf5-f538-465d-a979-3561447de2d4 · outbound

This paper cites On evaluating adversarial robustness of large vision-language models, 2023.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment On evaluating adversarial robustness of large vision-language models, 2023

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.520984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.326565Z digest=sha256:6cc1f687358e1c6b77899b036a77fcf9cc6e2a445a2a1e68d1eb8bfd8beeca0d

Observation 098cacae-68eb-49fe-875d-ea5afa1a7156 · outbound

This paper cites Manning, Christopher Potts, and Danqi Chen.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment Manning, Christopher Potts, and Danqi Chen

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.507548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.330724Z digest=sha256:7d482841db3cfeab865d6fed8741c9df084aace394bffa8bb0c41cf412bb266c

Observation 50236ac9-7aec-4818-a11d-6968cc81379e · outbound

This paper cites A" and "B.

PSA-VLM: Enhancing Vision-Language Model Safety through Progressive Concept-Bottleneck-Driven Alignment A" and "B

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:29:33.494338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T18:29:33.334887Z digest=sha256:6226b5d864b75840641589062128cf8a8c1c55354b2df548ff0a048633c42a35

Pith citing papers

No inbound Pith citation observations are available.