Pith. sign in

Paper Citation Record · LEDGER

Qwen3Guard Technical Report

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 81 inbound Pith citation observations for arXiv:2510.14276.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.14276 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 81 of 81 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 81 of 81 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:43:01.890905Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a2ee35e7-4f83-48c8-ba17-2020e2e83b0e · inbound

OI-Bench: An Option Injection Benchmark for Evaluating LLM Susceptibility to Directive Interference cites this paper.

OI-Bench: An Option Injection Benchmark for Evaluating LLM Susceptibility to Directive Interference Qwen3Guard Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T09:39:14.242702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:39:14.242702Z digest=sha256:19e8292d4e17b3be690c7140656f96783c6873c4f85fceeacc7de39742df22a5

Observation 3cf92171-df0c-43c1-af28-a441cfe0d64b · inbound

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking cites this paper.

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking Qwen3Guard Technical Report

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-03T05:10:16.641500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:10:16.641500Z digest=sha256:d32c6ae4d6231cf861aa52345802caeb621bd00b0da01beea9cfd2ebc3564b75

Observation 1192903c-7a2f-4236-8f6f-d311c236161e · inbound

Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation cites this paper.

Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation Qwen3Guard Technical Report

Reference 23

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T06:07:25.534823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T06:07:23.060966Z digest=sha256:851ee1a3ec51df7f200c1396f00b83fd7d6a5c73e68dbb8c4f69d7c13eafa29d

Observation ac3f4e19-64f3-49b7-ba25-df2614565de2 · inbound

When the Prompt Becomes Visual: Vision-Centric Jailbreak Attacks for Large Image Editing Models cites this paper.

When the Prompt Becomes Visual: Vision-Centric Jailbreak Attacks for Large Image Editing Models Qwen3Guard Technical Report

Reference 2025

Resolution
malformed identifier
no resolver link, observed 2026-08-03T01:22:20.614946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:22:20.614946Z digest=sha256:a8e2ba9bb41f8003c124ef6a3c939b85afa041e88d5d39cd2fcca076f815ecbd

Observation ef62d42f-acf5-4d33-a413-8fccf9ef0315 · inbound

Mitigating the Safety-utility Trade-off in LLM Alignment via Adaptive Safe Context Learning cites this paper.

Mitigating the Safety-utility Trade-off in LLM Alignment via Adaptive Safe Context Learning Qwen3Guard Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T23:32:42.145394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:32:42.145394Z digest=sha256:09dfcaa71464a8e78c10e33511a298f11aceeb5aa251d902d8ee9d23a33b12c8

Observation 953ff1f0-7746-4841-ad91-8df6107ca5f0 · inbound

Towards Cross-lingual Values Judgment: A Consensus-Pluralism Perspective cites this paper.

Towards Cross-lingual Values Judgment: A Consensus-Pluralism Perspective Qwen3Guard Technical Report

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-15T21:20:18.557992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T21:17:21.852857Z digest=sha256:81067ace1fa94893291e26831007cb611668b197d953de1fe2460caab43a0fb5

Observation d61d6741-9d06-49bb-8e60-f052ea538963 · inbound

FlexGuard: Continuous Risk Scoring for Strictness-Adaptive LLM Content Moderation cites this paper.

FlexGuard: Continuous Risk Scoring for Strictness-Adaptive LLM Content Moderation Qwen3Guard Technical Report

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T19:00:15.677990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T18:56:40.332443Z digest=sha256:ca887c5faf82a37aad98fa7b347122152e9493d58f21985f558fd586388c01e0

Observation 089d5599-51c5-4e85-863e-53799a233b26 · inbound

Adaptively Robust LLM Monitoring via Activation Watermarking cites this paper.

Adaptively Robust LLM Monitoring via Activation Watermarking Qwen3Guard Technical Report

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T17:39:43.471239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:39:43.471239Z digest=sha256:1d04254be52cc146659c9e21ee4a4418d8e85c06d0abd4caf968a309b181ff92

Observation 77548c0f-f68b-4302-b493-eeeabb6d6ba5 · inbound

Cooking Up Risks: Benchmarking and Reducing Food Safety Risks in Large Language Models cites this paper.

Cooking Up Risks: Benchmarking and Reducing Food Safety Risks in Large Language Models Qwen3Guard Technical Report

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T22:02:12.555644Z digest=sha256:be5b05b363e38bf79181d25fc6af3b13eceb7ea10096a6b6ee35a0ac52aba7b3

Observation ee9d3177-6c38-429a-a682-9f1fc0279b8e · inbound

ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis cites this paper.

ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis Qwen3Guard Technical Report

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T21:17:15.523370Z digest=sha256:8c5311fd396c0e2b1049ee92a831a4197bf325288a034609fe9ac6cd5eaa83eb

Observation 0100edde-10da-426a-9a50-ffa15b8d4a14 · inbound

ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis cites this paper.

ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis Qwen3Guard Technical Report

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-14T22:03:02.661178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-14T22:02:00.638549Z digest=sha256:2e1923060d4bdda878633f61a354175a5b747153f34a3ec1c8c7ff50bb221feb

Observation 6aae269a-4fcd-4655-b67d-1b764e5a7858 · inbound

AgentHazard: A Benchmark for Evaluating Harmful Behavior in Computer-Use Agents cites this paper.

AgentHazard: A Benchmark for Evaluating Harmful Behavior in Computer-Use Agents Qwen3Guard Technical Report

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T19:42:53.778980Z digest=sha256:462a2b538641e5eaf832aa570fda7c3bf65e89370f8e352e545041e15ee2a41d

Observation acb2e677-16d9-4437-bf87-1d765d248428 · inbound

Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming cites this paper.

Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming Qwen3Guard Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-13T11:43:31.089245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T11:43:31.089245Z digest=sha256:e5aee322e904a3cb3fa4465c3228acb0a83d6ba6642537778e53ab0e14fa0307

Observation 895d1e27-5556-4a2e-baa9-e6716d4636dc · inbound

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense cites this paper.

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense Qwen3Guard Technical Report

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T18:26:19.922383Z digest=sha256:701aa2a68bc49052f6147e5767ee131d1e326dc926fe7b8fa2f82356dd2d10c0

Observation c1b0cf5e-4bce-459e-90a4-579e2c7709b9 · inbound

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs cites this paper.

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs Qwen3Guard Technical Report

Reference 130

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:04:05.157103Z digest=sha256:0e98e4e755316059c301299650db8794fc3336ced411ac40c076314ef0f15d22

Observation 96198d4b-94d5-4a0a-9e6d-936350d542ae · inbound

Targeted Interpretable Safety Neuron Enhancement for Multilingual Vision-Language Large Models cites this paper.

Targeted Interpretable Safety Neuron Enhancement for Multilingual Vision-Language Large Models Qwen3Guard Technical Report

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:12:34.909229Z digest=sha256:7f6248acfefa33e6caca8b897d9de33d0df8b5667d4351a5a29f8f5dad80b684

Observation 5c2330a1-1240-412a-a106-d5c8f896d5c8 · inbound

Targeted Interpretable Safety Neuron Enhancement for Multilingual Vision-Language Large Models cites this paper.

Targeted Interpretable Safety Neuron Enhancement for Multilingual Vision-Language Large Models Qwen3Guard Technical Report

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T16:34:08.890416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:34:08.890416Z digest=sha256:74a1237acc92c1a2423ca79f99ea34fc3eb116830c1d6907a8b2ccdff0b1da08

Observation 59c4354f-ad08-4875-90f2-886c83a31690 · inbound

Conflicts Make Large Reasoning Models Vulnerable to Attacks cites this paper.

Conflicts Make Large Reasoning Models Vulnerable to Attacks Qwen3Guard Technical Report

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:24:25.255602Z digest=sha256:5853f8ab212ebbdc5c930399eb47bf1534c0e5bd5bffdc3a27f3a4ebcfb48e4e

Observation 7a8df8ac-0519-4525-b995-b6cff0ada22a · inbound

WebAgentGuard: A Reasoning-Driven Guard Model for Detecting Prompt Injection Attacks in Web Agents cites this paper.

WebAgentGuard: A Reasoning-Driven Guard Model for Detecting Prompt Injection Attacks in Web Agents Qwen3Guard Technical Report

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:12:51.787807Z digest=sha256:62a640dfa9549272f2555a3d86282cf9740097ff369e3293ccdac9ba958396f6

Observation 9907746f-35e6-4bc4-9247-2c87138b28bf · inbound

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs cites this paper.

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs Qwen3Guard Technical Report

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:06:43.140580Z digest=sha256:45d951f64c78952aad58c851c1e19c558e5a4672d4a162491063e95d6cba9b8d

Observation b97785af-a3bc-40ae-8253-aaa8cc275766 · inbound

TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts cites this paper.

TWGuard: A Case Study of LLM Safety Guardrails for Localized Linguistic Contexts Qwen3Guard Technical Report

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T09:07:57.713675Z digest=sha256:4ab75666b0ca740a0286aae880005515d982bf4b9cadc62086dcbfb43f16b2bc

Observation 7f1df501-f39a-4bbc-937e-38143a422378 · inbound

LLM Safety From Within: Detecting Harmful Content with Internal Representations cites this paper.

LLM Safety From Within: Detecting Harmful Content with Internal Representations Qwen3Guard Technical Report

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T04:33:54.058475Z digest=sha256:5c3b9f44c3a2c6ae837cba7578e6ccc1863ac19d265af4d969ba6b07fb379bcd

Observation cbb5fa31-2f11-4671-951b-e95dfa0cdeb3 · inbound

Cross-Lingual Jailbreak Detection via Semantic Codebooks cites this paper.

Cross-Lingual Jailbreak Detection via Semantic Codebooks Qwen3Guard Technical Report

Reference 20

Resolution
malformed identifier
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T16:22:33.567085Z digest=sha256:dc6373deccd4f4d539a31ae7839799eb2ed9016768408ad8f48e89a38311ef73

Observation 6ee575a7-71b0-483c-b980-ac83ad56f12c · inbound

ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models cites this paper.

ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models Qwen3Guard Technical Report

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T19:14:31.076067Z digest=sha256:5db959804250c0880858846ffd3a6691274a812f32d5f98ee57c5ea26eb57f98

Observation 741ebe01-1101-4823-9dcf-8eb789514f8f · inbound

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety cites this paper.

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety Qwen3Guard Technical Report

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T16:00:32.413225Z digest=sha256:2349906f7ddaaa86e06baa55adda39a64ef5b350cee61389bc3cd1bb71c1c770

Observation c20d8a6a-bf83-46fb-a9ad-df26c85d8ea5 · inbound

RouteHijack: Routing-Aware Attack on Mixture-of-Experts LLMs cites this paper.

RouteHijack: Routing-Aware Attack on Mixture-of-Experts LLMs Qwen3Guard Technical Report

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T19:22:00.217729Z digest=sha256:3d62f2809eda6b527675fa6a0b2ee18ebf7adc68cdb3bd46fb81422306ccad69

Observation a3e94ded-a735-4a6e-bb79-334960950d46 · inbound

One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue cites this paper.

One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue Qwen3Guard Technical Report

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T11:17:19.079380Z digest=sha256:235804c5e2092ea3ba96c437216aea1e24a1da76fab3087b963e03cc56d0e2f6

Observation dfe079d6-5def-4232-8c0e-88f2c01d0992 · inbound

One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue cites this paper.

One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue Qwen3Guard Technical Report

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:53:06.928500Z digest=sha256:1eaeecb2ecd0bf6048f164b4d6a2ab4a7f296b5266e61ef4ad28c27475ae84e8

Observation 8bf3d127-5e3c-477e-85fe-4539c4873800 · inbound

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence cites this paper.

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence Qwen3Guard Technical Report

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T10:12:58.421050Z digest=sha256:0c81e566d7514d826eb46c2f9078942ee07c238cf37cfb2fe3f8e68705009aba

Observation e2cba8ce-94e2-4df4-b2b4-7ed0663236e0 · inbound

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence cites this paper.

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence Qwen3Guard Technical Report

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T00:56:48.838028Z digest=sha256:5d83965a3882a9b14b3fee6dde5a21a6fd70d2dca203ce108641cfe78480fb79

Observation a69cbbfd-bae0-439e-81a0-ceffc8106d1d · inbound

GLiGuard: Schema-Conditioned Classification for LLM Safeguard cites this paper.

GLiGuard: Schema-Conditioned Classification for LLM Safeguard Qwen3Guard Technical Report

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T03:19:11.495230Z digest=sha256:9c728ab55d9b881991c57de6ba4b612c3e351c0b9561ad432ee9968b45351faf

Observation a91efb1d-f110-4ee8-98c6-6bce8c59cbdc · inbound

Why Do Aligned LLMs Remain Jailbreakable: Refusal-Escape Directions, Operator-Level Sources, and Safety-Utility Trade-off cites this paper.

Why Do Aligned LLMs Remain Jailbreakable: Refusal-Escape Directions, Operator-Level Sources, and Safety-Utility Trade-off Qwen3Guard Technical Report

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T01:36:41.651502Z digest=sha256:9e8afe6290a1e1694f4fd64c7f129ceabcaf2eddd5fc53027f4c3ba93f89846f

Observation 21944711-f9ad-4068-b05e-38b942855f32 · inbound

Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories cites this paper.

Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories Qwen3Guard Technical Report

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T02:07:04.364417Z digest=sha256:cbdbfe21bfddb2cf606c6af979273742ba13a80c6492b779a417d5c57a41c059

Observation f25c482f-3f03-461c-b915-41c5c9b2e414 · inbound

MT-JailBench: A Modular Benchmark for Understanding Multi-Turn Jailbreak Attacks cites this paper.

MT-JailBench: A Modular Benchmark for Understanding Multi-Turn Jailbreak Attacks Qwen3Guard Technical Report

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T01:24:24.242731Z digest=sha256:c37833157e8ffeb9b2d641640e9adab5cd0897b69ae1aaa27c0b3ea596bb7eb2

Observation 38c12b81-ec92-4a54-9533-a081addca1aa · inbound

On-Policy Self-Evolution via Failure Trajectories for Agentic Safety Alignment cites this paper.

On-Policy Self-Evolution via Failure Trajectories for Agentic Safety Alignment Qwen3Guard Technical Report

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:33:37.835217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T06:29:32.776647Z digest=sha256:3a995ec20b5ddbf0eb2b6ed6bb4bb239783281362ec624cffda4678b3a5a6d2e

Observation 6f7cd408-60e2-4749-a641-bda07f119b0b · inbound

PropGuard: Safeguarding LLM-MAS via Propagation-Aware Exploration and Remediation cites this paper.

PropGuard: Safeguarding LLM-MAS via Propagation-Aware Exploration and Remediation Qwen3Guard Technical Report

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:33:48.181891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T22:33:12.568495Z digest=sha256:16879017b4e19cb580e48b0ac8eb89a9e7f3359faa02de6528d960658d9e6fcc

Observation dc84d407-2b46-47fe-9f78-5f426fb8978c · inbound

LPG: Balancing Efficiency and Policy Reasoning in Latent Policy Guardrails cites this paper.

LPG: Balancing Efficiency and Policy Reasoning in Latent Policy Guardrails Qwen3Guard Technical Report

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-19T23:43:17.768883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T23:42:58.135553Z digest=sha256:d3657790f63a49f500bff77070bf4d66feec3e4570e1e7283e88a6a2f5b75656

Observation 58d4bf75-186e-4dc1-b5a8-bbb1bd9ba4a0 · inbound

ASEval: Automated Trajectory-Level Security Testing for Autonomous Agents cites this paper.

ASEval: Automated Trajectory-Level Security Testing for Autonomous Agents Qwen3Guard Technical Report

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-22T05:51:08.490229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T05:49:21.532568Z digest=sha256:e876ddc8988f06bcd81a9d2dc1620e31be47297bad0c739bff13dea7fe2ee9cb

Observation 22cbf2ee-a3bb-487a-9f31-5a3dc50a0cab · inbound

ASEval: Automated Trajectory-Level Security Testing for Autonomous Agents cites this paper.

ASEval: Automated Trajectory-Level Security Testing for Autonomous Agents Qwen3Guard Technical Report

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T13:28:38.489351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:28:38.489351Z digest=sha256:0d5dc2fa7b749cc65f43331b77e0ccb32e0ac4ba1aca370506898c95bb93c524

Observation addcadd6-44c4-40d7-88bf-852a058f9a43 · inbound

Boundary-targeted Membership Inference Attacks on Safety Classifiers cites this paper.

Boundary-targeted Membership Inference Attacks on Safety Classifiers Qwen3Guard Technical Report

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T08:11:17.564431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T08:06:45.896332Z digest=sha256:d52c9c527753b812e666ca150dc97fe7e188d7170c5d154ee2030b249401fa30

Observation 806d8ec1-9d40-4c23-a64f-9197ae35b049 · inbound

Boundary-targeted Membership Inference Attacks on Safety Classifiers cites this paper.

Boundary-targeted Membership Inference Attacks on Safety Classifiers Qwen3Guard Technical Report

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T05:45:24.364866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T05:42:55.999769Z digest=sha256:4ac45e2fc082fe889366994fb999ec64c2997ba4dd2eee8850474260c710aadf

Observation dfa74ac4-01f4-4a8b-aea0-3baaffb1bc6a · inbound

AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue cites this paper.

AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue Qwen3Guard Technical Report

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-06-30T21:35:04.717507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T21:28:25.577739Z digest=sha256:c98f525b523a29900afcab6ecfd8e70210b45f5a96b5cb30324c19422305aca2

Observation 102c8e1e-8392-4f15-8e9b-34698f4f9b99 · inbound

HARP: Measuring Harm Amplification in Multi-Agent LLM Systems cites this paper.

HARP: Measuring Harm Amplification in Multi-Agent LLM Systems Qwen3Guard Technical Report

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-06-29T17:23:44.848183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T17:20:09.363041Z digest=sha256:d193f6317327dca823844fe8574aa1e5d7bb90a6a0c7941c381d23087d4c3873

Observation f9c544c2-2970-4610-8795-75b28cedc916 · inbound

Can It Reach the Generator? Investigating the Survival of Prompt-Injection Attacks in Realistic RAG Settings cites this paper.

Can It Reach the Generator? Investigating the Survival of Prompt-Injection Attacks in Realistic RAG Settings Qwen3Guard Technical Report

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-06-29T12:03:24.289784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T11:55:34.911818Z digest=sha256:0102bac8de6197fe9ddc218fd2278b4f4b05ab8f27eec2aa0efd50f39e947c66

Observation ff6d13ed-0b0c-4314-91bd-5fe17d65b2c0 · inbound

Refusal Before Decoding: Detecting and Exploiting Refusal Signals in Intermediate LLM Activations cites this paper.

Refusal Before Decoding: Detecting and Exploiting Refusal Signals in Intermediate LLM Activations Qwen3Guard Technical Report

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-29T12:13:26.524287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T12:11:06.703658Z digest=sha256:a1982ca7d0a2982cfc1cbc5b806b9b44551b5da8f74804f45f75e3d64ab16c86

Observation 05e3e30e-2f45-44d1-b966-62c353dbe840 · inbound

Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models cites this paper.

Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models Qwen3Guard Technical Report

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:13.519898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T08:01:39.412431Z digest=sha256:bfded4c945df5fa63489e0469d4bf0a34607b1199d9a46e8eb837b7aa9b89c51

Observation a0d915d8-b543-402c-8299-c6941d727df3 · inbound

Triaging Threats to Specialized Guardrails cites this paper.

Triaging Threats to Specialized Guardrails Qwen3Guard Technical Report

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-06-28T22:32:44.290259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T22:27:44.703466Z digest=sha256:e7d67b46a53517cc3654b7f15330f6be216db7ff32588d6376138f147b788d7c

Observation 32818b3e-6f64-4ab0-8415-7e41f8c08743 · inbound

EvoDefense: Co-Evolving Black-Box Defense with Large Language Models cites this paper.

EvoDefense: Co-Evolving Black-Box Defense with Large Language Models Qwen3Guard Technical Report

Reference 5

Resolution
malformed identifier
local_arxiv, observed 2026-07-01T19:56:10.437273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:58:13.411712Z digest=sha256:8033c8d4bc8195c2735d7445218fd08523a284cbe932442bce67d1e98577289d

Observation 34a6731d-7ffc-462b-8994-6888630cd5de · inbound

Learning from Mistakes: Can LLM Self-Recover after Misalignment? cites this paper.

Learning from Mistakes: Can LLM Self-Recover after Misalignment? Qwen3Guard Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-13T18:51:10.298187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:51:10.298187Z digest=sha256:08f242e72599241ae8a2ccf4da100b75fd65f74470a23fded0b40800b78b353b

Observation 999dbe27-ee5c-4ed0-bbd7-4c2ac2b91cca · inbound

BraveGuard: From Open-World Threats to Safer Computer-Use Agents cites this paper.

BraveGuard: From Open-World Threats to Safer Computer-Use Agents Qwen3Guard Technical Report

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-07-01T21:26:14.200911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T17:04:46.912915Z digest=sha256:55adc9ee851b09c5f5ae6f373db6798fb016443b9e8079bc703adb0068ab4ceb

Observation a3361792-4000-4263-9bbe-dc4bf030ba50 · inbound

SentGuard: Sentence-Level Streaming Guardrails for Large Language Models cites this paper.

SentGuard: Sentence-Level Streaming Guardrails for Large Language Models Qwen3Guard Technical Report

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T22:56:20.738256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T14:51:17.667213Z digest=sha256:6a1d9714fced6aca99e00d2e1dca32dfe60d79e2c98d9d300db206ee3d852c7e

Observation 87f650aa-4c74-4889-9be3-07c16aecf429 · inbound

Investigating and Alleviating Harm Amplification in LLM Interactions cites this paper.

Investigating and Alleviating Harm Amplification in LLM Interactions Qwen3Guard Technical Report

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T23:16:24.051535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T14:31:52.027889Z digest=sha256:1403a7afb71ca5be00c5230ca11cb5b6a730d209cdca850d572d92ef8e9c29e5

Observation 7ca2c2a7-a9f7-47be-a0a5-efce46f2b391 · inbound

D-Judge: Disrupting Multi-Turn Jailbreaks using Semantics-Preserving Output Rewriting cites this paper.

D-Judge: Disrupting Multi-Turn Jailbreaks using Semantics-Preserving Output Rewriting Qwen3Guard Technical Report

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-01T21:16:14.792525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T17:14:15.815256Z digest=sha256:2689b718bc8cf2f177ea291ef4b990f7ac3260e53ac22e8366b361a9bcf7abea

Observation ea4e50c6-120e-4086-bb78-194ac08be0c5 · inbound

DDOR: Delta Debugging for Explainable Overrefusal Testing and Repair cites this paper.

DDOR: Delta Debugging for Explainable Overrefusal Testing and Repair Qwen3Guard Technical Report

Reference 27

Resolution
malformed identifier
local_arxiv, observed 2026-06-28T10:32:00.153885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T08:54:19.113501Z digest=sha256:ce4561be295a5ed77484619538482556d45c0257f8f063b3035f14057ec753e2

Observation 10a7bd08-76f6-408d-888f-9fac76c66858 · inbound

RUBAS: Rubric-Based Reinforcement Learning for Agent Safety cites this paper.

RUBAS: Rubric-Based Reinforcement Learning for Agent Safety Qwen3Guard Technical Report

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-02T02:16:26.540113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T11:07:47.814115Z digest=sha256:f7f8a047f3b1fcf5a17e3cdae4a1b01084063fbfa8eca7a73146f72fc87749ff

Observation 2e26bcb3-a5f4-4548-aa0a-58d142f7486f · inbound

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs cites this paper.

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs Qwen3Guard Technical Report

Reference 254

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T01:07:30.265011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T16:49:14.243931Z digest=sha256:24f3416518b081e98e1b2b08a8b4a2e4a5ad3d6ced2379be420e5ed1b8a215a8

Observation 5677da66-5809-4ddd-96d1-8d009b22d61b · inbound

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs cites this paper.

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs Qwen3Guard Technical Report

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T07:36:45.291932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T06:49:29.653094Z digest=sha256:963025ee59a2ac0a3bbd2f90da860fbff045eccdfd032f7ba8e85b70c9390e1e

Observation 48d2f00c-23e7-478f-84d4-d479a9a27744 · inbound

Greedy Coordinate Diffusion: Effective and Semantically Coherent Adversarial Attacks via Diffusion Guidance cites this paper.

Greedy Coordinate Diffusion: Effective and Semantically Coherent Adversarial Attacks via Diffusion Guidance Qwen3Guard Technical Report

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T17:08:43.472292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T04:35:35.594085Z digest=sha256:a5b5dde12a7a7b50e346e13136dd4a1ed11085c1a0f7af1fc9dd714a84d01916

Observation 7866f3d9-92a2-4f16-af08-7305969aeec6 · inbound

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling cites this paper.

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling Qwen3Guard Technical Report

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-04T03:59:33.713380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:23:06.382761Z digest=sha256:fe64f7280ea525f2cce282cdef46ba618c84cc3432b580346a4acf8a1bc5601c

Observation 919a3143-2067-4700-906d-12b4341afadc · inbound

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models cites this paper.

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models Qwen3Guard Technical Report

Reference 132

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T19:30:07.687112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-25T21:16:36.392606Z digest=sha256:7318b24184f463f7492c826b141e28f68317ce283ef125ca4e4b60be98c3b3c8

Observation 9b5ea55a-911f-4b47-aa53-578cc0a2dbb6 · inbound

PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models cites this paper.

PolicyAlign: Direct Policy-Based Safety Alignment for Large Language Models Qwen3Guard Technical Report

Reference 44

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T19:40:06.686892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-25T21:09:19.727723Z digest=sha256:4480f88543a814b713ed342de194848fc13b16b0f26d43ba887856e4cde8bc9b

Observation a9ec080c-6f31-4bee-85eb-fcd68ae0362a · inbound

RedVox: Safety and Fairness Gaps in Speech Models Across Languages cites this paper.

RedVox: Safety and Fairness Gaps in Speech Models Across Languages Qwen3Guard Technical Report

Reference 178

Resolution
verified exact
local_arxiv, observed 2026-07-04T13:59:52.792091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T04:37:00.399470Z digest=sha256:aaeaeed8fba4d05086f2ec0ec017fe981187ffa52e5016a460edaa6f8a20d64c

Observation 840de7ff-2bce-4795-9354-4a3ba6a3be6b · inbound

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety cites this paper.

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety Qwen3Guard Technical Report

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-06-29T00:52:55.753869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T00:46:03.210076Z digest=sha256:491445bda6418edb74610dc010a74a05724130eff37abc1cda74e47e49a57fb0

Observation 451b577f-e355-4210-91c9-1b1ae3569c9d · inbound

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models cites this paper.

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models Qwen3Guard Technical Report

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-07-01T17:35:51.339657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T03:35:34.594617Z digest=sha256:2d75f88d67c8e0f049ff8a7bebd86121775063c7b1d77a2c87891eb3792f6130

Observation 18f8e2e8-a9b5-4ac6-943e-d2f0b157d95a · inbound

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing cites this paper.

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing Qwen3Guard Technical Report

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-06-30T06:34:18.837556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T06:31:55.631719Z digest=sha256:f16b105580c9de7fdb10c6ea0fea5bf2bf20d4bfb68626e94495698a8973bd4c

Observation 5f22d630-37e7-4659-af2c-cde6ae8b2ec9 · inbound

Safety Testing LLM Agents at Scale: From Risk Discovery to Evidence-Grounded Verification cites this paper.

Safety Testing LLM Agents at Scale: From Risk Discovery to Evidence-Grounded Verification Qwen3Guard Technical Report

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-07-03T13:48:20.362672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T13:45:02.167211Z digest=sha256:f9acf29c87cf1f1b4bdbeaf829cd9e6714e10e341451dc161b5b0ba598500f45

Observation eca486b9-8716-48c0-a4c2-d5b3aa0e9e42 · inbound

HaloGuard 1.0: An Open Weights Constitutional Classifier for Multilingual AI Safety cites this paper.

HaloGuard 1.0: An Open Weights Constitutional Classifier for Multilingual AI Safety Qwen3Guard Technical Report

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T14:48:32.676927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T14:38:55.045628Z digest=sha256:80b585d257ad2770aedd99f2cc92cf2a2994e3e7dec1ff48f7cb16aa793db8ff

Observation a08cd464-6fe2-4b80-a965-4cee34df93bc · inbound

SafeGuard: A Multi-Agent Perception-Reasoning Framework for Social-Risk AI-Generated Video Detection cites this paper.

SafeGuard: A Multi-Agent Perception-Reasoning Framework for Social-Risk AI-Generated Video Detection Qwen3Guard Technical Report

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-12T05:07:18.364787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:07:18.364787Z digest=sha256:92b7937224b566bd42e41d6501c86dd82d6cb377a18c9fc10ec2360ff0b52327

Observation 33f58612-2cd1-4783-bbab-edb5e84dd497 · inbound

DT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety Guardrail cites this paper.

DT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety Guardrail Qwen3Guard Technical Report

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-08T10:04:51.438848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-08T10:00:02.677030Z digest=sha256:6c5d1dd81a8ac40dc5c900deec29b073120065a749d9bf3fa15db30adb204982

Observation 9b7e8732-b495-48ad-bfd8-cb0bf3f02011 · inbound

MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment cites this paper.

MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment Qwen3Guard Technical Report

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-14T07:16:57.009797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:16:57.009797Z digest=sha256:969e799aba2639ea83cc1a63300051c7bb66ec3a7c39a493a08ad6a39414da63

Observation a4eeda0a-0ae7-430e-ac72-7db9b183733a · inbound

TRACE: Trajectory-Based Safety Patch Learning for LLM Post-Training Realignment cites this paper.

TRACE: Trajectory-Based Safety Patch Learning for LLM Post-Training Realignment Qwen3Guard Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T10:00:11.484540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:00:11.484540Z digest=sha256:b23a78e940d8e5ab1784316411c13f307df1f51a92437ed4e44bd328ada9f415

Observation 74607a4b-5e4b-4619-96f8-bb475d4356f1 · inbound

Dynamic Defense Profiling Enables Cognitive Jailbreak of Text-to-Image Models cites this paper.

Dynamic Defense Profiling Enables Cognitive Jailbreak of Text-to-Image Models Qwen3Guard Technical Report

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T17:07:55.050266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:07:55.050266Z digest=sha256:ec6e14a8fd2da8f224ffbd49b01cbffe699f84697d985942f4e3e7be4969bf16

Observation 7aeaac08-ab9b-4bbb-be7c-2051410b33bf · inbound

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs cites this paper.

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs Qwen3Guard Technical Report

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T16:21:27.339175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:21:27.339175Z digest=sha256:d9443c53acac743e41ce31a1d29d837f91e54a43eb5c66e0476b51630735c856

Observation 394f09dc-e471-49c1-82ba-a426fa996511 · inbound

Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework cites this paper.

Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework Qwen3Guard Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T12:28:42.789128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:28:42.789128Z digest=sha256:a7bbc6375c3a5b7a36486554dc272687416aec613bae1057ee1f1fe69aa049c0

Observation 6d40fd64-6589-481e-9e53-08b0119ad890 · inbound

DARWIN: Evolving Jailbreak Adversary and Guardrail for LLM Safety Evaluation and Protection cites this paper.

DARWIN: Evolving Jailbreak Adversary and Guardrail for LLM Safety Evaluation and Protection Qwen3Guard Technical Report

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-01T11:40:29.208494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:40:29.208494Z digest=sha256:10ed13a345462c0a70819b11caa53b1376afb09cc0474976bdb48713971d0209

Observation a8dbd1ec-f3d7-4490-9551-643f179de68d · inbound

JANUS: Foreseeing Latent Risk for Long-Horizon Agent Safety cites this paper.

JANUS: Foreseeing Latent Risk for Long-Horizon Agent Safety Qwen3Guard Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T11:24:49.027826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T11:24:49.027826Z digest=sha256:f51f8082255a17e32431002e01b8f4c30700dff1758eb3348a6c668865348e92

Observation ed0c3849-3f89-460d-9de2-e723ffc92b01 · inbound

Forecasting Trajectory-Level Safety Risks in Black-Box Multi-Turn Interactions cites this paper.

Forecasting Trajectory-Level Safety Risks in Black-Box Multi-Turn Interactions Qwen3Guard Technical Report

Reference 88

Resolution
unresolved
no resolver link, observed 2026-07-30T20:11:00.888291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T20:11:00.888291Z digest=sha256:b7e290acde9f7bce623c2b40683e25a9f04437b1e4db8635d41e4a9a4a67a29f

Observation e2e490ba-21df-4a62-a7d0-872b3bd31cd8 · inbound

One Anchor for All: Unified Multilingual and Multimodal Safety Alignment for LVLMs cites this paper.

One Anchor for All: Unified Multilingual and Multimodal Safety Alignment for LVLMs Qwen3Guard Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-31T23:07:38.966539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T23:07:38.966539Z digest=sha256:f9e7ee9a2814e0df95ba76f65654a96f57095b41d682325ec423b656fff9687f

Observation 78c107ad-85f8-49d4-85c9-18d18ad73bde · inbound

SafeNexus: Discovering and Steering Modality-Universal Safety Neurons in MLLMs cites this paper.

SafeNexus: Discovering and Steering Modality-Universal Safety Neurons in MLLMs Qwen3Guard Technical Report

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T16:20:26.886091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T16:20:26.886091Z digest=sha256:fba0c58eb109c776d95df658fa83bf838e7c3c5c60dbf075c91f629b5fbdf1c5

Observation 9ce19e8a-ced2-4fe2-a713-61ce7b8e48ca · inbound

Mind the Gap: Zero-Query Jailbreaks via Filter-Generator Discrepancy in Text-to-Image Systems cites this paper.

Mind the Gap: Zero-Query Jailbreaks via Filter-Generator Discrepancy in Text-to-Image Systems Qwen3Guard Technical Report

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T00:43:01.890905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:43:01.890905Z digest=sha256:545dcefb0dd8b3bc3f2285782a9b2c31980b87726da4ad082133371e99e62837

Observation bc178fed-7253-42ef-866b-51cd6a4120b8 · inbound

When Refusal Looks Safe: The Refusal-Cue Shortcut in Safety Guard Models cites this paper.

When Refusal Looks Safe: The Refusal-Cue Shortcut in Safety Guard Models Qwen3Guard Technical Report

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T23:46:20.553544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:46:20.553544Z digest=sha256:0ed76dc2c08ec0765f1b0be2a7d9f7a32bae86cb77addaba4f2a26c9281d96f8