Pith. sign in

Paper Citation Record · LEDGER

Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2312.00029.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.00029 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:32:35.049834Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:38:54.929027Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 75a8c917-8db3-4c88-9c11-423168de2c56 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

Reference 119

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:26.023640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:d913703e9d90eb2d716a690983187becdc76e65630fb7ecd135583625c63cc21

Observation f7ffed2a-a3a6-40f1-b947-b3953e6135d3 · inbound

The Dark Side of Trust: Authority Citation-Driven Jailbreak Attacks on Large Language Models cites this paper.

The Dark Side of Trust: Authority Citation-Driven Jailbreak Attacks on Large Language Models Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T18:38:43.793388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:38:43.793388Z digest=sha256:098a736a25f5ed83427ac238705527e6b3e7d15da948db2c2871732808bdcfea

Observation 1b6ebc3f-f1d0-4aca-94ae-fcb85c3cac89 · inbound

Safeguarding Large Language Models in Real-time with Tunable Safety-Performance Trade-offs cites this paper.

Safeguarding Large Language Models in Real-time with Tunable Safety-Performance Trade-offs Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T22:35:26.948582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:35:26.948582Z digest=sha256:df9676606c319645022ee18f854dcbfbef7fe18a6427eadb236b6f853e27c153

Observation 871466a3-1766-4a9a-a112-e67accee42a5 · inbound

Layer-Level Self-Exposure and Patch: Affirmative Token Mitigation for Jailbreak Attack Defense cites this paper.

Layer-Level Self-Exposure and Patch: Affirmative Token Mitigation for Jailbreak Attack Defense Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-10T22:12:00.096522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:12:00.096522Z digest=sha256:80eb55ca57f25a14de2d512b51d5a314f196f9b9fb70b9fdb7c1dc5f2d2514f8

Observation 2a99f675-b42e-4a9d-8bdf-0e73256ea8a1 · inbound

A Survey on Responsible LLMs: Inherent Risk, Malicious Use, and Mitigation Strategy cites this paper.

A Survey on Responsible LLMs: Inherent Risk, Malicious Use, and Mitigation Strategy Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

Reference 157

Resolution
unresolved
no resolver link, observed 2026-08-10T20:05:12.494905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:05:12.494905Z digest=sha256:bfa24d7bacf1d192c09921b3f9a54926d1ef3e8c46f6ef4f7346210f54b563e9

Observation 42ac8c90-73cd-4576-a87d-4b9e99366f39 · inbound

Adversarial Suffix Filtering: a Defense Pipeline for LLMs cites this paper.

Adversarial Suffix Filtering: a Defense Pipeline for LLMs Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:32:35.049834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:32:35.049834Z digest=sha256:e6ab452798a6027d69e93b2bde106ddacabe7257ee8e629ce3ff64d28d9d0279

Observation 55568abc-3be0-4866-894d-b65a02215dee · inbound

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments cites this paper.

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-19T01:02:54.785859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-19T01:02:07.088724Z digest=sha256:ec714879b22bbb7a6f714da9ed717dd525b3c3ad77861bafc4f1637efce38c56

Observation 5914d115-67d8-40fc-9bc3-059d850101bb · inbound

A Real-Time, Self-Tuning Moderator Framework for Adversarial Prompt Detection cites this paper.

A Real-Time, Self-Tuning Moderator Framework for Adversarial Prompt Detection Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T22:24:23.412222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:24:23.412222Z digest=sha256:a275ae8963e811255dd5601360e339231f960bb0576de939f4e92950cedbba38

Observation befcbf67-bbdb-4427-8a02-9bb4c5ce48e4 · inbound

Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs cites this paper.

Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:46:34.417482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T17:27:13.339411Z digest=sha256:420a94765ab0c9d6c61e878248a03b299b8f5481ee7de15785ab91edf8a5c2f4

Observation eb265621-6396-4512-a35a-bdcfdb2513ed · inbound

Cognitive Firewall: A Proactive, Zero-Trust, Multi-Gate Framework for LLM Safety cites this paper.

Cognitive Firewall: A Proactive, Zero-Trust, Multi-Gate Framework for LLM Safety Bergeron: Combating Adversarial Attacks through a Conscience-Based Alignment Framework

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:38:54.930590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-03T20:38:16.610310Z digest=sha256:18330613b3fda47ba3550ccddf95c552d6df0cb7d5eee02ed9473ad07dfda4a0