Pith. sign in

Paper Citation Record · LEDGER

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs

As of 10 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 2 inbound Pith citation observations for arXiv:2502.01436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01436 v3

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T15:22:33.166652Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T05:41:16.033862Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact0
  • verified fuzzy56
  • unresolved18
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 019206a3-e398-48c3-8d2a-51fafcc6dc3f · outbound

This paper cites Gpt in sheep’s clothing: The risk of customized gpts, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gpt in sheep’s clothing: The risk of customized gpts, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.088821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.120408Z digest=sha256:10fb02d0ae654ad7cd8b02613275d2970a92f33082dbf3bb5aec453da7535602

Observation 8e3e88d3-3bd0-4a26-9847-994d1ad0006b · outbound

This paper cites Safety-tuned llamas: Lessons from improving the safety of large language models that follow instructions, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safety-tuned llamas: Lessons from improving the safety of large language models that follow instructions, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.149793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.149793Z digest=sha256:d7f578c2a6ef35687b1a41d1f0695080915bb030efd3b97bef75f77413479efd

Observation c8940e90-d379-4370-8a2d-5d2b0c6d8e49 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.189930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.189930Z digest=sha256:f0b52dab78773f0207730924921ee09f03a6bd90faea5f94f3c4daa024d8224f

Observation f0821e67-8002-4e16-9d3b-4271f9d07b60 · outbound

This paper cites Personal data flows and privacy policy traceability in third-party llm apps in the gpt ecosystem.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Personal data flows and privacy policy traceability in third-party llm apps in the gpt ecosystem

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.070316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.233461Z digest=sha256:ba9b70cd5eec88a075348cf08cbf1e8234d1541ffcc518df020b169058207621

Observation 128a0391-36be-457a-a81e-777611e808f1 · outbound

This paper cites From representational harms to quality-of-service harms: A case study on llama 2 safety safeguards, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs From representational harms to quality-of-service harms: A case study on llama 2 safety safeguards, 2024

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.061966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.255082Z digest=sha256:77de817a29ce8cc32845485b52285ba99e7a1642942d5fcdcf5574716449a55d

Observation e3d1aed1-d34f-452b-8650-61f0c4268baf · outbound

This paper cites Avoiding social engineering and phishing attacks, 2021.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Avoiding social engineering and phishing attacks, 2021

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.053425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.291674Z digest=sha256:82d1e8c478b7f657ca06a72deabf2a4db70fe5d9698bd18f80f9619094b7cde0

Observation efa79ddc-f1db-4c0b-b04b-2a0fa67bc754 · outbound

This paper cites A security risk taxonomy for prompt-based interaction with large language models.IEEE Access, 12:126176–126187, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A security risk taxonomy for prompt-based interaction with large language models.IEEE Access, 12:126176–126187, 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.045036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.313658Z digest=sha256:2583a6f8efcf368f5003e600b6cbd4e59e03a296d02b0c2a7392354a3883b6b1

Observation 75cf4072-82ea-414d-8abb-1431e869e3f4 · outbound

This paper cites Transferable adversarial distribution learning: Query-efficient adversarial attack against large language models.Computers & Security, 135:103482, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Transferable adversarial distribution learning: Query-efficient adversarial attack against large language models.Computers & Security, 135:103482, 2023

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.036294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.355045Z digest=sha256:f45a8175f863e481b1b681e1289bc8b06d7c0c62395d4dc9d234082ea31b7985

Observation c3abdff4-fc02-42f0-ad35-89e0c8167593 · outbound

This paper cites European network for academic integrity.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs European network for academic integrity

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.027983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.393483Z digest=sha256:49000efcb95e60ce1dba5de34afcbb935d0456ab92ee17cffb9128d0108bf469

Observation 3b5ffd19-d5d1-42ed-acb1-3e3693a8b468 · outbound

This paper cites Lessons From an App Update at Replika AI: Identity Discontinuity in Human-AI Relationships.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Lessons From an App Update at Replika AI: Identity Discontinuity in Human-AI Relationships

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.415231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.415231Z digest=sha256:bc1deb5a641d040140c66fe39c92e3f215cff5bce9b576952b53d7d7f4ac634f

Observation c517757a-4414-4006-927f-8246b2feb784 · outbound

This paper cites A brief survey on safety of large language models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A brief survey on safety of large language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.019991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.418570Z digest=sha256:878d4aa8aa86e8f56958dd998361533dc16b2877d9afee1785f61a6d297a16d4

Observation 79d098ab-4ccd-40de-93f9-fa2427add1cd · outbound

This paper cites A survey on llm-as-a-judge, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A survey on llm-as-a-judge, 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.011644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.422120Z digest=sha256:cbf50fb0c2d606fc6f843121655dc99700149ac5c983433217b7ec66b359bd1a

Observation 796c0bb4-531b-46fe-96f9-ace5c7cf1f53 · outbound

This paper cites Shin, and Karl Aberer.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Shin, and Karl Aberer

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.003005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.425212Z digest=sha256:c0e4255e9b7a1bf7ec7a60723a617c95ce5596749698e865ae2ccb96fffa4a39

Observation e91660af-e30e-403d-ab26-33ac5ca7e2ad · outbound

This paper cites Safe lora: the silver lining of reducing safety risks when fine-tuning large language models, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safe lora: the silver lining of reducing safety risks when fine-tuning large language models, 2025

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.994172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.427528Z digest=sha256:2a0bf170f8ce87bd3e0d73d12e79f6ce50ee0fda4182e9a3a6f3d2c3fb6ff922

Observation 0aa76448-e68f-426f-93c1-85c60258fd16 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs LoRA: Low-Rank Adaptation of Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.430229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.430229Z digest=sha256:706a285feef3a37e8a5375044550807ca5522cd7f6d49964344a625c0949ed81

Observation ff451910-41f3-412e-9f6a-80116996bd68 · outbound

This paper cites Chatgpt sets record for fastest-growing user base - analyst note.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Chatgpt sets record for fastest-growing user base - analyst note

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.985307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.432814Z digest=sha256:c2e7f0ce92470b829dc42dab94630eb7c38c5d8c3ff7171a957834290cbb790e

Observation 21ac8166-2f30-4f1b-8248-c3820b29a4c3 · outbound

This paper cites An empirical study of llm-as-a-judge for llm evaluation: Fine-tuned judge model is not a general substitute for gpt-4, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs An empirical study of llm-as-a-judge for llm evaluation: Fine-tuned judge model is not a general substitute for gpt-4, 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.976162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.435219Z digest=sha256:d85730f7090ebdabfcaa80df9093cb99716ea7f5454d8cb9db054e6ec7789f08

Observation d7cdd1e6-a34d-4bc5-97fe-8f18f3e7e334 · outbound

This paper cites Lisa: Lazy safety alignment for large language models against harmful fine-tuning attack, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Lisa: Lazy safety alignment for large language models against harmful fine-tuning attack, 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.966940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.437506Z digest=sha256:df330e2cac6d4a802cd3a10306a82710f74bfcf6ed9b7becf0a9d0080d8c022d

Observation 6deed3c1-589e-402b-934f-1ffa8cc4fd4e · outbound

This paper cites International center for academic integrity, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs International center for academic integrity, 2025

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.957596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.439853Z digest=sha256:17abce8d97f73e33d8cf241bc66fa0b6e824ae5be2e6a0419f67c205b6db7601

Observation 7d5e9a72-4357-46dd-aa09-3fd404d6c595 · outbound

This paper cites Social engineering scams, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Social engineering scams, 2025

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.949030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.442463Z digest=sha256:4d6c56ac63da13db06cc17ff0db4fe38c024a095e150735a40dc2d309cd34b66

Observation edcc5c7d-8010-4b88-b130-cfc2943327fe · outbound

This paper cites Knowledge Sanitization of Large Language Models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Knowledge Sanitization of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.444763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.444763Z digest=sha256:d20dfdbfd3d6ba532b9767313cfe1dcca8ffb7c7f040032e64265e0bc09b5588

Observation 5bc4e5a5-5835-4139-85f4-97b6575470f1 · outbound

This paper cites Unleashing offensive artificial intelligence: Automated attack technique code generation.Computers & Security, 147:104077, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unleashing offensive artificial intelligence: Automated attack technique code generation.Computers & Security, 147:104077, 2024

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.940245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.448021Z digest=sha256:633c4730273fc5c605084d9ac5ad16f38270ece2912dbd5e46c6b60e7168c533

Observation 8e27a0cf-fe33-488e-a78e-2b977708ff1a · outbound

This paper cites Springer Nature, Cham, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Springer Nature, Cham, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.931574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.450617Z digest=sha256:a46fe03131fa7137837c2302155bec2a3893ce56467f40fe511f78557ca2b7f2

Observation 1a80b2dd-f164-44fa-9063-f1c696aa02f2 · outbound

This paper cites Fine-tuning, quantization, and llms: Navigating unintended outcomes, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Fine-tuning, quantization, and llms: Navigating unintended outcomes, 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.922797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.453216Z digest=sha256:b5e55f4c77d7cd30c344f0b7ac5f325bfbf0f90b45bb5a977e25641945817217

Observation 7067d695-8ca5-4056-b53e-2e1116a87a37 · outbound

This paper cites From generation to judgment: Opportunities and challenges of llm-as-a-judge, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs From generation to judgment: Opportunities and challenges of llm-as-a-judge, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.456473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.456473Z digest=sha256:f7c418ab781609edecc14841e57a0cc8b363713263b8b5becb8a8d6bc04a911a

Observation 3d8041f8-1dac-40b8-b1df-e5b28aadaac6 · outbound

This paper cites Safety Layers in Aligned Large Language Models: The Key to LLM Security.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safety Layers in Aligned Large Language Models: The Key to LLM Security

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.459148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.459148Z digest=sha256:b03e9d276bf9e633e5538bda1c6354c8349181bd27b4f42187571c5604e0760f

Observation 17e77498-c930-49a3-aa8d-0280f046d738 · outbound

This paper cites A hitchhiker’s guide to jailbreaking chatgpt via prompt engineering.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A hitchhiker’s guide to jailbreaking chatgpt via prompt engineering

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.908975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.462378Z digest=sha256:9f1e4b8252b6235698ac58748c07f4e4febe8dbb34a6901bfbcb83eac7e62904

Observation c15a3b88-39fe-4493-a2fe-a84ad87a6159 · outbound

This paper cites Privacy perceptions of custom gpts by users and creators.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Privacy perceptions of custom gpts by users and creators

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.898997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.465169Z digest=sha256:ec3b40f8a45db505134f9f907a95ff740bcfd4e8abbd721ed37993279d887111

Observation dfb84552-fd2c-46d5-8965-c4c1f68d79b8 · outbound

This paper cites The trauma floor: The secret lives of facebook moderators in america, 2019.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs The trauma floor: The secret lives of facebook moderators in america, 2019

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.889995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.468202Z digest=sha256:8f67e79be2372573d817eb93c3b25608e25a1c3f0d9c133c1fc3dfe6b4adfe52

Observation 96fd7c20-2fab-4f8c-8317-a369e2051504 · outbound

This paper cites Academic integrity policy.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Academic integrity policy

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.881187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.471454Z digest=sha256:bbdbc13cb8773f9bcaf51e98e51dd65d8a12c067819f33a7073a2069903e4bad

Observation 59934e44-e09e-46bf-a8d2-cea1ace3b58a · outbound

This paper cites Introducing gpts: Custom versions of chatgpt for specific purposes.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Introducing gpts: Custom versions of chatgpt for specific purposes

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.872017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.473936Z digest=sha256:ebf20c23ae88935bbc8dffd4e73f6fd046cdbb5ad5956dc156675a6d1bdbd884

Observation fb0695bd-ad5d-403b-b064-258d06df640c · outbound

This paper cites Openai red teaming network.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Openai red teaming network

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.863044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.491594Z digest=sha256:6475d72a75be8c6b32f70a201f4a006cfb44f9736b7681446a76bdf70bc852b7

Observation 4c2b21dc-97c9-4206-99b4-be88a6d863e8 · outbound

This paper cites Usage policies.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Usage policies

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.848870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.573950Z digest=sha256:c4a852651d00cda4e8a21c77e223a97723eddf333b404c7e9c0acda46a54ae2f

Observation 29475659-529b-43bf-b717-be1ed399902b · outbound

This paper cites Openai safety.https://openai.com/safety/, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Openai safety.https://openai.com/safety/, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.840005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.594578Z digest=sha256:9a616e5c10ed26c9391c5c46bade353ce12d11fc7aa0265b162d3818ee15f47d

Observation 35111379-daf7-41b5-bf94-bd653498ac32 · outbound

This paper cites LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.635523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.635523Z digest=sha256:87f7ac69ff4aa0fdd584ae625a113c02ebd9a97dd25298a22431b282e578834c

Observation 2f31e543-4401-43be-a6c0-718aecaa9d5c · outbound

This paper cites Puppeteer documentation.https://pptr.dev/, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Puppeteer documentation.https://pptr.dev/, 2025

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.830737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.658304Z digest=sha256:9b9a11f73108abe5b373b8dde64ec0d42bc46de403c3db73ce8bc96d65c8db73

Observation 471e7bbf-93b1-453b-8bb5-c89f19ba5bf3 · outbound

This paper cites Fine-tuning aligned language models compromises safety, even when users do not intend to!, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Fine-tuning aligned language models compromises safety, even when users do not intend to!, 2023

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.674137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.674137Z digest=sha256:001a3f523fc5d5f8b73d2f8a677ae2c621591e077d881f59036791527f039894

Observation 98371aba-6aba-4d08-aa63-aa8d6c067bdd · outbound

This paper cites Improving language understanding by gen- erative pre-training.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Improving language understanding by gen- erative pre-training

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.816249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.713618Z digest=sha256:cbb3f39e2bf36f4ece075bb7d1bd7c0c82943d66f0a60d9ba157c4a44291df04

Observation 6b274457-b211-46de-8783-2f8b7612ce17 · outbound

This paper cites Language models are un- supervised multitask learners.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Language models are un- supervised multitask learners

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.807470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.756161Z digest=sha256:c9d8ff7e36aed843095f74743f235373ff6a913605048faf4a272f07f9509de5

Observation e44caf72-b81e-4df1-9f8d-730de30625f5 · outbound

This paper cites Safetyprompts: a systematic review of open datasets for evaluating and improving large language model safety, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safetyprompts: a systematic review of open datasets for evaluating and improving large language model safety, 2025

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.798685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.778313Z digest=sha256:1b60efa8bb5e73bbd1ce4b94f41ad749dae73f73c2b0db694070ac4971ad6ea5

Observation a42d0a49-0ef3-4b51-9697-fd1ed52b5c29 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.789711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.781392Z digest=sha256:d34e217012f5ba89bb499ace2f6a81672207649383ec11af3e997014e6771d10

Observation 03c98b62-6759-45ac-9759-b201b335595d · outbound

This paper cites Identifying the provision of choices in privacy policy text.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Identifying the provision of choices in privacy policy text

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.781589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.784352Z digest=sha256:ad4fda48749d370f7a5094517139954e812a86fb77f89ec811ef1d21e6dc35f8

Observation eae0ff55-cb52-405e-be6a-cf38a95f9a76 · outbound

This paper cites do anything now.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs do anything now

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.773120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.787981Z digest=sha256:c7fc0093c60a12c144b1750b3494eafba5715d9e8c109b2a74bc608b28d5fbc0

Observation 993b6e8e-dafc-400d-87b1-53fa64133403 · outbound

This paper cites Can people experience romantic love for artificial intelligence? an empirical study of intelligent assistants.Information & Management, 59(2):103595, 2022.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Can people experience romantic love for artificial intelligence? an empirical study of intelligent assistants.Information & Management, 59(2):103595, 2022

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.764155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.790816Z digest=sha256:dfc633572800622c8e21289e1d0e34712dea3bd0e8f73ecf1649a9d374aeb1f8

Observation a3796ab8-c6a7-41a8-8de5-c93cada3d45d · outbound

This paper cites Gpt store mining and analysis, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gpt store mining and analysis, 2024

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.754983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.793685Z digest=sha256:62091de60c7b1443d06a85b3bcef7865d19a4accd045ddc109d8c197b9583ab4

Observation 45517465-1b61-421f-97ad-2d73bfc9beb5 · outbound

This paper cites Safety assessment of chinese large language models, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safety assessment of chinese large language models, 2023

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.747004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.796557Z digest=sha256:b2160d760921848588624be1e1cbba1658599e01e087d6d1bda5d9e9b35f601d

Observation b3876033-f39d-4c7e-8ae9-5d732a3fb23e · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.739176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.799371Z digest=sha256:f8e529bc466f0bde8129722270f3310aa97af18a38070b3c214231bf29cf0375

Observation 347d0509-2d67-442b-8a25-a489982f549d · outbound

This paper cites Opening a pandora’s box: Things you should know in the era of custom gpts, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Opening a pandora’s box: Things you should know in the era of custom gpts, 2023

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.731819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.802293Z digest=sha256:b5c7a94f9805e98475e17213fbcb199043ee08e2f1b8633d37bf169bb539ced5

Observation 590064e0-d5f3-49ba-ac50-7dc2a95e2c60 · outbound

This paper cites Glossary for academic integrity.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Glossary for academic integrity

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.722660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.805047Z digest=sha256:e04bbb77ed4e4615062ee88621321012726761066e58ea829b0863e29988e9e5

Observation a00a6bcc-86a1-4a61-b8f3-6203474f4c87 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.713883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.807937Z digest=sha256:d3846ba8cfcc0f0a8f3b61ff5b16323398ac1abcf4b0a01f2bc77167038935d1

Observation 39df5f71-a3d6-49a8-9302-a7db996f07a3 · outbound

This paper cites Academic misconduct.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Academic misconduct

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.704661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.811029Z digest=sha256:f5641bde88947d8ec6774beecb8c08959d24e528bf0a80eaf61268ffd5819342

Observation 852c1533-8ec3-4296-ae2f-e8203cfbc0e8 · outbound

This paper cites Definition of academic dis- honesty.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Definition of academic dis- honesty

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.694955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.814385Z digest=sha256:6f685840e54786f9d227bd1c19317409401ef05c1b83129dfabcd24e560fb224

Observation a7607b6e-583a-47bc-8416-f8becb42a33f · outbound

This paper cites Academic integrity.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Academic integrity

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.685474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.817707Z digest=sha256:3431d3bb1994dc2b749e6c54f3ece7961fe5e9404cf930eb374ff8b7add79be9

Observation ab08b9b0-3099-4ccf-9055-761cb3d8d5da · outbound

This paper cites Gomez, Łukasz Kaiser, and Illia Polosukhin.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gomez, Łukasz Kaiser, and Illia Polosukhin

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.820995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.820995Z digest=sha256:556009869793b4f222c2142ff25f724261642ce1fe26645802a523fd313d760a

Observation f2444aa6-e5ce-4254-bd50-e16ffdd3972d · outbound

This paper cites Definitions of academic miscon- duct.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Definitions of academic miscon- duct

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.669906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.824075Z digest=sha256:2c6c8686c4e41895fd0ed506cd21d064bbca2d1a1b46f7f25d41d2bc054bc080

Observation 903a1fa3-1d5f-44e5-87bb-826851e03fc3 · outbound

This paper cites Learning from failure: Integrating negative examples when fine-tuning large language models as agents, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Learning from failure: Integrating negative examples when fine-tuning large language models as agents, 2024

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.827016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.827016Z digest=sha256:0d3a9df4cc2133158c220d15c8dad9869dda00665629082624dda9e3239b0b13

Observation f88f1040-69cf-4a09-9085-f496493a26ca · outbound

This paper cites Taxonomy of risks posed by language models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Taxonomy of risks posed by language models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.655644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.829688Z digest=sha256:b951d0382f209fa5d3e80387ec416498df5e93fe2fc57b3e9d46f33d596fefd2

Observation 6622c4fa-41f8-4106-a16a-9f8737f4da04 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.629857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.836245Z digest=sha256:c36a3fa1907b347246627edb3fc86e07177b17856bad830b7d3b4fe6de6e7460

Observation ea3f80e2-1f5d-47e5-bb37-c31f05405e55 · outbound

This paper cites Sorry-bench: Systematically evaluating large language model safety refusal behaviors, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Sorry-bench: Systematically evaluating large language model safety refusal behaviors, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.571732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.838631Z digest=sha256:963da018a2fef3eb524652efcd5360c6f7f62266c0f6399565d7bef336e9b694

Observation f2bf9242-4514-4d3c-b87d-4b4654aa72e2 · outbound

This paper cites Online safety analysis for llms: a benchmark, an assessment, and a path forward, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Online safety analysis for llms: a benchmark, an assessment, and a path forward, 2024

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.532103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.841731Z digest=sha256:5993324c0bcf7bbd6f1ec8a27a67f752fe3a0442958fe4dbea6dfbc69213f67a

Observation 70693349-8fb7-4f65-a289-70c29520a72d · outbound

This paper cites Gptfuzzer: Red teaming large language models with auto-generated jailbreak prompts, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gptfuzzer: Red teaming large language models with auto-generated jailbreak prompts, 2024

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.475401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.844756Z digest=sha256:06eae13cb80861ed65556e34d0d6de6498ca0048c560e96b4f16614b2f03ba68

Observation 1fed986f-f0ce-405e-8e2a-01d1cdb2c521 · outbound

This paper cites Assessing prompt injection risks in 200+ custom gpts, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Assessing prompt injection risks in 200+ custom gpts, 2024

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.459166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.847843Z digest=sha256:42c080351f8695a93e62d5cf5fffa44af2538bc2249c20327ab0b4d55ffe25c6

Observation 74d5d262-79e3-4316-9f81-9bcd7c1c4e59 · outbound

This paper cites Don’t listen to me: Understanding and exploring jailbreak prompts of large language models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Don’t listen to me: Understanding and exploring jailbreak prompts of large language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.450746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.851059Z digest=sha256:2aa8360de166a32839926e3cc78c8a1c3f23cf7621201e314da9e93bb47e743c

Observation 9ae6ee4a-25e3-4f19-a649-5e7562288a64 · outbound

This paper cites S-eval: Automatic and adaptive test generation for benchmarking safety evaluation of large language models, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs S-eval: Automatic and adaptive test generation for benchmarking safety evaluation of large language models, 2024

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.441334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.854144Z digest=sha256:c5e00a6bc532988f6d746f37e26a5c9ca8ecf17c667e2a2e898195c9abcd0e30

Observation 1ea83626-cb04-48aa-a480-7260b2e984c7 · outbound

This paper cites Defending against neural fake news.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Defending against neural fake news

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.432710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.857217Z digest=sha256:8ac68c0ed4b2a7f3d7cef92b61e6124ff507a47dac366b0b1914a86faf247d41

Observation ae2878a1-586c-4726-b0b2-e97773dde8b4 · outbound

This paper cites A first look at gpt apps: Landscape and vulnerability, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A first look at gpt apps: Landscape and vulnerability, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.423728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.861267Z digest=sha256:212c7877cc36eff1b8071eefcd9b3aa857158c5b1ebeb6da00ee19749a36dfc8

Observation b418e139-c5d8-4ac3-9d91-5e6a280b0bb2 · outbound

This paper cites Safetybench: Evaluating the safety of large language models, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safetybench: Evaluating the safety of large language models, 2024

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.415091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.863976Z digest=sha256:c6b0433883100bfd64e48a30f99e85ec436224f94bc8c8f48132935cae3571a5

Observation cad04cee-4903-4d93-9d9d-fc21efce5e0b · outbound

This paper cites Gonzalez, and Ion Stoica.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gonzalez, and Ion Stoica

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.405755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.867298Z digest=sha256:b7e1513bcb697f5f9ace035f974264a9724b0b7c00dc0fd885869617ec9d5cfd

Observation 8dd9fb83-606b-4073-8f8f-23e85d190ee6 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.396366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.871158Z digest=sha256:d840011af89e2631a86ff6c2a76bb54aed25828193f552b1ebb4cc76e30d72a8

Observation 3e7fa3cb-27a0-4d77-95ce-8059de71bb33 · outbound

This paper cites Sadeh, Steven M.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Sadeh, Steven M

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.386264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.907645Z digest=sha256:d8c1bc2996e19b346e181064a97512ce8393faa768aa8a6d138bf25f9958a4f4

Observation 4178d1f1-f767-48a0-85eb-09523cb82d00 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.368102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:33.047162Z digest=sha256:6e378af0c11cf761c22e9c27426b5824d6303b254c94d8a6113e5df2d34c2e74

Observation 21458664-362c-46a1-9fa6-49e4d973ec5c · outbound

This paper cites boyfriends,.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs boyfriends,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.335204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:33.166652Z digest=sha256:36ae0f2c38318a8437c0dc1346e97f72b7ca9764325639c32ddbea848da11bfb

Observation 193da6e3-9ac3-47c5-83cc-a17fe592b1d1 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 2017

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.377671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.980639Z digest=sha256:df6bf7539c6aad839a5333c28e205b5ccec91cb3b313b5b3267dda9472cbc714

Observation ae820769-76cd-4f49-b248-5ca52beabe61 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.832855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.832855Z digest=sha256:53c33628abba24e9aea16d8a18e1e3f175aa2d6a9fbe5458d84e4237af3c748d

Observation 88074c97-2878-4dd5-9eea-0e6e96702f2a · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 2023

Resolution
parse uncertain
no resolver link, observed 2026-08-09T15:22:32.530750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.530750Z digest=sha256:2e74b2955c479ce2a6fd5c39b8d891a14c8cd1b469ae7996b2d984bf07d47d6d

Pith citing papers

Observation 921eaf41-b4b4-40e5-86da-0f8c587bf484 · inbound

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models cites this paper.

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-14T02:21:16.410064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:49:55.738870Z digest=sha256:a865d8beff66391b4251639b8cfdd2f3c0aac9234b7ab248c2dc2814ec8fed40

Observation cc402f8c-443f-4399-bf1b-17e2c5a7b0bd · inbound

Beyond Single-Policy: Evaluating Composed Organization-Specific Policy Alignment in LLM Chatbots cites this paper.

Beyond Single-Policy: Evaluating Composed Organization-Specific Policy Alignment in LLM Chatbots Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-07-14T02:21:16.410064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T05:41:16.033862Z digest=sha256:b519a76c5ce0d3be34f8ccbbd5c25cddb6f48d227945a9fb36fa40a940c741fd