Pith. sign in

Paper Citation Record · LEDGER

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs

As of 10 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 2 inbound Pith citation observations for arXiv:2502.01436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01436 v3

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T15:22:33.166652Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T05:41:16.033862Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact0
  • verified fuzzy56
  • unresolved18
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 019206a3-e398-48c3-8d2a-51fafcc6dc3f · outbound

This paper cites Gpt in sheep’s clothing: The risk of customized gpts, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gpt in sheep’s clothing: The risk of customized gpts, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.088821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.120408Z digest=sha256:206086f7ca0e4002df81dba520b344ede8329751a3c16d6150fc84e61b0282f6

Observation 8e3e88d3-3bd0-4a26-9847-994d1ad0006b · outbound

This paper cites Safety-tuned llamas: Lessons from improving the safety of large language models that follow instructions, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safety-tuned llamas: Lessons from improving the safety of large language models that follow instructions, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.149793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.149793Z digest=sha256:f5ecf7b2d2f82236b202a5f410face703cd13703776bddc7e7a67263c17d43a5

Observation c8940e90-d379-4370-8a2d-5d2b0c6d8e49 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.189930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.189930Z digest=sha256:23b12a2ec7666d5d6784c989a2fdc3ef73a21ec23cbd78b498cd58a66fb93b61

Observation f0821e67-8002-4e16-9d3b-4271f9d07b60 · outbound

This paper cites Personal data flows and privacy policy traceability in third-party llm apps in the gpt ecosystem.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Personal data flows and privacy policy traceability in third-party llm apps in the gpt ecosystem

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.070316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.233461Z digest=sha256:bbb3811d04adb9b93836680fc2c903f5bf900adafa2cc9fcb45589aa8102f839

Observation 128a0391-36be-457a-a81e-777611e808f1 · outbound

This paper cites From representational harms to quality-of-service harms: A case study on llama 2 safety safeguards, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs From representational harms to quality-of-service harms: A case study on llama 2 safety safeguards, 2024

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.061966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.255082Z digest=sha256:bd5e29a22ce18a8e8a2a724e51a7a32b10d8e51dbe6769cb5c8efe1fcad0451b

Observation e3d1aed1-d34f-452b-8650-61f0c4268baf · outbound

This paper cites Avoiding social engineering and phishing attacks, 2021.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Avoiding social engineering and phishing attacks, 2021

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.053425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.291674Z digest=sha256:fb560ea8266f794d76b31df244b95545011c0ace4358d9d4884af8e6ae3f6a5a

Observation efa79ddc-f1db-4c0b-b04b-2a0fa67bc754 · outbound

This paper cites A security risk taxonomy for prompt-based interaction with large language models.IEEE Access, 12:126176–126187, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A security risk taxonomy for prompt-based interaction with large language models.IEEE Access, 12:126176–126187, 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.045036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.313658Z digest=sha256:bb1a1d9018ce9ba88278fd7290e87b439ec454b346af428e9a8dce5008bb3fee

Observation 75cf4072-82ea-414d-8abb-1431e869e3f4 · outbound

This paper cites Transferable adversarial distribution learning: Query-efficient adversarial attack against large language models.Computers & Security, 135:103482, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Transferable adversarial distribution learning: Query-efficient adversarial attack against large language models.Computers & Security, 135:103482, 2023

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.036294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.355045Z digest=sha256:53fb9e02c3f90ad26905de697c37d912388a9b1cee82423bfe6c156fdf1f62dd

Observation c3abdff4-fc02-42f0-ad35-89e0c8167593 · outbound

This paper cites European network for academic integrity.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs European network for academic integrity

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.027983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.393483Z digest=sha256:51a5db490375de44f99ea0cbdccb2fe89ba193609c6db186a7fe30af0036c6ac

Observation 3b5ffd19-d5d1-42ed-acb1-3e3693a8b468 · outbound

This paper cites Lessons From an App Update at Replika AI: Identity Discontinuity in Human-AI Relationships.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Lessons From an App Update at Replika AI: Identity Discontinuity in Human-AI Relationships

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.415231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.415231Z digest=sha256:cc56f2c43c25084aeaf2a73e9bb2c90df98f2215d43ddcfdec6c9700991752a6

Observation c517757a-4414-4006-927f-8246b2feb784 · outbound

This paper cites A brief survey on safety of large language models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A brief survey on safety of large language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.019991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.418570Z digest=sha256:e5541aab26fdd0b4026b349cd1488429af56def087b7a3d5b8d998ca963f95d8

Observation 79d098ab-4ccd-40de-93f9-fa2427add1cd · outbound

This paper cites A survey on llm-as-a-judge, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A survey on llm-as-a-judge, 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.011644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.422120Z digest=sha256:d4a5dcc52ff1440f1a459100c7fb2ef996744e28d9af7c0eb681a5a63e5760fa

Observation 796c0bb4-531b-46fe-96f9-ace5c7cf1f53 · outbound

This paper cites Shin, and Karl Aberer.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Shin, and Karl Aberer

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.003005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.425212Z digest=sha256:857753a263424923805540623259370620229feba22b3b7924b0ab12179b45c7

Observation e91660af-e30e-403d-ab26-33ac5ca7e2ad · outbound

This paper cites Safe lora: the silver lining of reducing safety risks when fine-tuning large language models, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safe lora: the silver lining of reducing safety risks when fine-tuning large language models, 2025

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.994172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.427528Z digest=sha256:250efce06aa6fcf2a687df175af1367731d74057682a450de8e18bbab371b59f

Observation 0aa76448-e68f-426f-93c1-85c60258fd16 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs LoRA: Low-Rank Adaptation of Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.430229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.430229Z digest=sha256:1f67ea698f69fcdfd04794f32a4cfa36bea84793ea6784722b35f9a0371fd6a0

Observation ff451910-41f3-412e-9f6a-80116996bd68 · outbound

This paper cites Chatgpt sets record for fastest-growing user base - analyst note.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Chatgpt sets record for fastest-growing user base - analyst note

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.985307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.432814Z digest=sha256:22f09be02795fb0acacd54fdf5a37d69f5585e538dd4b58adeb1ee4ca99a5f32

Observation 21ac8166-2f30-4f1b-8248-c3820b29a4c3 · outbound

This paper cites An empirical study of llm-as-a-judge for llm evaluation: Fine-tuned judge model is not a general substitute for gpt-4, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs An empirical study of llm-as-a-judge for llm evaluation: Fine-tuned judge model is not a general substitute for gpt-4, 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.976162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.435219Z digest=sha256:f1d66d0e51f7de4baca1151b79abffc89f9ec4861af220607a9bdccf8df8fb68

Observation d7cdd1e6-a34d-4bc5-97fe-8f18f3e7e334 · outbound

This paper cites Lisa: Lazy safety alignment for large language models against harmful fine-tuning attack, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Lisa: Lazy safety alignment for large language models against harmful fine-tuning attack, 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.966940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.437506Z digest=sha256:425662298dfde2fe557b6bc9869fa39d3f1540b4ccd46b73cdc948e88eead430

Observation 6deed3c1-589e-402b-934f-1ffa8cc4fd4e · outbound

This paper cites International center for academic integrity, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs International center for academic integrity, 2025

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.957596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.439853Z digest=sha256:64114f2f426ddad6c1f0908c5753329bf230432fa2ce04ee16796de934ed32b4

Observation 7d5e9a72-4357-46dd-aa09-3fd404d6c595 · outbound

This paper cites Social engineering scams, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Social engineering scams, 2025

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.949030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.442463Z digest=sha256:4e4b4a09e4e81248dc255fcf6f72d7e6111242f8245493e0ab36f3683bc18119

Observation edcc5c7d-8010-4b88-b130-cfc2943327fe · outbound

This paper cites Knowledge Sanitization of Large Language Models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Knowledge Sanitization of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.444763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.444763Z digest=sha256:de27fdd6db634326dc44076d08878d6e421d79dbffb272902a7eb0fb7e261701

Observation 5bc4e5a5-5835-4139-85f4-97b6575470f1 · outbound

This paper cites Unleashing offensive artificial intelligence: Automated attack technique code generation.Computers & Security, 147:104077, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unleashing offensive artificial intelligence: Automated attack technique code generation.Computers & Security, 147:104077, 2024

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.940245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.448021Z digest=sha256:d9b17d023e9446387b2dc8e74b8d8aaf31f4549afbf7057fdaf4e6b5ddf8d39b

Observation 8e27a0cf-fe33-488e-a78e-2b977708ff1a · outbound

This paper cites Springer Nature, Cham, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Springer Nature, Cham, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.931574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.450617Z digest=sha256:ede3d31447486d884b72a30a4c872217358243ee53f25c4fcf716b71b4a7ae85

Observation 1a80b2dd-f164-44fa-9063-f1c696aa02f2 · outbound

This paper cites Fine-tuning, quantization, and llms: Navigating unintended outcomes, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Fine-tuning, quantization, and llms: Navigating unintended outcomes, 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.922797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.453216Z digest=sha256:0bb8d67a240668bf079081f46ada0455f7c768234dac59695f14d9c45517db01

Observation 7067d695-8ca5-4056-b53e-2e1116a87a37 · outbound

This paper cites From generation to judgment: Opportunities and challenges of llm-as-a-judge, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs From generation to judgment: Opportunities and challenges of llm-as-a-judge, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.456473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.456473Z digest=sha256:af7d126087a0f7a0f9db61b91e9a86d5cb5fd45882102ab997bbb81e3667999a

Observation 3d8041f8-1dac-40b8-b1df-e5b28aadaac6 · outbound

This paper cites Safety Layers in Aligned Large Language Models: The Key to LLM Security.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safety Layers in Aligned Large Language Models: The Key to LLM Security

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.459148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.459148Z digest=sha256:ab6dbe05f2cd2cad1e1d48f3e219a8f930cff13c5607e7d8856a13c89cf8ba9a

Observation 17e77498-c930-49a3-aa8d-0280f046d738 · outbound

This paper cites A hitchhiker’s guide to jailbreaking chatgpt via prompt engineering.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A hitchhiker’s guide to jailbreaking chatgpt via prompt engineering

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.908975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.462378Z digest=sha256:a367001ab85bb2acee2f89d67c2222757b6a051d66e0cd40ce3d2d6feb6c43fe

Observation c15a3b88-39fe-4493-a2fe-a84ad87a6159 · outbound

This paper cites Privacy perceptions of custom gpts by users and creators.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Privacy perceptions of custom gpts by users and creators

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.898997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.465169Z digest=sha256:ae37090d4d704ed7349ff79d2ae62c37491dc3dbf21a577106a2c5625b0615ba

Observation dfb84552-fd2c-46d5-8965-c4c1f68d79b8 · outbound

This paper cites The trauma floor: The secret lives of facebook moderators in america, 2019.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs The trauma floor: The secret lives of facebook moderators in america, 2019

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.889995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.468202Z digest=sha256:d1230cb5bc3e5ceb7d6e6b06d3d2d58445cfeceaaba7c63560e7e25ebe687274

Observation 96fd7c20-2fab-4f8c-8317-a369e2051504 · outbound

This paper cites Academic integrity policy.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Academic integrity policy

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.881187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.471454Z digest=sha256:f151da32b4c696e2f1ac9aac9e9a1b25f1abdad929d15f145e36251438933cfc

Observation 59934e44-e09e-46bf-a8d2-cea1ace3b58a · outbound

This paper cites Introducing gpts: Custom versions of chatgpt for specific purposes.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Introducing gpts: Custom versions of chatgpt for specific purposes

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.872017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.473936Z digest=sha256:c30b0c900ee694c3c151c61f5226f60103217fddb27f976e90aaa91bc5346d68

Observation fb0695bd-ad5d-403b-b064-258d06df640c · outbound

This paper cites Openai red teaming network.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Openai red teaming network

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.863044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.491594Z digest=sha256:24cb8c1c08daee72677d6c0e0d407b3a40ffa6a73694086e6e2b2f4d0ed0d0c2

Observation 4c2b21dc-97c9-4206-99b4-be88a6d863e8 · outbound

This paper cites Usage policies.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Usage policies

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.848870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.573950Z digest=sha256:5ad5b4413b558d7f3e0b2203bbe2826c806cc66a9e6b0d003114b3f5a1f6b56d

Observation 29475659-529b-43bf-b717-be1ed399902b · outbound

This paper cites Openai safety.https://openai.com/safety/, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Openai safety.https://openai.com/safety/, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.840005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.594578Z digest=sha256:91268557e44ff89e183abf5255ad27c8498efd5c4c1a239619d32215c69d0bfe

Observation 35111379-daf7-41b5-bf94-bd653498ac32 · outbound

This paper cites LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.635523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.635523Z digest=sha256:c2e75ac4669326a90564d3a8b5e355f9a876a5cccff45362fbec04b31f96a670

Observation 2f31e543-4401-43be-a6c0-718aecaa9d5c · outbound

This paper cites Puppeteer documentation.https://pptr.dev/, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Puppeteer documentation.https://pptr.dev/, 2025

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.830737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.658304Z digest=sha256:b37671473b0dbbc04c8efc2b49fe79d46520bd81eb084e5342cbd7efb8042d2c

Observation 471e7bbf-93b1-453b-8bb5-c89f19ba5bf3 · outbound

This paper cites Fine-tuning aligned language models compromises safety, even when users do not intend to!, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Fine-tuning aligned language models compromises safety, even when users do not intend to!, 2023

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.674137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.674137Z digest=sha256:0c4838cf957eb52094f41bedd58952425e6382bc3195fc6f36b619947f618a80

Observation 98371aba-6aba-4d08-aa63-aa8d6c067bdd · outbound

This paper cites Improving language understanding by gen- erative pre-training.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Improving language understanding by gen- erative pre-training

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.816249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.713618Z digest=sha256:e371ba7a83d419ae0e8aab728be7a63456ef1b4c7204c1f8c4d81983d76dbdc7

Observation 6b274457-b211-46de-8783-2f8b7612ce17 · outbound

This paper cites Language models are un- supervised multitask learners.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Language models are un- supervised multitask learners

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.807470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.756161Z digest=sha256:d810f7e9601d001987fdf61d261e4b072bb557cd700bb426aaea6781ad60a63b

Observation e44caf72-b81e-4df1-9f8d-730de30625f5 · outbound

This paper cites Safetyprompts: a systematic review of open datasets for evaluating and improving large language model safety, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safetyprompts: a systematic review of open datasets for evaluating and improving large language model safety, 2025

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.798685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.778313Z digest=sha256:ee6ecb5824975bb1419aa47082087a2b227b50d561ee26dd9e8d9f260f3dcd42

Observation a42d0a49-0ef3-4b51-9697-fd1ed52b5c29 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.789711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.781392Z digest=sha256:8cba2b285d53ff8178fdf4512c05fdb0b7a8bc0a4655e6c29d367dbbbb4f163d

Observation 03c98b62-6759-45ac-9759-b201b335595d · outbound

This paper cites Identifying the provision of choices in privacy policy text.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Identifying the provision of choices in privacy policy text

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.781589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.784352Z digest=sha256:c25bbb29b894bf9d0ec6a74ac48f0b84e20ffe284d6a8fd5c74895c454168a9c

Observation eae0ff55-cb52-405e-be6a-cf38a95f9a76 · outbound

This paper cites do anything now.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs do anything now

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.773120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.787981Z digest=sha256:7220e1326c225f9a24566bbb393c2da96f4e735ecde9a36b4c90a35fa154b0ef

Observation 993b6e8e-dafc-400d-87b1-53fa64133403 · outbound

This paper cites Can people experience romantic love for artificial intelligence? an empirical study of intelligent assistants.Information & Management, 59(2):103595, 2022.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Can people experience romantic love for artificial intelligence? an empirical study of intelligent assistants.Information & Management, 59(2):103595, 2022

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.764155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.790816Z digest=sha256:01c2006c8583d0e00c22cdf163578cb2ce9017cbd56756b6affad6062459a30b

Observation a3796ab8-c6a7-41a8-8de5-c93cada3d45d · outbound

This paper cites Gpt store mining and analysis, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gpt store mining and analysis, 2024

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.754983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.793685Z digest=sha256:32d861bc08ebab009927c302c7d00bb906904ee9ff13c4ac8461023a88f9bf1d

Observation 45517465-1b61-421f-97ad-2d73bfc9beb5 · outbound

This paper cites Safety assessment of chinese large language models, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safety assessment of chinese large language models, 2023

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.747004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.796557Z digest=sha256:e510ba480a97b3d77396addb15744ae4c971b20809a45e35674abbb74f537a0a

Observation b3876033-f39d-4c7e-8ae9-5d732a3fb23e · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.739176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.799371Z digest=sha256:dd5bb61603c42e9c3b32155561ccf66bb393d5cf694dfed676ee2b1531c91109

Observation 347d0509-2d67-442b-8a25-a489982f549d · outbound

This paper cites Opening a pandora’s box: Things you should know in the era of custom gpts, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Opening a pandora’s box: Things you should know in the era of custom gpts, 2023

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.731819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.802293Z digest=sha256:475012aeaf9b4f0ae1beaf588993f76dbc1a15f6f379951377fdf69b85f4f762

Observation 590064e0-d5f3-49ba-ac50-7dc2a95e2c60 · outbound

This paper cites Glossary for academic integrity.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Glossary for academic integrity

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.722660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.805047Z digest=sha256:171a2e8db2ac78c1ab23bcbafa1ee1f3caa536d5ae3c1b79fb8b8d8b47608468

Observation a00a6bcc-86a1-4a61-b8f3-6203474f4c87 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.713883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.807937Z digest=sha256:ba2f6fbd38d18237d1be3deb14734adcbbddf1ee6fc83ae33d79e98ac95e71fe

Observation 39df5f71-a3d6-49a8-9302-a7db996f07a3 · outbound

This paper cites Academic misconduct.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Academic misconduct

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.704661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.811029Z digest=sha256:0e1b72f02ad7f98e1d1cf1d999ebd0b8fd16caf1b11707529a125bd48cf5a0b2

Observation 852c1533-8ec3-4296-ae2f-e8203cfbc0e8 · outbound

This paper cites Definition of academic dis- honesty.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Definition of academic dis- honesty

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.694955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.814385Z digest=sha256:0d656e19e7bf0299e93b3d29493ea2f36281b752085fdb482f541021a7509730

Observation a7607b6e-583a-47bc-8416-f8becb42a33f · outbound

This paper cites Academic integrity.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Academic integrity

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.685474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.817707Z digest=sha256:752fc6257e12fe3b5c57f287b0d5c704892c81c2f4d5b13749ab258fd830141d

Observation ab08b9b0-3099-4ccf-9055-761cb3d8d5da · outbound

This paper cites Gomez, Łukasz Kaiser, and Illia Polosukhin.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gomez, Łukasz Kaiser, and Illia Polosukhin

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.820995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.820995Z digest=sha256:598ff40392dc80e55fc51e734b2bcd74c3bb4337cdfc88cffa6b0b74373c56ff

Observation f2444aa6-e5ce-4254-bd50-e16ffdd3972d · outbound

This paper cites Definitions of academic miscon- duct.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Definitions of academic miscon- duct

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.669906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.824075Z digest=sha256:3ceb7a57675763859a5561431adf719dd0e4e523c8705f819b8ab998d76367a5

Observation 903a1fa3-1d5f-44e5-87bb-826851e03fc3 · outbound

This paper cites Learning from failure: Integrating negative examples when fine-tuning large language models as agents, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Learning from failure: Integrating negative examples when fine-tuning large language models as agents, 2024

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.827016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.827016Z digest=sha256:f2c36253c7833284b1dc568bd9b2997f54d4a6d942457e36fe75efeedd43f18e

Observation f88f1040-69cf-4a09-9085-f496493a26ca · outbound

This paper cites Taxonomy of risks posed by language models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Taxonomy of risks posed by language models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.655644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.829688Z digest=sha256:e7848952517fb6905c4aad8658c8b091d3ce6dacc8077c6a2939b1cae6ab4470

Observation 6622c4fa-41f8-4106-a16a-9f8737f4da04 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.629857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.836245Z digest=sha256:70192fd6a1586c9db358b18ac649e56a07dcb0d70da42905b55105afe4cbd05e

Observation ea3f80e2-1f5d-47e5-bb37-c31f05405e55 · outbound

This paper cites Sorry-bench: Systematically evaluating large language model safety refusal behaviors, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Sorry-bench: Systematically evaluating large language model safety refusal behaviors, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.571732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.838631Z digest=sha256:ebb3fb06f9e729cf310472727687082f08a0abececc71d451fd2f839c38b0217

Observation f2bf9242-4514-4d3c-b87d-4b4654aa72e2 · outbound

This paper cites Online safety analysis for llms: a benchmark, an assessment, and a path forward, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Online safety analysis for llms: a benchmark, an assessment, and a path forward, 2024

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.532103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.841731Z digest=sha256:e08bda1e73f5060d553523c5af2660fde7d11956598c2e36c99c24147239b92c

Observation 70693349-8fb7-4f65-a289-70c29520a72d · outbound

This paper cites Gptfuzzer: Red teaming large language models with auto-generated jailbreak prompts, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gptfuzzer: Red teaming large language models with auto-generated jailbreak prompts, 2024

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.475401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.844756Z digest=sha256:dcda7a5315d5cd02f65acc114f3219bc107979583424046cea4d8a8e30947d58

Observation 1fed986f-f0ce-405e-8e2a-01d1cdb2c521 · outbound

This paper cites Assessing prompt injection risks in 200+ custom gpts, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Assessing prompt injection risks in 200+ custom gpts, 2024

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.459166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.847843Z digest=sha256:09233c354606f7e572c6aa781cc94f4bce7d58bc570a8d4744fd9b1d51876a51

Observation 74d5d262-79e3-4316-9f81-9bcd7c1c4e59 · outbound

This paper cites Don’t listen to me: Understanding and exploring jailbreak prompts of large language models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Don’t listen to me: Understanding and exploring jailbreak prompts of large language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.450746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.851059Z digest=sha256:2c02f155e3bd1dd1d06c768ce91af080a97c9d6d9aa23c0c1483e623164624a0

Observation 9ae6ee4a-25e3-4f19-a649-5e7562288a64 · outbound

This paper cites S-eval: Automatic and adaptive test generation for benchmarking safety evaluation of large language models, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs S-eval: Automatic and adaptive test generation for benchmarking safety evaluation of large language models, 2024

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.441334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.854144Z digest=sha256:4eeac6ae1327de259fb44269a20d555284f3e2e00056286f4400af4702ba9072

Observation 1ea83626-cb04-48aa-a480-7260b2e984c7 · outbound

This paper cites Defending against neural fake news.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Defending against neural fake news

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.432710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.857217Z digest=sha256:8d143542b0e52c19eb4d05ae38b920925a58142e3f3a484368b4e41ec8825de2

Observation ae2878a1-586c-4726-b0b2-e97773dde8b4 · outbound

This paper cites A first look at gpt apps: Landscape and vulnerability, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A first look at gpt apps: Landscape and vulnerability, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.423728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.861267Z digest=sha256:77579f37f221fa26cc1c18d35da0523b0734de8c764488f80e9108a03e3920d7

Observation b418e139-c5d8-4ac3-9d91-5e6a280b0bb2 · outbound

This paper cites Safetybench: Evaluating the safety of large language models, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safetybench: Evaluating the safety of large language models, 2024

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.415091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.863976Z digest=sha256:34457a71505685c8815256bb7b6c7318f33d28debf46b4ef65d5871b0a77306c

Observation cad04cee-4903-4d93-9d9d-fc21efce5e0b · outbound

This paper cites Gonzalez, and Ion Stoica.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gonzalez, and Ion Stoica

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.405755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.867298Z digest=sha256:5f3996b90de720579ec7fd59a55d897a56a793f289f6f164d443586c3e4a13e8

Observation 8dd9fb83-606b-4073-8f8f-23e85d190ee6 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.396366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.871158Z digest=sha256:992a2fa077279b570d4e2c3690f031f48468629df21266d013f3c8ce83af8910

Observation 3e7fa3cb-27a0-4d77-95ce-8059de71bb33 · outbound

This paper cites Sadeh, Steven M.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Sadeh, Steven M

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.386264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.907645Z digest=sha256:ac3d5b5d99a70bb7d4b7317c98265570b31b944b62f8d9d32ac6dc7de0dad5d7

Observation 4178d1f1-f767-48a0-85eb-09523cb82d00 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.368102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:33.047162Z digest=sha256:5104be43335d70ed44dcc56b705ba0d1bcf7d8bb34a7d95bb0923f03386b2e48

Observation 21458664-362c-46a1-9fa6-49e4d973ec5c · outbound

This paper cites boyfriends,.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs boyfriends,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.335204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:33.166652Z digest=sha256:5a6a3fbcf726c76d058ae34b5b601e6d6e4b9eb588486955b98678f937d4e0c7

Observation 193da6e3-9ac3-47c5-83cc-a17fe592b1d1 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 2017

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.377671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T15:22:32.980639Z digest=sha256:101af49cded4f3f9b7f274df696dc96455e86a17f84daccf01f3c6ee4908fe19

Observation ae820769-76cd-4f49-b248-5ca52beabe61 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.832855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.832855Z digest=sha256:e71e4639c0f8519e2cf3cb7bd7a5b58da012c361b81f90f83b44425e6b560e62

Observation 88074c97-2878-4dd5-9eea-0e6e96702f2a · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 2023

Resolution
parse uncertain
no resolver link, observed 2026-08-09T15:22:32.530750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.530750Z digest=sha256:7486f08bdd5fdb1dfa3e169f12e0e7711a7e1e5b3c77a6dfdb52479fa94e45eb

Pith citing papers

Observation 921eaf41-b4b4-40e5-86da-0f8c587bf484 · inbound

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models cites this paper.

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-14T02:21:16.410064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:49:55.738870Z digest=sha256:4146353e68703ff8e48160c00df2dc99dde5abf77e4b0506a39874b3e4ef8111

Observation cc402f8c-443f-4399-bf1b-17e2c5a7b0bd · inbound

Beyond Single-Policy: Evaluating Composed Organization-Specific Policy Alignment in LLM Chatbots cites this paper.

Beyond Single-Policy: Evaluating Composed Organization-Specific Policy Alignment in LLM Chatbots Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-07-14T02:21:16.410064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T05:41:16.033862Z digest=sha256:07659b2d1ee07052f7289cf79a053f9be7f4055d9fcea8dc60d1ce25ef04b3ed