Pith. sign in

Paper Citation Record · LEDGER

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs

As of 11 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 2 inbound Pith citation observations for arXiv:2502.01436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01436 v3

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T15:22:33.166652Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T05:41:16.033862Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact0
  • verified fuzzy56
  • unresolved18
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 019206a3-e398-48c3-8d2a-51fafcc6dc3f · outbound

This paper cites Gpt in sheep’s clothing: The risk of customized gpts, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gpt in sheep’s clothing: The risk of customized gpts, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.088821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.120408Z digest=sha256:6cd6f8a7f99a5718de8c5529cc7d75f4597823de3bc33a2b0ef63bf968387926

Observation 8e3e88d3-3bd0-4a26-9847-994d1ad0006b · outbound

This paper cites Safety-tuned llamas: Lessons from improving the safety of large language models that follow instructions, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safety-tuned llamas: Lessons from improving the safety of large language models that follow instructions, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.149793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.149793Z digest=sha256:74e841eaec76144aabdd58673629aaa53c77f2bddfd7dc7efc5d20e713c6e894

Observation c8940e90-d379-4370-8a2d-5d2b0c6d8e49 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.189930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.189930Z digest=sha256:6212a0230d0a8aa9c750ddb7429593f854b3223da66984e3ff26a5d8e84bd78a

Observation f0821e67-8002-4e16-9d3b-4271f9d07b60 · outbound

This paper cites Personal data flows and privacy policy traceability in third-party llm apps in the gpt ecosystem.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Personal data flows and privacy policy traceability in third-party llm apps in the gpt ecosystem

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.070316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.233461Z digest=sha256:3379fc874c7076274d6e7f519b91a40ac545a80bc379a58c31ad5c39ad6087c1

Observation 128a0391-36be-457a-a81e-777611e808f1 · outbound

This paper cites From representational harms to quality-of-service harms: A case study on llama 2 safety safeguards, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs From representational harms to quality-of-service harms: A case study on llama 2 safety safeguards, 2024

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.061966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.255082Z digest=sha256:10bc9acfb5275af96d3e73fe8adbb6b8134af80c095ba2c16f99b57ee852cfa3

Observation e3d1aed1-d34f-452b-8650-61f0c4268baf · outbound

This paper cites Avoiding social engineering and phishing attacks, 2021.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Avoiding social engineering and phishing attacks, 2021

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.053425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.291674Z digest=sha256:95fa400474bfcbd28fe8b0d67fedbb5fcd1f6b1e6615d40c6d2a6f587d29238e

Observation efa79ddc-f1db-4c0b-b04b-2a0fa67bc754 · outbound

This paper cites A security risk taxonomy for prompt-based interaction with large language models.IEEE Access, 12:126176–126187, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A security risk taxonomy for prompt-based interaction with large language models.IEEE Access, 12:126176–126187, 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.045036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.313658Z digest=sha256:4634adb52c597945e6c9930045461e8e1293f583570e84c160d78832f703e6b8

Observation 75cf4072-82ea-414d-8abb-1431e869e3f4 · outbound

This paper cites Transferable adversarial distribution learning: Query-efficient adversarial attack against large language models.Computers & Security, 135:103482, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Transferable adversarial distribution learning: Query-efficient adversarial attack against large language models.Computers & Security, 135:103482, 2023

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.036294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.355045Z digest=sha256:8b85976a704589a84ff58792b94fa6b976e8edac263dae5b560767356a03b77c

Observation c3abdff4-fc02-42f0-ad35-89e0c8167593 · outbound

This paper cites European network for academic integrity.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs European network for academic integrity

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.027983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.393483Z digest=sha256:99e4aeb49fe6ddda2b269e0b4377f94d3003f29314d5ba58bae543071867c069

Observation 3b5ffd19-d5d1-42ed-acb1-3e3693a8b468 · outbound

This paper cites Lessons From an App Update at Replika AI: Identity Discontinuity in Human-AI Relationships.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Lessons From an App Update at Replika AI: Identity Discontinuity in Human-AI Relationships

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.415231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.415231Z digest=sha256:0d12a4a424118e131152acf26c6839e690d3c52dc98b28349630819b5c6c6e0d

Observation c517757a-4414-4006-927f-8246b2feb784 · outbound

This paper cites A brief survey on safety of large language models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A brief survey on safety of large language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.019991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.418570Z digest=sha256:74f7fcc7857ad8abeb673c20181b30ce6ebc8c72b79e8205d1adfd3934bf5b3a

Observation 79d098ab-4ccd-40de-93f9-fa2427add1cd · outbound

This paper cites A survey on llm-as-a-judge, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A survey on llm-as-a-judge, 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.011644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.422120Z digest=sha256:b13bf4596002a0d872de1f7bddead1b05c36bd201f015c5c1ea242020dfe9368

Observation 796c0bb4-531b-46fe-96f9-ace5c7cf1f53 · outbound

This paper cites Shin, and Karl Aberer.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Shin, and Karl Aberer

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:34.003005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.425212Z digest=sha256:6d38b7573eff69a20d04889adbe0943fa3baa046a9e9ce1634eafc7a7fa2135d

Observation e91660af-e30e-403d-ab26-33ac5ca7e2ad · outbound

This paper cites Safe lora: the silver lining of reducing safety risks when fine-tuning large language models, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safe lora: the silver lining of reducing safety risks when fine-tuning large language models, 2025

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.994172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.427528Z digest=sha256:52ce0269b79f2f7f7efd3e0978f397877d8ce3abfbdb95366527225e78a132b2

Observation 0aa76448-e68f-426f-93c1-85c60258fd16 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs LoRA: Low-Rank Adaptation of Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.430229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.430229Z digest=sha256:11aed1b7ec1a24bde77c6c5f7c3d03e5ebbe1e96f651e2efbcd87de80e9219aa

Observation ff451910-41f3-412e-9f6a-80116996bd68 · outbound

This paper cites Chatgpt sets record for fastest-growing user base - analyst note.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Chatgpt sets record for fastest-growing user base - analyst note

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.985307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.432814Z digest=sha256:33b6d62e8640a191c9ad8d8e3ebbabdd02317d987eecfd27c0c7df2567226f3d

Observation 21ac8166-2f30-4f1b-8248-c3820b29a4c3 · outbound

This paper cites An empirical study of llm-as-a-judge for llm evaluation: Fine-tuned judge model is not a general substitute for gpt-4, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs An empirical study of llm-as-a-judge for llm evaluation: Fine-tuned judge model is not a general substitute for gpt-4, 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.976162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.435219Z digest=sha256:496408b5a4ae5edc553dc6e907bf53f84292d62a800dbde2dcc63fbe5cd5dbe6

Observation d7cdd1e6-a34d-4bc5-97fe-8f18f3e7e334 · outbound

This paper cites Lisa: Lazy safety alignment for large language models against harmful fine-tuning attack, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Lisa: Lazy safety alignment for large language models against harmful fine-tuning attack, 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.966940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.437506Z digest=sha256:203b6dd754c71aec0f76d65ecb824547f805cefdbbffd283a1a4f69dce1fbf8a

Observation 6deed3c1-589e-402b-934f-1ffa8cc4fd4e · outbound

This paper cites International center for academic integrity, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs International center for academic integrity, 2025

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.957596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.439853Z digest=sha256:f58285d28a83ea0e6d95b0e06f0e57d748a83ee768d403fa1248d4fc8b974bf7

Observation 7d5e9a72-4357-46dd-aa09-3fd404d6c595 · outbound

This paper cites Social engineering scams, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Social engineering scams, 2025

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.949030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.442463Z digest=sha256:2fc38415bf8186ee289f709eccc38a83a33211542694dab16cb9dc16a2eb7f23

Observation edcc5c7d-8010-4b88-b130-cfc2943327fe · outbound

This paper cites Knowledge Sanitization of Large Language Models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Knowledge Sanitization of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.444763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.444763Z digest=sha256:648114a1538b963d46360f226f48111b583323c0990d93898d70cf2a26fcf635

Observation 5bc4e5a5-5835-4139-85f4-97b6575470f1 · outbound

This paper cites Unleashing offensive artificial intelligence: Automated attack technique code generation.Computers & Security, 147:104077, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unleashing offensive artificial intelligence: Automated attack technique code generation.Computers & Security, 147:104077, 2024

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.940245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.448021Z digest=sha256:253c9f313dd6d23fa058f7063ffee441522625d52ead4dce7e0e53cb1d60a4e5

Observation 8e27a0cf-fe33-488e-a78e-2b977708ff1a · outbound

This paper cites Springer Nature, Cham, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Springer Nature, Cham, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.931574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.450617Z digest=sha256:c31351e8ae3d5d6747c012ee0ed40329b8f2078e8395128244c40ff504c16a43

Observation 1a80b2dd-f164-44fa-9063-f1c696aa02f2 · outbound

This paper cites Fine-tuning, quantization, and llms: Navigating unintended outcomes, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Fine-tuning, quantization, and llms: Navigating unintended outcomes, 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.922797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.453216Z digest=sha256:d25c91c7149a7d0e37fd7ca3c47351d2c2c88dc26200c9f23dd3a30f586e4f3a

Observation 7067d695-8ca5-4056-b53e-2e1116a87a37 · outbound

This paper cites From generation to judgment: Opportunities and challenges of llm-as-a-judge, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs From generation to judgment: Opportunities and challenges of llm-as-a-judge, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.456473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.456473Z digest=sha256:7858fafc95cfabcef738c8ba66b66a4962a1d9b4353ad9e8d7f449765d49a3eb

Observation 3d8041f8-1dac-40b8-b1df-e5b28aadaac6 · outbound

This paper cites Safety Layers in Aligned Large Language Models: The Key to LLM Security.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safety Layers in Aligned Large Language Models: The Key to LLM Security

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.459148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.459148Z digest=sha256:8a8bdd3c0db2ca177d244b842f5e47fe2e375bf3172e3005d82572a1c80a30c5

Observation 17e77498-c930-49a3-aa8d-0280f046d738 · outbound

This paper cites A hitchhiker’s guide to jailbreaking chatgpt via prompt engineering.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A hitchhiker’s guide to jailbreaking chatgpt via prompt engineering

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.908975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.462378Z digest=sha256:96b0960730f70c858f3016799f00bec6617273fe453cd8772cefe2e020e67cb3

Observation c15a3b88-39fe-4493-a2fe-a84ad87a6159 · outbound

This paper cites Privacy perceptions of custom gpts by users and creators.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Privacy perceptions of custom gpts by users and creators

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.898997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.465169Z digest=sha256:ccf75a8826709cf4aaf27d53633c81286e6258ccd0717994bc87ea3cc72bcb50

Observation dfb84552-fd2c-46d5-8965-c4c1f68d79b8 · outbound

This paper cites The trauma floor: The secret lives of facebook moderators in america, 2019.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs The trauma floor: The secret lives of facebook moderators in america, 2019

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.889995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.468202Z digest=sha256:06d4878d3db176840464b235a092702f0f9bd6b997c6eec0e30cea081c989dbc

Observation 96fd7c20-2fab-4f8c-8317-a369e2051504 · outbound

This paper cites Academic integrity policy.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Academic integrity policy

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.881187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.471454Z digest=sha256:e748d22412dc37c4eb9e17a95eca2645b04a6c6418bae5f637979c8b456db15d

Observation 59934e44-e09e-46bf-a8d2-cea1ace3b58a · outbound

This paper cites Introducing gpts: Custom versions of chatgpt for specific purposes.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Introducing gpts: Custom versions of chatgpt for specific purposes

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.872017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.473936Z digest=sha256:75ba952149b384ab9ec89a5d73e3a006261d0c36d02af90d39f90e3b22ad8acc

Observation fb0695bd-ad5d-403b-b064-258d06df640c · outbound

This paper cites Openai red teaming network.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Openai red teaming network

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.863044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.491594Z digest=sha256:d65c848a4573551f6fb52a2bce5f75e3c953b7f1aa4434d64f5a0e77f17a8e02

Observation 4c2b21dc-97c9-4206-99b4-be88a6d863e8 · outbound

This paper cites Usage policies.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Usage policies

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.848870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.573950Z digest=sha256:8eab39b1be22b3784356ffe199380041e295a4b2693c98ad0eae2896e3ff3526

Observation 29475659-529b-43bf-b717-be1ed399902b · outbound

This paper cites Openai safety.https://openai.com/safety/, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Openai safety.https://openai.com/safety/, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.840005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.594578Z digest=sha256:41dab60220d92c8a417af357f25d0411bfb6e41ca7d54df3061cf6cdd08ec68a

Observation 35111379-daf7-41b5-bf94-bd653498ac32 · outbound

This paper cites LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.635523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.635523Z digest=sha256:f71df5c7ba6f07aec28fe055fa96370b5b7124d8703dd0689b1fa4c366956a64

Observation 2f31e543-4401-43be-a6c0-718aecaa9d5c · outbound

This paper cites Puppeteer documentation.https://pptr.dev/, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Puppeteer documentation.https://pptr.dev/, 2025

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.830737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.658304Z digest=sha256:7971762fb24361e94bef890775c6e6d765466f0d52c345192155d945e7920c43

Observation 471e7bbf-93b1-453b-8bb5-c89f19ba5bf3 · outbound

This paper cites Fine-tuning aligned language models compromises safety, even when users do not intend to!, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Fine-tuning aligned language models compromises safety, even when users do not intend to!, 2023

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.674137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.674137Z digest=sha256:ddcaf454ddf3cc15b6629c67fb00660a1d7d8fd3caec661f847615f2c800688e

Observation 98371aba-6aba-4d08-aa63-aa8d6c067bdd · outbound

This paper cites Improving language understanding by gen- erative pre-training.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Improving language understanding by gen- erative pre-training

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.816249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.713618Z digest=sha256:18a1d1f118c03e9f9f61f3375a02f325b21d2090b23a299aad16a600e4ed4d41

Observation 6b274457-b211-46de-8783-2f8b7612ce17 · outbound

This paper cites Language models are un- supervised multitask learners.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Language models are un- supervised multitask learners

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.807470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.756161Z digest=sha256:b56d1419c3350dc080a172ff3ff640bbf56f506f2610b80b7ffc3d26b9f05efb

Observation e44caf72-b81e-4df1-9f8d-730de30625f5 · outbound

This paper cites Safetyprompts: a systematic review of open datasets for evaluating and improving large language model safety, 2025.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safetyprompts: a systematic review of open datasets for evaluating and improving large language model safety, 2025

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.798685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.778313Z digest=sha256:f33f447073f459cd25370de15346a81ab9b9710a0af5a2609c875c33b3f0f7ce

Observation a42d0a49-0ef3-4b51-9697-fd1ed52b5c29 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.789711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.781392Z digest=sha256:a63adbe0e6b01b04293c44ba64f8dd794d342d8823a31e1f4d7b914065d103e5

Observation 03c98b62-6759-45ac-9759-b201b335595d · outbound

This paper cites Identifying the provision of choices in privacy policy text.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Identifying the provision of choices in privacy policy text

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.781589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.784352Z digest=sha256:7eb5849a5889f07285066b2bb1ab6d9a45f93947e6daeff2929de51db2637b92

Observation eae0ff55-cb52-405e-be6a-cf38a95f9a76 · outbound

This paper cites do anything now.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs do anything now

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.773120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.787981Z digest=sha256:ffd76b7c588b3fe76ea7d22c5aaf4b1fd889e25e32b462c0eaf87cb206c3a40f

Observation 993b6e8e-dafc-400d-87b1-53fa64133403 · outbound

This paper cites Can people experience romantic love for artificial intelligence? an empirical study of intelligent assistants.Information & Management, 59(2):103595, 2022.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Can people experience romantic love for artificial intelligence? an empirical study of intelligent assistants.Information & Management, 59(2):103595, 2022

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.764155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.790816Z digest=sha256:0dd7a3c48cb4794695b868de95315d40aacbf97b185561333dd448197df589f4

Observation a3796ab8-c6a7-41a8-8de5-c93cada3d45d · outbound

This paper cites Gpt store mining and analysis, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gpt store mining and analysis, 2024

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.754983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.793685Z digest=sha256:ee8542635f775adca2d6b540526d6452690e4e93d59a696f5aa01829c025958f

Observation 45517465-1b61-421f-97ad-2d73bfc9beb5 · outbound

This paper cites Safety assessment of chinese large language models, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safety assessment of chinese large language models, 2023

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.747004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.796557Z digest=sha256:fdb68914c06d34e7c0460377dc97111a7df5ea79386224c3eae4d572ce12a835

Observation b3876033-f39d-4c7e-8ae9-5d732a3fb23e · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.739176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.799371Z digest=sha256:c8a803e00ddfa7b424f6b5de9c14769b620efe5a1a65a783e2e57f5245a18288

Observation 347d0509-2d67-442b-8a25-a489982f549d · outbound

This paper cites Opening a pandora’s box: Things you should know in the era of custom gpts, 2023.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Opening a pandora’s box: Things you should know in the era of custom gpts, 2023

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.731819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.802293Z digest=sha256:b59988bfea3672c43e565e892a3bb3b4b246d91372604ac64492714cd4fed22d

Observation 590064e0-d5f3-49ba-ac50-7dc2a95e2c60 · outbound

This paper cites Glossary for academic integrity.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Glossary for academic integrity

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.722660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.805047Z digest=sha256:e062545959a3bf19e1c853b693f479feaf2843a2d3b424666f854a258a85dd73

Observation a00a6bcc-86a1-4a61-b8f3-6203474f4c87 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.713883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.807937Z digest=sha256:a76d9966c00b100a821d6a0ec16cdfe95c5018c8c702ebc9e9975520115cc080

Observation 39df5f71-a3d6-49a8-9302-a7db996f07a3 · outbound

This paper cites Academic misconduct.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Academic misconduct

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.704661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.811029Z digest=sha256:84b29289d2895c233a175d8a9ff01fd861547dcc9525f7f70e6462106d0bf6b6

Observation 852c1533-8ec3-4296-ae2f-e8203cfbc0e8 · outbound

This paper cites Definition of academic dis- honesty.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Definition of academic dis- honesty

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.694955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.814385Z digest=sha256:81cedddeabb7677d58c872a04cd8bfe90a8ddbcef8507993d3b32fdaf9d6b87a

Observation a7607b6e-583a-47bc-8416-f8becb42a33f · outbound

This paper cites Academic integrity.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Academic integrity

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.685474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.817707Z digest=sha256:76288c32190f280ae98a2efb94e5dce158c41dd12d7c0024c761f853206fd377

Observation ab08b9b0-3099-4ccf-9055-761cb3d8d5da · outbound

This paper cites Gomez, Łukasz Kaiser, and Illia Polosukhin.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gomez, Łukasz Kaiser, and Illia Polosukhin

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.820995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.820995Z digest=sha256:68a5825e18d2963c7d650b369c68d455795e4457513844444f8edac858b5c33a

Observation f2444aa6-e5ce-4254-bd50-e16ffdd3972d · outbound

This paper cites Definitions of academic miscon- duct.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Definitions of academic miscon- duct

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.669906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.824075Z digest=sha256:0931fc92aa01c4a303d0504dc8aef8bd388c83401c5044bfbe77908e4ef48b9d

Observation 903a1fa3-1d5f-44e5-87bb-826851e03fc3 · outbound

This paper cites Learning from failure: Integrating negative examples when fine-tuning large language models as agents, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Learning from failure: Integrating negative examples when fine-tuning large language models as agents, 2024

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.827016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.827016Z digest=sha256:caa4827547069c9b10efde7ea5d015b90e5265449df8ea24350cfcc34291049c

Observation f88f1040-69cf-4a09-9085-f496493a26ca · outbound

This paper cites Taxonomy of risks posed by language models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Taxonomy of risks posed by language models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.655644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.829688Z digest=sha256:a98717fc755bcb7c8284786f8de6cf869bc2d472e00bbca26b9064a80685c207

Observation 6622c4fa-41f8-4106-a16a-9f8737f4da04 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.629857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.836245Z digest=sha256:76e0b8fd094bb220399ab5a8ceb8ae227e55184cac1f64c874101bbeadde00e3

Observation ea3f80e2-1f5d-47e5-bb37-c31f05405e55 · outbound

This paper cites Sorry-bench: Systematically evaluating large language model safety refusal behaviors, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Sorry-bench: Systematically evaluating large language model safety refusal behaviors, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.571732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.838631Z digest=sha256:e35faf3c59c504d644063b6cc0e9b56ccd75eccb87d370abfc05dcc49c2e1763

Observation f2bf9242-4514-4d3c-b87d-4b4654aa72e2 · outbound

This paper cites Online safety analysis for llms: a benchmark, an assessment, and a path forward, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Online safety analysis for llms: a benchmark, an assessment, and a path forward, 2024

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.532103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.841731Z digest=sha256:2cd045896f4dca64dc263cf9d9aba6983e11266f0bf0bf7f93e8b083c702c17e

Observation 70693349-8fb7-4f65-a289-70c29520a72d · outbound

This paper cites Gptfuzzer: Red teaming large language models with auto-generated jailbreak prompts, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gptfuzzer: Red teaming large language models with auto-generated jailbreak prompts, 2024

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.475401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.844756Z digest=sha256:296e078388e5d3c184953e8971ad0ef354b6f09e5e494b000f68d1b9e71fb140

Observation 1fed986f-f0ce-405e-8e2a-01d1cdb2c521 · outbound

This paper cites Assessing prompt injection risks in 200+ custom gpts, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Assessing prompt injection risks in 200+ custom gpts, 2024

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.459166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.847843Z digest=sha256:41e32368455672fca90d96a950824ff8e4bc5307ff135521d15a7a7519744a73

Observation 74d5d262-79e3-4316-9f81-9bcd7c1c4e59 · outbound

This paper cites Don’t listen to me: Understanding and exploring jailbreak prompts of large language models.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Don’t listen to me: Understanding and exploring jailbreak prompts of large language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.450746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.851059Z digest=sha256:a8ec53f49c2ffb7b2b46c7bd94a9ac37a92206c482feff6a707194d0047a646b

Observation 9ae6ee4a-25e3-4f19-a649-5e7562288a64 · outbound

This paper cites S-eval: Automatic and adaptive test generation for benchmarking safety evaluation of large language models, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs S-eval: Automatic and adaptive test generation for benchmarking safety evaluation of large language models, 2024

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.441334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.854144Z digest=sha256:f7a83619852ada119cc22c7b890b224534d4acb8a58f1400156453ce486db01c

Observation 1ea83626-cb04-48aa-a480-7260b2e984c7 · outbound

This paper cites Defending against neural fake news.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Defending against neural fake news

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.432710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.857217Z digest=sha256:0ef1aa05d7fd8b40832cfea08b5dfd773568b2607f7fe11fe281eea71322ef56

Observation ae2878a1-586c-4726-b0b2-e97773dde8b4 · outbound

This paper cites A first look at gpt apps: Landscape and vulnerability, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs A first look at gpt apps: Landscape and vulnerability, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.423728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.861267Z digest=sha256:34e5cbe266861fcf16516e1bc5b33aa145ff5fba4284737a37ef07a779fc1ae5

Observation b418e139-c5d8-4ac3-9d91-5e6a280b0bb2 · outbound

This paper cites Safetybench: Evaluating the safety of large language models, 2024.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Safetybench: Evaluating the safety of large language models, 2024

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.415091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.863976Z digest=sha256:1abca9c0054537f6cdee0f9541b77dd5348ffaab59059624e8bb7f8f895e9ec0

Observation cad04cee-4903-4d93-9d9d-fc21efce5e0b · outbound

This paper cites Gonzalez, and Ion Stoica.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Gonzalez, and Ion Stoica

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.405755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.867298Z digest=sha256:f98a3158ac53d7e81ab2774a363022e3b6d2de9cb6acee320d33c7c72d7f08c8

Observation 8dd9fb83-606b-4073-8f8f-23e85d190ee6 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.396366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.871158Z digest=sha256:25b2eab591561495b3d771137800b77769907fec0753e8f99f87ee6a68260e0b

Observation 3e7fa3cb-27a0-4d77-95ce-8059de71bb33 · outbound

This paper cites Sadeh, Steven M.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Sadeh, Steven M

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.386264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.907645Z digest=sha256:cf4d4bc7e22aba795f421c0b501602923229dedc9eb47bc1dda72137d6c85c36

Observation 4178d1f1-f767-48a0-85eb-09523cb82d00 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.368102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:33.047162Z digest=sha256:fb3098ee4ec93291acc11fc465b181d37a436d73dd39468be54b446c7b996f26

Observation 21458664-362c-46a1-9fa6-49e4d973ec5c · outbound

This paper cites boyfriends,.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs boyfriends,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:22:33.335204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:33.166652Z digest=sha256:0b0c16d79fd614c50b21593edefc3300449d813106a0cde9143f5f0008152f0c

Observation 193da6e3-9ac3-47c5-83cc-a17fe592b1d1 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 2017

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:22:33.377671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T15:22:32.980639Z digest=sha256:83711a3ad27b9c2b4d3569263d109efc86227cde73385b42c503ce2edc5716d6

Observation ae820769-76cd-4f49-b248-5ca52beabe61 · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T15:22:32.832855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.832855Z digest=sha256:ac9b7193e6336d7c244b557d2f14edac111f98b24ad3769dc071257700aec71e

Observation 88074c97-2878-4dd5-9eea-0e6e96702f2a · outbound

This paper cites an unresolved cited work.

Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs Unresolved cited work

Reference 2023

Resolution
parse uncertain
no resolver link, observed 2026-08-09T15:22:32.530750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:22:32.530750Z digest=sha256:b57cc6707a7f469b77ec75f7cf9512232d44becaab744f44cea30837ac319b0a

Pith citing papers

Observation 921eaf41-b4b4-40e5-86da-0f8c587bf484 · inbound

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models cites this paper.

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-14T02:21:16.410064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T05:49:55.738870Z digest=sha256:3f4f90c826a6b9440ad9bf3f569d4b84a2877df3c9677a74f743d95fa41ae376

Observation cc402f8c-443f-4399-bf1b-17e2c5a7b0bd · inbound

Beyond Single-Policy: Evaluating Composed Organization-Specific Policy Alignment in LLM Chatbots cites this paper.

Beyond Single-Policy: Evaluating Composed Organization-Specific Policy Alignment in LLM Chatbots Towards Safer Chatbots: Automated Policy Compliance Evaluation of Custom GPTs

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-07-14T02:21:16.410064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T05:41:16.033862Z digest=sha256:9d2f668063612ccc1509dd628decf4d97f883e32876a14f31d813affe1608ac9