Pith. sign in

Paper Citation Record · LEDGER

TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2306.11507.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.11507 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T20:36:01.805843Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:50:11.115863Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eba4c398-411b-4bdb-b8d4-7a77ac51a6dd · inbound

Navigating the Risks: A Survey of Security, Privacy, and Ethics Threats in LLM-Based Agents cites this paper.

Navigating the Risks: A Survey of Security, Privacy, and Ethics Threats in LLM-Based Agents TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 110

Resolution
unresolved
no resolver link, observed 2026-08-12T20:36:01.805843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:36:01.805843Z digest=sha256:bfb1a4569c0e36865af50df78836a05a1dbd9306a6324e55bcb5b473248f508d

Observation 59f1e04a-4e79-4a20-a729-eae710f09163 · inbound

Moral Persuasion in Large Language Models: Evaluating Susceptibility and Ethical Alignment cites this paper.

Moral Persuasion in Large Language Models: Evaluating Susceptibility and Ethical Alignment TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T18:15:02.605649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:15:02.605649Z digest=sha256:3e655ab438909c5669b6cb37e4bc539d52aad3e39b1f3d8356efb4efb7bc2da3

Observation 250faf4d-a769-4771-a602-a0044531f234 · inbound

Observing Micromotives and Macrobehavior of Large Language Models cites this paper.

Observing Micromotives and Macrobehavior of Large Language Models TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T18:24:05.924949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:24:05.924949Z digest=sha256:7bf69785bd1545667032654e5ff2bd0c510451a731eec2c11d035a4bec88226e

Observation d2119278-9c00-4249-a600-99afeee24c86 · inbound

LangFair: A Python Package for Assessing Bias and Fairness in Large Language Model Use Cases cites this paper.

LangFair: A Python Package for Assessing Bias and Fairness in Large Language Model Use Cases TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:29.752606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:56:29.752606Z digest=sha256:86b705a7a305ff4c4a5fbe423cfb91b34b17f660d339a3a14d6c2abeef68c919

Observation a93fecbc-57ab-4676-a326-5f88daccd6ea · inbound

Value Compass Benchmarks: A Platform for Fundamental and Validated Evaluation of LLMs Values cites this paper.

Value Compass Benchmarks: A Platform for Fundamental and Validated Evaluation of LLMs Values TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T20:53:59.701166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:53:59.701166Z digest=sha256:35bd7f95ad82bfff8b818404ac69074e36480c7cd2d4aeb94f87f0868ba41a34

Observation a5beb4ef-85cb-4626-be2f-77ddb7b527a6 · inbound

Fairness through Difference Awareness: Measuring Desired Group Discrimination in LLMs cites this paper.

Fairness through Difference Awareness: Measuring Desired Group Discrimination in LLMs TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T14:03:44.434667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:03:44.434667Z digest=sha256:a1fa8d02952618f8cf0222668094b4ade3f478e6db3fb683e0d0265b0dfe9b97

Observation cd7d5d5c-ea78-42e4-b482-e3af89ecce4c · inbound

Artificial Intelligence in Spectroscopy: Advancing Chemistry from Prediction to Generation and Beyond cites this paper.

Artificial Intelligence in Spectroscopy: Advancing Chemistry from Prediction to Generation and Beyond TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T20:11:11.994909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:11:11.994909Z digest=sha256:2dd5a03512d5c87e22734f0632114a9651d6e0870ce380c85c419eb5b5db4437

Observation 162d771a-418c-4b4e-a320-87e3c17d1a75 · inbound

When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas cites this paper.

When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:22:55.580256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:22:55.580256Z digest=sha256:d8462c137e4d590ac334d830cd310e96211bab6f2e5804a81570a33bcb5cb3c1

Observation 69224819-50eb-4cda-833f-5ad01ab6c371 · inbound

OpenFActScore: Open-Source Atomic Evaluation of Factuality in Text Generation cites this paper.

OpenFActScore: Open-Source Atomic Evaluation of Factuality in Text Generation TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:20:16.536212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:20:16.536212Z digest=sha256:b1c6d4dc7efd905b33060ce75e7242563191c5aa4babb2eb20e47c5ea1aa3208

Observation 8b639068-14e2-43aa-bcb9-7271a1414879 · inbound

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM cites this paper.

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T23:13:00.005776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:13:00.005776Z digest=sha256:178d399b08567767f07accfd22d9898856a44f3b4b977dbec2c6757ab10b171d

Observation bb7343bc-d8fc-43bc-8867-3e4323e6d0e2 · inbound

EPT Benchmark: Evaluation of Persian Trustworthiness in Large Language Models cites this paper.

EPT Benchmark: Evaluation of Persian Trustworthiness in Large Language Models TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T23:03:16.421385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:03:16.421385Z digest=sha256:0e81c9fac16ad1355a20fff7cfc69c3c9436c1c4ab7ae1aa6dc493280d815a48

Observation 29799214-f1d3-4d1d-9649-d2b5ecd02cb2 · inbound

Safety Is Not Universal: The Selective Safety Trap in LLM Alignment cites this paper.

Safety Is Not Universal: The Selective Safety Trap in LLM Alignment TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:08:04.813096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T16:03:21.391883Z digest=sha256:8653b8e0206bd7a0a8b8779f78f433e29d62c57eeafe4070f801609e40a19bdb

Observation f9420c9b-c3fc-4a4f-a9f3-13835ab1c1fa · inbound

Safety Is Not Universal: The Selective Safety Trap in LLM Alignment cites this paper.

Safety Is Not Universal: The Selective Safety Trap in LLM Alignment TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T12:08:44.317070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:08:44.317070Z digest=sha256:36d34c5fc32fb55a9d53f2afab623fb6cf6d06bd041f2b34e91438c15d59b051

Observation 98745a4a-0145-48e9-b9ee-89b9cb6c67e5 · inbound

From RAG to Agentic RAG for Faithful Islamic Question Answering cites this paper.

From RAG to Agentic RAG for Faithful Islamic Question Answering TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T11:08:23.419771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:08:23.419771Z digest=sha256:ff1ea84bd3067aaa04110120cc28d85fd086e5bc76e8950b1ce4456f47bc1b6f

Observation 379c3a04-0776-476f-aa29-4ebb71d48d29 · inbound

Insight: Enhancing Mobile Accessibility for Blind and Visually Impaired Users with LLMs cites this paper.

Insight: Enhancing Mobile Accessibility for Blind and Visually Impaired Users with LLMs TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:15.302288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-12T02:11:09.787984Z digest=sha256:8593b1a1509febc48d9cfecec24b4678072a3ebb912433420451624e4b3e68a4

Observation 051922d9-5140-4680-97b0-23317a76243a · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.418348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-12T04:50:17.399580Z digest=sha256:fe414186800ba3322420f001965bd4d1084ba65cfa3e398db0bbf2c185543370

Observation 79ca632e-1935-4bf4-88a9-a6ef4e3eb072 · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:22:29.071324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T07:20:32.494840Z digest=sha256:e3b9e7fda6db3405577a91c2e7c92f0a4c59dfb2f9157f8864c18676370a16fd

Observation 1be6c546-25c7-4a0b-8668-7cfc56a920eb · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:50:11.118037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-25T20:58:53.119386Z digest=sha256:fdafaedaeba416291fd8577c685166905c37b6997276f83a7044139ebc342a71

Observation efcd11f5-2eff-45e5-985a-5ec9819e94c7 · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T10:16:36.314418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:16:36.314418Z digest=sha256:ce6cb5f6c0a995fa36e1d7d26a0d547160c506d2ac99b0c59e4bd66fa24f4453

Observation 726877a8-a677-4263-843b-061db4ead206 · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:28.889774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:28.889774Z digest=sha256:5a43bd0d6ee046d76e9e83eed1ba65034ee7ac2d407fdcb4dbda1ef1c211206c