Pith. sign in

Paper Citation Record · LEDGER

AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals

As of 18 August 2026, this Paper Citation Record lists 8 of 8 outbound references and 1 inbound Pith citation observation for arXiv:2505.15365.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15365 v1

Coverage vector

measured 8 of 8 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:24:32.458219Z

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T06:16:01.603353Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T06:20:59.053668Z

Reference resolution

8 of 8 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f0cc52de-7922-4f59-864e-6a82112d1228 · outbound

This paper cites I’m sorry, I can’t assist with that request because it involves unsafe content.

AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals I’m sorry, I can’t assist with that request because it involves unsafe content

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:24:34.230524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:24:31.986182Z digest=sha256:7f0b7dfeca1afc01ce642f0f6b24c875e5f424826f4f6e9418181349de613ba0

Observation 20a45c6c-7b1d-4b61-b4c0-7d00704ce979 · outbound

This paper cites control gap,.

AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals control gap,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:24:34.006660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:24:32.006207Z digest=sha256:b4b5eca7ec12d9255183fc09623da4056c1ad5fefffc98b40fd2481868876759

Observation df75ed2f-1818-4cbd-acbf-05a9de99cb94 · outbound

This paper cites I cannot help with that request as it may cause harm.

AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals I cannot help with that request as it may cause harm

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:24:33.739590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:24:32.048554Z digest=sha256:2711f2f87f03fb3b3104f83559ba0193494a94eca21a1f809039f1a3e6f90b57

Observation eb18988e-0f21-40ee-874e-2a2f9744bb91 · outbound

This paper cites Ethical 1235 1.2 0.38 0.34 0.28 1.12 0.46 0.46 0.08 1.01 0.49 0.49 0.02 1.00 Discl.

AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals Ethical 1235 1.2 0.38 0.34 0.28 1.12 0.46 0.46 0.08 1.01 0.49 0.49 0.02 1.00 Discl

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:24:33.307192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:24:32.200734Z digest=sha256:0592a3b1398ab8653c99bea9605081b9c74c29644c2ba6a41294cbe46f405228

Observation 0a201179-f7bf-4ee2-8882-38ee5fc6d22c · outbound

This paper cites an unresolved cited work.

AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:24:33.532250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:24:32.137296Z digest=sha256:a203b90bdcac660fb2a96296904453a950d313bd86b87b222c720d4c68e2174d

Observation 2ff4db94-9d60-40ff-8a84-f4ba87901eb6 · outbound

This paper cites safe” or “responsible.

AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals safe” or “responsible

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:24:33.011421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T15:24:32.272945Z digest=sha256:65d7b4ea0749f593b68bba31329a0a9ad4c2c398339845390ae0a8701d452886

Observation 3a05b733-71fc-4e3b-9ec9-263ea3257ddb · outbound

This paper cites an unresolved cited work.

AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:24:32.335157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:24:32.335157Z digest=sha256:f89bd8edc228b1cfc5a38ca10f5803cd8886a4390d247ad1ab90296576cfb8d1

Observation 701ad2f1-fb4f-4a3e-a3d0-0bbfb0027a9c · outbound

This paper cites Learning to Plan & Reason for Evaluation with Thinking-LLM-as-a-Judge.

AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals Learning to Plan & Reason for Evaluation with Thinking-LLM-as-a-Judge

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:24:32.458219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:24:32.458219Z digest=sha256:ec3b2fd20ca14b1169217120a25ba1d97e864c21a51d15efefee736a8daee747

Pith citing papers

Observation a4277555-1940-4204-90fe-015317f87bf5 · inbound

LLMs Judge Themselves: A Game-Theoretic Framework for Human-Aligned Evaluation cites this paper.

LLMs Judge Themselves: A Game-Theoretic Framework for Human-Aligned Evaluation AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-18T06:20:59.056117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T06:16:01.603353Z digest=sha256:651f2bc72ad56b40b79340bee6d527ad5e4877fd035ecd8597203e4a99c5406c