Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 41 inbound Pith citation observations for arXiv:2203.09509.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:52:27.948003Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
7
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 17fef00c-fb8a-49b0-8274-5f2caa630e97 · inbound
AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2b363b2-bcd9-4ec6-a74c-6b2cf0eba9b0 · inbound
Textbooks Are All You Need II: phi-1.5 technical report ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f22a3a6a-0f3e-4dc2-87c4-1671e09480b5 · inbound
Baichuan 2: Open Large-scale Language Models ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e29e0f7c-afdc-4ee8-9acd-78a2ec208bc1 · inbound
SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 768a96c8-2dc6-4f21-9cfd-faa16ccfa0ab · inbound
TrustLLM: Trustworthiness in Large Language Models ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 249
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 876e2dd4-3c44-4b80-b1e8-056742970f2a · inbound
Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 556dabd3-900e-4d59-a3f7-865468e7a661 · inbound
SweEval: Do LLMs Really Swear? A Safety Benchmark for Testing Limits for Enterprise Use ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dac24db-23b1-45a1-8b56-0a5a09ed3a42 · inbound
Lifelong Safety Alignment for Language Models ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21b1ad6c-6f7c-4cce-8b6a-29326355d961 · inbound
MCP Safety Training: Learning to Refuse Falsely Benign MCP Exploits using Improved Preference Alignment ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d05b928-c056-401c-bd28-2c7c2c051e3f · inbound
Large Language Models Often Know When They Are Being Evaluated ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19dd54ad-3ee4-4bb6-9dc0-5811637505d4 · inbound
AutoMixAlign: Adaptive Data Mixing for Multi-Task Preference Optimization in LLMs ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0dd774a-ea55-4e89-b2aa-b2286cbd45e5 · inbound
Surfer-H Meets Holo1: Cost-Efficient Web Agent Powered by Open Weights ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efb100cd-e091-4f8a-8418-59cf60dcab4f · inbound
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75904f59-fdc4-4187-9fda-2f3695cacbb4 · inbound
PL-Guard: Benchmarking Language Model Safety for Polish ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3249b7db-6ed0-435a-b294-600729d39252 · inbound
Fine-Grained Chinese Hate Speech Understanding: Span-Level Resources, Coded Term Lexicon, and Enhanced Detection Frameworks ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7216b0d2-e0d8-46c8-8a99-25df91ba84fe · inbound
Trade-offs in Image Generation: How Do Different Dimensions Interact? ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 113faed5-d820-4038-99c4-51063e4767e6 · inbound
Model Misalignment and Language Change: Traces of AI-Associated Language in Unscripted Spoken English ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1449341-6504-41bc-823a-8141d70d6479 · inbound
Decoding the Rule Book: Extracting Hidden Moderation Criteria from Reddit Communities ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ceb7cea7-f718-4b45-ae01-9dcfd5476e8a · inbound
YouthSafe: A Youth-Centric Safety Benchmark and Safeguard Model for Large Language Models ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04e703e6-f0d3-4b54-8fa1-f57a29d31de2 · inbound
Red-Bandit: Test-Time Adaptation for LLM Red-Teaming via Bandit-Guided LoRA Experts ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd1ee9e9-da8b-456a-bac8-6d22c45ce839 · inbound
AgentCrypt: Advancing Privacy and (Secure) Computation in AI Agent Collaboration ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb976230-0f8b-46af-ad25-8650ec1315fe · inbound
f-GRPO and Beyond: Divergence-Based Reinforcement Learning Algorithms for General LLM Alignment ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e162268f-bb73-4521-9222-f28e2bf6ece1 · inbound
Response-Based Knowledge Distillation for Multilingual Jailbreak Prevention Unwittingly Compromises Safety ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc6a5e7f-5ffb-4f45-935d-e121dc5fe7a7 · inbound
Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 199
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a5b8ed8-f3ab-4d06-a52d-f64160446278 · inbound
Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 192
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 193d677d-440b-46f3-b966-68430a4cfd68 · inbound
DRAFT: Task Decoupled Latent Reasoning for Agent Safety ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 920e93de-69f6-4240-9fd9-4133ed68bca4 · inbound
FlowGuard: Towards Lightweight In-Generation Safety Detection for Diffusion Models via Linear Latent Decoding ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b393628b-39dd-4540-9d8a-be8146657260 · inbound
MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9d1fd19c-21b3-4f7a-9e20-3a9533a509d7 · inbound
The Safety-Aware Denoiser for Text Diffusion Models ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ffdb036a-aed4-4429-a0d0-2856ee33e3f0 · inbound
Navigating the Sea of LLM Evaluation: Investigating Bias in Toxicity Benchmarks ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 67e2e034-cf4d-4998-a5e8-b807e5ca3df2 · inbound
Leveraging RAG for Training-Free Alignment of LLMs ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 30df317c-2213-4672-b830-bd304470482c · inbound
Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 252
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd82ca98-348a-40c4-8b0c-1437ba109620 · inbound
Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 60189dfc-fbbb-4a19-86cc-af48bc202230 · inbound
Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 50ad264b-4bc3-4fa5-a839-44c7d2b06bd7 · inbound
AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fcf05e21-ffa6-424c-aae3-4520b154222d · inbound
SafetyRepro: Configuration-Conditional Rank Instability on Alignment Benchmarks ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c7e67617-1ed6-4b1e-b2bd-3856f632c606 · inbound
UniSteer: Text-Guided Flow Matching in Activation Space for Versatile LLM Steering ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 055de9de-74fe-4e20-add7-badff5aa2d27 · inbound
Symmetric Divergence and Normalized Similarity: A Unified Topological Framework for Representation Analysis ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a400adf-39b8-4ab4-bddf-9a60ecd4fa53 · inbound
Distilling Safe LLM Systems via Soft Prompts for On Device Settings ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b1c598da-52c3-485c-99dc-b42c5ac3be2d · inbound
S2T-RLHF: Hierarchical Credit Assignment for Stable Preference-Based RLHF ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b99c374-9243-42d7-b248-ab1ad9176d5a · inbound
Fence: Specialized SLM Guardrails for LLM Applications ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.