Pith. sign in

Paper Citation Record · LEDGER

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace

As of 18 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 1 inbound Pith citation observation for arXiv:2505.17078.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17078 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:34.927348Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-25T21:16:36.392606Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:30:07.677913Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved44
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1d306a81-6087-4078-b15e-c9bb30816cd2 · outbound

This paper cites Language Models are Few-Shot Learners.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Language Models are Few-Shot Learners

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:29.664936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:29.664936Z digest=sha256:b487fa2f551d2e622b5f2ef2d5968141b48e36bfab6e4626f49b5c7e167d4448

Observation f7ea174d-9138-4304-bff3-45199293e2e5 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:29.746753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:29.746753Z digest=sha256:7821f3cf92c6481b45a45afb5c22911406adba709352c0875b0934f4f058cf8e

Observation 0030270b-98e4-4c25-ba40-54e871ad0c6b · outbound

This paper cites Everything is Editable: Extend Knowledge Editing to Unstructured Data in Large Language Models.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Everything is Editable: Extend Knowledge Editing to Unstructured Data in Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:29.873282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:29.873282Z digest=sha256:4d7fc9bc8cfff7f914f5b8b99587b573b6afebca4a1da439ec2dbb4098467fee

Observation 8bc8dc52-34b8-4476-9817-a1d9c6861d66 · outbound

This paper cites Related Knowledge Perturbation Matters: Rethinking Multiple Pieces of Knowledge Editing in Same-Subject.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Related Knowledge Perturbation Matters: Rethinking Multiple Pieces of Knowledge Editing in Same-Subject

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:42:35.937389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T15:42:29.987832Z digest=sha256:be525e8cca61160de87466d6e80f05015812a82e907090cc988507d6f7087595

Observation 329266cd-b7eb-4096-baf3-9c10cca7b507 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:30.146022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:30.146022Z digest=sha256:5a8a590e6f99ed3a0568d307e298beb0ea9a020d02b63e8ae59c258c43cfbd70

Observation 39c5fca4-d2dc-4489-847d-39654dee5017 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:30.234297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:30.234297Z digest=sha256:5d649cb7fdee975e20d4227e0ea5b9c1d9182f64f876501123fb2e6313f18275

Observation bf6c5531-3793-4654-8bf7-6ff5ccaa77cc · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:30.404467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:30.404467Z digest=sha256:e51dfd5f12ac01a33a091327fed19d02770113a8447e31ce1cbd3925e4676c81

Observation 641ea04d-3d6a-4670-92a9-b9e6ef1cb70b · outbound

This paper cites Transformer Feed-Forward Layers Are Key-Value Memories.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Transformer Feed-Forward Layers Are Key-Value Memories

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:30.575351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:30.575351Z digest=sha256:a760c195760c2ca7cf777f50a407058ee94df52d5064a7ea29e88ac7d9392b25

Observation d18b0f2a-41c8-422e-8a51-3e892a8222cc · outbound

This paper cites A Survey on LLM-as-a-Judge.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace A Survey on LLM-as-a-Judge

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:30.761330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:30.761330Z digest=sha256:030a04f1c3f113eade240d4d6a60906d70baea06807e2674a37152ed1a1500c4

Observation 9c46c601-880b-470b-8d8a-55e33dd258a3 · outbound

This paper cites Detoxifying Text with MaRCo: Controllable Revision with Experts and Anti-Experts.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Detoxifying Text with MaRCo: Controllable Revision with Experts and Anti-Experts

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:30.991971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:30.991971Z digest=sha256:37ab73fb9534c1f6d715630014554d5d09bcfa4c20870414bfc5328056391a31

Observation ad8fa427-1c89-40f0-a84c-0d01ea855c41 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace LoRA: Low-Rank Adaptation of Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:31.079246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:31.079246Z digest=sha256:b6614c4bdc274a102b17860395761a84470f33743c297f8f1aeb6ffadf85453c

Observation c9e0a81d-3400-4da3-9421-f438f68c5b97 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:31.187270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:31.187270Z digest=sha256:9ed43d82125c4095359ace0b1e992730460d3ca5e2fe79cb065c9f9bc44bc480

Observation a47f9934-f65f-439d-95bb-433dd09c4b1e · outbound

This paper cites CTRL: A Conditional Transformer Language Model for Controllable Generation.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace CTRL: A Conditional Transformer Language Model for Controllable Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:31.288935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:31.288935Z digest=sha256:701e56a287b6dade0068ac08d99862ea22b45ec4cbcfc07fbcfd2dcafbdf8925

Observation 51668387-e90c-4c23-9f51-2aa63e824fc6 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:31.478253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:31.478253Z digest=sha256:f231cdf887813d0a2d66fea78f22ee36977cfea2006d1a2687be06ff14dffe2e

Observation be266ba5-146b-4df2-bf88-7669b297273f · outbound

This paper cites Kummerfeld, and Rada Mihalcea.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Kummerfeld, and Rada Mihalcea

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:31.629100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:31.629100Z digest=sha256:ce3a3542b11212128e478954c82d84b71ed71ff7c4a5482c5ed061b503a57d2b

Observation 5458b9a1-d04a-4b6c-8d6f-a3446b74548a · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:31.866816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:31.866816Z digest=sha256:9a56517eb56ea5e5aaa4a7a0f5c85322cfac9693b47160252c5ecfa3b3592c23

Observation 923982c6-6001-4ee4-9417-33db731489ac · outbound

This paper cites Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:32.063185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:32.063185Z digest=sha256:9948feac71823d805d45cb51ad5fbe3fb683103e6a6cc19d530fd4ab54b49ba1

Observation 54f4af7d-478a-4730-a630-eaa271f23cca · outbound

This paper cites How Does DPO Reduce Toxicity? A Mechanistic Neuron-Level Analysis.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace How Does DPO Reduce Toxicity? A Mechanistic Neuron-Level Analysis

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:32.247738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:32.247738Z digest=sha256:303a61df617b184df4f7d738bfc77d5f0967d63bded99f6536ac1c74b32e4319

Observation f98c7e8c-78dd-4f6e-a4ea-985c269fd065 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:32.402656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:32.402656Z digest=sha256:34df939748d52974c42bf6bff9413e56d6edefdabc933dfeda946dd162f09b20

Observation 4a3cf486-9d7b-4c94-a396-953cb1cb176d · outbound

This paper cites Pointer Sentinel Mixture Models.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Pointer Sentinel Mixture Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:32.513753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:32.513753Z digest=sha256:d5ec5e572fd43e2c3e5f96901ef47c3f3e49ae7e5809f96ce466d8f3083e2665

Observation 1ae10703-9ec0-4440-9443-add8c5e5a0a6 · outbound

This paper cites How Do LLMs Acquire New Knowledge? A Knowledge Circuits Perspective on Continual Pre-Training.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace How Do LLMs Acquire New Knowledge? A Knowledge Circuits Perspective on Continual Pre-Training

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:32.659490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:32.659490Z digest=sha256:d6630a098b369179dc156d7cc7eab3be73b1f134eb832902d3583cdcaa76f2eb

Observation b2873002-6487-4545-88ad-cfda3ac77a59 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:32.737455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:32.737455Z digest=sha256:38f5b058207dade3669f151d27d9f675d35d25ff609f85b3ab9a8172b8b69be3

Observation 63ed797c-92c4-44b6-b78d-d3a1b93a6163 · outbound

This paper cites The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:32.812422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:32.812422Z digest=sha256:c116fce65fc1020a752d833706d9d08244718379e7e2737b10f29af06d9eb15e

Observation c4775586-43ab-4b75-9ed7-cd07cb1fb1b1 · outbound

This paper cites Back to the Future: Unsupervised Backprop-based Decoding for Counterfactual and Abductive Commonsense Reasoning.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Back to the Future: Unsupervised Backprop-based Decoding for Counterfactual and Abductive Commonsense Reasoning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:42:35.595656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T15:42:32.887910Z digest=sha256:a6c8a428be6cedc340f47509b5039dfdd712529cf907d4b3b39befc8a9d6938d

Observation f3e95b49-3a12-42ec-8797-3fac0f195282 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:32.984503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:32.984503Z digest=sha256:cc2364772f796da4299b74e8c15e361cf5ace9d3ef6e4b160b95e1e07498730b

Observation 1c257683-16cb-48fb-a62b-a6baca8c13bb · outbound

This paper cites Manning, and Chelsea Finn.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Manning, and Chelsea Finn

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.052940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.052940Z digest=sha256:a7d448181e3ead781db3fffd3c68144fd9a1845873c6713d663bf7b34171f90a

Observation 3747225d-75ba-47bf-a7c0-81c2fbf4c196 · outbound

This paper cites Whispering Experts: Neural Interventions for Toxicity Mitigation in Language Models.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Whispering Experts: Neural Interventions for Toxicity Mitigation in Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.134548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.134548Z digest=sha256:e7dcc61f3940785f444a71426da94ae8451e9b0380171214e2350fb2a8c53a20

Observation aafb8c31-4fb1-4b82-9d73-b1ad8801d2ee · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.226307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.226307Z digest=sha256:b2f4c892247e8627ea4292225d51cccff38af29d0c2dda7da7117420477fad2f

Observation d7a229c4-804a-43cd-ac26-2a4f81e96d8d · outbound

This paper cites Function Vectors in Large Language Models.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Function Vectors in Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.313206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.313206Z digest=sha256:f7d85619f727dc0f62325ffc2474f2389d5c16320e751c86ef65398e9ed5b319

Observation a342c3a4-e89b-4bca-86f7-afb188ca2318 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:36.407052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T15:42:33.391718Z digest=sha256:70d707037db58334defdf49e486760a818880fd05e3922d7f57f9ab64f163597

Observation bee9739f-8e95-4919-842f-7b17a6f03b7d · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.486225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.486225Z digest=sha256:3a74bf6221ac6eb3c7cac1bd6286495ae7c06123b9691ceb829de28306aff365

Observation 572592ee-afe2-490c-9246-afbd008ec68e · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.593951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.593951Z digest=sha256:a85d94e62afec5edf6a0abbe1850aaa2bbb224902bb30723b63c08a9ed47e858

Observation caf42a89-8f96-4b7a-b82b-b4f12b28c000 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.684125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.684125Z digest=sha256:cadc7c58eb8ddf6c24dd3615fde6f296ee28ab1be201e81572d3bff6a5f671ae

Observation e8d5339d-76e1-47eb-8184-3b2997e1852f · outbound

This paper cites MLaKE: Multilingual Knowledge Editing Benchmark for Large Language Models.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace MLaKE: Multilingual Knowledge Editing Benchmark for Large Language Models

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:42:35.368512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T15:42:33.775749Z digest=sha256:8bb6d16593b3d94d0d53faba6e99ce61b9cc591dea298c1012fe0e28a928b808

Observation fa8798e2-dfe8-4a2a-948a-3063f9ce1879 · outbound

This paper cites DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.854739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.854739Z digest=sha256:37665748d1f3b9341e2bacadf51dfa0f044266c1ae90cd097468d4efd3f1f94a

Observation 81774b8d-8151-4458-9f5b-abf3727b8984 · outbound

This paper cites SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:33.939169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:33.939169Z digest=sha256:802ae1b60a6922621e06ca39c594fb6f1f4552db7c1b81484dd0e665a09da13e

Observation 63c3cdbe-a1e5-456d-85ec-57ac075670c2 · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:34.038431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:34.038431Z digest=sha256:eb80646d0fcb2fbae1e11aa13575df80a6ce1f2b63f1424465873df1f73eaf61

Observation ba919713-722f-441b-acca-c035a371369a · outbound

This paper cites The Butterfly Effect of Model Editing: Few Edits Can Trigger Large Language Models Collapse.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace The Butterfly Effect of Model Editing: Few Edits Can Trigger Large Language Models Collapse

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:34.123986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:34.123986Z digest=sha256:bd71e76b082de81ffdc24840869826a130a3d5880e4bdfd5702bf1777a5226bc

Observation b50eff71-d362-4bc0-b505-bbc2919538a1 · outbound

This paper cites Knowledge Circuits in Pretrained Transformers.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Knowledge Circuits in Pretrained Transformers

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:34.207035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:34.207035Z digest=sha256:754862298e8c27343a5968e1d40450fee9ce14bcfd9718ae7836ac8b84fada75

Observation 5729fcf9-b6b0-4e6a-a4f9-8a32358e31cc · outbound

This paper cites Neuron-Level Knowledge Attribution in Large Language Models.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Neuron-Level Knowledge Attribution in Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:34.283123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:34.283123Z digest=sha256:df551c024bb7f2f906530ce4c49fac22f260b88084239028e96eb75bc07c83f2

Observation f7e0fd40-5c49-4eb0-b8eb-6d417cdea73a · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace OPT: Open Pre-trained Transformer Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:34.376831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:34.376831Z digest=sha256:0a42f42282c043498a3369d5cc79a25e3abcb97611e11515a1a3d84a22c38348

Observation f16bf62d-dd4e-4c87-83db-b92a1db5c31e · outbound

This paper cites an unresolved cited work.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:36.146355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T15:42:34.478018Z digest=sha256:d752ba7a14dac34ed2fcdbea23aff4e93569c633d4e433885da2a7431c80ec1f

Observation 011b081c-b633-4b13-b08b-47f08cc71277 · outbound

This paper cites Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:34.599908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:34.599908Z digest=sha256:e8c0314334612bf53e3d1b9ed096faf0dd96211adbb3000aa125cdc88fa70514

Observation 02a16147-cfe0-4bce-ad0c-128ffa0dddcf · outbound

This paper cites AutoDAN: Interpretable Gradient-Based Adversarial Attacks on Large Language Models.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace AutoDAN: Interpretable Gradient-Based Adversarial Attacks on Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:34.703077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:34.703077Z digest=sha256:3d772093e73ae620161521144cd8fa01fe11158ef6c1625858baff59e6ed342f

Observation 28e6be3c-d059-43de-8e0a-485b1903f590 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:34.774738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:34.774738Z digest=sha256:66ab9ed5334e62e2e0f5afcc5ac2d3c0df8ce6de507c773b4d89385340feb3a2

Observation ce3fa351-96ba-4940-b9f5-f6866efe69fb · outbound

This paper cites online" 'onlinestring :=.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace online" 'onlinestring :=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:34.854093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:34.854093Z digest=sha256:33832dae230e070228ea38fbba462438cda5c1f94f2c72d49b9b92a3e38ab893

Observation deca8457-878a-47e1-abdf-38d30f342aab · outbound

This paper cites write newline.

GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace write newline

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:34.927348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:34.927348Z digest=sha256:f2fa813f05a1c6318b06125524bada25cb6be158879e2c39bca647b9082df505

Pith citing papers

Observation 423b8540-b259-4a3e-b06f-2504de01f286 · inbound

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models cites this paper.

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:30:07.679358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-25T21:16:36.392606Z digest=sha256:f31d12ad2361b8cdd7028ac32122a13b6d2715b7518f079abbac22ef52f35f4a