Pith. sign in

Paper Citation Record · LEDGER

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study

As of 12 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2607.21988.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.21988 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T06:10:37.904767Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0f382788-691f-4708-b8ff-960063fc56a6 · outbound

This paper cites Qwen3 Technical Report.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study Qwen3 Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:36.999281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:36.999281Z digest=sha256:e9690b53085cdbf6e99a9775ff247b010dbb564c56998849fe4f6aa9ba253fc6

Observation e30fdb25-9eec-400f-a911-9a832bbaf29f · outbound

This paper cites In 2025 IEEE International Symposium on Technology and Society (ISTAS), pages 1–7.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study In 2025 IEEE International Symposium on Technology and Society (ISTAS), pages 1–7

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:37.058908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:37.058908Z digest=sha256:f595035dcd28a35b6acac2d3ba740204bc20341ed724cc9f015c149d154805e4

Observation 9709c33c-e441-4396-8d61-2ca1d114c4a2 · outbound

This paper cites MentaLLaMA: Interpretable Mental Health Analysis on Social Media with Large Language Models.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study MentaLLaMA: Interpretable Mental Health Analysis on Social Media with Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:37.441411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:37.441411Z digest=sha256:acc26783ebd7e02f1c1396421d407b0ffb8d0b431bf32c57c3078119819bd783

Observation cc88f229-edc1-4dd5-96c8-46390d1390d2 · outbound

This paper cites InProceedings of the 2017 Conference on Empirical Methods in Natural Language Processing, pages 2968–2978, Copenhagen, Denmark.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study InProceedings of the 2017 Conference on Empirical Methods in Natural Language Processing, pages 2968–2978, Copenhagen, Denmark

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:37.569169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:37.569169Z digest=sha256:bbe8b0bf716e64971facf32f909a80aba6f595c4764b6a6d6baa1d00c5d1d26d

Observation 14155928-89d1-45c8-83a1-6218f4bacf19 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study Representation Engineering: A Top-Down Approach to AI Transparency

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:37.904767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:37.904767Z digest=sha256:0fe2b2c80c31e544811a0cea6e35318778c74f1db2c3520ff654314a2cdd5b62

Observation 550e5d8a-eada-40e2-829c-128fb261cc49 · outbound

This paper cites Understanding intermediate layers using linear classifier probes.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study Understanding intermediate layers using linear classifier probes

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:36.658159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:36.658159Z digest=sha256:49f719dcb71a19341716e33ef25c43e933afc713684d559ad21a946ebfaf22e6

Observation 1a7f16ee-29c9-441f-ac8b-405a3d5a0624 · outbound

This paper cites InInternet Science - 4th International Confer- ence, INSCI 2017, Thessaloniki, Greece, November 22-24, 2017, Proceedings, Lecture Notes in Com- puter Science, pages 428–436.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study InInternet Science - 4th International Confer- ence, INSCI 2017, Thessaloniki, Greece, November 22-24, 2017, Proceedings, Lecture Notes in Com- puter Science, pages 428–436

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:36.928374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:36.928374Z digest=sha256:d911475e33bc7f7e7e14e59c11cd1526f1f954debc86e7cfd3f4643c4d8cfbe6

Observation 32a8b9aa-f41b-47e3-aa02-2e7baac08dc0 · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study Fine-Tuning Language Models from Human Preferences

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:37.729849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:37.729849Z digest=sha256:d06ed04bdeae016d141cc2107eee726cf59dc9b5e1835e6542f562cd392702b7

Observation 1b759818-d6f6-43df-aca8-b400f1969ebe · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:36.764038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:36.764038Z digest=sha256:d3cede835e4278173f00c32864a2ca907bf8c8c138f37b7ed8377973f2757b19

Observation 846bfb68-e57d-4e77-a4bf-61d37f65e40e · outbound

This paper cites InFind- ings of the Association for Computational Linguistics: ACL 2022, pages 566–581.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study InFind- ings of the Association for Computational Linguistics: ACL 2022, pages 566–581

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:37.186239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:37.186239Z digest=sha256:db359e2fafe0303a29eee3fcc7e9ef67917f8bf860375b98b7adbd6cf9a39c0c

Observation 624aee63-6284-4c55-81a3-024b296cc358 · outbound

This paper cites Steering Language Models With Activation Engineering.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study Steering Language Models With Activation Engineering

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:37.335164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:37.335164Z digest=sha256:6e7aee5f8f122c062a94b4f729500f631f186e780f8dc0ff73bc3fbe0683c84a

Observation 090a8873-243d-4235-a822-2053666b803a · outbound

This paper cites The Llama 3 Herd of Models.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study The Llama 3 Herd of Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:36.874629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:36.874629Z digest=sha256:c82176063a5e4a6adddc2bddf1348136f1c5182671eeb53fc41bf3b5d59fa6cd

Observation a61fee36-ccbe-4b64-a2a6-d5f5b461afe8 · outbound

This paper cites Gemma 3 Technical Report.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study Gemma 3 Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:36.824633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:36.824633Z digest=sha256:5ddbbffb0d7912aa30ba0b615aca431beae9af70fc23925bd08b690a5edc4461

Observation a5b0695d-6fd2-46f4-a6c6-33251d5ac91e · outbound

This paper cites In Findings of the Association for Computational Lin- guistics: EACL 2026, pages 809–820, Rabat, Mo- rocco.

Analysing Self-Harm Representations in Language Models: a Cross-Architecture Study In Findings of the Association for Computational Lin- guistics: EACL 2026, pages 809–820, Rabat, Mo- rocco

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-01T06:10:37.134717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:10:37.134717Z digest=sha256:af87de13b95e94511d1f557bdca1249e621b4fd08345fcf70ba53d1267e1fd46

Pith citing papers

No inbound Pith citation observations are available.