Pith. sign in

Paper Citation Record · LEDGER

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet

As of 18 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2502.05291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.05291 v2

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T19:55:43.420160Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b11e1da6-8dc2-453b-acec-94e85de860a4 · outbound

This paper cites Avoid areas with heavy police presence.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Avoid areas with heavy police presence

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:55:43.618677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T19:55:43.407539Z digest=sha256:7049c4024d582f2216f537520ba4c2c2fba418c464bbea1287eba54cbe39dd94

Observation ab7bf838-1006-4839-a28e-f3d107c9fc2c · outbound

This paper cites The Llama 3 Herd of Models.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet The Llama 3 Herd of Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.370629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.370629Z digest=sha256:8189e34d2f7c55fe433443d2503b60af646712fd5b422605f9fb326cc1bdae55

Observation 068a2ea1-8064-480c-9b56-0b76e9806fe3 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.392335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.392335Z digest=sha256:06f5396670e5de05c9a9ffd3ad305a5062aed7310d30091c6d6c911e38d3f7dc

Observation bd564e33-3a10-42ee-9c5b-676dbbee2368 · outbound

This paper cites SHAKTI: A 2.5 Billion Parameter Small Language Model Optimized for Edge AI and Low-Resource Environments.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet SHAKTI: A 2.5 Billion Parameter Small Language Model Optimized for Edge AI and Low-Resource Environments

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.397257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.397257Z digest=sha256:bf34bb0c787063b604db0bb13c7e3adcb4d31d25bca2a254c5999d01761d634b

Observation 949af65b-9330-47be-b971-b0c30670d6a4 · outbound

This paper cites an unresolved cited work.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-08T19:55:43.603638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T19:55:43.411737Z digest=sha256:1567d7708d1fc3ce7f83c579a187cf0ac869e2436156a79f566eb48194206a29

Observation 209e314d-0538-4fbd-a85a-7d9cb694633d · outbound

This paper cites Try to maintain eye contact and act natural.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Try to maintain eye contact and act natural

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:55:43.590031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T19:55:43.415971Z digest=sha256:00e925cf6b82ed6fd4171400ff35f3159d80f0e411cfdd11ef5a81be610a568b

Observation 491e9323-b43f-4f45-b17e-f7d6d07fab4b · outbound

This paper cites You would probably try to find a place where you could get close to your victim without being seen.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet You would probably try to find a place where you could get close to your victim without being seen

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:55:43.575662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T19:55:43.420160Z digest=sha256:a7a4cc0541b020edf42a3d6c9431fc90e5c77209ce4d2c407aeb6ca29dad2537

Observation 371885d4-e02c-4e9c-8b92-841e9565bd6d · outbound

This paper cites On-Device Language Models: A Comprehensive Review.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet On-Device Language Models: A Comprehensive Review

Reference 2010

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.402474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.402474Z digest=sha256:3e07bf4100619211dbacb46aebed03c0e70ec3e237ae2fa836468772ac5bed4e

Observation 1abd9e89-720a-410d-8317-5107d8d20229 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.365035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.365035Z digest=sha256:39fac4d830e3f55ef7e43f1d79a245467a70872952afb09bb52efcfc5d5a3565

Observation 875fe8b1-b9bd-4700-a9ff-7231136179cf · outbound

This paper cites Hey GPT, Can You be More Racist? Analysis from Crowdsourced Attempts to Elicit Biased Content from Generative AI.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Hey GPT, Can You be More Racist? Analysis from Crowdsourced Attempts to Elicit Biased Content from Generative AI

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.376744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.376744Z digest=sha256:b2555aa1cb83ef5d3cadb3938848639c0e44aa4ea9c620d071bba43cc2320c12

Observation 6aff4c95-b6ab-4d64-900c-b30c753babc7 · outbound

This paper cites WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.387259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.387259Z digest=sha256:29e22d080345d200f852b7a2e200496cdf51d4c154364ba863e645af68b17b89

Observation 984393f2-9028-4a00-98b8-316e8b7c20b1 · outbound

This paper cites Precision Knowledge Editing: Enhancing Safety in Large Language Models.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Precision Knowledge Editing: Enhancing Safety in Large Language Models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.382186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.382186Z digest=sha256:536b7dab5b50b6abee83e726dcd4082de1642bf32d88079d78f19b1989ec7014

Pith citing papers

No inbound Pith citation observations are available.