Pith. sign in

Paper Citation Record · LEDGER

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective

As of 10 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 4 inbound Pith citation observations for arXiv:2506.05166.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05166 v2

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:29:45.002961Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T10:04:06.753097Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-14T21:18:00.318407Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3772aed2-6acb-4a08-a69c-7a524d6ca671 · outbound

This paper cites online" 'onlinestring :=.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.768165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.768165Z digest=sha256:ee8171bdde2da5d6ec3e48add98c224808d8418556630a174696d95cc129edfe

Observation e06ece96-14c5-464e-834d-dc7ad7b63282 · outbound

This paper cites write newline.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.778131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.778131Z digest=sha256:8ee47973588f0d733bf69d39e5d965c60b07ed6027028f6310a0b1ad2b0e229a

Observation 39a7a4ba-b128-4fe7-9ec3-41d8d5ca2ab9 · outbound

This paper cites write newline.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective write newline

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.786484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.786484Z digest=sha256:e0b79abd1973f19f7754505f6fded3f6200bbd6061632231a010de51c99c3069

Observation fcfed349-23e6-4ade-af55-189c39c41452 · outbound

This paper cites Stubbersfield.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Stubbersfield

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.577242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.794108Z digest=sha256:e12d8d45340a48a2d4e863a9d810814cdd6a1e945533c80935728548a7df7a65

Observation ec4dd6b9-8fdc-4074-b803-81dd46760915 · outbound

This paper cites Science in the age of large language models.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Science in the age of large language models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.561584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.802060Z digest=sha256:68f2649e1b78cfc72a0817bbb8322bef19cdbb432bc63da2989b544b451aa817

Observation a5cd9e91-6d6b-45aa-9955-547dceb2b4b0 · outbound

This paper cites Quantifying and Reducing Stereotypes in Word Embeddings.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Quantifying and Reducing Stereotypes in Word Embeddings

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.807248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.807248Z digest=sha256:a79bd40d4e3bd41cc67781afcf8eb06cc9ab3573879e0b63086301995c73d2e5

Observation efc5a07e-07a5-4a00-aa93-730a8a3a21fd · outbound

This paper cites Man is to computer programmer as woman is to homemaker? debiasing word embeddings.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Man is to computer programmer as woman is to homemaker? debiasing word embeddings

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.546048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.814121Z digest=sha256:e9290244d2c12babc9b046425d6a7a2172c273c0cf751322a13c1dbf88499f2a

Observation a4c7aba8-cf3a-4836-ba85-db936d2ed7c8 · outbound

This paper cites Identifying and Adapting Transformer-Components Responsible for Gender Bias in an English Language Model.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Identifying and Adapting Transformer-Components Responsible for Gender Bias in an English Language Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.819483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.819483Z digest=sha256:003031c3f1d0ec76e53badfd6c6ba9f9b8e7e4b8c8d40a79213829d89baf8d7a

Observation 0b4416e2-4e1f-4970-a8e2-a74e81159d52 · outbound

This paper cites Towards automated circuit discovery for mechanistic interpretability.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Towards automated circuit discovery for mechanistic interpretability

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.827250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.827250Z digest=sha256:689d6fa083b7e8b33b5b085503db6d621686f0ff9e818d6e219c2a1b541de474

Observation 66ee3a01-cf46-4cfe-ae0f-d8639c625ca7 · outbound

This paper cites Gallegos, Ryan A.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Gallegos, Ryan A

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.520197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.834649Z digest=sha256:a410371d7d11023cc8b774219530fca600268a4ca10a5fffcfa7e264b8744a78

Observation 358dd5c8-2a99-49ff-879c-add0019436fa · outbound

This paper cites Causal abstractions of neural networks.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Causal abstractions of neural networks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.841155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.841155Z digest=sha256:ecd06219550391cd78fdbc08750fec478d904f53bf661fdfbec4f80ccb1c3eac

Observation 9e3a3c02-6824-4351-b7b7-f80224012029 · outbound

This paper cites Multimodal neurons in artificial neural networks.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Multimodal neurons in artificial neural networks

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.494054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.849885Z digest=sha256:2ff668a435f78c159718d72ebc773a1303edd70f3090011f15ee332486e721e0

Observation 9b87056f-b8f7-4591-ae43-061283bd12ea · outbound

This paper cites Localizing Model Behavior with Path Patching.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Localizing Model Behavior with Path Patching

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.855872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.855872Z digest=sha256:97bda988cd5ee8fef7a1ced56e23766a088b1518524a23264f714aa3c43ddeaa

Observation 97959f3a-54bf-4423-9ae8-50f52bc579d7 · outbound

This paper cites C hat GPT based data augmentation for improved parameter-efficient debiasing of LLM s.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective C hat GPT based data augmentation for improved parameter-efficient debiasing of LLM s

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.477923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.861398Z digest=sha256:9753ac32bf04eb5891b0c0b686d36037408e288c014da8f2da3ddf69f3993c65

Observation 78ad67f0-2229-4412-89cb-86380b5dfe33 · outbound

This paper cites distilbert-base-uncased-finetuned-sst-2-english (revision bfdd146), 2022.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective distilbert-base-uncased-finetuned-sst-2-english (revision bfdd146), 2022

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.461355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.867114Z digest=sha256:3373884f24938f069c64a1b45684f98bd05b7effe862f26c2b3c17fa18a5d4f9

Observation 16c76e03-68e9-4de1-9c26-09f0ed430230 · outbound

This paper cites Shovon, and Gene Kim.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Shovon, and Gene Kim

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.444004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.872004Z digest=sha256:885cd159bd5a33af308da79fb53aba04b82b368c7499fab0da2e410bee04bb7a

Observation c8a8dcd2-ad11-46bf-af70-bcb691cfda39 · outbound

This paper cites The impact of debiasing on the performance of language models in downstream tasks is underestimated.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective The impact of debiasing on the performance of language models in downstream tasks is underestimated

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.428479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.877309Z digest=sha256:cd6116a3e7edd12d061c7a098c5389dd878c378514a9d15c53c132485f4bb556

Observation d0f81c75-6dd2-4d9a-9ad1-a456da917fd3 · outbound

This paper cites Backward Lens: Projecting Language Model Gradients into the Vocabulary Space.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Backward Lens: Projecting Language Model Gradients into the Vocabulary Space

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.882463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.882463Z digest=sha256:e8e080b4f2245f55360bdde581d1bfb8b6573eb07c90809e012cd632227639dc

Observation f1d12fa1-2227-4053-9c63-9e7738c4a344 · outbound

This paper cites Linear Representations of Political Perspective Emerge in Large Language Models.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Linear Representations of Political Perspective Emerge in Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.887656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.887656Z digest=sha256:8d55e7c69e4e82a7f450c693fd895bf48418bb690f725473c49ae7903fb14a7a

Observation 02b82d85-c601-4faf-8e6c-89115c2f9854 · outbound

This paper cites Gender bias and stereotypes in large language models.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Gender bias and stereotypes in large language models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.411906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.893433Z digest=sha256:ef27447282be6cf85f981e9c25a26ce604d421b996479d31c732bcc2a0832b9c

Observation 791accd6-099f-44b4-a176-3cb679e6cb5b · outbound

This paper cites Sparse feature circuits: Discovering and editing interpretable causal graphs in language models.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Sparse feature circuits: Discovering and editing interpretable causal graphs in language models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.395503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.899079Z digest=sha256:c962785b86de81d8242bb6d61972fb794adf63c67f253c97d43f0d8c12545a83

Observation bf25f2f4-f439-4c5c-bcda-4c18322f57d3 · outbound

This paper cites Locating and editing factual associations in gpt.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Locating and editing factual associations in gpt

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.904438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.904438Z digest=sha256:8cbd428466823ef5911704c1b577afddb610c0e48af9a4500294fc90b402ff58

Observation 66267b2e-0673-442c-8df7-403023597f77 · outbound

This paper cites Progress measures for grokking via mechanistic interpretability.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Progress measures for grokking via mechanistic interpretability

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.911625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.911625Z digest=sha256:123591c724ffda7e1ec4b8ab8afb69665390dd8c2313f538b4e97afb6382d998

Observation bc928338-d764-4f8e-b201-698f2db36b69 · outbound

This paper cites Nationality bias in text generation.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Nationality bias in text generation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.369665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.917276Z digest=sha256:2febe605cbb02e8caec5dca0c5b2158601d228106521d4f16408672c919fc189

Observation 5ff7eb45-af62-47df-b791-e1d12b8744d1 · outbound

This paper cites Biases in large language models: Origins, inventory, and discussion.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Biases in large language models: Origins, inventory, and discussion

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.354171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.921954Z digest=sha256:f65ac03319ad3424f6fd5f592c762990e9e78cc1a0477e6ece2787f3a52b4cf5

Observation 847c158c-e188-4133-9580-7cf2bc72b50b · outbound

This paper cites Mechanistic interpretability, variables, and the importance of interpretable bases.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Mechanistic interpretability, variables, and the importance of interpretable bases

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.337615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.928556Z digest=sha256:759bb98dd1e473347a5cc84b5868c9c3065cfa4eb2dee1b819f1599827c473ef

Observation 9d893a85-de54-4037-a156-6432b71273b4 · outbound

This paper cites Zoom in: An introduction to circuits.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Zoom in: An introduction to circuits

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.934254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.934254Z digest=sha256:2ecfc932cf2839e30f7afc3e7738aa893133d77503c7e8bbfa0e001f59505cac

Observation 135a2cb1-9ab0-45a4-b8d0-a2cede383089 · outbound

This paper cites Gender biases in automatic evaluation metrics for image captioning.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Gender biases in automatic evaluation metrics for image captioning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.310828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.942668Z digest=sha256:f08cd1db36cf59bbf602ca335f37108d18af3ff31402ad8f072bafa329ecdedb

Observation 0cf90f9a-f802-4b7f-a6c8-98685c9a04c0 · outbound

This paper cites Language models are unsupervised multitask learners.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Language models are unsupervised multitask learners

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.948884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.948884Z digest=sha256:0c9a794c0492e1ef7fefa0e8bb8c67c5c4251335f9074e1028cd7410118d351e

Observation c7fc6bcb-07b2-4be0-9c34-a6165b7b0261 · outbound

This paper cites Introduction to the CoNLL-2003 Shared Task: Language-Independent Named Entity Recognition.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Introduction to the CoNLL-2003 Shared Task: Language-Independent Named Entity Recognition

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.953942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.953942Z digest=sha256:60a6321d9a092fc4daa2638b6ed09897f0d309d2f4b022b2085fd78f26c20363

Observation 8fede18b-8b4e-4c08-b56d-f4889fe45dbd · outbound

This paper cites Investigating gender bias in large language models through text generation.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Investigating gender bias in large language models through text generation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.285328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.958941Z digest=sha256:ed8ac767da0b9c81845213a1b0c1804adae142786102d8dfc9f36514edac821c

Observation ef6e9eec-cce5-4f92-9452-88191b5a20b1 · outbound

This paper cites Attribution Patching Outperforms Automated Circuit Discovery.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Attribution Patching Outperforms Automated Circuit Discovery

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.964417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.964417Z digest=sha256:a7ae23cbc7602d48d8cb54694466158b948a33b5a270cf6f7eead4188ecf3430

Observation 7bb28f1c-b4cf-4562-bbc5-b74d822a9173 · outbound

This paper cites Attribution patching outperforms automated circuit discovery.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Attribution patching outperforms automated circuit discovery

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:29:45.268077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:29:44.969616Z digest=sha256:199c449ea9645ad51adaf0c827f172aacb3ca0af5902012abdff43dce40d391b

Observation 6d1dcac1-8530-40f4-81a1-002bab812bb4 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.976406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.976406Z digest=sha256:5a92a6cbcd3971f32c3905fafa0f32f56ee309ec9ada96102ffdc0582e976afc

Observation 6e513610-6d58-4c62-a29b-0ed757fc2a36 · outbound

This paper cites Investigating gender bias in language models using causal mediation analysis.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Investigating gender bias in language models using causal mediation analysis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.981245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.981245Z digest=sha256:afb11d47fcf668e83dcb311c3ee42d0f97d5e69c496d057a6fbb71d364ca2bea

Observation 795ad8e5-0f9c-470e-b00f-ec2eff7fd9fe · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.987123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.987123Z digest=sha256:26d4433d05eb09ce1c3b6c875651f40e5162784b6a3c4e59173b133ebad8cb44

Observation 26a9f4e5-0caa-4643-80c8-0ee51cb33f37 · outbound

This paper cites Neural Network Acceptability Judgments.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Neural Network Acceptability Judgments

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.993184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.993184Z digest=sha256:f448233a64e2cf52022c8888b2154c9aa56a83d60449ab27915e90a22ff87d31

Observation 79525a86-b6b0-4a39-b3e9-d08605577f50 · outbound

This paper cites Interpretability at scale: Identifying causal mechanisms in alpaca.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Interpretability at scale: Identifying causal mechanisms in alpaca

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:44.998131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:44.998131Z digest=sha256:6d8288caeaad389521228c49f832587ebf7046f9a194b59b8dcc11a4d3c4f905

Observation 45d1bd76-2e0c-4611-b48c-608a999fa62a · outbound

This paper cites Towards Best Practices of Activation Patching in Language Models: Metrics and Methods.

Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective Towards Best Practices of Activation Patching in Language Models: Metrics and Methods

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:45.002961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:29:45.002961Z digest=sha256:e86c4a26c4184ec95e65fe293a5ba83a35e8fc077ec2f8414cbe0986ec1babe3

Pith citing papers

Observation 6a2ef335-0406-48f7-87ce-ccb2a49563dd · inbound

Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability cites this paper.

Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T10:04:06.753097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:04:06.753097Z digest=sha256:7db1ebc2c8150acc845d372707177f161446b0a48e2faee624db37483272ddea

Observation aa71bf9e-5315-4636-9005-727082cadf69 · inbound

Explainability of Large Language Models: Opportunities and Challenges toward Generating Trustworthy Explanations cites this paper.

Explainability of Large Language Models: Opportunities and Challenges toward Generating Trustworthy Explanations Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:58.465833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:58.465833Z digest=sha256:fa763428c9b3587db4010e99d6ebe2671ea7748b5cf0a4ac5dbf791d03d7573a

Observation 75533c7d-b5d6-4745-b05d-4ef1be771b61 · inbound

Attention, May I Have Your Decision? Localizing Generative Choices in Diffusion Models cites this paper.

Attention, May I Have Your Decision? Localizing Generative Choices in Diffusion Models Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:18:00.320812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T21:09:28.032659Z digest=sha256:2b88fb8047570457fc9ed8e4368bdbd44c0d7fe4aeace4889f7ffa7159e37005

Observation d9d1ab2f-48c3-4301-a161-5a866954f5e9 · inbound

GKnow: Measuring the Entanglement of Gender Bias and Factual Gender cites this paper.

GKnow: Measuring the Entanglement of Gender Bias and Factual Gender Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T04:52:16.728754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T04:50:43.547037Z digest=sha256:98ae66869d38e4cf60674ba5c481ed000547274f39dcdb8113c4d28615fcff04