Pith. sign in

Paper Citation Record · LEDGER

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems

As of 14 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 3 inbound Pith citation observations for arXiv:2411.15662.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15662 v1

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:06:05.920095Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:46:33.511140Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

72 of 72 outbound references displayed

  • verified exact2
  • verified fuzzy18
  • unresolved48
  • parse uncertain1
  • malformed identifier3
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 7fa1e43c-9c49-48e1-a24f-66ac1cf0e593 · outbound

This paper cites Measurement Validity: A Shared Standard for Qualitative and Quantitative Research.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Measurement Validity: A Shared Standard for Qualitative and Quantitative Research

Reference 1

Resolution
malformed identifier
raw_fallback, observed 2026-08-12T14:06:09.103738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.505775Z digest=sha256:1f22c278ab38e37281bd07ac24010c4f5fbb075da1e7177c56241f2c30f2a05a

Observation 3a4bf363-655f-4f09-ae4c-a3f8cbddfa5e · outbound

This paper cites Fairness Toolkits, A Checkbox Culture?.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Fairness Toolkits, A Checkbox Culture?

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.513936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.513936Z digest=sha256:9b9ec9549ad5e2628db88b6c1345d1e6a1c348474a42d1b0dc298a2ab5450825

Observation 770be39f-7417-4b85-a776-d02b694ac282 · outbound

This paper cites The problem with bias: Allocative versus representational harms in machine learning.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems The problem with bias: Allocative versus representational harms in machine learning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:09.085018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.519618Z digest=sha256:a900b97ba3f2bfbab0158b1114bec6752c0ff84b97d51566c36fcbb7978a209b

Observation e16fcdce-bef5-46bd-aad8-c586352856ec · outbound

This paper cites Bauer and JoAnn Kirchner.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Bauer and JoAnn Kirchner

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:09.061364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.525884Z digest=sha256:a52fbc7b8577fa93952f703dc15372bb1aed1f7c23816768280c68403a3f197b

Observation 2cb66b3c-f7a8-4a22-8a4e-e8793244c468 · outbound

This paper cites Bauer, Laura Damschroder, Hildi Hagedorn, Jeffrey Smith, and Amy M.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Bauer, Laura Damschroder, Hildi Hagedorn, Jeffrey Smith, and Amy M

Reference 5

Resolution
verified exact
doi, observed 2026-08-12T14:06:06.274410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.531201Z digest=sha256:bdbacf25d2fa96b581cc3ff894afde5f2d3cb177ef6bf72616e7753e23e3cd7b

Observation 3fd24fb9-d086-4000-a14d-d28cbe062457 · outbound

This paper cites A Scoping Study of Evaluation Practices for Responsible AI Tools: Steps Towards Effectiveness Evaluations.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems A Scoping Study of Evaluation Practices for Responsible AI Tools: Steps Towards Effectiveness Evaluations

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.537189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.537189Z digest=sha256:eadcd3db038d92ed846442264f5abf49ddb53501ac082e06052d0caaa0b6239d

Observation dcb0406a-391b-49d2-9398-5c63025a1043 · outbound

This paper cites an unresolved cited work.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.544070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.544070Z digest=sha256:258519a1919e27096575450f692fc1315bd282c937bfe6bae6bca87e458a75f6

Observation 2f6d15c2-aa48-48da-8b4f-bc194e3f52ae · outbound

This paper cites Language (Technology) is Power: A Critical Survey of “Bias” in NLP.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Language (Technology) is Power: A Critical Survey of “Bias” in NLP

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.549998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.549998Z digest=sha256:335a8b4901aa94ba5739d3df5817dcba9f262b8ea461259f0e500ea171a9e08d

Observation 78a12f29-21b3-4407-b6e3-3c19b0f2d32a · outbound

This paper cites Stereo- typing Norwegian Salmon: An Inventory of Pitfalls in Fairness Benchmark Datasets.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Stereo- typing Norwegian Salmon: An Inventory of Pitfalls in Fairness Benchmark Datasets

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.555246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.555246Z digest=sha256:e3e1c5c806a8664f573c595ae7e793e6f59ae813022228b99c7bed3de5b254ab

Observation 6c2a03b1-2448-4342-a1d0-18bae4f80001 · outbound

This paper cites Man is to Computer Programmer as Woman is to Homemaker? Debiasing Word Embeddings.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Man is to Computer Programmer as Woman is to Homemaker? Debiasing Word Embeddings

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.560549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.560549Z digest=sha256:631ddcb7790e9ed82f94a0301c7c65c703d6c385cb088546a2f614daf45d6188

Observation cc619101-72a7-42b4-92bb-3f6c69dda6f9 · outbound

This paper cites Trustworthy Social Bias Measurement.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Trustworthy Social Bias Measurement

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.566726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.566726Z digest=sha256:c516714326a2fd9a048a414d3769ea9c30e84a563077b49e64cd7f225db3fa36

Observation 45964897-0c37-48bc-8a3b-6cbfcc00c8fb · outbound

This paper cites Using thematic analysis in psychology.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Using thematic analysis in psychology

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.572662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.572662Z digest=sha256:22b9f15a9a35e1a08e1210d91e1d275f9cd5b48a8a9b340c01437786f94d92a7

Observation 057f2753-d2d6-4dee-ae6e-4afb1f44eef1 · outbound

This paper cites Reflecting on reflexive thematic analysis.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Reflecting on reflexive thematic analysis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.578396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.578396Z digest=sha256:b00e44842d7e865a4c52fdbb93305633e5a14f082f2770a988d64a6b517e2e97

Observation 2f80ac12-16b7-48e1-996b-00df5f463565 · outbound

This paper cites AHA!: Facilitating AI Impact Assessment by Generating Examples of Harms.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems AHA!: Facilitating AI Impact Assessment by Generating Examples of Harms

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.583696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.583696Z digest=sha256:0e79a8cfb393f5ba26a686c9932189ec3550d07eed377727f6cf28a1eb6d17cc

Observation 5dda9893-99c7-4a4d-b591-13d17a565a74 · outbound

This paper cites Bryson, and Arvind Narayanan.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Bryson, and Arvind Narayanan

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.590180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.590180Z digest=sha256:52e9d61c6f6fdffb9b752c8fa71a7cc9668eefe2f8f5f20f1026db7dc53117d7

Observation 29a0094f-65dc-48f2-88e4-91e5263e22fc · outbound

This paper cites Representational harms through the lens of speech act theory.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Representational harms through the lens of speech act theory

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:09.042541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.601717Z digest=sha256:9e94d0e871a8c100caf06bf850a459d93f11b3423e12b9f2fe8d3579e5549181

Observation 39831700-a0d7-4001-b564-d1cc247850f8 · outbound

This paper cites The trouble with bias.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems The trouble with bias

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:09.025631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.607336Z digest=sha256:bb520ba7495ef01f6ea2b57f3f782aa1b0dcaaced20505cf912f40c4e46eae72

Observation effda52b-0e8e-4faf-b648-4a8cd68fa4fc · outbound

This paper cites Exploring How Machine Learning Practitioners (Try To) Use Fairness Toolkits.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Exploring How Machine Learning Practitioners (Try To) Use Fairness Toolkits

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:09.006613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.612992Z digest=sha256:c3004f3fd5fbabc4bf6714096e175630d53df4cae6b7071288e549bc226a02be

Observation ef59007e-2b0d-425e-92d5-325b7b7f6852 · outbound

This paper cites URL https://dl.acm.org/doi/10.1145/3531146.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems URL https://dl.acm.org/doi/10.1145/3531146

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.618443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.618443Z digest=sha256:ee91a4aa10e612fa2f17601b61b07f5623aa9e692bb9b25bf57974af12e4fd09

Observation 5254ffa9-dd04-4f23-bb6a-46cd687f8138 · outbound

This paper cites Understanding Practices, Challenges, and Opportunities for User-Engaged Algorithm Auditing in Industry Practice.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Understanding Practices, Challenges, and Opportunities for User-Engaged Algorithm Auditing in Industry Practice

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.626341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.626341Z digest=sha256:39a0c077f010717b732298b833bab22070af8aeb399c193fdf63a7aa9d5131cb

Observation 42b3932a-a2f3-4fc4-8c74-449a724c6aa7 · outbound

This paper cites Building Stereotype Repositories with Complementary Approaches for Scale and Depth.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Building Stereotype Repositories with Complementary Approaches for Scale and Depth

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.635869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.635869Z digest=sha256:69323853abd077ac892cb50754c0c899a53fbe5ee55290cd0ee6b71f1d896123

Observation 98036ae4-6148-4081-ac59-93cbfe02deda · outbound

This paper cites Multi-Dimensional Gender Bias Classification.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Multi-Dimensional Gender Bias Classification

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.646497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.646497Z digest=sha256:25b2d4a4458af596ca45698afc0fce4f1003410a54501857d82c19b004a329a3

Observation 6694a024-51ea-472f-8eb9-f831bb255a99 · outbound

This paper cites Enkin and A.R.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Enkin and A.R

Reference 26

Resolution
verified exact
doi, observed 2026-08-12T14:06:06.180161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.652168Z digest=sha256:7c6c49e6d1b256c8ab549dc87f6f9889540b12d1f8aeea604a72eb005c9f7977

Observation 60fe9edd-a4a4-4a55-8726-4b54bb987b65 · outbound

This paper cites ROBBIE: Robust Bias Evaluation of Large Generative Language Models.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems ROBBIE: Robust Bias Evaluation of Large Generative Language Models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.989054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.657225Z digest=sha256:582c4dd6425c44dcb2f960a662adafa42dfdf95770596056933d3deb57f99653

Observation a73ff923-3c12-42b9-aff5-4625752cadf3 · outbound

This paper cites an unresolved cited work.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.679091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.679091Z digest=sha256:913b6d477be34be6eec2203f1144e8e81fa5c24d1287a58e1765fc59122116a7

Observation d01501f1-4562-43a3-a110-bc9748aa9c41 · outbound

This paper cites FairPrism: Evaluating Fairness-Related Harms in Text Generation.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems FairPrism: Evaluating Fairness-Related Harms in Text Generation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.969082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.668215Z digest=sha256:b6f592c3cd1b46ce6ae2168389cae5ad3d6d90b586542626fa5296c2ecfe28bc

Observation 713e8bad-f04c-4695-8a6a-92c5f78395fe · outbound

This paper cites doi: 10.18653/v1/2023.acl-long.343.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems doi: 10.18653/v1/2023.acl-long.343

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.673506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.673506Z digest=sha256:75b503e2a76f800dfd9efa6256373788ca2933ba92d2c8c3331008dbaedf810d

Observation 7bb5e344-201e-4789-bcb0-cf51a23df1cf · outbound

This paper cites Automatically Identifying Gender Issues in Machine Trans- lation using Perturbations.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Automatically Identifying Gender Issues in Machine Trans- lation using Perturbations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.693921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.693921Z digest=sha256:8346bd780b7ccc3b45bc203fb544150da561fa46ec17707012674f5062d0a93a

Observation 60e9571b-19aa-4c75-9d79-8dffabffe8c5 · outbound

This paper cites Repairing the Cracked Foundation: A Survey of Obstacles in Evaluation Practices for Generated Text.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Repairing the Cracked Foundation: A Survey of Obstacles in Evaluation Practices for Generated Text

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.683931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.683931Z digest=sha256:f95e0002dcdda90ea780cbce1836473ac80dcc56b70589e0ebbf43662975aeda

Observation 372a9503-5f22-45f1-86bd-bb16f164e40f · outbound

This paper cites Glasgow and William T.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Glasgow and William T

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.688752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.688752Z digest=sha256:02d748b6378d57a7c6b646e451e241d529eae7f0b06b5eeb6d8935bf084bdfad

Observation 83357607-7bb4-44d0-9673-e2460437d1b4 · outbound

This paper cites ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.715377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.715377Z digest=sha256:5d846dc688cb844b20cdf649a1c36fb952a0dcdf011de75fced94857a88a8ff9

Observation 4d82746e-a102-4656-9fc7-2c0fc926892a · outbound

This paper cites Fifty Shades of Bias.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Fifty Shades of Bias

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.952387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.698681Z digest=sha256:6d0b27dcf251e3fefa7fa5982a8493b29f29d18a4393b992cccccc5c82058ec3

Observation 5634176d-c81f-497d-8285-167f88d1623f · outbound

This paper cites doi: 10.18653/v1/2023.emnlp-main.115.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems doi: 10.18653/v1/2023.emnlp-main.115

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.703573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.703573Z digest=sha256:9f403de6cbec307a06b094f0382837123174046c9f89ab21b218ef276c864212

Observation cd7048f6-4a7e-4653-8a88-2ef89ca441e8 · outbound

This paper cites Measurement: A very short introduction.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Measurement: A very short introduction

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.937658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.709681Z digest=sha256:e9fa1889779a4e4a34fa55ee1d9be6511947dbf5b277ed9923c9ac0702ceda3e

Observation dc05cc3e-0e6e-4063-b483-a3e57d6903fd · outbound

This paper cites Jacobs and Hanna Wallach.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Jacobs and Hanna Wallach

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.902886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.743538Z digest=sha256:fe766a0f04d3fdca90305307ba6796f688427cbfceaf055c4432f683fbafa598

Observation 2f713165-3bb3-4df5-8f0e-d52d1e4fe1cc · outbound

This paper cites an unresolved cited work.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.720420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.720420Z digest=sha256:281359f432c6d34df507e82f95e9dc3222bbd504a2c6dd695d0d8dd583b57bd1

Observation 0dab1596-a496-403f-8b9e-a76dced07c02 · outbound

This paper cites Ai generates covertly racist decisions about people based on their dialect.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Ai generates covertly racist decisions about people based on their dialect

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.921143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.725078Z digest=sha256:ef12c52e3b23e4e1b7a8a6a0025cec6e88a56d0deaacac53829df525e696e398

Observation 30492223-f4e6-4d24-8017-8270b9a9df25 · outbound

This paper cites Madaio, Luke Stark, Jennifer Wortman Vaughan, and Hanna Wallach.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Madaio, Luke Stark, Jennifer Wortman Vaughan, and Hanna Wallach

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.768236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.768236Z digest=sha256:e5147ddb5b01ba5af574ea6d3def447d78c4ca86e90079e1b65cfc0a8d271fa0

Observation d26118cf-d67f-476c-a230-62b2a960efa4 · outbound

This paper cites an unresolved cited work.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.737058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.737058Z digest=sha256:deb9f2720940982ccb8c41eeafd65fb7744786dc70a0acd78a26bc5af9bb0edb

Observation b48b72cb-0ca0-4d7c-ab37-3f24cfdd677d · outbound

This paper cites Fair Without Leveling Down: A New Intersectional Fairness Definition.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Fair Without Leveling Down: A New Intersectional Fairness Definition

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.886947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.778975Z digest=sha256:4a90ce9662c38ec434369796e57bdf2630212724ab8e8c327d38f7fcee35c08d

Observation 707a2d13-66d1-4ec9-a2cc-34355e9bb316 · outbound

This paper cites URL https://dl.acm.org/doi/10.1145/3442188.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems URL https://dl.acm.org/doi/10.1145/3442188

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.752027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.752027Z digest=sha256:ddfa7842ed90ba3585a73b834cb6280b9a3a0484effccbc4fa5f019f8f9e3045

Observation d363af89-8640-4147-a9eb-7a6edb84f908 · outbound

This paper cites The Landscape and Gaps in Open Source Fairness Toolkits.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems The Landscape and Gaps in Open Source Fairness Toolkits

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.758166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.758166Z digest=sha256:65e420aec77068c4d49daee07f7ae22573a773658baf96733ddbb395e8c3ebac

Observation 6e882cd1-843d-4745-b460-2ff1c4de00db · outbound

This paper cites Tran, Yi Tay, Jeffrey Sorensen, Jai Gupta, Donald Metzler, and Lucy Vasserman.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Tran, Yi Tay, Jeffrey Sorensen, Jai Gupta, Donald Metzler, and Lucy Vasserman

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.763090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.763090Z digest=sha256:ca148649012d175da2b497e2e177596722ea33a2d13823905d1f22502eb64bdd

Observation a441145a-5d45-442d-9af5-c5896e89f3f7 · outbound

This paper cites Towards AI Accountability Infrastructure: Gaps and Opportunities in AI Audit Tooling, February.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Towards AI Accountability Infrastructure: Gaps and Opportunities in AI Audit Tooling, February

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.852787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.804843Z digest=sha256:740c2fdda9ba7306a6fb0d89b93313f304a2181c4ced08ecc06179ac3f7e52a0

Observation 34020ced-0be3-4a93-b684-d3f4a9131a71 · outbound

This paper cites A Framework for Automated Measurement of Responsible AI Harms in Generative AI Applications.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems A Framework for Automated Measurement of Responsible AI Harms in Generative AI Applications

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.773247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.773247Z digest=sha256:1ff67ccea0067dd1f0ce08c742d5156da1de1bfcd2ce9dc957a8724c60c4d2a8

Observation cd86a0f8-e2e4-4ad7-b266-bc543211268e · outbound

This paper cites Disentangling and Operationalizing AI Fair- ness at LinkedIn.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Disentangling and Operationalizing AI Fair- ness at LinkedIn

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.822434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.822434Z digest=sha256:85a298a2312fa3db7869bee4073064b4a62c74966d3660cb69fb51f20d19c17c

Observation 13d65f9a-cc74-4355-abe8-beb8747db49a · outbound

This paper cites doi: 10.18653/v1/2023.emnlp-main.558.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems doi: 10.18653/v1/2023.emnlp-main.558

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.783728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.783728Z digest=sha256:83c3100e3f6704faacdfbc6b5376255afd74ff32d7dca39a0a3c47c135ed4451

Observation cd1c386d-8140-434e-8dac-70de1bacfdd5 · outbound

This paper cites Snowball sampling.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Snowball sampling

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.869651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.788555Z digest=sha256:e55ed68338a2acab4e245448b8c4e84f16ffc757093ac5aee94f6b38072849b7

Observation 85f0439a-1c5f-4b76-8231-c6e6ff621c13 · outbound

This paper cites StereoSet: Measuring stereotypical bias in pretrained language models.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems StereoSet: Measuring stereotypical bias in pretrained language models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.793709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.793709Z digest=sha256:833e6950289018a6ff811fb17d18271e331b798f87b98892866e3d8e42a205fd

Observation a2399eb7-ee6f-4898-8631-d8edd65fef18 · outbound

This paper cites an unresolved cited work.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.799486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.799486Z digest=sha256:e54bba26a682efa8d2225db503df89af14ee3126329c5eefe898f14af10d6d41

Observation fa6041b1-2b52-412d-9a30-c11f093cedfd · outbound

This paper cites an unresolved cited work.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:06:08.814796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.851823Z digest=sha256:be654474fa8c25507689d1969ab875a5ed01b74103a730a62207dd4b2e9587f8

Observation 3fe7e3ba-1ad0-4f9f-b870-68899f54a660 · outbound

This paper cites Towards AI Accountability Infrastructure: Gaps and Opportunities in AI Audit Tooling.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Towards AI Accountability Infrastructure: Gaps and Opportunities in AI Audit Tooling

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.810904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.810904Z digest=sha256:1a364a363a6e1c9296f1b0624fc360bdeec6dea392c3f3da85ec15d861a2e1f5

Observation 06d6f96f-6639-457a-a389-5948f1dbebef · outbound

This paper cites BBQ: A hand-built bias benchmark for question answering.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems BBQ: A hand-built bias benchmark for question answering

Reference 56

Resolution
malformed identifier
raw_fallback, observed 2026-08-12T14:06:08.835496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.816648Z digest=sha256:66c1032903c797c2ca1edf0306868024b5c35bc44e783c6c5b4cc42a64ea29b7

Observation 4e7ca84b-7e51-4e45-967c-4a055534f744 · outbound

This paper cites In the Walled Garden: Challenges and Opportunities for Re- search on the Practices of the AI Tech Industry.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems In the Walled Garden: Challenges and Opportunities for Re- search on the Practices of the AI Tech Industry

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.872388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.872388Z digest=sha256:b36aab5ef72266827c36e1f6995056c5a30b1d33cc24cadd284bb54e79639f01

Observation a09c6fe0-e5a7-44c3-beef-a345072d46fa · outbound

This paper cites White, Margaret Mitchell, Timnit Gebru, Ben Hutchinson, Jamila Smith-Loud, Daniel Theron, and Parker Barnes.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems White, Margaret Mitchell, Timnit Gebru, Ben Hutchinson, Jamila Smith-Loud, Daniel Theron, and Parker Barnes

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.829555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.829555Z digest=sha256:240903e8f05927f6d523c9c95e215f89500531ed678e70f667a53cec362cde35

Observation 0f6e6dee-088a-41ca-8898-e4729e752ae2 · outbound

This paper cites Where Responsi- ble AI meets Reality: Practitioner Perspectives on Enablers for Shifting Organizational Practices.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Where Responsi- ble AI meets Reality: Practitioner Perspectives on Enablers for Shifting Organizational Practices

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.835163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.835163Z digest=sha256:52e576f2cf12ceb29d6d71a8cd7bbd39fa70e182174c4fd4f9f3241d53258602

Observation fbf51ca8-7bc6-40e4-94d4-261160a62406 · outbound

This paper cites Beyond accuracy: Behavioral testing of NLP models with CheckList.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Beyond accuracy: Behavioral testing of NLP models with CheckList

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.841376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.841376Z digest=sha256:af64242ea5e21ac3326409f60beeb8a9ef9b0753459546cf3a0484cda696fb4f

Observation e1493355-9cf6-4503-9245-6c864fa27423 · outbound

This paper cites Way, Jennifer Thom, and Henriette Cramer.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Way, Jennifer Thom, and Henriette Cramer

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.846510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.846510Z digest=sha256:aa8f7b16aae97f3c69a374d52141edd57475c4fb668cc822bd5e37585a9954bb

Observation 2eb183f6-14af-4885-a548-79309716ef1c · outbound

This paper cites Measuring stereotype harm from machine learning errors requires understanding who is being harmed by which errors in what ways.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Measuring stereotype harm from machine learning errors requires understanding who is being harmed by which errors in what ways

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.758514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.904229Z digest=sha256:d4d40601bd141c17adad240672e87e7df9b3235b7cc40423832a94445e32d69d

Observation 490a2d80-c4b8-4cb0-bd57-60eb26b050d2 · outbound

This paper cites Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.742384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.910186Z digest=sha256:b399b39e9f6f270d2f462ca2e1c0b506b76008497cffd3d621979a33e9b5ebdd

Observation 742dfca1-06ab-4079-831b-e5c61582549a · outbound

This paper cites Gender Bias in Coreference Resolution.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Gender Bias in Coreference Resolution

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.862280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.862280Z digest=sha256:45c90ba57d2d6e37d164a3568a6fe30bf2bbfdb3105ebadbb467aa48f02fce57

Observation aa93a974-25a8-4ae8-b0ae-dfdc11ccc8a5 · outbound

This paper cites A Rose by Any Other Name would not Smell as Sweet: Social Bias in Names Mistranslation.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems A Rose by Any Other Name would not Smell as Sweet: Social Bias in Names Mistranslation

Reference 65

Resolution
malformed identifier
no resolver link, observed 2026-08-12T14:06:05.867203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.867203Z digest=sha256:f0bdf66241109a03b250b396e8d12dd8ab7cad05ae5f9923dedd372398d63df5

Observation 0824278a-2f50-4552-b32f-610493099c6f · outbound

This paper cites The Woman Worked as a Babysitter: On Biases in Language Generation.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems The Woman Worked as a Babysitter: On Biases in Language Generation

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.777432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.877583Z digest=sha256:0c1674a9e39f04a9c91da0d78fd8f3f58fed85cc34f224e686c93d002d03eb50

Observation 827d7529-d802-4ab6-8447-b025dba76aac · outbound

This paper cites ‘how many cases do i need?’: On science and the logic of case selection in field-based research.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems ‘how many cases do i need?’: On science and the logic of case selection in field-based research

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.889157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.889157Z digest=sha256:d7bb0b78494a43405a95fdbdcfdc0b5fdd8e5bcdeae6f310755e8ebe45e8dddf

Observation 8af82e98-a953-4641-a760-6c3c5843716d · outbound

This paper cites Evaluating the Social Impact of Generative AI Systems in Systems and Society.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Evaluating the Social Impact of Generative AI Systems in Systems and Society

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.893874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.893874Z digest=sha256:5dd2924bfe4df0c439d042010e838ecff66e23fa1bf2d2aacebbbf6a7b9ef710

Observation 859ed729-07f3-4caf-8d4d-8ecc454eadf3 · outbound

This paper cites Measuring Representational Harms in Image Captioning.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Measuring Representational Harms in Image Captioning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.899167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.899167Z digest=sha256:c96e621e8f58480b6c599c5467ce689d5cf46048cf19b342960678beb42239c0

Observation e846f195-ff25-47b1-aafd-67506e23e18e · outbound

This paper cites Wildchat: 1m chatGPT interaction logs in the wild.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Wildchat: 1m chatGPT interaction logs in the wild

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:06:08.714743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.920095Z digest=sha256:1633abc2ef7e159392f1024f147edfda33ef57c8fa62253a9d4117b0d51bf724

Observation a00bd9d8-eca4-4463-8765-68af4a5d600d · outbound

This paper cites an unresolved cited work.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems Unresolved cited work

Reference 1962

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T14:06:08.795098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T14:06:05.857019Z digest=sha256:81b8e7ec57431be33d4fdcd1cbb843b6f7285ca916b1ac88925246811efaadfa

Observation 0bee4036-a124-4011-98d3-dc25760eef18 · outbound

This paper cites doi: 10.18653/v1/N18-2003.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems doi: 10.18653/v1/N18-2003

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.914864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.914864Z digest=sha256:f5041793f69102990101dca0f4f04f66779c802f90295801be092ca668fc4b59

Observation 46e75d1e-2d08-4840-8f31-298377ba1f0a · outbound

This paper cites doi: 10.18653/v1/D19-1339.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems doi: 10.18653/v1/D19-1339

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.883467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.883467Z digest=sha256:820501118ccb3cd623c2d984d7e97c851b08d578dc0c8a8c2a93dd564533aa28

Observation 22156abe-5141-4504-bee5-e3c0df620d08 · outbound

This paper cites doi: 10.18653/v1/2023.emnlp-main.230.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems doi: 10.18653/v1/2023.emnlp-main.230

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.662969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.662969Z digest=sha256:4f0949b1f7f7d12ad5139cab3c992359c21f608ef5c8b987c6c06ef919d877a9

Observation 36a32167-5ed1-4aef-92dc-a3e6e6216fe9 · outbound

This paper cites doi: 10.1038/s41586-024-07856-5.

Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems doi: 10.1038/s41586-024-07856-5

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T14:06:05.730447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:06:05.730447Z digest=sha256:ef82544007f28decdbe55451e726cb260004b79d52db348a5dc9e792e73d52a5

Pith citing papers

Observation b650bb70-8265-4251-a831-0ca5fc41754c · inbound

A Framework for Auditing Chatbots for Dialect-Based Quality-of-Service Harms cites this paper.

A Framework for Auditing Chatbots for Dialect-Based Quality-of-Service Harms Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T10:46:33.511140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:46:33.511140Z digest=sha256:d62d92b7da116f268b7486885e3fb437c4542d2a05685278a4df4a0277bb51c0

Observation b99f0607-5bde-4907-983e-bf5ae46662e4 · inbound

Toward Valid Measurement Of (Un)fairness For Generative AI: A Proposal For Systematization Through The Lens Of Fair Equality of Chances cites this paper.

Toward Valid Measurement Of (Un)fairness For Generative AI: A Proposal For Systematization Through The Lens Of Fair Equality of Chances Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T19:47:45.252044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:47:45.252044Z digest=sha256:b691b196251b835849f5d33eb55bb4e02cdd314e5ec9d8959784317267ee953a

Observation 18c81d29-605b-4224-9efc-87910cd11e6a · inbound

Discriminatory Compliance: How LLMs Answer Queries from Protected Groups cites this paper.

Discriminatory Compliance: How LLMs Answer Queries from Protected Groups Gaps Between Research and Practice When Measuring Representational Harms Caused by LLM-Based Systems

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-06-26T12:59:29.367774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-26T12:58:09.227648Z digest=sha256:29643c7e7b3354f0aee92915583227b280ba1eb7b096e2546d433647613ccb26