Pith. sign in

Paper Citation Record · LEDGER

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks

As of 1 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 4 inbound Pith citation observations for arXiv:2502.04419.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04419 v3

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-23T03:53:00.509414Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-01T06:32:01.292127+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T22:49:03.124658Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-06-28T22:52:45.523134Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact21
  • verified fuzzy21
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f49515d-d79e-4eb4-8cfb-9ebc312b69ad · outbound

This paper cites Phi-4 Technical Report.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Phi-4 Technical Report

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T03:55:21.911213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:274c8840a99edb18a09d76fa36538e23ceff3ad84759274171c87c00963d2da4

Observation 922ede37-ebb6-4baf-8953-181b1dd35375 · outbound

This paper cites Physics of Language Models: Part 3.1, Knowledge Storage and Extraction.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Physics of Language Models: Part 3.1, Knowledge Storage and Extraction

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.891117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:35c3de37f956a1acc574a1ea74893e37d854d0ba8904443f28fe7fe1ebc7cb67

Observation 4747ffc8-e229-4155-b11e-fd746416bd47 · outbound

This paper cites Overview of mex-a3t at ibereval 2018: Authorship and aggressiveness analysis in mexican spanish tweets.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Overview of mex-a3t at ibereval 2018: Authorship and aggressiveness analysis in mexican spanish tweets

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.343476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:f60e941b670915fbcc75cf231da372059f1cfff8e1ee4a65ad59bbf8f91b09da

Observation 384891be-0c0d-47ff-ba75-1a9a50e9e6ad · outbound

This paper cites Measuring Implicit Bias in Explicitly Unbiased Large Language Models.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Measuring Implicit Bias in Explicitly Unbiased Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.879588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:3f7e852b5f0e80da8047c80d40b5e80936f0d5e2db2a0e895c056f3111133902

Observation f27ff4a1-536a-4f01-b641-1df362127b57 · outbound

This paper cites Semeval-2019 task 5: Multilingual detection of hate speech against immigrants and women in twitter.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Semeval-2019 task 5: Multilingual detection of hate speech against immigrants and women in twitter

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.353664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:93b8de5e2727b8257d6aefe36d55ce6b982f6cd3f912b1a096abb839527e9696

Observation 6107f8dc-2780-441a-8cc2-d62e17b18b43 · outbound

This paper cites Revealing Hidden Bias in AI: Lessons from Large Language Models.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Revealing Hidden Bias in AI: Lessons from Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.918365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:e9db5d096254ee033507c5e96d3dd5a67fff0d7bc977cb1bafafdb1c1eb17d4f

Observation 4b771d01-1630-4144-a638-b6d8975d1442 · outbound

This paper cites Language models are few-shot learners.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Language models are few-shot learners

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.356213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:643e8a21c5e6eba278a579b9707fefe7bc507a2417994545a257767b42df5df4

Observation c6b32a61-7784-426f-8313-d647dd1ef884 · outbound

This paper cites AGR: Age group fairness reward for bias mitigation in LLMs.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks AGR: Age group fairness reward for bias mitigation in LLMs

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.388992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:2759ce127cced9e7726017c7622d4be06b6afbd454562d86c7f55dd38c050941

Observation e56fdd23-24ca-46ad-8e30-d5f34d4c0296 · outbound

This paper cites URLhttps://aclanthology.org/2020.lrec-1.761/.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks URLhttps://aclanthology.org/2020.lrec-1.761/

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.345648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:2c1cea9f72fe04a6f3cb36244635b356978b75f92933daec52101aec28a2e9ff

Observation 9416f78d-23e2-4137-b857-6067f5c979f4 · outbound

This paper cites Data augmentation using llms: Data perspectives, learning paradigms and challenges.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Data augmentation using llms: Data perspectives, learning paradigms and challenges

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.348447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:53cb8772ae945f244b421204bbe1db6e374851e5c3310f26e20c7150645ca34d

Observation 6d06e287-e6b7-40db-8c58-ecbbabd2e380 · outbound

This paper cites Sina at FigNews 2024: Multilingual Datasets Annotated with Bias and Propaganda.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Sina at FigNews 2024: Multilingual Datasets Annotated with Bias and Propaganda

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.900421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:197fdc38ef9bcbf53b8685ed796726e2ddc899e6ec0aac24ab13ae5f512ac26a

Observation 149613b7-e315-40ca-a6d7-6f6d4082f1b5 · outbound

This paper cites Towards Measuring the Representation of Subjective Global Opinions in Language Models.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Towards Measuring the Representation of Subjective Global Opinions in Language Models

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-23T03:55:21.894076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:f68aba208b5a68b7d206d4e91354416cd4d5a6962247fa86e1d30d40f050817e

Observation 90ae74cf-1825-4618-8036-92ebbba00b33 · outbound

This paper cites First-Person Fairness in Chatbots.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks First-Person Fairness in Chatbots

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.921936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:fb3543dce62fee9320202db5eae44dd65f957ca2b24bd669d5d95eb5db410608

Observation e12e55e4-c0a9-4c41-a113-4849c413a4cb · outbound

This paper cites Jillian Fisher, Shangbin Feng, Robert Aron, Thomas Richardson, Yejin Choi, Daniel W Fisher, Jennifer Pan, Yulia Tsvetkov, and Katharina Reinecke.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Jillian Fisher, Shangbin Feng, Robert Aron, Thomas Richardson, Yejin Choi, Daniel W Fisher, Jennifer Pan, Yulia Tsvetkov, and Katharina Reinecke

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.873189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:c0d33359915a56496ac58b3ac36d382cbe546cf6d94662bb252ad101e17f342e

Observation e4d1c963-08e5-423a-b865-fd059a5455e5 · outbound

This paper cites Modeling Human Subjectivity in LLMs Using Explicit and Implicit Human Factors in Personas.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Modeling Human Subjectivity in LLMs Using Explicit and Implicit Human Factors in Personas

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.919993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:45e93eeabe8fa48a764febd40492639962f4150f0d0eb66a14c3350749257efb

Observation 8587fe5f-b75f-48c9-8ee0-027f8ebc1599 · outbound

This paper cites Hey GPT, Can You be More Racist? Analysis from Crowdsourced Attempts to Elicit Biased Content from Generative AI.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Hey GPT, Can You be More Racist? Analysis from Crowdsourced Attempts to Elicit Biased Content from Generative AI

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.889220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:398e1258127bd70cfac79e09afb7069799a25908a2c3d969a9228dede35d8375

Observation f84421a6-54ef-48c6-976e-c4c7a7c0671d · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks LoRA: Low-Rank Adaptation of Large Language Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-23T03:55:21.925097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:ae7c1736951f4e80307e9b8c354f47161d73698b372929c4d0d5ecdce3e6e6e1

Observation 59833ce8-cf64-4aa6-91a2-7ba64860836b · outbound

This paper cites Huggingface.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Huggingface

Reference 18

Resolution
verified exact
doi, observed 2026-05-23T03:55:21.700811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:5b748d061e1ebcb946b05886225c335792511907e857e68b51fa9a65d49277f8

Observation 5c4e54b9-aa56-40f0-97bf-8c2a19258704 · outbound

This paper cites URLhttps://aclanthology.org/2020.osact-1.8/.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks URLhttps://aclanthology.org/2020.osact-1.8/

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.376715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:f0824a067a52ac39ccdc10d885d95b55d5427cb8b9ee5c3214bed10c1a71ee4f

Observation 472c9b56-d825-48ae-a341-0b4d197db65c · outbound

This paper cites Shovon, and Gene Kim.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Shovon, and Gene Kim

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.371395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:fa14c17292e71ffff2d04546624883cd69542fa0dd109a6eba63cc9d61641b80

Observation 327b9d34-50e1-45b0-878c-37264f8b9b93 · outbound

This paper cites doi: 10.18653/v1/2024.findings-acl.530.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks doi: 10.18653/v1/2024.findings-acl.530

Reference 21

Resolution
verified exact
doi, observed 2026-05-23T03:55:21.715398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:33e1cfe242acc99c869b21f1b78e7ecdc96ebfa4758531201747bd76f1d80852

Observation 0d9ace79-0cc5-4cd2-afe8-975950fefa54 · outbound

This paper cites Subtle biases need subtler measures: Dual metrics for evaluating representative and affinity bias in large language models.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Subtle biases need subtler measures: Dual metrics for evaluating representative and affinity bias in large language models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.379725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:a7c805b45c448bc2d20782464a68d8eb846f79a6c27f8377ca4bfe529c43d611

Observation 0513b694-593d-4f8d-a854-956a086452a3 · outbound

This paper cites Ruleprompt: Weakly supervised text classification with prompting plms and self-iterative logical rules.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Ruleprompt: Weakly supervised text classification with prompting plms and self-iterative logical rules

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.382969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:5ecd3a2feed2f0a5a9a856c3d8d3215ef281342ca36e0839099a1417870352b7

Observation 3b837be3-5946-469f-ab75-6761d8314b29 · outbound

This paper cites Chen Liu, Fajri Koto, Timothy Baldwin, and Iryna Gurevych.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Chen Liu, Fajri Koto, Timothy Baldwin, and Iryna Gurevych

Reference 24

Resolution
verified exact
doi, observed 2026-05-23T03:55:21.712194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:afddf58109f447071fb951e5a3f19a5b6e562e6a5ae3b8fc0f7c472dc3a44689

Observation 1d96d095-7c70-4498-bc0e-8e6dd08ad316 · outbound

This paper cites The generation gap: Exploring age bias in the value systems of large language models.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks The generation gap: Exploring age bias in the value systems of large language models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.371164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:9c13891b6f235e664584b0c19c2bbbb2c410e9317ee93f02e9e4fc2446864a9d

Observation cde44b65-113e-4441-8ddc-d5a3d2f603bd · outbound

This paper cites AI-UPV at IberLEF-2021 DETOXIS task: Toxicity Detection in Immigration-Related Web News Comments Using Transformers and Statistical Models.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks AI-UPV at IberLEF-2021 DETOXIS task: Toxicity Detection in Immigration-Related Web News Comments Using Transformers and Statistical Models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.368933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:dca07f2a39161d49ed530abb32ba7acdbb9a2a8a298118fbaf5b4184fa3d21b8

Observation 84a321ac-336e-47b5-bf48-e2db02a37647 · outbound

This paper cites AI-UPV at IberLEF-2021 DETOXIS task: Toxicity Detection in Immigration-Related Web News Comments Using Transformers and Statistical Models.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks AI-UPV at IberLEF-2021 DETOXIS task: Toxicity Detection in Immigration-Related Web News Comments Using Transformers and Statistical Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.863950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:5a96ab746becd6f5b90efc7c27765bf523557e9897bccf180381bf7449020c2a

Observation 3d3edaee-d305-4a89-9b71-13bccf96f3bd · outbound

This paper cites Text classification using label names only: A language model self-training approach.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Text classification using label names only: A language model self-training approach

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.358689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:49871d73e50a45749abef7afe4b1f2380d465a8e0775f5b3f982623bda0dbe73

Observation 390c7e7c-09b2-4edb-aeaf-47d8be059c25 · outbound

This paper cites Tarek Naous, Michael J Ryan, Alan Ritter, and Wei Xu.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Tarek Naous, Michael J Ryan, Alan Ritter, and Wei Xu

Reference 29

Resolution
verified exact
doi, observed 2026-05-23T03:55:21.704546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:cf992f6695ed2370a2b48837b55a4da1975c0547413fadb2259db7663f06aa0f

Observation 75cd83ad-17da-4e4d-8de7-cc20d7be1c16 · outbound

This paper cites you gotta be a doctor, lin.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks you gotta be a doctor, lin

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.365617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:bdc03d0c36e03e9625b34e406af2803c1f928e876c50d9efa1d215d254589494

Observation 5b779e7e-bcca-4e19-b052-bca647e56469 · outbound

This paper cites Steven Rogulsky, Nicholas Popovic, and Michael Färber.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Steven Rogulsky, Nicholas Popovic, and Michael Färber

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.718989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:93ab658101bb38b8d7cdc1bd62bf81c0284bb94f271b88c27709ade6d28dacde

Observation 7555ef31-e65f-4d3e-873c-20ff10642be6 · outbound

This paper cites The bias amplification paradox in text-to-image generation.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks The bias amplification paradox in text-to-image generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.365768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:51fa6b7855fcca17c271a455459f6ecfc3d72598be7e0dcdae20b8d5dd4c6e81

Observation 6071184a-91f5-4464-a28c-871ae81e33aa · outbound

This paper cites Detection and measurement of syntactic templates in generated text.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Detection and measurement of syntactic templates in generated text

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.374304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:f77c5a03c93f901c37cd9ea921aed2150350614545e00d791814c9650958f8f8

Observation c85d53b5-131c-4da6-9f8d-b98ced7b3352 · outbound

This paper cites LLM Theory of Mind and Alignment: Opportunities and Risks.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks LLM Theory of Mind and Alignment: Opportunities and Risks

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.914832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:c4866602a4810896e6d6426d23b7677c521d4c8e706ff6e6a4519d063e1ad93a

Observation c7ca237a-e0d2-4117-a607-fc3f3d0c8afa · outbound

This paper cites Will we run out of data? Limits of LLM scaling based on human-generated data.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Will we run out of data? Limits of LLM scaling based on human-generated data

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T03:55:21.898749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:d0e8e4767cb6385322731a5592ca7709499dc35445704ff33195e9470e5466b2

Observation e128e0e0-77ba-49c6-b899-176f93d44837 · outbound

This paper cites Unraveling downstream gender bias from large language models: A study on AI educational writing assistance.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Unraveling downstream gender bias from large language models: A study on AI educational writing assistance

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.346583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:9af4cb2de6134880ee2db842e925b485f0b3ea306f8f95da02182cb4aa0dc9e0

Observation a903054b-3a5a-41cf-8385-4cfbe6bb697e · outbound

This paper cites kelly is a warm person, joseph is a role model.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks kelly is a warm person, joseph is a role model

Reference 37

Resolution
verified exact
doi, observed 2026-05-23T03:55:21.709591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:36d4f7444b920e050a2ca8743484ceb5b9a1e5ca698c77d65157c25dac360e98

Observation 280703ef-5dd4-46e8-bf82-c5857f9d9d99 · outbound

This paper cites doi: 10.18653/v1/2023.findings-emnlp.243.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks doi: 10.18653/v1/2023.findings-emnlp.243

Reference 38

Resolution
verified exact
doi, observed 2026-05-23T03:55:21.707100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:65cbd49c709eb7f3865b4bd131aae3bc4cf79eeee5fb39f3b25c2c2fb630f2c5

Observation 3f3165c5-8934-48e8-baef-aba5c0e41a84 · outbound

This paper cites Will the Inclusion of Generated Data Amplify Bias Across Generations in Future Image Classification Models?.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Will the Inclusion of Generated Data Amplify Bias Across Generations in Future Image Classification Models?

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.904083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:64071e744f60f99c9bebb2c6f1746747ab1bc9403f6470de02c3f2cb74d64540

Observation c21a2c00-e6bd-4078-82a0-274f7877c901 · outbound

This paper cites How to Synthesize Text Data without Model Collapse?.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks How to Synthesize Text Data without Model Collapse?

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:55:21.915792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:f93fa03c04a544ddbe976ccbee7fe296c7f4911e176a8f82215f162284bef3ff

Observation 5932ab21-1744-415e-a359-fb6e04749a56 · outbound

This paper cites You are a person influenced by Spanish culture responding to the following question.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks You are a person influenced by Spanish culture responding to the following question

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.340533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:7ce285a9d36e5b96fba7a1960d2748471eebdcab7f4cc3ac1444370da610b888

Observation d2a107ed-7d92-4194-9017-f6caffdf2de3 · outbound

This paper cites Generate a story about a character named [NAME].

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks Generate a story about a character named [NAME]

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.355405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:6e7e3231458b0746510a51650483cb0cce9826420f371a0e2c12825116580378

Observation 09e64b81-178f-4662-8029-34aae8802a90 · outbound

This paper cites We extract these adjectives from the generated stories, analyzing the frequency of adjectives used to describe the characters.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks We extract these adjectives from the generated stories, analyzing the frequency of adjectives used to describe the characters

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.386123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:d089c3f4fe22d7a4598bb2229ddcc36d67563ebf0c1be33d35aa9d9d641089b4

Observation 08175fe2-5dc9-49f4-9168-9ccedb52751d · outbound

This paper cites To ensure gender balance, we sample 600 examples for each profession, with an equal split between male and female data.

Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks To ensure gender balance, we sample 600 examples for each profession, with an equal split between male and female data

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:55:22.362281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-23T03:53:00.509414Z digest=sha256:a1cc13e542d1816c0590c5b40767dc422b3cb9b4f52a306cd2c54deea07cd450

Pith citing papers

Observation 9ef16029-748a-49a0-bd4a-b239353d1850 · inbound

Inertia in Moral and Value Judgments of Large Language Models cites this paper.

Inertia in Moral and Value Judgments of Large Language Models Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-23T22:05:50.183243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-23T22:05:17.165589Z digest=sha256:b13f93259090f2d491eb4c504e7a1bcdb69062e5f8e9341da7a57df43a310d6f

Observation 3cb4a72c-281a-4e42-a32b-11a22413fc77 · inbound

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs cites this paper.

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-19T11:07:15.444706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-19T11:03:59.222849Z digest=sha256:4f26e9941610040c2dee97de36d3a1eec425081da315a33029817c684b9f689b

Observation 964e8a14-58d4-4f7d-a979-1579318fc093 · inbound

Safe for Whom? Rethinking How We Evaluate the Safety of LLMs for Real Users cites this paper.

Safe for Whom? Rethinking How We Evaluate the Safety of LLMs for Real Users Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-16T23:23:39.906906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-16T23:22:42.431997Z digest=sha256:042627ec06ab78d0adb1a587849cb576b1d67bb85432e45df1b4deff7708ec13

Observation bc8ea647-1a8b-4bc8-b3a0-eaa065ae412c · inbound

Every Act Has Its Price: Compressed Moral Composition in Frontier LLMs cites this paper.

Every Act Has Its Price: Compressed Moral Composition in Frontier LLMs Understanding and Mitigating Bias Inheritance in LLM-based Data Augmentation on Downstream Tasks

Reference 145

Resolution
verified exact
local_arxiv, observed 2026-06-28T22:52:45.524323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-06-28T22:49:03.124658Z digest=sha256:a98b6d48f624811fba9a398d91d6a9fe497cc0dedc3f8b8b20f9be85eca9962d