Pith. sign in

Paper Citation Record · LEDGER

ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2303.15056.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.15056 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:59:22.843126Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

71
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 86697ead-1457-4f21-851c-7e73d68b855e · inbound

Visual Instruction Tuning cites this paper.

Visual Instruction Tuning ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:22:03.749556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T08:22:03.403362Z digest=sha256:0dfc5fe7c85730a187b2646d2deec5da9b37b4c86cc0af3f0e2ac325f2f7e127

Observation 39ff66de-2c9a-48ef-91c3-6229f6954662 · inbound

Enhancing Chat Language Models by Scaling High-quality Instructional Conversations cites this paper.

Enhancing Chat Language Models by Scaling High-quality Instructional Conversations ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 240

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:25:08.120802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T17:25:07.730933Z digest=sha256:103f16e99f3352631a525edf4a4de4faad2f2a74bdd6121822c95ae8623a2a07

Observation 882919d0-3361-467d-8025-cc4e9587a3b4 · inbound

Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena cites this paper.

Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T18:52:59.071717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:52:59.033645Z digest=sha256:bd52d90f9fdca16df1a794a86e9c94efc40d2ee1e9ca4261931c713b8e45db0f

Observation 2b3ed6db-cbb8-4c8b-871d-9110df7a8b91 · inbound

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning cites this paper.

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:34:56.917191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T17:34:56.836034Z digest=sha256:e882af25daa95d0232265494e2339ee6159405e3f004ffff6905917cf026e56e

Observation 14cd40a8-3af4-452f-8b6b-c0d0e0a491e1 · inbound

RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback cites this paper.

RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:32:28.019478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T21:32:27.806494Z digest=sha256:582890b8f584a59fc004136ac68374ec2463ea540ee916b93015f68f6e9669c1

Observation df608a20-961d-4a1a-a590-0248b4ba1c00 · inbound

Chain-of-Verification Reduces Hallucination in Large Language Models cites this paper.

Chain-of-Verification Reduces Hallucination in Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 140

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:06:50.336442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T01:06:49.811982Z digest=sha256:bb3c24e8b9f2ce9ec2a61eb38d0e0badd10044852464620122c36d40b2cd19a6

Observation 6692d942-ac0f-47e3-bdc8-a178f2e22c9d · inbound

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models cites this paper.

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 287

Resolution
verified exact
arxiv_id, observed 2026-05-18T06:38:37.160208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T06:38:36.517935Z digest=sha256:11d92efe899a6927c6b80f36751e687b6d0d0e3778407b11054d3747797e9c96

Observation e74d8083-0eb7-4527-8b75-5757ca1d7b26 · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-17T12:04:10.577423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:9e60c97d6bcbbdae12fb1e894ce9d6f7058b69f676ec57ee77224fc4544b51ff

Observation c2d74b0e-2cfc-490e-b576-61cb250bd370 · inbound

Simple Prompt Injection Attacks Can Leak Personal Data Observed by LLM Agents During Task Execution cites this paper.

Simple Prompt Injection Attacks Can Leak Personal Data Observed by LLM Agents During Task Execution ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:22.843126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:22.843126Z digest=sha256:2d7d37c9e663e57cf746bf9d0e36f204b50174c5958060468b753d7793462aa5

Observation 5e8af0c7-399c-4a34-a539-655d3c378448 · inbound

Prompt Candidates, then Distill: A Teacher-Student Framework for LLM-driven Data Annotation cites this paper.

Prompt Candidates, then Distill: A Teacher-Student Framework for LLM-driven Data Annotation ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:01:31.478037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:01:31.478037Z digest=sha256:6d033a9bcdb72dc1a2a7bb3c4c0aa8c4efa6a2f6e60fe215850f6ce7aa052197

Observation 05ffb229-30b1-4835-97db-4bc9ddd22642 · inbound

Towards Efficient and Effective Alignment of Large Language Models cites this paper.

Towards Efficient and Effective Alignment of Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T04:55:39.571857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:55:39.571857Z digest=sha256:700678853a9f731828cd26ba107be37dbad115094f9ac3ecb87584c56486c582

Observation 4de7c867-4b92-4e56-8f9f-f55197ffec7d · inbound

Using AI to replicate human experimental results: a motion study cites this paper.

Using AI to replicate human experimental results: a motion study ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T17:37:30.227785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:37:30.227785Z digest=sha256:a53f75e9d07f44017f87aaab0f1f29f983ae9f8dff5aed45b72fa18e030b8a33

Observation c2b645b4-8125-4f13-aa1f-d0ae38348701 · inbound

Just Put a Human in the Loop? Investigating LLM-Assisted Annotation for Subjective Tasks cites this paper.

Just Put a Human in the Loop? Investigating LLM-Assisted Annotation for Subjective Tasks ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:04.188426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:04.188426Z digest=sha256:62d83d9ef3192ac826b44cd043aacd49878d396006f3e72c1a0d715766b1cc2e

Observation 6cf5b94e-3daf-44dc-8c0d-f10f6e1250aa · inbound

Guidelines for Empirical Studies in Software Engineering involving Large Language Models cites this paper.

Guidelines for Empirical Studies in Software Engineering involving Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:02:52.213502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T22:02:36.307598Z digest=sha256:80baa1840bb9c56b6745ea57d6fe47855d67e53f5acd6aec898211d024bc1cfb

Observation 5887031b-d97b-4b9b-b220-606a97bca22b · inbound

Guidelines for Empirical Studies in Software Engineering involving Large Language Models cites this paper.

Guidelines for Empirical Studies in Software Engineering involving Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:20:31.709513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T08:18:18.448122Z digest=sha256:dfe761476319b260cd6f1300c8ad615ba72a77369804b9af94c4aafdb9f6957f

Observation 7167ba93-7cdb-48c7-be19-2973bd35ae55 · inbound

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection cites this paper.

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:40:27.025055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T13:38:02.450128Z digest=sha256:4fd2572fdf731dedb23267f7d856ec2c6921ee436a80cbf504bfa5b39221ee45

Observation cfc0df83-aab3-43b8-861d-d3db85060f38 · inbound

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection cites this paper.

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T20:35:52.598345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:35:52.598345Z digest=sha256:872cad63148391ec7ea596bf4d2919049f8261cb44215bcd88d46a57a60f6d3e

Observation f26bc985-4ba5-4c0a-8145-b2e5de255ef6 · inbound

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages cites this paper.

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 150

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:38:21.738599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T14:33:36.100966Z digest=sha256:8eaa04aaaa306f212ba74aeeb9c1f90752a7700a63dbb3f4c533304121e65c34

Observation 2a0ad235-2a2f-4444-bec6-1eaa18c479a0 · inbound

Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization cites this paper.

Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:13:27.570712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T13:05:15.701729Z digest=sha256:8340c3f17e3142d50f3aa0094abb285fa38c1aa4fb646f53f94208f0e72dcf8e

Observation 4f8d98ad-81b0-416e-be19-0a571b5ebf90 · inbound

Structure Before Collapse: Transient semantic geometry in next-token prediction cites this paper.

Structure Before Collapse: Transient semantic geometry in next-token prediction ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.131669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T05:14:07.208255Z digest=sha256:d0b0b5eeb50bf334dd2b1acc3fdd719015c20bc41e692210e7938bef3e854589

Observation b2c27081-31b0-4912-b352-d4f76c96a608 · inbound

CORTEX: High-Quality Cross-Domain Organization of Web-Scale Corpora through Ontological Corpus Graph cites this paper.

CORTEX: High-Quality Cross-Domain Organization of Web-Scale Corpora through Ontological Corpus Graph ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:14:18.532747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-30T06:04:40.684934Z digest=sha256:bac88c356948f9078b2cfb8861d6b817f41a80dbdb81f5fb26d34a53f26ab7b2

Observation 1c8a3931-ef25-47f7-8f0d-3cb8305b55e4 · inbound

SERUM: State Extraction and Refinement for User Modeling cites this paper.

SERUM: State Extraction and Refinement for User Modeling ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T12:13:37.936172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:13:37.936172Z digest=sha256:855e80e84e64d10c7d10050308f2c90d0b9b706a907f405170186a0939a4a704