Pith. sign in

Paper Citation Record · LEDGER

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis

As of 8 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 1 inbound Pith citation observation for arXiv:2508.14764.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.14764 v2

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T18:22:38.879800Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:27:21.014118Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T06:16:20.379382Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact1
  • verified fuzzy20
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0cd3559e-7c47-4d36-a9a4-66e5ed25346b · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:36.079033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:36.079033Z digest=sha256:9abce0cc0acbe7b5406787782c2d3f6a6a83590bb216c369d9ce5d61d91de937

Observation fd02565e-4b8c-45ba-b8e0-e410f1fb4001 · outbound

This paper cites To find agreement between human raters and LLMs, we FIG.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis To find agreement between human raters and LLMs, we FIG

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.738538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:35.961318Z digest=sha256:74a8d716a7315ad92eaf60a92ddc040bcba21119852eaa1faf6d6eac76831a7c

Observation 860c4662-d38d-4cfc-b09c-f9601c390cc1 · outbound

This paper cites Honey, G.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Honey, G

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.709235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.136603Z digest=sha256:105bb3654ac1458516508994247de9c6860eae032987979809622fcb43e4b0da

Observation 14fd7bd6-97ad-4816-9e25-04ad74c54f50 · outbound

This paper cites Fischer, C.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Fischer, C

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.689283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.194604Z digest=sha256:59fbeeb6a415411a173ffbfec31a54b3ab7dd3e9401467d098cd3bddaca14abb

Observation 168f0429-891a-4b22-9c06-4a66eef1d926 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.668790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.281607Z digest=sha256:2481d9435f30a629214f9a2a0a9ec7ea66f2a7b748337d8237d613c8de034283

Observation 040f22c4-a813-4b66-9aae-1b20385d975f · outbound

This paper cites Dalal, A.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Dalal, A

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.651125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.372775Z digest=sha256:938c8a93578ccd026772780a508308413891f0664a86d3437ef6b88dd612c31e

Observation ca62b77a-8bb9-4bdc-b056-1eed9bb0fd1d · outbound

This paper cites Slavit, E.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Slavit, E

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.632136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.454909Z digest=sha256:a90dd12ac5fcf0955ff2af891fb5989b16d0920b46ab92a56092e4745368551b

Observation 7394c228-c935-4989-b59f-9bf1af8b8b1e · outbound

This paper cites Slavit, E.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Slavit, E

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.611080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.533925Z digest=sha256:dd3deef45e3c0b53ae26f6811d4b2643136ec69c01d826a88409ea0b6f50e6e4

Observation dc0323ed-ec0c-4345-b45e-e458d747f176 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.593548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.616010Z digest=sha256:8dfea177157aadbc10a076d09c302057b2bb9009d0b44740f8f317471d107716

Observation 6efca26d-85ee-47a4-839f-1f001dd2acc7 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.576100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.680585Z digest=sha256:28673e5bf2c68a3fe07cc8a0d2b5f27dbfc8602dd886797813f690015553465d

Observation 4e06193e-d724-4040-95eb-28b9667b0c85 · outbound

This paper cites Talanquer and J.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Talanquer and J

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.559225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.742077Z digest=sha256:9393d126aae765ca46d5b35276969a0df1a4e07e138545ba4c6354494d63612b

Observation cbcd871d-2e95-4ab9-9183-1c23dcaeefc9 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.542863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.810808Z digest=sha256:3b218ee045ad41db6f881d545f2ad6a5a2903c3c63ff6719b92e7be74235f9ab

Observation 1c33057e-cc00-46b8-86bb-3de6c7597b72 · outbound

This paper cites Wilkinson, A.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Wilkinson, A

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.527850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.872197Z digest=sha256:cabe8bfb5e2438b34703c8bba8695a236ad83fe5353685e97407d5fa6d9a431c

Observation 28256f66-6d61-45f8-b4c1-0cb65c0a880e · outbound

This paper cites Etkina and G.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Etkina and G

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.509389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:36.963718Z digest=sha256:1caf102d75417b153da5308179ac67bf9ad637ebd2b7525d067c68443b566b8d

Observation e2cbc877-a394-463a-a7a0-1f0ab554990a · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.493736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:37.004072Z digest=sha256:4e991618be3fe32356cddceda03998a6a9b21c7788671007a25e16627a981c22

Observation 39f0ff9c-2fb9-4455-adcd-75612cd3dd4f · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.479473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:37.042177Z digest=sha256:da1bb4d252172051e7903e8e437bd66877ab921e43e91077fb239cbcd198dee5

Observation 579f2ea1-48c7-4a39-94d9-4460f43e44eb · outbound

This paper cites Erickson et al., Qualitative methods in research on teaching (Institute for Research on Teaching East Lansing, MI, 1985).

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Erickson et al., Qualitative methods in research on teaching (Institute for Research on Teaching East Lansing, MI, 1985)

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.465025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:37.101955Z digest=sha256:cd0d1c4e2614305104f33eea51ae3bfb436ee9370093ea3136d27603d2f3886f

Observation d602f4a1-0dfc-4d58-a647-1fbbb5566f86 · outbound

This paper cites Braun and V.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Braun and V

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.450146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:37.267943Z digest=sha256:b35b59ce31d92cd134c90ea44aa41edea3bbd2379aa265f28b92997ea67eda79

Observation 38c3cbe9-82a8-427b-803b-ec5ba860b07e · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.435180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:37.372269Z digest=sha256:5b87c1e5afa02534e93bb4388115a4d1951dab3d2262747305f82ad64ddc3fba

Observation c180b053-addb-4f86-a0c2-947d47547ee9 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.420040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:37.488537Z digest=sha256:9a92be48eb64c055569b68b0ff95956f8fff03b0df93265271b0bc233ae798c6

Observation aa043082-b914-4871-881c-ad720b52b1dd · outbound

This paper cites Saldana, The coding manual for qualitative researchers , V ol.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Saldana, The coding manual for qualitative researchers , V ol

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.403882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:37.586251Z digest=sha256:0a5ef75ddf25d86ec957526b1e5bc8e4f01a3eae75668d656477d32c2bfc59b3

Observation 673e88e1-8bc9-49e9-9677-20bf4ea635f6 · outbound

This paper cites Houghton, K.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Houghton, K

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.384281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:37.677001Z digest=sha256:bb8be343cba5865b69e02413596700ed102938719c9ff4580d17b4cb200a0ee5

Observation fee14e93-29f6-4660-b040-9a7e7dcb63d3 · outbound

This paper cites Jackson and P.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Jackson and P

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.368611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:37.773548Z digest=sha256:bf8f0c1341e755b55fd875a500f75c59af6b08b520df835cd663fc64d5809d98

Observation 4fb51306-6fbd-400d-84aa-9c1814f7e45c · outbound

This paper cites Uysal and N.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Uysal and N

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.352912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:37.866609Z digest=sha256:ea991e54b024eeec6e94ff1a3d2dc3d46994e04f2421deac9318efa94cd3ff31

Observation d045538f-c993-46f5-9e2c-424cdeba319a · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.336619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.001074Z digest=sha256:52b2b6db593f1e627190f688451a464eccca27435cf50de8260154b2bfeff87f

Observation c68b5e03-f09f-4cd2-ae36-987cacdb21ba · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.321165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.135492Z digest=sha256:0a446ad6d982f7a005a193d4eab4c95e7bbb264351b55e1c8cff85cd4e20a04d

Observation 412135b2-2063-4713-9220-7a325364a447 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.228019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.228019Z digest=sha256:f7a9c7f7b851aa6d9bfaeb3492e9858af69ac69d6c03443eaf3093c70b550673

Observation 55eb4a2e-a401-48d8-a9a0-5af891485742 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.295854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.319873Z digest=sha256:ddd6359099098dde9628ee77c9f16a53f95c4e15ed79e7d4088c8bf2d025c232

Observation acca8eee-682e-423e-bf31-19db028dbc75 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.280200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.482804Z digest=sha256:17a8caa4485645570752687716e2887aeebf16fb4dc19f806abe4eec7e9eee72

Observation 13eb7344-ee42-4540-a345-b76a39ec6199 · outbound

This paper cites Exploring the Efficacy of ChatGPT in Analyzing Student Teamwork Feedback with an Existing Taxonomy.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Exploring the Efficacy of ChatGPT in Analyzing Student Teamwork Feedback with an Existing Taxonomy

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.608842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.608842Z digest=sha256:f62167cda17b69fbfd07ec94d87de6e6d57a62fbe4903f7b0c3cb1e528966834

Observation ee7926f1-8ba1-497c-8505-f09b31213a0e · outbound

This paper cites Hitch, Artificial intelligence augmented qualitative analysis: the way of the future?, Qualitative Health Research 34, 595 (2024).

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Hitch, Artificial intelligence augmented qualitative analysis: the way of the future?, Qualitative Health Research 34, 595 (2024)

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.265877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.688426Z digest=sha256:f46f2ae36c1ed6b2780dace92c0100a34336ad45d27a73e75304d091fa462c28

Observation 5a7d55a1-ff41-457d-b963-04c8feb8903e · outbound

This paper cites Tabone and J.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Tabone and J

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.252102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.732854Z digest=sha256:78bb32f305723c46e138c4477b25ee43a23ed1238f658af650d2e1fd2d5e4c32

Observation a6aafd8f-42f5-4d83-9d5d-123f3d742975 · outbound

This paper cites Redefining Qualitative Analysis in the AI Era: Utilizing ChatGPT for Efficient Thematic Analysis.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Redefining Qualitative Analysis in the AI Era: Utilizing ChatGPT for Efficient Thematic Analysis

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.788741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.788741Z digest=sha256:a2900ace17e7711c15469a3f607eb96b25b4d8cea727712af69fe604d2dc7812

Observation 1bf2da57-7cf4-485b-b228-bf35e0831d77 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.238145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.801848Z digest=sha256:a7b839757885ccd909147ff3f01112c7fc0873a0d405c832f996fa3e1250b779

Observation a46cf742-9688-4c6d-b747-32993b68e106 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.223869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.805840Z digest=sha256:afc0db3ec60cfb70d7c42aa77bf9d6b6e9e950a1a877b39d2e5e73fea8d8fec9

Observation e63cbe10-d1dd-484a-ab06-47408f9f0435 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.209920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.809762Z digest=sha256:0c53a0ca14402ff0367f0fa50c8a0744dcb3ade1cde67be57ffd31fef5c750c6

Observation d01810d8-3b3e-4b67-bb33-6dc1326e7abd · outbound

This paper cites GPT-4 Technical Report.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis GPT-4 Technical Report

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.813821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.813821Z digest=sha256:9eb55883102f3746d0b5c5cbc8d072bd7c42e3271854d13ed286429f2a715f0d

Observation c9889a5f-d283-4e59-9115-b98a93a5aefe · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.195389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.818785Z digest=sha256:d55f27fe9b2ae15af8eaff5b529820578e09cb4ae343ff17abe3f074029b392a

Observation eb5bfd5a-3f86-48fa-9eb3-4369bd977176 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.179312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.823114Z digest=sha256:2e8360cae9ffe58f6dca6470efd90de9d31c1bbfca24803ba0289cd1bb6eb869

Observation 061a9329-e136-4d17-b7a9-30c276c9a24d · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.164102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.828118Z digest=sha256:64b19a1142bd1211b7d51cecbe5577bc4871dfe912307cbb22bc30ab547ad443

Observation 5e445e4a-ce0b-49ea-86d2-77a01b2a7879 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.149368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.832151Z digest=sha256:525aee684374e3d2d03dad0b50935ce9cd847a615d64fb9ec6dc97ba8f9c945d

Observation 93731589-e9a2-4cfe-aa43-802585d2a1be · outbound

This paper cites Applying a STEM Ways of Thinking Framework for Student-generated Engineering Design-based Physics Problems.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Applying a STEM Ways of Thinking Framework for Student-generated Engineering Design-based Physics Problems

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-05T18:22:38.957039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.836255Z digest=sha256:d65220166918121a6658e40a3e8232efd0358521458cc2754c1250b8473ab013

Observation 750d32c8-dbc0-486b-a058-395f0527c9ee · outbound

This paper cites Bijker, S.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Bijker, S

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.133612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.840450Z digest=sha256:1fe5f630d93cc721419fc9f80ea8229f4a5108c0ae06c6db87ad85a763898c42

Observation 4264472f-d91f-45ba-ad8d-217bce3634ca · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.116170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.845001Z digest=sha256:36d5886142f324cf953bf30b6661420ee4be2401c04f23cb45b61e2114bab3ff

Observation f228e63f-10c1-4881-89a2-33e3e30c84e5 · outbound

This paper cites Mizumoto and M.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Mizumoto and M

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.099592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.849419Z digest=sha256:1eecc0e0237c9659881734c11208ba180eb13886066ecac8555ea39f1c55bbaa

Observation 9ab8cb29-450f-4460-9d6e-759f20f9b023 · outbound

This paper cites Unleashing the potential of prompt engineering for large language models.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unleashing the potential of prompt engineering for large language models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.853411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.853411Z digest=sha256:5b3abbce85206ab33b307cd36972a0be50c32f8ee7bc05e6bb55f05268ee7348

Observation 916401d2-7ab3-4ccd-9f9d-c2fccf468c75 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.083092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.857964Z digest=sha256:2c8dfb9903c120335a11b990721fb47bdac2450f5be506df92873abcef816d29

Observation c6c5f65d-8733-46b0-b53e-b45f25d46936 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.067668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.862085Z digest=sha256:3aa2e13355e644eb0b0edfe0cfb271549394e66119dcbf1d88fd65592f8ffe12

Observation 671f9955-5cd0-4760-8162-6d91ee302200 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.051844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.866440Z digest=sha256:2705c141549643e1e7ad1bafab0ea9785a67aea9d2b511ae5e0de789b3eb511d

Observation 8ba0ae9c-e47d-46b6-b589-8a8684f917d2 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.036749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.870936Z digest=sha256:0064c0d9d9a66573f949204ec907cd5008da4e8c57ece22e68507d7a2fd140ec

Observation f45be533-4e8b-4905-b790-667213e86783 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.875155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.875155Z digest=sha256:9696e6f5e11fca4f074a4600ae89f5a6348b6ebb936138f8e753fbb49c56a24e

Observation 8c9748e9-2495-4725-86ad-04e6627ce478 · outbound

This paper cites Tschisgale, P.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Tschisgale, P

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.022025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T18:22:38.879800Z digest=sha256:03a479ed50281579f323ba3eeaaa5232c8176fda306a7ade6f5e8a1d07575226

Pith citing papers

Observation a725a0bc-af27-40f1-a102-f2567f27ae3c · inbound

Me and My Bot: What Users Talk About in AI Companion Communities on Reddit cites this paper.

Me and My Bot: What Users Talk About in AI Companion Communities on Reddit Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis

Reference 2026

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T00:27:21.045369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T00:27:21.014118Z digest=sha256:b22ce64be3813cb0b624022fab957274bd5232cfa09db69d216494df89c69494