Pith. sign in

Paper Citation Record · LEDGER

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis

As of 10 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 1 inbound Pith citation observation for arXiv:2508.14764.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.14764 v2

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T18:22:38.879800Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:27:21.014118Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T05:30:23.456663Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact1
  • verified fuzzy20
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
pith, observed 2026-08-10T05:30:23.456663Z

Outbound references

Observation 0cd3559e-7c47-4d36-a9a4-66e5ed25346b · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:36.079033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:36.079033Z digest=sha256:1e4e1389795f9db3fd3c02903e52fa257f6c44300d4436c5200631dba56d9785

Observation fd02565e-4b8c-45ba-b8e0-e410f1fb4001 · outbound

This paper cites To find agreement between human raters and LLMs, we FIG.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis To find agreement between human raters and LLMs, we FIG

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.738538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:35.961318Z digest=sha256:7f2074035f27b49ee20e74d9eb09830c81b467f0a721e83f28fe83954bbf2e81

Observation 860c4662-d38d-4cfc-b09c-f9601c390cc1 · outbound

This paper cites Honey, G.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Honey, G

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.709235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.136603Z digest=sha256:406567a2791bcdbcc28db2e732f4724cedc5ae182d30281ef923ca65f138b753

Observation 14fd7bd6-97ad-4816-9e25-04ad74c54f50 · outbound

This paper cites Fischer, C.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Fischer, C

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.689283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.194604Z digest=sha256:081eef945d22755f310f09a7320d7cd4d1d24f142424dfbed8ae266dc9b8e74f

Observation 168f0429-891a-4b22-9c06-4a66eef1d926 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.668790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.281607Z digest=sha256:9670c333c72615ceca51a76f64516dcf8789db5098ddf441c536874691a2037a

Observation 040f22c4-a813-4b66-9aae-1b20385d975f · outbound

This paper cites Dalal, A.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Dalal, A

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.651125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.372775Z digest=sha256:3cfc7cb176ad7d5d314d289d02be979d4dfa3d1cfb31f9ee8cc229b79e2e3f27

Observation ca62b77a-8bb9-4bdc-b056-1eed9bb0fd1d · outbound

This paper cites Slavit, E.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Slavit, E

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.632136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.454909Z digest=sha256:3fd31bcbd49b7fd4e4cfe121d8d02d092d32684a68bb850a29a02974f42524b2

Observation 7394c228-c935-4989-b59f-9bf1af8b8b1e · outbound

This paper cites Slavit, E.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Slavit, E

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.611080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.533925Z digest=sha256:01067c44168c209dc9f870c9d2c1446a82c4d85d52ea8b59bb4665f4a6f278af

Observation dc0323ed-ec0c-4345-b45e-e458d747f176 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.593548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.616010Z digest=sha256:d4869d99fda7b4c8185445952ee712628af0ff5957d38e887b0345b7c7847009

Observation 6efca26d-85ee-47a4-839f-1f001dd2acc7 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.576100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.680585Z digest=sha256:ea8c484fe0da67266299dedc89a2ef2abb541f1e8b6b442c9f11f26e0765dc8d

Observation 4e06193e-d724-4040-95eb-28b9667b0c85 · outbound

This paper cites Talanquer and J.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Talanquer and J

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.559225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.742077Z digest=sha256:074e6a307a145340305454d03cf01b6d636fcb4b14bfa9fb3cabcf40c8fe4d38

Observation cbcd871d-2e95-4ab9-9183-1c23dcaeefc9 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.542863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.810808Z digest=sha256:2efd0ed0651ee875760a10544e85ec1923a31b0b51b67b38ee0438fb7bcde71b

Observation 1c33057e-cc00-46b8-86bb-3de6c7597b72 · outbound

This paper cites Wilkinson, A.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Wilkinson, A

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.527850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.872197Z digest=sha256:588e2558c43aa82eb7c09469b1ac478627a7d1c6ce4ddd562742b3dcf4252544

Observation 28256f66-6d61-45f8-b4c1-0cb65c0a880e · outbound

This paper cites Etkina and G.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Etkina and G

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.509389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:36.963718Z digest=sha256:bb163b5a6aa57dbe390746cfc1b77a2e595b8faa558c4d8a30f6b758524d94ca

Observation e2cbc877-a394-463a-a7a0-1f0ab554990a · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.493736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:37.004072Z digest=sha256:034b1a00e0671a7ce9d387047191869dd53229d1ef09d9dd9762fae1616fea88

Observation 39f0ff9c-2fb9-4455-adcd-75612cd3dd4f · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.479473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:37.042177Z digest=sha256:5605430308f9c428cee2d37569bdf136d5d98b582a10666489850628c064b0ad

Observation 579f2ea1-48c7-4a39-94d9-4460f43e44eb · outbound

This paper cites Erickson et al., Qualitative methods in research on teaching (Institute for Research on Teaching East Lansing, MI, 1985).

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Erickson et al., Qualitative methods in research on teaching (Institute for Research on Teaching East Lansing, MI, 1985)

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.465025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:37.101955Z digest=sha256:7634357bea6db05c28a186db030bb81a9143f07c92ac7978a22fd52b02ce854d

Observation d602f4a1-0dfc-4d58-a647-1fbbb5566f86 · outbound

This paper cites Braun and V.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Braun and V

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.450146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:37.267943Z digest=sha256:dcb8d604c95233d44d36e8d9fb1a160135735457aec2b896cf645c6d62d319b1

Observation 38c3cbe9-82a8-427b-803b-ec5ba860b07e · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.435180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:37.372269Z digest=sha256:b0bb0999e74d5f81b661876bb2daaad1d957f16b1e75746335f722c8d5ccd8a0

Observation c180b053-addb-4f86-a0c2-947d47547ee9 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.420040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:37.488537Z digest=sha256:15a49e08c6891468d725d627d090cccfccc8536d2d7f72356985f61d6860fb86

Observation aa043082-b914-4871-881c-ad720b52b1dd · outbound

This paper cites Saldana, The coding manual for qualitative researchers , V ol.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Saldana, The coding manual for qualitative researchers , V ol

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.403882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:37.586251Z digest=sha256:e9a25cbf2f75901a78e39b399df6d8ed4e98a9a9af879dcd37433cdc38ec5d5e

Observation 673e88e1-8bc9-49e9-9677-20bf4ea635f6 · outbound

This paper cites Houghton, K.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Houghton, K

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.384281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:37.677001Z digest=sha256:a2a077c7e0809a68765e85cb8c4e692992c47690cd8d02abc8b30cea19e1ff91

Observation fee14e93-29f6-4660-b040-9a7e7dcb63d3 · outbound

This paper cites Jackson and P.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Jackson and P

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.368611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:37.773548Z digest=sha256:b51173d0ebbb6a26cb64d8d3a767e0967bded8c7ad0d301fe802aa09b34d99aa

Observation 4fb51306-6fbd-400d-84aa-9c1814f7e45c · outbound

This paper cites Uysal and N.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Uysal and N

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.352912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:37.866609Z digest=sha256:5207a11ab849e673ecdfa3dc635ac4aef8045268046c3c4ac05dc975a2351280

Observation d045538f-c993-46f5-9e2c-424cdeba319a · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.336619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.001074Z digest=sha256:54844132710dfff08ca81b681fbe87bd843631783c7a6ba24b5ebc380cabac8a

Observation c68b5e03-f09f-4cd2-ae36-987cacdb21ba · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.321165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.135492Z digest=sha256:7b29994a57d60d07c3a566b9f315e79e2819fd17d4a38205e9685d35903ebbd3

Observation 412135b2-2063-4713-9220-7a325364a447 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.228019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.228019Z digest=sha256:11fbcaf9d66de2616d970aa31946b12d590b07c9abe4c1ffe3a72bd4a291ac24

Observation 55eb4a2e-a401-48d8-a9a0-5af891485742 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.295854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.319873Z digest=sha256:49a32442e11d4d5b66f25960863f361871a33794668af3d12f56cc162bb7128f

Observation acca8eee-682e-423e-bf31-19db028dbc75 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.280200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.482804Z digest=sha256:d95c252d63d7eb0a8c579344ff7e7c165360c71e85a55712116d51d9f66b2e86

Observation 13eb7344-ee42-4540-a345-b76a39ec6199 · outbound

This paper cites Exploring the Efficacy of ChatGPT in Analyzing Student Teamwork Feedback with an Existing Taxonomy.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Exploring the Efficacy of ChatGPT in Analyzing Student Teamwork Feedback with an Existing Taxonomy

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.608842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.608842Z digest=sha256:366dd1f90cd9585ff73a910a5a7ec2a5dd20b6971e6f2ccfc2de7e947d2f8acd

Observation ee7926f1-8ba1-497c-8505-f09b31213a0e · outbound

This paper cites Hitch, Artificial intelligence augmented qualitative analysis: the way of the future?, Qualitative Health Research 34, 595 (2024).

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Hitch, Artificial intelligence augmented qualitative analysis: the way of the future?, Qualitative Health Research 34, 595 (2024)

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.265877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.688426Z digest=sha256:cb67ea60545d4a8bd9492bb311e78529af04d0f79ccdf42c4bf1ca53958aef61

Observation 5a7d55a1-ff41-457d-b963-04c8feb8903e · outbound

This paper cites Tabone and J.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Tabone and J

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.252102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.732854Z digest=sha256:f807fa6867819a7613bf7f2c2b5dfb86aca05ee71d1c4ca09f99143853de8b02

Observation a6aafd8f-42f5-4d83-9d5d-123f3d742975 · outbound

This paper cites Redefining Qualitative Analysis in the AI Era: Utilizing ChatGPT for Efficient Thematic Analysis.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Redefining Qualitative Analysis in the AI Era: Utilizing ChatGPT for Efficient Thematic Analysis

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.788741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.788741Z digest=sha256:30d0eb208f56059d15fd18d63a1ba0dffe440599bd538b34795d78a77724bda9

Observation 1bf2da57-7cf4-485b-b228-bf35e0831d77 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.238145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.801848Z digest=sha256:08d48afff97b38e57cd183079c612acd6f245bf732dfcf7cb681cce438f68f59

Observation a46cf742-9688-4c6d-b747-32993b68e106 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.223869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.805840Z digest=sha256:e7ff6181a3d21428118d72985e4afaac6cee6e43e79d14d592a08df95a3a8843

Observation e63cbe10-d1dd-484a-ab06-47408f9f0435 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.209920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.809762Z digest=sha256:f27fb7603814fac00f26867e00618158b241d325988131553eb90c26fbf16876

Observation d01810d8-3b3e-4b67-bb33-6dc1326e7abd · outbound

This paper cites GPT-4 Technical Report.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis GPT-4 Technical Report

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.813821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.813821Z digest=sha256:cac1244cb6b6a93d320d45d2837e48a2bac812098903fb36d6f6f70171f5fc65

Observation c9889a5f-d283-4e59-9115-b98a93a5aefe · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.195389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.818785Z digest=sha256:260b438b492b32433597331b81bf6a451a07dd20aff317399370f3b2c190ac43

Observation eb5bfd5a-3f86-48fa-9eb3-4369bd977176 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.179312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.823114Z digest=sha256:582e0e76654476d76bbeb396c7ba759f438893bac589ea985008345d5c55cace

Observation 061a9329-e136-4d17-b7a9-30c276c9a24d · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.164102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.828118Z digest=sha256:c9ecde6ce98eb625b819a4e31ebaffeff4573e030a8048c9c72ecb752fe29b64

Observation 5e445e4a-ce0b-49ea-86d2-77a01b2a7879 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.149368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.832151Z digest=sha256:b23526d191c8fcca688ad53b337551e27401bac3fd2ab7cdfc2597212e0ff66d

Observation 93731589-e9a2-4cfe-aa43-802585d2a1be · outbound

This paper cites Applying a STEM Ways of Thinking Framework for Student-generated Engineering Design-based Physics Problems.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Applying a STEM Ways of Thinking Framework for Student-generated Engineering Design-based Physics Problems

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-05T18:22:38.957039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.836255Z digest=sha256:cc96772a1ba9437cb2bda288cade683df617b3dd81e0689a99a6fa974e613ef0

Observation 750d32c8-dbc0-486b-a058-395f0527c9ee · outbound

This paper cites Bijker, S.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Bijker, S

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.133612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.840450Z digest=sha256:515f4dd748a0055c98fa8edd88bab08dcafc6cd61dc6e9131b9317b2fcfb891c

Observation 4264472f-d91f-45ba-ad8d-217bce3634ca · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.116170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.845001Z digest=sha256:f167031c89c715b0138cd285efd8cb13d7900fe40c3dcc4e41ec8fc22c6233d1

Observation f228e63f-10c1-4881-89a2-33e3e30c84e5 · outbound

This paper cites Mizumoto and M.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Mizumoto and M

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.099592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.849419Z digest=sha256:c935adda63f51d81c27a395aaca7a0f2517cf98a58d4374e0fcba88816c98224

Observation 9ab8cb29-450f-4460-9d6e-759f20f9b023 · outbound

This paper cites Unleashing the potential of prompt engineering for large language models.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unleashing the potential of prompt engineering for large language models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.853411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.853411Z digest=sha256:4a75b009922613ac4717e7ec27daa55ee9b0a324be0b9b44f20d5b845d2a87ce

Observation 916401d2-7ab3-4ccd-9f9d-c2fccf468c75 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.083092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.857964Z digest=sha256:232c7205b8e6941c636969223ad5dbf9726381d208f57fb32f610132c7479bf8

Observation c6c5f65d-8733-46b0-b53e-b45f25d46936 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.067668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.862085Z digest=sha256:3d14e643ced422455ee0a1b27e7a14df98be44182cce6220dcdc6ea1b9f7ca32

Observation 671f9955-5cd0-4760-8162-6d91ee302200 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.051844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.866440Z digest=sha256:7892829c9d93bcd74f83232829966c7013a507fc9998789488f81df3328a3c8a

Observation 8ba0ae9c-e47d-46b6-b589-8a8684f917d2 · outbound

This paper cites an unresolved cited work.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:22:39.036749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.870936Z digest=sha256:0a907a410b9a10c33bed5862881da13accb0ee068853e8f197fca0cd0b236902

Observation f45be533-4e8b-4905-b790-667213e86783 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T18:22:38.875155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:22:38.875155Z digest=sha256:a1348f8b2df3284391f9f6eefeedc0b4c4b981f3005f4c10361a85fb5b1752d9

Observation 8c9748e9-2495-4725-86ad-04e6627ce478 · outbound

This paper cites Tschisgale, P.

Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis Tschisgale, P

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:22:39.022025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T18:22:38.879800Z digest=sha256:cf6ea6382a9bb05fd0779ddb0b3afef619dbbba204ef49cbf4889008c1deb93c

Pith citing papers

Observation a725a0bc-af27-40f1-a102-f2567f27ae3c · inbound

Me and My Bot: What Users Talk About in AI Companion Communities on Reddit cites this paper.

Me and My Bot: What Users Talk About in AI Companion Communities on Reddit Investigation of the Inter-Rater Reliability between Large Language Models and Human Raters in Qualitative Analysis

Reference 2026

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T00:27:21.045369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T00:27:21.014118Z digest=sha256:b30e8ab1ba853a179a73674cf86604fa0e17ebeaa8cc2de1321f57197b0a62f1