Pith. sign in

Paper Citation Record · LEDGER

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension

As of 13 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2412.00314.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.00314 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:35:51.340765Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:24:28.140998Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T16:24:28.418595Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 752027d0-541b-4e23-9dde-d15fe79095bc · outbound

This paper cites Who judges the judge: An empirical study on online judge tests,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Who judges the judge: An empirical study on online judge tests,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:52.039866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.117340Z digest=sha256:f15b67187f07d1927a2b80c41169b7ce445e51326d6f7969d9e67cce2b1003a9

Observation f02cab78-82c1-40ef-93de-817ed722b6ba · outbound

This paper cites Automatic source code evaluation test develop- ment in programming education using grey-box methods,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Automatic source code evaluation test develop- ment in programming education using grey-box methods,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:52.022557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.123217Z digest=sha256:f582388e31a4885c929c85eea746cbbdef8eb60d91eab332f7a3cd5435a8d141

Observation e6d004e8-46e5-4149-a0f6-ebf5bbbd1cc2 · outbound

This paper cites CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.128443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.128443Z digest=sha256:43a85f79f36cab99b39cc96fd240d129e0144e60c29cda5b24c71fec82579e39

Observation f6f2a5a7-d864-428e-b609-d78100806a24 · outbound

This paper cites Coderl: Mastering code generation through pretrained models and deep reinforcement learning,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Coderl: Mastering code generation through pretrained models and deep reinforcement learning,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.133846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.133846Z digest=sha256:c807b8fc14f41b358dc3274001cf70643838b899537b74911dc3c0d3e84ffcf3

Observation deb65a93-2211-412d-927b-f870e39d239f · outbound

This paper cites Large language models meet nl2code: A survey,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Large language models meet nl2code: A survey,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.993817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.139127Z digest=sha256:27120535f7338433af654590ac4b481f56ad3c826658f5ec400173eb8557a9fc

Observation 1cd7d9a5-1049-42db-9821-3db990e28abc · outbound

This paper cites CodeScore: Evaluating Code Generation by Learning Code Execution.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeScore: Evaluating Code Generation by Learning Code Execution

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.144229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.144229Z digest=sha256:7cc0d7038753571e610b5b441dd75d88d37e89bc911832321729f299a9e8515f

Observation 3992f1cc-c07c-4b5e-aa1b-3b96311eb678 · outbound

This paper cites Spoc: Search-based pseudocode to code,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Spoc: Search-based pseudocode to code,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.150080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.150080Z digest=sha256:f299017189eab016e4d5c3e5a36e5a62c0c5fd022479f5592245424fcb57ba24

Observation c0bd294d-bee7-4eb8-a74d-6da231ae2dd3 · outbound

This paper cites Measuring Coding Challenge Competence With APPS.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Measuring Coding Challenge Competence With APPS

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.154868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.154868Z digest=sha256:46249ac2a49c1bad399ca04e12ebea70146211b7bcfafa237f9eb7ec70174bf0

Observation 1a5342d1-b826-470c-ab36-4e52373ada91 · outbound

This paper cites Out of the bleu: how should we assess quality of the code generation models?.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Out of the bleu: how should we assess quality of the code generation models?

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.966553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.160198Z digest=sha256:38a424765d45945f6055d81ad43d11ae1759f4f934392f6297d04fc1a57b8647

Observation 1287646d-da8b-41d5-8c8c-112888733f82 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Bleu: a method for automatic evaluation of machine translation,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.164815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.164815Z digest=sha256:7687b0163778267e8c102aea4c4025bdc3ade7ef819ae7fa82c33b0906eda535

Observation c60260b5-3f25-40ea-a253-c423b5844166 · outbound

This paper cites A package for automatic evaluation of summaries,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension A package for automatic evaluation of summaries,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.938577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.169661Z digest=sha256:ed5879f70a0753a4f9364cb7f54b53de98e3cd68d5381b51ecbc1b5b472dc818

Observation cdd5a12a-ef2c-468e-ac8c-754a3cc599ff · outbound

This paper cites Does bleu score work for code migration?.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Does bleu score work for code migration?

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.921968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.174483Z digest=sha256:05e0efc9d20990529284ec68673f2bae14d4d4419e8e487caa1569be0cd18d2f

Observation f5dc2978-eab1-4b94-8c2f-b83e9be49e99 · outbound

This paper cites CodeBERTScore: Evaluating Code Generation with Pretrained Models of Code.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeBERTScore: Evaluating Code Generation with Pretrained Models of Code

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.179925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.179925Z digest=sha256:7aed7d3a9cc1b22ff5a962d3a33e5c259f10bf01f8eafda8342c91b4e693d95a

Observation 068580ea-0365-4a79-8662-dc320204b5af · outbound

This paper cites CodeBLEU: a Method for Automatic Evaluation of Code Synthesis.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeBLEU: a Method for Automatic Evaluation of Code Synthesis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.185496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.185496Z digest=sha256:1c2f8f32747fbf00c7c85f34e383f44faa56133ce768b887657f6baa7cebf720

Observation b00d2833-ff52-4abf-b2af-97f2860c9cfc · outbound

This paper cites chrf: character n-gram f-score for automatic mt evalu- ation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension chrf: character n-gram f-score for automatic mt evalu- ation,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.190751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.190751Z digest=sha256:92ff324c8cf21edb46fc1ead84de57aecd7b494ed287fd4c3e70acaa1e0ffb3f

Observation 526d3951-d78a-4f68-a0a5-34915fb4f34a · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.195631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.195631Z digest=sha256:06bed64cb41f36cca19aef9896b39c24cf1316b167826ebaf84248e3b7ab0d61

Observation a931a814-1e50-4419-8454-83330b2aedd9 · outbound

This paper cites ICE-Score: Instructing Large Language Models to Evaluate Code.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension ICE-Score: Instructing Large Language Models to Evaluate Code

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.200654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.200654Z digest=sha256:0a607f6565fbc60f383a07ae213ac96c3686df877aa2dcfe60533f69887874d3

Observation fc582aaf-96ce-489c-bc10-abc4196325ca · outbound

This paper cites Learning to mine aligned code and natural language pairs from stack overflow,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Learning to mine aligned code and natural language pairs from stack overflow,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.206144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.206144Z digest=sha256:0c08158852b172e488dac46b512af137db29c422f440edfec66d0498f4861100

Observation dc7885f2-297a-4488-b3a4-02136c718eaa · outbound

This paper cites Latent Predictor Networks for Code Generation.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Latent Predictor Networks for Code Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.211234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.211234Z digest=sha256:d6f2cc9bfb82be8aca7cd81bfea27167a58c3b2b96cd55b37f570c46daa9263d

Observation 7b7dda7c-3b6a-4183-a7ec-cd239a5faef4 · outbound

This paper cites Towards a Unified Multi-Dimensional Evaluator for Text Generation.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Towards a Unified Multi-Dimensional Evaluator for Text Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.217643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.217643Z digest=sha256:af22291645287296798c2361518bad4bb0911088945e5cd651cf4761f3d02391

Observation 9b1269a7-94a0-469d-ba27-29e26808068e · outbound

This paper cites Survey of hallucination in natural language generation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Survey of hallucination in natural language generation,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.223383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.223383Z digest=sha256:d933b9e3550e04453c44a04b1217d5f53f49341aa7163e237d93a01a00e6d31d

Observation 911bbbca-0825-4cfd-aebe-06eb822312c4 · outbound

This paper cites Faithful Reasoning Using Large Language Models.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Faithful Reasoning Using Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.228537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.228537Z digest=sha256:c4d5cc1184ccfab7bbd04533cbb86e4fd6c5187ed97d3c4518ffd294e46e2993

Observation 38d12603-7d58-4eef-b6d8-a1b436fc1f86 · outbound

This paper cites CodeNet: A Large-Scale AI for Code Dataset for Learning a Diversity of Coding Tasks.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeNet: A Large-Scale AI for Code Dataset for Learning a Diversity of Coding Tasks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.233779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.233779Z digest=sha256:2de80a00a1d9da81f44274a09d599905cb2883cd435127f79df21ca1c901765a

Observation 15094510-9d6d-4464-adee-e218b43d67e8 · outbound

This paper cites Increasing employability of indian engineering graduates through experiential learning programs and competitive programming: Case study,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Increasing employability of indian engineering graduates through experiential learning programs and competitive programming: Case study,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.871489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.239284Z digest=sha256:2100a6bacfd9c184231ed4633aa94b9eee648f5f59dcf2d861ae5fb1c9a2ab57

Observation d119ebda-43a2-480f-8312-099250ea6d70 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Evaluating Large Language Models Trained on Code

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.244040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.244040Z digest=sha256:87c5179061516a772e97adba9ba92b24c476fd7c0cea44c26b33b92add6424a8

Observation a7846e2d-37fc-4255-bb1b-7869fa9d4059 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Chain-of-thought prompting elicits reasoning in large language models,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.249521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.249521Z digest=sha256:f5f8f41c04c834b620fd1ee4ccaebe2aefadfd8a17f0b95a6bab2ff7931e43bb

Observation faa84c3d-c7b7-4c30-b14c-c3d58d8170ad · outbound

This paper cites Fine-grained code clone detection with block-based splitting of abstract syntax tree,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Fine-grained code clone detection with block-based splitting of abstract syntax tree,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.843479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.254513Z digest=sha256:97d3706ea77a802daa1a2499e87a9a4d5ae0e06857c3a994e8a16c26d5828c59

Observation 4cdc6928-b329-4be6-a59f-0a6e3d58de07 · outbound

This paper cites Cocoast: Representing source code via hierarchical splitting and re- construction of abstract syntax trees,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Cocoast: Representing source code via hierarchical splitting and re- construction of abstract syntax trees,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.827289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.259535Z digest=sha256:dc6da483c87420af4d7da3b590bfb14a67088326a721e53a1d707ab24d2eda27

Observation d0f395a4-6c88-4b23-8d2e-8e074dec6c5f · outbound

This paper cites Blocsum: Block scope-based source code summarization via shared block representation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Blocsum: Block scope-based source code summarization via shared block representation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.809874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.264836Z digest=sha256:80fcb3fa5908d5405785d982e66c66c89bdc69113b075dbf27ca841c3c198c39

Observation 674e58ec-730e-4b40-b7df-0ba2037027af · outbound

This paper cites Openai gpt-3.5 turbo,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Openai gpt-3.5 turbo,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.791428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.270072Z digest=sha256:7568c0fa050dc0452bbe93b8b1dd1e06c94667c5667caabfca19fe062bc91679

Observation 2aeb5347-8c25-45fc-bcd5-f766c271daf2 · outbound

This paper cites Exploring security vulnerabilities in competitive programming: An empirical study,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Exploring security vulnerabilities in competitive programming: An empirical study,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.774773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.275147Z digest=sha256:1e13c4dc05db1eb97bfa6cbc748da666d43be20e94d212ef6d3d1bec2e45e140

Observation 3ae686d0-9386-4566-837b-0b9fd98375a2 · outbound

This paper cites Multipl-e: A scalable and polyglot approach to benchmarking neural code generation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Multipl-e: A scalable and polyglot approach to benchmarking neural code generation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.756582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.280106Z digest=sha256:b333d8e56c4e5d7434ff47da0cdd01ff6cb5260746096a7757c7cefc8b8fd454

Observation 69553955-b7c8-4108-a109-f3161c4fe0c0 · outbound

This paper cites Gpt-4: Language models at scale,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Gpt-4: Language models at scale,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.739906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.285090Z digest=sha256:7bf083a25958907ae71982c71db5910de045de74ada8e25d9431a3c6c21c76a7

Observation 7d8c3c11-2f7e-4ff2-a96a-7a5e878a1e38 · outbound

This paper cites Openai gpt-4 turbo,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Openai gpt-4 turbo,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.722040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.289829Z digest=sha256:b428a80295db466263317a69d909300f626296526df000e4b86cf99da014a037

Observation 7c7178db-2b09-486a-aa18-98ba39063488 · outbound

This paper cites Meteor: An automatic metric for mt evalua- tion with improved correlation with human judgments,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Meteor: An automatic metric for mt evalua- tion with improved correlation with human judgments,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.294747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.294747Z digest=sha256:d008df8237cf7f5670cd2ebdce5285253b264ef48fccb82a73f872fa3885845e

Observation 76779232-788b-4f44-8ea5-5f673283996a · outbound

This paper cites Large lan- guage models are zero-shot reasoners,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Large lan- guage models are zero-shot reasoners,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.299866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.299866Z digest=sha256:5a121c211a594404a89b611e50b42be32ad804717b7278e8771a3c1537f6887d

Observation df9a7451-b3c2-4b4f-919b-4f0f479b29ab · outbound

This paper cites Language mod- els are few-shot learners,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Language mod- els are few-shot learners,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.305123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.305123Z digest=sha256:eec4b5b2cd65b31762ceb48cca53866dea691cf9b83b8e99c8bc04072308ec3b

Observation 4cfeeb3d-d209-4133-bc10-1c835bd9c235 · outbound

This paper cites A new measure of rank correlation,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension A new measure of rank correlation,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.310133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.310133Z digest=sha256:6cf56ebb9b5a437b327da92f94bc94c8a72f6b7095a647bfb4f23dd43694be01

Observation 58cd1b00-3a99-4ab5-b1d2-a38d1dd020a7 · outbound

This paper cites Pearson correlation coefficient,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Pearson correlation coefficient,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.315250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.315250Z digest=sha256:b66804fe761ea5f4c5d5be4bc634a071f49fc61f35f0e00c69829ea1855429fc

Observation 31a71bb3-22ba-4f7b-ba02-01f535efb7b9 · outbound

This paper cites Likert scale: Explored and explained,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Likert scale: Explored and explained,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.321002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.321002Z digest=sha256:824640871cfcd8d03774a160c5b054c248601471ecf9ff02db2f00cbaf360e69

Observation 65067bc0-9be7-446d-9b5c-1c0d2785c538 · outbound

This paper cites Unsu- pervised translation of programming languages,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Unsu- pervised translation of programming languages,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.641453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.326026Z digest=sha256:6f90fd80ecc85b083d47fab6c040c8f7c4fdaeca00d03085593db14a6316255b

Observation 98cca902-aa89-44ac-80d1-ee826433265b · outbound

This paper cites An empirical study of auto- mated unit test generation for python,.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension An empirical study of auto- mated unit test generation for python,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:51.623676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:51.331017Z digest=sha256:4f994c79d8a9b0772cc9b3c372d16121e8c097029c8cf47b8c50dc8340d018e0

Observation 24059aa6-2376-47a7-9ebe-1ebf7aeeb741 · outbound

This paper cites Can Large Language Models Be an Alternative to Human Evaluations?.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Can Large Language Models Be an Alternative to Human Evaluations?

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.335834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.335834Z digest=sha256:65382519964c0d71035548ca533ccfd88ed8083292762f06d98262f98988dc99

Observation b4e5cb43-f826-4735-be77-653f9dbce629 · outbound

This paper cites G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment.

Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:51.340765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:51.340765Z digest=sha256:5eff15371c0d7e0082fec1dd05fc7e53f9c6fa46f3810987fde0ef1d89851cc2

Pith citing papers

Observation 107c2bc4-bd65-498b-8a61-d813d5839c35 · inbound

FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection cites this paper.

FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:24:28.422461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T16:24:28.140998Z digest=sha256:3b4ddca487796788b333ef8fa948c0338c17bd400ace85455bf58868f33b99ff