Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T05:35:51.340765Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2412.00314.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T05:35:51.340765Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T16:24:28.140998Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T16:24:28.418595Z
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 752027d0-541b-4e23-9dde-d15fe79095bc · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Who judges the judge: An empirical study on online judge tests,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f02cab78-82c1-40ef-93de-817ed722b6ba · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Automatic source code evaluation test develop- ment in programming education using grey-box methods,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e6d004e8-46e5-4149-a0f6-ebf5bbbd1cc2 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6f2a5a7-d864-428e-b609-d78100806a24 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Coderl: Mastering code generation through pretrained models and deep reinforcement learning,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation deb65a93-2211-412d-927b-f870e39d239f · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Large language models meet nl2code: A survey,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1cd7d9a5-1049-42db-9821-3db990e28abc · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeScore: Evaluating Code Generation by Learning Code Execution
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3992f1cc-c07c-4b5e-aa1b-3b96311eb678 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Spoc: Search-based pseudocode to code,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0bd294d-bee7-4eb8-a74d-6da231ae2dd3 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Measuring Coding Challenge Competence With APPS
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a5342d1-b826-470c-ab36-4e52373ada91 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Out of the bleu: how should we assess quality of the code generation models?
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1287646d-da8b-41d5-8c8c-112888733f82 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Bleu: a method for automatic evaluation of machine translation,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c60260b5-3f25-40ea-a253-c423b5844166 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension A package for automatic evaluation of summaries,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation cdd5a12a-ef2c-468e-ac8c-754a3cc599ff · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Does bleu score work for code migration?
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f5dc2978-eab1-4b94-8c2f-b83e9be49e99 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeBERTScore: Evaluating Code Generation with Pretrained Models of Code
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 068580ea-0365-4a79-8662-dc320204b5af · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeBLEU: a Method for Automatic Evaluation of Code Synthesis
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b00d2833-ff52-4abf-b2af-97f2860c9cfc · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension chrf: character n-gram f-score for automatic mt evalu- ation,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 526d3951-d78a-4f68-a0a5-34915fb4f34a · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a931a814-1e50-4419-8454-83330b2aedd9 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension ICE-Score: Instructing Large Language Models to Evaluate Code
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc582aaf-96ce-489c-bc10-abc4196325ca · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Learning to mine aligned code and natural language pairs from stack overflow,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc7885f2-297a-4488-b3a4-02136c718eaa · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Latent Predictor Networks for Code Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b7dda7c-3b6a-4183-a7ec-cd239a5faef4 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Towards a Unified Multi-Dimensional Evaluator for Text Generation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b1269a7-94a0-469d-ba27-29e26808068e · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Survey of hallucination in natural language generation,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 911bbbca-0825-4cfd-aebe-06eb822312c4 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Faithful Reasoning Using Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38d12603-7d58-4eef-b6d8-a1b436fc1f86 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension CodeNet: A Large-Scale AI for Code Dataset for Learning a Diversity of Coding Tasks
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15094510-9d6d-4464-adee-e218b43d67e8 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Increasing employability of indian engineering graduates through experiential learning programs and competitive programming: Case study,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d119ebda-43a2-480f-8312-099250ea6d70 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Evaluating Large Language Models Trained on Code
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7846e2d-37fc-4255-bb1b-7869fa9d4059 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Chain-of-thought prompting elicits reasoning in large language models,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faa84c3d-c7b7-4c30-b14c-c3d58d8170ad · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Fine-grained code clone detection with block-based splitting of abstract syntax tree,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4cdc6928-b329-4be6-a59f-0a6e3d58de07 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Cocoast: Representing source code via hierarchical splitting and re- construction of abstract syntax trees,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d0f395a4-6c88-4b23-8d2e-8e074dec6c5f · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Blocsum: Block scope-based source code summarization via shared block representation,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 674e58ec-730e-4b40-b7df-0ba2037027af · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Openai gpt-3.5 turbo,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2aeb5347-8c25-45fc-bcd5-f766c271daf2 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Exploring security vulnerabilities in competitive programming: An empirical study,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 3ae686d0-9386-4566-837b-0b9fd98375a2 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Multipl-e: A scalable and polyglot approach to benchmarking neural code generation,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 69553955-b7c8-4108-a109-f3161c4fe0c0 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Gpt-4: Language models at scale,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7d8c3c11-2f7e-4ff2-a96a-7a5e878a1e38 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Openai gpt-4 turbo,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7c7178db-2b09-486a-aa18-98ba39063488 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Meteor: An automatic metric for mt evalua- tion with improved correlation with human judgments,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76779232-788b-4f44-8ea5-5f673283996a · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Large lan- guage models are zero-shot reasoners,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df9a7451-b3c2-4b4f-919b-4f0f479b29ab · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Language mod- els are few-shot learners,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cfeeb3d-d209-4133-bc10-1c835bd9c235 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension A new measure of rank correlation,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58cd1b00-3a99-4ab5-b1d2-a38d1dd020a7 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Pearson correlation coefficient,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31a71bb3-22ba-4f7b-ba02-01f535efb7b9 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Likert scale: Explored and explained,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65067bc0-9be7-446d-9b5c-1c0d2785c538 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Unsu- pervised translation of programming languages,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 98cca902-aa89-44ac-80d1-ee826433265b · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension An empirical study of auto- mated unit test generation for python,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 24059aa6-2376-47a7-9ebe-1ebf7aeeb741 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension Can Large Language Models Be an Alternative to Human Evaluations?
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4e5cb43-f826-4735-be77-653f9dbce629 · outbound
Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 107c2bc4-bd65-498b-8a61-d813d5839c35 · inbound
FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection Human-Like Code Quality Evaluation through LLM-based Recursive Semantic Comprehension
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.