Pith. sign in

Paper Citation Record · LEDGER

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs

As of 8 August 2026, this Paper Citation Record lists 100 of 111 outbound references and 1 inbound Pith citation observation for arXiv:2506.11059.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11059 v1

Coverage vector

measured 100 of 111 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:55:15.691408Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:54:51.091578Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T23:54:56.588252Z

Reference resolution

100 of 111 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved59
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 65788c0f-6182-4abf-b3a8-2ea9b74c138c · outbound

This paper cites GPT-4 Technical Report.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.476756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.476756Z digest=sha256:518d305995c86aa9bbe3481ef3faed70784a68a13b5fc362a93f31f1aa3e3c24

Observation 75df5980-73d5-411e-8495-bbdb621b2640 · outbound

This paper cites An Empirical Study of AI Generated Text Detection Tools.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs An Empirical Study of AI Generated Text Detection Tools

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.570106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.570106Z digest=sha256:d20a7b7a671fb6f1cb74205f8e009f758f935efa5a8d19cc3afcfa0ee04fa66d

Observation 908da8b7-9eeb-4478-9439-254c9cab0d7e · outbound

This paper cites Introducing computer use, a new claude 3.5 sonnet, and claude 3.5 haiku, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Introducing computer use, a new claude 3.5 sonnet, and claude 3.5 haiku, 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.633741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.633741Z digest=sha256:c957270bc847028eea6c1145abf1cac08ca0e3c6922983e1cff8cfbe00c46690

Observation a7846d60-9b79-475a-9822-16282d44c1d2 · outbound

This paper cites Claude 3.7 Sonnet and Claude Code, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Claude 3.7 Sonnet and Claude Code, 2025

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.720560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.720560Z digest=sha256:d8881a846c53e4c26c13cdf24fb2a6a8c0879b12565336c268b29677791856dc

Observation c5a100e3-a86e-4296-89d1-260fca603978 · outbound

This paper cites Is github’s copilot as bad as humans at introducing vulnerabilities in code?Empirical Software Engineering, 28(6):129, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Is github’s copilot as bad as humans at introducing vulnerabilities in code?Empirical Software Engineering, 28(6):129, 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.811473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.811473Z digest=sha256:bdbe42ac3bce5137e359941de6f283089e4d617e12e785d9256144d8c3f2565e

Observation c0a0349f-1f09-4419-bb99-36ab95a46fc6 · outbound

This paper cites Random forests.Machine learning, 45:5–32, 2001.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Random forests.Machine learning, 45:5–32, 2001

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.895270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.895270Z digest=sha256:5cde47caad65f0bbf9a641a18990a5ec34ba58921a98cba2b26edb33a69a2063

Observation f391913e-6860-425b-8fa2-3debfa1511e7 · outbound

This paper cites Membership inference attacks from first principles.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Membership inference attacks from first principles

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.976947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.976947Z digest=sha256:db34b432a08d6f4e79302c3054cb3664f087985e9e049e485b9d8be066ddefc2

Observation 4e3022a3-1094-44a2-af28-e5090a1fb735 · outbound

This paper cites R package version 3.0.1.1.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs R package version 3.0.1.1

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.063202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.063202Z digest=sha256:46e724d61924ac84ba37f7be40ceacf79e74fc236ed8ad6d58b006e7364cf41f

Observation 553cb2d9-df04-4555-913c-b75ef7780dab · outbound

This paper cites Github code clean dataset, 2022.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Github code clean dataset, 2022

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.129620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.129620Z digest=sha256:92967db24cd3c4f899b60a23da89e5b044706cefbdbbcab77ffa2f05a3350686

Observation 02afa950-9b55-48b6-b04f-744abf07e8c5 · outbound

This paper cites Github code dataset, 2022.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Github code dataset, 2022

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.200058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.200058Z digest=sha256:bc9fb5723464d436ddc0b40fd1e16918db3f2a4173c0ee237f36bb0778ae4d0c

Observation f25dda91-c58c-43df-8202-91a4e16cfd1a · outbound

This paper cites Vulnerabilities in ai code generators: Exploring targeted data poisoning attacks.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Vulnerabilities in ai code generators: Exploring targeted data poisoning attacks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.294524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.294524Z digest=sha256:9ec1df32965d8c0096705bf0388b45d7a805d7ff9854e604c11125710249e0a9

Observation 29e82846-deea-4d38-bd2e-2d9edfe7f36b · outbound

This paper cites Cursor: The AI Code Editor, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Cursor: The AI Code Editor, 2023

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.393332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.393332Z digest=sha256:7ab8d5c789bce5b7c3ecb769ee9753b8011e1c47ef335494a8d710999b6d368e

Observation 1b0a11d0-a7f0-462f-8638-6c5988df77d1 · outbound

This paper cites Plagiarism in the age of massive generative pre-trained transformers (gpt-3).Ethics in Science and Environmental Politics, 21:17–23, 2021.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Plagiarism in the age of massive generative pre-trained transformers (gpt-3).Ethics in Science and Environmental Politics, 21:17–23, 2021

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.467782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.467782Z digest=sha256:2f9fb2414caad8e74c3c9be9012a27d23fdeb03d5ba9629d413013123e293822

Observation 9338dc5a-b143-4bc3-821b-57e1450a2458 · outbound

This paper cites AIGCodeSet: A New Annotated Dataset for AI Generated Code Detection.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs AIGCodeSet: A New Annotated Dataset for AI Generated Code Detection

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.550790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.550790Z digest=sha256:420993e77056b45c1ccf33cbc7ec3e30567f22ba304b259042b198d90be5a1fa

Observation d7d6f9c8-d5dd-4eab-8e0f-659fe2ae56a6 · outbound

This paper cites The DeepFake Detection Challenge (DFDC) Dataset.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs The DeepFake Detection Challenge (DFDC) Dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.640002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.640002Z digest=sha256:1032c88191b433a9324a140dbfbd99348f3894da07b53f42b1dd9d639e602f29

Observation 8d87665b-4f0b-46c3-8a0e-7128c09be92f · outbound

This paper cites Codep: grammatical seq2seq model for general-purpose code generation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Codep: grammatical seq2seq model for general-purpose code generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.701676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.701676Z digest=sha256:aff8a9790af67316ec5c3f48aa3fb48768a0229c834e4961a7813f9a90217e0a

Observation 08fb92ab-8479-4797-be7c-5184de1e9f47 · outbound

This paper cites JaCoText: A Pretrained Model for Java Code-Text Generation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs JaCoText: A Pretrained Model for Java Code-Text Generation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:55:17.120274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:47.773655Z digest=sha256:b997abbb79a9d3a3587a1418c846302f8c3f6bbf9061dbdd438e0ec69c8f988f

Observation cc471623-5e94-487d-a09f-a0dd57b31ff3 · outbound

This paper cites Out of the bleu: how should we assess quality of the code generation models?Journal of Systems and Software, 203:111741, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Out of the bleu: how should we assess quality of the code generation models?Journal of Systems and Software, 203:111741, 2023

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.862883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.862883Z digest=sha256:8d3600a926c3393483607b0b493ca08fc4f4a73b1d9379b73dc4ff8cead7fed3

Observation 11467041-3d7d-44fc-acf8-15269374ff63 · outbound

This paper cites CodeBERT: A Pre-Trained Model for Programming and Natural Languages.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs CodeBERT: A Pre-Trained Model for Programming and Natural Languages

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.920730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.920730Z digest=sha256:a33d70e8d49b4cfcaedd3c4ed50312fea68c891ab64f2007d182c0e29fa5e77d

Observation 21847269-8275-48a4-b353-0e7bb165121f · outbound

This paper cites Introducing GitHub Copilot: your AI pair programmer, 2022.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Introducing GitHub Copilot: your AI pair programmer, 2022

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.011427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.011427Z digest=sha256:8b014ca9a72ddd661022b1f50f01b2709f71d8a186e559fb29a6ec5c521901e0

Observation d7e6ddef-b792-4499-8ca8-05d7ac0df019 · outbound

This paper cites What makes good in-context demonstrations for code intelligence tasks with llms? InIEEE/ACM International Conference on Automated Software Engineering (ASE), pages 761–773, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs What makes good in-context demonstrations for code intelligence tasks with llms? InIEEE/ACM International Conference on Automated Software Engineering (ASE), pages 761–773, 2023

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.070030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.070030Z digest=sha256:42a59ee75b468a757781163a71288ea8927d6c5b325d5f5853d9dec8f93a5e15

Observation 1528c165-1662-4f98-a82a-3165eaa55899 · outbound

This paper cites Gltr: Statistical detection and visualization of generated text.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Gltr: Statistical detection and visualization of generated text

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.124138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.124138Z digest=sha256:68aa9e2feda1d9f7d24cc88a61fc88aced1225db3e4dbb1f57037001e0063df4

Observation 314297cd-8c44-4ef9-b4a0-4a1e554b2176 · outbound

This paper cites A survey on the possibilities & impossibilities of ai-generated text detection.Transactions on Machine Learning Research (TMLR), 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs A survey on the possibilities & impossibilities of ai-generated text detection.Transactions on Machine Learning Research (TMLR), 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.246520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.246520Z digest=sha256:2139d981800bca133743e14bf4a8472dbd7de83448bb93468d3c8606446f83e4

Observation b5aac917-e193-4767-ad10-2fbec12deecd · outbound

This paper cites Deepfake video detection using recurrent neural networks.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Deepfake video detection using recurrent neural networks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.359564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.359564Z digest=sha256:26596ff2b92ce1f36fdd2714ce15472099a9c46cbf1a0437d267343f0af3844f

Observation f13457b0-5431-42a2-b472-218cb58cd458 · outbound

This paper cites Graphcodebert: Pre-training code representations with data flow.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Graphcodebert: Pre-training code representations with data flow

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.449654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.449654Z digest=sha256:ad7fccc14dcf28a175abcebcfd3513d5a6d0836d5d5c9197b5411021d57c5a48

Observation 05003157-91eb-4798-8360-04d2c36081b1 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.545437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.545437Z digest=sha256:89e80ac5d3aed834af889fed0b6770042ee5c08168c4122dc0707aa8fbeffa2c

Observation eb20765c-9e15-4507-8edc-b618ad05b088 · outbound

This paper cites Biscope: Ai-generated text detection by checking memorization of preceding tokens.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Biscope: Ai-generated text detection by checking memorization of preceding tokens

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.665601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.665601Z digest=sha256:b268975ab8e3d2b17320b4a02e63423bf2e630d248a1f0c45ee7f91dd3966589

Observation c66b5304-5d20-4b43-abe3-3212d8cf7a90 · outbound

This paper cites Spotting llms with binoculars: Zero-shot detection of machine-generated text.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Spotting llms with binoculars: Zero-shot detection of machine-generated text

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.809631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.809631Z digest=sha256:6bc63840a31595bb3d5b2d7b52615dc1a33810a9d5bddf50d85452d2eb79cf23

Observation 5195c632-aa5c-4485-855a-96398f705a7d · outbound

This paper cites Denoising diffusion probabilistic models.Advances in Neural Information Processing Systems (NeurIPS), 33:6840–6851, 2020.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Denoising diffusion probabilistic models.Advances in Neural Information Processing Systems (NeurIPS), 33:6840–6851, 2020

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.908202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.908202Z digest=sha256:2f7c87f211cebe0928c0ac0d88af21b94f8a3028a4fdaf638d86122c755752b4

Observation 83737ca2-14d0-46cf-85c5-01beb47d2d14 · outbound

This paper cites Qwen2.5-Coder Technical Report.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Qwen2.5-Coder Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.989561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.989561Z digest=sha256:90e9b5adfaf49aa04ea4fd694ec88de1a0661b94073966086e213046c29e8df2

Observation 77d0cb4e-9f09-4066-ab9d-84f237151c12 · outbound

This paper cites Rethinking plagiarism in the era of generative ai.Journal of Intelligent Communication, 3(2):20–31, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Rethinking plagiarism in the era of generative ai.Journal of Intelligent Communication, 3(2):20–31, 2024

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.120454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.120454Z digest=sha256:8d5cf614826f3e15b3dbba915caeb54acc5b7e56122fba02c769216777757f19

Observation 3207f12b-4c39-4ebc-951d-9d6ed0e8ceb6 · outbound

This paper cites Whodunit: Classifying code as human authored or gpt-4 generated-a case study on codechef problems.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Whodunit: Classifying code as human authored or gpt-4 generated-a case study on codechef problems

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.242525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.242525Z digest=sha256:3a588c04a943b9ea98bf44bbd8e14eb4b7fe8460df4d90f6e198622ca82556c6

Observation bc477be6-3ca6-452d-a785-58be0314634a · outbound

This paper cites Automatic detection of generated text is easiest when humans are fooled.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Automatic detection of generated text is easiest when humans are fooled

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.340671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.340671Z digest=sha256:d0ee059c6aacd21f2ef5b79f29a12518077d5651bd12d9b0160e3398950c5594

Observation c4238f08-8892-4f25-8231-5600621c9da7 · outbound

This paper cites Swe-bench: Can language models resolve real-world github issues? InInternational Conference on Learning Representations (ICLR), 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Swe-bench: Can language models resolve real-world github issues? InInternational Conference on Learning Representations (ICLR), 2024

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.537272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.537272Z digest=sha256:20bff6e909700da377d786e871e1cea139179df1c3f783d78074e757b7223823

Observation bae908c6-127e-4bda-a198-82489ae72e37 · outbound

This paper cites Access the latest 2.0 experimental models in the gemini app., 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Access the latest 2.0 experimental models in the gemini app., 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.661149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.661149Z digest=sha256:9e5794c8811577c6e4ca7e7d053ecf0a153f74b6b21cd2db2d1efc066be05037

Observation 605bc256-bacb-4710-8e72-c6021ee8bb78 · outbound

This paper cites Vulnerability handling of ai- generated code-existing solutions and open challenges.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Vulnerability handling of ai- generated code-existing solutions and open challenges

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.780282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.780282Z digest=sha256:e01b5025af3700ca1f508513cb7fa6ae9622ab34709ec72edb2db8647751ffe7

Observation d508cc1b-25f4-4f13-bd1b-d3a77ffe37be · outbound

This paper cites Gemini 2.0 is now available to everyone, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Gemini 2.0 is now available to everyone, 2025

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.898415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.898415Z digest=sha256:a9529f98b2e5d50e02dac54c65a883b19df3f5036f7aa456f612bfe51bc701ac

Observation dbda737d-454c-453b-ba59-b83e84ff4e8e · outbound

This paper cites Does attitude towards plagiarism predict aigiarism using chatgpt?AI and Ethics, 5(1):677–688, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Does attitude towards plagiarism predict aigiarism using chatgpt?AI and Ethics, 5(1):677–688, 2025

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:50.058627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:50.058627Z digest=sha256:52646029302ed5c3d654ed73e5efe41ec230ce228dbea9602700a2fff4665b37

Observation ef0bb54f-4125-47f8-89f4-f384d8462faf · outbound

This paper cites Will chatgpt g et you caught? rethinking of plagiarism detection.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Will chatgpt g et you caught? rethinking of plagiarism detection

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:50.183926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:50.183926Z digest=sha256:f63d993e76c3fd6a98404ef5efecd0f102069f257dc8217c253d5f637a47ed9e

Observation 64ffc721-c503-4bb6-a5a5-533ecc3ba99f · outbound

This paper cites How secure is code generated by chatgpt? InIEEE international conference on systems, man, and cybernetics (SMC), pages 2445–2451.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs How secure is code generated by chatgpt? InIEEE international conference on systems, man, and cybernetics (SMC), pages 2445–2451

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:26.240995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:50.269865Z digest=sha256:198416ad1adf09fec02d8fbd135329aa15fe148777d904074204217f4a1db36c

Observation 03f30dd7-d73a-47ee-bd15-91d0063106e0 · outbound

This paper cites Paraphrasing evades detectors of ai-generated text, but retrieval is an effective defense.Advances in Neural Information Processing Systems (NeurIPS), 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Paraphrasing evades detectors of ai-generated text, but retrieval is an effective defense.Advances in Neural Information Processing Systems (NeurIPS), 2023

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:26.074964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:50.386281Z digest=sha256:2c324b82e655ba7289ec62fd33e70b83796f86ab75b1992f6504d1aa60f69d8e

Observation 4271435f-c3b5-455a-8867-637eaf646188 · outbound

This paper cites Detecting fake content with relative entropy scoring.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Detecting fake content with relative entropy scoring

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.908463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:50.553406Z digest=sha256:9033523635fe319986cd3a02db71175517dc656983986f55ea7e7ab1825db303

Observation 50499960-1047-4a1f-b5a3-ade3debd7f65 · outbound

This paper cites Protecting intellectual property of large language model-based code generation apis via watermarks.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Protecting intellectual property of large language model-based code generation apis via watermarks

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.821311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:50.664897Z digest=sha256:fc65c15b937b4ae3500cfd7cc245f377785fc11c9dad890aa1aa0f80fdaee60f

Observation cabc3883-4bb6-4818-80ea-24f5fff2cfc9 · outbound

This paper cites DeepSeek-V3 Technical Report.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs DeepSeek-V3 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:50.744812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:50.744812Z digest=sha256:7adfd2900d8ccd558a7e355ba4fdd9f7ea36d5413e8e01191e039af907b9706d

Observation 1da2f7b4-1105-407c-a195-1bc12dcf90ad · outbound

This paper cites an unresolved cited work.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:25.697427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:50.865882Z digest=sha256:54d383e5e9d6982077b7973ddc55426fcca61a1449e58e70af4f6b5a35d8f6a5

Observation 04279a04-6e3d-4f41-8e20-d41d409b3ac1 · outbound

This paper cites CodeXEmbed: A Generalist Embedding Model Family for Multiligual and Multi-task Code Retrieval.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs CodeXEmbed: A Generalist Embedding Model Family for Multiligual and Multi-task Code Retrieval

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.004303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.004303Z digest=sha256:1cb4a855d98024b827b90c628fc215fe0d10228e958968749489046e1d75af86

Observation 2ce6caf2-b835-4f6e-80a1-10112003ecd9 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.120028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.120028Z digest=sha256:a548b69353f4db9dc0a229abce7f820d0744b1fb0efbbe603a0dc3c73845d232

Observation c6f9638a-d993-4bd4-bf4d-944c3afec798 · outbound

This paper cites Raidar: generative ai detection via rewriting.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Raidar: generative ai detection via rewriting

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.537792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.218996Z digest=sha256:804e701c9a3e71d96778718857435862ad3da095ff6eb62bcf48d5297d67a87b

Observation db5bea98-ac35-408c-8e6a-9907e16e7a86 · outbound

This paper cites On the robustness of code generation techniques: An empirical study on github copilot.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs On the robustness of code generation techniques: An empirical study on github copilot

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.415448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.348636Z digest=sha256:b0172e17f0c906f3c190d1f2a72883cb9e01578c7b17a5da597f7baed9310a5b

Observation 47c2b334-6ec8-4be7-84ef-3bb770d4ba7f · outbound

This paper cites Llama 3.3: Model cards & prompt formats, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Llama 3.3: Model cards & prompt formats, 2024

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.275821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.489432Z digest=sha256:132894f43a74a35834c2420c1cd8b0716226c19eaaf33d82002177452e0f52b0

Observation a1845f18-307f-48db-99d2-ef1065386dff · outbound

This paper cites Detectgpt: Zero-shot machine-generated text detection using probability curvature.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Detectgpt: Zero-shot machine-generated text detection using probability curvature

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.167755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.641080Z digest=sha256:fdccee6b398e6f7e69ea8208f4bc5d3bcf688b68389a5dc4074df79967fc6a52

Observation 369ee951-331b-4b17-a544-0fc92734f160 · outbound

This paper cites Is this Snippet Written by ChatGPT? An Empirical Study with a CodeBERT-Based Classifier.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Is this Snippet Written by ChatGPT? An Empirical Study with a CodeBERT-Based Classifier

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.700600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.700600Z digest=sha256:f79f73fcc1d457ed4c617a61d4a81565b5b78b5aff96fbe6a65b94f80872b570

Observation e054e314-9f13-4178-bf36-13dc638c012a · outbound

This paper cites Gptsniffer: A codebert-based classifier to detect source code written by chatgpt.Journal of Systems and Software, 214:112059, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Gptsniffer: A codebert-based classifier to detect source code written by chatgpt.Journal of Systems and Software, 214:112059, 2024

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.981311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.759655Z digest=sha256:fdf8244a5baf319d83f305f485c7c7bb7fb6d66e3ef1eaa1847d0c651f85968a

Observation 5607cbcf-e488-48c7-945f-c638c8485b33 · outbound

This paper cites Poisoned chatgpt finds work for idle hands: Exploring developers’ coding practices with insecure suggestions from poisoned ai models.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Poisoned chatgpt finds work for idle hands: Exploring developers’ coding practices with insecure suggestions from poisoned ai models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.839371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.838559Z digest=sha256:9e77780e7d8542d7de3e8f07a2144bf73cb94753350804f78e45ef83bb899921

Observation 16b5d11e-f0ca-4610-be35-b2a5a824779c · outbound

This paper cites Introducing ChatGPT, 2022.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Introducing ChatGPT, 2022

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.919051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.919051Z digest=sha256:ec3b3a6b3dab9bc51d419b2526f48d8fb015340eed0e71fc35bd170a1908d910

Observation 9e47c2de-2fd7-4582-a635-8af1d6a61da5 · outbound

This paper cites Gpt-4o mini: advancing cost-efficient intelligence, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Gpt-4o mini: advancing cost-efficient intelligence, 2024

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.992202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.992202Z digest=sha256:8594d04744b51b133e6c6e5a984169b66683d42637ca0c244aa7ee4f168ebcae

Observation adbea689-e8b3-492f-b51e-222122821c97 · outbound

This paper cites Openai o3-mini: Pushing the frontier of cost-effective reasoning, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Openai o3-mini: Pushing the frontier of cost-effective reasoning, 2025

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.666988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.043090Z digest=sha256:805aef5cf823130f40d7b9b1e4cbb2a90040c7e11610e0c701cc26f500b998db

Observation 3f3a22e2-6ddf-4c0d-8dc1-a910825615d8 · outbound

This paper cites CoDet-M4: Detecting Machine-Generated Code in Multi-Lingual, Multi-Generator and Multi-Domain Settings.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs CoDet-M4: Detecting Machine-Generated Code in Multi-Lingual, Multi-Generator and Multi-Domain Settings

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:52.108811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:52.108811Z digest=sha256:af293b0b04efa9ca00a32216f65b9bc93a673eb20797bb58a358158726349e38

Observation a74c8129-f622-4c8e-855d-711bc48a174f · outbound

This paper cites Assessing ai detectors in identifying ai-generated code: Implications for education.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Assessing ai detectors in identifying ai-generated code: Implications for education

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.515159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.157680Z digest=sha256:a49696ca43404ca674302f99c1d6f287536437cc29a16836a1d37c33bb58844c

Observation 1b147bce-2743-4336-be30-1d54c61c8701 · outbound

This paper cites Bleu: A method for automatic evaluation of machine translation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Bleu: A method for automatic evaluation of machine translation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.371342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.221757Z digest=sha256:01c2a12bca1850f2e617a90d7fbcc1672a14f2620c4bbe60f231914bcc728b7d

Observation 68814478-2094-4fa8-9715-03362c8d9dfa · outbound

This paper cites Asleep at the keyboard? assessing the security of github copilot’s code contributions.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Asleep at the keyboard? assessing the security of github copilot’s code contributions

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.265456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.278896Z digest=sha256:bcf37aa36269986b8f26fe515d394df95ccc36214507c9d7d344793c2e4c9ee7

Observation b8253700-57d5-448a-87ba-54f616a19d6b · outbound

This paper cites Magecode: Machine- generated code detection method using large language models.IEEE Access, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Magecode: Machine- generated code detection method using large language models.IEEE Access, 2024

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.178275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.343700Z digest=sha256:ca75385283227ac61ebd748671a019f1684c13a090f5a1f04309c1a15a973ff9

Observation b9be9d4a-a780-403d-b1ff-d1c738eadafe · outbound

This paper cites Introducing gemini 2.0: our new ai model for the agentic era, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Introducing gemini 2.0: our new ai model for the agentic era, 2024

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.021607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.428624Z digest=sha256:79572e8ea51622acfe215b75323afd4e70c686d6dc6d9b358819954155e8b8ba

Observation 353bcb8a-68f2-4f89-85f2-ebea9ed390cf · outbound

This paper cites Using tf-idf to determine word relevance in document queries.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Using tf-idf to determine word relevance in document queries

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:23.877765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.507552Z digest=sha256:cb9a8336980977eea8bc7fac10d9d271f37882bdddb291636cf93d21538d30de

Observation 77607fbd-f9cc-4b54-9e84-827390f386e8 · outbound

This paper cites CodeBLEU: a Method for Automatic Evaluation of Code Synthesis.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs CodeBLEU: a Method for Automatic Evaluation of Code Synthesis

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:52.586166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:52.586166Z digest=sha256:3036d0cbdaba997cae3f493346e78b43d51dfe78e675c8e9ff9b7bfd61354fe4

Observation 0aea7182-b91b-4944-a0e6-0b80e6e140b3 · outbound

This paper cites The perceptron: a probabilistic model for information storage and organization in the brain.Psychological review, 65(6):386, 1958.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs The perceptron: a probabilistic model for information storage and organization in the brain.Psychological review, 65(6):386, 1958

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:52.664974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:52.664974Z digest=sha256:d9030c20b10efa7cc2d355a4f0379166fe1fa5a0b1a7ede7b0e156c1ce3d8013

Observation 01af329a-b3ef-47a6-bfcf-632bd019bb96 · outbound

This paper cites FaceForensics: A Large-scale Video Dataset for Forgery Detection in Human Faces.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs FaceForensics: A Large-scale Video Dataset for Forgery Detection in Human Faces

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:52.729191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:52.729191Z digest=sha256:e1e159cd289131916f18293faa19f2f643a714b205819c6b16cd344b0f142f93

Observation dc0d9a15-6988-4450-9b53-7e21e5f7b7c4 · outbound

This paper cites Can AI-Generated Text be Reliably Detected?.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Can AI-Generated Text be Reliably Detected?

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:52.803803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:52.803803Z digest=sha256:7601786cd6e5d8da9cd907242efa27d11985ce99f090245aaef95984804aa12c

Observation 2f655c91-4889-4576-9e70-790a4b834762 · outbound

This paper cites Automated detection of ai-obfuscated plagiarism in modeling assignments.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Automated detection of ai-obfuscated plagiarism in modeling assignments

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:23.647131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.878609Z digest=sha256:fd8e435f1648e3eb5955afe980f077969cb9152f22e135ebc2199c821d6d406a

Observation 7c3dd540-98d5-48df-8e2d-9f9244f49fba · outbound

This paper cites Between lines of code: Unraveling the distinct patterns of machine and human programmers.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Between lines of code: Unraveling the distinct patterns of machine and human programmers

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:23.470078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.959556Z digest=sha256:b94ed978411e73f6231b31302a9772987d84874bfcdd4ad18162e307b91db59c

Observation 0c1ca927-6b1f-407a-a316-2911d4455c3b · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Deep unsupervised learning using nonequilibrium thermodynamics

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:23.301606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.995946Z digest=sha256:e608665dfb12be7f3f48c3e8577b90e7c56965f447d5321f45662a695d7b1890

Observation 78b2e5a4-a0c9-442b-8f13-4145bc6d8800 · outbound

This paper cites 2024 Stack Overflow Developer Survey, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs 2024 Stack Overflow Developer Survey, 2024

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:23.023700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.024513Z digest=sha256:b467a8a380883a89ad4c4994c6c4099d97c42f91c6fa65ca23393111d42ea6a1

Observation c3bd3099-8ece-4dc6-bf6f-a010a95a4eb1 · outbound

This paper cites Developer Perspectives on Licensing and Copyright Issues Arising from Generative AI for Software Development.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Developer Perspectives on Licensing and Copyright Issues Arising from Generative AI for Software Development

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:14.040744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:14.040744Z digest=sha256:8ef96fae7b93a6387da73e63cdb8aaa3fd2a7d69338b2dd7bad293c3f3a78641

Observation 3c1c00e4-1d96-44f0-bb43-aa7ee7d7cd28 · outbound

This paper cites Plagiarism in ai empowered world.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Plagiarism in ai empowered world

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:22.748784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.056712Z digest=sha256:4ebd97741f4d92039d1b69ec6f39e3edfb79045c5fff35af72202b44c6c6c889

Observation ee171013-724e-4eb6-a7ac-0ff12e27303b · outbound

This paper cites An empirical study on automatically detecting ai-generated source code: How far are we? InInternational Conference on Software Engineering (ICSE), 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs An empirical study on automatically detecting ai-generated source code: How far are we? InInternational Conference on Software Engineering (ICSE), 2025

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:22.569105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.076339Z digest=sha256:3249ec3529b2f6b2e4a3742e32c0ec5f7aef1266df78d117ea8dc221dd2d6186

Observation 0fa2b0ba-a5c4-4881-8712-9fa7925527eb · outbound

This paper cites Bugs in large language models generated code: An empirical study.Empirical Software Engineering, 30(3):1–48, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Bugs in large language models generated code: An empirical study.Empirical Software Engineering, 30(3):1–48, 2025

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:22.322393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.096731Z digest=sha256:7540062f18f1249b54ee85ea2ca626e817920bce445f562ce350cbe0cdc9aa33

Observation ff6127b3-03dc-40ee-8637-e058f3d68f20 · outbound

This paper cites How secure is ai-generated code: a large-scale comparison of large language models.Empirical Software Engineering, 30(2):1–42, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs How secure is ai-generated code: a large-scale comparison of large language models.Empirical Software Engineering, 30(2):1–42, 2025

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:22.045275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.118637Z digest=sha256:723a5084f361f245522044293fedaf7785009789c523a1f63db9a496aa006cf3

Observation 6587dbf4-3410-4e7e-8012-c34f119d7781 · outbound

This paper cites Llms in web development: Evaluating llm-generated php code unveiling vulnerabilities and limitations.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Llms in web development: Evaluating llm-generated php code unveiling vulnerabilities and limitations

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.795506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.179810Z digest=sha256:34f32a5762056248f42988bd149082fb2895311c2964ee669e5b4f81203fa90a

Observation 281323d5-087a-459e-a11e-f0f9bc864f63 · outbound

This paper cites Turingbench: A benchmark environ- ment for turing test in the age of neural text generation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Turingbench: A benchmark environ- ment for turing test in the age of neural text generation

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.551090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.217593Z digest=sha256:e310e65a86237889a123d888c1a5307bef543fe4ab2e33f1e3f5abf9d3c112e9

Observation be9e6b2a-63ed-4c8d-abd5-81e52907e025 · outbound

This paper cites A critical look at ai-generate software: Coding with the new ai tools is both irresistible and dangerous.IEEE Spectrum, 60(7):34–39, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs A critical look at ai-generate software: Coding with the new ai tools is both irresistible and dangerous.IEEE Spectrum, 60(7):34–39, 2023

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.423137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.277784Z digest=sha256:dd09acd892080e5486eb5370cb1ce05048993e1b2ec4a777516c80447e9dc6aa

Observation b88f5d05-c87f-4528-b0b1-f8c8780fe4c7 · outbound

This paper cites Attention is all you need.Advances in Neural Information Processing Systems (NeurIPS), 30, 2017.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Attention is all you need.Advances in Neural Information Processing Systems (NeurIPS), 30, 2017

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.314436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.310837Z digest=sha256:1c1802ebe71e1c3d09fd3d06b45bf998ab299302fe3cff6dbd1471f919a0e236

Observation 9d273f42-c329-43ba-bd73-61fc2f03955b · outbound

This paper cites Is Your AI-Generated Code Really Safe? Evaluating Large Language Models on Secure Code Generation with CodeSecEval.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Is Your AI-Generated Code Really Safe? Evaluating Large Language Models on Secure Code Generation with CodeSecEval

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:14.361719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:14.361719Z digest=sha256:e29be3f47cc38c0ff3fba035b6b07bec88f2bcb16c68510aac25e8fdac15f407

Observation 5a4c4378-4bcb-4f06-b2ce-31c15e7b716b · outbound

This paper cites Openhands: An open platform for ai software developers as generalist agents.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Openhands: An open platform for ai software developers as generalist agents

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.222402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.456862Z digest=sha256:0d6038e0aa297d2268c3c8e64522d3dd36aa3cd9e81c37f8f9db4ba0641049a5

Observation d7b5277a-0773-4316-95da-07206b479896 · outbound

This paper cites Codet5+: Open code large language models for code understanding and generation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Codet5+: Open code large language models for code understanding and generation

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.117660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.546271Z digest=sha256:6ec5deb3753005dff6b2d793955cd0a1da1a2cc2a39ed02574c47fcd33cf6c34

Observation 75d1735f-1810-49cb-ab70-67083d7d1190 · outbound

This paper cites A new era of plagiarism the danger of cheating using ai.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs A new era of plagiarism the danger of cheating using ai

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.003759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.634932Z digest=sha256:24aaeb95bcffccca8f92c964e0402c17c50d5476cbb5e570fc4e99124f6d503f

Observation 17920c00-40ea-4080-908b-2bac02b2510b · outbound

This paper cites Do llms know to respect copyright notice? InConference on Empirical Methods in Natural Language Processing (EMNLP), pages 20604–20619, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Do llms know to respect copyright notice? InConference on Empirical Methods in Natural Language Processing (EMNLP), pages 20604–20619, 2024

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.853496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.723234Z digest=sha256:72f6592b0bbe78f34a55a447bcdc32ec8a69ce088b28aecd703b531ecea5b80a

Observation 4dbac241-c745-401e-a5f6-e9b4a06dcd99 · outbound

This paper cites One Size Does Not Fit All: Investigating Efficacy of Perplexity in Detecting LLM-Generated Code.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs One Size Does Not Fit All: Investigating Efficacy of Perplexity in Detecting LLM-Generated Code

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:14.798043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:14.798043Z digest=sha256:717a7349943c0736cc29d8132dd62866cab3427ef2de3f9df4dd5088a533c485

Observation e1ac7c31-4310-47fd-85cb-983ba9055911 · outbound

This paper cites LiCoEval: Evaluating LLMs on License Compliance in Code Generation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs LiCoEval: Evaluating LLMs on License Compliance in Code Generation

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:14.880420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:14.880420Z digest=sha256:b5599785d550a78c9a4852aea7cbc5439c3c92e3569114e0659c92f36e09c830

Observation 6e66fc8d-f6ad-4b95-b5bb-5793c48afe27 · outbound

This paper cites Distin- guishing llm-generated from human-written code by contrastive learning.ACM Transactions on Software Engineering and Methodology, 34(4):1–31, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Distin- guishing llm-generated from human-written code by contrastive learning.ACM Transactions on Software Engineering and Methodology, 34(4):1–31, 2025

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.760397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.956953Z digest=sha256:2504b696c2af11d37316361d7d9e153b72e51b3b99a6fd0f897c79cb0488082c

Observation 559512f9-8e72-460e-8ef4-c0308bf77398 · outbound

This paper cites Detecting ai-generated code assignments using perplexity of large language models.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Detecting ai-generated code assignments using perplexity of large language models

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.608036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.024786Z digest=sha256:0822b92874b0540842d4ce93c79a56b78bae6d7df164a62368bf03c0ad84e771

Observation 1a0ffe97-c900-4d1d-a030-c9b766ecbe73 · outbound

This paper cites An {LLM-Assisted}{Easy-to-Trigger} backdoor attack on code completion models: Injecting disguised vulnerabilities against strong detection.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs An {LLM-Assisted}{Easy-to-Trigger} backdoor attack on code completion models: Injecting disguised vulnerabilities against strong detection

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.428612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.059826Z digest=sha256:b49a341fe6db82ceb36aea806a20a554aea86cde5a08d73e0d8dba21ebd3174d

Observation 646f7d6b-37ce-431e-8a7a-e7e6bf2dcc18 · outbound

This paper cites Zero-Shot Detection of Machine-Generated Codes.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Zero-Shot Detection of Machine-Generated Codes

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:15.071294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:15.071294Z digest=sha256:5dbc0437500cd68a5534a99470209d6feb8c2464871c9eaa46d9c13339e51552

Observation 98df593a-6b3d-46a5-be04-6eb4eac9b636 · outbound

This paper cites Uncovering llm-generated code: A zero-shot synthetic code detector via code rewriting.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Uncovering llm-generated code: A zero-shot synthetic code detector via code rewriting

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.249361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.099323Z digest=sha256:0c3e7f1186424c324fc87ce8b42de12f59c0542285c2a184e40def03de8db9bb

Observation 6b711f23-85f8-44f8-929f-41adc1564bf7 · outbound

This paper cites Codeipprompt: intellectual property infringement assessment of code language models.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Codeipprompt: intellectual property infringement assessment of code language models

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.053757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.206200Z digest=sha256:e64ef6478d923e8f0975049132ea95e17edc27d8a0dd5d38eb9531026e3e518a

Observation 16b48c01-9f40-41ff-888e-384a39c9a208 · outbound

This paper cites Inducing Vulnerable Code Generation in LLM Coding Assistants.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Inducing Vulnerable Code Generation in LLM Coding Assistants

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:15.340847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:15.340847Z digest=sha256:2ffc4ac4a2b33e28b5030e6e8e32e4ce77d9de9530d633d56401b11344e4c5a0

Observation c76f7210-55c6-4567-aaf4-cb8d912ee442 · outbound

This paper cites How well does llm generate security tests?arXiv preprint arXiv:2310.00710, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs How well does llm generate security tests?arXiv preprint arXiv:2310.00710, 2023

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:15.407231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:15.407231Z digest=sha256:b7552ee4258c4fd35bf192928a7d88844efbd39b177ba3290894e684eb2cda25

Observation 31be87fc-8af6-437f-815a-8d84cee8a050 · outbound

This paper cites Genimage: A million-scale benchmark for detecting ai-generated image.Advances in Neural Information Processing Systems (NeurIPS), 36:77771–77782, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Genimage: A million-scale benchmark for detecting ai-generated image.Advances in Neural Information Processing Systems (NeurIPS), 36:77771–77782, 2023

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:19.823765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.510378Z digest=sha256:c6efa59011911f71e3a2a547de9416871b195724ad845d2a4a0953b4e4ea8874

Observation fa6ba879-fa02-4f20-a6f2-f0e7caf8aec1 · outbound

This paper cites Wilddeepfake: A challenging real-world dataset for deepfake detection.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Wilddeepfake: A challenging real-world dataset for deepfake detection

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:19.529874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.611545Z digest=sha256:253e6a4f31fe3a1399c04ad445cc4b01080f23d400fd7f1b8073b73ce0bab3fa

Observation 3b4e48ee-263b-4fa0-ba3e-0eda070ab4b1 · outbound

This paper cites an unresolved cited work.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:19.259887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.642841Z digest=sha256:70b6e16de1e2dbee33ae079fbd797234c78c7f5a93b312076b2df68569d30871

Observation 83acb047-23d9-4acc-b227-7eb9507bc60f · outbound

This paper cites an unresolved cited work.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Unresolved cited work

Reference 100

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:19.095212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.691408Z digest=sha256:a04cfc4d1dc49fb5427ca35439d6200d79e56a727439d297a1d19f6297ddd0e7

Pith citing papers

Observation 1ed2e67d-6f24-4557-8c0c-03900d823793 · inbound

I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution cites this paper.

I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs

Reference 2025

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T23:54:56.716112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:54:51.091578Z digest=sha256:747a91bf90155d44aab20eb39005337b722180361e2fd7328a696751dfafcd94