Pith. sign in

Paper Citation Record · LEDGER

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs

As of 8 August 2026, this Paper Citation Record lists 100 of 111 outbound references and 1 inbound Pith citation observation for arXiv:2506.11059.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11059 v1

Coverage vector

measured 100 of 111 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:55:15.691408Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:54:51.091578Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T23:54:56.588252Z

Reference resolution

100 of 111 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved59
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 65788c0f-6182-4abf-b3a8-2ea9b74c138c · outbound

This paper cites GPT-4 Technical Report.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.476756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.476756Z digest=sha256:5df9e050d0ba82eefbb5142bbfcb52497b136f545da01c30bd0c6b96fc6b8177

Observation 75df5980-73d5-411e-8495-bbdb621b2640 · outbound

This paper cites An Empirical Study of AI Generated Text Detection Tools.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs An Empirical Study of AI Generated Text Detection Tools

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.570106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.570106Z digest=sha256:4a852474722d6f889ce638ccb705d88114cfd223c69739575ed878f406793920

Observation 908da8b7-9eeb-4478-9439-254c9cab0d7e · outbound

This paper cites Introducing computer use, a new claude 3.5 sonnet, and claude 3.5 haiku, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Introducing computer use, a new claude 3.5 sonnet, and claude 3.5 haiku, 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.633741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.633741Z digest=sha256:9fb801d1884a2b861279014d3e504d9d96d332c4095d02bbae831c6e8c2937cf

Observation a7846d60-9b79-475a-9822-16282d44c1d2 · outbound

This paper cites Claude 3.7 Sonnet and Claude Code, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Claude 3.7 Sonnet and Claude Code, 2025

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.720560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.720560Z digest=sha256:87757921a2994d1b485f84a350a48d6fb8f4190f8a9991fa7cb6f47d6a406155

Observation c5a100e3-a86e-4296-89d1-260fca603978 · outbound

This paper cites Is github’s copilot as bad as humans at introducing vulnerabilities in code?Empirical Software Engineering, 28(6):129, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Is github’s copilot as bad as humans at introducing vulnerabilities in code?Empirical Software Engineering, 28(6):129, 2023

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.811473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.811473Z digest=sha256:c6ffe9d2e069c43d89a12fdb9b7d1f412f463579e47740fff02156c22c17b150

Observation c0a0349f-1f09-4419-bb99-36ab95a46fc6 · outbound

This paper cites Random forests.Machine learning, 45:5–32, 2001.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Random forests.Machine learning, 45:5–32, 2001

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.895270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.895270Z digest=sha256:a753b62d41526632275a9c527c5bcffea3a5766cf76443cd28f6786eab00b1a0

Observation f391913e-6860-425b-8fa2-3debfa1511e7 · outbound

This paper cites Membership inference attacks from first principles.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Membership inference attacks from first principles

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:46.976947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:46.976947Z digest=sha256:9b771366cdcb4731fbf6ecaf7d072362dfa72615aa784afbd7d354988e7934ab

Observation 4e3022a3-1094-44a2-af28-e5090a1fb735 · outbound

This paper cites R package version 3.0.1.1.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs R package version 3.0.1.1

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.063202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.063202Z digest=sha256:b9d538f59ae1fb2563952ddf20b1565f931b0b570712833799920ad70af8d6a1

Observation 553cb2d9-df04-4555-913c-b75ef7780dab · outbound

This paper cites Github code clean dataset, 2022.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Github code clean dataset, 2022

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.129620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.129620Z digest=sha256:0025a9b5c5e6954a2a671a05ed9dae1a02b13c38250b76ba12e9d4fd20cf8c15

Observation 02afa950-9b55-48b6-b04f-744abf07e8c5 · outbound

This paper cites Github code dataset, 2022.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Github code dataset, 2022

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.200058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.200058Z digest=sha256:fcc7d95f996469a2c31ead16eaaad0918495145ab4605cf32fd031eb553f1c67

Observation f25dda91-c58c-43df-8202-91a4e16cfd1a · outbound

This paper cites Vulnerabilities in ai code generators: Exploring targeted data poisoning attacks.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Vulnerabilities in ai code generators: Exploring targeted data poisoning attacks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.294524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.294524Z digest=sha256:b030a5e430ca0aae1a437270c5548debd58b38e5917c664a0372f696b2a2a4dc

Observation 29e82846-deea-4d38-bd2e-2d9edfe7f36b · outbound

This paper cites Cursor: The AI Code Editor, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Cursor: The AI Code Editor, 2023

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.393332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.393332Z digest=sha256:d5406ba5b3b9ee9a625f5c9ae66603596899881c2d8a69d24ec03624ae585152

Observation 1b0a11d0-a7f0-462f-8638-6c5988df77d1 · outbound

This paper cites Plagiarism in the age of massive generative pre-trained transformers (gpt-3).Ethics in Science and Environmental Politics, 21:17–23, 2021.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Plagiarism in the age of massive generative pre-trained transformers (gpt-3).Ethics in Science and Environmental Politics, 21:17–23, 2021

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.467782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.467782Z digest=sha256:393f51d8d5797a36b9888152477372b44c9319b05bb800208e61600a9f63e419

Observation 9338dc5a-b143-4bc3-821b-57e1450a2458 · outbound

This paper cites AIGCodeSet: A New Annotated Dataset for AI Generated Code Detection.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs AIGCodeSet: A New Annotated Dataset for AI Generated Code Detection

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.550790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.550790Z digest=sha256:dc4e6a50f1473d1ee94412bf9766d1d1a8ec4a40a0dea922d453c915ea370fee

Observation d7d6f9c8-d5dd-4eab-8e0f-659fe2ae56a6 · outbound

This paper cites The DeepFake Detection Challenge (DFDC) Dataset.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs The DeepFake Detection Challenge (DFDC) Dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.640002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.640002Z digest=sha256:ef8b904f8579c210e644f0826cba007e205afacc690a531739a7d254082164be

Observation 8d87665b-4f0b-46c3-8a0e-7128c09be92f · outbound

This paper cites Codep: grammatical seq2seq model for general-purpose code generation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Codep: grammatical seq2seq model for general-purpose code generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.701676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.701676Z digest=sha256:f076d689893014abcd66a2b8477986f1e9663fd13317347e2078987035fa2def

Observation 08fb92ab-8479-4797-be7c-5184de1e9f47 · outbound

This paper cites JaCoText: A Pretrained Model for Java Code-Text Generation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs JaCoText: A Pretrained Model for Java Code-Text Generation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:55:17.120274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:47.773655Z digest=sha256:7b261cde10ec518e8dfc84a3fc81065c46908b07c03ba736571c4fc390fe10a8

Observation cc471623-5e94-487d-a09f-a0dd57b31ff3 · outbound

This paper cites Out of the bleu: how should we assess quality of the code generation models?Journal of Systems and Software, 203:111741, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Out of the bleu: how should we assess quality of the code generation models?Journal of Systems and Software, 203:111741, 2023

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.862883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.862883Z digest=sha256:108a1414b91b258fe068d798974fa4a7d818a32bb4975fdcd5c5765e549a553d

Observation 11467041-3d7d-44fc-acf8-15269374ff63 · outbound

This paper cites CodeBERT: A Pre-Trained Model for Programming and Natural Languages.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs CodeBERT: A Pre-Trained Model for Programming and Natural Languages

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:47.920730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:47.920730Z digest=sha256:58db038a4f32f5be5190ee4e841852e915d856e01f092ef2123caa0432cafd32

Observation 21847269-8275-48a4-b353-0e7bb165121f · outbound

This paper cites Introducing GitHub Copilot: your AI pair programmer, 2022.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Introducing GitHub Copilot: your AI pair programmer, 2022

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.011427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.011427Z digest=sha256:0a31a134e41208b0df9d8220cc1d403912f963e05c9659cb20af325831facc8c

Observation d7e6ddef-b792-4499-8ca8-05d7ac0df019 · outbound

This paper cites What makes good in-context demonstrations for code intelligence tasks with llms? InIEEE/ACM International Conference on Automated Software Engineering (ASE), pages 761–773, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs What makes good in-context demonstrations for code intelligence tasks with llms? InIEEE/ACM International Conference on Automated Software Engineering (ASE), pages 761–773, 2023

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.070030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.070030Z digest=sha256:85ef045b85eda6f227d09755970b3881e5100a8998cf4614973201036fa95078

Observation 1528c165-1662-4f98-a82a-3165eaa55899 · outbound

This paper cites Gltr: Statistical detection and visualization of generated text.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Gltr: Statistical detection and visualization of generated text

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.124138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.124138Z digest=sha256:d6753506aed49f112b4f5344df7251a0a1f2d091e6838fca7ad6d3965df2ed77

Observation 314297cd-8c44-4ef9-b4a0-4a1e554b2176 · outbound

This paper cites A survey on the possibilities & impossibilities of ai-generated text detection.Transactions on Machine Learning Research (TMLR), 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs A survey on the possibilities & impossibilities of ai-generated text detection.Transactions on Machine Learning Research (TMLR), 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.246520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.246520Z digest=sha256:477f1402296ee669f01d1bcd43dce2cbd6e58d13d92e7d9d75004137cc5cac8e

Observation b5aac917-e193-4767-ad10-2fbec12deecd · outbound

This paper cites Deepfake video detection using recurrent neural networks.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Deepfake video detection using recurrent neural networks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.359564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.359564Z digest=sha256:4154e3f67050bd5114284c7b65d72420257ace53d8577cf18ebe0f9d2b4a2cdc

Observation f13457b0-5431-42a2-b472-218cb58cd458 · outbound

This paper cites Graphcodebert: Pre-training code representations with data flow.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Graphcodebert: Pre-training code representations with data flow

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.449654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.449654Z digest=sha256:a7482614826609a003530c6265e19f1c16ec48c461fff6ba0e8d70867d2a0852

Observation 05003157-91eb-4798-8360-04d2c36081b1 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.545437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.545437Z digest=sha256:bda770b616de9954c504d44567e8420dd83211c3462fea43e73b08967ce0dde8

Observation eb20765c-9e15-4507-8edc-b618ad05b088 · outbound

This paper cites Biscope: Ai-generated text detection by checking memorization of preceding tokens.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Biscope: Ai-generated text detection by checking memorization of preceding tokens

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.665601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.665601Z digest=sha256:2829f92b4ec0705c1a8a705ad3af154e6a6f2ed302b84bdf1fe938f11f0c7a44

Observation c66b5304-5d20-4b43-abe3-3212d8cf7a90 · outbound

This paper cites Spotting llms with binoculars: Zero-shot detection of machine-generated text.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Spotting llms with binoculars: Zero-shot detection of machine-generated text

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.809631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.809631Z digest=sha256:43480d8632090f645314c033f5113c53074b3203e656fa9a66baf856e2b28ddc

Observation 5195c632-aa5c-4485-855a-96398f705a7d · outbound

This paper cites Denoising diffusion probabilistic models.Advances in Neural Information Processing Systems (NeurIPS), 33:6840–6851, 2020.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Denoising diffusion probabilistic models.Advances in Neural Information Processing Systems (NeurIPS), 33:6840–6851, 2020

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.908202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.908202Z digest=sha256:fc422a488642314dc01cfb9be6c8deab1dac8baff4d42d2d325d99b2d1d2de17

Observation 83737ca2-14d0-46cf-85c5-01beb47d2d14 · outbound

This paper cites Qwen2.5-Coder Technical Report.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Qwen2.5-Coder Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:48.989561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:48.989561Z digest=sha256:f1679f92359c5e4f14cb3fc245a6974e87243732df63b826d2cef5cac3790ae0

Observation 77d0cb4e-9f09-4066-ab9d-84f237151c12 · outbound

This paper cites Rethinking plagiarism in the era of generative ai.Journal of Intelligent Communication, 3(2):20–31, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Rethinking plagiarism in the era of generative ai.Journal of Intelligent Communication, 3(2):20–31, 2024

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.120454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.120454Z digest=sha256:d02fa98314c99d78fdcd68ee8096c9a52f77b09bbe2f58650f349331a1fdeddf

Observation 3207f12b-4c39-4ebc-951d-9d6ed0e8ceb6 · outbound

This paper cites Whodunit: Classifying code as human authored or gpt-4 generated-a case study on codechef problems.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Whodunit: Classifying code as human authored or gpt-4 generated-a case study on codechef problems

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.242525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.242525Z digest=sha256:713461709dbf878eb3fa533ec61d65df447297f9fbca35a3d0003a02fdec4e02

Observation bc477be6-3ca6-452d-a785-58be0314634a · outbound

This paper cites Automatic detection of generated text is easiest when humans are fooled.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Automatic detection of generated text is easiest when humans are fooled

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.340671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.340671Z digest=sha256:27ad6a4c7ebf0acae1e8b1628fba76455ebde4cf23ffba73a6f92b45623e927e

Observation c4238f08-8892-4f25-8231-5600621c9da7 · outbound

This paper cites Swe-bench: Can language models resolve real-world github issues? InInternational Conference on Learning Representations (ICLR), 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Swe-bench: Can language models resolve real-world github issues? InInternational Conference on Learning Representations (ICLR), 2024

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.537272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.537272Z digest=sha256:89ac641a47296444f9ec4a28062b40e1ea417b6a48c2a8d471c096d09c487e4e

Observation bae908c6-127e-4bda-a198-82489ae72e37 · outbound

This paper cites Access the latest 2.0 experimental models in the gemini app., 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Access the latest 2.0 experimental models in the gemini app., 2025

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.661149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.661149Z digest=sha256:a7e3cb847ea30fed7e6bcdc9b73cac4c5bcc909e1041526e7bb5b9046dddedc1

Observation 605bc256-bacb-4710-8e72-c6021ee8bb78 · outbound

This paper cites Vulnerability handling of ai- generated code-existing solutions and open challenges.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Vulnerability handling of ai- generated code-existing solutions and open challenges

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.780282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.780282Z digest=sha256:970207ea2b54091f3d68ce9cab35cfcca60a380f828a99b6d75102ece86832e2

Observation d508cc1b-25f4-4f13-bd1b-d3a77ffe37be · outbound

This paper cites Gemini 2.0 is now available to everyone, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Gemini 2.0 is now available to everyone, 2025

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:49.898415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:49.898415Z digest=sha256:2ee20eff9f4bf18ab24d260634f059f340b0c775d7a02652416d13b35bc57f66

Observation dbda737d-454c-453b-ba59-b83e84ff4e8e · outbound

This paper cites Does attitude towards plagiarism predict aigiarism using chatgpt?AI and Ethics, 5(1):677–688, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Does attitude towards plagiarism predict aigiarism using chatgpt?AI and Ethics, 5(1):677–688, 2025

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:50.058627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:50.058627Z digest=sha256:a8876e5e4d31e744789bd9670e961f71c606470aba4d86bb52fadf9eed0fe505

Observation ef0bb54f-4125-47f8-89f4-f384d8462faf · outbound

This paper cites Will chatgpt g et you caught? rethinking of plagiarism detection.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Will chatgpt g et you caught? rethinking of plagiarism detection

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:50.183926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:50.183926Z digest=sha256:5c40764da1649097c96b8c638e816c758dee8974c78b519e3600ad0f15c77e83

Observation 64ffc721-c503-4bb6-a5a5-533ecc3ba99f · outbound

This paper cites How secure is code generated by chatgpt? InIEEE international conference on systems, man, and cybernetics (SMC), pages 2445–2451.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs How secure is code generated by chatgpt? InIEEE international conference on systems, man, and cybernetics (SMC), pages 2445–2451

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:26.240995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:50.269865Z digest=sha256:26f103c804c976904f01d97cae88f6b7ad63879bbe2d45e5e34d2676147077d2

Observation 03f30dd7-d73a-47ee-bd15-91d0063106e0 · outbound

This paper cites Paraphrasing evades detectors of ai-generated text, but retrieval is an effective defense.Advances in Neural Information Processing Systems (NeurIPS), 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Paraphrasing evades detectors of ai-generated text, but retrieval is an effective defense.Advances in Neural Information Processing Systems (NeurIPS), 2023

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:26.074964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:50.386281Z digest=sha256:6935645ce1a3843a250f4f048b3f37ee5f61b10668718d536c69bb558281bcfe

Observation 4271435f-c3b5-455a-8867-637eaf646188 · outbound

This paper cites Detecting fake content with relative entropy scoring.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Detecting fake content with relative entropy scoring

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.908463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:50.553406Z digest=sha256:f849dc549af76f04f67ecdf2890baa315ae7a9dcee41feb15c9551cc560d4625

Observation 50499960-1047-4a1f-b5a3-ade3debd7f65 · outbound

This paper cites Protecting intellectual property of large language model-based code generation apis via watermarks.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Protecting intellectual property of large language model-based code generation apis via watermarks

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.821311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:50.664897Z digest=sha256:e0a1d0b6b7a252016d253efb2695536ccbf3e10c93b6557d9881c99163c6bef9

Observation cabc3883-4bb6-4818-80ea-24f5fff2cfc9 · outbound

This paper cites DeepSeek-V3 Technical Report.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs DeepSeek-V3 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:50.744812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:50.744812Z digest=sha256:01af8e09cd9cb34ac85282c520402590bd8bf78a439f595b8ac43da1241cd232

Observation 1da2f7b4-1105-407c-a195-1bc12dcf90ad · outbound

This paper cites an unresolved cited work.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:25.697427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:50.865882Z digest=sha256:ffa5c968b64ad12b64aa7d4de97ea487e1b9906a3efacf36cbe6dda1b3294f9e

Observation 04279a04-6e3d-4f41-8e20-d41d409b3ac1 · outbound

This paper cites CodeXEmbed: A Generalist Embedding Model Family for Multiligual and Multi-task Code Retrieval.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs CodeXEmbed: A Generalist Embedding Model Family for Multiligual and Multi-task Code Retrieval

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.004303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.004303Z digest=sha256:0caf8288ee11f043e89520c6d78ff40cae5ed03f560ae3603fa1bf7306ad684f

Observation 2ce6caf2-b835-4f6e-80a1-10112003ecd9 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.120028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.120028Z digest=sha256:23dfaeb71e3f4ab799c8ff546f2f78c9dbede96d69cc3ed5af43998f8f08b7be

Observation c6f9638a-d993-4bd4-bf4d-944c3afec798 · outbound

This paper cites Raidar: generative ai detection via rewriting.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Raidar: generative ai detection via rewriting

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.537792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.218996Z digest=sha256:e5b5b2785d8dd55fbc86935d7f403839506768c771a50cd1d0f5dd1f57f1e026

Observation db5bea98-ac35-408c-8e6a-9907e16e7a86 · outbound

This paper cites On the robustness of code generation techniques: An empirical study on github copilot.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs On the robustness of code generation techniques: An empirical study on github copilot

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.415448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.348636Z digest=sha256:3ecbceb2fa3128429fe0377587cb414b34ce7730bb6668fe1cb801e0c1fd1597

Observation 47c2b334-6ec8-4be7-84ef-3bb770d4ba7f · outbound

This paper cites Llama 3.3: Model cards & prompt formats, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Llama 3.3: Model cards & prompt formats, 2024

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.275821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.489432Z digest=sha256:ec170643c7f14d074c4d844466d493c5c1f03c22e4ee4ff2d18ed32059b68680

Observation a1845f18-307f-48db-99d2-ef1065386dff · outbound

This paper cites Detectgpt: Zero-shot machine-generated text detection using probability curvature.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Detectgpt: Zero-shot machine-generated text detection using probability curvature

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:25.167755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.641080Z digest=sha256:078ac881ee5c880fb87882a72f350a2a61ed7ee8a4bd6f2e9bd9142d9257fef4

Observation 369ee951-331b-4b17-a544-0fc92734f160 · outbound

This paper cites Is this Snippet Written by ChatGPT? An Empirical Study with a CodeBERT-Based Classifier.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Is this Snippet Written by ChatGPT? An Empirical Study with a CodeBERT-Based Classifier

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.700600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.700600Z digest=sha256:b61f6f512152d209a6e104bdd804f36dbd46d34a78d7803d9067abdd43818104

Observation e054e314-9f13-4178-bf36-13dc638c012a · outbound

This paper cites Gptsniffer: A codebert-based classifier to detect source code written by chatgpt.Journal of Systems and Software, 214:112059, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Gptsniffer: A codebert-based classifier to detect source code written by chatgpt.Journal of Systems and Software, 214:112059, 2024

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.981311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.759655Z digest=sha256:26885817f0beefe87d8460e6a3de6164f9fbd26a9a9b68a209c6d210179b5e73

Observation 5607cbcf-e488-48c7-945f-c638c8485b33 · outbound

This paper cites Poisoned chatgpt finds work for idle hands: Exploring developers’ coding practices with insecure suggestions from poisoned ai models.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Poisoned chatgpt finds work for idle hands: Exploring developers’ coding practices with insecure suggestions from poisoned ai models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.839371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:51.838559Z digest=sha256:5b3ae6d7e0ea3df7fddd790df3834d06bd002e006d2c45bae839b72ca9be5dfc

Observation 16b5d11e-f0ca-4610-be35-b2a5a824779c · outbound

This paper cites Introducing ChatGPT, 2022.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Introducing ChatGPT, 2022

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.919051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.919051Z digest=sha256:7a93bf23da63d1935e6e5b7edf13830a8a5beec0fc0dad70f83940404e3a1760

Observation 9e47c2de-2fd7-4582-a635-8af1d6a61da5 · outbound

This paper cites Gpt-4o mini: advancing cost-efficient intelligence, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Gpt-4o mini: advancing cost-efficient intelligence, 2024

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.992202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.992202Z digest=sha256:3d70a4d984e7528e270b20135f076d6c51c5de662929fc8db72dd8e21f3b7710

Observation adbea689-e8b3-492f-b51e-222122821c97 · outbound

This paper cites Openai o3-mini: Pushing the frontier of cost-effective reasoning, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Openai o3-mini: Pushing the frontier of cost-effective reasoning, 2025

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.666988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.043090Z digest=sha256:6412c36472b073c417c291f1d76d442effa492acbc5fac9be71149f7ba18bed2

Observation 3f3a22e2-6ddf-4c0d-8dc1-a910825615d8 · outbound

This paper cites CoDet-M4: Detecting Machine-Generated Code in Multi-Lingual, Multi-Generator and Multi-Domain Settings.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs CoDet-M4: Detecting Machine-Generated Code in Multi-Lingual, Multi-Generator and Multi-Domain Settings

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:52.108811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:52.108811Z digest=sha256:b4e96957a0a863582e04abe85dc74222dae5717c5fe751b433d051aa9d369e38

Observation a74c8129-f622-4c8e-855d-711bc48a174f · outbound

This paper cites Assessing ai detectors in identifying ai-generated code: Implications for education.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Assessing ai detectors in identifying ai-generated code: Implications for education

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.515159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.157680Z digest=sha256:f7655379cbab3676185463f2d941289b0dc2b11ddddfac4d4980dc1a3964ccd4

Observation 1b147bce-2743-4336-be30-1d54c61c8701 · outbound

This paper cites Bleu: A method for automatic evaluation of machine translation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Bleu: A method for automatic evaluation of machine translation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.371342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.221757Z digest=sha256:4a2ea973800016e3403507e367185f52d00cf554f15443f82045e3475c633452

Observation 68814478-2094-4fa8-9715-03362c8d9dfa · outbound

This paper cites Asleep at the keyboard? assessing the security of github copilot’s code contributions.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Asleep at the keyboard? assessing the security of github copilot’s code contributions

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.265456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.278896Z digest=sha256:08e3a7ace50b15189ac53e55168ca8c902c65e2758599717945408aadde5fba8

Observation b8253700-57d5-448a-87ba-54f616a19d6b · outbound

This paper cites Magecode: Machine- generated code detection method using large language models.IEEE Access, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Magecode: Machine- generated code detection method using large language models.IEEE Access, 2024

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.178275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.343700Z digest=sha256:136e5aa875117053d33827482572ac3ba1043b566cf358022864ad1a94111e7d

Observation b9be9d4a-a780-403d-b1ff-d1c738eadafe · outbound

This paper cites Introducing gemini 2.0: our new ai model for the agentic era, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Introducing gemini 2.0: our new ai model for the agentic era, 2024

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:24.021607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.428624Z digest=sha256:41f614e4563ece1c640481cd8456df6bf23844bb840f67da5a3ea31b9d560c11

Observation 353bcb8a-68f2-4f89-85f2-ebea9ed390cf · outbound

This paper cites Using tf-idf to determine word relevance in document queries.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Using tf-idf to determine word relevance in document queries

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:23.877765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.507552Z digest=sha256:bc66973fe0ed5cec8d9870ea51914df4620a859b59043e27cfa84b61d3348350

Observation 77607fbd-f9cc-4b54-9e84-827390f386e8 · outbound

This paper cites CodeBLEU: a Method for Automatic Evaluation of Code Synthesis.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs CodeBLEU: a Method for Automatic Evaluation of Code Synthesis

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:52.586166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:52.586166Z digest=sha256:609531f6d2c177c48b10708b669da884cc8feb766dc29abe1007db6b095e3117

Observation 0aea7182-b91b-4944-a0e6-0b80e6e140b3 · outbound

This paper cites The perceptron: a probabilistic model for information storage and organization in the brain.Psychological review, 65(6):386, 1958.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs The perceptron: a probabilistic model for information storage and organization in the brain.Psychological review, 65(6):386, 1958

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:52.664974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:52.664974Z digest=sha256:b83222e2e36f6c901f14ce648778fd89eb89ec617f45014051a9a5838f2dcac2

Observation 01af329a-b3ef-47a6-bfcf-632bd019bb96 · outbound

This paper cites FaceForensics: A Large-scale Video Dataset for Forgery Detection in Human Faces.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs FaceForensics: A Large-scale Video Dataset for Forgery Detection in Human Faces

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:52.729191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:52.729191Z digest=sha256:ef4cbf34d6ed6f7a3cbf4b3f39f47d68425d390148b279438f0aa08b28290e6f

Observation dc0d9a15-6988-4450-9b53-7e21e5f7b7c4 · outbound

This paper cites Can AI-Generated Text be Reliably Detected?.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Can AI-Generated Text be Reliably Detected?

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:52.803803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:52.803803Z digest=sha256:444509f3a200bbd8fa703846b7ce87a19ece5416bb16cf980366ca6ced1e0569

Observation 2f655c91-4889-4576-9e70-790a4b834762 · outbound

This paper cites Automated detection of ai-obfuscated plagiarism in modeling assignments.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Automated detection of ai-obfuscated plagiarism in modeling assignments

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:23.647131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.878609Z digest=sha256:ceb94086a8dc92a4aebda07459f79546d6a9c6cd508df9ce26c13ab08bfac67f

Observation 7c3dd540-98d5-48df-8e2d-9f9244f49fba · outbound

This paper cites Between lines of code: Unraveling the distinct patterns of machine and human programmers.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Between lines of code: Unraveling the distinct patterns of machine and human programmers

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:23.470078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.959556Z digest=sha256:b600b83ea23631c5b8e09bca9941bb5e3f05e4056dbd72569c2441ac0b3695de

Observation 0c1ca927-6b1f-407a-a316-2911d4455c3b · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Deep unsupervised learning using nonequilibrium thermodynamics

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:23.301606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:54:52.995946Z digest=sha256:f4016a475ee5075f16c02b9626500370eaf80a76901916f389efcdfffbdcc944

Observation 78b2e5a4-a0c9-442b-8f13-4145bc6d8800 · outbound

This paper cites 2024 Stack Overflow Developer Survey, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs 2024 Stack Overflow Developer Survey, 2024

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:23.023700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.024513Z digest=sha256:0da2e356b493a09ebe44993d7467ef2d7e447be89e7b29b45e08849f9ff4bcaa

Observation c3bd3099-8ece-4dc6-bf6f-a010a95a4eb1 · outbound

This paper cites Developer Perspectives on Licensing and Copyright Issues Arising from Generative AI for Software Development.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Developer Perspectives on Licensing and Copyright Issues Arising from Generative AI for Software Development

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:14.040744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:14.040744Z digest=sha256:58720810fc783884195b95390884764cfe145c6a15f8f5f45b7007901dc8c9da

Observation 3c1c00e4-1d96-44f0-bb43-aa7ee7d7cd28 · outbound

This paper cites Plagiarism in ai empowered world.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Plagiarism in ai empowered world

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:22.748784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.056712Z digest=sha256:f496a3bcb6e42c97ba81cf3fb98de8c68fc0ef2c254916839ceb76e1404ebb7e

Observation ee171013-724e-4eb6-a7ac-0ff12e27303b · outbound

This paper cites An empirical study on automatically detecting ai-generated source code: How far are we? InInternational Conference on Software Engineering (ICSE), 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs An empirical study on automatically detecting ai-generated source code: How far are we? InInternational Conference on Software Engineering (ICSE), 2025

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:22.569105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.076339Z digest=sha256:a83ec25c7c57c2e17d91d0e351e38c5e1c892936bf9cf8e710a57a063710c385

Observation 0fa2b0ba-a5c4-4881-8712-9fa7925527eb · outbound

This paper cites Bugs in large language models generated code: An empirical study.Empirical Software Engineering, 30(3):1–48, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Bugs in large language models generated code: An empirical study.Empirical Software Engineering, 30(3):1–48, 2025

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:22.322393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.096731Z digest=sha256:90bf4e65a49aa0562a4149fde23fd4c619fbc9790192ea659526b17719d54c97

Observation ff6127b3-03dc-40ee-8637-e058f3d68f20 · outbound

This paper cites How secure is ai-generated code: a large-scale comparison of large language models.Empirical Software Engineering, 30(2):1–42, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs How secure is ai-generated code: a large-scale comparison of large language models.Empirical Software Engineering, 30(2):1–42, 2025

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:22.045275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.118637Z digest=sha256:284ff240ad1e264dac2a1e57ecd991726da7ad1835795e1dad7dbf89fc99a4b9

Observation 6587dbf4-3410-4e7e-8012-c34f119d7781 · outbound

This paper cites Llms in web development: Evaluating llm-generated php code unveiling vulnerabilities and limitations.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Llms in web development: Evaluating llm-generated php code unveiling vulnerabilities and limitations

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.795506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.179810Z digest=sha256:81806a31c03cc502d8d54f0ba7d182b80ef648e609a7fbe31a09bf7243d493fd

Observation 281323d5-087a-459e-a11e-f0f9bc864f63 · outbound

This paper cites Turingbench: A benchmark environ- ment for turing test in the age of neural text generation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Turingbench: A benchmark environ- ment for turing test in the age of neural text generation

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.551090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.217593Z digest=sha256:75bb75d1e61fee497360ac3344f7e305881ea3c10933d35199708357de290470

Observation be9e6b2a-63ed-4c8d-abd5-81e52907e025 · outbound

This paper cites A critical look at ai-generate software: Coding with the new ai tools is both irresistible and dangerous.IEEE Spectrum, 60(7):34–39, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs A critical look at ai-generate software: Coding with the new ai tools is both irresistible and dangerous.IEEE Spectrum, 60(7):34–39, 2023

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.423137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.277784Z digest=sha256:6c20c65c9265c234a6dc13a5cc07b09274b55e859daea354c779b7945917bd15

Observation b88f5d05-c87f-4528-b0b1-f8c8780fe4c7 · outbound

This paper cites Attention is all you need.Advances in Neural Information Processing Systems (NeurIPS), 30, 2017.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Attention is all you need.Advances in Neural Information Processing Systems (NeurIPS), 30, 2017

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.314436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.310837Z digest=sha256:ca7e892a6cf7820f9a674c00b9cedad4cd18b89166b27f010f412ff4be3c64ee

Observation 9d273f42-c329-43ba-bd73-61fc2f03955b · outbound

This paper cites Is Your AI-Generated Code Really Safe? Evaluating Large Language Models on Secure Code Generation with CodeSecEval.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Is Your AI-Generated Code Really Safe? Evaluating Large Language Models on Secure Code Generation with CodeSecEval

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:14.361719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:14.361719Z digest=sha256:47ce7884749ea57d078ea88961c554d53babf50db9c897884800053d91921e2a

Observation 5a4c4378-4bcb-4f06-b2ce-31c15e7b716b · outbound

This paper cites Openhands: An open platform for ai software developers as generalist agents.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Openhands: An open platform for ai software developers as generalist agents

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.222402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.456862Z digest=sha256:33d738a70a04d7499d1c8e3f48d70d272ba04ab93a1b7188252d5ee7066056b1

Observation d7b5277a-0773-4316-95da-07206b479896 · outbound

This paper cites Codet5+: Open code large language models for code understanding and generation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Codet5+: Open code large language models for code understanding and generation

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.117660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.546271Z digest=sha256:18f2415ce4fb57b762f2248b18d76c1941490b908c14dc35bc4bbf421d6f1590

Observation 75d1735f-1810-49cb-ab70-67083d7d1190 · outbound

This paper cites A new era of plagiarism the danger of cheating using ai.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs A new era of plagiarism the danger of cheating using ai

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:21.003759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.634932Z digest=sha256:7a5b01c3fe3a8dbee5f6375d415b5a2bfe07943a874d767496a181152c89c150

Observation 17920c00-40ea-4080-908b-2bac02b2510b · outbound

This paper cites Do llms know to respect copyright notice? InConference on Empirical Methods in Natural Language Processing (EMNLP), pages 20604–20619, 2024.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Do llms know to respect copyright notice? InConference on Empirical Methods in Natural Language Processing (EMNLP), pages 20604–20619, 2024

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.853496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.723234Z digest=sha256:c543664da2c0c7fe039a7576595f7a995f809964fe54db57f53a99f4234250cb

Observation 4dbac241-c745-401e-a5f6-e9b4a06dcd99 · outbound

This paper cites One Size Does Not Fit All: Investigating Efficacy of Perplexity in Detecting LLM-Generated Code.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs One Size Does Not Fit All: Investigating Efficacy of Perplexity in Detecting LLM-Generated Code

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:14.798043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:14.798043Z digest=sha256:fe3a0b68f96f1a0b9a24996fa862a762e4858d194d699d486509fde4b6a897b1

Observation e1ac7c31-4310-47fd-85cb-983ba9055911 · outbound

This paper cites LiCoEval: Evaluating LLMs on License Compliance in Code Generation.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs LiCoEval: Evaluating LLMs on License Compliance in Code Generation

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:14.880420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:14.880420Z digest=sha256:8bcb966cd184152add591172779ac3d5e57cc18e4f6cce266c19bb41460e6612

Observation 6e66fc8d-f6ad-4b95-b5bb-5793c48afe27 · outbound

This paper cites Distin- guishing llm-generated from human-written code by contrastive learning.ACM Transactions on Software Engineering and Methodology, 34(4):1–31, 2025.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Distin- guishing llm-generated from human-written code by contrastive learning.ACM Transactions on Software Engineering and Methodology, 34(4):1–31, 2025

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.760397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:14.956953Z digest=sha256:65a47d38dae1fef715c6bbd39ef5c3505b833dcd9f48e154a879901d91c8b55e

Observation 559512f9-8e72-460e-8ef4-c0308bf77398 · outbound

This paper cites Detecting ai-generated code assignments using perplexity of large language models.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Detecting ai-generated code assignments using perplexity of large language models

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.608036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.024786Z digest=sha256:a47d4903f3c97cc00d536046f2a2c58f33dc2aeca188afca703e001a839993d5

Observation 1a0ffe97-c900-4d1d-a030-c9b766ecbe73 · outbound

This paper cites An {LLM-Assisted}{Easy-to-Trigger} backdoor attack on code completion models: Injecting disguised vulnerabilities against strong detection.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs An {LLM-Assisted}{Easy-to-Trigger} backdoor attack on code completion models: Injecting disguised vulnerabilities against strong detection

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.428612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.059826Z digest=sha256:8e969f4856e1a487296658bf8aeaeaeb7cd05d431c0a716208a7469865341963

Observation 646f7d6b-37ce-431e-8a7a-e7e6bf2dcc18 · outbound

This paper cites Zero-Shot Detection of Machine-Generated Codes.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Zero-Shot Detection of Machine-Generated Codes

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:15.071294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:15.071294Z digest=sha256:f505e63c58951ed044313e18a672efefd62b138dab66307ff574075589d86f60

Observation 98df593a-6b3d-46a5-be04-6eb4eac9b636 · outbound

This paper cites Uncovering llm-generated code: A zero-shot synthetic code detector via code rewriting.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Uncovering llm-generated code: A zero-shot synthetic code detector via code rewriting

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.249361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.099323Z digest=sha256:defa9ec70d2c566faa914f90db6c614f2790779c82ceb53517c7ca09aea1eaef

Observation 6b711f23-85f8-44f8-929f-41adc1564bf7 · outbound

This paper cites Codeipprompt: intellectual property infringement assessment of code language models.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Codeipprompt: intellectual property infringement assessment of code language models

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:20.053757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.206200Z digest=sha256:25a43aef2609292ff2a03d9801fb9c463bd8333688b486c4999cd856f4ee20db

Observation 16b48c01-9f40-41ff-888e-384a39c9a208 · outbound

This paper cites Inducing Vulnerable Code Generation in LLM Coding Assistants.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Inducing Vulnerable Code Generation in LLM Coding Assistants

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:15.340847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:15.340847Z digest=sha256:d9d36cf229c73423045875396f47d07056de4474ae7fdc843748113df20e0b9c

Observation c76f7210-55c6-4567-aaf4-cb8d912ee442 · outbound

This paper cites How well does llm generate security tests?arXiv preprint arXiv:2310.00710, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs How well does llm generate security tests?arXiv preprint arXiv:2310.00710, 2023

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:15.407231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:15.407231Z digest=sha256:040f7c441be029ff55d85407bc6c018b9da452f77113c2344165dbf3e8227461

Observation 31be87fc-8af6-437f-815a-8d84cee8a050 · outbound

This paper cites Genimage: A million-scale benchmark for detecting ai-generated image.Advances in Neural Information Processing Systems (NeurIPS), 36:77771–77782, 2023.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Genimage: A million-scale benchmark for detecting ai-generated image.Advances in Neural Information Processing Systems (NeurIPS), 36:77771–77782, 2023

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:19.823765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.510378Z digest=sha256:e4b6f2e9e2e86a24733981866ae98f8da453a29c01d56cbb6d1b1c3f8274680c

Observation fa6ba879-fa02-4f20-a6f2-f0e7caf8aec1 · outbound

This paper cites Wilddeepfake: A challenging real-world dataset for deepfake detection.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Wilddeepfake: A challenging real-world dataset for deepfake detection

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:19.529874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.611545Z digest=sha256:ed39f682965bc5681453579e069efe19e818de5063656ebc6bbc21a0b1c57e11

Observation 3b4e48ee-263b-4fa0-ba3e-0eda070ab4b1 · outbound

This paper cites an unresolved cited work.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:19.259887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.642841Z digest=sha256:1c40277fd6a62cc16824ee4fa28b270dcd607bf8d604ab13e1ee9e6454aaf052

Observation 83acb047-23d9-4acc-b227-7eb9507bc60f · outbound

This paper cites an unresolved cited work.

CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs Unresolved cited work

Reference 100

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:19.095212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:55:15.691408Z digest=sha256:f0c7687f0bdbda1960cb1d38c264fccf9e9f217cfbec35e58c796f9622b84096

Pith citing papers

Observation 1ed2e67d-6f24-4557-8c0c-03900d823793 · inbound

I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution cites this paper.

I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution CodeMirage: A Multi-Lingual Benchmark for Detecting AI-Generated and Paraphrased Source Code from Production-Level LLMs

Reference 2025

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T23:54:56.716112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:54:51.091578Z digest=sha256:0c16106ba63be66fd46ba9d26373434778f5a0d3d61a67b3c119f17937bff820