Pith. sign in

REVIEW 5 cited by

RADAR: Robust AI-Text Detection via Adversarial Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2307.03838 v2 pith:5VIYBJVZ submitted 2023-07-07 cs.CL cs.AIcs.LG

classification cs.CLcs.AIcs.LG
keywords radarai-textllmsadversarialdetectiondetectorparaphraserrobust
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recent advances in large language models (LLMs) and the intensifying popularity of ChatGPT-like applications have blurred the boundary of high-quality text generation between humans and machines. However, in addition to the anticipated revolutionary changes to our technology and society, the difficulty of distinguishing LLM-generated texts (AI-text) from human-generated texts poses new challenges of misuse and fairness, such as fake content generation, plagiarism, and false accusations of innocent writers. While existing works show that current AI-text detectors are not robust to LLM-based paraphrasing, this paper aims to bridge this gap by proposing a new framework called RADAR, which jointly trains a robust AI-text detector via adversarial learning. RADAR is based on adversarial training of a paraphraser and a detector. The paraphraser's goal is to generate realistic content to evade AI-text detection. RADAR uses the feedback from the detector to update the paraphraser, and vice versa. Evaluated with 8 different LLMs (Pythia, Dolly 2.0, Palmyra, Camel, GPT-J, Dolly 1.0, LLaMA, and Vicuna) across 4 datasets, experimental results show that RADAR significantly outperforms existing AI-text detection methods, especially when paraphrasing is in place. We also identify the strong transferability of RADAR from instruction-tuned LLMs to other LLMs, and evaluate the improved capability of RADAR via GPT-3.5-Turbo.

Discussion (0). Sign in to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. PeerPrism: Peer Evaluation Expertise vs Review-writing AI

    cs.CL 2026-04 unverdicted novelty 7.0 of 10

    PeerPrism benchmark demonstrates that state-of-the-art LLM detectors conflate surface text style with intellectual contribution and fail on hybrid human-AI peer reviews.

  2. Performance Analysis and Optimization for Laser-Phase-Noise based Quantum Random Number Generation

    quant-ph 2026-04 unverdicted novelty 6.0 of 10

    A comprehensive physical model predicts power spectrum and probability distribution in laser phase noise QRNG, enabling quantitative optimization of generation rate and quantum min-entropy.

  3. DEER: Disentangled Mixture of Experts with Instance-Adaptive Routing for Generalizable Machine-Generated Text Detection

    cs.CL 2025-11 conditional novelty 6.0 of 10

    DEER, a disentangled mixture-of-experts detector with RL-based instance routing, reports F1 gains of about 1.4 in-domain and 5.3 points out-of-domain over prior MGT detectors.

  4. Paraphrasing Attack Resilience of Various AI-Generated Text Detection Methods

    cs.LG 2026-05 unverdicted novelty 5.0 of 10

    Binoculars-inclusive ensembles detect AI text best overall but suffer the largest performance drops under paraphrasing attacks.

  5. Performance Analysis and Optimization for Laser-Phase-Noise based Quantum Random Number Generation

    quant-ph 2026-04 unverdicted novelty 4.0 of 10

    A validated physical model predicts power spectrum and raw-data distributions for laser-phase-noise QRNGs, enabling quantitative rate optimization and proactive photonic-integrated design.

Pith tools