Awesome LLM-generated Text Detection

August 3, 2026 · View on GitHub

Logo

Awesome License-MIT PRs-Welcome last-commit

The powerful ability of large language models (LLMs) to understand, follow, and generate complex languages has enabled LLM-generated texts to flood many areas of our daily lives at an incredible rate, with potentially negative impacts and risks on society and academia. As LLMs continue to expand, how can we detect LLM-generated texts to help minimize the threat posed by the misuse of LLMs?

cover

¹ Junchao Wu, ¹ Shu Yang, ¹ Runzhe Zhan, ¹ ² Yulin Yuan, ¹ Derek Fai Wong, ¹ Lidia Sam Chao

¹ University of Macau, ² Peking University

📢 News

🔍 Table of Contents

📃 Papers

Overview

A survey and reflection on the latest research breakthroughs in LLM-generated Text detection, including data, detectors, metrics, current issues and future directions. Please refer to our article/paper for more details.

Datasets

Benchmarks

Benchmarks / DatasetsVenueDateUseHumanLLMs
HC3arXiv2023-01train58k26k
HC3-ChinesearXiv2023-01train22k17k
CHEATarXiv2023-04train15k35k
GROVER DatasetNeurIPS 20192019-05train valid test5k 2k 8k5k 1k 4k
TweepFakePLoS ONE2020-07train12k12k
GPT-2 Output DatasetGitHub-train250k250k
TuringBenchEMNLP 2021 Findings2021train10k190k
MGTBencharXiv2023-03train test2k 56313k 3k
ArguGPTarXiv2023-04train valid test3k 350 3503k 350 350
DeepfakeText-DatasetACL 20242023-05train valid test95k 29k 29k236k 29k 28k
M4arXiv2023-05train valid test122k 500 500122k 500 500
GPABenchmarkarXiv2023-06train600k600k
Scientific-articles BenchmarkTrustNLP 20232023train test8k 4k8k 4k
DetectRLNeurIPS 2024 D&B2024-10train test101k134k (4 LLMs, 4 domains, 4 attacks)
DetectRL-XACL 20262026-05test3.46M (8 langs, 6 domains, 4 LLMs, 8 attacks)
MULTITuDEEMNLP 20232023-10test-56k (7 langs, 8 LLMs)
RAIDACL 20242024-05test-6M (11 models)
M4GT-BenchACL 20242024-02test-- (4 domains, multi-lingual)
MultiSocialACL 20252025-07test58k414k (22 langs, 5 platforms, 7 LLMs)
Detecting the MachinearXiv2026-03test23k15k (HC3 + ELI5, multiple LLMs)

Potential Datasets

TasksDatasets
Questions AnsweringPubMedQA, Children book corpus (CBT), ELI5, TruthfulQA, NarrativeQA
Scientific writingPeer Read, arXiv, TOEFL11
Story generationWritingPrompts
News Article writingXSum
Web TextWiki40b, WebText, Avax tweets dataset, Climate Change Tweets Ids
Opinion statementsr/ChangeMyView (CMV) Reddit subcommunity, Yelp , IMDB Dataset
Comprehension and ReasoningSciGen, ROCStories Corpora, HellaSwag, SQuAD

Detectors

Watermarking Technology

PaperVenueDateLink
A watermark for large language models.ICML 20232023-01Static Badge Static Badge
On the Reliability of Watermarks for Large Language ModelsICLR 20242023-06Static Badge Static Badge
A Private Watermark for Large Language ModelsICLR 20242023-07Static Badge Static Badge
Distillation-Resistant Watermarking for Model Protection in NLParXiv2022-10Static Badge Static Badge
Watermarking Pre-trained Language Models with BackdooringarXiv2022-10Static Badge
CATER: Intellectual Property Protection on Text Generation APIs via Conditional WatermarksNeurIPS 20222022-09Static Badge Static Badge
An Adaptive Watermark for Large Language ModelsICML 20242024-01Static Badge
WaterBench: A Benchmark for LLM WatermarkingACL 20242023-11Static Badge
Watermarking Makes Language Models RadioactivearXiv2024-02Static Badge
SynthID-Text: Identifying AI-Generated Text ContentNature 20242024-10Static Badge

Statistics-based Detectors

PaperVenueDateLink
DetectGPT: Zero-Shot Machine-Generated Text Detection using Probability CurvatureICML 20232023-01Static Badge Static Badge
Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability CurvatureICLR 20242023-10Static Badge Static Badge
Efficient Detection of LLM-generated Texts with a Bayesian Surrogate ModelarXiv2023-05Static Badge
DetectLLM: Leveraging Log Rank Information for Zero-Shot Detection of Machine-Generated TextarXiv2023-05Static Badge Static Badge
GLTR: Statistical Detection and Visualization of Generated TextACL 2019 Demo2019-06Static Badge Static Badge
HowkGPT: Investigating the Detection of ChatGPT-generated University Student Homework through Context-Aware Perplexity AnalysisarXiv2023-05Static Badge
Intrinsic Dimension Estimation for Robust Detection of AI-Generated TextsarXiv2023-06Static Badge
Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScoreCOLING 20252024-05Static Badge Static Badge
RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation PatternsTACL 20252025-08Static Badge Static Badge
Is ChatGPT Involved in Texts? Measure the Polish Ratio to Detect ChatGPT-Generated TextarXiv2023-07Static Badge
DNA-GPT: Divergent N-Gram Analysis for Training-Free Detection of GPT-Generated TextICLR 20242023-05Static Badge Static Badge
Binoculars: Zero-Shot Detection of Machine-Generated TextICML 20242024-01Static Badge
LLMDet: A Third Party Large Language Models Generated Text Detection ToolEMNLP 2023 Findings2023-05Static Badge
Ghostbuster: Detecting Text Ghostwritten by Large Language ModelsarXiv2023-05Static Badge
GPT-who: An Information Density-based Machine-Generated Text DetectorarXiv2023-10Static Badge
BiScope: AI-generated Text Detection by Checking Memorization of Preceding TokensNeurIPS 20242024-06Static Badge Static Badge
Ten Words Only Still Help: Improving Black-Box AI-Generated Text Detection via Proxy-Guided Efficient Re-Sampling (POGER)IJCAI 20242024-02Static Badge Static Badge
AdaDetectGPT: Adaptive Detection of LLM-Generated Text with Statistical GuaranteesNeurIPS 20252025-10Static Badge Static Badge
Exons-Detect: Identifying and Amplifying Exonic Tokens via Hidden-State Discrepancy for Robust AI-Generated Text DetectionACL 20262026-03Static Badge
Segmenting Human–LLM Co-authored Text via Change Point DetectionarXiv2026-05Static Badge Static Badge
Triospect: A Three-Dimensional Framework for Robust Statistical AI-Generated Text Detection Against Diverse AttacksTACL 20262026-06Static Badge Static Badge
Detecting LLM-Generated Tokens in Human–LLM Coauthored TextarXiv2026-07Static Badge
Detecting LLM-Generated Text with Performance GuaranteesarXiv2026-01Static Badge

Neural-based Detectors

PaperVenueDateLink
How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and DetectionarXiv2023-01Static Badge Static Badge
Multiscale Positive-Unlabeled Detection of AI-Generated TextsICLR 2024 (Spotlight)2023-05Static Badge Static Badge
Real or fake? Learning to discriminate machine from human generated textarXiv2019-06Static Badge
Automatic Detection of Generated Text is Easiest when Humans are FooledACL 20202019-11Static Badge
Stylometric Detection of AI-Generated Text in Twitter TimelinesarXiv2023-03Static Badge
TweepFake: about Detecting Deepfake TweetsPLoS ONE2020-07Static Badge Static Badge
Towards a Robust Detection of Language Model Generated Text: Is ChatGPT that Easy to Detect?TALN 20232023-06Static Badge
Deepfake Text Detection in the WildACL 20242023-05Static Badge Static Badge
ArguGPT: evaluating, understanding and identifying argumentative essays generated by GPT modelsarXiv2023-04Static Badge Static Badge
Check Me If You Can: Detecting ChatGPT-Generated Academic Writing using CheckGPTarXiv2023-06Static Badge
GPT-Sentinel: Distinguishing Human and ChatGPT Generated ContentarXiv2023-05Static Badge
Neural Deepfake Detection with Factual Structure of TextEMNLP 20202020-10Static Badge
ConDA: Contrastive Domain Adaptation for AI-generated Text DetectionIJCNLP-AACL 20232023-09Static Badge Static Badge
RADAR: Robust AI-Text Detection via Adversarial LearningNeurIPS 20232023-07Static Badge
OUTFOX: LLM-generated Essay Detection through In-context Learning with Adversarially Generated ExamplesAAAI 20242023-07Static Badge Static Badge
Fighting fire with fire: Can chatgpt detect ai-generated text?SIGKDD Explorations2023-08Static Badge Static Badge
GPT Paternity Test: GPT Generated Text Detection with GPT Genetic InheritancearXiv2023-05Static Badge
Raidar: geneRative AI Detection viA RewritingICLR 20242024-01Static Badge
Beat LLMs at Their Own Game: Zero-Shot LLM-Generated Text Detection via Querying ChatGPTEMNLP 20232023-12Static Badge Static Badge
Defending Against Neural Fake News by Removing LLMs' Greatest WeaknessNeurIPS 20192019-05Static Badge Static Badge
CoCo: Coherence-Enhanced Machine-Generated Text DetectionEMNLP 20232022-12Static Badge
SeqXGPT: Sentence-Level AI-Generated Text DetectionEMNLP 20232023-10Static Badge
J-Guard: Robust Guardrails against Unreliable Text Generation by LLMsIJCNLP-AACL 20232023-09Static Badge
DEMASQ: Unmasking the ChatGPT WordsmithNDSS 20242023-11Static Badge
Smaller Language Models are Better Black-box DetectorsEACL 20242023-05Static Badge
DeTeCtive: Detecting AI-generated Text via Multi-Level Contrastive LearningNeurIPS 20242024-10Static Badge Static Badge
A Ship of Theseus: Paraphrasing Capabilities for Detecting LLM-Generated TextACL 20242023-11Static Badge
ReMoDetect: Reward Modified Detection of LLM-Generated TextarXiv2024-05Static Badge
Origin Tracing and Detecting of Large Language ModelsarXiv2023-04Static Badge
Text Fluoroscopy: Detecting LLM-Generated Text through Intrinsic FeaturesEMNLP 20242024-11Static Badge Static Badge
Human Texts Are Outliers: Detecting LLM-generated Texts via Out-of-distribution DetectionNeurIPS 20252025-10Static Badge Static Badge
DAMAGE: Detecting Adversarially Modified AI Generated TextarXiv2025-01Static Badge
NOTAI.AI: Explainable Detection of Machine-Generated Text via Curvature and Feature AttributionarXiv2026-03Static Badge Static Badge
Breaking the Generator Barrier: Disentangled Representation for Generalizable AI-Text Detection (DRGD)arXiv2026-04Static Badge Static Badge
GigaCheck: Detecting LLM-generated Content via Object-Centric Span LocalizationACL 2026 Findings2026-07Static Badge

Human-assisted Methods

PaperVenueDateLink
RoFT: A Tool for Real vs Fake Text DetectionEMNLP 2020 Demo2020-10Static Badge
Human Heuristics for AI-Generated Language Are FlawedarXiv2022-06Static Badge
Real or Fake Text? Investigating the Abilities of Language Models to Detect AI-Generated TextAAAI 20232022-12Static Badge
Does Human Collaboration Enhance the Accuracy and Reliability of AI-Generated Text Detection?HCOMP 20232023-04Static Badge
People cannot distinguish GPT-4 from a human in a Turing testarXiv2024-05Static Badge

Detector Attack

PaperVenueDateLink
Paraphrasing evades detectors of AI-generated text, but retrieval is an effective defenseNeurIPS 20232023-03Static Badge Static Badge
Can AI-Generated Text be Reliably Detected?arXiv2023-03Static Badge Static Badge
Red Teaming Language Model Detectors with Language ModelsTACL 20232023-05Static Badge Static Badge
Humanizing Machine-Generated Content: Evading AI-Text Detection through Adversarial AttackLREC-COLING 20242024-05Static Badge
RAFT: Realistic Attacks to Fool Text DetectorsEMNLP 20242024-10Static Badge Static Badge
Adversarial Paraphrasing: A Universal Attack for Humanizing AI-Generated TextNeurIPS 20252025-06Static Badge
CoPA: Contrastive Paraphrase Attacks on LLM-Generated Text DetectorsEMNLP 20252025-05Static Badge
Attacks on Machine-Text Detectors Retain Stylistic FingerprintsarXiv2025-05Static Badge
Revealing Weaknesses in Text Watermarking Through Self-Information Rewrite Attacks (SIRA)ICML 20252025-05Static Badge Static Badge
DE-MARK: Watermark Removal in Large Language ModelsICML 20252024-10Static Badge Static Badge
MASH: Evading Black-Box AI-Generated Text Detectors via Style HumanizationACL 2026 Findings2026-01Static Badge Static Badge
Vaporizer: Breaking Watermarking Schemes for Large Language Model OutputsarXiv2026-05Static Badge
AI Watermark Evidence Fails Forensic Readiness: An Empirical EvaluationarXiv2026-07Static Badge

Shared Task

PaperVenueDateLink
Findings of the RuATD Shared Task 2022 on Artificial Text Detection in RussianDialogue 20222022-06Static Badge Static Badge
SemEval-2024 Task 8: Multigenerator, Multidomain, and Multilingual Black-Box Machine-Generated Text DetectionSemEval 20242024-04Static Badge Static Badge
GenAI Content Detection Task 1: English and Multilingual Machine-Generated Text DetectionCOLING 20252025-01Static Badge Static Badge
Overview of the NLPCC 2025 Shared Task 1: LLM-Generated Text DetectionNLPCC 20252025-08Static Badge Static Badge
Findings of the Counter Turing Test (CT2): AI-Generated Text DetectionDeFactify 20252026-05Static Badge
The Second Shared Task on LLM-Generated Text Detection (NLPCC 2026 Task 6)NLPCC 20262026-11Static Badge

Other Surveys

PaperVenueDateLink
Automatic Detection of Machine Generated Text: A Critical SurveyCOLING 20202020-11Static Badge
The Science of Detecting LLM-Generated TextsarXiv2023-02Static Badge
Machine Generated Text: A Comprehensive Survey of Threat Models and Detection MethodsACM Trustworthy AI2022-10Static Badge
Computer-Generated Text Detection Using Machine Learning: A Systematic ReviewSpringer-Static Badge
Attribution and Obfuscation of Neural Text Authorship: A Data Mining PerspectiveACM SIGKDD Explorations2022-10Static Badge
Deepfake Text Detection: Limitations and OpportunitiesS&P 20232022-10Static Badge
A Survey of AI-generated Text Forensic SystemsarXiv2024-03Static Badge
SoK: Watermarking for AI-Generated ContentarXiv2024-11Static Badge
The Imitation Game Revisited: A Comprehensive Survey on Recent Advances in AI-generated Text DetectionESWA 20252025-01Static Badge Static Badge
Watermarking for AI Content Detection: A ReviewICLR 2025 Workshop2025-04Static Badge
A Survey on LLM Watermarking: Theory and DeploymentarXiv2026-07Static Badge

🚩 Citation

If our research helps you, please kindly cite our paper.

@article{wu2025survey,
      title={A Survey on LLM-Generated Text Detection: Necessity, Methods, and Future Directions}, 
      author={Junchao Wu and Shu Yang and Runzhe Zhan and Yulin Yuan and Lidia Sam Chao and Derek Fai Wong},
      journal      = {Computational Linguistics},
      volume       = {51},
      number       = {1},
      year         = {2025},
      pages        = {275--338},
      url          = {https://aclanthology.org/2025.cl-1.8/},
}

@inproceedings{wu2025GECScore,
      title={Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScore}, 
      author={Junchao Wu and Runzhe Zhan and Derek F. Wong and Shu Yang and Xuebo Liu and Lidia S. Chao and Min Zhang},
      booktitle    = {Proceedings of the 31st International Conference on Computational Linguistics},
      year         = {2025},
      pages        = {10275--10292},
      url          = {https://aclanthology.org/2025.coling-main.684/},
}

@article{chen2025RepreGuard,
      title={RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation Patterns}, 
      author={Xin Chen and Junchao Wu and Shu Yang and Runzhe Zhan and Zeyu Wu and Ziyang Luo and Di Wang and Min Yang and Lidia S. Chao and Derek F. Wong},
      journal      = {Transactions of the Association for Computational Linguistics},
      volume       = {13},
      year         = {2025},
      pages        = {1812--1831},
      url          = {https://aclanthology.org/2025.tacl-1.81/},
}

@inproceedings{wu2026DetectRLX,
      title={DetectRL-X: Towards Reliable Multilingual and Real-World LLM-Generated Text Detection}, 
      author={Junchao Wu and Yefeng Liu and Chenyu Zhu and Hao Zhang and Zeyu Wu and Tianqi Shi and Yichao Du and Longyue Wang and Weihua Luo and Jinsong Su and Derek F. Wong},
      booktitle    = {Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)},
      year         = {2026},
      pages        = {38247--38294},
      url          = {https://aclanthology.org/2026.acl-long.1773/},
}

@inproceedings{wu2024DetectRL,
      title={DetectRL: Benchmarking LLM-Generated Text Detection in Real-World Scenarios}, 
      author={Junchao Wu and Runzhe Zhan and Derek F. Wong and Shu Yang and Xinyi Yang and Yulin Yuan and Lidia S. Chao},
      booktitle    = {Advances in Neural Information Processing Systems 37 (NeurIPS 2024) Datasets and Benchmarks Track},
      year         = {2024},
      url          = {https://proceedings.neurips.cc/paper_files/paper/2024/hash/b61bdf7e9f64c04ec75a26e781e2ad51-Abstract-Datasets_and_Benchmarks_Track.html},
}

Contributing

Contributions are welcome! If you have any ideas, suggestions, or bug reports, please open an issue or submit a pull request. We appreciate your contributions to making LLM-generated Text Detection work even better.