Natural Language Processing Faculty Publications

DetectLLM: Leveraging Log-Rank Information for Zero-Shot Detection of Machine-Generated Text

Jinyan Su, Mohamed Bin Zayed University of Artificial Intelligence
Terry Yue Zhuo, Monash University
Di Wang, King Abdullah University of Science and Technology
Preslav Nakov, Mohamed Bin Zayed University of Artificial IntelligenceFollow

Document Type

Conference Proceeding

Publication Title

Findings of the Association for Computational Linguistics: EMNLP 2023

Abstract

With the rapid progress of Large language models (LLMs) and the huge amount of text they generate, it becomes impractical to manually distinguish whether a text is machine-generated. The growing use of LLMs in social media and education, prompts us to develop methods to detect machine-generated text, preventing malicious use such as plagiarism, misinformation, and propaganda. In this paper, we introduce two novel zero-shot methods for detecting machine-generated text by leveraging the Log-Rank information. One is called DetectLLM-LRR, which is fast and efficient, and the other is called DetectLLM-NPR, which is more accurate, but slower due to the need for perturbations. Our experiments on three datasets and seven language models show that our proposed methods improve over the state of the art by 3.9 and 1.75 AUROC points absolute. Moreover, DetectLLM-NPR needs fewer perturbations than previous work to achieve the same level of performance, which makes it more practical for real-world use. We also investigate the efficiency-performance trade-off based on users' preference for these two measures and provide intuition for using them in practice effectively. We release the data and the code of both methods in https://github.com/mbzuai-nlp/DetectLLM.

First Page

12395

Last Page

12412

Publication Date

1-1-2023

Recommended Citation

J. Su et al., "DetectLLM: Leveraging Log-Rank Information for Zero-Shot Detection of Machine-Generated Text," Findings of the Association for Computational Linguistics: EMNLP 2023, pp. 12395 - 12412, Jan 2023.

This document is currently not available here.

COinS

Natural Language Processing Faculty Publications

DetectLLM: Leveraging Log-Rank Information for Zero-Shot Detection of Machine-Generated Text

Document Type

Publication Title

Abstract

First Page

Last Page

Publication Date

Recommended Citation

Browse

Contribute

Links

Natural Language Processing Faculty Publications

DetectLLM: Leveraging Log-Rank Information for Zero-Shot Detection of Machine-Generated Text

Authors

Document Type

Publication Title

Abstract

First Page

Last Page

Publication Date

Recommended Citation

Share

Browse

Contribute

Links