ResearchTrend.AI
  • Papers
  • Communities
  • Events
  • Blog
  • Pricing
Papers
Communities
Social Events
Terms and Conditions
Pricing
Parameter LabParameter LabTwitterGitHubLinkedInBlueskyYoutube

© 2025 ResearchTrend.AI, All rights reserved.

  1. Home
  2. Papers
  3. 2504.21019
23
0

Kill two birds with one stone: generalized and robust AI-generated text detection via dynamic perturbations

22 April 2025
Y. Zhou
Juan Wen
Wanli Peng
Yiming Xue
Ziwei Zhang
Zhengxian Wu
    AAML
ArXivPDFHTML
Abstract

The growing popularity of large language models has raised concerns regarding the potential to misuse AI-generated text (AIGT). It becomes increasingly critical to establish an excellent AIGT detection method with high generalization and robustness. However, existing methods either focus on model generalization or concentrate on robustness. The unified mechanism, to simultaneously address the challenges of generalization and robustness, is less explored. In this paper, we argue that robustness can be view as a specific form of domain shift, and empirically reveal an intrinsic mechanism for model generalization of AIGT detection task. Then, we proposed a novel AIGT detection method (DP-Net) via dynamic perturbations introduced by a reinforcement learning with elaborated reward and action. Experimentally, extensive results show that the proposed DP-Net significantly outperforms some state-of-the-art AIGT detection methods for generalization capacity in three cross-domain scenarios. Meanwhile, the DP-Net achieves best robustness under two text adversarial attacks. The code is publicly available atthis https URL.

View on arXiv
@article{zhou2025_2504.21019,
  title={ Kill two birds with one stone: generalized and robust AI-generated text detection via dynamic perturbations },
  author={ Yinghan Zhou and Juan Wen and Wanli Peng and Yiming Xue and Ziwei Zhang and Zhengxian Wu },
  journal={arXiv preprint arXiv:2504.21019},
  year={ 2025 }
}
Comments on this paper