DetectRL: Benchmarking LLM-Generated Text Detection in Real-World Scenarios
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wu, Junchao, Zhan, Runzhe, Wong, Derek F., Yang, Shu, Yang, Xinyi, Yuan, Yulin, Chao, Lidia S. |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
A Survey on LLM-Generated Text Detection: Necessity, Methods, and Future Directions
par: Wu, Junchao, et autres
Publié: (2023)
par: Wu, Junchao, et autres
Publié: (2023)
DetectRL-X: Towards Reliable Multilingual and Real-World LLM-Generated Text Detection
par: Wu, Junchao, et autres
Publié: (2026)
par: Wu, Junchao, et autres
Publié: (2026)
Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScore
par: Wu, Junchao, et autres
Publié: (2024)
par: Wu, Junchao, et autres
Publié: (2024)
Rethinking Prompt-based Debiasing in Large Language Models
par: Yang, Xinyi, et autres
Publié: (2025)
par: Yang, Xinyi, et autres
Publié: (2025)
RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation Patterns
par: Chen, Xin, et autres
Publié: (2025)
par: Chen, Xin, et autres
Publié: (2025)
Prefix Text as a Yarn: Eliciting Non-English Alignment in Foundation Language Model
par: Zhan, Runzhe, et autres
Publié: (2024)
par: Zhan, Runzhe, et autres
Publié: (2024)
Benchmarking the Detection of LLMs-Generated Modern Chinese Poetry
par: Wang, Shanshan, et autres
Publié: (2025)
par: Wang, Shanshan, et autres
Publié: (2025)
Are Large Reasoning Models Good Translation Evaluators? Analysis and Performance Boost
par: Zhan, Runzhe, et autres
Publié: (2025)
par: Zhan, Runzhe, et autres
Publié: (2025)
Neuron-Aware Data Selection In Instruction Tuning For Large Language Models
par: Chen, Xin, et autres
Publié: (2026)
par: Chen, Xin, et autres
Publié: (2026)
Let's Focus on Neuron: Neuron-Level Supervised Fine-tuning for Large Language Model
par: Xu, Haoyun, et autres
Publié: (2024)
par: Xu, Haoyun, et autres
Publié: (2024)
Path Drift in Large Reasoning Models:How First-Person Commitments Override Safety
par: Huang, Yuyi, et autres
Publié: (2025)
par: Huang, Yuyi, et autres
Publié: (2025)
Intrinsic Model Weaknesses: How Priming Attacks Unveil Vulnerabilities in Large Language Models
par: Huang, Yuyi, et autres
Publié: (2025)
par: Huang, Yuyi, et autres
Publié: (2025)
VisAidMath: Benchmarking Visual-Aided Mathematical Reasoning
par: Ma, Jingkun, et autres
Publié: (2024)
par: Ma, Jingkun, et autres
Publié: (2024)
Understanding and Mitigating Political Stance Cross-topic Generalization in Large Language Models
par: Zhang, Jiayi, et autres
Publié: (2025)
par: Zhang, Jiayi, et autres
Publié: (2025)
Understanding Aha Moments: from External Observations to Internal Mechanisms
par: Yang, Shu, et autres
Publié: (2025)
par: Yang, Shu, et autres
Publié: (2025)
Exposing the Cracks: Vulnerabilities of Retrieval-Augmented LLM-based Machine Translation
par: Sun, Yanming, et autres
Publié: (2025)
par: Sun, Yanming, et autres
Publié: (2025)
Fraud-R1 : A Multi-Round Benchmark for Assessing the Robustness of LLM Against Augmented Fraud and Phishing Inducements
par: Yang, Shu, et autres
Publié: (2025)
par: Yang, Shu, et autres
Publié: (2025)
Worlds Within Words: Translating Culture in Ancient Chinese Texts with Multi-Agent Coordination
par: He, Xiaoqi, et autres
Publié: (2026)
par: He, Xiaoqi, et autres
Publié: (2026)
Is Long-to-Short a Free Lunch? Investigating Inconsistency and Reasoning Efficiency in LRMs
par: Yang, Shu, et autres
Publié: (2025)
par: Yang, Shu, et autres
Publié: (2025)
C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts
par: Qing, Chenxi, et autres
Publié: (2026)
par: Qing, Chenxi, et autres
Publié: (2026)
Investigating CoT Monitorability in Large Reasoning Models
par: Yang, Shu, et autres
Publié: (2025)
par: Yang, Shu, et autres
Publié: (2025)
LLMs are Superior Feedback Providers: Bootstrapping Reasoning for Lie Detection with Self-Generated Feedback
par: Banerjee, Tanushree, et autres
Publié: (2024)
par: Banerjee, Tanushree, et autres
Publié: (2024)
DV-World: Benchmarking Data Visualization Agents in Real-World Scenarios
par: Meng, Jinxiang, et autres
Publié: (2026)
par: Meng, Jinxiang, et autres
Publié: (2026)
Learning to Rewrite: Generalized LLM-Generated Text Detection
par: Li, Ran, et autres
Publié: (2024)
par: Li, Ran, et autres
Publié: (2024)
Detecting LLM-Generated Text with Performance Guarantees
par: Zhou, Hongyi, et autres
Publié: (2026)
par: Zhou, Hongyi, et autres
Publié: (2026)
RuleArena: A Benchmark for Rule-Guided Reasoning with LLMs in Real-World Scenarios
par: Zhou, Ruiwen, et autres
Publié: (2024)
par: Zhou, Ruiwen, et autres
Publié: (2024)
Seeing the Poem: Image-Semantic Detection of AI-Generated Modern Chinese Poetry with MLLMs
par: Wang, Shanshan, et autres
Publié: (2026)
par: Wang, Shanshan, et autres
Publié: (2026)
FinSafetyBench: Evaluating LLM Safety in Real-World Financial Scenarios
par: Hou, Yutao, et autres
Publié: (2026)
par: Hou, Yutao, et autres
Publié: (2026)
A Two-Stage Prediction-Aware Contrastive Learning Framework for Multi-Intent NLU
par: Chen, Guanhua, et autres
Publié: (2024)
par: Chen, Guanhua, et autres
Publié: (2024)
$k$NNProxy: Efficient Training-Free Proxy Alignment for Black-Box Zero-Shot LLM-Generated Text Detection
par: Wong, Kahim, et autres
Publié: (2026)
par: Wong, Kahim, et autres
Publié: (2026)
Robust Detection of LLM-Generated Text: A Comparative Analysis
par: Su, Yongye, et autres
Publié: (2024)
par: Su, Yongye, et autres
Publié: (2024)
PhoStream: Benchmarking Real-World Streaming for Omnimodal Assistants in Mobile Scenarios
par: Lu, Xudong, et autres
Publié: (2026)
par: Lu, Xudong, et autres
Publié: (2026)
Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry
par: Li, Jiang, et autres
Publié: (2026)
par: Li, Jiang, et autres
Publié: (2026)
Can ChatGPT Really Understand Modern Chinese Poetry?
par: Wang, Shanshan, et autres
Publié: (2026)
par: Wang, Shanshan, et autres
Publié: (2026)
What is the Best Way for ChatGPT to Translate Poetry?
par: Wang, Shanshan, et autres
Publié: (2024)
par: Wang, Shanshan, et autres
Publié: (2024)
Unveiling LLMs' Metaphorical Understanding: Exploring Conceptual Irrelevance, Context Leveraging and Syntactic Influence
par: Ye, Fengying, et autres
Publié: (2025)
par: Ye, Fengying, et autres
Publié: (2025)
AdaDetectGPT: Adaptive Detection of LLM-Generated Text with Statistical Guarantees
par: Zhou, Hongyi, et autres
Publié: (2025)
par: Zhou, Hongyi, et autres
Publié: (2025)
Learn-to-Distance: Distance Learning for Detecting LLM-Generated Text
par: Zhou, Hongyi, et autres
Publié: (2026)
par: Zhou, Hongyi, et autres
Publié: (2026)
SGIC: A Self-Guided Iterative Calibration Framework for RAG
par: Chen, Guanhua, et autres
Publié: (2025)
par: Chen, Guanhua, et autres
Publié: (2025)
Can Large Language Models Identify Implicit Suicidal Ideation? An Empirical Evaluation
par: Li, Tong, et autres
Publié: (2025)
par: Li, Tong, et autres
Publié: (2025)
Documents similaires
-
A Survey on LLM-Generated Text Detection: Necessity, Methods, and Future Directions
par: Wu, Junchao, et autres
Publié: (2023) -
DetectRL-X: Towards Reliable Multilingual and Real-World LLM-Generated Text Detection
par: Wu, Junchao, et autres
Publié: (2026) -
Who Wrote This? The Key to Zero-Shot LLM-Generated Text Detection Is GECScore
par: Wu, Junchao, et autres
Publié: (2024) -
Rethinking Prompt-based Debiasing in Large Language Models
par: Yang, Xinyi, et autres
Publié: (2025) -
RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation Patterns
par: Chen, Xin, et autres
Publié: (2025)