Towards Reliable Detection of LLM-Generated Texts: A Comprehensive Evaluation Framework with CUDRT
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tao, Zhen, Chen, Yanfang, Xi, Dinghao, Li, Zhiyu, Xu, Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fake Artificial Intelligence Generated Contents (FAIGC): A Survey of Theories, Detection Methods, and Opportunities
von: Yu, Xiaomin, et al.
Veröffentlicht: (2024)
von: Yu, Xiaomin, et al.
Veröffentlicht: (2024)
Unveiling Large Language Models Generated Texts: A Multi-Level Fine-Grained Detection Framework
von: Tao, Zhen, et al.
Veröffentlicht: (2024)
von: Tao, Zhen, et al.
Veröffentlicht: (2024)
CAT-LLM: Style-enhanced Large Language Models with Text Style Definition for Chinese Article-style Transfer
von: Tao, Zhen, et al.
Veröffentlicht: (2024)
von: Tao, Zhen, et al.
Veröffentlicht: (2024)
SurveyEval: Towards Comprehensive Evaluation of LLM-Generated Academic Surveys
von: Zhao, Jiahao, et al.
Veröffentlicht: (2025)
von: Zhao, Jiahao, et al.
Veröffentlicht: (2025)
StyleDecipher: Robust and Explainable Detection of LLM-Generated Texts with Stylistic Analysis
von: Li, Siyuan, et al.
Veröffentlicht: (2025)
von: Li, Siyuan, et al.
Veröffentlicht: (2025)
Variation is the Key: A Variation-Based Framework for LLM-Generated Text Detection
von: Li, Xuecong, et al.
Veröffentlicht: (2026)
von: Li, Xuecong, et al.
Veröffentlicht: (2026)
DSIPA: Detecting LLM-Generated Texts via Sentiment-Invariant Patterns Divergence Analysis
von: Li, Siyuan, et al.
Veröffentlicht: (2026)
von: Li, Siyuan, et al.
Veröffentlicht: (2026)
MemBench: Towards More Comprehensive Evaluation on the Memory of LLM-based Agents
von: Tan, Haoran, et al.
Veröffentlicht: (2025)
von: Tan, Haoran, et al.
Veröffentlicht: (2025)
Towards Fair and Comprehensive Evaluation of Routers in Collaborative LLM Systems
von: Wu, Wanxing, et al.
Veröffentlicht: (2026)
von: Wu, Wanxing, et al.
Veröffentlicht: (2026)
From Rubrics to Reliable Scores: Evidence-Grounded Text Evaluation with LLM Judges
von: Hong, Yihan, et al.
Veröffentlicht: (2026)
von: Hong, Yihan, et al.
Veröffentlicht: (2026)
XMark: Reliable Multi-Bit Watermarking for LLM-Generated Texts
von: Xu, Jiahao, et al.
Veröffentlicht: (2026)
von: Xu, Jiahao, et al.
Veröffentlicht: (2026)
DetectAnyLLM: Towards Generalizable and Robust Detection of Machine-Generated Text Across Domains and Models
von: Fu, Jiachen, et al.
Veröffentlicht: (2025)
von: Fu, Jiachen, et al.
Veröffentlicht: (2025)
SEFD: Semantic-Enhanced Framework for Detecting LLM-Generated Text
von: He, Weiqing, et al.
Veröffentlicht: (2024)
von: He, Weiqing, et al.
Veröffentlicht: (2024)
CoSER: A Comprehensive Literary Dataset and Framework for Training and Evaluating LLM Role-Playing and Persona Simulation
von: Wang, Xintao, et al.
Veröffentlicht: (2025)
von: Wang, Xintao, et al.
Veröffentlicht: (2025)
STED and Consistency Scoring: A Framework for Evaluating LLM Structured Output Reliability
von: Wang, Guanghui, et al.
Veröffentlicht: (2025)
von: Wang, Guanghui, et al.
Veröffentlicht: (2025)
Can AI-Generated Text be Reliably Detected?
von: Sadasivan, Vinu Sankar, et al.
Veröffentlicht: (2023)
von: Sadasivan, Vinu Sankar, et al.
Veröffentlicht: (2023)
DPIC: Decoupling Prompt and Intrinsic Characteristics for LLM Generated Text Detection
von: Yu, Xiao, et al.
Veröffentlicht: (2023)
von: Yu, Xiao, et al.
Veröffentlicht: (2023)
C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts
von: Qing, Chenxi, et al.
Veröffentlicht: (2026)
von: Qing, Chenxi, et al.
Veröffentlicht: (2026)
An LLM Maturity Model for Reliable and Transparent Text-to-Query
von: Yu, Lei, et al.
Veröffentlicht: (2024)
von: Yu, Lei, et al.
Veröffentlicht: (2024)
Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey
von: Zhu, Jiachen, et al.
Veröffentlicht: (2025)
von: Zhu, Jiachen, et al.
Veröffentlicht: (2025)
TAD-Bench: A Comprehensive Benchmark for Embedding-Based Text Anomaly Detection
von: Cao, Yang, et al.
Veröffentlicht: (2025)
von: Cao, Yang, et al.
Veröffentlicht: (2025)
Falcon: A Comprehensive Chinese Text-to-SQL Benchmark for Enterprise-Grade Evaluation
von: Luo, Wenzhen, et al.
Veröffentlicht: (2025)
von: Luo, Wenzhen, et al.
Veröffentlicht: (2025)
Well Begun, Half Done: Reinforcement Learning with Prefix Optimization for LLM Reasoning
von: Sun, Yiliu, et al.
Veröffentlicht: (2025)
von: Sun, Yiliu, et al.
Veröffentlicht: (2025)
SQLBench: A Comprehensive Evaluation for Text-to-SQL Capabilities of Large Language Models
von: Zhang, Bin, et al.
Veröffentlicht: (2024)
von: Zhang, Bin, et al.
Veröffentlicht: (2024)
Towards Human-Like Grading: A Unified LLM-Enhanced Framework for Subjective Question Evaluation
von: Zhua, Fanwei, et al.
Veröffentlicht: (2025)
von: Zhua, Fanwei, et al.
Veröffentlicht: (2025)
DivScore: Zero-Shot Detection of LLM-Generated Text in Specialized Domains
von: Chen, Zhihui, et al.
Veröffentlicht: (2025)
von: Chen, Zhihui, et al.
Veröffentlicht: (2025)
Towards Comprehensive Detection of Chinese Harmful Memes
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
IMGTB: A Framework for Machine-Generated Text Detection Benchmarking
von: Spiegel, Michal, et al.
Veröffentlicht: (2023)
von: Spiegel, Michal, et al.
Veröffentlicht: (2023)
On the Detectability of LLM-Generated Text: What Exactly Is LLM-Generated Text?
von: Geng, Mingmeng, et al.
Veröffentlicht: (2025)
von: Geng, Mingmeng, et al.
Veröffentlicht: (2025)
GEAR: A General Evaluation Framework for Abductive Reasoning
von: He, Kaiyu, et al.
Veröffentlicht: (2025)
von: He, Kaiyu, et al.
Veröffentlicht: (2025)
Conan-Embedding-v2: Training an LLM from Scratch for Text Embeddings
von: Li, Shiyu, et al.
Veröffentlicht: (2025)
von: Li, Shiyu, et al.
Veröffentlicht: (2025)
LLM-Powered Benchmark Factory: Reliable, Generic, and Efficient
von: Yuan, Peiwen, et al.
Veröffentlicht: (2025)
von: Yuan, Peiwen, et al.
Veröffentlicht: (2025)
MERA: A Comprehensive LLM Evaluation in Russian
von: Fenogenova, Alena, et al.
Veröffentlicht: (2024)
von: Fenogenova, Alena, et al.
Veröffentlicht: (2024)
SHIELD: Evaluation and Defense Strategies for Copyright Compliance in LLM Text Generation
von: Liu, Xiaoze, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoze, et al.
Veröffentlicht: (2024)
DetectRL: Benchmarking LLM-Generated Text Detection in Real-World Scenarios
von: Wu, Junchao, et al.
Veröffentlicht: (2024)
von: Wu, Junchao, et al.
Veröffentlicht: (2024)
Integrated Framework for LLM Evaluation with Answer Generation
von: Lee, Sujeong, et al.
Veröffentlicht: (2025)
von: Lee, Sujeong, et al.
Veröffentlicht: (2025)
Can LLMs Evaluate What They Cannot Annotate? Revisiting LLM Reliability in Hate Speech Detection
von: Piot, Paloma, et al.
Veröffentlicht: (2025)
von: Piot, Paloma, et al.
Veröffentlicht: (2025)
TruthEval: A Dataset to Evaluate LLM Truthfulness and Reliability
von: Khatun, Aisha, et al.
Veröffentlicht: (2024)
von: Khatun, Aisha, et al.
Veröffentlicht: (2024)
Multiscale Positive-Unlabeled Detection of AI-Generated Texts
von: Tian, Yuchuan, et al.
Veröffentlicht: (2023)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2023)
RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation Patterns
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Fake Artificial Intelligence Generated Contents (FAIGC): A Survey of Theories, Detection Methods, and Opportunities
von: Yu, Xiaomin, et al.
Veröffentlicht: (2024) -
Unveiling Large Language Models Generated Texts: A Multi-Level Fine-Grained Detection Framework
von: Tao, Zhen, et al.
Veröffentlicht: (2024) -
CAT-LLM: Style-enhanced Large Language Models with Text Style Definition for Chinese Article-style Transfer
von: Tao, Zhen, et al.
Veröffentlicht: (2024) -
SurveyEval: Towards Comprehensive Evaluation of LLM-Generated Academic Surveys
von: Zhao, Jiahao, et al.
Veröffentlicht: (2025) -
StyleDecipher: Robust and Explainable Detection of LLM-Generated Texts with Stylistic Analysis
von: Li, Siyuan, et al.
Veröffentlicht: (2025)