MAGA-Bench: Machine-Augment-Generated Text via Alignment Detection Benchmark
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Anyang, Cheng, Ying, Xu, Yiqian, Feng, Rui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
M4GT-Bench: Evaluation Benchmark for Black-Box Machine-Generated Text Detection
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
MGT-Prism: Enhancing Domain Generalization for Machine-Generated Text Detection via Spectral Alignment
von: Liu, Shengchao, et al.
Veröffentlicht: (2025)
von: Liu, Shengchao, et al.
Veröffentlicht: (2025)
R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
von: Chen, Kaijie, et al.
Veröffentlicht: (2025)
LCTG Bench: LLM Controlled Text Generation Benchmark
von: Kurihara, Kentaro, et al.
Veröffentlicht: (2025)
von: Kurihara, Kentaro, et al.
Veröffentlicht: (2025)
RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment
von: Jin, Zhuoran, et al.
Veröffentlicht: (2024)
von: Jin, Zhuoran, et al.
Veröffentlicht: (2024)
OpenTuringBench: An Open-Model-based Benchmark and Framework for Machine-Generated Text Detection and Attribution
von: La Cava, Lucio, et al.
Veröffentlicht: (2025)
von: La Cava, Lucio, et al.
Veröffentlicht: (2025)
MAD: Multi-Alignment MEG-to-Text Decoding
von: Yang, Yiqian, et al.
Veröffentlicht: (2024)
von: Yang, Yiqian, et al.
Veröffentlicht: (2024)
BenchBench: Benchmarking Automated Benchmark Generation
von: Zheng, Yandan, et al.
Veröffentlicht: (2026)
von: Zheng, Yandan, et al.
Veröffentlicht: (2026)
CLaS-Bench: A Cross-Lingual Alignment and Steering Benchmark
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2026)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2026)
SPADE: Structured Prompting Augmentation for Dialogue Enhancement in Machine-Generated Text Detection
von: Li, Haoyi, et al.
Veröffentlicht: (2025)
von: Li, Haoyi, et al.
Veröffentlicht: (2025)
IMGTB: A Framework for Machine-Generated Text Detection Benchmarking
von: Spiegel, Michal, et al.
Veröffentlicht: (2023)
von: Spiegel, Michal, et al.
Veröffentlicht: (2023)
WETBench: A Benchmark for Detecting Task-Specific Machine-Generated Text on Wikipedia
von: Quaremba, Gerrit, et al.
Veröffentlicht: (2025)
von: Quaremba, Gerrit, et al.
Veröffentlicht: (2025)
ChronoMagic-Bench: A Benchmark for Metamorphic Evaluation of Text-to-Time-lapse Video Generation
von: Yuan, Shenghai, et al.
Veröffentlicht: (2024)
von: Yuan, Shenghai, et al.
Veröffentlicht: (2024)
RAGCap-Bench: Benchmarking Capabilities of LLMs in Agentic Retrieval Augmented Generation Systems
von: Lin, Jingru, et al.
Veröffentlicht: (2025)
von: Lin, Jingru, et al.
Veröffentlicht: (2025)
VLR-Bench: Multilingual Benchmark Dataset for Vision-Language Retrieval Augmented Generation
von: Lim, Hyeonseok, et al.
Veröffentlicht: (2024)
von: Lim, Hyeonseok, et al.
Veröffentlicht: (2024)
MultiSocial: Multilingual Benchmark of Machine-Generated Text Detection of Social-Media Texts
von: Macko, Dominik, et al.
Veröffentlicht: (2024)
von: Macko, Dominik, et al.
Veröffentlicht: (2024)
AlignBench: Benchmarking Chinese Alignment of Large Language Models
von: Liu, Xiao, et al.
Veröffentlicht: (2023)
von: Liu, Xiao, et al.
Veröffentlicht: (2023)
MCiteBench: A Multimodal Benchmark for Generating Text with Citations
von: Hu, Caiyu, et al.
Veröffentlicht: (2025)
von: Hu, Caiyu, et al.
Veröffentlicht: (2025)
TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation
von: Feng, Weixi, et al.
Veröffentlicht: (2024)
von: Feng, Weixi, et al.
Veröffentlicht: (2024)
LM$^2$otifs : An Explainable Framework for Machine-Generated Texts Detection
von: Zheng, Xu, et al.
Veröffentlicht: (2025)
von: Zheng, Xu, et al.
Veröffentlicht: (2025)
TransBench: Benchmarking Machine Translation for Industrial-Scale Applications
von: Li, Haijun, et al.
Veröffentlicht: (2025)
von: Li, Haijun, et al.
Veröffentlicht: (2025)
Benchmarking and Improving Compositional Generalization of Multi-aspect Controllable Text Generation
von: Zhong, Tianqi, et al.
Veröffentlicht: (2024)
von: Zhong, Tianqi, et al.
Veröffentlicht: (2024)
PodBench: A Comprehensive Benchmark for Instruction-Aware Audio-Oriented Podcast Script Generation
von: Xu, Chenning, et al.
Veröffentlicht: (2026)
von: Xu, Chenning, et al.
Veröffentlicht: (2026)
Struct-Bench: A Benchmark for Differentially Private Structured Text Generation
von: Wang, Shuaiqi, et al.
Veröffentlicht: (2025)
von: Wang, Shuaiqi, et al.
Veröffentlicht: (2025)
TAD-Bench: A Comprehensive Benchmark for Embedding-Based Text Anomaly Detection
von: Cao, Yang, et al.
Veröffentlicht: (2025)
von: Cao, Yang, et al.
Veröffentlicht: (2025)
MULTITuDE: Large-Scale Multilingual Machine-Generated Text Detection Benchmark
von: Macko, Dominik, et al.
Veröffentlicht: (2023)
von: Macko, Dominik, et al.
Veröffentlicht: (2023)
MIRAGE-Bench: Automatic Multilingual Benchmark Arena for Retrieval-Augmented Generation Systems
von: Thakur, Nandan, et al.
Veröffentlicht: (2024)
von: Thakur, Nandan, et al.
Veröffentlicht: (2024)
ReDeEP: Detecting Hallucination in Retrieval-Augmented Generation via Mechanistic Interpretability
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2024)
von: Sun, Zhongxiang, et al.
Veröffentlicht: (2024)
Exploring the Limitations of Detecting Machine-Generated Text
von: Doughman, Jad, et al.
Veröffentlicht: (2024)
von: Doughman, Jad, et al.
Veröffentlicht: (2024)
AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment
von: Xiao, Jianfei, et al.
Veröffentlicht: (2026)
von: Xiao, Jianfei, et al.
Veröffentlicht: (2026)
CEAID: Benchmark of Multilingual Machine-Generated Text Detection Methods for Central European Languages
von: Macko, Dominik, et al.
Veröffentlicht: (2025)
von: Macko, Dominik, et al.
Veröffentlicht: (2025)
T$^3$Bench: Benchmarking Current Progress in Text-to-3D Generation
von: He, Yuze, et al.
Veröffentlicht: (2023)
von: He, Yuze, et al.
Veröffentlicht: (2023)
MMKU-Bench: A Multimodal Update Benchmark for Diverse Visual Knowledge
von: Fu, Baochen, et al.
Veröffentlicht: (2026)
von: Fu, Baochen, et al.
Veröffentlicht: (2026)
StreamBench: Towards Benchmarking Continuous Improvement of Language Agents
von: Wu, Cheng-Kuang, et al.
Veröffentlicht: (2024)
von: Wu, Cheng-Kuang, et al.
Veröffentlicht: (2024)
AI Idea Bench 2025: AI Research Idea Generation Benchmark
von: Qiu, Yansheng, et al.
Veröffentlicht: (2025)
von: Qiu, Yansheng, et al.
Veröffentlicht: (2025)
Detecting Hallucinations in Graph Retrieval-Augmented Generation via Attention Patterns and Semantic Alignment
von: Li, Shanghao, et al.
Veröffentlicht: (2025)
von: Li, Shanghao, et al.
Veröffentlicht: (2025)
CameraBench: Benchmarking Visual Reasoning in MLLMs via Photography
von: Fang, I-Sheng, et al.
Veröffentlicht: (2025)
von: Fang, I-Sheng, et al.
Veröffentlicht: (2025)
AVGen-Bench: A Task-Driven Benchmark for Multi-Granular Evaluation of Text-to-Audio-Video Generation
von: Zhou, Ziwei, et al.
Veröffentlicht: (2026)
von: Zhou, Ziwei, et al.
Veröffentlicht: (2026)
When Personalization Tricks Detectors: The Feature-Inversion Trap in Machine-Generated Text Detection
von: Gao, Lang, et al.
Veröffentlicht: (2025)
von: Gao, Lang, et al.
Veröffentlicht: (2025)
SciRerankBench: Benchmarking Rerankers Towards Scientific Retrieval-Augmented Generated LLMs
von: Chen, Haotian, et al.
Veröffentlicht: (2025)
von: Chen, Haotian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
M4GT-Bench: Evaluation Benchmark for Black-Box Machine-Generated Text Detection
von: Wang, Yuxia, et al.
Veröffentlicht: (2024) -
MGT-Prism: Enhancing Domain Generalization for Machine-Generated Text Detection via Spectral Alignment
von: Liu, Shengchao, et al.
Veröffentlicht: (2025) -
R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation
von: Chen, Kaijie, et al.
Veröffentlicht: (2025) -
LCTG Bench: LLM Controlled Text Generation Benchmark
von: Kurihara, Kentaro, et al.
Veröffentlicht: (2025) -
RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment
von: Jin, Zhuoran, et al.
Veröffentlicht: (2024)