Advancing Machine-Generated Text Detection from an Easy to Hard Supervision Perspective
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Chenwang, Cheung, Yiu-ming, Han, Bo, Lian, Defu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hidden Human-Like Nature of Machine-Generated Texts: Theory and Detection Enhancement
by: Wu, Chenwang, et al.
Published: (2026)
by: Wu, Chenwang, et al.
Published: (2026)
Beyond Raw Detection Scores: Markov-Informed Calibration for Boosting Machine-Generated Text Detection
by: Wu, Chenwang, et al.
Published: (2026)
by: Wu, Chenwang, et al.
Published: (2026)
Multi-Level Contextual Token Relation Modeling for Machine-Generated Text Detection
by: Wu, Chenwang, et al.
Published: (2026)
by: Wu, Chenwang, et al.
Published: (2026)
Learning to Substitute Components for Compositional Generalization
by: Li, Zhaoyi, et al.
Published: (2025)
by: Li, Zhaoyi, et al.
Published: (2025)
Understanding Privacy Risks of Embeddings Induced by Large Language Models
by: Zhu, Zhihao, et al.
Published: (2024)
by: Zhu, Zhihao, et al.
Published: (2024)
Beyond Easy Wins: A Text Hardness-Aware Benchmark for LLM-generated Text Detection
by: Ayoobi, Navid, et al.
Published: (2025)
by: Ayoobi, Navid, et al.
Published: (2025)
Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision
by: Sun, Zhiqing, et al.
Published: (2024)
by: Sun, Zhiqing, et al.
Published: (2024)
Bridging the Semantic Gap for Categorical Data Clustering via Large Language Models
by: Yang, Zihua, et al.
Published: (2026)
by: Yang, Zihua, et al.
Published: (2026)
When Large Language Models Meet Personalization: Perspectives of Challenges and Opportunities
by: Chen, Jin, et al.
Published: (2023)
by: Chen, Jin, et al.
Published: (2023)
Epistemic Uncertainty for Generated Image Detection
by: Nie, Jun, et al.
Published: (2024)
by: Nie, Jun, et al.
Published: (2024)
Efficient Machine Unlearning via Influence Approximation
by: Liu, Jiawei, et al.
Published: (2025)
by: Liu, Jiawei, et al.
Published: (2025)
Benchmarking and Improving Compositional Generalization of Multi-aspect Controllable Text Generation
by: Zhong, Tianqi, et al.
Published: (2024)
by: Zhong, Tianqi, et al.
Published: (2024)
Interpreting and Improving Large Language Models in Arithmetic Calculation
by: Zhang, Wei, et al.
Published: (2024)
by: Zhang, Wei, et al.
Published: (2024)
TYrPPG: Uncomplicated and Enhanced Learning Capability rPPG for Remote Heart Rate Estimation
by: Chen, Taixi, et al.
Published: (2025)
by: Chen, Taixi, et al.
Published: (2025)
MetaLint: Easy-to-Hard Generalization for Code Linting
by: Naik, Atharva, et al.
Published: (2025)
by: Naik, Atharva, et al.
Published: (2025)
AI "News" Content Farms Are Easy to Make and Hard to Detect: A Case Study in Italian
by: Puccetti, Giovanni, et al.
Published: (2024)
by: Puccetti, Giovanni, et al.
Published: (2024)
Emergent Misalignment is Easy, Narrow Misalignment is Hard
by: Soligo, Anna, et al.
Published: (2026)
by: Soligo, Anna, et al.
Published: (2026)
On the Role of Reasoning Patterns in the Generalization Discrepancy of Long Chain-of-Thought Supervised Fine-Tuning
by: Li, Zhaoyi, et al.
Published: (2026)
by: Li, Zhaoyi, et al.
Published: (2026)
Detecting Machine-Generated Texts by Multi-Population Aware Optimization for Maximum Mean Discrepancy
by: Zhang, Shuhai, et al.
Published: (2024)
by: Zhang, Shuhai, et al.
Published: (2024)
O1 Embedder: Let Retrievers Think Before Action
by: Yan, Ruiran, et al.
Published: (2025)
by: Yan, Ruiran, et al.
Published: (2025)
ReasonEmbed: Enhanced Text Embeddings for Reasoning-Intensive Document Retrieval
by: Chen, Jianlyu, et al.
Published: (2025)
by: Chen, Jianlyu, et al.
Published: (2025)
Automating Easy Read Text Segmentation
by: Calleja, Jesús, et al.
Published: (2024)
by: Calleja, Jesús, et al.
Published: (2024)
Mitigate Negative Transfer with Similarity Heuristic Lifelong Prompt Tuning
by: Wu, Chenyuan, et al.
Published: (2024)
by: Wu, Chenyuan, et al.
Published: (2024)
Lighter And Better: Towards Flexible Context Adaptation For Retrieval Augmented Generation
by: Liu, Zheng, et al.
Published: (2024)
by: Liu, Zheng, et al.
Published: (2024)
Easy2Hard-Bench: Standardized Difficulty Labels for Profiling LLM Performance and Generalization
by: Ding, Mucong, et al.
Published: (2024)
by: Ding, Mucong, et al.
Published: (2024)
SPADE: Structured Prompting Augmentation for Dialogue Enhancement in Machine-Generated Text Detection
by: Li, Haoyi, et al.
Published: (2025)
by: Li, Haoyi, et al.
Published: (2025)
Exploring the Limitations of Detecting Machine-Generated Text
by: Doughman, Jad, et al.
Published: (2024)
by: Doughman, Jad, et al.
Published: (2024)
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
by: Ling, Zehui, et al.
Published: (2025)
by: Ling, Zehui, et al.
Published: (2025)
Rethinking Easy-to-Hard: Limits of Curriculum Learning in Post-Training for Deductive Reasoning
by: Mordig, Maximilian, et al.
Published: (2026)
by: Mordig, Maximilian, et al.
Published: (2026)
Ask, Attend, Attack: A Effective Decision-Based Black-Box Targeted Attack for Image-to-Text Models
by: Zeng, Qingyuan, et al.
Published: (2024)
by: Zeng, Qingyuan, et al.
Published: (2024)
EMMM, Explain Me My Model! Explainable Machine Generated Text Detection in Dialogues
by: Yuan, Angela Yifei, et al.
Published: (2025)
by: Yuan, Angela Yifei, et al.
Published: (2025)
The Unreasonable Effectiveness of Easy Training Data for Hard Tasks
by: Hase, Peter, et al.
Published: (2024)
by: Hase, Peter, et al.
Published: (2024)
Authorship Obfuscation in Multilingual Machine-Generated Text Detection
by: Macko, Dominik, et al.
Published: (2024)
by: Macko, Dominik, et al.
Published: (2024)
Semantic-guided Fine-tuning of Foundation Model for Long-tailed Visual Recognition
by: Peng, Yufei, et al.
Published: (2025)
by: Peng, Yufei, et al.
Published: (2025)
EH-MAM: Easy-to-Hard Masked Acoustic Modeling for Self-Supervised Speech Representation Learning
by: Seth, Ashish, et al.
Published: (2024)
by: Seth, Ashish, et al.
Published: (2024)
EasyGen: Easing Multimodal Generation with BiDiffuser and LLMs
by: Zhao, Xiangyu, et al.
Published: (2023)
by: Zhao, Xiangyu, et al.
Published: (2023)
LLM Cache Bandit Revisited: Addressing Query Heterogeneity for Cost-Effective LLM Inference
by: Yang, Hantao, et al.
Published: (2025)
by: Yang, Hantao, et al.
Published: (2025)
RAPID: Efficient Retrieval-Augmented Long Text Generation with Writing Planning and Information Discovery
by: Gu, Hongchao, et al.
Published: (2025)
by: Gu, Hongchao, et al.
Published: (2025)
Language Models Do Hard Arithmetic Tasks Easily and Hardly Do Easy Arithmetic Tasks
by: Gambardella, Andrew, et al.
Published: (2024)
by: Gambardella, Andrew, et al.
Published: (2024)
Detection of Machine-Generated Text: Literature Survey
by: Valiaiev, Dmytro
Published: (2024)
by: Valiaiev, Dmytro
Published: (2024)
Similar Items
-
Hidden Human-Like Nature of Machine-Generated Texts: Theory and Detection Enhancement
by: Wu, Chenwang, et al.
Published: (2026) -
Beyond Raw Detection Scores: Markov-Informed Calibration for Boosting Machine-Generated Text Detection
by: Wu, Chenwang, et al.
Published: (2026) -
Multi-Level Contextual Token Relation Modeling for Machine-Generated Text Detection
by: Wu, Chenwang, et al.
Published: (2026) -
Learning to Substitute Components for Compositional Generalization
by: Li, Zhaoyi, et al.
Published: (2025) -
Understanding Privacy Risks of Embeddings Induced by Large Language Models
by: Zhu, Zhihao, et al.
Published: (2024)