CaughtCheating: Is Your MLLM a Good Cheating Detective? Exploring the Boundary of Visual Perception and Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Ming, Wang, Chenguang, Liang, Yijun, Wang, Xiyao, Zhou, Yuhang, Wu, Xiyang, Zhang, Yuqing, Zhang, Ruiyi, Zhou, Tianyi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Perception of Prevalence of Cheating and Usage of Generative AI
by: Denkin, Roman
Published: (2024)
by: Denkin, Roman
Published: (2024)
Cheating the Kids.
by: Meltzer, Bonnie
Published: (2000)
by: Meltzer, Bonnie
Published: (2000)
Is It Thinking or Cheating? Detecting Implicit Reward Hacking by Measuring Reasoning Effort
by: Wang, Xinpeng, et al.
Published: (2025)
by: Wang, Xinpeng, et al.
Published: (2025)
Distilling Many-Shot In-Context Learning into a Cheat Sheet
by: Honda, Ukyo, et al.
Published: (2025)
by: Honda, Ukyo, et al.
Published: (2025)
Human-in-the-Loop AI for Cheating Ring Detection
by: Shih, Yong-Siang, et al.
Published: (2024)
by: Shih, Yong-Siang, et al.
Published: (2024)
Balancing The Perception of Cheating Detection, Privacy and Fairness: A Mixed-Methods Study of Visual Data Obfuscation in Remote Proctoring
by: Mukherjee, Suvadeep, et al.
Published: (2024)
by: Mukherjee, Suvadeep, et al.
Published: (2024)
Towards Visual Text Grounding of Multimodal Large Language Model
by: Li, Ming, et al.
Published: (2025)
by: Li, Ming, et al.
Published: (2025)
Out-of-Bounding-Box Triggers: A Stealthy Approach to Cheat Object Detectors
by: Lin, Tao, et al.
Published: (2024)
by: Lin, Tao, et al.
Published: (2024)
LLM Use, Cheating, and Academic Integrity in Software Engineering Education
by: Santos, Ronnie de Souza, et al.
Published: (2026)
by: Santos, Ronnie de Souza, et al.
Published: (2026)
Impeding LLM-assisted Cheating in Introductory Programming Assignments via Adversarial Perturbation
by: Salim, Saiful Islam, et al.
Published: (2024)
by: Salim, Saiful Islam, et al.
Published: (2024)
Synopticon: Consensus-Based Cheating Detection System for Competitive Games
by: Kang, Jeuk, et al.
Published: (2025)
by: Kang, Jeuk, et al.
Published: (2025)
VisR-Bench: An Empirical Study on Visual Retrieval-Augmented Generation for Multilingual Long Document Understanding
by: Chen, Jian, et al.
Published: (2025)
by: Chen, Jian, et al.
Published: (2025)
Assessing the Prevalence of AI-assisted Cheating in Programming Courses: A Pilot Study
by: Delphino, Kaléu
Published: (2025)
by: Delphino, Kaléu
Published: (2025)
A Study on the Vulnerability of Test Questions against ChatGPT-based Cheating
by: Ram, Shanker, et al.
Published: (2024)
by: Ram, Shanker, et al.
Published: (2024)
Multiple Instance Learning for Cheating Detection and Localization in Online Examinations
by: Liu, Yemeng, et al.
Published: (2024)
by: Liu, Yemeng, et al.
Published: (2024)
LLM-Assisted Cheating Detection in Korean Language via Keystrokes
by: Roh, Dong Hyun, et al.
Published: (2025)
by: Roh, Dong Hyun, et al.
Published: (2025)
Ref-Adv: Exploring MLLM Visual Reasoning in Referring Expression Tasks
by: Dong, Qihua, et al.
Published: (2026)
by: Dong, Qihua, et al.
Published: (2026)
Detecting AI-Assisted Cheating in Online Exams through Behavior Analytics
by: Akçapınar, Gökhan
Published: (2025)
by: Akçapınar, Gökhan
Published: (2025)
What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding
by: Li, Ming, et al.
Published: (2025)
by: Li, Ming, et al.
Published: (2025)
WIP: Exploring the Value of a Debugging Cheat Sheet and Mini Lecture in Improving Undergraduate Debugging Skills and Mindset
by: Ash, Andrew, et al.
Published: (2025)
by: Ash, Andrew, et al.
Published: (2025)
V-REX: Benchmarking Exploratory Visual Reasoning via Chain-of-Questions
by: Fan, Chenrui, et al.
Published: (2025)
by: Fan, Chenrui, et al.
Published: (2025)
How Much Do Large Language Model Cheat on Evaluation? Benchmarking Overestimation under the One-Time-Pad-Based Framework
by: Liang, Zi, et al.
Published: (2025)
by: Liang, Zi, et al.
Published: (2025)
Leak, Cheat, Repeat: Data Contamination and Evaluation Malpractices in Closed-Source LLMs
by: Balloccu, Simone, et al.
Published: (2024)
by: Balloccu, Simone, et al.
Published: (2024)
Mosaic-IT: Cost-Free Compositional Data Synthesis for Instruction Tuning
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement
by: Wang, Xiyao, et al.
Published: (2024)
by: Wang, Xiyao, et al.
Published: (2024)
Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates
by: Zheng, Xiaosen, et al.
Published: (2024)
by: Zheng, Xiaosen, et al.
Published: (2024)
GoT-R1: Unleashing Reasoning Capability of MLLM for Visual Generation with Reinforcement Learning
by: Duan, Chengqi, et al.
Published: (2025)
by: Duan, Chengqi, et al.
Published: (2025)
AbductiveMLLM: Boosting Visual Abductive Reasoning Within MLLMs
by: Chang, Boyu, et al.
Published: (2026)
by: Chang, Boyu, et al.
Published: (2026)
Time Travel is Cheating: Going Live with DeepFund for Real-Time Fund Investment Benchmarking
by: Li, Changlun, et al.
Published: (2025)
by: Li, Changlun, et al.
Published: (2025)
Cheating. An Annotated Bibliography.
by: Wildemuth, Barbara M., Comp.
Published: (1976)
by: Wildemuth, Barbara M., Comp.
Published: (1976)
Is It Cheating if Everybody Does It?
by: Gustafon, Chris
Published: (2004)
by: Gustafon, Chris
Published: (2004)
Is Your Model Really A Good Math Reasoner? Evaluating Mathematical Reasoning with Checklist
by: Zhou, Zihao, et al.
Published: (2024)
by: Zhou, Zihao, et al.
Published: (2024)
ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs
by: Wang, Xiyao, et al.
Published: (2025)
by: Wang, Xiyao, et al.
Published: (2025)
Elysium: Exploring Object-level Perception in Videos via MLLM
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
Treat Visual Tokens as Text? But Your MLLM Only Needs Fewer Efforts to See
by: Zhang, Zeliang, et al.
Published: (2024)
by: Zhang, Zeliang, et al.
Published: (2024)
Identify As A Human Does: A Pathfinder of Next-Generation Anti-Cheat Framework for First-Person Shooter Games
by: Zhang, Jiayi, et al.
Published: (2024)
by: Zhang, Jiayi, et al.
Published: (2024)
Syn-GRPO: Self-Evolving Data Synthesis for MLLM Perception Reasoning
by: Huang, Qihan, et al.
Published: (2025)
by: Huang, Qihan, et al.
Published: (2025)
ColorBench: Can VLMs See and Understand the Colorful World? A Comprehensive Benchmark for Color Perception, Reasoning, and Robustness
by: Liang, Yijun, et al.
Published: (2025)
by: Liang, Yijun, et al.
Published: (2025)
Your Reasoning Benchmark May Not Test Reasoning: Revealing Perception Bottleneck in Abstract Reasoning Benchmarks
by: Wang, Xinhe, et al.
Published: (2025)
by: Wang, Xinhe, et al.
Published: (2025)
What does RL improve for Visual Reasoning? A Frankenstein-Style Analysis
by: Li, Xirui, et al.
Published: (2026)
by: Li, Xirui, et al.
Published: (2026)
Similar Items
-
On Perception of Prevalence of Cheating and Usage of Generative AI
by: Denkin, Roman
Published: (2024) -
Cheating the Kids.
by: Meltzer, Bonnie
Published: (2000) -
Is It Thinking or Cheating? Detecting Implicit Reward Hacking by Measuring Reasoning Effort
by: Wang, Xinpeng, et al.
Published: (2025) -
Distilling Many-Shot In-Context Learning into a Cheat Sheet
by: Honda, Ukyo, et al.
Published: (2025) -
Human-in-the-Loop AI for Cheating Ring Detection
by: Shih, Yong-Siang, et al.
Published: (2024)