Is Your Paper Being Reviewed by an LLM? Benchmarking AI Text Detection in Peer Review
Fuente:
arXiv
Guardado en:
| Autores principales: | Yu, Sungduk, Luo, Man, Madasu, Avinash, Lal, Vasudev, Howard, Phillip |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Is Your Paper Being Reviewed by an LLM? Investigating AI Text Detectability in Peer Review
por: Yu, Sungduk, et al.
Publicado: (2024)
por: Yu, Sungduk, et al.
Publicado: (2024)
Cultural Awareness in Vision-Language Models: A Cross-Country Exploration
por: Madasu, Avinash, et al.
Publicado: (2025)
por: Madasu, Avinash, et al.
Publicado: (2025)
ICSVR: Investigating Compositional and Syntactic Understanding in Video Retrieval Models
por: Madasu, Avinash, et al.
Publicado: (2023)
por: Madasu, Avinash, et al.
Publicado: (2023)
Pruning the Paradox: How CLIP's Most Informative Heads Enhance Performance While Amplifying Bias
por: Madasu, Avinash, et al.
Publicado: (2025)
por: Madasu, Avinash, et al.
Publicado: (2025)
Quantifying and Enabling the Interpretability of CLIP-like Models
por: Madasu, Avinash, et al.
Publicado: (2024)
por: Madasu, Avinash, et al.
Publicado: (2024)
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression
por: Kundu, Souvik, et al.
Publicado: (2025)
por: Kundu, Souvik, et al.
Publicado: (2025)
SocialCounterfactuals: Probing and Mitigating Intersectional Social Biases in Vision-Language Models with Counterfactual Examples
por: Howard, Phillip, et al.
Publicado: (2023)
por: Howard, Phillip, et al.
Publicado: (2023)
Probing Semantic Routing in Large Mixture-of-Expert Models
por: Olson, Matthew Lyle, et al.
Publicado: (2025)
por: Olson, Matthew Lyle, et al.
Publicado: (2025)
Learning from Reasoning Failures via Synthetic Data Generation
por: Stan, Gabriela Ben Melech, et al.
Publicado: (2025)
por: Stan, Gabriela Ben Melech, et al.
Publicado: (2025)
PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewing
por: Żurawicki, Krzysztof, et al.
Publicado: (2026)
por: Żurawicki, Krzysztof, et al.
Publicado: (2026)
A Causal World Model Underlying Next Token Prediction: Exploring GPT in a Controlled Environment
por: Rohekar, Raanan Y., et al.
Publicado: (2024)
por: Rohekar, Raanan Y., et al.
Publicado: (2024)
NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
por: Rosenman, Shachar, et al.
Publicado: (2023)
por: Rosenman, Shachar, et al.
Publicado: (2023)
Training-Free Mitigation of Language Reasoning Degradation After Multimodal Instruction Tuning
por: Ratzlaff, Neale, et al.
Publicado: (2024)
por: Ratzlaff, Neale, et al.
Publicado: (2024)
Can AI Be a Good Peer Reviewer? A Survey of Peer Review Process, Evaluation, and the Future
por: Wu, Sihong, et al.
Publicado: (2026)
por: Wu, Sihong, et al.
Publicado: (2026)
Benchmark on Peer Review Toxic Detection: A Challenging Task with a New Dataset
por: Luo, Man, et al.
Publicado: (2025)
por: Luo, Man, et al.
Publicado: (2025)
Sarang at DEFACTIFY 4.0: Detecting AI-Generated Text Using Noised Data and an Ensemble of DeBERTa Models
por: Trivedi, Avinash, et al.
Publicado: (2025)
por: Trivedi, Avinash, et al.
Publicado: (2025)
Dual Optimal: Make Your LLM Peer-like with Dignity
por: Wang, Xiangqi, et al.
Publicado: (2026)
por: Wang, Xiangqi, et al.
Publicado: (2026)
ReviewerToo: Should AI Join The Program Committee? A Look At The Future of Peer Review
por: Sahu, Gaurav, et al.
Publicado: (2025)
por: Sahu, Gaurav, et al.
Publicado: (2025)
Is Peer-Reviewing Worth the Effort?
por: Church, Kenneth, et al.
Publicado: (2024)
por: Church, Kenneth, et al.
Publicado: (2024)
EchoReview: Learning Peer Review from the Echoes of Scientific Citations
por: Zhang, Yinuo, et al.
Publicado: (2026)
por: Zhang, Yinuo, et al.
Publicado: (2026)
LLM Review: Enhancing Creative Writing via Blind Peer Review Feedback
por: Li, Weiyue, et al.
Publicado: (2026)
por: Li, Weiyue, et al.
Publicado: (2026)
Detecting AI-Generated Content in Academic Peer Reviews
por: Shen, Siyuan, et al.
Publicado: (2026)
por: Shen, Siyuan, et al.
Publicado: (2026)
p-less Sampling: A Robust Hyperparameter-Free Approach for LLM Decoding
por: Tan, Runyan, et al.
Publicado: (2025)
por: Tan, Runyan, et al.
Publicado: (2025)
CoCoNUTS: Concentrating on Content while Neglecting Uninformative Textual Styles for AI-Generated Peer Review Detection
por: Chen, Yihan, et al.
Publicado: (2025)
por: Chen, Yihan, et al.
Publicado: (2025)
When Reviews Disagree: Fine-Grained Contradiction Analysis in Scientific Peer Reviews
por: Kumar, Sandeep, et al.
Publicado: (2026)
por: Kumar, Sandeep, et al.
Publicado: (2026)
DPO Learning with LLMs-Judge Signal for Computer Use Agents
por: Luo, Man, et al.
Publicado: (2025)
por: Luo, Man, et al.
Publicado: (2025)
Breaking the Reviewer: Assessing the Vulnerability of Large Language Models in Automated Peer Review Under Textual Adversarial Attacks
por: Lin, Tzu-Ling, et al.
Publicado: (2025)
por: Lin, Tzu-Ling, et al.
Publicado: (2025)
Cognitive-Mental-LLM: Evaluating Reasoning in Large Language Models for Mental Health Prediction via Online Text
por: Patil, Avinash, et al.
Publicado: (2025)
por: Patil, Avinash, et al.
Publicado: (2025)
Support-Contra Asymmetry in LLM Explanations
por: Patil, Avinash
Publicado: (2025)
por: Patil, Avinash
Publicado: (2025)
DetectRL: Benchmarking LLM-Generated Text Detection in Real-World Scenarios
por: Wu, Junchao, et al.
Publicado: (2024)
por: Wu, Junchao, et al.
Publicado: (2024)
CoCoReviewBench: A Completeness- and Correctness-Oriented Benchmark for AI Reviewers
por: Deng, Hexuan, et al.
Publicado: (2026)
por: Deng, Hexuan, et al.
Publicado: (2026)
Hidden Prompts in Manuscripts Exploit AI-Assisted Peer Review
por: Lin, Zhicheng
Publicado: (2025)
por: Lin, Zhicheng
Publicado: (2025)
Policies Permitting LLM Use for Polishing Peer Reviews Are Currently Not Enforceable
por: Saha, Rounak, et al.
Publicado: (2026)
por: Saha, Rounak, et al.
Publicado: (2026)
Steering Large Language Models to Evaluate and Amplify Creativity
por: Olson, Matthew Lyle, et al.
Publicado: (2024)
por: Olson, Matthew Lyle, et al.
Publicado: (2024)
LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Model
por: Hinck, Musashi, et al.
Publicado: (2024)
por: Hinck, Musashi, et al.
Publicado: (2024)
ReviewRobot: Explainable Paper Review Generation based on Knowledge Synthesis
por: Wang, Qingyun, et al.
Publicado: (2020)
por: Wang, Qingyun, et al.
Publicado: (2020)
SEAGraph: Unveiling the Whole Story of Paper Review Comments
por: Yu, Jianxiang, et al.
Publicado: (2024)
por: Yu, Jianxiang, et al.
Publicado: (2024)
When Reject Turns into Accept: Quantifying the Vulnerability of LLM-Based Scientific Reviewers to Indirect Prompt Injection
por: Sahoo, Devanshu, et al.
Publicado: (2025)
por: Sahoo, Devanshu, et al.
Publicado: (2025)
PiCO: Peer Review in LLMs based on the Consistency Optimization
por: Ning, Kun-Peng, et al.
Publicado: (2024)
por: Ning, Kun-Peng, et al.
Publicado: (2024)
When Informal Text Breaks NLI: Tokenization Failure, Distribution Shift, and Targeted Mitigations
por: Aluguvelly, Avinash Goutham
Publicado: (2026)
por: Aluguvelly, Avinash Goutham
Publicado: (2026)
Ejemplares similares
-
Is Your Paper Being Reviewed by an LLM? Investigating AI Text Detectability in Peer Review
por: Yu, Sungduk, et al.
Publicado: (2024) -
Cultural Awareness in Vision-Language Models: A Cross-Country Exploration
por: Madasu, Avinash, et al.
Publicado: (2025) -
ICSVR: Investigating Compositional and Syntactic Understanding in Video Retrieval Models
por: Madasu, Avinash, et al.
Publicado: (2023) -
Pruning the Paradox: How CLIP's Most Informative Heads Enhance Performance While Amplifying Bias
por: Madasu, Avinash, et al.
Publicado: (2025) -
Quantifying and Enabling the Interpretability of CLIP-like Models
por: Madasu, Avinash, et al.
Publicado: (2024)