Detecting LLM-Generated Peer Reviews
Fuente:
arXiv
Saved in:
| Main Authors: | Rao, Vishisht, Kumar, Aounon, Lakkaraju, Himabindu, Shah, Nihar B. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Identity Theft in AI Conference Peer Review
by: Shah, Nihar B., et al.
Published: (2025)
by: Shah, Nihar B., et al.
Published: (2025)
Vulnerability of Text-Matching in ML/AI Conference Reviewer Assignments to Collusions
by: Hsieh, Jhih-Yi, et al.
Published: (2024)
by: Hsieh, Jhih-Yi, et al.
Published: (2024)
ML Researchers Support Openness in Peer Review But Are Concerned About Resubmission Bias
by: Rao, Vishisht, et al.
Published: (2025)
by: Rao, Vishisht, et al.
Published: (2025)
Certifying LLM Safety against Adversarial Prompting
by: Kumar, Aounon, et al.
Published: (2023)
by: Kumar, Aounon, et al.
Published: (2023)
FLAWS: A Benchmark for Error Identification and Localization in Scientific Papers
by: Xi, Sarina, et al.
Published: (2025)
by: Xi, Sarina, et al.
Published: (2025)
From Content Creation to Citation Inflation: A GenAI Case Study
by: Al-Sinani, Haitham S., et al.
Published: (2025)
by: Al-Sinani, Haitham S., et al.
Published: (2025)
Enhancing Peer Review in Astronomy: A Machine Learning and Optimization Approach to Reviewer Assignments for ALMA
by: Carpenter, John M., et al.
Published: (2024)
by: Carpenter, John M., et al.
Published: (2024)
In-Context Unlearning: Language Models as Few Shot Unlearners
by: Pawelczyk, Martin, et al.
Published: (2023)
by: Pawelczyk, Martin, et al.
Published: (2023)
AoI-Guided Client Selection for Robust and Timely Federated Intrusion Detection in Cloud-Edge Security Analytics
by: Chiu, Chun Yin
Published: (2026)
by: Chiu, Chun Yin
Published: (2026)
De-identification of clinical free text using natural language processing: A systematic review of current approaches
by: Kovačević, Aleksandar, et al.
Published: (2023)
by: Kovačević, Aleksandar, et al.
Published: (2023)
Manipulating Large Language Models to Increase Product Visibility
by: Kumar, Aounon, et al.
Published: (2024)
by: Kumar, Aounon, et al.
Published: (2024)
The More You Automate, the Less You See: Hidden Pitfalls of AI Scientist Systems
by: Luo, Ziming, et al.
Published: (2025)
by: Luo, Ziming, et al.
Published: (2025)
Mens Sana In Corpore Sano: Sound Firmware Corpora for Vulnerability Research
by: Helmke, René, et al.
Published: (2024)
by: Helmke, René, et al.
Published: (2024)
Data Poisoning Attacks on Off-Policy Policy Evaluation Methods
by: Lobo, Elita, et al.
Published: (2024)
by: Lobo, Elita, et al.
Published: (2024)
Follow My Instruction and Spill the Beans: Scalable Data Extraction from Retrieval-Augmented Generation Systems
by: Qi, Zhenting, et al.
Published: (2024)
by: Qi, Zhenting, et al.
Published: (2024)
IANEC: Digital Forensic Investigation of Contemporary Writers' Archives
by: Giguet, Emmanuel
Published: (2025)
by: Giguet, Emmanuel
Published: (2025)
The Feasibility of Topic-Based Watermarking on Academic Peer Reviews
by: Nemecek, Alexander, et al.
Published: (2025)
by: Nemecek, Alexander, et al.
Published: (2025)
Shattering the Echo Chamber: Hidden Safeguards in Manuscripts Against the AI Takeover of Peer Review
by: Ma, Oubo, et al.
Published: (2026)
by: Ma, Oubo, et al.
Published: (2026)
MedSafetyBench: Evaluating and Improving the Medical Safety of Large Language Models
by: Han, Tessa, et al.
Published: (2024)
by: Han, Tessa, et al.
Published: (2024)
Compound Deception in Elite Peer Review: A Failure Mode Taxonomy of 100 Fabricated Citations at NeurIPS 2025
by: Ansari, Samar
Published: (2026)
by: Ansari, Samar
Published: (2026)
Explaining the Model, Protecting Your Data: Revealing and Mitigating the Data Privacy Risks of Post-Hoc Model Explanations via Membership Inference
by: Huang, Catherine, et al.
Published: (2024)
by: Huang, Catherine, et al.
Published: (2024)
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities
by: Ferrag, Mohamed Amine, et al.
Published: (2024)
by: Ferrag, Mohamed Amine, et al.
Published: (2024)
Modification and Generated-Text Detection: Achieving Dual Detection Capabilities for the Outputs of LLM by Watermark
by: Cai, Yuhang, et al.
Published: (2025)
by: Cai, Yuhang, et al.
Published: (2025)
VerilogLAVD: LLM-Aided Rule Generation for Vulnerability Detection in Verilog
by: Long, Xiang, et al.
Published: (2025)
by: Long, Xiang, et al.
Published: (2025)
EXPLICATE: Enhancing Phishing Detection through Explainable AI and LLM-Powered Interpretability
by: Lim, Bryan, et al.
Published: (2025)
by: Lim, Bryan, et al.
Published: (2025)
Causal Effect of Group Diversity on Redundancy and Coverage in Peer-Reviewing
by: Goyal, Navita, et al.
Published: (2024)
by: Goyal, Navita, et al.
Published: (2024)
Marking Code Without Breaking It: Code Watermarking for Detecting LLM-Generated Code
by: Kim, Jungin, et al.
Published: (2025)
by: Kim, Jungin, et al.
Published: (2025)
Peer Reviews of Peer Reviews: A Randomized Controlled Trial and Other Experiments
by: Goldberg, Alexander, et al.
Published: (2023)
by: Goldberg, Alexander, et al.
Published: (2023)
SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code
by: Li, Xinghang, et al.
Published: (2025)
by: Li, Xinghang, et al.
Published: (2025)
Pre-review to Peer review: Pitfalls of Automating Reviews using Large Language Models
by: Akella, Akhil Pandey, et al.
Published: (2025)
by: Akella, Akhil Pandey, et al.
Published: (2025)
Information Security Based on LLM Approaches: A Review
by: Gong, Chang, et al.
Published: (2025)
by: Gong, Chang, et al.
Published: (2025)
VulnLLM-R: Specialized Reasoning LLM with Agent Scaffold for Vulnerability Detection
by: Nie, Yuzhou, et al.
Published: (2025)
by: Nie, Yuzhou, et al.
Published: (2025)
A Randomized Controlled Trial on Anonymizing Reviewers to Each Other in Peer Review Discussions
by: Rastogi, Charvi, et al.
Published: (2024)
by: Rastogi, Charvi, et al.
Published: (2024)
LProtector: An LLM-driven Vulnerability Detection System
by: Sheng, Ze, et al.
Published: (2024)
by: Sheng, Ze, et al.
Published: (2024)
CiteCheck: Retrieval-Grounded Detection of LLM Citation Hallucinations in Scientific Text
by: Khajavi, Khashayar, et al.
Published: (2026)
by: Khajavi, Khashayar, et al.
Published: (2026)
AiRacleX: Automated Detection of Price Oracle Manipulations via LLM-Driven Knowledge Mining and Prompt Generation
by: Gao, Bo, et al.
Published: (2025)
by: Gao, Bo, et al.
Published: (2025)
Generalizing Trust: Weak-to-Strong Trustworthiness in Language Models
by: Pawelczyk, Martin, et al.
Published: (2024)
by: Pawelczyk, Martin, et al.
Published: (2024)
Latent Adversarial Detection: Adaptive Probing of LLM Activations for Multi-Turn Attack Detection
by: Kulkarni, Prashant
Published: (2026)
by: Kulkarni, Prashant
Published: (2026)
An LLM Framework For Cryptography Over Chat Channels
by: Gligoroski, Danilo, et al.
Published: (2025)
by: Gligoroski, Danilo, et al.
Published: (2025)
BadLLM-TG: A Backdoor Defender powered by LLM Trigger Generator
by: Zhang, Ruyi, et al.
Published: (2026)
by: Zhang, Ruyi, et al.
Published: (2026)
Similar Items
-
Identity Theft in AI Conference Peer Review
by: Shah, Nihar B., et al.
Published: (2025) -
Vulnerability of Text-Matching in ML/AI Conference Reviewer Assignments to Collusions
by: Hsieh, Jhih-Yi, et al.
Published: (2024) -
ML Researchers Support Openness in Peer Review But Are Concerned About Resubmission Bias
by: Rao, Vishisht, et al.
Published: (2025) -
Certifying LLM Safety against Adversarial Prompting
by: Kumar, Aounon, et al.
Published: (2023) -
FLAWS: A Benchmark for Error Identification and Localization in Scientific Papers
by: Xi, Sarina, et al.
Published: (2025)