Policies Permitting LLM Use for Polishing Peer Reviews Are Currently Not Enforceable
Fuente:
arXiv
Saved in:
| Main Authors: | Saha, Rounak, Juneja, Gurusha, Chaudhuri, Dayita, Sajeevan, Naveeja, Shah, Nihar B, Pruthi, Danish |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Richer Output for Richer Countries: Uncovering Geographical Disparities in Generated Stories and Travel Recommendations
by: Bhagat, Kirti, et al.
Published: (2024)
by: Bhagat, Kirti, et al.
Published: (2024)
TALES: A Taxonomy and Analysis of Cultural Representations in LLM-generated Stories
by: Bhagat, Kirti, et al.
Published: (2025)
by: Bhagat, Kirti, et al.
Published: (2025)
Reviewing the Reviewer: Elevating Peer Review Quality through LLM-Guided Feedback
by: Purkayastha, Sukannya, et al.
Published: (2026)
by: Purkayastha, Sukannya, et al.
Published: (2026)
A Principled Approach to Randomized Selection under Uncertainty: Applications to Peer Review and Grant Funding
by: Goldberg, Alexander, et al.
Published: (2025)
by: Goldberg, Alexander, et al.
Published: (2025)
$\texttt{LM}^\texttt{2}$: A Simple Society of Language Models Solves Complex Reasoning
by: Juneja, Gurusha, et al.
Published: (2024)
by: Juneja, Gurusha, et al.
Published: (2024)
All That Glitters is Not Novel: Plagiarism in AI Generated Research
by: Gupta, Tarun, et al.
Published: (2025)
by: Gupta, Tarun, et al.
Published: (2025)
Benchmarking LLMs for Environmental Review and Permitting
by: Meyur, Rounak, et al.
Published: (2024)
by: Meyur, Rounak, et al.
Published: (2024)
MAGPIE: A dataset for Multi-AGent contextual PrIvacy Evaluation
by: Juneja, Gurusha, et al.
Published: (2025)
by: Juneja, Gurusha, et al.
Published: (2025)
A Randomized Controlled Trial on Anonymizing Reviewers to Each Other in Peer Review Discussions
by: Rastogi, Charvi, et al.
Published: (2024)
by: Rastogi, Charvi, et al.
Published: (2024)
Benchmark on Peer Review Toxic Detection: A Challenging Task with a New Dataset
by: Luo, Man, et al.
Published: (2025)
by: Luo, Man, et al.
Published: (2025)
Revisiting the Robustness of Watermarking to Paraphrasing Attacks
by: Rastogi, Saksham, et al.
Published: (2024)
by: Rastogi, Saksham, et al.
Published: (2024)
Evaluating Reasoning Models for Queries with Presuppositions
by: Sathyanathan, Rose, et al.
Published: (2026)
by: Sathyanathan, Rose, et al.
Published: (2026)
Towards an Enforceable GDPR Specification
by: Hublet, François, et al.
Published: (2024)
by: Hublet, François, et al.
Published: (2024)
Small Language Models Fine-tuned to Coordinate Larger Language Models improve Complex Reasoning
by: Juneja, Gurusha, et al.
Published: (2023)
by: Juneja, Gurusha, et al.
Published: (2023)
Commitment Checklist: Auditing Author Commitments in Peer Review
by: Chen, Chung-Chi, et al.
Published: (2026)
by: Chen, Chung-Chi, et al.
Published: (2026)
Causal Effect of Group Diversity on Redundancy and Coverage in Peer-Reviewing
by: Goyal, Navita, et al.
Published: (2024)
by: Goyal, Navita, et al.
Published: (2024)
SafeMath: Inference-time Safety improves Math Accuracy
by: Basu, Sagnik, et al.
Published: (2026)
by: Basu, Sagnik, et al.
Published: (2026)
What Can Natural Language Processing Do for Peer Review?
by: Kuznetsov, Ilia, et al.
Published: (2024)
by: Kuznetsov, Ilia, et al.
Published: (2024)
PolicyLLM: Towards Excellent Comprehension of Public Policy for Large Language Models
by: Bao, Han, et al.
Published: (2026)
by: Bao, Han, et al.
Published: (2026)
DeepReviewer 2.0: A Traceable Agentic System for Auditable Scientific Peer Review
by: Weng, Yixuan, et al.
Published: (2026)
by: Weng, Yixuan, et al.
Published: (2026)
The Good, the Bad and the Constructive: Automatically Measuring Peer Review's Utility for Authors
by: Sadallah, Abdelrahman, et al.
Published: (2025)
by: Sadallah, Abdelrahman, et al.
Published: (2025)
Do Voters Get the Information They Want? Understanding Authentic Voter FAQs in the US and How to Improve for Informed Electoral Participation
by: Rawte, Vipula, et al.
Published: (2024)
by: Rawte, Vipula, et al.
Published: (2024)
LLM-Generated Feedback Supports Learning If Learners Choose to Use It
by: Thomas, Danielle R., et al.
Published: (2025)
by: Thomas, Danielle R., et al.
Published: (2025)
Knowledge Graph Guided Evaluation of Abstention Techniques
by: Vasisht, Kinshuk, et al.
Published: (2024)
by: Vasisht, Kinshuk, et al.
Published: (2024)
"Amazing, They All Lean Left" -- Analyzing the Political Temperaments of Current LLMs
by: Neuman, W. Russell, et al.
Published: (2025)
by: Neuman, W. Russell, et al.
Published: (2025)
Hidden Prompts in Manuscripts Exploit AI-Assisted Peer Review
by: Lin, Zhicheng
Published: (2025)
by: Lin, Zhicheng
Published: (2025)
Use Me Wisely: AI-Driven Assessment for LLM Prompting Skills Development
by: Ognibene, Dimitri, et al.
Published: (2025)
by: Ognibene, Dimitri, et al.
Published: (2025)
Evaluating Large Language Models for Health-related Queries with Presuppositions
by: Kaur, Navreet, et al.
Published: (2023)
by: Kaur, Navreet, et al.
Published: (2023)
Who is a Better Matchmaker? Human vs. Algorithmic Judge Assignment in a High-Stakes Startup Competition
by: Xi, Sarina, et al.
Published: (2025)
by: Xi, Sarina, et al.
Published: (2025)
Beyond the Rubric: Cultural Misalignment in LLM Benchmarks for Sexual and Reproductive Health
by: Dey, Sumon Kanti, et al.
Published: (2025)
by: Dey, Sumon Kanti, et al.
Published: (2025)
STAMP Your Content: Proving Dataset Membership via Watermarked Rephrasings
by: Rastogi, Saksham, et al.
Published: (2025)
by: Rastogi, Saksham, et al.
Published: (2025)
Clinical Note Bloat Reduction for Efficient LLM Use
by: Cahoon, Jordan L., et al.
Published: (2026)
by: Cahoon, Jordan L., et al.
Published: (2026)
MAGPIE: A benchmark for Multi-AGent contextual PrIvacy Evaluation
by: Juneja, Gurusha, et al.
Published: (2025)
by: Juneja, Gurusha, et al.
Published: (2025)
AgentPeerTalk: Empowering Students through Agentic-AI-Driven Discernment of Bullying and Joking in Peer Interactions in Schools
by: Paul, Aditya, et al.
Published: (2024)
by: Paul, Aditya, et al.
Published: (2024)
Task Facet Learning: A Structured Approach to Prompt Optimization
by: Juneja, Gurusha, et al.
Published: (2024)
by: Juneja, Gurusha, et al.
Published: (2024)
Downstream Trade-offs of a Family of Text Watermarks
by: Ajith, Anirudh, et al.
Published: (2023)
by: Ajith, Anirudh, et al.
Published: (2023)
LLM-Generated or Human-Written? Comparing Review and Non-Review Papers on ArXiv
by: Elazar, Yanai, et al.
Published: (2026)
by: Elazar, Yanai, et al.
Published: (2026)
LLM-as-a-Reviewer: Benchmarking Their Ability, Divergence, and Prompt Injection Resistance as Paper Reviewers
by: Li, Lingyao, et al.
Published: (2026)
by: Li, Lingyao, et al.
Published: (2026)
Computational Analysis of Climate Policy
by: Hicks, Carolyn
Published: (2025)
by: Hicks, Carolyn
Published: (2025)
Virtual Agent-Based Communication Skills Training to Facilitate Health Persuasion Among Peers
by: Nouraei, Farnaz, et al.
Published: (2024)
by: Nouraei, Farnaz, et al.
Published: (2024)
Similar Items
-
Richer Output for Richer Countries: Uncovering Geographical Disparities in Generated Stories and Travel Recommendations
by: Bhagat, Kirti, et al.
Published: (2024) -
TALES: A Taxonomy and Analysis of Cultural Representations in LLM-generated Stories
by: Bhagat, Kirti, et al.
Published: (2025) -
Reviewing the Reviewer: Elevating Peer Review Quality through LLM-Guided Feedback
by: Purkayastha, Sukannya, et al.
Published: (2026) -
A Principled Approach to Randomized Selection under Uncertainty: Applications to Peer Review and Grant Funding
by: Goldberg, Alexander, et al.
Published: (2025) -
$\texttt{LM}^\texttt{2}$: A Simple Society of Language Models Solves Complex Reasoning
by: Juneja, Gurusha, et al.
Published: (2024)