LazyReview A Dataset for Uncovering Lazy Thinking in NLP Peer Reviews
Fuente:
arXiv
Saved in:
| Main Authors: | Purkayastha, Sukannya, Li, Zhuang, Lauscher, Anne, Qu, Lizhen, Gurevych, Iryna |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reviewing the Reviewer: Elevating Peer Review Quality through LLM-Guided Feedback
by: Purkayastha, Sukannya, et al.
Published: (2026)
by: Purkayastha, Sukannya, et al.
Published: (2026)
Decision-Making with Deliberation: Meta-reviewing as a Document-grounded Dialogue
by: Purkayastha, Sukannya, et al.
Published: (2025)
by: Purkayastha, Sukannya, et al.
Published: (2025)
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
by: Baumgärtner, Tim, et al.
Published: (2025)
by: Baumgärtner, Tim, et al.
Published: (2025)
The Lou Dataset -- Exploring the Impact of Gender-Fair Language in German Text Classification
by: Waldis, Andreas, et al.
Published: (2024)
by: Waldis, Andreas, et al.
Published: (2024)
Identifying Aspects in Peer Reviews
by: Lu, Sheng, et al.
Published: (2025)
by: Lu, Sheng, et al.
Published: (2025)
Commitment Checklist: Auditing Author Commitments in Peer Review
by: Chen, Chung-Chi, et al.
Published: (2026)
by: Chen, Chung-Chi, et al.
Published: (2026)
Is Peer Review Really in Decline? Analyzing Review Quality across Venues and Time
by: Kuznetsov, Ilia, et al.
Published: (2026)
by: Kuznetsov, Ilia, et al.
Published: (2026)
Sensitivity, Performance, Robustness: Deconstructing the Effect of Sociodemographic Prompting
by: Beck, Tilman, et al.
Published: (2023)
by: Beck, Tilman, et al.
Published: (2023)
The Good, the Bad and the Constructive: Automatically Measuring Peer Review's Utility for Authors
by: Sadallah, Abdelrahman, et al.
Published: (2025)
by: Sadallah, Abdelrahman, et al.
Published: (2025)
Aligned Probing: Relating Toxic Behavior and Model Internals
by: Waldis, Andreas, et al.
Published: (2025)
by: Waldis, Andreas, et al.
Published: (2025)
Automatic Reviewers Fail to Detect Faulty Reasoning in Research Papers: A New Counterfactual Evaluation Framework
by: Dycke, Nils, et al.
Published: (2025)
by: Dycke, Nils, et al.
Published: (2025)
The Nature of NLP: Analyzing Contributions in NLP Papers
by: Pramanick, Aniket, et al.
Published: (2024)
by: Pramanick, Aniket, et al.
Published: (2024)
STRICTA: Structured Reasoning in Critical Text Assessment for Peer Review and Beyond
by: Dycke, Nils, et al.
Published: (2024)
by: Dycke, Nils, et al.
Published: (2024)
What Can Natural Language Processing Do for Peer Review?
by: Kuznetsov, Ilia, et al.
Published: (2024)
by: Kuznetsov, Ilia, et al.
Published: (2024)
A Lazy, Concurrent Convertibility Checker
by: Courant, Nathanaëlle, et al.
Published: (2025)
by: Courant, Nathanaëlle, et al.
Published: (2025)
Culturally Aware and Adapted NLP: A Taxonomy and a Survey of the State of the Art
by: Liu, Chen Cecilia, et al.
Published: (2024)
by: Liu, Chen Cecilia, et al.
Published: (2024)
Common Sense vs. Morality: The Curious Case of Narrative Focus Bias in LLMs
by: Purkayastha, Saugata, et al.
Published: (2026)
by: Purkayastha, Saugata, et al.
Published: (2026)
ClaimFlow: Tracing the Evolution of Scientific Claims in NLP
by: Pramanick, Aniket, et al.
Published: (2026)
by: Pramanick, Aniket, et al.
Published: (2026)
A Comprehensive Review of Datasets for Clinical Mental Health AI Systems
by: Mandal, Aishik, et al.
Published: (2025)
by: Mandal, Aishik, et al.
Published: (2025)
Re3: A Holistic Framework and Dataset for Modeling Collaborative Document Revision
by: Ruan, Qian, et al.
Published: (2024)
by: Ruan, Qian, et al.
Published: (2024)
Exposía: Teaching and Assessment of Academic Writing Skills for Research Project Proposals and Peer Feedback
by: Zyska, Dennis, et al.
Published: (2026)
by: Zyska, Dennis, et al.
Published: (2026)
On Group Relative Policy Optimization Collapse in Agent Search: The Lazy Likelihood-Displacement
by: Deng, Wenlong, et al.
Published: (2025)
by: Deng, Wenlong, et al.
Published: (2025)
Story of Your Lazy Function's Life: A Bidirectional Demand Semantics for Mechanized Cost Analysis of Lazy Programs
by: Xia, Li-yao, et al.
Published: (2024)
by: Xia, Li-yao, et al.
Published: (2024)
M4FC: a Multimodal, Multilingual, Multicultural, Multitask Real-World Fact-Checking Dataset
by: Geng, Jiahui, et al.
Published: (2025)
by: Geng, Jiahui, et al.
Published: (2025)
Stop! In the Name of Flaws: Disentangling Personal Names and Sociodemographic Attributes in NLP
by: Gautam, Vagrant, et al.
Published: (2024)
by: Gautam, Vagrant, et al.
Published: (2024)
Building Bridges: A Dataset for Evaluating Gender-Fair Machine Translation into German
by: Lardelli, Manuel, et al.
Published: (2024)
by: Lardelli, Manuel, et al.
Published: (2024)
Difficult Task Yes but Simple Task No: Unveiling the Laziness in Multimodal LLMs
by: Zhao, Sihang, et al.
Published: (2024)
by: Zhao, Sihang, et al.
Published: (2024)
IDALC: A Semi-Supervised Framework for Intent Detection and Active Learning based Correction
by: Mullick, Ankan, et al.
Published: (2025)
by: Mullick, Ankan, et al.
Published: (2025)
Citation Failure: Definition, Analysis and Efficient Mitigation
by: Buchmann, Jan, et al.
Published: (2025)
by: Buchmann, Jan, et al.
Published: (2025)
Like a Good Nearest Neighbor: Practical Content Moderation and Text Classification
by: Bates, Luke, et al.
Published: (2023)
by: Bates, Luke, et al.
Published: (2023)
Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to Deliberation
by: Zhang, Zhiwei, et al.
Published: (2025)
by: Zhang, Zhiwei, et al.
Published: (2025)
ConspirED: A Dataset for Cognitive Traits of Conspiracy Theories and Large Language Model Safety
by: Bates, Luke, et al.
Published: (2025)
by: Bates, Luke, et al.
Published: (2025)
Shorten After You're Right: Lazy Length Penalties for Reasoning RL
by: Yuan, Danlong, et al.
Published: (2025)
by: Yuan, Danlong, et al.
Published: (2025)
Analyzing Dataset Annotation Quality Management in the Wild
by: Klie, Jan-Christoph, et al.
Published: (2023)
by: Klie, Jan-Christoph, et al.
Published: (2023)
Diversity Over Size: On the Effect of Sample and Topic Sizes for Topic-Dependent Argument Mining Datasets
by: Schiller, Benjamin, et al.
Published: (2022)
by: Schiller, Benjamin, et al.
Published: (2022)
L-RAG: Balancing Context and Retrieval with Entropy-Based Lazy Loading
by: Voloshyn, Sergii
Published: (2026)
by: Voloshyn, Sergii
Published: (2026)
SciCoQA: Quality Assurance for Scientific Paper--Code Alignment
by: Baumgärtner, Tim, et al.
Published: (2026)
by: Baumgärtner, Tim, et al.
Published: (2026)
Lazy Linearity for a Core Functional Language
by: Mesquita, Rodrigo, et al.
Published: (2025)
by: Mesquita, Rodrigo, et al.
Published: (2025)
Lexicographic Ranking Supermartingales with Lazy Lower Bounds
by: Takisaka, Toru, et al.
Published: (2023)
by: Takisaka, Toru, et al.
Published: (2023)
Are Non-English Papers Reviewed Fairly? Language-of-Study Bias in NLP Peer Reviews
by: Barkhordar, Ehsan, et al.
Published: (2026)
by: Barkhordar, Ehsan, et al.
Published: (2026)
Similar Items
-
Reviewing the Reviewer: Elevating Peer Review Quality through LLM-Guided Feedback
by: Purkayastha, Sukannya, et al.
Published: (2026) -
Decision-Making with Deliberation: Meta-reviewing as a Document-grounded Dialogue
by: Purkayastha, Sukannya, et al.
Published: (2025) -
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
by: Baumgärtner, Tim, et al.
Published: (2025) -
The Lou Dataset -- Exploring the Impact of Gender-Fair Language in German Text Classification
by: Waldis, Andreas, et al.
Published: (2024) -
Identifying Aspects in Peer Reviews
by: Lu, Sheng, et al.
Published: (2025)