Allocate Marginal Reviews to Borderline Papers Using LLM Comparative Ranking
Fuente:
arXiv
Salvato in:
| Autori principali: | Epstein, Elliot L., Dwaraknath, Rajat, Winnicki, John, Sornwanee, Thanawat |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LLMs are Overconfident: Evaluating Confidence Interval Calibration with FermiEval
di: Epstein, Elliot L., et al.
Pubblicazione: (2025)
di: Epstein, Elliot L., et al.
Pubblicazione: (2025)
Agreement Between Large Language Models, Human Reviewers, and Authors in Evaluating STROBE Checklists for Observational Studies in Rheumatology
di: Bilgin, Emre, et al.
Pubblicazione: (2026)
di: Bilgin, Emre, et al.
Pubblicazione: (2026)
DNB-AI-Project at SemEval-2025 Task 5: An LLM-Ensemble Approach for Automated Subject Indexing
di: Kluge, Lisa, et al.
Pubblicazione: (2025)
di: Kluge, Lisa, et al.
Pubblicazione: (2025)
HLM-Cite: Hybrid Language Model Workflow for Text-based Scientific Citation Prediction
di: Hao, Qianyue, et al.
Pubblicazione: (2024)
di: Hao, Qianyue, et al.
Pubblicazione: (2024)
Historical Ink: Exploring Large Language Models for Irony Detection in 19th-Century Spanish
di: Cohen, Kevin, et al.
Pubblicazione: (2025)
di: Cohen, Kevin, et al.
Pubblicazione: (2025)
Citation Grounding: Detecting and Reducing LLM Citation Hallucinations via Legal Citation Graphs
di: Ovcharov, Volodymyr
Pubblicazione: (2026)
di: Ovcharov, Volodymyr
Pubblicazione: (2026)
Historical Ink: 19th Century Latin American Spanish Newspaper Corpus with LLM OCR Correction
di: Manrique-Gómez, Laura, et al.
Pubblicazione: (2024)
di: Manrique-Gómez, Laura, et al.
Pubblicazione: (2024)
LLM-ReSum: A Framework for LLM Reflective Summarization through Self-Evaluation
di: Nguyen, Huyen, et al.
Pubblicazione: (2026)
di: Nguyen, Huyen, et al.
Pubblicazione: (2026)
2024 Google Scholar Research Interest Ranking for Top 3260 Computer Science Authors
di: Rasane, Atharva
Pubblicazione: (2024)
di: Rasane, Atharva
Pubblicazione: (2024)
LCA and energy efficiency in buildings: mapping more than twenty years of research
di: Asdrubali, F., et al.
Pubblicazione: (2024)
di: Asdrubali, F., et al.
Pubblicazione: (2024)
The Journal of Prompt-Engineered (Moral) Philosophy Or: Why AI-Assisted Ethics Research Requires Process Transparency
di: Loi, Michele
Pubblicazione: (2025)
di: Loi, Michele
Pubblicazione: (2025)
Annif at SemEval-2025 Task 5: Traditional XMTC augmented by LLMs
di: Suominen, Osma, et al.
Pubblicazione: (2025)
di: Suominen, Osma, et al.
Pubblicazione: (2025)
A Survey on Spoken Italian Datasets and Corpora
di: Giordano, Marco, et al.
Pubblicazione: (2025)
di: Giordano, Marco, et al.
Pubblicazione: (2025)
AI-Powered Citation Auditing: A Zero-Assumption Protocol for Systematic Reference Verification in Academic Research
di: van Rensburg, L. J. Janse
Pubblicazione: (2025)
di: van Rensburg, L. J. Janse
Pubblicazione: (2025)
Reconnecting Fragmented Citation Networks with Semantic Augmentation
di: Huong, Vu Thi, et al.
Pubblicazione: (2026)
di: Huong, Vu Thi, et al.
Pubblicazione: (2026)
CitePrism: Human-in-the-Loop AI for Citation Auditing and Editorial Integrity
di: Mahesh, Gowrika, et al.
Pubblicazione: (2026)
di: Mahesh, Gowrika, et al.
Pubblicazione: (2026)
Research status of the Mendeleev Periodic Table: a bibliometric analysis
di: Sharma, Kamna, et al.
Pubblicazione: (2024)
di: Sharma, Kamna, et al.
Pubblicazione: (2024)
Autonomous Editorial Systems and Computational Investigation with Artificial Intelligence
di: Banafea, Ahmed
Pubblicazione: (2026)
di: Banafea, Ahmed
Pubblicazione: (2026)
ARISE: Agentic Rubric-Guided Iterative Survey Engine for Automated Scholarly Paper Generation
di: Wang, Zi, et al.
Pubblicazione: (2025)
di: Wang, Zi, et al.
Pubblicazione: (2025)
SD-KDE: Score-Debiased Kernel Density Estimation
di: Epstein, Elliot L., et al.
Pubblicazione: (2025)
di: Epstein, Elliot L., et al.
Pubblicazione: (2025)
EVINCE: Optimizing Multi-LLM Dialogues Using Conditional Statistics and Information Theory
di: Chang, Edward Y.
Pubblicazione: (2024)
di: Chang, Edward Y.
Pubblicazione: (2024)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
di: Peters, Sydney, et al.
Pubblicazione: (2025)
di: Peters, Sydney, et al.
Pubblicazione: (2025)
Flash-SD-KDE: Accelerating SD-KDE with Tensor Cores
di: Epstein, Elliot L., et al.
Pubblicazione: (2026)
di: Epstein, Elliot L., et al.
Pubblicazione: (2026)
SagaLLM: Context Management, Validation, and Transaction Guarantees for Multi-Agent LLM Planning
di: Chang, Edward Y., et al.
Pubblicazione: (2025)
di: Chang, Edward Y., et al.
Pubblicazione: (2025)
A Library of LLM Intrinsics for Retrieval-Augmented Generation
di: Danilevsky, Marina, et al.
Pubblicazione: (2025)
di: Danilevsky, Marina, et al.
Pubblicazione: (2025)
Evaluating LLM Metrics Through Real-World Capabilities
di: Miller, Justin K, et al.
Pubblicazione: (2025)
di: Miller, Justin K, et al.
Pubblicazione: (2025)
Diagnosing and Mitigating Sycophancy and Skepticism in LLM Causal Judgment
di: Chang, Edward Y.
Pubblicazione: (2026)
di: Chang, Edward Y.
Pubblicazione: (2026)
Applying Cognitive Design Patterns to General LLM Agents
di: Wray, Robert E., et al.
Pubblicazione: (2025)
di: Wray, Robert E., et al.
Pubblicazione: (2025)
ALAS: A Stateful Multi-LLM Agent Framework for Disruption-Aware Planning
di: Chang, Edward Y., et al.
Pubblicazione: (2025)
di: Chang, Edward Y., et al.
Pubblicazione: (2025)
LLM-based Automated Theorem Proving Hinges on Scalable Synthetic Data Generation
di: Lai, Junyu, et al.
Pubblicazione: (2025)
di: Lai, Junyu, et al.
Pubblicazione: (2025)
Understanding LLM Evaluator Behavior: A Structured Multi-Evaluator Framework for Merchant Risk Assessment
di: Wang, Liang, et al.
Pubblicazione: (2026)
di: Wang, Liang, et al.
Pubblicazione: (2026)
Permanent Data Encoding (PDE): A Visual Language for Semantic Compression and Knowledge Preservation in 3-Character Units
di: Tsuyuki, Yoshiharu, et al.
Pubblicazione: (2025)
di: Tsuyuki, Yoshiharu, et al.
Pubblicazione: (2025)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
di: Saji, Alan, et al.
Pubblicazione: (2025)
di: Saji, Alan, et al.
Pubblicazione: (2025)
AI Predicts AGI: Leveraging AGI Forecasting and Peer Review to Explore LLMs' Complex Reasoning Capabilities
di: Davide, Fabrizio, et al.
Pubblicazione: (2024)
di: Davide, Fabrizio, et al.
Pubblicazione: (2024)
Learning Software Bug Reports: A Systematic Literature Review
di: Long, Guoming, et al.
Pubblicazione: (2025)
di: Long, Guoming, et al.
Pubblicazione: (2025)
Challenges and Opportunities of NLP for HR Applications: A Discussion Paper
di: Leidner, Jochen L., et al.
Pubblicazione: (2024)
di: Leidner, Jochen L., et al.
Pubblicazione: (2024)
Meta-Learning at Scale for Large Language Models via Low-Rank Amortized Bayesian Meta-Learning
di: Zhang, Liyi, et al.
Pubblicazione: (2025)
di: Zhang, Liyi, et al.
Pubblicazione: (2025)
Enhancing Automatic PT Tagging for MEDLINE Citations Using Transformer-Based Models
di: Cid, Victor H., et al.
Pubblicazione: (2025)
di: Cid, Victor H., et al.
Pubblicazione: (2025)
What Makes a Good AI Review? Concern-Level Diagnostics for AI Peer Review
di: Jin, Ming
Pubblicazione: (2026)
di: Jin, Ming
Pubblicazione: (2026)
Is Our Chatbot Telling Lies? Assessing Correctness of an LLM-based Dutch Support Chatbot
di: Lassche, Herman, et al.
Pubblicazione: (2024)
di: Lassche, Herman, et al.
Pubblicazione: (2024)
Documenti analoghi
-
LLMs are Overconfident: Evaluating Confidence Interval Calibration with FermiEval
di: Epstein, Elliot L., et al.
Pubblicazione: (2025) -
Agreement Between Large Language Models, Human Reviewers, and Authors in Evaluating STROBE Checklists for Observational Studies in Rheumatology
di: Bilgin, Emre, et al.
Pubblicazione: (2026) -
DNB-AI-Project at SemEval-2025 Task 5: An LLM-Ensemble Approach for Automated Subject Indexing
di: Kluge, Lisa, et al.
Pubblicazione: (2025) -
HLM-Cite: Hybrid Language Model Workflow for Text-based Scientific Citation Prediction
di: Hao, Qianyue, et al.
Pubblicazione: (2024) -
Historical Ink: Exploring Large Language Models for Irony Detection in 19th-Century Spanish
di: Cohen, Kevin, et al.
Pubblicazione: (2025)