ScholarEval: Research Idea Evaluation Grounded in Literature
Fuente:
arXiv
Saved in:
| Main Authors: | Moussa, Hanane Nour, Da Silva, Patrick Queiroz, Adu-Ampratwum, Daniel, East, Alyson, Lu, Zitong, Puccetti, Nikki, Xue, Mingyi, Sun, Huan, Majumder, Bodhisattwa Prasad, Kumar, Sachin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Accepted with Minor Revisions: Value of AI-Assisted Scientific Writing
by: Hazra, Sanchaita, et al.
Published: (2025)
by: Hazra, Sanchaita, et al.
Published: (2025)
To Tell The Truth: Language of Deception and Language Models
by: Hazra, Sanchaita, et al.
Published: (2023)
by: Hazra, Sanchaita, et al.
Published: (2023)
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
by: Kalluri, Tarun, et al.
Published: (2024)
by: Kalluri, Tarun, et al.
Published: (2024)
AI Safety Should Prioritize the Future of Work
by: Hazra, Sanchaita, et al.
Published: (2025)
by: Hazra, Sanchaita, et al.
Published: (2025)
The Good, the Bad, and the Ugly: The Role of AI Quality Disclosure in Lie Detection
by: Bhattacharya, Haimanti, et al.
Published: (2024)
by: Bhattacharya, Haimanti, et al.
Published: (2024)
ArtifactLinker: Linking Scientific Artifacts for Automatic State-of-the-Art Discovery
by: Yu, Haofei, et al.
Published: (2026)
by: Yu, Haofei, et al.
Published: (2026)
Put Your Money Where Your Mouth Is: Evaluating Strategic Planning and Execution of LLM Agents in an Auction Arena
by: Chen, Jiangjie, et al.
Published: (2023)
by: Chen, Jiangjie, et al.
Published: (2023)
RLSynC: Offline-Online Reinforcement Learning for Synthon Completion
by: Baker, Frazier N., et al.
Published: (2023)
by: Baker, Frazier N., et al.
Published: (2023)
LARC: Towards Human-level Constrained Retrosynthesis Planning through an Agentic Framework
by: Baker, Frazier N., et al.
Published: (2025)
by: Baker, Frazier N., et al.
Published: (2025)
BeetleVerse: A Study on Taxonomic Classification of Ground Beetles
by: Rayeed, S M, et al.
Published: (2025)
by: Rayeed, S M, et al.
Published: (2025)
Comprehensive framework for assessing and optimizing existing research networks
by: Alyson East, et al.
Published: (2026)
by: Alyson East, et al.
Published: (2026)
log-RRIM: Yield Prediction via Local-to-global Reaction Representation Learning and Interaction Modeling
by: Hu, Xiao, et al.
Published: (2024)
by: Hu, Xiao, et al.
Published: (2024)
Generating 3D Binding Molecules Using Shape-Conditioned Diffusion Models with Guidance
by: Chen, Ziqi, et al.
Published: (2025)
by: Chen, Ziqi, et al.
Published: (2025)
DiffER: Categorical Diffusion for Chemical Retrosynthesis
by: Current, Sean, et al.
Published: (2025)
by: Current, Sean, et al.
Published: (2025)
ARMOR: An Agentic Framework for Reaction Feasibility Prediction via Adaptive Utility-aware Multi-tool Reasoning
by: Liu, Ye, et al.
Published: (2026)
by: Liu, Ye, et al.
Published: (2026)
Data-driven Discovery with Large Generative Models
by: Majumder, Bodhisattwa Prasad, et al.
Published: (2024)
by: Majumder, Bodhisattwa Prasad, et al.
Published: (2024)
MMORF: A Multi-agent Framework for Designing Multi-objective Retrosynthesis Planning Systems
by: Baker, Frazier N., et al.
Published: (2026)
by: Baker, Frazier N., et al.
Published: (2026)
Tailoring with Targeted Precision: Edit-Based Agents for Open-Domain Procedure Customization
by: Lal, Yash Kumar, et al.
Published: (2023)
by: Lal, Yash Kumar, et al.
Published: (2023)
Skill Set Optimization: Reinforcing Language Model Behavior via Transferable Skills
by: Nottingham, Kolby, et al.
Published: (2024)
by: Nottingham, Kolby, et al.
Published: (2024)
Literature-Grounded Novelty Assessment of Scientific Ideas
by: Shahid, Simra, et al.
Published: (2025)
by: Shahid, Simra, et al.
Published: (2025)
ChemToolAgent: The Impact of Tools on Language Agents for Chemistry Problem Solving
by: Yu, Botao, et al.
Published: (2024)
by: Yu, Botao, et al.
Published: (2024)
IdeaSynth: Iterative Research Idea Development Through Evolving and Composing Idea Facets with Literature-Grounded Feedback
by: Pu, Kevin, et al.
Published: (2024)
by: Pu, Kevin, et al.
Published: (2024)
Few-shot Dialogue Strategy Learning for Motivational Interviewing via Inductive Reasoning
by: Xie, Zhouhang, et al.
Published: (2024)
by: Xie, Zhouhang, et al.
Published: (2024)
Age of Patients in Trials Submitted to the FDA Versus Age in Average Patients With Cancer
by: Alyson Haslam, et al.
Published: (2025)
by: Alyson Haslam, et al.
Published: (2025)
InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem
by: Qiao, Shuofei, et al.
Published: (2026)
by: Qiao, Shuofei, et al.
Published: (2026)
D3-Gym: Constructing Real-World Verifiable Environments for Data-Driven Discovery
by: Moussa, Hanane Nour, et al.
Published: (2026)
by: Moussa, Hanane Nour, et al.
Published: (2026)
GroundCocoa: A Benchmark for Evaluating Compositional & Conditional Reasoning in Language Models
by: Kohli, Harsh, et al.
Published: (2024)
by: Kohli, Harsh, et al.
Published: (2024)
Latent Factor Models Meets Instructions: Goal-conditioned Latent Factor Discovery without Task Supervision
by: Xie, Zhouhang, et al.
Published: (2025)
by: Xie, Zhouhang, et al.
Published: (2025)
AutoSDT: Scaling Data-Driven Discovery Tasks Toward Open Co-Scientists
by: Li, Yifei, et al.
Published: (2025)
by: Li, Yifei, et al.
Published: (2025)
Automatic Image-Level Morphological Trait Annotation for Organismal Images
by: Pahuja, Vardaan, et al.
Published: (2026)
by: Pahuja, Vardaan, et al.
Published: (2026)
Steering off Course: Reliability Challenges in Steering Language Models
by: Da Silva, Patrick Queiroz, et al.
Published: (2025)
by: Da Silva, Patrick Queiroz, et al.
Published: (2025)
A Survey on Personalized and Pluralistic Preference Alignment in Large Language Models
by: Xie, Zhouhang, et al.
Published: (2025)
by: Xie, Zhouhang, et al.
Published: (2025)
MB-ORES: A Multi-Branch Object Reasoner for Visual Grounding in Remote Sensing
by: Radouane, Karim, et al.
Published: (2025)
by: Radouane, Karim, et al.
Published: (2025)
DISCOVERYWORLD: A Virtual Environment for Developing and Evaluating Automated Scientific Discovery Agents
by: Jansen, Peter, et al.
Published: (2024)
by: Jansen, Peter, et al.
Published: (2024)
Self-Learning Longitudinal Control for On-Road Vehicles
by: Puccetti, Luca
Published: (2023)
by: Puccetti, Luca
Published: (2023)
Articulando: arte, ensino e produção para uma educação especial
by: Roberta Puccetti
Published: (2005)
by: Roberta Puccetti
Published: (2005)
ScienceAgentBench: Toward Rigorous Assessment of Language Agents for Data-Driven Scientific Discovery
by: Chen, Ziru, et al.
Published: (2024)
by: Chen, Ziru, et al.
Published: (2024)
Characterization of bridging therapies in clinical trials leading to FDA approval of CAR‐T cell therapies
by: Victoria Kaestner, et al.
Published: (2025)
by: Victoria Kaestner, et al.
Published: (2025)
How many people in the US are eligible for and respond to checkpoint inhibitors: An empirical analysis
by: Alyson Haslam, et al.
Published: (2025)
by: Alyson Haslam, et al.
Published: (2025)
Using and Reading Scholarly Literature.
by: King, Donald W., et al.
Published: (1999)
by: King, Donald W., et al.
Published: (1999)
Similar Items
-
Accepted with Minor Revisions: Value of AI-Assisted Scientific Writing
by: Hazra, Sanchaita, et al.
Published: (2025) -
To Tell The Truth: Language of Deception and Language Models
by: Hazra, Sanchaita, et al.
Published: (2023) -
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
by: Kalluri, Tarun, et al.
Published: (2024) -
AI Safety Should Prioritize the Future of Work
by: Hazra, Sanchaita, et al.
Published: (2025) -
The Good, the Bad, and the Ugly: The Role of AI Quality Disclosure in Lie Detection
by: Bhattacharya, Haimanti, et al.
Published: (2024)