AInstein: Can LLMs Solve Research Problems From Parametric Memory Alone?
Fuente:
arXiv
Saved in:
| Main Authors: | Mishra, Shambhavi, Sahu, Gaurav, Pedersoli, Marco, Charlin, Laurent, Dolz, Jose, Pal, Christopher |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking Literature Search Evaluation: Deep Research Helps, and Human Citation Lists Are Not a Ground Truth
by: Sahu, Gaurav, et al.
Published: (2026)
by: Sahu, Gaurav, et al.
Published: (2026)
ReviewerToo: Should AI Join The Program Committee? A Look At The Future of Peer Review
by: Sahu, Gaurav, et al.
Published: (2025)
by: Sahu, Gaurav, et al.
Published: (2025)
LitLLMs, LLMs for Literature Review: Are we there yet?
by: Agarwal, Shubham, et al.
Published: (2024)
by: Agarwal, Shubham, et al.
Published: (2024)
LitLLM: A Toolkit for Scientific Literature Review
by: Agarwal, Shubham, et al.
Published: (2024)
by: Agarwal, Shubham, et al.
Published: (2024)
Do not trust what you trust: Miscalibration in Semi-supervised Learning
by: Mishra, Shambhavi, et al.
Published: (2024)
by: Mishra, Shambhavi, et al.
Published: (2024)
AInsteinBench: Benchmarking Coding Agents on Scientific Repositories
by: Duston, Titouan, et al.
Published: (2025)
by: Duston, Titouan, et al.
Published: (2025)
Can Graph Descriptive Order Affect Solving Graph Problems with LLMs?
by: Ge, Yuyao, et al.
Published: (2024)
by: Ge, Yuyao, et al.
Published: (2024)
Semantic Anchor Transport: Robust Test-Time Adaptation for Vision-Language Models
by: Mishra, Shambhavi, et al.
Published: (2024)
by: Mishra, Shambhavi, et al.
Published: (2024)
From Problem-Solving to Teaching Problem-Solving: Aligning LLMs with Pedagogy using Reinforcement Learning
by: Dinucu-Jianu, David, et al.
Published: (2025)
by: Dinucu-Jianu, David, et al.
Published: (2025)
From Logic to Language: A Trust Index for Problem Solving with LLMs
by: Rug, Tehseen, et al.
Published: (2025)
by: Rug, Tehseen, et al.
Published: (2025)
Can LLMs Solve ASP Problems? Insights from a Benchmarking Study (Extended Version)
by: Ren, Lin, et al.
Published: (2025)
by: Ren, Lin, et al.
Published: (2025)
A Guide To Effectively Leveraging LLMs for Low-Resource Text Summarization: Data Augmentation and Semi-supervised Approaches
by: Sahu, Gaurav, et al.
Published: (2024)
by: Sahu, Gaurav, et al.
Published: (2024)
Can Large Language Models Solve Robot Routing?
by: Huang, Zhehui, et al.
Published: (2024)
by: Huang, Zhehui, et al.
Published: (2024)
Can LLMs Reconcile Knowledge Conflicts in Counterfactual Reasoning
by: Yamin, Khurram, et al.
Published: (2025)
by: Yamin, Khurram, et al.
Published: (2025)
OpInf-LLM: Parametric PDE Solving with LLMs via Operator Inference
by: Wang, Zhuoyuan, et al.
Published: (2026)
by: Wang, Zhuoyuan, et al.
Published: (2026)
An Agentic Framework with LLMs for Solving Complex Vehicle Routing Problems
by: Zhang, Ni, et al.
Published: (2025)
by: Zhang, Ni, et al.
Published: (2025)
TEARS: Textual Representations for Scrutable Recommendations
by: Penaloza, Emiliano, et al.
Published: (2024)
by: Penaloza, Emiliano, et al.
Published: (2024)
ProSEA: Problem Solving via Exploration Agents
by: Nguyen, William, et al.
Published: (2025)
by: Nguyen, William, et al.
Published: (2025)
Problems With Large Language Models for Learner Modelling: Why LLMs Alone Fall Short for Responsible Tutoring in K--12 Education
by: Hooshyar, Danial, et al.
Published: (2025)
by: Hooshyar, Danial, et al.
Published: (2025)
Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
Optimization Problem Solving Can Transition to Evolutionary Agentic Workflows
by: Li, Wenhao, et al.
Published: (2025)
by: Li, Wenhao, et al.
Published: (2025)
Can Language Models Solve Graph Problems in Natural Language?
by: Wang, Heng, et al.
Published: (2023)
by: Wang, Heng, et al.
Published: (2023)
Addressing Concept Mislabeling in Concept Bottleneck Models Through Preference Optimization
by: Penaloza, Emiliano, et al.
Published: (2025)
by: Penaloza, Emiliano, et al.
Published: (2025)
CLARity: Reasoning Consistency Alone Can Teach Reinforced Experts
by: Lin, Jiuheng, et al.
Published: (2025)
by: Lin, Jiuheng, et al.
Published: (2025)
Performance of LLMs on Stochastic Modeling Operations Research Problems: From Theory to Practice
by: Kumar, Akshit, et al.
Published: (2025)
by: Kumar, Akshit, et al.
Published: (2025)
Plan before Solving: Problem-Aware Strategy Routing for Mathematical Reasoning with LLMs
by: Qi, Shihao, et al.
Published: (2025)
by: Qi, Shihao, et al.
Published: (2025)
Improving Multimodal LLMs Ability In Geometry Problem Solving, Reasoning, And Multistep Scoring
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
EduBot -- Can LLMs Solve Personalized Learning and Programming Assignments?
by: Wang, Yibin, et al.
Published: (2025)
by: Wang, Yibin, et al.
Published: (2025)
SMDD-Bench: Can LLMs Solve Real-World Small Molecule Drug Design Tasks?
by: Han, Kevin, et al.
Published: (2026)
by: Han, Kevin, et al.
Published: (2026)
Principled Data Augmentation for Learning to Solve Quadratic Programming Problems
by: Qian, Chendi, et al.
Published: (2025)
by: Qian, Chendi, et al.
Published: (2025)
Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving
by: Zhou, Yuxuan, et al.
Published: (2025)
by: Zhou, Yuxuan, et al.
Published: (2025)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
by: Pramanik, Rishav, et al.
Published: (2024)
by: Pramanik, Rishav, et al.
Published: (2024)
Large Vision Models Can Solve Mental Rotation Problems
by: Mason, Sebastian Ray, et al.
Published: (2025)
by: Mason, Sebastian Ray, et al.
Published: (2025)
MLRC-Bench: Can Language Agents Solve Machine Learning Research Challenges?
by: Zhang, Yunxiang, et al.
Published: (2025)
by: Zhang, Yunxiang, et al.
Published: (2025)
Probing Embodied LLMs: When Higher Observation Fidelity Hurts Problem Solving
by: Zenkri, Oussama, et al.
Published: (2026)
by: Zenkri, Oussama, et al.
Published: (2026)
Climbing the Ladder of Reasoning: What LLMs Can-and Still Can't-Solve after SFT?
by: Sun, Yiyou, et al.
Published: (2025)
by: Sun, Yiyou, et al.
Published: (2025)
Teaching LLMs According to Their Aptitude: Adaptive Reasoning for Mathematical Problem Solving
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
Privileged Information Distillation for Language Models
by: Penaloza, Emiliano, et al.
Published: (2026)
by: Penaloza, Emiliano, et al.
Published: (2026)
SMART: Self-Generating and Self-Validating Multi-Dimensional Assessment for LLMs' Mathematical Problem Solving
by: Hou, Yujie, et al.
Published: (2025)
by: Hou, Yujie, et al.
Published: (2025)
From RAG to Memory: Non-Parametric Continual Learning for Large Language Models
by: Gutiérrez, Bernal Jiménez, et al.
Published: (2025)
by: Gutiérrez, Bernal Jiménez, et al.
Published: (2025)
Similar Items
-
Rethinking Literature Search Evaluation: Deep Research Helps, and Human Citation Lists Are Not a Ground Truth
by: Sahu, Gaurav, et al.
Published: (2026) -
ReviewerToo: Should AI Join The Program Committee? A Look At The Future of Peer Review
by: Sahu, Gaurav, et al.
Published: (2025) -
LitLLMs, LLMs for Literature Review: Are we there yet?
by: Agarwal, Shubham, et al.
Published: (2024) -
LitLLM: A Toolkit for Scientific Literature Review
by: Agarwal, Shubham, et al.
Published: (2024) -
Do not trust what you trust: Miscalibration in Semi-supervised Learning
by: Mishra, Shambhavi, et al.
Published: (2024)