ACADREASON: Exploring the Limits of Reasoning Models with Academic Research Problems
Fuente:
arXiv
Saved in:
| Main Authors: | Gui, Xin, Zhu, King, Ren, JinCheng, Chen, Qianben, Wang, Zekun Moore, LI, Yizhi, Liu, Xinpeng, Li, Xiaowan, Ren, Wenli, Miao, Linyu, Qin, Tianrui, Shu, Ziqi, Zhu, He, Tang, Xiangru, Shi, Dingfeng, Liu, Jiaheng, Jiang, Yuchen Eleanor, Liu, Minghao, Zhang, Ge, Zhou, Wangchunshu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A$^2$FM: An Adaptive Agent Foundation Model for Tool-Aware Hybrid Reasoning
by: Chen, Qianben, et al.
Published: (2025)
by: Chen, Qianben, et al.
Published: (2025)
Flash-Searcher: Fast and Effective Web Agents via DAG-Based Parallel Execution
by: Qin, Tianrui, et al.
Published: (2025)
by: Qin, Tianrui, et al.
Published: (2025)
O-Researcher: An Open Ended Deep Research Model via Multi-Agent Distillation and Agentic RL
by: Yao, Yi, et al.
Published: (2026)
by: Yao, Yi, et al.
Published: (2026)
TaskCraft: Automated Generation of Agentic Tasks
by: Shi, Dingfeng, et al.
Published: (2025)
by: Shi, Dingfeng, et al.
Published: (2025)
Search More, Think Less: Rethinking Long-Horizon Agentic Search for Efficiency and Generalization
by: Chen, Qianben, et al.
Published: (2026)
by: Chen, Qianben, et al.
Published: (2026)
Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL
by: Li, Weizhen, et al.
Published: (2025)
by: Li, Weizhen, et al.
Published: (2025)
How Far Are We from Genuinely Useful Deep Research Agents?
by: Zhang, Dingling, et al.
Published: (2025)
by: Zhang, Dingling, et al.
Published: (2025)
OAgents: An Empirical Study of Building Effective Agents
by: Zhu, He, et al.
Published: (2025)
by: Zhu, He, et al.
Published: (2025)
EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies
by: Hu, Xavier, et al.
Published: (2026)
by: Hu, Xavier, et al.
Published: (2026)
Scaling Test-time Compute for LLM Agents
by: Zhu, King, et al.
Published: (2025)
by: Zhu, King, et al.
Published: (2025)
PopAlign: Diversifying Contrasting Patterns for a More Comprehensive Alignment
by: Wang, Zekun Moore, et al.
Published: (2024)
by: Wang, Zekun Moore, et al.
Published: (2024)
Ten‐word Test:An Effective Tool to Differential Mild Cognitive Impairment from Subjective Cognitive Decline
by: Hua Ren, et al.
Published: (2024)
by: Hua Ren, et al.
Published: (2024)
Depression Detection Using Digital Traces on Social Media: A Knowledge-aware Deep Learning Approach
by: Zhang, Wenli, et al.
Published: (2023)
by: Zhang, Wenli, et al.
Published: (2023)
MiCoTA: Bridging the Learnability Gap with Intermediate CoT and Teacher Assistants
by: Ding, Dongyi, et al.
Published: (2025)
by: Ding, Dongyi, et al.
Published: (2025)
PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization
by: Tao, Meiling, et al.
Published: (2025)
by: Tao, Meiling, et al.
Published: (2025)
Towards Faithful and Controllable Personalization via Critique-Post-Edit Reinforcement Learning
by: Zhu, Chenghao, et al.
Published: (2025)
by: Zhu, Chenghao, et al.
Published: (2025)
O-Mem: Omni Memory System for Personalized, Long Horizon, Self-Evolving Agents
by: Wang, Piaohong, et al.
Published: (2025)
by: Wang, Piaohong, et al.
Published: (2025)
SEAL: Can Saturated Benchmarks Be Revived by LLM-as-a-Meta-Judge?
by: Chen, Jiamin, et al.
Published: (2026)
by: Chen, Jiamin, et al.
Published: (2026)
MIO: A Foundation Model on Multimodal Tokens
by: Wang, Zekun, et al.
Published: (2024)
by: Wang, Zekun, et al.
Published: (2024)
Human and AI Perceptual Differences in Image Classification Errors
by: Liu, Minghao, et al.
Published: (2023)
by: Liu, Minghao, et al.
Published: (2023)
Efficient Agents: Building Effective Agents While Reducing Cost
by: Wang, Ningning, et al.
Published: (2025)
by: Wang, Ningning, et al.
Published: (2025)
Towards Personalized Deep Research: Benchmarks and Evaluations
by: Liang, Yuan, et al.
Published: (2025)
by: Liang, Yuan, et al.
Published: (2025)
The finite basis problem for additively idempotent semirings of order four, III
by: Ren, Miaomiao, et al.
Published: (2025)
by: Ren, Miaomiao, et al.
Published: (2025)
Few-Shot Learning for Mental Disorder Detection: A Continuous Multi-Prompt Engineering Approach with Medical Knowledge Injection
by: Liu, Haoxin, et al.
Published: (2024)
by: Liu, Haoxin, et al.
Published: (2024)
Can AI automatically analyze public opinion? A LLM agents-based agentic pipeline for timely public opinion analysis
by: Liu, Jing, et al.
Published: (2025)
by: Liu, Jing, et al.
Published: (2025)
Avoiding Premature Collapse: Adaptive Annealing for Entropy-Regularized Structural Inference
by: Liu, Yizhi
Published: (2026)
by: Liu, Yizhi
Published: (2026)
The Homogeneity Trap: Spectral Collapse in Doubly-Stochastic Deep Networks
by: Liu, Yizhi
Published: (2026)
by: Liu, Yizhi
Published: (2026)
Antibiotic resistance and pathogenicity-related genes
by: Liu, Tianrui
Published: (2025)
by: Liu, Tianrui
Published: (2025)
Evidence for the suppression of the hybrid skin-topological effect by fragile topology
by: Liu, Tianrui
Published: (2026)
by: Liu, Tianrui
Published: (2026)
Enhanced active disturbance rejection speed controller for permanent magnet synchronous motors using virtual friction feedback technique
by: Dingfeng Dong, et al.
Published: (2024)
by: Dingfeng Dong, et al.
Published: (2024)
Agent KB: Leveraging Cross-Domain Experience for Agentic Problem Solving
by: Tang, Xiangru, et al.
Published: (2025)
by: Tang, Xiangru, et al.
Published: (2025)
Multiple soliton solutions and similarity reduction of a (2+1)-dimensional variable-coefficient Korteweg-de Vries system
by: Liu, Yaqing, et al.
Published: (2022)
by: Liu, Yaqing, et al.
Published: (2022)
Unveiling Memorization-Generalization Coexistence: A Case Study on Arithmetic Tasks with Label Noise
by: Liu, Linyu, et al.
Published: (2026)
by: Liu, Linyu, et al.
Published: (2026)
Engineering Fractional Chern Insulators through Periodic Strain in Monolayer Graphene and Transition Metal Dichalcogenides
by: Liu, Yuchen, et al.
Published: (2024)
by: Liu, Yuchen, et al.
Published: (2024)
An Extended ADMM for 3-Block Nonconvex Nonseparable Problems with Applications
by: Liu, Zekun
Published: (2024)
by: Liu, Zekun
Published: (2024)
PhysicsSolver: Transformer-Enhanced Physics-Informed Neural Networks for Forward and Forecasting Problems in Partial Differential Equations
by: Zhu, Zhenyi, et al.
Published: (2025)
by: Zhu, Zhenyi, et al.
Published: (2025)
DLMMPR:Deep Learning-based Measurement Matrix for Phase Retrieval
by: Liu, Jing, et al.
Published: (2025)
by: Liu, Jing, et al.
Published: (2025)
Distributed Invariant Kalman Filter for Cooperative Localization using Matrix Lie Groups
by: Zhou, Yizhi, et al.
Published: (2024)
by: Zhou, Yizhi, et al.
Published: (2024)
Benchmarking Table Comprehension In The Wild
by: Pan, Yikang, et al.
Published: (2024)
by: Pan, Yikang, et al.
Published: (2024)
AI PERSONA: Towards Life-long Personalization of LLMs
by: Wang, Tiannan, et al.
Published: (2024)
by: Wang, Tiannan, et al.
Published: (2024)
Similar Items
-
A$^2$FM: An Adaptive Agent Foundation Model for Tool-Aware Hybrid Reasoning
by: Chen, Qianben, et al.
Published: (2025) -
Flash-Searcher: Fast and Effective Web Agents via DAG-Based Parallel Execution
by: Qin, Tianrui, et al.
Published: (2025) -
O-Researcher: An Open Ended Deep Research Model via Multi-Agent Distillation and Agentic RL
by: Yao, Yi, et al.
Published: (2026) -
TaskCraft: Automated Generation of Agentic Tasks
by: Shi, Dingfeng, et al.
Published: (2025) -
Search More, Think Less: Rethinking Long-Horizon Agentic Search for Efficiency and Generalization
by: Chen, Qianben, et al.
Published: (2026)