Similar Items
Graph Neural Networks Are More Than Filters: Revisiting and Benchmarking from A Spectral Perspective
by: Dong, Yushun, et al.
Published: (2024)
by: Dong, Yushun, et al.
Published: (2024)
PreScience: A Benchmark for Forecasting Scientific Contributions
by: Ajith, Anirudh, et al.
Published: (2026)
by: Ajith, Anirudh, et al.
Published: (2026)
SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images
by: Shen, Jialu, et al.
Published: (2026)
by: Shen, Jialu, et al.
Published: (2026)
Are Biological Systems More Intelligent Than Artificial Intelligence?
by: Bennett, Michael Timothy
Published: (2024)
by: Bennett, Michael Timothy
Published: (2024)
Deep Ideation: Designing LLM Agents to Generate Novel Research Ideas on Scientific Concept Network
by: Zhao, Keyu, et al.
Published: (2025)
by: Zhao, Keyu, et al.
Published: (2025)
Measuring Political Bias in Large Language Models: What Is Said and How It Is Said
by: Bang, Yejin, et al.
Published: (2024)
by: Bang, Yejin, et al.
Published: (2024)
More Than Sum of Its Parts: Deciphering Intent Shifts in Multimodal Hate Speech Detection
by: Sun, Runze, et al.
Published: (2026)
by: Sun, Runze, et al.
Published: (2026)
Evolving Idea Graphs with Learnable Edits-and-Commits for Multi-Agent Scientific Ideation
by: Dong, Jiangwen, et al.
Published: (2026)
by: Dong, Jiangwen, et al.
Published: (2026)
IRIS: Interactive Research Ideation System for Accelerating Scientific Discovery
by: Garikaparthi, Aniketh, et al.
Published: (2025)
by: Garikaparthi, Aniketh, et al.
Published: (2025)
AI Needs Physics More Than Physics Needs AI
by: Coveney, Peter, et al.
Published: (2025)
by: Coveney, Peter, et al.
Published: (2025)
SceMQA: A Scientific College Entrance Level Multimodal Question Answering Benchmark
by: Liang, Zhenwen, et al.
Published: (2024)
by: Liang, Zhenwen, et al.
Published: (2024)
Saying More Than They Know: A Framework for Quantifying Epistemic-Rhetorical Miscalibration in Large Language Models
by: Bakhshi, Asim D.
Published: (2026)
by: Bakhshi, Asim D.
Published: (2026)
Debate as Reward: A Multi-Agent Reward System for Scientific Ideation via RL Post-Training
by: Salimi, Moein, et al.
Published: (2026)
by: Salimi, Moein, et al.
Published: (2026)
More Than A Shortcut: A Hyperbolic Approach To Early-Exit Networks
by: Bhosale, Swapnil, et al.
Published: (2025)
by: Bhosale, Swapnil, et al.
Published: (2025)
Can AI Validate Science? Benchmarking LLMs for Accurate Scientific Claim $\rightarrow$ Evidence Reasoning
by: Javaji, Shashidhar Reddy, et al.
Published: (2025)
by: Javaji, Shashidhar Reddy, et al.
Published: (2025)
Tokenization Is More Than Compression
by: Schmidt, Craig W., et al.
Published: (2024)
by: Schmidt, Craig W., et al.
Published: (2024)
Read, Grep, and Synthesize: Diagnosing Cross-Domain Seed Exposure for LLM Research Ideation
by: Choi, Yunju, et al.
Published: (2026)
by: Choi, Yunju, et al.
Published: (2026)
Aggregate-Combine-Readout GNNs Are More Expressive Than Logic C2
by: Hauke, Stan P, et al.
Published: (2025)
by: Hauke, Stan P, et al.
Published: (2025)
Navigating Ideation Space: Decomposed Conceptual Representations for Positioning Scientific Ideas
by: Shen, Yuexi, et al.
Published: (2026)
by: Shen, Yuexi, et al.
Published: (2026)
HARPA: A Testability-Driven, Literature-Grounded Framework for Research Ideation
by: Vasu, Rosni, et al.
Published: (2025)
by: Vasu, Rosni, et al.
Published: (2025)
When Models Know More Than They Can Explain: Quantifying Knowledge Transfer in Human-AI Collaboration
by: Shi, Quan, et al.
Published: (2025)
by: Shi, Quan, et al.
Published: (2025)
IRPAPERS: A Visual Document Benchmark for Scientific Retrieval and Question Answering
by: Shorten, Connor, et al.
Published: (2026)
by: Shorten, Connor, et al.
Published: (2026)
LLMs Struggle with Abstract Meaning Comprehension More Than Expected
by: Alhazmi, Hamoud, et al.
Published: (2026)
by: Alhazmi, Hamoud, et al.
Published: (2026)
A Concept is More Than a Word: Diversified Unlearning in Text-to-Image Diffusion Models
by: Pham, Duc Hao, et al.
Published: (2026)
by: Pham, Duc Hao, et al.
Published: (2026)
Can LLMs Ask Good Questions?
by: Zhang, Yueheng, et al.
Published: (2025)
by: Zhang, Yueheng, et al.
Published: (2025)
Aster: Autonomous Scientific Discovery over 20x Faster Than Existing Methods
by: Bicker, Emmett
Published: (2026)
by: Bicker, Emmett
Published: (2026)
Who Do LLMs Trust? Human Experts Matter More Than Other LLMs
by: Bajaj, Anooshka, et al.
Published: (2026)
by: Bajaj, Anooshka, et al.
Published: (2026)
MuISQA: Multi-Intent Retrieval-Augmented Generation for Scientific Question Answering
by: Li, Zhiyuan, et al.
Published: (2025)
by: Li, Zhiyuan, et al.
Published: (2025)
More Than Routing: Joint GPS and Route Modeling for Refine Trajectory Representation Learning
by: Ma, Zhipeng, et al.
Published: (2024)
by: Ma, Zhipeng, et al.
Published: (2024)
Anchorless Diversification for Parallel LLM Ideation
by: Ibrahim, Fares Nabil, et al.
Published: (2026)
by: Ibrahim, Fares Nabil, et al.
Published: (2026)
Solving for X and Beyond: Can Large Language Models Solve Complex Math Problems with More-Than-Two Unknowns?
by: Kao, Kuei-Chun, et al.
Published: (2024)
by: Kao, Kuei-Chun, et al.
Published: (2024)
Labels Matter More Than Models: Rethinking the Unsupervised Paradigm in Time Series Anomaly Detection
by: Zhong, Zhijie, et al.
Published: (2025)
by: Zhong, Zhijie, et al.
Published: (2025)
More Than Meets the Eye? Uncovering the Reasoning-Planning Disconnect in Training Vision-Language Driving Models
by: Song, Xurui, et al.
Published: (2025)
by: Song, Xurui, et al.
Published: (2025)
Right this way: Can VLMs Guide Us to See More to Answer Questions?
by: Liu, Li, et al.
Published: (2024)
by: Liu, Li, et al.
Published: (2024)
MEQA: A Meta-Evaluation Framework for Question & Answer LLM Benchmarks
by: Veuthey, Jaime Raldua, et al.
Published: (2025)
by: Veuthey, Jaime Raldua, et al.
Published: (2025)
Why DDIM Hallucinates More Than DDPM: A Theoretical Analysis of Reverse Dynamics
by: Ashiq, Muhammad H., et al.
Published: (2026)
by: Ashiq, Muhammad H., et al.
Published: (2026)
LLMs Know More Than Words: A Genre Study with Syntax, Metaphor & Phonetics
by: Shi, Weiye, et al.
Published: (2025)
by: Shi, Weiye, et al.
Published: (2025)
TSAQA: Time Series Analysis Question And Answering Benchmark
by: Jing, Baoyu, et al.
Published: (2026)
by: Jing, Baoyu, et al.
Published: (2026)
More Than One Teacher: Adaptive Multi-Guidance Policy Optimization for Diverse Exploration
by: Yuan, Xiaoyang, et al.
Published: (2025)
by: Yuan, Xiaoyang, et al.
Published: (2025)
Is Pre-training Truly Better Than Meta-Learning?
by: Miranda, Brando, et al.
Published: (2023)
by: Miranda, Brando, et al.
Published: (2023)
Similar Items
-
Graph Neural Networks Are More Than Filters: Revisiting and Benchmarking from A Spectral Perspective
by: Dong, Yushun, et al.
Published: (2024) -
PreScience: A Benchmark for Forecasting Scientific Contributions
by: Ajith, Anirudh, et al.
Published: (2026) -
SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images
by: Shen, Jialu, et al.
Published: (2026) -
Are Biological Systems More Intelligent Than Artificial Intelligence?
by: Bennett, Michael Timothy
Published: (2024) -
Deep Ideation: Designing LLM Agents to Generate Novel Research Ideas on Scientific Concept Network
by: Zhao, Keyu, et al.
Published: (2025)