PRELUDE: A Benchmark Designed to Require Global Comprehension and Reasoning over Long Contexts
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Mo, Chung, Tsz Ting, Zhou, Chulun, Li, Tong, Lu, Rui, Li, Jiangnan, Xu, Liyan, Lu, Haoshu, Zhang, Ning, Li, Jing, Zhou, Jie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fine-Grained Modeling of Narrative Context: A Coherence Perspective via Retrospective Questions
by: Xu, Liyan, et al.
Published: (2024)
by: Xu, Liyan, et al.
Published: (2024)
MiA-Signature: Approximating Global Activation for Long-Context Understanding
by: Li, Yuqing, et al.
Published: (2026)
by: Li, Yuqing, et al.
Published: (2026)
SitEmb-v1.5: Improved Context-Aware Dense Retrieval for Semantic Association and Long Story Comprehension
by: Wu, Junjie, et al.
Published: (2025)
by: Wu, Junjie, et al.
Published: (2025)
The Essence of Contextual Understanding in Theory of Mind: A Study on Question Answering with Story Characters
by: Zhou, Chulun, et al.
Published: (2025)
by: Zhou, Chulun, et al.
Published: (2025)
HGMEM: Hypergraph-based Working Memory to Improve Multi-step RAG for Long-Context Complex Relational Modeling
by: Zhou, Chulun, et al.
Published: (2025)
by: Zhou, Chulun, et al.
Published: (2025)
Mindscape-Aware Retrieval Augmented Generation for Improved Long Context Understanding
by: Li, Yuqing, et al.
Published: (2025)
by: Li, Yuqing, et al.
Published: (2025)
Query-focused and Memory-aware Reranker for Long Context Processing
by: Li, Yuqing, et al.
Published: (2026)
by: Li, Yuqing, et al.
Published: (2026)
Dense Retrievers Can Fail on Simple Queries: Revealing The Granularity Dilemma of Embeddings
by: Xu, Liyan, et al.
Published: (2025)
by: Xu, Liyan, et al.
Published: (2025)
The Stochastic Parrot on LLM's Shoulder: A Summative Assessment of Physical Concept Understanding
by: Yu, Mo, et al.
Published: (2025)
by: Yu, Mo, et al.
Published: (2025)
DivLogicEval: A Framework for Benchmarking Logical Reasoning Evaluation in Large Language Models
by: Chung, Tsz Ting, et al.
Published: (2025)
by: Chung, Tsz Ting, et al.
Published: (2025)
Beyond characteristic equations: A unified one-dimensional non-Bloch band theory via wavefunction data
by: Li, Haoshu
Published: (2025)
by: Li, Haoshu
Published: (2025)
Previously on the Stories: Recap Snippet Identification for Story Reading
by: Li, Jiangnan, et al.
Published: (2024)
by: Li, Jiangnan, et al.
Published: (2024)
Test of partial effects for Frechet regression on Bures-Wasserstein manifolds
by: Xu, Haoshu, et al.
Published: (2025)
by: Xu, Haoshu, et al.
Published: (2025)
Wasserstein F-tests for Fréchet regression on Bures-Wasserstein manifolds
by: Xu, Haoshu, et al.
Published: (2024)
by: Xu, Haoshu, et al.
Published: (2024)
LoongRL: Reinforcement Learning for Advanced Reasoning over Long Contexts
by: Wang, Siyuan, et al.
Published: (2025)
by: Wang, Siyuan, et al.
Published: (2025)
LongReason: A Synthetic Long-Context Reasoning Benchmark via Context Expansion
by: Ling, Zhan, et al.
Published: (2025)
by: Ling, Zhan, et al.
Published: (2025)
A Policy Gradient Framework for Stochastic Optimal Control Problems with Global Convergence Guarantee
by: Zhou, Mo, et al.
Published: (2023)
by: Zhou, Mo, et al.
Published: (2023)
ComoRAG: A Cognitive-Inspired Memory-Organized RAG for Stateful Long Narrative Reasoning
by: Wang, Juyuan, et al.
Published: (2025)
by: Wang, Juyuan, et al.
Published: (2025)
Mementos: A Comprehensive Benchmark for Multimodal Large Language Model Reasoning over Image Sequences
by: Wang, Xiyao, et al.
Published: (2024)
by: Wang, Xiyao, et al.
Published: (2024)
Dexterous Manipulation Based on Prior Dexterous Grasp Pose Knowledge
by: Yan, Hengxu, et al.
Published: (2024)
by: Yan, Hengxu, et al.
Published: (2024)
Incentivizing In-depth Reasoning over Long Contexts with Process Advantage Shaping
by: Peng, Miao, et al.
Published: (2026)
by: Peng, Miao, et al.
Published: (2026)
AV-Reasoner: Improving and Benchmarking Clue-Grounded Audio-Visual Counting for MLLMs
by: Lu, Lidong, et al.
Published: (2025)
by: Lu, Lidong, et al.
Published: (2025)
SIG: Speaker Identification in Literature via Prompt-Based Generation
by: Su, Zhenlin, et al.
Published: (2023)
by: Su, Zhenlin, et al.
Published: (2023)
Who Gets Cited Most? Benchmarking Long-Context Numerical Reasoning on Scientific Articles
by: Li, Miao, et al.
Published: (2025)
by: Li, Miao, et al.
Published: (2025)
LongDocURL: a Comprehensive Multimodal Long Document Benchmark Integrating Understanding, Reasoning, and Locating
by: Deng, Chao, et al.
Published: (2024)
by: Deng, Chao, et al.
Published: (2024)
Controlling atom-photon bound states in a coupled resonator array with a two-level quantum emitter
by: Lu, Zelin, et al.
Published: (2024)
by: Lu, Zelin, et al.
Published: (2024)
Aberrant dynamic properties of whole‐brain functional connectivity in acute mild traumatic brain injury revealed by hidden Markov models
by: Liyan Lu, et al.
Published: (2024)
by: Liyan Lu, et al.
Published: (2024)
ContextLens: Modeling Imperfect Privacy and Safety Context for Legal Compliance
by: Li, Haoran, et al.
Published: (2026)
by: Li, Haoran, et al.
Published: (2026)
Efficient Size Constraint Community Search over Heterogeneous Information Networks
by: Zhang, Xinjian, et al.
Published: (2025)
by: Zhang, Xinjian, et al.
Published: (2025)
Latent Context Compilation: Distilling Long Context into Compact Portable Memory
by: Li, Zeju, et al.
Published: (2026)
by: Li, Zeju, et al.
Published: (2026)
PunchBench: Benchmarking MLLMs in Multimodal Punchline Comprehension
by: Ouyang, Kun, et al.
Published: (2024)
by: Ouyang, Kun, et al.
Published: (2024)
LONGCODEU: Benchmarking Long-Context Language Models on Long Code Understanding
by: Li, Jia, et al.
Published: (2025)
by: Li, Jia, et al.
Published: (2025)
MTR-Bench: A Comprehensive Benchmark for Multi-Turn Reasoning Evaluation
by: Li, Xiaoyuan, et al.
Published: (2025)
by: Li, Xiaoyuan, et al.
Published: (2025)
Dissecting Long-Chain-of-Thought Reasoning Models: An Empirical Study
by: Mu, Yongyu, et al.
Published: (2025)
by: Mu, Yongyu, et al.
Published: (2025)
Evidence-Augmented Policy Optimization with Reward Co-Evolution for Long-Context Reasoning
by: Guan, Xin, et al.
Published: (2026)
by: Guan, Xin, et al.
Published: (2026)
WhenLoss: Diagnosing Write and Retrieval Bottlenecks in Long-Context Memory Systems
by: Yu, Jiangnan, et al.
Published: (2026)
by: Yu, Jiangnan, et al.
Published: (2026)
Judge Like Human Examiners: A Weighted Importance Multi-Point Evaluation Framework for Generative Tasks with Long-form Answers
by: Yu, Guoxin, et al.
Published: (2026)
by: Yu, Guoxin, et al.
Published: (2026)
Recent Progress in Cathode Materials for Ca‐Ion Batteries
by: Chong‐Yu Du, et al.
Published: (2025)
by: Chong‐Yu Du, et al.
Published: (2025)
QwenLong-L1.5: Post-Training Recipe for Long-Context Reasoning and Memory Management
by: Shen, Weizhou, et al.
Published: (2025)
by: Shen, Weizhou, et al.
Published: (2025)
AudioMarathon: A Comprehensive Benchmark for Long-Context Audio Understanding and Efficiency in Audio LLMs
by: He, Peize, et al.
Published: (2025)
by: He, Peize, et al.
Published: (2025)
Similar Items
-
Fine-Grained Modeling of Narrative Context: A Coherence Perspective via Retrospective Questions
by: Xu, Liyan, et al.
Published: (2024) -
MiA-Signature: Approximating Global Activation for Long-Context Understanding
by: Li, Yuqing, et al.
Published: (2026) -
SitEmb-v1.5: Improved Context-Aware Dense Retrieval for Semantic Association and Long Story Comprehension
by: Wu, Junjie, et al.
Published: (2025) -
The Essence of Contextual Understanding in Theory of Mind: A Study on Question Answering with Story Characters
by: Zhou, Chulun, et al.
Published: (2025) -
HGMEM: Hypergraph-based Working Memory to Improve Multi-step RAG for Long-Context Complex Relational Modeling
by: Zhou, Chulun, et al.
Published: (2025)