BAMBOO: A Comprehensive Benchmark for Evaluating Long Text Modeling Capacities of Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dong, Zican, Tang, Tianyi, Li, Junyi, Zhao, Wayne Xin, Wen, Ji-Rong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Survey on Long Text Modeling with Transformers
von: Dong, Zican, et al.
Veröffentlicht: (2023)
von: Dong, Zican, et al.
Veröffentlicht: (2023)
LongReD: Mitigating Short-Text Degradation of Long-Context Large Language Models via Restoration Distillation
von: Dong, Zican, et al.
Veröffentlicht: (2025)
von: Dong, Zican, et al.
Veröffentlicht: (2025)
Exploring Context Window of Large Language Models via Decomposed Positional Vectors
von: Dong, Zican, et al.
Veröffentlicht: (2024)
von: Dong, Zican, et al.
Veröffentlicht: (2024)
ChainLM: Empowering Large Language Models with Improved Chain-of-Thought Prompting
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2024)
How Efficient Are Diffusion Language Models? A Critical Examination of Efficiency Evaluation Practices
von: Peng, Han, et al.
Veröffentlicht: (2025)
von: Peng, Han, et al.
Veröffentlicht: (2025)
LLMBox: A Comprehensive Library for Large Language Models
von: Tang, Tianyi, et al.
Veröffentlicht: (2024)
von: Tang, Tianyi, et al.
Veröffentlicht: (2024)
Neuron-based Personality Trait Induction in Large Language Models
von: Deng, Jia, et al.
Veröffentlicht: (2024)
von: Deng, Jia, et al.
Veröffentlicht: (2024)
Rethinking the Evaluation for Conversational Recommendation in the Era of Large Language Models
von: Wang, Xiaolei, et al.
Veröffentlicht: (2023)
von: Wang, Xiaolei, et al.
Veröffentlicht: (2023)
Towards Coarse-to-Fine Evaluation of Inference Efficiency for Large Language Models
von: Chen, Yushuo, et al.
Veröffentlicht: (2024)
von: Chen, Yushuo, et al.
Veröffentlicht: (2024)
ForesightKV: Optimizing KV Cache Eviction for Reasoning Models by Learning Long-Term Contribution
von: Dong, Zican, et al.
Veröffentlicht: (2026)
von: Dong, Zican, et al.
Veröffentlicht: (2026)
Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models
von: Sun, Haoxiang, et al.
Veröffentlicht: (2025)
von: Sun, Haoxiang, et al.
Veröffentlicht: (2025)
The Dawn After the Dark: An Empirical Study on Factuality Hallucination in Large Language Models
von: Li, Junyi, et al.
Veröffentlicht: (2024)
von: Li, Junyi, et al.
Veröffentlicht: (2024)
Unleashing the Potential of Large Language Models as Prompt Optimizers: Analogical Analysis with Gradient-based Model Optimizers
von: Tang, Xinyu, et al.
Veröffentlicht: (2024)
von: Tang, Xinyu, et al.
Veröffentlicht: (2024)
A Survey of Large Language Models
von: Zhao, Wayne Xin, et al.
Veröffentlicht: (2023)
von: Zhao, Wayne Xin, et al.
Veröffentlicht: (2023)
Think More, Hallucinate Less: Mitigating Hallucinations via Dual Process of Fast and Slow Thinking
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2025)
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2025)
YuLan-Mini: An Open Data-efficient Language Model
von: Hu, Yiwen, et al.
Veröffentlicht: (2024)
von: Hu, Yiwen, et al.
Veröffentlicht: (2024)
Language-Specific Neurons: The Key to Multilingual Capabilities in Large Language Models
von: Tang, Tianyi, et al.
Veröffentlicht: (2024)
von: Tang, Tianyi, et al.
Veröffentlicht: (2024)
Not Everything is All You Need: Toward Low-Redundant Optimization for Large Language Model Alignment
von: Chen, Zhipeng, et al.
Veröffentlicht: (2024)
von: Chen, Zhipeng, et al.
Veröffentlicht: (2024)
Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models
von: Li, Yifan, et al.
Veröffentlicht: (2024)
von: Li, Yifan, et al.
Veröffentlicht: (2024)
Beyond Imitation: Leveraging Fine-grained Quality Signals for Alignment
von: Guo, Geyang, et al.
Veröffentlicht: (2023)
von: Guo, Geyang, et al.
Veröffentlicht: (2023)
Domain-Specific Pruning of Large Mixture-of-Experts Models with Few-shot Demonstrations
von: Dong, Zican, et al.
Veröffentlicht: (2025)
von: Dong, Zican, et al.
Veröffentlicht: (2025)
Enhancing Cross-task Transfer of Large Language Models via Activation Steering
von: Tang, Xinyu, et al.
Veröffentlicht: (2025)
von: Tang, Xinyu, et al.
Veröffentlicht: (2025)
Small Agent Can Also Rock! Empowering Small Language Models as Hallucination Detector
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2024)
Unveiling the Flaws: Exploring Imperfections in Synthetic Data and Mitigation Strategies for Large Language Models
von: Chen, Jie, et al.
Veröffentlicht: (2024)
von: Chen, Jie, et al.
Veröffentlicht: (2024)
Incentivizing Dual Process Thinking for Efficient Large Language Model Reasoning
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2025)
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2025)
Sticker-TTS: Learn to Utilize Historical Experience with a Sticker-driven Test-Time Scaling Framework
von: Chen, Jie, et al.
Veröffentlicht: (2025)
von: Chen, Jie, et al.
Veröffentlicht: (2025)
Extracting and Combining Abilities For Building Multi-lingual Ability-enhanced Large Language Models
von: Chen, Zhipeng, et al.
Veröffentlicht: (2024)
von: Chen, Zhipeng, et al.
Veröffentlicht: (2024)
CAFE: Retrieval Head-based Coarse-to-Fine Information Seeking to Enhance Multi-Document QA Capability
von: Peng, Han, et al.
Veröffentlicht: (2025)
von: Peng, Han, et al.
Veröffentlicht: (2025)
Do we Really Need Visual Instructions? Towards Visual Instruction-Free Fine-tuning for Large Vision-Language Models
von: Liu, Zikang, et al.
Veröffentlicht: (2025)
von: Liu, Zikang, et al.
Veröffentlicht: (2025)
Improving Large Language Models via Fine-grained Reinforcement Learning with Minimum Editing Constraint
von: Chen, Zhipeng, et al.
Veröffentlicht: (2024)
von: Chen, Zhipeng, et al.
Veröffentlicht: (2024)
Experience-Guided Reflective Co-Evolution of Prompts and Heuristics for Automatic Algorithm Design
von: Liu, Yihong, et al.
Veröffentlicht: (2025)
von: Liu, Yihong, et al.
Veröffentlicht: (2025)
Adaptive Ability Decomposing for Unlocking Large Reasoning Model Effective Reinforcement Learning
von: Chen, Zhipeng, et al.
Veröffentlicht: (2026)
von: Chen, Zhipeng, et al.
Veröffentlicht: (2026)
ReasoningLM: Enabling Structural Subgraph Reasoning in Pre-trained Language Models for Question Answering over Knowledge Graph
von: Jiang, Jinhao, et al.
Veröffentlicht: (2023)
von: Jiang, Jinhao, et al.
Veröffentlicht: (2023)
MMATH: A Multilingual Benchmark for Mathematical Reasoning
von: Luo, Wenyang, et al.
Veröffentlicht: (2025)
von: Luo, Wenyang, et al.
Veröffentlicht: (2025)
Mix-CPT: A Domain Adaptation Framework via Decoupling Knowledge Learning and Format Alignment
von: Jiang, Jinhao, et al.
Veröffentlicht: (2024)
von: Jiang, Jinhao, et al.
Veröffentlicht: (2024)
DAWN-ICL: Strategic Planning of Problem-solving Trajectories for Zero-Shot In-Context Learning
von: Tang, Xinyu, et al.
Veröffentlicht: (2024)
von: Tang, Xinyu, et al.
Veröffentlicht: (2024)
Self-Calibrated Listwise Reranking with Large Language Models
von: Ren, Ruiyang, et al.
Veröffentlicht: (2024)
von: Ren, Ruiyang, et al.
Veröffentlicht: (2024)
Estimating the Error of Large Language Models at Pairwise Text Comparison
von: Li, Tianyi
Veröffentlicht: (2025)
von: Li, Tianyi
Veröffentlicht: (2025)
A Comprehensive Evaluation of Large Language Models on Benchmark Biomedical Text Processing Tasks
von: Jahan, Israt, et al.
Veröffentlicht: (2023)
von: Jahan, Israt, et al.
Veröffentlicht: (2023)
BASES: Large-scale Web Search User Simulation with Large Language Model based Agents
von: Ren, Ruiyang, et al.
Veröffentlicht: (2024)
von: Ren, Ruiyang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Survey on Long Text Modeling with Transformers
von: Dong, Zican, et al.
Veröffentlicht: (2023) -
LongReD: Mitigating Short-Text Degradation of Long-Context Large Language Models via Restoration Distillation
von: Dong, Zican, et al.
Veröffentlicht: (2025) -
Exploring Context Window of Large Language Models via Decomposed Positional Vectors
von: Dong, Zican, et al.
Veröffentlicht: (2024) -
ChainLM: Empowering Large Language Models with Improved Chain-of-Thought Prompting
von: Cheng, Xiaoxue, et al.
Veröffentlicht: (2024) -
How Efficient Are Diffusion Language Models? A Critical Examination of Efficiency Evaluation Practices
von: Peng, Han, et al.
Veröffentlicht: (2025)