Hierarchical Context Merging: Better Long Context Understanding for Pre-trained LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Song, Woomin, Oh, Seunghyuk, Mo, Sangwoo, Kim, Jaehyung, Yun, Sukmin, Ha, Jung-Woo, Shin, Jinwoo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Sparsified State-Space Models are Efficient Highway Networks
di: Song, Woomin, et al.
Pubblicazione: (2025)
di: Song, Woomin, et al.
Pubblicazione: (2025)
Tabular Transfer Learning via Prompting LLMs
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)
Optimized Feature Generation for Tabular Data via LLMs with Decision Tree Reasoning
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)
di: Nam, Jaehyun, et al.
Pubblicazione: (2024)
SuRe: Summarizing Retrievals using Answer Candidates for Open-domain QA of LLMs
di: Kim, Jaehyung, et al.
Pubblicazione: (2024)
di: Kim, Jaehyung, et al.
Pubblicazione: (2024)
ReVISE: Learning to Refine at Test-Time via Intrinsic Self-Verification
di: Lee, Hyunseok, et al.
Pubblicazione: (2025)
di: Lee, Hyunseok, et al.
Pubblicazione: (2025)
LVRPO: Language-Visual Alignment with GRPO for Multimodal Understanding and Generation
di: Mo, Shentong, et al.
Pubblicazione: (2026)
di: Mo, Shentong, et al.
Pubblicazione: (2026)
RoboAlign: Learning Test-Time Reasoning for Language-Action Alignment in Vision-Language-Action Models
di: Kim, Dongyoung, et al.
Pubblicazione: (2026)
di: Kim, Dongyoung, et al.
Pubblicazione: (2026)
Personalized Language Models via Privacy-Preserving Evolutionary Model Merging
di: Kim, Kyuyoung, et al.
Pubblicazione: (2025)
di: Kim, Kyuyoung, et al.
Pubblicazione: (2025)
EMCEE: Improving Multilingual Capability of LLMs via Bridging Knowledge and Reasoning with Extracted Synthetic Multilingual Context
di: Koo, Hamin, et al.
Pubblicazione: (2025)
di: Koo, Hamin, et al.
Pubblicazione: (2025)
Mamba Drafters for Speculative Decoding
di: Choi, Daewon, et al.
Pubblicazione: (2025)
di: Choi, Daewon, et al.
Pubblicazione: (2025)
Training Text-to-Molecule Models with Context-Aware Tokenization
di: Kim, Seojin, et al.
Pubblicazione: (2025)
di: Kim, Seojin, et al.
Pubblicazione: (2025)
Improving Visual Representation Alignment Generation with GRPO
di: Mo, Shentong, et al.
Pubblicazione: (2026)
di: Mo, Shentong, et al.
Pubblicazione: (2026)
GMAIL: Generative Modality Alignment for generated Image Learning
di: Mo, Shentong, et al.
Pubblicazione: (2026)
di: Mo, Shentong, et al.
Pubblicazione: (2026)
DMT-JEPA: Discriminative Masked Targets for Joint-Embedding Predictive Architecture
di: Mo, Shentong, et al.
Pubblicazione: (2024)
di: Mo, Shentong, et al.
Pubblicazione: (2024)
Debiasing Online Preference Learning via Preference Feature Preservation
di: Kim, Dongyoung, et al.
Pubblicazione: (2025)
di: Kim, Dongyoung, et al.
Pubblicazione: (2025)
Spread Preference Annotation: Direct Preference Judgment for Efficient LLM Alignment
di: Kim, Dongyoung, et al.
Pubblicazione: (2024)
di: Kim, Dongyoung, et al.
Pubblicazione: (2024)
APCE: Adaptive Progressive Context Expansion for Long Context Processing
di: Lee, Baisub, et al.
Pubblicazione: (2025)
di: Lee, Baisub, et al.
Pubblicazione: (2025)
Long-Tailed Recognition on Binary Networks by Calibrating A Pre-trained Model
di: Kim, Jihun, et al.
Pubblicazione: (2024)
di: Kim, Jihun, et al.
Pubblicazione: (2024)
ExComm: Exploration-Stage Communication for Error-Resilient Agentic Test-Time Scaling
di: Song, Woomin, et al.
Pubblicazione: (2026)
di: Song, Woomin, et al.
Pubblicazione: (2026)
Compress, Gather, and Recompute: REFORMing Long-Context Processing in Transformers
di: Song, Woomin, et al.
Pubblicazione: (2025)
di: Song, Woomin, et al.
Pubblicazione: (2025)
Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding
di: Woo, Byeongju, et al.
Pubblicazione: (2026)
di: Woo, Byeongju, et al.
Pubblicazione: (2026)
Double-P: Hierarchical Top-P Sparse Attention for Long-Context LLMs
di: Ni, Wentao, et al.
Pubblicazione: (2026)
di: Ni, Wentao, et al.
Pubblicazione: (2026)
Robot-R1: Reinforcement Learning for Enhanced Embodied Reasoning in Robotics
di: Kim, Dongyoung, et al.
Pubblicazione: (2025)
di: Kim, Dongyoung, et al.
Pubblicazione: (2025)
IdleSpec: Exploiting Idle Time via Speculative Planning for LLM Agents
di: Choi, Daewon, et al.
Pubblicazione: (2026)
di: Choi, Daewon, et al.
Pubblicazione: (2026)
ContextVLA: Vision-Language-Action Model with Amortized Multi-Frame Context
di: Jang, Huiwon, et al.
Pubblicazione: (2025)
di: Jang, Huiwon, et al.
Pubblicazione: (2025)
Feature Unlearning for Pre-trained GANs and VAEs
di: Moon, Saemi, et al.
Pubblicazione: (2023)
di: Moon, Saemi, et al.
Pubblicazione: (2023)
Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL
di: Choi, Jinwoo, et al.
Pubblicazione: (2026)
di: Choi, Jinwoo, et al.
Pubblicazione: (2026)
CURIE: Evaluating LLMs On Multitask Scientific Long Context Understanding and Reasoning
di: Cui, Hao, et al.
Pubblicazione: (2025)
di: Cui, Hao, et al.
Pubblicazione: (2025)
TiTok: Transfer Token-level Knowledge via Contrastive Excess to Transplant LoRA
di: Jung, Chanjoo, et al.
Pubblicazione: (2025)
di: Jung, Chanjoo, et al.
Pubblicazione: (2025)
Comparing Pre-trained Human Language Models: Is it Better with Human Context as Groups, Individual Traits, or Both?
di: Soni, Nikita, et al.
Pubblicazione: (2024)
di: Soni, Nikita, et al.
Pubblicazione: (2024)
MultiDocFusion: Hierarchical and Multimodal Chunking Pipeline for Enhanced RAG on Long Industrial Documents
di: Shin, Joongmin, et al.
Pubblicazione: (2026)
di: Shin, Joongmin, et al.
Pubblicazione: (2026)
One-shot Imitation in a Non-Stationary Environment via Multi-Modal Skill
di: Shin, Sangwoo, et al.
Pubblicazione: (2024)
di: Shin, Sangwoo, et al.
Pubblicazione: (2024)
Efficient LLM Collaboration via Planning
di: Lee, Byeongchan, et al.
Pubblicazione: (2025)
di: Lee, Byeongchan, et al.
Pubblicazione: (2025)
MERIT Feedback Elicits Better Bargaining in LLM Negotiators
di: Oh, Jihwan, et al.
Pubblicazione: (2026)
di: Oh, Jihwan, et al.
Pubblicazione: (2026)
Rethinking Prompt Design for Inference-time Scaling in Text-to-Visual Generation
di: Kim, Subin, et al.
Pubblicazione: (2025)
di: Kim, Subin, et al.
Pubblicazione: (2025)
LooGLE: Can Long-Context Language Models Understand Long Contexts?
di: Li, Jiaqi, et al.
Pubblicazione: (2023)
di: Li, Jiaqi, et al.
Pubblicazione: (2023)
Enhancing Instruction Following of LLMs via Activation Steering with Dynamic Rejection
di: Kang, Minjae, et al.
Pubblicazione: (2026)
di: Kang, Minjae, et al.
Pubblicazione: (2026)
EXIT: Context-Aware Extractive Compression for Enhancing Retrieval-Augmented Generation
di: Hwang, Taeho, et al.
Pubblicazione: (2024)
di: Hwang, Taeho, et al.
Pubblicazione: (2024)
Online Adaptation of Language Models with a Memory of Amortized Contexts
di: Tack, Jihoon, et al.
Pubblicazione: (2024)
di: Tack, Jihoon, et al.
Pubblicazione: (2024)
Few-shot Personalization of LLMs with Mis-aligned Responses
di: Kim, Jaehyung, et al.
Pubblicazione: (2024)
di: Kim, Jaehyung, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Sparsified State-Space Models are Efficient Highway Networks
di: Song, Woomin, et al.
Pubblicazione: (2025) -
Tabular Transfer Learning via Prompting LLMs
di: Nam, Jaehyun, et al.
Pubblicazione: (2024) -
Optimized Feature Generation for Tabular Data via LLMs with Decision Tree Reasoning
di: Nam, Jaehyun, et al.
Pubblicazione: (2024) -
SuRe: Summarizing Retrievals using Answer Candidates for Open-domain QA of LLMs
di: Kim, Jaehyung, et al.
Pubblicazione: (2024) -
ReVISE: Learning to Refine at Test-Time via Intrinsic Self-Verification
di: Lee, Hyunseok, et al.
Pubblicazione: (2025)