Learning to Decide with Just Enough: Information-Theoretic Context Summarization for CMDPs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Peidong, Lin, Junjiang, Wang, Shaowen, Xu, Yao, Li, Haiqing, Xie, Xuhao, Wu, Siyi, Li, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Just Enough Thinking: Efficient Reasoning with Adaptive Length Penalties Reinforcement Learning
von: Xiang, Violet, et al.
Veröffentlicht: (2025)
von: Xiang, Violet, et al.
Veröffentlicht: (2025)
Information-Theoretic Distillation for Reference-less Summarization
von: Jung, Jaehun, et al.
Veröffentlicht: (2024)
von: Jung, Jaehun, et al.
Veröffentlicht: (2024)
Escaping the Context Bottleneck: Active Context Curation for LLM Agents via Reinforcement Learning
von: Li, Xiaozhe, et al.
Veröffentlicht: (2026)
von: Li, Xiaozhe, et al.
Veröffentlicht: (2026)
Context-CoT: Enhancing Context Learning via High-Quality Reasoning Synthesis
von: Jin, Hongbo, et al.
Veröffentlicht: (2026)
von: Jin, Hongbo, et al.
Veröffentlicht: (2026)
Language Models for Text Classification: Is In-Context Learning Enough?
von: Edwards, Aleksandra, et al.
Veröffentlicht: (2024)
von: Edwards, Aleksandra, et al.
Veröffentlicht: (2024)
One Token Is Enough: Improving Diffusion Language Models with a Sink Token
von: Zhang, Zihou, et al.
Veröffentlicht: (2026)
von: Zhang, Zihou, et al.
Veröffentlicht: (2026)
To Theoretically Understand Transformer-Based In-Context Learning for Optimizing CSMA
von: Hao, Shugang, et al.
Veröffentlicht: (2025)
von: Hao, Shugang, et al.
Veröffentlicht: (2025)
Innate Reasoning is Not Enough: In-Context Learning Enhances Reasoning Large Language Models with Less Overthinking
von: Ge, Yuyao, et al.
Veröffentlicht: (2025)
von: Ge, Yuyao, et al.
Veröffentlicht: (2025)
DeepContext: A Context-aware, Cross-platform, and Cross-framework Tool for Performance Profiling and Analysis of Deep Learning Workloads
von: Zhao, Qidong, et al.
Veröffentlicht: (2024)
von: Zhao, Qidong, et al.
Veröffentlicht: (2024)
Fitting Is Not Enough: Smoothness in Extremely Quantized LLMs
von: Xu, Yuzhuang, et al.
Veröffentlicht: (2026)
von: Xu, Yuzhuang, et al.
Veröffentlicht: (2026)
Think Just Enough: Sequence-Level Entropy as a Confidence Signal for LLM Reasoning
von: Sharma, Aman, et al.
Veröffentlicht: (2025)
von: Sharma, Aman, et al.
Veröffentlicht: (2025)
Context-Aware Pseudo-Label Scoring for Zero-Shot Video Summarization
von: Wu, Yuanli, et al.
Veröffentlicht: (2025)
von: Wu, Yuanli, et al.
Veröffentlicht: (2025)
Enhancing In-Context Learning Performance with just SVD-Based Weight Pruning: A Theoretical Perspective
von: Yao, Xinhao, et al.
Veröffentlicht: (2024)
von: Yao, Xinhao, et al.
Veröffentlicht: (2024)
Improving Factual Consistency of News Summarization by Contrastive Preference Optimization
von: Feng, Huawen, et al.
Veröffentlicht: (2023)
von: Feng, Huawen, et al.
Veröffentlicht: (2023)
Is Distance Matrix Enough for Geometric Deep Learning?
von: Li, Zian, et al.
Veröffentlicht: (2023)
von: Li, Zian, et al.
Veröffentlicht: (2023)
KITE: Kernelized and Information Theoretic Exemplars for In-Context Learning
von: Singh, Vaibhav, et al.
Veröffentlicht: (2025)
von: Singh, Vaibhav, et al.
Veröffentlicht: (2025)
Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning
von: Dabas, Mahavir, et al.
Veröffentlicht: (2025)
von: Dabas, Mahavir, et al.
Veröffentlicht: (2025)
On the Size Complexity and Decidability of First-Order Progression
von: Classen, Jens, et al.
Veröffentlicht: (2026)
von: Classen, Jens, et al.
Veröffentlicht: (2026)
Tracking vs. Deciding: The Dual-Capability Bottleneck in Searchless Chess Transformers
von: Li, Quanhao, et al.
Veröffentlicht: (2026)
von: Li, Quanhao, et al.
Veröffentlicht: (2026)
Shortcut Learning in In-Context Learning: A Survey
von: Song, Rui, et al.
Veröffentlicht: (2024)
von: Song, Rui, et al.
Veröffentlicht: (2024)
OptiLeak: Efficient Prompt Reconstruction via Reinforcement Learning in Multi-tenant LLM Services
von: Wang, Longxiang, et al.
Veröffentlicht: (2026)
von: Wang, Longxiang, et al.
Veröffentlicht: (2026)
RLPF: Reinforcement Learning from Prediction Feedback for User Summarization with LLMs
von: Wu, Jiaxing, et al.
Veröffentlicht: (2024)
von: Wu, Jiaxing, et al.
Veröffentlicht: (2024)
Learning Agent-Compatible Context Management for Long-Horizon Tasks
von: Yi, Lu, et al.
Veröffentlicht: (2026)
von: Yi, Lu, et al.
Veröffentlicht: (2026)
Quotient DAGs for Off-Policy Evaluation:Forward-Flow Importance Sampling and Exact Slate Propensities
von: Xie, Ziwen, et al.
Veröffentlicht: (2026)
von: Xie, Ziwen, et al.
Veröffentlicht: (2026)
Beyond Prompting: Efficient and Robust Contextual Biasing for Speech LLMs via Logit-Space Integration (LOGIC)
von: Wang, Peidong
Veröffentlicht: (2026)
von: Wang, Peidong
Veröffentlicht: (2026)
Reward Is Enough: LLMs Are In-Context Reinforcement Learners
von: Song, Kefan, et al.
Veröffentlicht: (2025)
von: Song, Kefan, et al.
Veröffentlicht: (2025)
Disentangling Instructive Information from Ranked Multiple Candidates for Multi-Document Scientific Summarization
von: Wang, Pancheng, et al.
Veröffentlicht: (2024)
von: Wang, Pancheng, et al.
Veröffentlicht: (2024)
Empowering In-Browser Deep Learning Inference on Edge Devices with Just-in-Time Kernel Optimizations
von: Jia, Fucheng, et al.
Veröffentlicht: (2023)
von: Jia, Fucheng, et al.
Veröffentlicht: (2023)
Memorizing is Not Enough: Deep Knowledge Injection Through Reasoning
von: Xu, Ruoxi, et al.
Veröffentlicht: (2025)
von: Xu, Ruoxi, et al.
Veröffentlicht: (2025)
Scaling LLM Multi-turn RL with End-to-end Summarization-based Context Management
von: Lu, Miao, et al.
Veröffentlicht: (2025)
von: Lu, Miao, et al.
Veröffentlicht: (2025)
GeoBS: Information-Theoretic Quantification of Geographic Bias in AI Models
von: Wang, Zhangyu, et al.
Veröffentlicht: (2025)
von: Wang, Zhangyu, et al.
Veröffentlicht: (2025)
JustEva: A Toolkit to Evaluate LLM Fairness in Legal Knowledge Inference
von: Xue, Zongyue, et al.
Veröffentlicht: (2025)
von: Xue, Zongyue, et al.
Veröffentlicht: (2025)
Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work
von: Wu, Zihao, et al.
Veröffentlicht: (2026)
von: Wu, Zihao, et al.
Veröffentlicht: (2026)
S^2tory: Story Spine Distillation for Movie Script Summarization
von: Lu, Mingzhe, et al.
Veröffentlicht: (2026)
von: Lu, Mingzhe, et al.
Veröffentlicht: (2026)
Enhancing Video Summarization with Context Awareness
von: Huynh-Lam, Hai-Dang, et al.
Veröffentlicht: (2024)
von: Huynh-Lam, Hai-Dang, et al.
Veröffentlicht: (2024)
Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates
von: Li, Yibo, et al.
Veröffentlicht: (2026)
von: Li, Yibo, et al.
Veröffentlicht: (2026)
Hallucination Diversity-Aware Active Learning for Text Summarization
von: Xia, Yu, et al.
Veröffentlicht: (2024)
von: Xia, Yu, et al.
Veröffentlicht: (2024)
Generative Pre-Trained Transformer for Symbolic Regression Base In-Context Reinforcement Learning
von: Li, Yanjie, et al.
Veröffentlicht: (2024)
von: Li, Yanjie, et al.
Veröffentlicht: (2024)
Evaluating Small Language Models for News Summarization: Implications and Factors Influencing Performance
von: Xu, Borui, et al.
Veröffentlicht: (2025)
von: Xu, Borui, et al.
Veröffentlicht: (2025)
Precision in Practice: Knowledge Guided Code Summarizing Grounded in Industrial Expectations
von: Li, Jintai, et al.
Veröffentlicht: (2026)
von: Li, Jintai, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Just Enough Thinking: Efficient Reasoning with Adaptive Length Penalties Reinforcement Learning
von: Xiang, Violet, et al.
Veröffentlicht: (2025) -
Information-Theoretic Distillation for Reference-less Summarization
von: Jung, Jaehun, et al.
Veröffentlicht: (2024) -
Escaping the Context Bottleneck: Active Context Curation for LLM Agents via Reinforcement Learning
von: Li, Xiaozhe, et al.
Veröffentlicht: (2026) -
Context-CoT: Enhancing Context Learning via High-Quality Reasoning Synthesis
von: Jin, Hongbo, et al.
Veröffentlicht: (2026) -
Language Models for Text Classification: Is In-Context Learning Enough?
von: Edwards, Aleksandra, et al.
Veröffentlicht: (2024)