Learning What Matters Now: A Dual-Critic Context-Aware RL Framework for Priority-Driven Information Gain
Fuente:
arXiv
Saved in:
| Main Authors: | Panagopoulos, Dimitris, Perrusquia, Adolfo, Guo, Weisi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Signal Contract for Online Language Grounding and Discovery in Decision-Making
by: Panagopoulos, Dimitris, et al.
Published: (2025)
by: Panagopoulos, Dimitris, et al.
Published: (2025)
Dialogue Telemetry: Turn-Level Instrumentation for Autonomous Information Gathering
by: Panagopoulos, Dimitris, et al.
Published: (2026)
by: Panagopoulos, Dimitris, et al.
Published: (2026)
Selective Exploration and Information Gathering in Search and Rescue Using Hierarchical Learning Guided by Natural Language Input
by: Panagopoulos, Dimitrios, et al.
Published: (2024)
by: Panagopoulos, Dimitrios, et al.
Published: (2024)
Curriculum-Guided Antifragile Reinforcement Learning for Secure UAV Deconfliction under Observation-Space Attacks
by: Panda, Deepak Kumar, et al.
Published: (2025)
by: Panda, Deepak Kumar, et al.
Published: (2025)
Explainable Interface for Human-Autonomy Teaming: A Survey
by: Kong, Xiangqi, et al.
Published: (2024)
by: Kong, Xiangqi, et al.
Published: (2024)
Learning What Matters Now: Dynamic Preference Inference under Contextual Shifts
by: Cao, Xianwei, et al.
Published: (2026)
by: Cao, Xianwei, et al.
Published: (2026)
Where and What Matters: Sensitivity-Aware Task Vectors for Many-Shot Multimodal In-Context Learning
by: Ma, Ziyu, et al.
Published: (2025)
by: Ma, Ziyu, et al.
Published: (2025)
Real vs. Semi-Simulated: Rethinking Evaluation for Treatment Effect Estimation
by: Panagopoulos, George
Published: (2026)
by: Panagopoulos, George
Published: (2026)
Guaranteeing and Explaining Stability across Heterogeneous Load Balancing using Calculus Network Dynamics
by: Zou, Mengbang, et al.
Published: (2025)
by: Zou, Mengbang, et al.
Published: (2025)
RL-Driven Security-Aware Resource Allocation Framework for UAV-Assisted O-RAN
by: Abughazzah, Zaineh, et al.
Published: (2025)
by: Abughazzah, Zaineh, et al.
Published: (2025)
ProCeedRL: Process Critic with Exploratory Demonstration Reinforcement Learning for LLM Agentic Reasoning
by: Gao, Jingyue, et al.
Published: (2026)
by: Gao, Jingyue, et al.
Published: (2026)
Mixed-Precision Federated Learning via Multi-Precision Over-The-Air Aggregation
by: Yuan, Jinsheng, et al.
Published: (2024)
by: Yuan, Jinsheng, et al.
Published: (2024)
Minimize Control Inputs for Strong Structural Controllability Using Reinforcement Learning with Graph Neural Network
by: Zou, Mengbang, et al.
Published: (2024)
by: Zou, Mengbang, et al.
Published: (2024)
Time-Scaling Is What Agents Need Now
by: Liu, Zhi, et al.
Published: (2026)
by: Liu, Zhi, et al.
Published: (2026)
Towards Empowerment Gain through Causal Structure Learning in Model-Based RL
by: Cao, Hongye, et al.
Published: (2025)
by: Cao, Hongye, et al.
Published: (2025)
Dual-Prototype Disentanglement: A Context-Aware Enhancement Framework for Time Series Forecasting
by: Yang, Haonan, et al.
Published: (2026)
by: Yang, Haonan, et al.
Published: (2026)
Explainable Adversarial Learning Framework on Physical Layer Secret Keys Combating Malicious Reconfigurable Intelligent Surface
by: Wei, Zhuangkun, et al.
Published: (2024)
by: Wei, Zhuangkun, et al.
Published: (2024)
A Dual-Directional Context-Aware Test-Time Learning for Text Classification
by: Xu, Dong, et al.
Published: (2025)
by: Xu, Dong, et al.
Published: (2025)
Generalized Priority-Aware Shapley Value
by: Lee, Kiljae, et al.
Published: (2026)
by: Lee, Kiljae, et al.
Published: (2026)
Critic-Driven Voronoi-Quantization for Distilling Deep RL Policies to Explainable Models
by: Deproost, Senne, et al.
Published: (2026)
by: Deproost, Senne, et al.
Published: (2026)
Meta Policy Switching for Secure UAV Deconfliction in Adversarial Airspace
by: Panda, Deepak Kumar, et al.
Published: (2025)
by: Panda, Deepak Kumar, et al.
Published: (2025)
Generative Adversarial Evasion and Out-of-Distribution Detection for UAV Cyber-Attacks
by: Panda, Deepak Kumar, et al.
Published: (2025)
by: Panda, Deepak Kumar, et al.
Published: (2025)
PickLLM: Context-Aware RL-Assisted Large Language Model Routing
by: Sikeridis, Dimitrios, et al.
Published: (2024)
by: Sikeridis, Dimitrios, et al.
Published: (2024)
AIOT based Smart Education System: A Dual Layer Authentication and Context-Aware Tutoring Framework for Learning Environments
by: Neelakantan, Adithya, et al.
Published: (2025)
by: Neelakantan, Adithya, et al.
Published: (2025)
Optimizing What Matters: AUC-Driven Learning for Robust Neural Retrieval
by: Sheikholeslami, Nima, et al.
Published: (2025)
by: Sheikholeslami, Nima, et al.
Published: (2025)
AwareCompiler: Agentic Context-Aware Compiler Optimization via a Synergistic Knowledge-Data Driven Framework
by: Lin, Hongyu, et al.
Published: (2025)
by: Lin, Hongyu, et al.
Published: (2025)
RL-VLA$^3$: A Flexible and Asynchronous Reinforcement Learning Framework for VLA Training
by: Sun, Haoran, et al.
Published: (2026)
by: Sun, Haoran, et al.
Published: (2026)
ZipRL: Adaptive Multi-Turn Context Compression with Hindsight Response Replay
by: Hu, Zhexin, et al.
Published: (2026)
by: Hu, Zhexin, et al.
Published: (2026)
The Gaining Paths to Investment Success: Information-Driven LLM Graph Reasoning for Venture Capital Prediction
by: Pei, Haoyu, et al.
Published: (2025)
by: Pei, Haoyu, et al.
Published: (2025)
From Coordination to Personalization: A Trust-Aware Simulation Framework for Emergency Department Decision Support
by: Lygizou, Zoi, et al.
Published: (2025)
by: Lygizou, Zoi, et al.
Published: (2025)
Focus On What Matters: Separated Models For Visual-Based RL Generalization
by: Zhang, Di, et al.
Published: (2024)
by: Zhang, Di, et al.
Published: (2024)
Code Driven Planning with Domain-Adaptive Critic
by: Tian, Zikang, et al.
Published: (2025)
by: Tian, Zikang, et al.
Published: (2025)
Dual Alignment Maximin Optimization for Offline Model-based RL
by: Zhou, Chi, et al.
Published: (2025)
by: Zhou, Chi, et al.
Published: (2025)
Context and Diversity Matter: The Emergence of In-Context Learning in World Models
by: Wang, Fan, et al.
Published: (2025)
by: Wang, Fan, et al.
Published: (2025)
ContextRL: Enhancing MLLM's Knowledge Discovery Efficiency with Context-Augmented RL
by: Lu, Xingyu, et al.
Published: (2026)
by: Lu, Xingyu, et al.
Published: (2026)
Asking What Matters: Reward-Driven Clarification for Software Engineering Tasks
by: Vijayvargiya, Sanidhya, et al.
Published: (2026)
by: Vijayvargiya, Sanidhya, et al.
Published: (2026)
RLPO: Residual Listwise Preference Optimization for Long-Context Review Ranking
by: Jiang, Hao, et al.
Published: (2026)
by: Jiang, Hao, et al.
Published: (2026)
Context Matters: Incorporating Target Awareness in Conversational Abusive Language Detection
by: Alharthi, Raneem, et al.
Published: (2025)
by: Alharthi, Raneem, et al.
Published: (2025)
Context Matters: Repository-Aware Security Analysis of the Agent Skill Ecosystem
by: Holzbauer, Florian, et al.
Published: (2026)
by: Holzbauer, Florian, et al.
Published: (2026)
AdapShot: Adaptive Many-Shot In-Context Learning with Semantic-Aware KV Cache Reuse
by: Ou, Jie, et al.
Published: (2026)
by: Ou, Jie, et al.
Published: (2026)
Similar Items
-
A Signal Contract for Online Language Grounding and Discovery in Decision-Making
by: Panagopoulos, Dimitris, et al.
Published: (2025) -
Dialogue Telemetry: Turn-Level Instrumentation for Autonomous Information Gathering
by: Panagopoulos, Dimitris, et al.
Published: (2026) -
Selective Exploration and Information Gathering in Search and Rescue Using Hierarchical Learning Guided by Natural Language Input
by: Panagopoulos, Dimitrios, et al.
Published: (2024) -
Curriculum-Guided Antifragile Reinforcement Learning for Secure UAV Deconfliction under Observation-Space Attacks
by: Panda, Deepak Kumar, et al.
Published: (2025) -
Explainable Interface for Human-Autonomy Teaming: A Survey
by: Kong, Xiangqi, et al.
Published: (2024)