Online Learning for Multi-Layer Hierarchical Inference under Partial and Policy-Dependent Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Haoran, Cha, Seohyeon, Beytur, Hasan Burhan, Chan, Kevin S, de Veciana, Gustavo, Vikalo, Haris |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimal Resource Allocation for ML Model Training and Deployment under Concept Drift
by: Beytur, Hasan Burhan, et al.
Published: (2025)
by: Beytur, Hasan Burhan, et al.
Published: (2025)
Batching-Aware Joint Model Onloading and Offloading for Hierarchical Multi-Task Inference
by: Cha, Seohyeon, et al.
Published: (2025)
by: Cha, Seohyeon, et al.
Published: (2025)
CoreQ: Learning-Free Mismatch Correction and Successive Rounding for Quantization
by: Cha, Seohyeon, et al.
Published: (2026)
by: Cha, Seohyeon, et al.
Published: (2026)
FedRot-LoRA: Mitigating Rotational Misalignment in Federated LoRA
by: Zhang, Haoran, et al.
Published: (2026)
by: Zhang, Haoran, et al.
Published: (2026)
Task-Agnostic Federated Continual Learning via Replay-Free Gradient Projection
by: Cha, Seohyeon, et al.
Published: (2025)
by: Cha, Seohyeon, et al.
Published: (2025)
Federated Self-Supervised Learning for Automatic Modulation Classification under Non-IID and Class-Imbalanced Data
by: Akram, Usman, et al.
Published: (2025)
by: Akram, Usman, et al.
Published: (2025)
GeFL: Model-Agnostic Federated Learning with Generative Models
by: Kang, Honggu, et al.
Published: (2024)
by: Kang, Honggu, et al.
Published: (2024)
When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search
by: Robertson, John T., et al.
Published: (2026)
by: Robertson, John T., et al.
Published: (2026)
NeFL: Nested Model Scaling for Federated Learning with System Heterogeneous Clients
by: Kang, Honggu, et al.
Published: (2023)
by: Kang, Honggu, et al.
Published: (2023)
Foundation-Preserving Adaptation via Generalized Rayleigh-Quotient Optimization
by: Kim, Dongjun, et al.
Published: (2026)
by: Kim, Dongjun, et al.
Published: (2026)
Hierarchically Gated Experts for Efficient Online Continual Learning
by: Luong, Kevin, et al.
Published: (2024)
by: Luong, Kevin, et al.
Published: (2024)
Your Agents Are Aging Too: Agent Lifespan Engineering for Deployed Systems
by: Zhu, Jianing, et al.
Published: (2026)
by: Zhu, Jianing, et al.
Published: (2026)
Online Feedback Efficient Active Target Discovery in Partially Observable Environments
by: Sarkar, Anindya, et al.
Published: (2025)
by: Sarkar, Anindya, et al.
Published: (2025)
Safeguard Text-to-Image Diffusion Models with Human Feedback Inversion
by: Kim, Sanghyun, et al.
Published: (2024)
by: Kim, Sanghyun, et al.
Published: (2024)
Uncertainty Representations in State-Space Layers for Deep Reinforcement Learning under Partial Observability
by: Luis, Carlos E., et al.
Published: (2024)
by: Luis, Carlos E., et al.
Published: (2024)
Guided Policy Optimization under Partial Observability
by: Li, Yueheng, et al.
Published: (2025)
by: Li, Yueheng, et al.
Published: (2025)
DarkPatterns-LLM: A Multi-Layer Benchmark for Detecting Manipulative and Harmful AI Behavior
by: Asif, Sadia, et al.
Published: (2025)
by: Asif, Sadia, et al.
Published: (2025)
Recovering Labels from Local Updates in Federated Learning
by: Chen, Huancheng, et al.
Published: (2024)
by: Chen, Huancheng, et al.
Published: (2024)
Transformers as Implicit State Estimators: In-Context Learning in Dynamical Systems
by: Akram, Usman, et al.
Published: (2024)
by: Akram, Usman, et al.
Published: (2024)
Negative Feedback System as Optimizer for Machine Learning Systems
by: Hasan, Md Munir, et al.
Published: (2021)
by: Hasan, Md Munir, et al.
Published: (2021)
The Role of Artificial Intelligence and Machine Learning in Software Testing
by: Ramadan, Ahmed, et al.
Published: (2024)
by: Ramadan, Ahmed, et al.
Published: (2024)
Heterogeneity-Guided Client Sampling: Towards Fast and Efficient Non-IID Federated Learning
by: Chen, Huancheng, et al.
Published: (2023)
by: Chen, Huancheng, et al.
Published: (2023)
Near-Optimal Partially Observable Reinforcement Learning with Partial Online State Information
by: Shi, Ming, et al.
Published: (2023)
by: Shi, Ming, et al.
Published: (2023)
Versatile Navigation under Partial Observability via Value-guided Diffusion Policy
by: Zhang, Gengyu, et al.
Published: (2024)
by: Zhang, Gengyu, et al.
Published: (2024)
Decentralized Consensus Inference-based Hierarchical Reinforcement Learning for Multi-Constrained UAV Pursuit-Evasion Game
by: Yuming, Xiang, et al.
Published: (2025)
by: Yuming, Xiang, et al.
Published: (2025)
Deep Hierarchical Reinforcement Learning Algorithm in Partially Observable Markov Decision Processes
by: Tuyen, Le Pham, et al.
Published: (2018)
by: Tuyen, Le Pham, et al.
Published: (2018)
PIMSM: Physics-Informed Multi-Scale Mamba for Stable Neural Representations under Distribution Shift
by: Bae, Sangyoon, et al.
Published: (2026)
by: Bae, Sangyoon, et al.
Published: (2026)
Belief States for Cooperative Multi-Agent Reinforcement Learning under Partial Observability
by: Pritz, Paul J., et al.
Published: (2025)
by: Pritz, Paul J., et al.
Published: (2025)
Proposing Hierarchical Goal-Conditioned Policy Planning in Multi-Goal Reinforcement Learning
by: Rens, Gavin B.
Published: (2025)
by: Rens, Gavin B.
Published: (2025)
Sim-to-Real of Humanoid Locomotion Policies via Joint Torque Space Perturbation Injection
by: Cha, Junhyeok Rui, et al.
Published: (2026)
by: Cha, Junhyeok Rui, et al.
Published: (2026)
Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies
by: Li, Zhuoran, et al.
Published: (2026)
by: Li, Zhuoran, et al.
Published: (2026)
Towards Off-Policy Reinforcement Learning for Ranking Policies with Human Feedback
by: Xiao, Teng, et al.
Published: (2024)
by: Xiao, Teng, et al.
Published: (2024)
DepthKV: Layer-Dependent KV Cache Pruning for Long-Context LLM Inference
by: Dehghanighobadi, Zahra, et al.
Published: (2026)
by: Dehghanighobadi, Zahra, et al.
Published: (2026)
Policy Expansion for Bridging Offline-to-Online Reinforcement Learning
by: Zhang, Haichao, et al.
Published: (2023)
by: Zhang, Haichao, et al.
Published: (2023)
Key-Conditioned Orthonormal Transform Gating (K-OTG): Multi-Key Access Control with Hidden-State Scrambling for LoRA-Tuned Models
by: Khan, Muhammad Haris
Published: (2025)
by: Khan, Muhammad Haris
Published: (2025)
Who Uses AI? Platform Selection and the Measurement of Occupational AI Exposure
by: Yin, Michelle, et al.
Published: (2026)
by: Yin, Michelle, et al.
Published: (2026)
Success in Humanoid Reinforcement Learning under Partial Observation
by: Wang, Wuhao, et al.
Published: (2025)
by: Wang, Wuhao, et al.
Published: (2025)
Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning
by: Bozkurt, Alper Kamil, et al.
Published: (2026)
by: Bozkurt, Alper Kamil, et al.
Published: (2026)
ABC3: Active Bayesian Causal Inference with Cohn Criteria in Randomized Experiments
by: Cha, Taehun, et al.
Published: (2024)
by: Cha, Taehun, et al.
Published: (2024)
Hybrid AI for Responsive Multi-Turn Online Conversations with Novel Dynamic Routing and Feedback Adaptation
by: Pattnayak, Priyaranjan, et al.
Published: (2025)
by: Pattnayak, Priyaranjan, et al.
Published: (2025)
Similar Items
-
Optimal Resource Allocation for ML Model Training and Deployment under Concept Drift
by: Beytur, Hasan Burhan, et al.
Published: (2025) -
Batching-Aware Joint Model Onloading and Offloading for Hierarchical Multi-Task Inference
by: Cha, Seohyeon, et al.
Published: (2025) -
CoreQ: Learning-Free Mismatch Correction and Successive Rounding for Quantization
by: Cha, Seohyeon, et al.
Published: (2026) -
FedRot-LoRA: Mitigating Rotational Misalignment in Federated LoRA
by: Zhang, Haoran, et al.
Published: (2026) -
Task-Agnostic Federated Continual Learning via Replay-Free Gradient Projection
by: Cha, Seohyeon, et al.
Published: (2025)