Low-Regret and Low-Complexity Learning for Hierarchical Inference
Fuente:
arXiv
Guardado en:
| Autores principales: | Chattopadhyay, Sameep, Sutar, Vinay, Champati, Jaya Prakash, Moharir, Sharayu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Observation-Free Attacks on Online Learning to Rank
por: Chattopadhyay, Sameep, et al.
Publicado: (2025)
por: Chattopadhyay, Sameep, et al.
Publicado: (2025)
Inference Offloading for Cost-Sensitive Binary Classification at the Edge
por: Moothedath, Vishnu Narayanan, et al.
Publicado: (2025)
por: Moothedath, Vishnu Narayanan, et al.
Publicado: (2025)
Online Algorithms for Hierarchical Inference in Deep Learning applications at the Edge
por: Moothedath, Vishnu Narayanan, et al.
Publicado: (2023)
por: Moothedath, Vishnu Narayanan, et al.
Publicado: (2023)
Fixed-Budget Constrained Best Arm Identification in Grouped Bandits
por: Mukherjee, Raunak, et al.
Publicado: (2026)
por: Mukherjee, Raunak, et al.
Publicado: (2026)
Cascading Bandits With Feedback
por: Prakash, R Sri, et al.
Publicado: (2025)
por: Prakash, R Sri, et al.
Publicado: (2025)
Towards Efficient Multi-LLM Inference: Characterization and Analysis of LLM Routing and Hierarchical Techniques
por: Behera, Adarsh Prasad, et al.
Publicado: (2025)
por: Behera, Adarsh Prasad, et al.
Publicado: (2025)
Pay for Hints, Not Answers: LLM Shepherding for Cost-Efficient Inference
por: Dong, Ziming, et al.
Publicado: (2026)
por: Dong, Ziming, et al.
Publicado: (2026)
On-line Learning in Tree MDPs by Treating Policies as Bandit Arms
por: Shah, Anvay, et al.
Publicado: (2026)
por: Shah, Anvay, et al.
Publicado: (2026)
Unreliable Multi-Armed Bandits: A Novel Approach to Recommendation Systems
por: Ravi, Aditya Narayan, et al.
Publicado: (2019)
por: Ravi, Aditya Narayan, et al.
Publicado: (2019)
Exploring the Boundaries of On-Device Inference: When Tiny Falls Short, Go Hierarchical
por: Behera, Adarsh Prasad, et al.
Publicado: (2024)
por: Behera, Adarsh Prasad, et al.
Publicado: (2024)
Influencing Bandits: Arm Selection for Preference Shaping
por: Nadkarni, Viraj, et al.
Publicado: (2024)
por: Nadkarni, Viraj, et al.
Publicado: (2024)
Constrained Best Arm Identification in Grouped Bandits
por: Dharod, Sahil, et al.
Publicado: (2024)
por: Dharod, Sahil, et al.
Publicado: (2024)
Parameter-efficient Adaptation of Multilingual Multimodal Models for Low-resource ASR
por: Gupta, Abhishek, et al.
Publicado: (2024)
por: Gupta, Abhishek, et al.
Publicado: (2024)
Fixed-Confidence Best Arm Identification with Decreasing Variance
por: Roychowdhury, Tamojeet, et al.
Publicado: (2025)
por: Roychowdhury, Tamojeet, et al.
Publicado: (2025)
Minimizing Age of Detection for a Markov Source over a Lossy Channel
por: Garde, Shivang, et al.
Publicado: (2025)
por: Garde, Shivang, et al.
Publicado: (2025)
Improved Decision Module Selection for Hierarchical Inference in Resource-Constrained Edge Devices
por: Behera, Adarsh Prasad, et al.
Publicado: (2024)
por: Behera, Adarsh Prasad, et al.
Publicado: (2024)
Fast and Regret Optimal Best Arm Identification: Fundamental Limits and Low-Complexity Algorithms
por: Zhang, Qining, et al.
Publicado: (2023)
por: Zhang, Qining, et al.
Publicado: (2023)
Federated Q-Learning: Linear Regret Speedup with Low Communication Cost
por: Zheng, Zhong, et al.
Publicado: (2023)
por: Zheng, Zhong, et al.
Publicado: (2023)
The Perils of Optimizing Learned Reward Functions: Low Training Error Does Not Guarantee Low Regret
por: Fluri, Lukas, et al.
Publicado: (2024)
por: Fluri, Lukas, et al.
Publicado: (2024)
AMPS: ASR with Multimodal Paraphrase Supervision
por: Gupta, Abhishek, et al.
Publicado: (2024)
por: Gupta, Abhishek, et al.
Publicado: (2024)
Regret-Optimal Q-Learning with Low Cost for Single-Agent and Federated Reinforcement Learning
por: Zhang, Haochen, et al.
Publicado: (2025)
por: Zhang, Haochen, et al.
Publicado: (2025)
Federated Learning of Binary Neural Networks: Enabling Low-Cost Inference
por: Shankar, Nitin Priyadarshini, et al.
Publicado: (2026)
por: Shankar, Nitin Priyadarshini, et al.
Publicado: (2026)
Efficient, Low-Regret, Online Reinforcement Learning for Linear MDPs
por: John, Philips George, et al.
Publicado: (2024)
por: John, Philips George, et al.
Publicado: (2024)
Low-Complexity Inference in Continual Learning via Compressed Knowledge Transfer
por: Liu, Zhenrong, et al.
Publicado: (2025)
por: Liu, Zhenrong, et al.
Publicado: (2025)
Provably Efficient Exploration in Reward Machines with Low Regret
por: Bourel, Hippolyte, et al.
Publicado: (2024)
por: Bourel, Hippolyte, et al.
Publicado: (2024)
Context Matters: Leveraging Contextual Features for Time Series Forecasting
por: Chattopadhyay, Sameep, et al.
Publicado: (2024)
por: Chattopadhyay, Sameep, et al.
Publicado: (2024)
Hierarchical Deep Counterfactual Regret Minimization
por: Chen, Jiayu, et al.
Publicado: (2023)
por: Chen, Jiayu, et al.
Publicado: (2023)
Regret-Oracle Complexity Tradeoffs in Agnostic Online Learning
por: Attias, Idan, et al.
Publicado: (2026)
por: Attias, Idan, et al.
Publicado: (2026)
PEAR: Primitive Enabled Adaptive Relabeling for Boosting Hierarchical Reinforcement Learning
por: Singh, Utsav, et al.
Publicado: (2023)
por: Singh, Utsav, et al.
Publicado: (2023)
FOSSIL: Regret-Minimizing Curriculum Learning for Metadata-Free and Low-Data Mpox Diagnosis
por: Han, Sahng-Min, et al.
Publicado: (2025)
por: Han, Sahng-Min, et al.
Publicado: (2025)
CRISP: Curriculum Inducing Primitive Informed Subgoal Prediction for Hierarchical Reinforcement Learning
por: Singh, Utsav, et al.
Publicado: (2023)
por: Singh, Utsav, et al.
Publicado: (2023)
DHP: Discrete Hierarchical Planning for Hierarchical Reinforcement Learning Agents
por: Sharma, Shashank, et al.
Publicado: (2025)
por: Sharma, Shashank, et al.
Publicado: (2025)
Curiosity is Knowledge: Self-Consistent Learning and No-Regret Optimization with Active Inference
por: Li, Yingke, et al.
Publicado: (2026)
por: Li, Yingke, et al.
Publicado: (2026)
Quantum Machine Learning with HQC Architectures using non-Classically Simulable Feature Maps
por: Ahmad, Syed Farhan, et al.
Publicado: (2021)
por: Ahmad, Syed Farhan, et al.
Publicado: (2021)
COLA: Continual Learning via Autoencoder Retrieval of Adapters
por: Mandivarapu, Jaya Krishna
Publicado: (2025)
por: Mandivarapu, Jaya Krishna
Publicado: (2025)
Renting Edge Computing Resources for Service Hosting
por: Madnaik, Aadesh, et al.
Publicado: (2022)
por: Madnaik, Aadesh, et al.
Publicado: (2022)
On the Low-Complexity of Fair Learning for Combinatorial Multi-Armed Bandit
por: Wu, Xiaoyi, et al.
Publicado: (2025)
por: Wu, Xiaoyi, et al.
Publicado: (2025)
A Study on Regularization-Based Continual Learning Methods for Indic ASR
por: T, Gokul Adethya, et al.
Publicado: (2025)
por: T, Gokul Adethya, et al.
Publicado: (2025)
Spatiotemporal deep learning models for detection of rapid intensification in cyclones
por: Sutar, Vamshika, et al.
Publicado: (2025)
por: Sutar, Vamshika, et al.
Publicado: (2025)
SHARP-QoS: Sparsely-gated Hierarchical Adaptive Routing for joint Prediction of QoS
por: Kumar, Suraj, et al.
Publicado: (2025)
por: Kumar, Suraj, et al.
Publicado: (2025)
Ejemplares similares
-
Observation-Free Attacks on Online Learning to Rank
por: Chattopadhyay, Sameep, et al.
Publicado: (2025) -
Inference Offloading for Cost-Sensitive Binary Classification at the Edge
por: Moothedath, Vishnu Narayanan, et al.
Publicado: (2025) -
Online Algorithms for Hierarchical Inference in Deep Learning applications at the Edge
por: Moothedath, Vishnu Narayanan, et al.
Publicado: (2023) -
Fixed-Budget Constrained Best Arm Identification in Grouped Bandits
por: Mukherjee, Raunak, et al.
Publicado: (2026) -
Cascading Bandits With Feedback
por: Prakash, R Sri, et al.
Publicado: (2025)