CAESAR: Enhancing Federated RL in Heterogeneous MDPs through Convergence-Aware Sampling with Screening
Fuente:
arXiv
Saved in:
| Main Authors: | Mak, Hei Yi, Fan, Flint Xiaofeng, Lanzendörfer, Luca A., Tan, Cheston, Ooi, Wei Tsang, Wattenhofer, Roger |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF
by: Fan, Flint Xiaofeng, et al.
Published: (2024)
by: Fan, Flint Xiaofeng, et al.
Published: (2024)
Position Paper: Rethinking Privacy in RL for Sequential Decision-making in the Age of LLMs
by: Fan, Flint Xiaofeng, et al.
Published: (2025)
by: Fan, Flint Xiaofeng, et al.
Published: (2025)
SUBER: An RL Environment with Simulated Human Behavior for Recommender Systems
by: Corecco, Nathan, et al.
Published: (2024)
by: Corecco, Nathan, et al.
Published: (2024)
Information Fidelity in Tool-Using LLM Agents: A Martingale Analysis of the Model Context Protocol
by: Fan, Flint Xiaofeng, et al.
Published: (2026)
by: Fan, Flint Xiaofeng, et al.
Published: (2026)
Decentralized Federated Policy Gradient with Byzantine Fault-Tolerance and Provably Fast Convergence
by: Jordan, Philip, et al.
Published: (2024)
by: Jordan, Philip, et al.
Published: (2024)
FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation
by: Jiang, Wenzheng, et al.
Published: (2025)
by: Jiang, Wenzheng, et al.
Published: (2025)
Alignment-Aware Decoding
by: Berdoz, Frédéric, et al.
Published: (2025)
by: Berdoz, Frédéric, et al.
Published: (2025)
Cue Point Estimation using Object Detection
by: Argüello, Giulia, et al.
Published: (2024)
by: Argüello, Giulia, et al.
Published: (2024)
Towards Leveraging Contrastively Pretrained Neural Audio Embeddings for Recommender Tasks
by: Grötschla, Florian, et al.
Published: (2024)
by: Grötschla, Florian, et al.
Published: (2024)
MaskBeat: Loopable Drum Beat Generation
by: Lanzendörfer, Luca A., et al.
Published: (2025)
by: Lanzendörfer, Luca A., et al.
Published: (2025)
Benchmarking Diarization Models
by: Lanzendörfer, Luca A., et al.
Published: (2025)
by: Lanzendörfer, Luca A., et al.
Published: (2025)
AEye: A Visualization Tool for Image Datasets
by: Grötschla, Florian, et al.
Published: (2024)
by: Grötschla, Florian, et al.
Published: (2024)
Benchmarking Music Generation Models and Metrics via Human Preference Studies
by: Grötschla, Florian, et al.
Published: (2025)
by: Grötschla, Florian, et al.
Published: (2025)
Parametric Neural Amp Modeling with Active Learning
by: Grötschla, Florian, et al.
Published: (2025)
by: Grötschla, Florian, et al.
Published: (2025)
Parametric Neural Amp Modeling with Active Learning
by: Grötschla, Florian, et al.
Published: (2025)
by: Grötschla, Florian, et al.
Published: (2025)
Multi-bit Audio Watermarking
by: Lanzendörfer, Luca A., et al.
Published: (2025)
by: Lanzendörfer, Luca A., et al.
Published: (2025)
High-Fidelity Speech Enhancement via Discrete Audio Tokens
by: Lanzendörfer, Luca A., et al.
Published: (2025)
by: Lanzendörfer, Luca A., et al.
Published: (2025)
Audio Atlas: Visualizing and Exploring Audio Datasets
by: Lanzendörfer, Luca A., et al.
Published: (2024)
by: Lanzendörfer, Luca A., et al.
Published: (2024)
Inductive Transfer Learning for Graph-Based Recommenders
by: Grötschla, Florian, et al.
Published: (2025)
by: Grötschla, Florian, et al.
Published: (2025)
Bias beyond Borders: Global Inequalities in AI-Generated Music
by: Solak, Ahmet, et al.
Published: (2025)
by: Solak, Ahmet, et al.
Published: (2025)
Text-to-Scene with Large Reasoning Models
by: Berdoz, Frédéric, et al.
Published: (2025)
by: Berdoz, Frédéric, et al.
Published: (2025)
High-Fidelity Music Vocoder using Neural Audio Codecs
by: Lanzendörfer, Luca A., et al.
Published: (2025)
by: Lanzendörfer, Luca A., et al.
Published: (2025)
SALSA-V: Shortcut-Augmented Long-form Synchronized Audio from Videos
by: Dellali, Amir, et al.
Published: (2025)
by: Dellali, Amir, et al.
Published: (2025)
WorldSpeech: A Multilingual Speech Corpus from Around the World
by: Asonitis, Antonis, et al.
Published: (2026)
by: Asonitis, Antonis, et al.
Published: (2026)
PUZZLES: A Benchmark for Neural Algorithmic Reasoning
by: Estermann, Benjamin, et al.
Published: (2024)
by: Estermann, Benjamin, et al.
Published: (2024)
EuroSpeech: A Multilingual Speech Corpus
by: Pfisterer, Samuel, et al.
Published: (2025)
by: Pfisterer, Samuel, et al.
Published: (2025)
Adapting Neural Audio Codecs to EEG
by: Kastrati, Ard, et al.
Published: (2025)
by: Kastrati, Ard, et al.
Published: (2025)
Low-Resource NMT: A Case Study on the Written and Spoken Languages in Hong Kong
by: Mak, Hei Yi, et al.
Published: (2025)
by: Mak, Hei Yi, et al.
Published: (2025)
SAO-Instruct: Free-form Audio Editing using Natural Language Instructions
by: Ungersböck, Michael, et al.
Published: (2025)
by: Ungersböck, Michael, et al.
Published: (2025)
Why Do We Suffer for Fun? Ordeal Pleasure in Souls-like Games
by: Fan, Flint Xiaofeng
Published: (2026)
by: Fan, Flint Xiaofeng
Published: (2026)
Hybrid Reinforcement Learning Breaks Sample Size Barriers in Linear MDPs
by: Tan, Kevin, et al.
Published: (2024)
by: Tan, Kevin, et al.
Published: (2024)
Social Learning through Interactions with Other Agents: A Survey
by: Hillier, Dylan, et al.
Published: (2024)
by: Hillier, Dylan, et al.
Published: (2024)
TangramSR: Can Vision-Language Models Reason in Continuous Geometric Space?
by: Zong, Yikun, et al.
Published: (2026)
by: Zong, Yikun, et al.
Published: (2026)
Evaluating Objective Speech Quality Metrics for Neural Audio Codecs
by: Lanzendörfer, Luca A., et al.
Published: (2025)
by: Lanzendörfer, Luca A., et al.
Published: (2025)
Planning for Dexterous Ungrasping: Secure Ungrasping through Dexterous Manipulation
by: Kim, Chung Hee, et al.
Published: (2021)
by: Kim, Chung Hee, et al.
Published: (2021)
Adaptive Federated Learning in Heterogeneous Wireless Networks with Independent Sampling
by: Geng, Jiaxiang, et al.
Published: (2024)
by: Geng, Jiaxiang, et al.
Published: (2024)
Convergence of processes time-changed by Gaussian multiplicative chaos
by: Ooi, Takumu
Published: (2023)
by: Ooi, Takumu
Published: (2023)
Sketch&Patch++: Efficient Structure-Aware 3D Gaussian Representation
by: Shi, Yuang, et al.
Published: (2026)
by: Shi, Yuang, et al.
Published: (2026)
On Value Iteration Convergence in Connected MDPs
by: Mustafin, Arsenii, et al.
Published: (2024)
by: Mustafin, Arsenii, et al.
Published: (2024)
Prior-Aligned Meta-RL: Thompson Sampling with Learned Priors and Guarantees in Finite-Horizon MDPs
by: Zhou, Runlin, et al.
Published: (2025)
by: Zhou, Runlin, et al.
Published: (2025)
Similar Items
-
FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF
by: Fan, Flint Xiaofeng, et al.
Published: (2024) -
Position Paper: Rethinking Privacy in RL for Sequential Decision-making in the Age of LLMs
by: Fan, Flint Xiaofeng, et al.
Published: (2025) -
SUBER: An RL Environment with Simulated Human Behavior for Recommender Systems
by: Corecco, Nathan, et al.
Published: (2024) -
Information Fidelity in Tool-Using LLM Agents: A Martingale Analysis of the Model Context Protocol
by: Fan, Flint Xiaofeng, et al.
Published: (2026) -
Decentralized Federated Policy Gradient with Byzantine Fault-Tolerance and Provably Fast Convergence
by: Jordan, Philip, et al.
Published: (2024)