Gespeichert in:
| Hauptverfasser: | Jedra, Yassir, Réveillard, William, Stojanovic, Stefan, Proutiere, Alexandre |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2402.15739 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Model-free Low-Rank Reinforcement Learning via Leveraged Entry-wise Matrix Estimation
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2024)
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2024)
Near-optimal Rank Adaptive Inference of High Dimensional Matrices
von: Zheng, Frédéric, et al.
Veröffentlicht: (2025)
von: Zheng, Frédéric, et al.
Veröffentlicht: (2025)
Minimal Order Recovery through Rank-adaptive Identification
von: Zheng, Frédéric, et al.
Veröffentlicht: (2025)
von: Zheng, Frédéric, et al.
Veröffentlicht: (2025)
Shift Before You Learn: Enabling Low-Rank Representations in Reinforcement Learning
von: Dubail, Bastien, et al.
Veröffentlicht: (2025)
von: Dubail, Bastien, et al.
Veröffentlicht: (2025)
Near-Optimal Clustering in Mixture of Markov Chains
von: Lee, Junghyun, et al.
Veröffentlicht: (2025)
von: Lee, Junghyun, et al.
Veröffentlicht: (2025)
Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2026)
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2026)
Minimizing Human Intervention in Online Classification
von: Réveillard, William, et al.
Veröffentlicht: (2025)
von: Réveillard, William, et al.
Veröffentlicht: (2025)
Instance-Optimal Estimation with Multiple LLM Judges on a Budget
von: Lee, Junghyun, et al.
Veröffentlicht: (2026)
von: Lee, Junghyun, et al.
Veröffentlicht: (2026)
Multimodal Bandits: Regret Lower Bounds and Optimal Algorithms
von: Réveillard, William, et al.
Veröffentlicht: (2025)
von: Réveillard, William, et al.
Veröffentlicht: (2025)
$k$-SVD with Gradient Descent
von: Jedra, Yassir, et al.
Veröffentlicht: (2025)
von: Jedra, Yassir, et al.
Veröffentlicht: (2025)
Explore-then-Commit for Nonstationary Linear Bandits with Latent Dynamics
von: Choi, Sunmook, et al.
Veröffentlicht: (2025)
von: Choi, Sunmook, et al.
Veröffentlicht: (2025)
Learning Linear Dynamics from Bilinear Observations
von: Sattar, Yahya, et al.
Veröffentlicht: (2024)
von: Sattar, Yahya, et al.
Veröffentlicht: (2024)
Exploiting Observation Bias to Improve Matrix Completion
von: Jedra, Yassir, et al.
Veröffentlicht: (2023)
von: Jedra, Yassir, et al.
Veröffentlicht: (2023)
Finite Sample Identification of Partially Observed Bilinear Dynamical Systems
von: Sattar, Yahya, et al.
Veröffentlicht: (2025)
von: Sattar, Yahya, et al.
Veröffentlicht: (2025)
Revisiting Instance-Optimal Cluster Recovery in the Labeled Stochastic Block Model
von: Ariu, Kaito, et al.
Veröffentlicht: (2023)
von: Ariu, Kaito, et al.
Veröffentlicht: (2023)
Optimal Transfer Learning for Missing Not-at-Random Matrix Completion
von: Jalan, Akhil, et al.
Veröffentlicht: (2025)
von: Jalan, Akhil, et al.
Veröffentlicht: (2025)
Model-Free Active Exploration in Reinforcement Learning
von: Russo, Alessio, et al.
Veröffentlicht: (2024)
von: Russo, Alessio, et al.
Veröffentlicht: (2024)
Curvature-Guided LoRA: Steering in the pretrained NTK subspace
von: Zheng, Frédéric, et al.
Veröffentlicht: (2026)
von: Zheng, Frédéric, et al.
Veröffentlicht: (2026)
Sub-optimality of the Separation Principle for Quadratic Control from Bilinear Observations
von: Sattar, Yahya, et al.
Veröffentlicht: (2025)
von: Sattar, Yahya, et al.
Veröffentlicht: (2025)
Catching a Moving Subspace: Low-Rank Bandits Beyond Stationarity
von: Khosravi, Hamed, et al.
Veröffentlicht: (2026)
von: Khosravi, Hamed, et al.
Veröffentlicht: (2026)
Conformal Predictions under Markovian Data
von: Zheng, Frédéric, et al.
Veröffentlicht: (2024)
von: Zheng, Frédéric, et al.
Veröffentlicht: (2024)
Optimal Centered Active Excitation in Linear System Identification
von: Ito, Kaito, et al.
Veröffentlicht: (2026)
von: Ito, Kaito, et al.
Veröffentlicht: (2026)
A Tutorial on the Non-Asymptotic Theory of System Identification
von: Ziemann, Ingvar, et al.
Veröffentlicht: (2023)
von: Ziemann, Ingvar, et al.
Veröffentlicht: (2023)
On Universally Optimal Algorithms for A/B Testing
von: Wang, Po-An, et al.
Veröffentlicht: (2023)
von: Wang, Po-An, et al.
Veröffentlicht: (2023)
Adversarial Diffusion for Robust Reinforcement Learning
von: Foffano, Daniele, et al.
Veröffentlicht: (2025)
von: Foffano, Daniele, et al.
Veröffentlicht: (2025)
Best Arm Identification with Fixed Budget: A Large Deviation Perspective
von: Wang, Po-An, et al.
Veröffentlicht: (2023)
von: Wang, Po-An, et al.
Veröffentlicht: (2023)
Adaptive Reinforcement Learning for Unobservable Random Delays
von: Wikman, John, et al.
Veröffentlicht: (2025)
von: Wikman, John, et al.
Veröffentlicht: (2025)
Conformal Off-Policy Evaluation in Markov Decision Processes
von: Foffano, Daniele, et al.
Veröffentlicht: (2023)
von: Foffano, Daniele, et al.
Veröffentlicht: (2023)
Mixture-of-Subspaces in Low-Rank Adaptation
von: Wu, Taiqiang, et al.
Veröffentlicht: (2024)
von: Wu, Taiqiang, et al.
Veröffentlicht: (2024)
From Low Rank Gradient Subspace Stabilization to Low-Rank Weights: Observations, Theories, and Applications
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2024)
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2024)
Optimal Clustering from Noisy Binary Feedback
von: Ariu, Kaito, et al.
Veröffentlicht: (2019)
von: Ariu, Kaito, et al.
Veröffentlicht: (2019)
Policy Testing in Markov Decision Processes
von: Ariu, Kaito, et al.
Veröffentlicht: (2025)
von: Ariu, Kaito, et al.
Veröffentlicht: (2025)
Efficient Generalized Low-Rank Tensor Contextual Bandits
von: Yi, Qianxin, et al.
Veröffentlicht: (2023)
von: Yi, Qianxin, et al.
Veröffentlicht: (2023)
Low-Rank Compression of Pretrained Models via Randomized Subspace Iteration
von: Pourkamali-Anaraki, Farhad
Veröffentlicht: (2026)
von: Pourkamali-Anaraki, Farhad
Veröffentlicht: (2026)
Efficient Low-Rank Matrix Estimation, Experimental Design, and Arm-Set-Dependent Low-Rank Bandits
von: Jang, Kyoungseok, et al.
Veröffentlicht: (2024)
von: Jang, Kyoungseok, et al.
Veröffentlicht: (2024)
Online Minimization of Polarization and Disagreement via Low-Rank Matrix Bandits
von: Cinus, Federico, et al.
Veröffentlicht: (2025)
von: Cinus, Federico, et al.
Veröffentlicht: (2025)
Efficient Frameworks for Generalized Low-Rank Matrix Bandit Problems
von: Kang, Yue, et al.
Veröffentlicht: (2024)
von: Kang, Yue, et al.
Veröffentlicht: (2024)
Generalized Low-Rank Matrix Contextual Bandits with Graph Information
von: Wang, Yao, et al.
Veröffentlicht: (2025)
von: Wang, Yao, et al.
Veröffentlicht: (2025)
Tight Rates for Bandit Control Beyond Quadratics
von: Sun, Y. Jennifer, et al.
Veröffentlicht: (2024)
von: Sun, Y. Jennifer, et al.
Veröffentlicht: (2024)
Interpretable Safety Alignment via SAE-Constructed Low-Rank Subspace Adaptation
von: Wang, Dianyun, et al.
Veröffentlicht: (2025)
von: Wang, Dianyun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Model-free Low-Rank Reinforcement Learning via Leveraged Entry-wise Matrix Estimation
von: Stojanovic, Stefan, et al.
Veröffentlicht: (2024) -
Near-optimal Rank Adaptive Inference of High Dimensional Matrices
von: Zheng, Frédéric, et al.
Veröffentlicht: (2025) -
Minimal Order Recovery through Rank-adaptive Identification
von: Zheng, Frédéric, et al.
Veröffentlicht: (2025) -
Shift Before You Learn: Enabling Low-Rank Representations in Reinforcement Learning
von: Dubail, Bastien, et al.
Veröffentlicht: (2025) -
Near-Optimal Clustering in Mixture of Markov Chains
von: Lee, Junghyun, et al.
Veröffentlicht: (2025)