Minimizing Human Intervention in Online Classification
Fuente:
arXiv
Guardado en:
| Autores principales: | Réveillard, William, Saketos, Vasileios, Proutiere, Alexandre, Combes, Richard |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multimodal Bandits: Regret Lower Bounds and Optimal Algorithms
por: Réveillard, William, et al.
Publicado: (2025)
por: Réveillard, William, et al.
Publicado: (2025)
Low-Rank Bandits via Tight Two-to-Infinity Singular Subspace Recovery
por: Jedra, Yassir, et al.
Publicado: (2024)
por: Jedra, Yassir, et al.
Publicado: (2024)
The Kalman Evolve: Closing the Gap in Kalman Filtering via Interpretable Algorithm Discovery
por: Saketos, Vasileios, et al.
Publicado: (2026)
por: Saketos, Vasileios, et al.
Publicado: (2026)
The Large Language Model GreekLegalRoBERTa
por: Saketos, Vasileios, et al.
Publicado: (2024)
por: Saketos, Vasileios, et al.
Publicado: (2024)
Data-Driven Discovery of Interpretable Kalman Filter Variants through Large Language Models and Genetic Programming
por: Saketos, Vasileios, et al.
Publicado: (2025)
por: Saketos, Vasileios, et al.
Publicado: (2025)
Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning
por: Stojanovic, Stefan, et al.
Publicado: (2026)
por: Stojanovic, Stefan, et al.
Publicado: (2026)
Model-Free Active Exploration in Reinforcement Learning
por: Russo, Alessio, et al.
Publicado: (2024)
por: Russo, Alessio, et al.
Publicado: (2024)
Curvature-Guided LoRA: Steering in the pretrained NTK subspace
por: Zheng, Frédéric, et al.
Publicado: (2026)
por: Zheng, Frédéric, et al.
Publicado: (2026)
Conformal Predictions under Markovian Data
por: Zheng, Frédéric, et al.
Publicado: (2024)
por: Zheng, Frédéric, et al.
Publicado: (2024)
Optimal Centered Active Excitation in Linear System Identification
por: Ito, Kaito, et al.
Publicado: (2026)
por: Ito, Kaito, et al.
Publicado: (2026)
Shift Before You Learn: Enabling Low-Rank Representations in Reinforcement Learning
por: Dubail, Bastien, et al.
Publicado: (2025)
por: Dubail, Bastien, et al.
Publicado: (2025)
On Universally Optimal Algorithms for A/B Testing
por: Wang, Po-An, et al.
Publicado: (2023)
por: Wang, Po-An, et al.
Publicado: (2023)
Model-free Low-Rank Reinforcement Learning via Leveraged Entry-wise Matrix Estimation
por: Stojanovic, Stefan, et al.
Publicado: (2024)
por: Stojanovic, Stefan, et al.
Publicado: (2024)
Best Arm Identification with Fixed Budget: A Large Deviation Perspective
por: Wang, Po-An, et al.
Publicado: (2023)
por: Wang, Po-An, et al.
Publicado: (2023)
Adversarial Diffusion for Robust Reinforcement Learning
por: Foffano, Daniele, et al.
Publicado: (2025)
por: Foffano, Daniele, et al.
Publicado: (2025)
Near-optimal Rank Adaptive Inference of High Dimensional Matrices
por: Zheng, Frédéric, et al.
Publicado: (2025)
por: Zheng, Frédéric, et al.
Publicado: (2025)
Adaptive Reinforcement Learning for Unobservable Random Delays
por: Wikman, John, et al.
Publicado: (2025)
por: Wikman, John, et al.
Publicado: (2025)
Conformal Off-Policy Evaluation in Markov Decision Processes
por: Foffano, Daniele, et al.
Publicado: (2023)
por: Foffano, Daniele, et al.
Publicado: (2023)
Revisiting Instance-Optimal Cluster Recovery in the Labeled Stochastic Block Model
por: Ariu, Kaito, et al.
Publicado: (2023)
por: Ariu, Kaito, et al.
Publicado: (2023)
Optimal Clustering from Noisy Binary Feedback
por: Ariu, Kaito, et al.
Publicado: (2019)
por: Ariu, Kaito, et al.
Publicado: (2019)
Policy Testing in Markov Decision Processes
por: Ariu, Kaito, et al.
Publicado: (2025)
por: Ariu, Kaito, et al.
Publicado: (2025)
An extension of McDiarmid's inequality
por: Combes, Richard
Publicado: (2015)
por: Combes, Richard
Publicado: (2015)
Thompson Sampling For Combinatorial Bandits: Polynomial Regret and Mismatched Sampling Paradox
por: Zhang, Raymond, et al.
Publicado: (2024)
por: Zhang, Raymond, et al.
Publicado: (2024)
Near-Optimal Clustering in Mixture of Markov Chains
por: Lee, Junghyun, et al.
Publicado: (2025)
por: Lee, Junghyun, et al.
Publicado: (2025)
Autonomous Algorithm for Training Autonomous Vehicles with Minimal Human Intervention
por: Lee, Sang-Hyun, et al.
Publicado: (2024)
por: Lee, Sang-Hyun, et al.
Publicado: (2024)
Advantage-Guided Diffusion for Model-Based Reinforcement Learning
por: Foffano, Daniele, et al.
Publicado: (2026)
por: Foffano, Daniele, et al.
Publicado: (2026)
Online Learning for Function Placement in Serverless Computing
por: Huang, Wei, et al.
Publicado: (2024)
por: Huang, Wei, et al.
Publicado: (2024)
Tractable Instances of Bilinear Maximization: Implementing LinUCB on Ellipsoids
por: Zhang, Raymond, et al.
Publicado: (2025)
por: Zhang, Raymond, et al.
Publicado: (2025)
Linear Bandits on Ellipsoids: Minimax Optimal Algorithms
por: Zhang, Raymond, et al.
Publicado: (2025)
por: Zhang, Raymond, et al.
Publicado: (2025)
Instance-Optimal Estimation with Multiple LLM Judges on a Budget
por: Lee, Junghyun, et al.
Publicado: (2026)
por: Lee, Junghyun, et al.
Publicado: (2026)
Distilling On-device Language Models for Robot Planning with Minimal Human Intervention
por: Ravichandran, Zachary, et al.
Publicado: (2025)
por: Ravichandran, Zachary, et al.
Publicado: (2025)
Online Cluster-Based Parameter Control for Metaheuristic
por: Tatsis, Vasileios A., et al.
Publicado: (2025)
por: Tatsis, Vasileios A., et al.
Publicado: (2025)
Online $\mathrm{L}^{\natural}$-Convex Minimization
por: Yokoyama, Ken, et al.
Publicado: (2024)
por: Yokoyama, Ken, et al.
Publicado: (2024)
Kalman Filter for Online Classification of Non-Stationary Data
por: Titsias, Michalis K., et al.
Publicado: (2023)
por: Titsias, Michalis K., et al.
Publicado: (2023)
Fair Classification by Direct Intervention on Operating Characteristics
por: Jiang, Kevin, et al.
Publicado: (2025)
por: Jiang, Kevin, et al.
Publicado: (2025)
PCARNN-DCBF: Minimal-Intervention Geofence Enforcement for Ground Vehicles
por: Yu, Yinan, et al.
Publicado: (2025)
por: Yu, Yinan, et al.
Publicado: (2025)
Smoothed Online Classification can be Harder than Batch Classification
por: Raman, Vinod, et al.
Publicado: (2024)
por: Raman, Vinod, et al.
Publicado: (2024)
Learning-augmented Online Minimization of Age of Information and Transmission Costs
por: Liu, Zhongdong, et al.
Publicado: (2024)
por: Liu, Zhongdong, et al.
Publicado: (2024)
Online Deterministic Annealing for Classification and Clustering
por: Mavridis, Christos, et al.
Publicado: (2021)
por: Mavridis, Christos, et al.
Publicado: (2021)
Distributed Online Submodular Maximization under Communication Delays: A Simultaneous Decision-Making Approach
por: Xu, Zirui, et al.
Publicado: (2026)
por: Xu, Zirui, et al.
Publicado: (2026)
Ejemplares similares
-
Multimodal Bandits: Regret Lower Bounds and Optimal Algorithms
por: Réveillard, William, et al.
Publicado: (2025) -
Low-Rank Bandits via Tight Two-to-Infinity Singular Subspace Recovery
por: Jedra, Yassir, et al.
Publicado: (2024) -
The Kalman Evolve: Closing the Gap in Kalman Filtering via Interpretable Algorithm Discovery
por: Saketos, Vasileios, et al.
Publicado: (2026) -
The Large Language Model GreekLegalRoBERTa
por: Saketos, Vasileios, et al.
Publicado: (2024) -
Data-Driven Discovery of Interpretable Kalman Filter Variants through Large Language Models and Genetic Programming
por: Saketos, Vasileios, et al.
Publicado: (2025)