BanditQ: Fair Bandits with Guaranteed Rewards
Fuente:
arXiv
Guardado en:
| Autor principal: | Sinha, Abhishek |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HPC Application Parameter Autotuning on Edge Devices: A Bandit Learning Approach
por: Hossain, Abrar, et al.
Publicado: (2025)
por: Hossain, Abrar, et al.
Publicado: (2025)
Multi-agent Multi-armed Bandits with Minimum Reward Guarantee Fairness
por: Manupriya, Piyushi, et al.
Publicado: (2025)
por: Manupriya, Piyushi, et al.
Publicado: (2025)
Constrained Contextual Bandits with Adversarial Contexts
por: Sarkar, Dhruv, et al.
Publicado: (2026)
por: Sarkar, Dhruv, et al.
Publicado: (2026)
Fairness and Privacy Guarantees in Federated Contextual Bandits
por: Solanki, Sambhav, et al.
Publicado: (2024)
por: Solanki, Sambhav, et al.
Publicado: (2024)
Adversarial Network Optimization under Bandit Feedback: Maximizing Utility in Non-Stationary Multi-Hop Networks
por: Dai, Yan, et al.
Publicado: (2024)
por: Dai, Yan, et al.
Publicado: (2024)
A Simple Reduction Scheme for Constrained Contextual Bandits with Adversarial Contexts via Regression
por: Sarkar, Dhruv, et al.
Publicado: (2026)
por: Sarkar, Dhruv, et al.
Publicado: (2026)
Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback
por: Goyal, Tanmay, et al.
Publicado: (2025)
por: Goyal, Tanmay, et al.
Publicado: (2025)
Fairness in Serving Large Language Models
por: Sheng, Ying, et al.
Publicado: (2023)
por: Sheng, Ying, et al.
Publicado: (2023)
A2Q+: Improving Accumulator-Aware Weight Quantization
por: Colbert, Ian, et al.
Publicado: (2024)
por: Colbert, Ian, et al.
Publicado: (2024)
Finite-Time Guarantees for Multi-Agent Combinatorial Bandits with Nonstationary Rewards
por: Adams, Katherine B., et al.
Publicado: (2025)
por: Adams, Katherine B., et al.
Publicado: (2025)
Score-Aware Policy-Gradient and Performance Guarantees using Local Lyapunov Stability
por: Comte, Céline, et al.
Publicado: (2023)
por: Comte, Céline, et al.
Publicado: (2023)
Stochastic Q-learning for Large Discrete Action Spaces
por: Fourati, Fares, et al.
Publicado: (2024)
por: Fourati, Fares, et al.
Publicado: (2024)
Single Index Bandits: Generalized Linear Contextual Bandits with Unknown Reward Functions
por: Kang, Yue, et al.
Publicado: (2025)
por: Kang, Yue, et al.
Publicado: (2025)
Bandit Simulation for Average Reward Inference
por: Praharaj, Samya, et al.
Publicado: (2026)
por: Praharaj, Samya, et al.
Publicado: (2026)
Bandit Max-Min Fair Allocation
por: Harada, Tsubasa, et al.
Publicado: (2025)
por: Harada, Tsubasa, et al.
Publicado: (2025)
LiveTune: Dynamic Parameter Tuning for Feedback-Driven Optimization
por: Shabgahi, Soheil Zibakhsh, et al.
Publicado: (2023)
por: Shabgahi, Soheil Zibakhsh, et al.
Publicado: (2023)
Towards Computational Performance Engineering for Unsupervised Concept Drift Detection -- Complexities, Benchmarking, Performance Analysis
por: Werner, Elias, et al.
Publicado: (2023)
por: Werner, Elias, et al.
Publicado: (2023)
Benchmarking GPUs on SVBRDF Extractor Model
por: Kandel, Narayan, et al.
Publicado: (2023)
por: Kandel, Narayan, et al.
Publicado: (2023)
SENSEi: Input-Sensitive Compilation for Accelerating GNNs
por: Lenadora, Damitha, et al.
Publicado: (2023)
por: Lenadora, Damitha, et al.
Publicado: (2023)
oneDNN Graph Compiler: A Hybrid Approach for High-Performance Deep Learning Compilation
por: Li, Jianhui, et al.
Publicado: (2023)
por: Li, Jianhui, et al.
Publicado: (2023)
U-TOE: Universal TinyML On-board Evaluation Toolkit for Low-Power IoT
por: Huang, Zhaolan, et al.
Publicado: (2023)
por: Huang, Zhaolan, et al.
Publicado: (2023)
Conformer-Based Speech Recognition On Extreme Edge-Computing Devices
por: Xu, Mingbin, et al.
Publicado: (2023)
por: Xu, Mingbin, et al.
Publicado: (2023)
A Scalable k-Medoids Clustering via Whale Optimization Algorithm
por: Chenan, Huang, et al.
Publicado: (2024)
por: Chenan, Huang, et al.
Publicado: (2024)
CPINN-ABPI: Physics-Informed Neural Networks for Accurate Power Estimation in MPSoCs
por: Elshamy, Mohamed R., et al.
Publicado: (2025)
por: Elshamy, Mohamed R., et al.
Publicado: (2025)
GPU-Accelerated INT8 Quantization for KV Cache Compression in Large Language Models
por: Taneja, Maanas, et al.
Publicado: (2026)
por: Taneja, Maanas, et al.
Publicado: (2026)
DistZO2: High-Throughput and Memory-Efficient Zeroth-Order Fine-tuning LLMs with Distributed Parallel Computing
por: Wang, Liangyu, et al.
Publicado: (2025)
por: Wang, Liangyu, et al.
Publicado: (2025)
PARD: Accelerating LLM Inference with Low-Cost PARallel Draft Model Adaptation
por: An, Zihao, et al.
Publicado: (2025)
por: An, Zihao, et al.
Publicado: (2025)
Flashlight: PyTorch Compiler Extensions to Accelerate Attention Variants
por: You, Bozhi, et al.
Publicado: (2025)
por: You, Bozhi, et al.
Publicado: (2025)
MarginGate: Sparse Margin-Triggered Verification for Batch-Invariant LLM Inference
por: Chu, Kexin, et al.
Publicado: (2026)
por: Chu, Kexin, et al.
Publicado: (2026)
Parallel Implementations Assessment of a Spatial-Spectral Classifier for Hyperspectral Clinical Applications
por: Lazcano, Raquel, et al.
Publicado: (2024)
por: Lazcano, Raquel, et al.
Publicado: (2024)
A Structure-Aware Framework for Learning Device Placements on Computation Graphs
por: Duan, Shukai, et al.
Publicado: (2024)
por: Duan, Shukai, et al.
Publicado: (2024)
Automating Energy-Efficient GPU Kernel Generation: A Fast Search-Based Compilation Approach
por: Zhang, Yijia, et al.
Publicado: (2024)
por: Zhang, Yijia, et al.
Publicado: (2024)
Enhancing Tropical Cyclone Path Forecasting with an Improved Transformer Network
por: Van Thanh, Nguyen, et al.
Publicado: (2025)
por: Van Thanh, Nguyen, et al.
Publicado: (2025)
lm-Meter: Unveiling Runtime Inference Latency for On-Device Language Models
por: Wang, Haoxin, et al.
Publicado: (2025)
por: Wang, Haoxin, et al.
Publicado: (2025)
Reducing Compute Waste in LLMs through Kernel-Level DVFS
por: Spaan, Jeffrey, et al.
Publicado: (2026)
por: Spaan, Jeffrey, et al.
Publicado: (2026)
WCDT: Systematic WCET Optimization for Decision Tree Implementations
por: Hölscher, Nils, et al.
Publicado: (2025)
por: Hölscher, Nils, et al.
Publicado: (2025)
KernelBenchX: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernels
por: Wang, Han, et al.
Publicado: (2026)
por: Wang, Han, et al.
Publicado: (2026)
Single-Thread JPEG Decoder Benchmarks Mis-Evaluate ML Data Loaders
por: Iglovikov, Vladimir, et al.
Publicado: (2026)
por: Iglovikov, Vladimir, et al.
Publicado: (2026)
Plug-and-Play Performance Estimation for LLM Services without Relying on Labeled Data
por: Wang, Can, et al.
Publicado: (2024)
por: Wang, Can, et al.
Publicado: (2024)
MoE-Inference-Bench: Performance Evaluation of Mixture of Expert Large Language and Vision Models
por: Chitty-Venkata, Krishna Teja, et al.
Publicado: (2025)
por: Chitty-Venkata, Krishna Teja, et al.
Publicado: (2025)
Ejemplares similares
-
HPC Application Parameter Autotuning on Edge Devices: A Bandit Learning Approach
por: Hossain, Abrar, et al.
Publicado: (2025) -
Multi-agent Multi-armed Bandits with Minimum Reward Guarantee Fairness
por: Manupriya, Piyushi, et al.
Publicado: (2025) -
Constrained Contextual Bandits with Adversarial Contexts
por: Sarkar, Dhruv, et al.
Publicado: (2026) -
Fairness and Privacy Guarantees in Federated Contextual Bandits
por: Solanki, Sambhav, et al.
Publicado: (2024) -
Adversarial Network Optimization under Bandit Feedback: Maximizing Utility in Non-Stationary Multi-Hop Networks
por: Dai, Yan, et al.
Publicado: (2024)