Sparse Linear Bandits with Blocking Constraints
Fuente:
arXiv
Guardado en:
| Autores principales: | Jain, Adit, Pal, Soumyabrata, Choudhary, Sunav, Narayanam, Ramasuri, Chopra, Harshita, Krishnamurthy, Vikram |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Structured Reinforcement Learning for Incentivized Stochastic Covert Optimization
por: Jain, Adit, et al.
Publicado: (2024)
por: Jain, Adit, et al.
Publicado: (2024)
From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning
por: Purohit, Kiran, et al.
Publicado: (2026)
por: Purohit, Kiran, et al.
Publicado: (2026)
Interacting Large Language Model Agents. Interpretable Models and Social Learning
por: Jain, Adit, et al.
Publicado: (2024)
por: Jain, Adit, et al.
Publicado: (2024)
Sequential Causal Discovery with Noisy Language Model Priors
por: Verma, Prakhar, et al.
Publicado: (2025)
por: Verma, Prakhar, et al.
Publicado: (2025)
Delivery Optimized Discovery in Behavioral User Segmentation under Budget Constraint
por: Chopra, Harshita, et al.
Publicado: (2024)
por: Chopra, Harshita, et al.
Publicado: (2024)
Tab-Shapley: Identifying Top-k Tabular Data Quality Insights
por: Padala, Manisha, et al.
Publicado: (2025)
por: Padala, Manisha, et al.
Publicado: (2025)
Regret minimization in Linear Bandits with offline data via extended D-optimal exploration
por: Vijayan, Sushant, et al.
Publicado: (2025)
por: Vijayan, Sushant, et al.
Publicado: (2025)
Why Most Optimism Bandit Algorithms Have the Same Regret Analysis: A Simple Unifying Theorem
por: Krishnamurthy, Vikram
Publicado: (2025)
por: Krishnamurthy, Vikram
Publicado: (2025)
CAFIN: Centrality Aware Fairness inducing IN-processing for Unsupervised Representation Learning on Graphs
por: Arun, Arvindh, et al.
Publicado: (2023)
por: Arun, Arvindh, et al.
Publicado: (2023)
Identifying Hate Speech Peddlers in Online Platforms. A Bayesian Social Learning Approach for Large Language Model Driven Decision-Makers
por: Jain, Adit, et al.
Publicado: (2024)
por: Jain, Adit, et al.
Publicado: (2024)
Fake or Compromised? Making Sense of Malicious Clients in Federated Learning
por: Mozaffari, Hamid, et al.
Publicado: (2024)
por: Mozaffari, Hamid, et al.
Publicado: (2024)
Improving Selective Classification with Pairwise Queries for Binary Classification
por: Vardhan, Harsh, et al.
Publicado: (2026)
por: Vardhan, Harsh, et al.
Publicado: (2026)
The Cost of Avoiding Backpropagation
por: Panchal, Kunjal, et al.
Publicado: (2025)
por: Panchal, Kunjal, et al.
Publicado: (2025)
Inferring Group Intent as a Cooperative Game. An NLP-based Framework for Trajectory Analysis
por: Zhang, Yiming, et al.
Publicado: (2025)
por: Zhang, Yiming, et al.
Publicado: (2025)
Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization
por: Krishnamurthy, Vikram
Publicado: (2025)
por: Krishnamurthy, Vikram
Publicado: (2025)
Feedback-Aware Monte Carlo Tree Search for Efficient Information Seeking in Goal-Oriented Conversations
por: Chopra, Harshita, et al.
Publicado: (2025)
por: Chopra, Harshita, et al.
Publicado: (2025)
Pure Exploration in Bandits with Linear Constraints
por: Carlsson, Emil, et al.
Publicado: (2023)
por: Carlsson, Emil, et al.
Publicado: (2023)
Flow: Per-Instance Personalized Federated Learning Through Dynamic Routing
por: Panchal, Kunjal, et al.
Publicado: (2022)
por: Panchal, Kunjal, et al.
Publicado: (2022)
Online Matrix Completion: A Collaborative Approach with Hott Items
por: Baby, Dheeraj, et al.
Publicado: (2024)
por: Baby, Dheeraj, et al.
Publicado: (2024)
Information Diffusion and Preferential Attachment in a Network of Large Language Models
por: Jain, Adit, et al.
Publicado: (2025)
por: Jain, Adit, et al.
Publicado: (2025)
Collaborative QA using Interacting LLMs. Impact of Network Structure, Node Capability and Distributed Data
por: Jain, Adit, et al.
Publicado: (2025)
por: Jain, Adit, et al.
Publicado: (2025)
Distributed Linear Bandits under Communication Constraints
por: Salgia, Sudeep, et al.
Publicado: (2022)
por: Salgia, Sudeep, et al.
Publicado: (2022)
Learning to Reason with Mixture of Tokens
por: Jain, Adit, et al.
Publicado: (2025)
por: Jain, Adit, et al.
Publicado: (2025)
Malliavin Calculus for Counterfactual Gradient Estimation in Adaptive Inverse Reinforcement Learning
por: Krishnamurthy, Vikram, et al.
Publicado: (2026)
por: Krishnamurthy, Vikram, et al.
Publicado: (2026)
Finite-Sample Bounds for Adaptive Inverse Reinforcement Learning using Passive Langevin Dynamics
por: Snow, Luke, et al.
Publicado: (2023)
por: Snow, Luke, et al.
Publicado: (2023)
Efficient Neural SDE Training using Wiener-Space Cubature
por: Snow, Luke, et al.
Publicado: (2025)
por: Snow, Luke, et al.
Publicado: (2025)
LLMs as High-Dimensional Nonlinear Autoregressive Models with Attention: Training, Alignment and Inference
por: Krishnamurthy, Vikram
Publicado: (2026)
por: Krishnamurthy, Vikram
Publicado: (2026)
Blocking Bandits
por: Basu, Soumya, et al.
Publicado: (2019)
por: Basu, Soumya, et al.
Publicado: (2019)
Thinking Forward: Memory-Efficient Federated Finetuning of Language Models
por: Panchal, Kunjal, et al.
Publicado: (2024)
por: Panchal, Kunjal, et al.
Publicado: (2024)
Classifier Language Models: Unifying Sparse Finetuning and Adaptive Tokenization for Specialized Classification Tasks
por: Krishnan, Adit, et al.
Publicado: (2025)
por: Krishnan, Adit, et al.
Publicado: (2025)
Finite Sample and Large Deviations Analysis of Stochastic Gradient Algorithm with Correlated Noise
por: Yin, George, et al.
Publicado: (2024)
por: Yin, George, et al.
Publicado: (2024)
Malliavin Calculus with Weak Derivatives for Counterfactual Stochastic Optimization
por: Krishnamurthy, Vikram, et al.
Publicado: (2025)
por: Krishnamurthy, Vikram, et al.
Publicado: (2025)
Distributionally Robust Inverse Reinforcement Learning for Identifying Multi-Agent Coordinated Sensing
por: Snow, Luke, et al.
Publicado: (2024)
por: Snow, Luke, et al.
Publicado: (2024)
Slow Convergence of Interacting Kalman Filters in Word-of-Mouth Social Learning
por: Krishnamurthy, Vikram, et al.
Publicado: (2024)
por: Krishnamurthy, Vikram, et al.
Publicado: (2024)
Optimal Multitask Linear Regression and Contextual Bandits under Sparse Heterogeneity
por: Huang, Xinmeng, et al.
Publicado: (2023)
por: Huang, Xinmeng, et al.
Publicado: (2023)
LEWIS (LayEr WIse Sparsity) -- A Training Free Guided Model Merging Approach
por: Chopra, Hetarth, et al.
Publicado: (2025)
por: Chopra, Hetarth, et al.
Publicado: (2025)
Language Models Entangle Language and Culture
por: Jain, Shourya, et al.
Publicado: (2026)
por: Jain, Shourya, et al.
Publicado: (2026)
Detecting Structural Shifts in Multivariate Hawkes Processes with Fréchet Statistics
por: Luo, Rui, et al.
Publicado: (2023)
por: Luo, Rui, et al.
Publicado: (2023)
Data Wrangling Task Automation Using Code-Generating Language Models
por: Akella, Ashlesha, et al.
Publicado: (2025)
por: Akella, Ashlesha, et al.
Publicado: (2025)
Revisiting Matrix Sketching in Linear Bandits: Achieving Sublinear Regret via Dyadic Block Sketching
por: Wen, Dongxie, et al.
Publicado: (2024)
por: Wen, Dongxie, et al.
Publicado: (2024)
Ejemplares similares
-
Structured Reinforcement Learning for Incentivized Stochastic Covert Optimization
por: Jain, Adit, et al.
Publicado: (2024) -
From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning
por: Purohit, Kiran, et al.
Publicado: (2026) -
Interacting Large Language Model Agents. Interpretable Models and Social Learning
por: Jain, Adit, et al.
Publicado: (2024) -
Sequential Causal Discovery with Noisy Language Model Priors
por: Verma, Prakhar, et al.
Publicado: (2025) -
Delivery Optimized Discovery in Behavioral User Segmentation under Budget Constraint
por: Chopra, Harshita, et al.
Publicado: (2024)