Sparse Linear Bandits with Blocking Constraints
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jain, Adit, Pal, Soumyabrata, Choudhary, Sunav, Narayanam, Ramasuri, Chopra, Harshita, Krishnamurthy, Vikram |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Structured Reinforcement Learning for Incentivized Stochastic Covert Optimization
von: Jain, Adit, et al.
Veröffentlicht: (2024)
von: Jain, Adit, et al.
Veröffentlicht: (2024)
From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning
von: Purohit, Kiran, et al.
Veröffentlicht: (2026)
von: Purohit, Kiran, et al.
Veröffentlicht: (2026)
Interacting Large Language Model Agents. Interpretable Models and Social Learning
von: Jain, Adit, et al.
Veröffentlicht: (2024)
von: Jain, Adit, et al.
Veröffentlicht: (2024)
Sequential Causal Discovery with Noisy Language Model Priors
von: Verma, Prakhar, et al.
Veröffentlicht: (2025)
von: Verma, Prakhar, et al.
Veröffentlicht: (2025)
Delivery Optimized Discovery in Behavioral User Segmentation under Budget Constraint
von: Chopra, Harshita, et al.
Veröffentlicht: (2024)
von: Chopra, Harshita, et al.
Veröffentlicht: (2024)
Tab-Shapley: Identifying Top-k Tabular Data Quality Insights
von: Padala, Manisha, et al.
Veröffentlicht: (2025)
von: Padala, Manisha, et al.
Veröffentlicht: (2025)
Regret minimization in Linear Bandits with offline data via extended D-optimal exploration
von: Vijayan, Sushant, et al.
Veröffentlicht: (2025)
von: Vijayan, Sushant, et al.
Veröffentlicht: (2025)
Why Most Optimism Bandit Algorithms Have the Same Regret Analysis: A Simple Unifying Theorem
von: Krishnamurthy, Vikram
Veröffentlicht: (2025)
von: Krishnamurthy, Vikram
Veröffentlicht: (2025)
CAFIN: Centrality Aware Fairness inducing IN-processing for Unsupervised Representation Learning on Graphs
von: Arun, Arvindh, et al.
Veröffentlicht: (2023)
von: Arun, Arvindh, et al.
Veröffentlicht: (2023)
Identifying Hate Speech Peddlers in Online Platforms. A Bayesian Social Learning Approach for Large Language Model Driven Decision-Makers
von: Jain, Adit, et al.
Veröffentlicht: (2024)
von: Jain, Adit, et al.
Veröffentlicht: (2024)
Fake or Compromised? Making Sense of Malicious Clients in Federated Learning
von: Mozaffari, Hamid, et al.
Veröffentlicht: (2024)
von: Mozaffari, Hamid, et al.
Veröffentlicht: (2024)
Improving Selective Classification with Pairwise Queries for Binary Classification
von: Vardhan, Harsh, et al.
Veröffentlicht: (2026)
von: Vardhan, Harsh, et al.
Veröffentlicht: (2026)
The Cost of Avoiding Backpropagation
von: Panchal, Kunjal, et al.
Veröffentlicht: (2025)
von: Panchal, Kunjal, et al.
Veröffentlicht: (2025)
Inferring Group Intent as a Cooperative Game. An NLP-based Framework for Trajectory Analysis
von: Zhang, Yiming, et al.
Veröffentlicht: (2025)
von: Zhang, Yiming, et al.
Veröffentlicht: (2025)
Inverse Reinforcement Learning using Revealed Preferences and Passive Stochastic Optimization
von: Krishnamurthy, Vikram
Veröffentlicht: (2025)
von: Krishnamurthy, Vikram
Veröffentlicht: (2025)
Feedback-Aware Monte Carlo Tree Search for Efficient Information Seeking in Goal-Oriented Conversations
von: Chopra, Harshita, et al.
Veröffentlicht: (2025)
von: Chopra, Harshita, et al.
Veröffentlicht: (2025)
Pure Exploration in Bandits with Linear Constraints
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)
von: Carlsson, Emil, et al.
Veröffentlicht: (2023)
Flow: Per-Instance Personalized Federated Learning Through Dynamic Routing
von: Panchal, Kunjal, et al.
Veröffentlicht: (2022)
von: Panchal, Kunjal, et al.
Veröffentlicht: (2022)
Online Matrix Completion: A Collaborative Approach with Hott Items
von: Baby, Dheeraj, et al.
Veröffentlicht: (2024)
von: Baby, Dheeraj, et al.
Veröffentlicht: (2024)
Information Diffusion and Preferential Attachment in a Network of Large Language Models
von: Jain, Adit, et al.
Veröffentlicht: (2025)
von: Jain, Adit, et al.
Veröffentlicht: (2025)
Collaborative QA using Interacting LLMs. Impact of Network Structure, Node Capability and Distributed Data
von: Jain, Adit, et al.
Veröffentlicht: (2025)
von: Jain, Adit, et al.
Veröffentlicht: (2025)
Distributed Linear Bandits under Communication Constraints
von: Salgia, Sudeep, et al.
Veröffentlicht: (2022)
von: Salgia, Sudeep, et al.
Veröffentlicht: (2022)
Learning to Reason with Mixture of Tokens
von: Jain, Adit, et al.
Veröffentlicht: (2025)
von: Jain, Adit, et al.
Veröffentlicht: (2025)
Malliavin Calculus for Counterfactual Gradient Estimation in Adaptive Inverse Reinforcement Learning
von: Krishnamurthy, Vikram, et al.
Veröffentlicht: (2026)
von: Krishnamurthy, Vikram, et al.
Veröffentlicht: (2026)
Finite-Sample Bounds for Adaptive Inverse Reinforcement Learning using Passive Langevin Dynamics
von: Snow, Luke, et al.
Veröffentlicht: (2023)
von: Snow, Luke, et al.
Veröffentlicht: (2023)
Efficient Neural SDE Training using Wiener-Space Cubature
von: Snow, Luke, et al.
Veröffentlicht: (2025)
von: Snow, Luke, et al.
Veröffentlicht: (2025)
LLMs as High-Dimensional Nonlinear Autoregressive Models with Attention: Training, Alignment and Inference
von: Krishnamurthy, Vikram
Veröffentlicht: (2026)
von: Krishnamurthy, Vikram
Veröffentlicht: (2026)
Blocking Bandits
von: Basu, Soumya, et al.
Veröffentlicht: (2019)
von: Basu, Soumya, et al.
Veröffentlicht: (2019)
Thinking Forward: Memory-Efficient Federated Finetuning of Language Models
von: Panchal, Kunjal, et al.
Veröffentlicht: (2024)
von: Panchal, Kunjal, et al.
Veröffentlicht: (2024)
Classifier Language Models: Unifying Sparse Finetuning and Adaptive Tokenization for Specialized Classification Tasks
von: Krishnan, Adit, et al.
Veröffentlicht: (2025)
von: Krishnan, Adit, et al.
Veröffentlicht: (2025)
Finite Sample and Large Deviations Analysis of Stochastic Gradient Algorithm with Correlated Noise
von: Yin, George, et al.
Veröffentlicht: (2024)
von: Yin, George, et al.
Veröffentlicht: (2024)
Malliavin Calculus with Weak Derivatives for Counterfactual Stochastic Optimization
von: Krishnamurthy, Vikram, et al.
Veröffentlicht: (2025)
von: Krishnamurthy, Vikram, et al.
Veröffentlicht: (2025)
Distributionally Robust Inverse Reinforcement Learning for Identifying Multi-Agent Coordinated Sensing
von: Snow, Luke, et al.
Veröffentlicht: (2024)
von: Snow, Luke, et al.
Veröffentlicht: (2024)
Slow Convergence of Interacting Kalman Filters in Word-of-Mouth Social Learning
von: Krishnamurthy, Vikram, et al.
Veröffentlicht: (2024)
von: Krishnamurthy, Vikram, et al.
Veröffentlicht: (2024)
Optimal Multitask Linear Regression and Contextual Bandits under Sparse Heterogeneity
von: Huang, Xinmeng, et al.
Veröffentlicht: (2023)
von: Huang, Xinmeng, et al.
Veröffentlicht: (2023)
LEWIS (LayEr WIse Sparsity) -- A Training Free Guided Model Merging Approach
von: Chopra, Hetarth, et al.
Veröffentlicht: (2025)
von: Chopra, Hetarth, et al.
Veröffentlicht: (2025)
Language Models Entangle Language and Culture
von: Jain, Shourya, et al.
Veröffentlicht: (2026)
von: Jain, Shourya, et al.
Veröffentlicht: (2026)
Detecting Structural Shifts in Multivariate Hawkes Processes with Fréchet Statistics
von: Luo, Rui, et al.
Veröffentlicht: (2023)
von: Luo, Rui, et al.
Veröffentlicht: (2023)
Data Wrangling Task Automation Using Code-Generating Language Models
von: Akella, Ashlesha, et al.
Veröffentlicht: (2025)
von: Akella, Ashlesha, et al.
Veröffentlicht: (2025)
Revisiting Matrix Sketching in Linear Bandits: Achieving Sublinear Regret via Dyadic Block Sketching
von: Wen, Dongxie, et al.
Veröffentlicht: (2024)
von: Wen, Dongxie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Structured Reinforcement Learning for Incentivized Stochastic Covert Optimization
von: Jain, Adit, et al.
Veröffentlicht: (2024) -
From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning
von: Purohit, Kiran, et al.
Veröffentlicht: (2026) -
Interacting Large Language Model Agents. Interpretable Models and Social Learning
von: Jain, Adit, et al.
Veröffentlicht: (2024) -
Sequential Causal Discovery with Noisy Language Model Priors
von: Verma, Prakhar, et al.
Veröffentlicht: (2025) -
Delivery Optimized Discovery in Behavioral User Segmentation under Budget Constraint
von: Chopra, Harshita, et al.
Veröffentlicht: (2024)