Efficient Adversarial Attacks on High-dimensional Offline Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Hosseini, Seyed Mohammad Hadi, Najafi, Amir, Baghshah, Mahdieh Soleymani |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SUSD: Structured Unsupervised Skill Discovery through State Factorization
by: Hosseini, Seyed Mohammad Hadi, et al.
Published: (2026)
by: Hosseini, Seyed Mohammad Hadi, et al.
Published: (2026)
CER: Confidence Enhanced Reasoning in LLMs
by: Razghandi, Ali, et al.
Published: (2025)
by: Razghandi, Ali, et al.
Published: (2025)
LibraGrad: Balancing Gradient Flow for Universally Better Vision Transformer Attributions
by: Mehri, Faridoun, et al.
Published: (2024)
by: Mehri, Faridoun, et al.
Published: (2024)
Dilated Balanced Cross Entropy Loss for Medical Image Segmentation
by: Hosseini, Seyed Mohsen, et al.
Published: (2024)
by: Hosseini, Seyed Mohsen, et al.
Published: (2024)
Inductive Biases for Zero-shot Systematic Generalization in Language-informed Reinforcement Learning
by: Dijujin, Negin Hashemi, et al.
Published: (2025)
by: Dijujin, Negin Hashemi, et al.
Published: (2025)
MEENA (PersianMMMU): Multimodal-Multilingual Educational Exams for N-level Assessment
by: Ghahroodi, Omid, et al.
Published: (2025)
by: Ghahroodi, Omid, et al.
Published: (2025)
CAREL: Instruction-guided reinforcement learning with cross-modal auxiliary objectives
by: Saghafian, Armin, et al.
Published: (2024)
by: Saghafian, Armin, et al.
Published: (2024)
Improving 3D Few-Shot Segmentation with Inference-Time Pseudo-Labeling
by: Mozafari, Mohammad, et al.
Published: (2024)
by: Mozafari, Mohammad, et al.
Published: (2024)
Learning to Attack: A Bandit Approach to Adversarial Context Poisoning
by: Telikani, Ray, et al.
Published: (2026)
by: Telikani, Ray, et al.
Published: (2026)
Bridging Reasoning to Learning: Unmasking Illusions using Complexity Out of Distribution Generalization
by: Paqaleh, Mohammad Mahdi Samiei, et al.
Published: (2025)
by: Paqaleh, Mohammad Mahdi Samiei, et al.
Published: (2025)
Fine-Grained Alignment and Noise Refinement for Compositional Text-to-Image Generation
by: Izadi, Amir Mohammad, et al.
Published: (2025)
by: Izadi, Amir Mohammad, et al.
Published: (2025)
Visual Structures Helps Visual Reasoning: Addressing the Binding Problem in VLMs
by: Izadi, Amirmohammad, et al.
Published: (2025)
by: Izadi, Amirmohammad, et al.
Published: (2025)
Trained Models Tell Us How to Make Them Robust to Spurious Correlation without Group Annotation
by: Ghaznavi, Mahdi, et al.
Published: (2024)
by: Ghaznavi, Mahdi, et al.
Published: (2024)
Exploiting Layer-Specific Vulnerabilities to Backdoor Attack in Federated Learning
by: Foroughi, Mohammad Hadi, et al.
Published: (2026)
by: Foroughi, Mohammad Hadi, et al.
Published: (2026)
The Illusion of Procedural Reasoning: Measuring Long-Horizon FSM Execution in LLMs
by: Samiei, Mahdi, et al.
Published: (2025)
by: Samiei, Mahdi, et al.
Published: (2025)
Robust Satisficing Gaussian Process Bandits Under Adversarial Attacks
by: Saday, Artun, et al.
Published: (2025)
by: Saday, Artun, et al.
Published: (2025)
Optimizing Warfarin Dosing Using Contextual Bandit: An Offline Policy Learning and Evaluation Method
by: Huang, Yong, et al.
Published: (2024)
by: Huang, Yong, et al.
Published: (2024)
VQEL: Enabling Self-Play in Emergent Language Games via Agent-Internal Vector Quantization
by: Paqaleh, Mohammad Mahdi Samiei, et al.
Published: (2025)
by: Paqaleh, Mohammad Mahdi Samiei, et al.
Published: (2025)
Exploiting Expertise of Non-Expert and Diverse Agents in Social Bandit Learning: A Free Energy Approach
by: Mirzaei, Erfan, et al.
Published: (2026)
by: Mirzaei, Erfan, et al.
Published: (2026)
Adaptive Locally Linear Embedding
by: Goli, Ali, et al.
Published: (2025)
by: Goli, Ali, et al.
Published: (2025)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
by: Xiong, Guojun, et al.
Published: (2024)
by: Xiong, Guojun, et al.
Published: (2024)
Attention and Autoencoder Hybrid Model for Unsupervised Online Anomaly Detection
by: Najafi, Seyed Amirhossein, et al.
Published: (2024)
by: Najafi, Seyed Amirhossein, et al.
Published: (2024)
GABInsight: Exploring Gender-Activity Binding Bias in Vision-Language Models
by: Abdollahi, Ali, et al.
Published: (2024)
by: Abdollahi, Ali, et al.
Published: (2024)
Leveraging Offline Data in Linear Latent Contextual Bandits
by: Kausik, Chinmaya, et al.
Published: (2024)
by: Kausik, Chinmaya, et al.
Published: (2024)
Towards Robust Policy: Enhancing Offline Reinforcement Learning with Adversarial Attacks and Defenses
by: Nguyen, Thanh, et al.
Published: (2024)
by: Nguyen, Thanh, et al.
Published: (2024)
Practical Adversarial Attacks on Stochastic Bandits via Fake Data Injection
by: Zeng, Qirun, et al.
Published: (2025)
by: Zeng, Qirun, et al.
Published: (2025)
PreND: Enhancing Intrinsic Motivation in Reinforcement Learning through Pre-trained Network Distillation
by: Davoodabadi, Mohammadamin, et al.
Published: (2024)
by: Davoodabadi, Mohammadamin, et al.
Published: (2024)
Feature-to-Image Data Augmentation: Improving Model Feature Extraction with Cluster-Guided Synthetic Samples
by: Haghbin, Yasaman, et al.
Published: (2024)
by: Haghbin, Yasaman, et al.
Published: (2024)
Latent Adversarial Regularization for Offline Preference Optimization
by: Jiang, Enyi, et al.
Published: (2026)
by: Jiang, Enyi, et al.
Published: (2026)
Language Plays a Pivotal Role in the Object-Attribute Compositional Generalization of CLIP
by: Abbasi, Reza, et al.
Published: (2024)
by: Abbasi, Reza, et al.
Published: (2024)
Adversarial Attacks to Latent Representations of Distributed Neural Networks in Split Computing
by: Zhang, Milin, et al.
Published: (2023)
by: Zhang, Milin, et al.
Published: (2023)
T2I-FineEval: Fine-Grained Compositional Metric for Text-to-Image Evaluation
by: Hosseini, Seyed Mohammad Hadi, et al.
Published: (2025)
by: Hosseini, Seyed Mohammad Hadi, et al.
Published: (2025)
Decompose-and-Compose: A Compositional Approach to Mitigating Spurious Correlation
by: Noohdani, Fahimeh Hosseini, et al.
Published: (2024)
by: Noohdani, Fahimeh Hosseini, et al.
Published: (2024)
Classification of Breast Cancer Histopathology Images using a Modified Supervised Contrastive Learning Method
by: Sani, Matina Mahdizadeh, et al.
Published: (2024)
by: Sani, Matina Mahdizadeh, et al.
Published: (2024)
Model-Based Offline Reinforcement Learning with Adversarial Data Augmentation
by: Cao, Hongye, et al.
Published: (2025)
by: Cao, Hongye, et al.
Published: (2025)
Adversarial Policy Optimization for Offline Preference-based Reinforcement Learning
by: Kang, Hyungkyu, et al.
Published: (2025)
by: Kang, Hyungkyu, et al.
Published: (2025)
Adversarial Attacks on Hyperbolic Networks
by: van Spengler, Max, et al.
Published: (2024)
by: van Spengler, Max, et al.
Published: (2024)
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
by: Ji, Kaixuan, et al.
Published: (2026)
by: Ji, Kaixuan, et al.
Published: (2026)
LLM-Agent-Controller: A Universal Multi-Agent Large Language Model System as a Control Engineer
by: Zahedifar, Rasoul, et al.
Published: (2025)
by: Zahedifar, Rasoul, et al.
Published: (2025)
Trust, Don't Trust, or Flip: Robust Preference-Based Reinforcement Learning with Multi-Expert Feedback
by: Hosseini, Seyed Amir, et al.
Published: (2026)
by: Hosseini, Seyed Amir, et al.
Published: (2026)
Similar Items
-
SUSD: Structured Unsupervised Skill Discovery through State Factorization
by: Hosseini, Seyed Mohammad Hadi, et al.
Published: (2026) -
CER: Confidence Enhanced Reasoning in LLMs
by: Razghandi, Ali, et al.
Published: (2025) -
LibraGrad: Balancing Gradient Flow for Universally Better Vision Transformer Attributions
by: Mehri, Faridoun, et al.
Published: (2024) -
Dilated Balanced Cross Entropy Loss for Medical Image Segmentation
by: Hosseini, Seyed Mohsen, et al.
Published: (2024) -
Inductive Biases for Zero-shot Systematic Generalization in Language-informed Reinforcement Learning
by: Dijujin, Negin Hashemi, et al.
Published: (2025)