Saved in:
| Main Authors: | Malloy, Tailia, Seow, Roderick, Gonzalez, Cleotilde |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.11161 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Defend by Attacking (and Vice-Versa): Transfer of Learning in Cybersecurity Games
by: Malloy, Tailia, et al.
Published: (2023)
by: Malloy, Tailia, et al.
Published: (2023)
Improving the Prediction of Individual Engagement in Recommendations Using Cognitive Models
by: Seow, Roderick, et al.
Published: (2024)
by: Seow, Roderick, et al.
Published: (2024)
Leveraging a Cognitive Model to Measure Subjective Similarity of Human and GPT-4 Written Content
by: Malloy, Tailia, et al.
Published: (2024)
by: Malloy, Tailia, et al.
Published: (2024)
Training Users Against Human and GPT-4 Generated Social Engineering Attacks
by: Malloy, Tailia, et al.
Published: (2025)
by: Malloy, Tailia, et al.
Published: (2025)
Deep RL With Information Constrained Policies: Generalization in Continuous Control
by: Malloy, Tailia, et al.
Published: (2020)
by: Malloy, Tailia, et al.
Published: (2020)
Assessing Spear-Phishing Website Generation in Large Language Model Coding Agents
by: Malloy, Tailia, et al.
Published: (2026)
by: Malloy, Tailia, et al.
Published: (2026)
Reasoning Elicitation in Language Models via Counterfactual Feedback
by: Hüyük, Alihan, et al.
Published: (2024)
by: Hüyük, Alihan, et al.
Published: (2024)
Improving Sequential Query Recommendation with Immediate User Feedback
by: Parambath, Shameem A Puthiya, et al.
Published: (2022)
by: Parambath, Shameem A Puthiya, et al.
Published: (2022)
Budgeted Recommendation with Delayed Feedback
by: Liu, Kweiguu, et al.
Published: (2024)
by: Liu, Kweiguu, et al.
Published: (2024)
Unsupervised Structural-Counterfactual Generation under Domain Shift
by: Kher, Krishn Vishwas, et al.
Published: (2025)
by: Kher, Krishn Vishwas, et al.
Published: (2025)
Towards Neural Network based Cognitive Models of Dynamic Decision-Making by Humans
by: Chen, Changyu, et al.
Published: (2024)
by: Chen, Changyu, et al.
Published: (2024)
Delayed Feedback Modeling with Influence Functions
by: Ding, Chenlu, et al.
Published: (2025)
by: Ding, Chenlu, et al.
Published: (2025)
Lipschitz Bandits with Stochastic Delayed Feedback
by: Liu, Zhongxuan, et al.
Published: (2025)
by: Liu, Zhongxuan, et al.
Published: (2025)
Verified Training for Counterfactual Explanation Robustness under Data Shift
by: Meyer, Anna P., et al.
Published: (2024)
by: Meyer, Anna P., et al.
Published: (2024)
Cyclic Counterfactuals under Shift-Scale Interventions
by: Saha, Saptarshi, et al.
Published: (2025)
by: Saha, Saptarshi, et al.
Published: (2025)
Choice-Model-Assisted Q-learning for Delayed-Feedback Revenue Management
by: Shen, Owen, et al.
Published: (2026)
by: Shen, Owen, et al.
Published: (2026)
Bandit and Delayed Feedback in Online Structured Prediction
by: Shibukawa, Yuki, et al.
Published: (2025)
by: Shibukawa, Yuki, et al.
Published: (2025)
Biased Dueling Bandits with Stochastic Delayed Feedback
by: Yi, Bongsoo, et al.
Published: (2024)
by: Yi, Bongsoo, et al.
Published: (2024)
A Best-of-both-worlds Algorithm for Bandits with Delayed Feedback with Robustness to Excessive Delays
by: Masoudian, Saeed, et al.
Published: (2023)
by: Masoudian, Saeed, et al.
Published: (2023)
Differentiable Attenuation Filters for Feedback Delay Networks
by: Ibnyahya, Ilias, et al.
Published: (2025)
by: Ibnyahya, Ilias, et al.
Published: (2025)
iFlip: Iterative Feedback-driven Counterfactual Example Refinement
by: Wang, Yilong, et al.
Published: (2026)
by: Wang, Yilong, et al.
Published: (2026)
Exploiting Curvature in Online Convex Optimization with Delayed Feedback
by: Qiu, Hao, et al.
Published: (2025)
by: Qiu, Hao, et al.
Published: (2025)
Online Nonsubmodular Optimization with Delayed Feedback in the Bandit Setting
by: Yang, Sifan, et al.
Published: (2025)
by: Yang, Sifan, et al.
Published: (2025)
Neural Contextual Bandits Under Delayed Feedback Constraints
by: Moghimi, Mohammadali, et al.
Published: (2025)
by: Moghimi, Mohammadali, et al.
Published: (2025)
Improved Regret for Bandit Convex Optimization with Delayed Feedback
by: Wan, Yuanyu, et al.
Published: (2024)
by: Wan, Yuanyu, et al.
Published: (2024)
Linear and Neural Dueling Bandits with Delayed Feedback
by: Wang, Xiangyi, et al.
Published: (2026)
by: Wang, Xiangyi, et al.
Published: (2026)
Improved Best-of-Both-Worlds Regret for Bandits with Delayed Feedback
by: Schlisselberg, Ofir, et al.
Published: (2025)
by: Schlisselberg, Ofir, et al.
Published: (2025)
Adversarial Bandits with Multi-User Delayed Feedback: Theory and Application
by: Li, Yandi, et al.
Published: (2023)
by: Li, Yandi, et al.
Published: (2023)
Introducing User Feedback-based Counterfactual Explanations (UFCE)
by: Suffian, Muhammad, et al.
Published: (2024)
by: Suffian, Muhammad, et al.
Published: (2024)
Data-Driven Room Acoustic Modeling Via Differentiable Feedback Delay Networks With Learnable Delay Lines
by: Mezza, Alessandro Ilic, et al.
Published: (2024)
by: Mezza, Alessandro Ilic, et al.
Published: (2024)
Learning Counterfactually Decoupled Attention for Open-World Model Attribution
by: Zheng, Yu, et al.
Published: (2025)
by: Zheng, Yu, et al.
Published: (2025)
Delayed Feedback Modeling for Post-Click Gross Merchandise Volume Prediction: Benchmark, Insights and Approaches
by: Li, Xinyu, et al.
Published: (2026)
by: Li, Xinyu, et al.
Published: (2026)
Modeling Cascaded Delay Feedback for Online Net Conversion Rate Prediction: Benchmark, Insights and Solutions
by: Luo, Mingxuan, et al.
Published: (2026)
by: Luo, Mingxuan, et al.
Published: (2026)
Decentralized Online Convex Optimization with Unknown Feedback Delays
by: Qiu, Hao, et al.
Published: (2026)
by: Qiu, Hao, et al.
Published: (2026)
CoFi-PGMA: Counterfactual Policy Gradients under Filtered Feedback for Multi-Agent LLMs
by: Tong, Stela, et al.
Published: (2026)
by: Tong, Stela, et al.
Published: (2026)
A Reduction-based Framework for Sequential Decision Making with Delayed Feedback
by: Yang, Yunchang, et al.
Published: (2023)
by: Yang, Yunchang, et al.
Published: (2023)
Merit-based Fair Combinatorial Semi-Bandit with Unrestricted Feedback Delays
by: Chen, Ziqun, et al.
Published: (2024)
by: Chen, Ziqun, et al.
Published: (2024)
Boosting Reinforcement Learning with Strongly Delayed Feedback Through Auxiliary Short Delays
by: Wu, Qingyuan, et al.
Published: (2024)
by: Wu, Qingyuan, et al.
Published: (2024)
Feed-Forward Optimization With Delayed Feedback for Neural Network Training
by: Flügel, Katharina, et al.
Published: (2023)
by: Flügel, Katharina, et al.
Published: (2023)
Regret Bounds for Adversarial Contextual Bandits with General Function Approximation and Delayed Feedback
by: Levy, Orin, et al.
Published: (2025)
by: Levy, Orin, et al.
Published: (2025)
Similar Items
-
Learning to Defend by Attacking (and Vice-Versa): Transfer of Learning in Cybersecurity Games
by: Malloy, Tailia, et al.
Published: (2023) -
Improving the Prediction of Individual Engagement in Recommendations Using Cognitive Models
by: Seow, Roderick, et al.
Published: (2024) -
Leveraging a Cognitive Model to Measure Subjective Similarity of Human and GPT-4 Written Content
by: Malloy, Tailia, et al.
Published: (2024) -
Training Users Against Human and GPT-4 Generated Social Engineering Attacks
by: Malloy, Tailia, et al.
Published: (2025) -
Deep RL With Information Constrained Policies: Generalization in Continuous Control
by: Malloy, Tailia, et al.
Published: (2020)