Learning in Position-Aware Multinomial Logit Bandits: From Multiplicative to General Position Effects
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Xi, Dai, Shibo, Lyu, Jiameng, Zhou, Yuan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Contextual Multinomial Logit Bandits with General Value Functions
by: Zhang, Mengxiao, et al.
Published: (2024)
by: Zhang, Mengxiao, et al.
Published: (2024)
A Tractable Online Learning Algorithm for the Multinomial Logit Contextual Bandit
by: Agrawal, Priyank, et al.
Published: (2020)
by: Agrawal, Priyank, et al.
Published: (2020)
Tractable Multinomial Logit Contextual Bandits with Non-Linear Utilities
by: Hwang, Taehyun, et al.
Published: (2026)
by: Hwang, Taehyun, et al.
Published: (2026)
Provably Efficient Reinforcement Learning with Multinomial Logit Function Approximation
by: Li, Long-Fei, et al.
Published: (2024)
by: Li, Long-Fei, et al.
Published: (2024)
Deconstructing Positional Information: From Attention Logits to Training Biases
by: Gu, Zihan, et al.
Published: (2025)
by: Gu, Zihan, et al.
Published: (2025)
Learning When to Restart: Nonstationary Newsvendor from Uncensored to Censored Demand
by: Chen, Xin, et al.
Published: (2025)
by: Chen, Xin, et al.
Published: (2025)
Variance-Adaptive Optimal Algorithm for Reinforcement Learning with Multinomial Logit Function Approximation
by: Kim, Wonyoung, et al.
Published: (2026)
by: Kim, Wonyoung, et al.
Published: (2026)
Learning Multinomial Logits in $O(n \log n)$ time
by: Chierichetti, Flavio, et al.
Published: (2026)
by: Chierichetti, Flavio, et al.
Published: (2026)
Optimal Design for Multinomial Logit Model with Applications to Best Assortment Identification
by: Lee, Joongkyu, et al.
Published: (2026)
by: Lee, Joongkyu, et al.
Published: (2026)
Achieving Limited Adaptivity for Multinomial Logistic Bandits
by: Midigeshi, Sukruta Prakash, et al.
Published: (2025)
by: Midigeshi, Sukruta Prakash, et al.
Published: (2025)
A Minibatch-SGD-Based Learning Meta-Policy for Inventory Systems with Myopic Optimal Policy
by: Lyu, Jiameng, et al.
Published: (2024)
by: Lyu, Jiameng, et al.
Published: (2024)
Improved Online Confidence Bounds for Multinomial Logistic Bandits
by: Lee, Joongkyu, et al.
Published: (2025)
by: Lee, Joongkyu, et al.
Published: (2025)
Nearly Minimax Optimal Regret for Multinomial Logistic Bandit
by: Lee, Joongkyu, et al.
Published: (2024)
by: Lee, Joongkyu, et al.
Published: (2024)
Closing the Gaps: Optimality of Sample Average Approximation for Data-Driven Newsvendor Problems
by: Lyu, Jiameng, et al.
Published: (2024)
by: Lyu, Jiameng, et al.
Published: (2024)
ELODI: Ensemble Logit Difference Inhibition for Positive-Congruent Training
by: Zhao, Yue, et al.
Published: (2022)
by: Zhao, Yue, et al.
Published: (2022)
Improved Regret Bounds of (Multinomial) Logistic Bandits via Regret-to-Confidence-Set Conversion
by: Lee, Junghyun, et al.
Published: (2023)
by: Lee, Junghyun, et al.
Published: (2023)
Multiplicative Logit Adjustment Approximates Neural-Collapse-Aware Decision Boundary Adjustment
by: Hasegawa, Naoya, et al.
Published: (2024)
by: Hasegawa, Naoya, et al.
Published: (2024)
Adaptive Budget Allocation in LLM-Augmented Surveys
by: Ye, Zikun, et al.
Published: (2026)
by: Ye, Zikun, et al.
Published: (2026)
Single Index Bandits: Generalized Linear Contextual Bandits with Unknown Reward Functions
by: Kang, Yue, et al.
Published: (2025)
by: Kang, Yue, et al.
Published: (2025)
Enjoying Non-linearity in Multinomial Logistic Bandits: A Minimax-Optimal Algorithm
by: Boudart, Pierre, et al.
Published: (2025)
by: Boudart, Pierre, et al.
Published: (2025)
RMLR: Extending Multinomial Logistic Regression into General Geometries
by: Chen, Ziheng, et al.
Published: (2024)
by: Chen, Ziheng, et al.
Published: (2024)
Online Continual Learning via Logit Adjusted Softmax
by: Huang, Zhehao, et al.
Published: (2023)
by: Huang, Zhehao, et al.
Published: (2023)
Rethinking Molecular OOD Generalization via Target-Aware Source Selection
by: Lin, Zhuohao, et al.
Published: (2026)
by: Lin, Zhuohao, et al.
Published: (2026)
Set2Seq Transformer: Temporal and Position-Aware Set Representations for Sequential Multiple-Instance Learning
by: Efthymiou, Athanasios, et al.
Published: (2024)
by: Efthymiou, Athanasios, et al.
Published: (2024)
Mutual Information Multinomial Estimation
by: Chen, Yanzhi, et al.
Published: (2024)
by: Chen, Yanzhi, et al.
Published: (2024)
Positional Knowledge is All You Need: Position-induced Transformer (PiT) for Operator Learning
by: Chen, Junfeng, et al.
Published: (2024)
by: Chen, Junfeng, et al.
Published: (2024)
SCALA: Split Federated Learning with Concatenated Activations and Logit Adjustments
by: Yang, Jiarong, et al.
Published: (2024)
by: Yang, Jiarong, et al.
Published: (2024)
FIRAL: An Active Learning Algorithm for Multinomial Logistic Regression
by: Chen, Youguang, et al.
Published: (2024)
by: Chen, Youguang, et al.
Published: (2024)
Position-Aware Parameter Efficient Fine-Tuning Approach for Reducing Positional Bias in LLMs
by: Zhang, Zheng, et al.
Published: (2024)
by: Zhang, Zheng, et al.
Published: (2024)
Learning High-Frequency Functions Made Easy with Sinusoidal Positional Encoding
by: Sun, Chuanhao, et al.
Published: (2024)
by: Sun, Chuanhao, et al.
Published: (2024)
Position-Sensing Graph Neural Networks: Proactively Learning Nodes Relative Positions
by: Qin, Zhenyue, et al.
Published: (2021)
by: Qin, Zhenyue, et al.
Published: (2021)
Learning Laplacian Positional Encodings for Heterophilous Graphs
by: Ito, Michael, et al.
Published: (2025)
by: Ito, Michael, et al.
Published: (2025)
Online Bandits with (Biased) Offline Data: Adaptive Learning under Distribution Mismatch
by: Cheung, Wang Chi, et al.
Published: (2024)
by: Cheung, Wang Chi, et al.
Published: (2024)
Model-Based Reinforcement Learning with Multinomial Logistic Function Approximation
by: Hwang, Taehyun, et al.
Published: (2022)
by: Hwang, Taehyun, et al.
Published: (2022)
Randomized Exploration for Reinforcement Learning with Multinomial Logistic Function Approximation
by: Cho, Wooseong, et al.
Published: (2024)
by: Cho, Wooseong, et al.
Published: (2024)
MEP: Multiple Kernel Learning Enhancing Relative Positional Encoding Length Extrapolation
by: Gao, Weiguo
Published: (2024)
by: Gao, Weiguo
Published: (2024)
Pessimistic Risk-Aware Policy Learning in Contextual Bandits
by: Wan, Yilong, et al.
Published: (2026)
by: Wan, Yilong, et al.
Published: (2026)
Consistency Regularization for Domain Generalization with Logit Attribution Matching
by: Gao, Han, et al.
Published: (2023)
by: Gao, Han, et al.
Published: (2023)
RoTHP: Rotary Position Embedding-based Transformer Hawkes Process
by: Gao, Anningzhe, et al.
Published: (2024)
by: Gao, Anningzhe, et al.
Published: (2024)
Negative as Positive: Enhancing Out-of-distribution Generalization for Graph Contrastive Learning
by: Wang, Zixu, et al.
Published: (2024)
by: Wang, Zixu, et al.
Published: (2024)
Similar Items
-
Contextual Multinomial Logit Bandits with General Value Functions
by: Zhang, Mengxiao, et al.
Published: (2024) -
A Tractable Online Learning Algorithm for the Multinomial Logit Contextual Bandit
by: Agrawal, Priyank, et al.
Published: (2020) -
Tractable Multinomial Logit Contextual Bandits with Non-Linear Utilities
by: Hwang, Taehyun, et al.
Published: (2026) -
Provably Efficient Reinforcement Learning with Multinomial Logit Function Approximation
by: Li, Long-Fei, et al.
Published: (2024) -
Deconstructing Positional Information: From Attention Logits to Training Biases
by: Gu, Zihan, et al.
Published: (2025)