NDCG-Consistent Softmax Approximation with Accelerated Convergence
Fuente:
arXiv
Saved in:
| Main Authors: | Pu, Yuanhao, Lian, Defu, Chen, Xiaolong, Huang, Xu, Chen, Jin, Chen, Enhong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Is Softmax Loss All You Need? A Principled Analysis of Softmax-family Loss
by: Pu, Yuanhao, et al.
Published: (2026)
by: Pu, Yuanhao, et al.
Published: (2026)
Beyond Surrogates: A Quantitative Analysis for Inter-Metric Relationships
by: Pu, Yuanhao, et al.
Published: (2026)
by: Pu, Yuanhao, et al.
Published: (2026)
Adaptive Sampled Softmax with Inverted Multi-Index: Methods, Theory and Applications
by: Chen, Jin, et al.
Published: (2025)
by: Chen, Jin, et al.
Published: (2025)
Efficient Machine Unlearning via Influence Approximation
by: Liu, Jiawei, et al.
Published: (2025)
by: Liu, Jiawei, et al.
Published: (2025)
Provably Convergent Subgraph-wise Sampling for Fast GNN Training
by: Wang, Jie, et al.
Published: (2023)
by: Wang, Jie, et al.
Published: (2023)
Understanding the planning of LLM agents: A survey
by: Huang, Xu, et al.
Published: (2024)
by: Huang, Xu, et al.
Published: (2024)
Breaking Determinism: Fuzzy Modeling of Sequential Recommendation Using Discrete State Space Diffusion Model
by: Xie, Wenjia, et al.
Published: (2024)
by: Xie, Wenjia, et al.
Published: (2024)
Universal Approximation with Softmax Attention
by: Hu, Jerry Yao-Chieh, et al.
Published: (2025)
by: Hu, Jerry Yao-Chieh, et al.
Published: (2025)
Rethinking the Global Convergence of Softmax Policy Gradient with Linear Function Approximation
by: Lin, Max Qiushi, et al.
Published: (2025)
by: Lin, Max Qiushi, et al.
Published: (2025)
PiXTime: A Model for Federated Time Series Forecasting with Heterogeneous Data across Nodes
by: Zhou, Yiming, et al.
Published: (2026)
by: Zhou, Yiming, et al.
Published: (2026)
MDAP: A Multi-view Disentangled and Adaptive Preference Learning Framework for Cross-Domain Recommendation
by: Tong, Junxiong, et al.
Published: (2024)
by: Tong, Junxiong, et al.
Published: (2024)
Model Stealing Attack against Graph Classification with Authenticity, Uncertainty and Diversity
by: Zhu, Zhihao, et al.
Published: (2023)
by: Zhu, Zhihao, et al.
Published: (2023)
Convergence Rates for Softmax Gating Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2025)
by: Nguyen, Huy, et al.
Published: (2025)
Federated Contextual Cascading Bandits with Asynchronous Communication and Heterogeneous Users
by: Yang, Hantao, et al.
Published: (2024)
by: Yang, Hantao, et al.
Published: (2024)
Exploring User Retrieval Integration towards Large Language Models for Cross-Domain Sequential Recommendation
by: Shen, Tingjia, et al.
Published: (2024)
by: Shen, Tingjia, et al.
Published: (2024)
Demystifying Design Choices of Reinforcement Fine-tuning: A Batched Contextual Bandit Learning Perspective
by: Xie, Hong, et al.
Published: (2026)
by: Xie, Hong, et al.
Published: (2026)
GTM: A General Time-series Model for Enhanced Representation Learning of Time-Series Data
by: He, Cheng, et al.
Published: (2025)
by: He, Cheng, et al.
Published: (2025)
A Unified Frequency Domain Decomposition Framework for Interpretable and Robust Time Series Forecasting
by: He, Cheng, et al.
Published: (2025)
by: He, Cheng, et al.
Published: (2025)
Composable Score-based Graph Diffusion Model for Multi-Conditional Molecular Generation
by: Qiao, Anjie, et al.
Published: (2025)
by: Qiao, Anjie, et al.
Published: (2025)
Fast Convergence of Softmax Policy Mirror Ascent
by: Asad, Reza, et al.
Published: (2024)
by: Asad, Reza, et al.
Published: (2024)
Foundations and Frontiers of Graph Learning Theory
by: Huang, Yu, et al.
Published: (2024)
by: Huang, Yu, et al.
Published: (2024)
CastFlow: Learning Role-Specialized Agentic Workflows for Time Series Forecasting
by: Pan, Bokai, et al.
Published: (2026)
by: Pan, Bokai, et al.
Published: (2026)
Training Dynamics of Multi-Head Softmax Attention for In-Context Learning: Emergence, Convergence, and Optimality
by: Chen, Siyu, et al.
Published: (2024)
by: Chen, Siyu, et al.
Published: (2024)
Adam-HNAG: A Convergent Reformulation of Adam with Accelerated Rate
by: Yu, Yaxin, et al.
Published: (2026)
by: Yu, Yaxin, et al.
Published: (2026)
Approximate Global Convergence of Independent Learning in Multi-Agent Systems
by: Jin, Ruiyang, et al.
Published: (2024)
by: Jin, Ruiyang, et al.
Published: (2024)
Evaluation of Neural Networks Defenses and Attacks using NDCG and Reciprocal Rank Metrics
by: Brama, Haya, et al.
Published: (2022)
by: Brama, Haya, et al.
Published: (2022)
MetaLA: Unified Optimal Linear Approximation to Softmax Attention Map
by: Chou, Yuhong, et al.
Published: (2024)
by: Chou, Yuhong, et al.
Published: (2024)
On Global Convergence Rates for Federated Softmax Policy Gradient under Heterogeneous Environments
by: Labbi, Safwan, et al.
Published: (2025)
by: Labbi, Safwan, et al.
Published: (2025)
Training Dynamics of Softmax Self-Attention: Fast Global Convergence via Preconditioning
by: Goel, Gautam, et al.
Published: (2026)
by: Goel, Gautam, et al.
Published: (2026)
Convergence Rate in Nonlinear Two-Time-Scale Stochastic Approximation with State (Time)-Dependence
by: Chen, Zixi, et al.
Published: (2025)
by: Chen, Zixi, et al.
Published: (2025)
Accelerated Policy Gradient: On the Convergence Rates of the Nesterov Momentum for Reinforcement Learning
by: Chen, Yen-Ju, et al.
Published: (2023)
by: Chen, Yen-Ju, et al.
Published: (2023)
Adaptive Sparse Softmax: An Effective and Efficient Softmax Variant
by: Lv, Qi, et al.
Published: (2025)
by: Lv, Qi, et al.
Published: (2025)
Beyond Stationarity: Convergence Analysis of Stochastic Softmax Policy Gradient Methods
by: Klein, Sara, et al.
Published: (2023)
by: Klein, Sara, et al.
Published: (2023)
Learning Theory of Transformers: Local-to-Global Approximation via Softmax Partition of Unity
by: Shi, Zhongjie, et al.
Published: (2026)
by: Shi, Zhongjie, et al.
Published: (2026)
Entropy Law: The Story Behind Data Compression and LLM Performance
by: Yin, Mingjia, et al.
Published: (2024)
by: Yin, Mingjia, et al.
Published: (2024)
Convergence Of Consistency Model With Multistep Sampling Under General Data Assumptions
by: Chen, Yiding, et al.
Published: (2025)
by: Chen, Yiding, et al.
Published: (2025)
ITA: An Energy-Efficient Attention and Softmax Accelerator for Quantized Transformers
by: İslamoğlu, Gamze, et al.
Published: (2023)
by: İslamoğlu, Gamze, et al.
Published: (2023)
Beyond State Consistency: Behavior Consistency in Text-Based World Models
by: Huang, Youling, et al.
Published: (2026)
by: Huang, Youling, et al.
Published: (2026)
ReNF: Rethinking the Design of Neural Long-Term Time Series Forecasters
by: Lu, Yihang, et al.
Published: (2025)
by: Lu, Yihang, et al.
Published: (2025)
When Large Language Models Meet Personalization: Perspectives of Challenges and Opportunities
by: Chen, Jin, et al.
Published: (2023)
by: Chen, Jin, et al.
Published: (2023)
Similar Items
-
Is Softmax Loss All You Need? A Principled Analysis of Softmax-family Loss
by: Pu, Yuanhao, et al.
Published: (2026) -
Beyond Surrogates: A Quantitative Analysis for Inter-Metric Relationships
by: Pu, Yuanhao, et al.
Published: (2026) -
Adaptive Sampled Softmax with Inverted Multi-Index: Methods, Theory and Applications
by: Chen, Jin, et al.
Published: (2025) -
Efficient Machine Unlearning via Influence Approximation
by: Liu, Jiawei, et al.
Published: (2025) -
Provably Convergent Subgraph-wise Sampling for Fast GNN Training
by: Wang, Jie, et al.
Published: (2023)