Adaptive Sampled Softmax with Inverted Multi-Index: Methods, Theory and Applications
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Jin, Zhang, Jin, huang, Xu, Yang, Yi, Lian, Defu, Chen, Enhong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Is Softmax Loss All You Need? A Principled Analysis of Softmax-family Loss
by: Pu, Yuanhao, et al.
Published: (2026)
by: Pu, Yuanhao, et al.
Published: (2026)
NDCG-Consistent Softmax Approximation with Accelerated Convergence
by: Pu, Yuanhao, et al.
Published: (2025)
by: Pu, Yuanhao, et al.
Published: (2025)
Beyond Surrogates: A Quantitative Analysis for Inter-Metric Relationships
by: Pu, Yuanhao, et al.
Published: (2026)
by: Pu, Yuanhao, et al.
Published: (2026)
MDAP: A Multi-view Disentangled and Adaptive Preference Learning Framework for Cross-Domain Recommendation
by: Tong, Junxiong, et al.
Published: (2024)
by: Tong, Junxiong, et al.
Published: (2024)
Efficient Machine Unlearning via Influence Approximation
by: Liu, Jiawei, et al.
Published: (2025)
by: Liu, Jiawei, et al.
Published: (2025)
Provably Convergent Subgraph-wise Sampling for Fast GNN Training
by: Wang, Jie, et al.
Published: (2023)
by: Wang, Jie, et al.
Published: (2023)
Breaking Determinism: Fuzzy Modeling of Sequential Recommendation Using Discrete State Space Diffusion Model
by: Xie, Wenjia, et al.
Published: (2024)
by: Xie, Wenjia, et al.
Published: (2024)
Model Stealing Attack against Graph Classification with Authenticity, Uncertainty and Diversity
by: Zhu, Zhihao, et al.
Published: (2023)
by: Zhu, Zhihao, et al.
Published: (2023)
Foundations and Frontiers of Graph Learning Theory
by: Huang, Yu, et al.
Published: (2024)
by: Huang, Yu, et al.
Published: (2024)
PiXTime: A Model for Federated Time Series Forecasting with Heterogeneous Data across Nodes
by: Zhou, Yiming, et al.
Published: (2026)
by: Zhou, Yiming, et al.
Published: (2026)
Adaptive Sparse Softmax: An Effective and Efficient Softmax Variant
by: Lv, Qi, et al.
Published: (2025)
by: Lv, Qi, et al.
Published: (2025)
Adaptive Inverted-Index Routing for Granular Mixtures-of-Experts
by: Kladny, Klaus-Rudolf, et al.
Published: (2026)
by: Kladny, Klaus-Rudolf, et al.
Published: (2026)
TDDBench: A Benchmark for Training data detection
by: Zhu, Zhihao, et al.
Published: (2024)
by: Zhu, Zhihao, et al.
Published: (2024)
Federated Contextual Cascading Bandits with Asynchronous Communication and Heterogeneous Users
by: Yang, Hantao, et al.
Published: (2024)
by: Yang, Hantao, et al.
Published: (2024)
Composable Score-based Graph Diffusion Model for Multi-Conditional Molecular Generation
by: Qiao, Anjie, et al.
Published: (2025)
by: Qiao, Anjie, et al.
Published: (2025)
Learning Deep Tree-based Retriever for Efficient Recommendation: Theory and Method
by: Liu, Ze, et al.
Published: (2024)
by: Liu, Ze, et al.
Published: (2024)
Exploring User Retrieval Integration towards Large Language Models for Cross-Domain Sequential Recommendation
by: Shen, Tingjia, et al.
Published: (2024)
by: Shen, Tingjia, et al.
Published: (2024)
Demystifying Design Choices of Reinforcement Fine-tuning: A Batched Contextual Bandit Learning Perspective
by: Xie, Hong, et al.
Published: (2026)
by: Xie, Hong, et al.
Published: (2026)
Understanding the planning of LLM agents: A survey
by: Huang, Xu, et al.
Published: (2024)
by: Huang, Xu, et al.
Published: (2024)
In-Context Linear Regression Demystified: Training Dynamics and Mechanistic Interpretability of Multi-Head Softmax Attention
by: He, Jianliang, et al.
Published: (2025)
by: He, Jianliang, et al.
Published: (2025)
Data Augmentation in Time Series Forecasting through Inverted Framework
by: Tan, Hongming, et al.
Published: (2025)
by: Tan, Hongming, et al.
Published: (2025)
A Latent Variable Approach for Non-Hierarchical Multi-Fidelity Adaptive Sampling
by: Chen, Yi-Ping, et al.
Published: (2023)
by: Chen, Yi-Ping, et al.
Published: (2023)
Is Temperature Sample Efficient for Softmax Gaussian Mixture of Experts?
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Softmax is not Enough (for Adaptive Conformal Classification)
by: Attar, Navid Akhavan, et al.
Published: (2026)
by: Attar, Navid Akhavan, et al.
Published: (2026)
M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation
by: Chen, Jianlv, et al.
Published: (2024)
by: Chen, Jianlv, et al.
Published: (2024)
Multi-Label Adaptive Batch Selection by Highlighting Hard and Imbalanced Samples
by: Zhou, Ao, et al.
Published: (2024)
by: Zhou, Ao, et al.
Published: (2024)
GTM: A General Time-series Model for Enhanced Representation Learning of Time-Series Data
by: He, Cheng, et al.
Published: (2025)
by: He, Cheng, et al.
Published: (2025)
A Unified Frequency Domain Decomposition Framework for Interpretable and Robust Time Series Forecasting
by: He, Cheng, et al.
Published: (2025)
by: He, Cheng, et al.
Published: (2025)
CastFlow: Learning Role-Specialized Agentic Workflows for Time Series Forecasting
by: Pan, Bokai, et al.
Published: (2026)
by: Pan, Bokai, et al.
Published: (2026)
Scaling Federated Linear Contextual Bandits via Sketching
by: Yang, Hantao, et al.
Published: (2026)
by: Yang, Hantao, et al.
Published: (2026)
Dual Test-time Training for Out-of-distribution Recommender System
by: Yang, Xihong, et al.
Published: (2024)
by: Yang, Xihong, et al.
Published: (2024)
Logit Dynamics in Softmax Policy Gradient Methods
by: Li, Yingru
Published: (2025)
by: Li, Yingru
Published: (2025)
A General Theory for Softmax Gating Multinomial Logistic Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2023)
by: Nguyen, Huy, et al.
Published: (2023)
Cross-Domain Pre-training with Language Models for Transferable Time Series Representations
by: Cheng, Mingyue, et al.
Published: (2024)
by: Cheng, Mingyue, et al.
Published: (2024)
Exploring the Frontiers of Softmax: Provable Optimization, Applications in Diffusion Model, and Beyond
by: Cao, Yang, et al.
Published: (2024)
by: Cao, Yang, et al.
Published: (2024)
Universal Approximation with Softmax Attention
by: Hu, Jerry Yao-Chieh, et al.
Published: (2025)
by: Hu, Jerry Yao-Chieh, et al.
Published: (2025)
Sigmoid Gating is More Sample Efficient than Softmax Gating in Mixture of Experts
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Entropy Law: The Story Behind Data Compression and LLM Performance
by: Yin, Mingjia, et al.
Published: (2024)
by: Yin, Mingjia, et al.
Published: (2024)
Learning Theory of Transformers: Local-to-Global Approximation via Softmax Partition of Unity
by: Shi, Zhongjie, et al.
Published: (2026)
by: Shi, Zhongjie, et al.
Published: (2026)
Inverting the Leverage Score Gradient: An Efficient Approximate Newton Method
by: Li, Chenyang, et al.
Published: (2024)
by: Li, Chenyang, et al.
Published: (2024)
Similar Items
-
Is Softmax Loss All You Need? A Principled Analysis of Softmax-family Loss
by: Pu, Yuanhao, et al.
Published: (2026) -
NDCG-Consistent Softmax Approximation with Accelerated Convergence
by: Pu, Yuanhao, et al.
Published: (2025) -
Beyond Surrogates: A Quantitative Analysis for Inter-Metric Relationships
by: Pu, Yuanhao, et al.
Published: (2026) -
MDAP: A Multi-view Disentangled and Adaptive Preference Learning Framework for Cross-Domain Recommendation
by: Tong, Junxiong, et al.
Published: (2024) -
Efficient Machine Unlearning via Influence Approximation
by: Liu, Jiawei, et al.
Published: (2025)