Online Continual Learning via Logit Adjusted Softmax
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Zhehao, Li, Tao, Yuan, Chenhe, Wu, Yingwen, Huang, Xiaolin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Compensation-free Machine Unlearning in Text-to-Image Diffusion Models by Eliminating the Mutual Information
by: Cheng, Xinwen, et al.
Published: (2026)
by: Cheng, Xinwen, et al.
Published: (2026)
Trainable Weight Averaging: Accelerating Training and Improving Generalization
by: Li, Tao, et al.
Published: (2022)
by: Li, Tao, et al.
Published: (2022)
Remaining-data-free Machine Unlearning by Suppressing Sample Contribution
by: Cheng, Xinwen, et al.
Published: (2024)
by: Cheng, Xinwen, et al.
Published: (2024)
Logit Dynamics in Softmax Policy Gradient Methods
by: Li, Yingru
Published: (2025)
by: Li, Yingru
Published: (2025)
A Unified Gradient-based Framework for Task-agnostic Continual Learning-Unlearning
by: Huang, Zhehao, et al.
Published: (2025)
by: Huang, Zhehao, et al.
Published: (2025)
Towards Natural Machine Unlearning
by: He, Zhengbao, et al.
Published: (2024)
by: He, Zhengbao, et al.
Published: (2024)
Bi-LoRA: Efficient Sharpness-Aware Minimization for Fine-Tuning Large-Scale Models
by: Liu, Yuhang, et al.
Published: (2025)
by: Liu, Yuhang, et al.
Published: (2025)
Multi-head Ensemble of Smoothed Classifiers for Certified Robustness
by: Fang, Kun, et al.
Published: (2022)
by: Fang, Kun, et al.
Published: (2022)
Towards Robust Neural Networks via Orthogonal Diversity
by: Fang, Kun, et al.
Published: (2020)
by: Fang, Kun, et al.
Published: (2020)
Revisiting Random Weight Perturbation for Efficiently Improving Generalization
by: Li, Tao, et al.
Published: (2024)
by: Li, Tao, et al.
Published: (2024)
AnomalyAID: Reliable Interpretation for Semi-supervised Network Anomaly Detection
by: Yuan, Yachao, et al.
Published: (2024)
by: Yuan, Yachao, et al.
Published: (2024)
Pursuing Feature Separation based on Neural Collapse for Out-of-Distribution Detection
by: Wu, Yingwen, et al.
Published: (2024)
by: Wu, Yingwen, et al.
Published: (2024)
T2I-ConBench: Text-to-Image Benchmark for Continual Post-training
by: Huang, Zhehao, et al.
Published: (2025)
by: Huang, Zhehao, et al.
Published: (2025)
Unified Gradient-Based Machine Unlearning with Remain Geometry Enhancement
by: Huang, Zhehao, et al.
Published: (2024)
by: Huang, Zhehao, et al.
Published: (2024)
BeeTLe: An Imbalance-Aware Deep Sequence Model for Linear B-Cell Epitope Prediction and Classification with Logit-Adjusted Losses
by: Yuan, Xiao
Published: (2023)
by: Yuan, Xiao
Published: (2023)
RAIN-Merging: A Gradient-Free Method to Enhance Instruction Following in Large Reasoning Models with Preserved Thinking Format
by: Huang, Zhehao, et al.
Published: (2026)
by: Huang, Zhehao, et al.
Published: (2026)
Adaptive Sparse Softmax: An Effective and Efficient Softmax Variant
by: Lv, Qi, et al.
Published: (2025)
by: Lv, Qi, et al.
Published: (2025)
Divide, Reweight, and Conquer: A Logit Arithmetic Approach for In-Context Learning
by: Huang, Chengsong, et al.
Published: (2024)
by: Huang, Chengsong, et al.
Published: (2024)
A Tractable Online Learning Algorithm for the Multinomial Logit Contextual Bandit
by: Agrawal, Priyank, et al.
Published: (2020)
by: Agrawal, Priyank, et al.
Published: (2020)
VL-RouterBench: A Benchmark for Vision-Language Model Routing
by: Huang, Zhehao, et al.
Published: (2025)
by: Huang, Zhehao, et al.
Published: (2025)
NDCG-Consistent Softmax Approximation with Accelerated Convergence
by: Pu, Yuanhao, et al.
Published: (2025)
by: Pu, Yuanhao, et al.
Published: (2025)
Vertex-Softmax: Tight Transformer Verification via Exact Softmax Optimization
by: Rezazadeh, Navid, et al.
Published: (2026)
by: Rezazadeh, Navid, et al.
Published: (2026)
Logits-Based Finetuning
by: Li, Jingyao, et al.
Published: (2025)
by: Li, Jingyao, et al.
Published: (2025)
Top-$nσ$: Not All Logits Are You Need
by: Tang, Chenxia, et al.
Published: (2024)
by: Tang, Chenxia, et al.
Published: (2024)
Learning Scalable Model Soup on a Single GPU: An Efficient Subspace Training Strategy
by: Li, Tao, et al.
Published: (2024)
by: Li, Tao, et al.
Published: (2024)
Stabilizing Policy Optimization via Logits Convexity
by: Chen, Hongzhan, et al.
Published: (2026)
by: Chen, Hongzhan, et al.
Published: (2026)
Solving Continual Offline Reinforcement Learning with Decision Transformer
by: Huang, Kaixin, et al.
Published: (2024)
by: Huang, Kaixin, et al.
Published: (2024)
Online Continual Learning with Dynamic Label Hierarchies
by: Wang, Xinrui, et al.
Published: (2026)
by: Wang, Xinrui, et al.
Published: (2026)
End-to-end Kernel Learning via Generative Random Fourier Features
by: Fang, Kun, et al.
Published: (2020)
by: Fang, Kun, et al.
Published: (2020)
Enhancing Certified Robustness via Block Reflector Orthogonal Layers and Logit Annealing Loss
by: Lai, Bo-Han, et al.
Published: (2025)
by: Lai, Bo-Han, et al.
Published: (2025)
SCALA: Split Federated Learning with Concatenated Activations and Logit Adjustments
by: Yang, Jiarong, et al.
Published: (2024)
by: Yang, Jiarong, et al.
Published: (2024)
Dimension Reduction with Locally Adjusted Graphs
by: Wang, Yingfan, et al.
Published: (2024)
by: Wang, Yingfan, et al.
Published: (2024)
Learning Theory of Transformers: Local-to-Global Approximation via Softmax Partition of Unity
by: Shi, Zhongjie, et al.
Published: (2026)
by: Shi, Zhongjie, et al.
Published: (2026)
Revisiting Softmax Masking: Stop Gradient for Enhancing Stability in Replay-based Continual Learning
by: Kim, Hoyong, et al.
Published: (2023)
by: Kim, Hoyong, et al.
Published: (2023)
Consistency Regularization for Domain Generalization with Logit Attribution Matching
by: Gao, Han, et al.
Published: (2023)
by: Gao, Han, et al.
Published: (2023)
Learning with Selectively Labeled Data from Multiple Decision-makers
by: Chen, Jian, et al.
Published: (2023)
by: Chen, Jian, et al.
Published: (2023)
Logit Distillation on Manifolds: Mapping by Learning
by: Yang, Yiru, et al.
Published: (2026)
by: Yang, Yiru, et al.
Published: (2026)
ALSA: Anchors in Logit Space for Out-of-Distribution Accuracy Estimation
by: Liu, Chenzhi, et al.
Published: (2025)
by: Liu, Chenzhi, et al.
Published: (2025)
Softmax-free Linear Transformers
by: Lu, Jiachen, et al.
Published: (2022)
by: Lu, Jiachen, et al.
Published: (2022)
To Softmax, or not to Softmax: that is the question when applying Active Learning for Transformer Models
by: Gonsior, Julius, et al.
Published: (2022)
by: Gonsior, Julius, et al.
Published: (2022)
Similar Items
-
Compensation-free Machine Unlearning in Text-to-Image Diffusion Models by Eliminating the Mutual Information
by: Cheng, Xinwen, et al.
Published: (2026) -
Trainable Weight Averaging: Accelerating Training and Improving Generalization
by: Li, Tao, et al.
Published: (2022) -
Remaining-data-free Machine Unlearning by Suppressing Sample Contribution
by: Cheng, Xinwen, et al.
Published: (2024) -
Logit Dynamics in Softmax Policy Gradient Methods
by: Li, Yingru
Published: (2025) -
A Unified Gradient-based Framework for Task-agnostic Continual Learning-Unlearning
by: Huang, Zhehao, et al.
Published: (2025)