Adaptive Regularization for Large-Scale Sparse Feature Embedding Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Mang, Lyu, Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Proximal Projection for Doubly Sparse Regularized Models
von: He, Jia Wei, et al.
Veröffentlicht: (2026)
von: He, Jia Wei, et al.
Veröffentlicht: (2026)
Tree-Regularized Tabular Embeddings
von: Li, Xuan, et al.
Veröffentlicht: (2024)
von: Li, Xuan, et al.
Veröffentlicht: (2024)
SAE-FD: Sparse Autoencoder Feature Distillation for Continual Learning of Large Language Models
von: Zhang, Mingxu, et al.
Veröffentlicht: (2026)
von: Zhang, Mingxu, et al.
Veröffentlicht: (2026)
Adaptive Block Sparse Regularization under Arbitrary Linear Transform
von: Furuhashi, Takanobu, et al.
Veröffentlicht: (2024)
von: Furuhashi, Takanobu, et al.
Veröffentlicht: (2024)
Adaptive Sparse Allocation with Mutual Choice & Feature Choice Sparse Autoencoders
von: Ayonrinde, Kola
Veröffentlicht: (2024)
von: Ayonrinde, Kola
Veröffentlicht: (2024)
Evaluating Model Robustness Using Adaptive Sparse L0 Regularization
von: Liu, Weiyou, et al.
Veröffentlicht: (2024)
von: Liu, Weiyou, et al.
Veröffentlicht: (2024)
Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models
von: Deng, Boyi, et al.
Veröffentlicht: (2026)
von: Deng, Boyi, et al.
Veröffentlicht: (2026)
Efficient Graph Encoder Embedding for Large Sparse Graphs in Python
von: Qin, Xihan, et al.
Veröffentlicht: (2024)
von: Qin, Xihan, et al.
Veröffentlicht: (2024)
Efficient Traffic Forecasting on Large-Scale Road Network by Regularized Adaptive Graph Convolution
von: Wu, Kaiqi, et al.
Veröffentlicht: (2025)
von: Wu, Kaiqi, et al.
Veröffentlicht: (2025)
Dual Feature Reduction for the Sparse-group Lasso and its Adaptive Variant
von: Feser, Fabio, et al.
Veröffentlicht: (2024)
von: Feser, Fabio, et al.
Veröffentlicht: (2024)
Embedded Federated Feature Selection with Dynamic Sparse Training: Balancing Accuracy-Cost Tradeoffs
von: Mahanipour, Afsaneh, et al.
Veröffentlicht: (2025)
von: Mahanipour, Afsaneh, et al.
Veröffentlicht: (2025)
Route Sparse Autoencoder to Interpret Large Language Models
von: Shi, Wei, et al.
Veröffentlicht: (2025)
von: Shi, Wei, et al.
Veröffentlicht: (2025)
On the SDEs and Scaling Rules for Adaptive Gradient Algorithms
von: Malladi, Sadhika, et al.
Veröffentlicht: (2022)
von: Malladi, Sadhika, et al.
Veröffentlicht: (2022)
Scaling Embeddings Outperforms Scaling Experts in Language Models
von: Liu, Hong, et al.
Veröffentlicht: (2026)
von: Liu, Hong, et al.
Veröffentlicht: (2026)
Enhanced High-Dimensional Data Visualization through Adaptive Multi-Scale Manifold Embedding
von: Ni, Tianhao, et al.
Veröffentlicht: (2025)
von: Ni, Tianhao, et al.
Veröffentlicht: (2025)
CAFE: Towards Compact, Adaptive, and Fast Embedding for Large-scale Recommendation Models
von: Zhang, Hailin, et al.
Veröffentlicht: (2023)
von: Zhang, Hailin, et al.
Veröffentlicht: (2023)
Dynamic Reserve Price Design with Distributed Solving Algorithm
von: Li, Mang
Veröffentlicht: (2022)
von: Li, Mang
Veröffentlicht: (2022)
Accelerating Large-Scale Regularized High-Order Tensor Recovery
von: Qin, Wenjin, et al.
Veröffentlicht: (2025)
von: Qin, Wenjin, et al.
Veröffentlicht: (2025)
Scaling Sparse Fine-Tuning to Large Language Models
von: Ansell, Alan, et al.
Veröffentlicht: (2024)
von: Ansell, Alan, et al.
Veröffentlicht: (2024)
Generalizing Scaling Laws for Dense and Sparse Large Language Models
von: Hossain, Md Arafat, et al.
Veröffentlicht: (2025)
von: Hossain, Md Arafat, et al.
Veröffentlicht: (2025)
Sparse Deep Learning Models with the $\ell_1$ Regularization
von: Shen, Lixin, et al.
Veröffentlicht: (2024)
von: Shen, Lixin, et al.
Veröffentlicht: (2024)
NuTime: Numerically Multi-Scaled Embedding for Large-Scale Time-Series Pretraining
von: Lin, Chenguo, et al.
Veröffentlicht: (2023)
von: Lin, Chenguo, et al.
Veröffentlicht: (2023)
Stable and Steerable Sparse Autoencoders with Weight Regularization
von: Jedryszek, Piotr, et al.
Veröffentlicht: (2026)
von: Jedryszek, Piotr, et al.
Veröffentlicht: (2026)
Multi-Probe Zero Collision Hash (MPZCH): Mitigating Embedding Collisions and Enhancing Model Freshness in Large-Scale Recommenders
von: Zhao, Ziliang, et al.
Veröffentlicht: (2026)
von: Zhao, Ziliang, et al.
Veröffentlicht: (2026)
Q-Sparse: All Large Language Models can be Fully Sparsely-Activated
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)
von: Wang, Hongyu, et al.
Veröffentlicht: (2024)
Row-wise Fusion Regularization: An Interpretable Personalized Federated Learning Framework in Large-Scale Scenarios
von: Zhou, Runlin, et al.
Veröffentlicht: (2025)
von: Zhou, Runlin, et al.
Veröffentlicht: (2025)
Disentangling Dense Embeddings with Sparse Autoencoders
von: O'Neill, Charles, et al.
Veröffentlicht: (2024)
von: O'Neill, Charles, et al.
Veröffentlicht: (2024)
Kernel Learning with Adversarial Features: Numerical Efficiency and Adaptive Regularization
von: Ribeiro, Antônio H., et al.
Veröffentlicht: (2025)
von: Ribeiro, Antônio H., et al.
Veröffentlicht: (2025)
Asymmetric Scaling Laws from Sparse Features
von: Sous, John, et al.
Veröffentlicht: (2026)
von: Sous, John, et al.
Veröffentlicht: (2026)
Accelerating Large-Scale Reasoning Model Inference with Sparse Self-Speculative Decoding
von: Zhao, Yilong, et al.
Veröffentlicht: (2025)
von: Zhao, Yilong, et al.
Veröffentlicht: (2025)
Supporting Evidence for the Adaptive Feature Program across Diverse Models
von: Li, Yicheng, et al.
Veröffentlicht: (2025)
von: Li, Yicheng, et al.
Veröffentlicht: (2025)
Time-Aware Feature Selection: Adaptive Temporal Masking for Stable Sparse Autoencoder Training
von: Li, T. Ed, et al.
Veröffentlicht: (2025)
von: Li, T. Ed, et al.
Veröffentlicht: (2025)
Adaptive Divergence Regularized Policy Optimization for Fine-tuning Generative Models
von: Fan, Jiajun, et al.
Veröffentlicht: (2025)
von: Fan, Jiajun, et al.
Veröffentlicht: (2025)
Do Sparse Autoencoders Identify Reasoning Features in Language Models?
von: Ma, George, et al.
Veröffentlicht: (2026)
von: Ma, George, et al.
Veröffentlicht: (2026)
Diagonal Over-parameterization in Reproducing Kernel Hilbert Spaces as an Adaptive Feature Model: Generalization and Adaptivity
von: Li, Yicheng, et al.
Veröffentlicht: (2025)
von: Li, Yicheng, et al.
Veröffentlicht: (2025)
Sparse Partial Optimal Transport via Quadratic Regularization
von: Tran, Khang, et al.
Veröffentlicht: (2025)
von: Tran, Khang, et al.
Veröffentlicht: (2025)
Sparsity via Sparse Group $k$-max Regularization
von: Tao, Qinghua, et al.
Veröffentlicht: (2024)
von: Tao, Qinghua, et al.
Veröffentlicht: (2024)
Sparse NMF with Archetypal Regularization: Computational and Robustness Properties
von: Behdin, Kayhan, et al.
Veröffentlicht: (2021)
von: Behdin, Kayhan, et al.
Veröffentlicht: (2021)
Mixture of Scales: Memory-Efficient Token-Adaptive Binarization for Large Language Models
von: Jo, Dongwon, et al.
Veröffentlicht: (2024)
von: Jo, Dongwon, et al.
Veröffentlicht: (2024)
MLKV: Efficiently Scaling up Large Embedding Model Training with Disk-based Key-Value Storage
von: He, Yongjun, et al.
Veröffentlicht: (2025)
von: He, Yongjun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Proximal Projection for Doubly Sparse Regularized Models
von: He, Jia Wei, et al.
Veröffentlicht: (2026) -
Tree-Regularized Tabular Embeddings
von: Li, Xuan, et al.
Veröffentlicht: (2024) -
SAE-FD: Sparse Autoencoder Feature Distillation for Continual Learning of Large Language Models
von: Zhang, Mingxu, et al.
Veröffentlicht: (2026) -
Adaptive Block Sparse Regularization under Arbitrary Linear Transform
von: Furuhashi, Takanobu, et al.
Veröffentlicht: (2024) -
Adaptive Sparse Allocation with Mutual Choice & Feature Choice Sparse Autoencoders
von: Ayonrinde, Kola
Veröffentlicht: (2024)