Guardado en:
| Autor principal: | Liu, Yizhi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2601.02080 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Avoiding Premature Collapse: Adaptive Annealing for Entropy-Regularized Structural Inference
por: Liu, Yizhi
Publicado: (2026)
por: Liu, Yizhi
Publicado: (2026)
Partially Observable Gaussian Process Network and Doubly Stochastic Variational Inference
por: Kiroriwal, Saksham, et al.
Publicado: (2025)
por: Kiroriwal, Saksham, et al.
Publicado: (2025)
Doubly Stochastic Adaptive Neighbors Clustering via the Marcus Mapping
por: Yuan, Jinghui, et al.
Publicado: (2024)
por: Yuan, Jinghui, et al.
Publicado: (2024)
Collapsed Inference for Bayesian Deep Learning
por: Zeng, Zhe, et al.
Publicado: (2023)
por: Zeng, Zhe, et al.
Publicado: (2023)
Doubly Robust Monte Carlo Tree Search
por: Liu, Manqing, et al.
Publicado: (2025)
por: Liu, Manqing, et al.
Publicado: (2025)
Quantum Doubly Stochastic Transformers
por: Born, Jannis, et al.
Publicado: (2025)
por: Born, Jannis, et al.
Publicado: (2025)
Addressing Spectral Bias of Deep Neural Networks by Multi-Grade Deep Learning
por: Fang, Ronglong, et al.
Publicado: (2024)
por: Fang, Ronglong, et al.
Publicado: (2024)
Doubly Inhomogeneous Reinforcement Learning
por: Hu, Liyuan, et al.
Publicado: (2022)
por: Hu, Liyuan, et al.
Publicado: (2022)
Gene Regulatory Network Inference from Pre-trained Single-Cell Transcriptomics Transformer with Joint Graph Learning
por: Kommu, Sindhura, et al.
Publicado: (2024)
por: Kommu, Sindhura, et al.
Publicado: (2024)
Stochastic Collapse: How Gradient Noise Attracts SGD Dynamics Towards Simpler Subnetworks
por: Chen, Feng, et al.
Publicado: (2023)
por: Chen, Feng, et al.
Publicado: (2023)
Spectral Entropy Collapse as a Phase Transition in Delayed Generalisation: An Interventional and Predictive Framework for Grokkin
por: Khanh, Truong Xuan, et al.
Publicado: (2026)
por: Khanh, Truong Xuan, et al.
Publicado: (2026)
Has the Deep Neural Network learned the Stochastic Process? An Evaluation Viewpoint
por: Kumar, Harshit, et al.
Publicado: (2024)
por: Kumar, Harshit, et al.
Publicado: (2024)
Deep Neural Regression Collapse
por: Rangamani, Akshay, et al.
Publicado: (2026)
por: Rangamani, Akshay, et al.
Publicado: (2026)
Neural Collapse is Globally Optimal in Deep Regularized ResNets and Transformers
por: Súkeník, Peter, et al.
Publicado: (2025)
por: Súkeník, Peter, et al.
Publicado: (2025)
Understanding Representation of Deep Equilibrium Models from Neural Collapse Perspective
por: Sun, Haixiang, et al.
Publicado: (2024)
por: Sun, Haixiang, et al.
Publicado: (2024)
Doubly Mild Generalization for Offline Reinforcement Learning
por: Mao, Yixiu, et al.
Publicado: (2024)
por: Mao, Yixiu, et al.
Publicado: (2024)
Doubly Robust Alignment for Large Language Models
por: Xu, Erhan, et al.
Publicado: (2025)
por: Xu, Erhan, et al.
Publicado: (2025)
Clarifying Shampoo: Adapting Spectral Descent to Stochasticity and the Parameter Trajectory
por: Eschenhagen, Runa, et al.
Publicado: (2026)
por: Eschenhagen, Runa, et al.
Publicado: (2026)
Distribution Fitting for Combating Mode Collapse in Generative Adversarial Networks
por: Gong, Yanxiang, et al.
Publicado: (2022)
por: Gong, Yanxiang, et al.
Publicado: (2022)
Spectral Collapse in Diffusion Inversion
por: Bourriez, Nicolas, et al.
Publicado: (2026)
por: Bourriez, Nicolas, et al.
Publicado: (2026)
Spectral Gating Networks
por: Zhang, Jusheng, et al.
Publicado: (2026)
por: Zhang, Jusheng, et al.
Publicado: (2026)
Representation Collapsing Problems in Vector Quantization
por: Zhao, Wenhao, et al.
Publicado: (2024)
por: Zhao, Wenhao, et al.
Publicado: (2024)
Spectral Convolution on Orbifolds for Geometric Deep Learning
por: Mangliers, Tim, et al.
Publicado: (2026)
por: Mangliers, Tim, et al.
Publicado: (2026)
Late Breaking Results: Quamba-SE: Soft-edge Quantizer for Activations in State Space Models
por: Chen, Yizhi, et al.
Publicado: (2026)
por: Chen, Yizhi, et al.
Publicado: (2026)
Safe RLHF Beyond Expectation: Stochastic Dominance for Universal Spectral Risk Control
por: Chittepu, Yaswanth, et al.
Publicado: (2026)
por: Chittepu, Yaswanth, et al.
Publicado: (2026)
Training Multimodal Large Reasoning Models Needs Better Thoughts: A Three-Stage Framework for Long Chain-of-Thought Synthesis and Selection
por: Wang, Yizhi, et al.
Publicado: (2025)
por: Wang, Yizhi, et al.
Publicado: (2025)
Why Do Neural Networks Forget: A Study of Collapse in Continual Learning
por: Zhu, Yunqin, et al.
Publicado: (2026)
por: Zhu, Yunqin, et al.
Publicado: (2026)
Rank Collapse Causes Over-Smoothing and Over-Correlation in Graph Neural Networks
por: Roth, Andreas, et al.
Publicado: (2023)
por: Roth, Andreas, et al.
Publicado: (2023)
Cross-Entropy Optimization for Hyperparameter Optimization in Stochastic Gradient-based Approaches to Train Deep Neural Networks
por: Li, Kevin, et al.
Publicado: (2024)
por: Li, Kevin, et al.
Publicado: (2024)
Approaching Deep Learning through the Spectral Dynamics of Weights
por: Yunis, David, et al.
Publicado: (2024)
por: Yunis, David, et al.
Publicado: (2024)
Spectral Higher-Order Neural Networks
por: Peri, Gianluca, et al.
Publicado: (2026)
por: Peri, Gianluca, et al.
Publicado: (2026)
Spatio-Spectral Graph Neural Networks
por: Geisler, Simon, et al.
Publicado: (2024)
por: Geisler, Simon, et al.
Publicado: (2024)
TriCon-SF: A Triple-Shuffle and Contribution-Aware Serial Federated Learning Framework for Heterogeneous Healthcare Data
por: Yan, Yuping, et al.
Publicado: (2025)
por: Yan, Yuping, et al.
Publicado: (2025)
Stochastic Deep Graph Clustering for Practical Group Formation
por: Park, Junhyung, et al.
Publicado: (2025)
por: Park, Junhyung, et al.
Publicado: (2025)
Equivariant Eikonal Neural Networks: Grid-Free, Scalable Travel-Time Prediction on Homogeneous Spaces
por: García-Castellanos, Alejandro, et al.
Publicado: (2025)
por: García-Castellanos, Alejandro, et al.
Publicado: (2025)
LiPUP-MA: A Residential Experience-centric Multi-Agent Framework for Living-in-the-loop Participatory Urban Planning
por: Ni, Hang, et al.
Publicado: (2024)
por: Ni, Hang, et al.
Publicado: (2024)
Scale-Dropout: Estimating Uncertainty in Deep Neural Networks Using Stochastic Scale
por: Ahmed, Soyed Tuhin, et al.
Publicado: (2023)
por: Ahmed, Soyed Tuhin, et al.
Publicado: (2023)
Spectral Greedy Coresets for Graph Neural Networks
por: Ding, Mucong, et al.
Publicado: (2024)
por: Ding, Mucong, et al.
Publicado: (2024)
Piecewise Constant Spectral Graph Neural Network
por: Martirosyan, Vahan, et al.
Publicado: (2025)
por: Martirosyan, Vahan, et al.
Publicado: (2025)
Tequila: Trapping-free Ternary Quantization for Large Language Models
por: Huang, Hong, et al.
Publicado: (2025)
por: Huang, Hong, et al.
Publicado: (2025)
Ejemplares similares
-
Avoiding Premature Collapse: Adaptive Annealing for Entropy-Regularized Structural Inference
por: Liu, Yizhi
Publicado: (2026) -
Partially Observable Gaussian Process Network and Doubly Stochastic Variational Inference
por: Kiroriwal, Saksham, et al.
Publicado: (2025) -
Doubly Stochastic Adaptive Neighbors Clustering via the Marcus Mapping
por: Yuan, Jinghui, et al.
Publicado: (2024) -
Collapsed Inference for Bayesian Deep Learning
por: Zeng, Zhe, et al.
Publicado: (2023) -
Doubly Robust Monte Carlo Tree Search
por: Liu, Manqing, et al.
Publicado: (2025)