LaCoOT: Layer Collapse through Optimal Transport
Fuente:
arXiv
Saved in:
| Main Authors: | Quétu, Victor, Liao, Zhu, Hezbri, Nour, Pizzati, Fabio, Tartaglione, Enzo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Till the Layers Collapse: Compressing a Deep Neural Network through the Lenses of Batch Normalization Layers
by: Liao, Zhu, et al.
Published: (2024)
by: Liao, Zhu, et al.
Published: (2024)
Layer Collapse Can be Induced by Unstructured Pruning
by: Liao, Zhu, et al.
Published: (2024)
by: Liao, Zhu, et al.
Published: (2024)
The Simpler The Better: An Entropy-Based Importance Metric To Reduce Neural Networks' Depth
by: Quétu, Victor, et al.
Published: (2024)
by: Quétu, Victor, et al.
Published: (2024)
DSD$^2$: Can We Dodge Sparse Double Descent and Compress the Neural Network Worry-Free?
by: Quétu, Victor, et al.
Published: (2023)
by: Quétu, Victor, et al.
Published: (2023)
Study of Training Dynamics for Memory-Constrained Fine-Tuning
by: Quélennec, Aël, et al.
Published: (2025)
by: Quélennec, Aël, et al.
Published: (2025)
Memory-Optimized Once-For-All Network
by: Girard, Maxime, et al.
Published: (2024)
by: Girard, Maxime, et al.
Published: (2024)
GradNetOT: Learning Optimal Transport Maps with GradNets
by: Chaudhari, Shreyas, et al.
Published: (2025)
by: Chaudhari, Shreyas, et al.
Published: (2025)
Statistical Guarantees for Distributionally Robust Optimization with Optimal Transport and OT-Regularized Divergences
by: Birrell, Jeremiah, et al.
Published: (2026)
by: Birrell, Jeremiah, et al.
Published: (2026)
React-OT: Optimal Transport for Generating Transition State in Chemical Reactions
by: Duan, Chenru, et al.
Published: (2024)
by: Duan, Chenru, et al.
Published: (2024)
GCL-OT: Graph Contrastive Learning with Optimal Transport for Heterophilic Text-Attributed Graphs
by: Ren, Yating, et al.
Published: (2025)
by: Ren, Yating, et al.
Published: (2025)
Efficient Resource-Constrained Training of Transformers via Subspace Optimization
by: Nguyen, Le-Trung, et al.
Published: (2025)
by: Nguyen, Le-Trung, et al.
Published: (2025)
OT-Transformer: A Continuous-time Transformer Architecture with Optimal Transport Regularization
by: Kan, Kelvin, et al.
Published: (2025)
by: Kan, Kelvin, et al.
Published: (2025)
Multi-robot Path Planning and Scheduling via Model Predictive Optimal Transport (MPC-OT)
by: Khan, Usman A., et al.
Published: (2025)
by: Khan, Usman A., et al.
Published: (2025)
Hoeffding Concept Bottleneck Models with Applications to Overhead Images
by: Bénard, Clément, et al.
Published: (2026)
by: Bénard, Clément, et al.
Published: (2026)
OT-VP: Optimal Transport-guided Visual Prompting for Test-Time Adaptation
by: Zhang, Yunbei, et al.
Published: (2024)
by: Zhang, Yunbei, et al.
Published: (2024)
Efficient Adaptation of Deep Neural Networks for Semantic Segmentation in Space Applications
by: Olivi, Leonardo, et al.
Published: (2025)
by: Olivi, Leonardo, et al.
Published: (2025)
Bispectral OT: Dataset Comparison using Symmetry-Aware Optimal Transport
by: Ma, Annabel, et al.
Published: (2025)
by: Ma, Annabel, et al.
Published: (2025)
SP$^2$OT: Semantic-Regularized Progressive Partial Optimal Transport for Imbalanced Clustering
by: Zhang, Chuyu, et al.
Published: (2024)
by: Zhang, Chuyu, et al.
Published: (2024)
Unsupervised Learning of Unbiased Visual Representations
by: Barbano, Carlo Alberto, et al.
Published: (2022)
by: Barbano, Carlo Alberto, et al.
Published: (2022)
OT-MeanFlow3D: Bridging Optimal Transport and Meanflow for Efficient 3D Point Cloud Generation
by: Akbari, Elaheh, et al.
Published: (2025)
by: Akbari, Elaheh, et al.
Published: (2025)
Activation Map Compression through Tensor Decomposition for Deep Learning
by: Nguyen, Le-Trung, et al.
Published: (2024)
by: Nguyen, Le-Trung, et al.
Published: (2024)
HYGENE: A Diffusion-based Hypergraph Generation Method
by: Gailhard, Dorian, et al.
Published: (2024)
by: Gailhard, Dorian, et al.
Published: (2024)
Feature-Aware (Hyper)graph Generation via Next-Scale Prediction
by: Gailhard, Dorian, et al.
Published: (2025)
by: Gailhard, Dorian, et al.
Published: (2025)
Memory Constrained Dynamic Subnetwork Update for Transfer Learning
by: Quélennec, Aël, et al.
Published: (2025)
by: Quélennec, Aël, et al.
Published: (2025)
cuRegOT: A GPU-Accelerated Solver for Entropic-Regularized Optimal Transport
by: Qiu, Yixuan
Published: (2026)
by: Qiu, Yixuan
Published: (2026)
Beyond Low-rank Decomposition: A Shortcut Approach for Efficient On-Device Learning
by: Nguyen, Le-Trung, et al.
Published: (2025)
by: Nguyen, Le-Trung, et al.
Published: (2025)
The silence of the weights: a structural pruning strategy for attention-based audio signal architectures with second order metrics
by: Diecidue, Andrea, et al.
Published: (2025)
by: Diecidue, Andrea, et al.
Published: (2025)
Weighted Ensemble Models Are Strong Continual Learners
by: Marouf, Imad Eddine, et al.
Published: (2023)
by: Marouf, Imad Eddine, et al.
Published: (2023)
How I Met Your Bias: Investigating Bias Amplification in Diffusion Models
by: Roos, Nathan, et al.
Published: (2025)
by: Roos, Nathan, et al.
Published: (2025)
Layer Collapse in Diffusion Language Models
by: Conzelmann, Alexander, et al.
Published: (2026)
by: Conzelmann, Alexander, et al.
Published: (2026)
Packed-Ensembles for Efficient Uncertainty Estimation
by: Laurent, Olivier, et al.
Published: (2022)
by: Laurent, Olivier, et al.
Published: (2022)
Specify and Edit: Overcoming Ambiguity in Text-Based Image Editing
by: Iakovleva, Ekaterina, et al.
Published: (2024)
by: Iakovleva, Ekaterina, et al.
Published: (2024)
Simplifying Optimal Transport through Schatten-$p$ Regularization
by: Maunu, Tyler
Published: (2025)
by: Maunu, Tyler
Published: (2025)
Embedding the MLOps Lifecycle into OT Reference Models
by: Schindler, Simon, et al.
Published: (2025)
by: Schindler, Simon, et al.
Published: (2025)
OT Score: An OT based Confidence Score for Prototype-Assisted Source Free Unsupervised Domain Adaptation
by: Zhang, Yiming, et al.
Published: (2025)
by: Zhang, Yiming, et al.
Published: (2025)
Counterfactual Identifiability via Dynamic Optimal Transport
by: Ribeiro, Fabio De Sousa, et al.
Published: (2025)
by: Ribeiro, Fabio De Sousa, et al.
Published: (2025)
Dynamic Conditional Optimal Transport through Simulation-Free Flows
by: Kerrigan, Gavin, et al.
Published: (2024)
by: Kerrigan, Gavin, et al.
Published: (2024)
What's Behind PPO's Collapse in Long-CoT? Value Optimization Holds the Secret
by: Yuan, Yufeng, et al.
Published: (2025)
by: Yuan, Yufeng, et al.
Published: (2025)
OT on the Map: Quantifying Domain Shifts in Geographic Space
by: Zhang, Haoran, et al.
Published: (2026)
by: Zhang, Haoran, et al.
Published: (2026)
WaterMAS: Sharpness-Aware Maximization for Neural Network Watermarking
by: Trias, Carl De Sousa, et al.
Published: (2024)
by: Trias, Carl De Sousa, et al.
Published: (2024)
Similar Items
-
Till the Layers Collapse: Compressing a Deep Neural Network through the Lenses of Batch Normalization Layers
by: Liao, Zhu, et al.
Published: (2024) -
Layer Collapse Can be Induced by Unstructured Pruning
by: Liao, Zhu, et al.
Published: (2024) -
The Simpler The Better: An Entropy-Based Importance Metric To Reduce Neural Networks' Depth
by: Quétu, Victor, et al.
Published: (2024) -
DSD$^2$: Can We Dodge Sparse Double Descent and Compress the Neural Network Worry-Free?
by: Quétu, Victor, et al.
Published: (2023) -
Study of Training Dynamics for Memory-Constrained Fine-Tuning
by: Quélennec, Aël, et al.
Published: (2025)