Slicing Mutual Information Generalization Bounds for Neural Networks
Fuente:
arXiv
Guardado en:
| Autores principales: | Nadjahi, Kimia, Greenewald, Kristjan, Gabrielsson, Rickard Brüel, Solomon, Justin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Asymmetry in Low-Rank Adapters of Foundation Models
por: Zhu, Jiacheng, et al.
Publicado: (2024)
por: Zhu, Jiacheng, et al.
Publicado: (2024)
Compress then Serve: Serving Thousands of LoRA Adapters with Little Overhead
por: Brüel-Gabrielsson, Rickard, et al.
Publicado: (2024)
por: Brüel-Gabrielsson, Rickard, et al.
Publicado: (2024)
Deep Augmentation: Dropout as Augmentation for Self-Supervised Learning
por: Brüel-Gabrielsson, Rickard, et al.
Publicado: (2023)
por: Brüel-Gabrielsson, Rickard, et al.
Publicado: (2023)
Convergence Rates for Distribution Matching with Sliced Optimal Transport
por: Thurin, Gauthier, et al.
Publicado: (2026)
por: Thurin, Gauthier, et al.
Publicado: (2026)
Tighter CMI-Based Generalization Bounds via Stochastic Projection and Quantization
por: Sefidgaran, Milad, et al.
Publicado: (2025)
por: Sefidgaran, Milad, et al.
Publicado: (2025)
Privacy without Noisy Gradients: Slicing Mechanism for Generative Model Training
por: Greenewald, Kristjan, et al.
Publicado: (2024)
por: Greenewald, Kristjan, et al.
Publicado: (2024)
Slicing Unbalanced Optimal Transport
por: Bonet, Clément, et al.
Publicado: (2023)
por: Bonet, Clément, et al.
Publicado: (2023)
Private Continuous-Time Synthetic Trajectory Generation via Mean-Field Langevin Dynamics
por: Gu, Anming, et al.
Publicado: (2025)
por: Gu, Anming, et al.
Publicado: (2025)
Optimal Transport-based Conformal Prediction
por: Thurin, Gauthier, et al.
Publicado: (2025)
por: Thurin, Gauthier, et al.
Publicado: (2025)
Expected Batch Optimal Transport Plans and Consequences for Flow Matching
por: Boïté, Samuel, et al.
Publicado: (2026)
por: Boïté, Samuel, et al.
Publicado: (2026)
Neural Estimation for Scaling Entropic Multimarginal Optimal Transport
por: Tsur, Dor, et al.
Publicado: (2025)
por: Tsur, Dor, et al.
Publicado: (2025)
Partially Observed Trajectory Inference using Optimal Transport and a Dynamics Prior
por: Gu, Anming, et al.
Publicado: (2024)
por: Gu, Anming, et al.
Publicado: (2024)
Balanced LoRA: Removing Parameter Invariance to Accelerate Convergence
por: Castin, Valérie, et al.
Publicado: (2026)
por: Castin, Valérie, et al.
Publicado: (2026)
Distributional Process Reward Models: Calibrated Prediction of Future Rewards via Conditional Optimal Transport
por: Ma, Rachel, et al.
Publicado: (2026)
por: Ma, Rachel, et al.
Publicado: (2026)
Entropic Causal Inference: Graph Identifiability
por: Compton, Spencer, et al.
Publicado: (2025)
por: Compton, Spencer, et al.
Publicado: (2025)
Differentially Private Wasserstein Barycenters
por: Gu, Anming, et al.
Publicado: (2025)
por: Gu, Anming, et al.
Publicado: (2025)
Efficient Multi-Adapter LLM Serving via Cross-Model KV-Cache Reuse with Activated LoRA
por: Li, Allison, et al.
Publicado: (2025)
por: Li, Allison, et al.
Publicado: (2025)
Score Distillation via Reparametrized DDIM
por: Lukoianov, Artem, et al.
Publicado: (2024)
por: Lukoianov, Artem, et al.
Publicado: (2024)
Multivariate Stochastic Dominance via Optimal Transport and Applications to Models Benchmarking
por: Rioux, Gabriel, et al.
Publicado: (2024)
por: Rioux, Gabriel, et al.
Publicado: (2024)
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
por: Park, Young-Jin, et al.
Publicado: (2025)
por: Park, Young-Jin, et al.
Publicado: (2025)
Mutual Information Preserving Neural Network Pruning
por: Westphal, Charles, et al.
Publicado: (2024)
por: Westphal, Charles, et al.
Publicado: (2024)
Thermometer: Towards Universal Calibration for Large Language Models
por: Shen, Maohao, et al.
Publicado: (2024)
por: Shen, Maohao, et al.
Publicado: (2024)
Gradient Flows and Riemannian Structure in the Gromov-Wasserstein Geometry
por: Zhang, Zhengxin, et al.
Publicado: (2024)
por: Zhang, Zhengxin, et al.
Publicado: (2024)
Pruning Deep Convolutional Neural Network Using Conditional Mutual Information
por: Vu-Van, Tien, et al.
Publicado: (2024)
por: Vu-Van, Tien, et al.
Publicado: (2024)
Information-Theoretic Generalization Bounds for Deep Neural Networks
por: He, Haiyun, et al.
Publicado: (2024)
por: He, Haiyun, et al.
Publicado: (2024)
Improving Mutual Information Estimation with Annealed and Energy-Based Bounds
por: Brekelmans, Rob, et al.
Publicado: (2023)
por: Brekelmans, Rob, et al.
Publicado: (2023)
Compressing Deep Neural Networks Using Explainable AI
por: Soroush, Kimia, et al.
Publicado: (2025)
por: Soroush, Kimia, et al.
Publicado: (2025)
Neural Mutual Information Estimation with Vector Copulas
por: Chen, Yanzhi, et al.
Publicado: (2025)
por: Chen, Yanzhi, et al.
Publicado: (2025)
MINDE: Mutual Information Neural Diffusion Estimation
por: Franzese, Giulio, et al.
Publicado: (2023)
por: Franzese, Giulio, et al.
Publicado: (2023)
Convergence of Statistical Estimators via Mutual Information Bounds
por: Khribch, El Mahdi, et al.
Publicado: (2024)
por: Khribch, El Mahdi, et al.
Publicado: (2024)
Heterogeneous Sheaf Neural Networks
por: Braithwaite, Luke, et al.
Publicado: (2024)
por: Braithwaite, Luke, et al.
Publicado: (2024)
Distributional Preference Alignment of LLMs via Optimal Transport
por: Melnyk, Igor, et al.
Publicado: (2024)
por: Melnyk, Igor, et al.
Publicado: (2024)
Revisiting Group Relative Policy Optimization: Insights into On-Policy and Off-Policy Training
por: Mroueh, Youssef, et al.
Publicado: (2025)
por: Mroueh, Youssef, et al.
Publicado: (2025)
Generalization and Risk Bounds for Recurrent Neural Networks
por: Cheng, Xuewei, et al.
Publicado: (2024)
por: Cheng, Xuewei, et al.
Publicado: (2024)
Generalization Bounds for Rank-sparse Neural Networks
por: Ledent, Antoine, et al.
Publicado: (2025)
por: Ledent, Antoine, et al.
Publicado: (2025)
Slice and Explain: Logic-Based Explanations for Neural Networks through Domain Slicing
por: Queiroz, Luiz Fernando Paulino, et al.
Publicado: (2026)
por: Queiroz, Luiz Fernando Paulino, et al.
Publicado: (2026)
Nuclear Norm Regularization for Deep Learning
por: Scarvelis, Christopher, et al.
Publicado: (2024)
por: Scarvelis, Christopher, et al.
Publicado: (2024)
Sensitivity Analysis for Diffusion Models
por: Scarvelis, Christopher, et al.
Publicado: (2025)
por: Scarvelis, Christopher, et al.
Publicado: (2025)
MPruner: Optimizing Neural Network Size with CKA-Based Mutual Information Pruning
por: Hu, Seungbeom, et al.
Publicado: (2024)
por: Hu, Seungbeom, et al.
Publicado: (2024)
On Generalization Bounds for Neural Networks with Low Rank Layers
por: Pinto, Andrea, et al.
Publicado: (2024)
por: Pinto, Andrea, et al.
Publicado: (2024)
Ejemplares similares
-
Asymmetry in Low-Rank Adapters of Foundation Models
por: Zhu, Jiacheng, et al.
Publicado: (2024) -
Compress then Serve: Serving Thousands of LoRA Adapters with Little Overhead
por: Brüel-Gabrielsson, Rickard, et al.
Publicado: (2024) -
Deep Augmentation: Dropout as Augmentation for Self-Supervised Learning
por: Brüel-Gabrielsson, Rickard, et al.
Publicado: (2023) -
Convergence Rates for Distribution Matching with Sliced Optimal Transport
por: Thurin, Gauthier, et al.
Publicado: (2026) -
Tighter CMI-Based Generalization Bounds via Stochastic Projection and Quantization
por: Sefidgaran, Milad, et al.
Publicado: (2025)