Salvato in:
| Autori principali: | Hor, Soheil, Qian, Ying, Pilanci, Mert, Arbabian, Amin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2402.04359 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Adaptive Large Language Models By Layerwise Attention Shortcuts
di: Verma, Prateek, et al.
Pubblicazione: (2024)
di: Verma, Prateek, et al.
Pubblicazione: (2024)
From Complexity to Clarity: Analytical Expressions of Deep Neural Network Weights via Clifford's Geometric Algebra and Convexity
di: Pilanci, Mert
Pubblicazione: (2023)
di: Pilanci, Mert
Pubblicazione: (2023)
Convex Distillation: Efficient Compression of Deep Networks via Convex Optimization
di: Varshney, Prateek, et al.
Pubblicazione: (2024)
di: Varshney, Prateek, et al.
Pubblicazione: (2024)
Black Boxes and Looking Glasses: Multilevel Symmetries, Reflection Planes, and Convex Optimization in Deep Networks
di: Zeger, Emi, et al.
Pubblicazione: (2024)
di: Zeger, Emi, et al.
Pubblicazione: (2024)
Optimal Sets and Solution Paths of ReLU Networks
di: Mishkin, Aaron, et al.
Pubblicazione: (2023)
di: Mishkin, Aaron, et al.
Pubblicazione: (2023)
AdaPTwin: Low-Cost Adaptive Compression of Product Twins in Transformers
di: Biju, Emil, et al.
Pubblicazione: (2024)
di: Biju, Emil, et al.
Pubblicazione: (2024)
Analyzing Neural Network-Based Generative Diffusion Models through Convex Optimization
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
Convex Relaxations of ReLU Neural Networks Approximate Global Optima in Polynomial Time
di: Kim, Sungyoon, et al.
Pubblicazione: (2024)
di: Kim, Sungyoon, et al.
Pubblicazione: (2024)
Spectral Adapter: Fine-Tuning in Spectral Space
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
Unveiling Hidden Convexity in Deep Learning: a Sparse Signal Processing Perspective
di: Zeger, Emi, et al.
Pubblicazione: (2026)
di: Zeger, Emi, et al.
Pubblicazione: (2026)
Riemannian Preconditioned LoRA for Fine-Tuning Foundation Models
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
A Recovery Guarantee for Sparse Neural Networks
di: Fridovich-Keil, Sara, et al.
Pubblicazione: (2025)
di: Fridovich-Keil, Sara, et al.
Pubblicazione: (2025)
Thinking While Listening: Simple Test Time Scaling For Audio Classification
di: Verma, Prateek, et al.
Pubblicazione: (2025)
di: Verma, Prateek, et al.
Pubblicazione: (2025)
Exploring the loss landscape of regularized neural networks via convex duality
di: Kim, Sungyoon, et al.
Pubblicazione: (2024)
di: Kim, Sungyoon, et al.
Pubblicazione: (2024)
Fast Convex Optimization for Two-Layer ReLU Networks: Equivalent Model Classes and Cone Decompositions
di: Mishkin, Aaron, et al.
Pubblicazione: (2022)
di: Mishkin, Aaron, et al.
Pubblicazione: (2022)
Towards Signal Processing In Large Language Models
di: Verma, Prateek, et al.
Pubblicazione: (2024)
di: Verma, Prateek, et al.
Pubblicazione: (2024)
Faster Convergence of Stochastic Accelerated Gradient Descent under Interpolation
di: Mishkin, Aaron, et al.
Pubblicazione: (2024)
di: Mishkin, Aaron, et al.
Pubblicazione: (2024)
Active Learning of Deep Neural Networks via Gradient-Free Cutting Planes
di: Zhang, Erica, et al.
Pubblicazione: (2024)
di: Zhang, Erica, et al.
Pubblicazione: (2024)
CRONOS: Enhancing Deep Learning with Scalable GPU Accelerated Convex Neural Networks
di: Feng, Miria, et al.
Pubblicazione: (2024)
di: Feng, Miria, et al.
Pubblicazione: (2024)
Adversarial Training of Two-Layer Polynomial and ReLU Activation Networks via Convex Optimization
di: Kuelbs, Daniel, et al.
Pubblicazione: (2024)
di: Kuelbs, Daniel, et al.
Pubblicazione: (2024)
Large Language Models Implicitly Learn to See and Hear Just By Reading
di: Verma, Prateek, et al.
Pubblicazione: (2025)
di: Verma, Prateek, et al.
Pubblicazione: (2025)
Convex Optimization for Alignment and Preference Learning on a Single GPU
di: Feng, Miria, et al.
Pubblicazione: (2026)
di: Feng, Miria, et al.
Pubblicazione: (2026)
Convex Low-resource Accent-Robust Language Detection in Speech Recognition
di: Feng, Miria, et al.
Pubblicazione: (2026)
di: Feng, Miria, et al.
Pubblicazione: (2026)
Newton Meets Marchenko-Pastur: Massively Parallel Second-Order Optimization with Hessian Sketching and Debiasing
di: Romanov, Elad, et al.
Pubblicazione: (2024)
di: Romanov, Elad, et al.
Pubblicazione: (2024)
MatRL: Provably Generalizable Iterative Algorithm Discovery via Monte-Carlo Tree Search
di: Kim, Sungyoon, et al.
Pubblicazione: (2025)
di: Kim, Sungyoon, et al.
Pubblicazione: (2025)
NanoFlux: Adversarial Dual-LLM Evaluation and Distillation For Multi-Domain Reasoning
di: Anantha, Raviteja, et al.
Pubblicazione: (2025)
di: Anantha, Raviteja, et al.
Pubblicazione: (2025)
Randomized Geometric Algebra Methods for Convex Neural Networks
di: Wang, Yifei, et al.
Pubblicazione: (2024)
di: Wang, Yifei, et al.
Pubblicazione: (2024)
Learning When to Trust LLM Priors: A Validated Framework for Semantic Prior Integration
di: Zhang, Erica, et al.
Pubblicazione: (2026)
di: Zhang, Erica, et al.
Pubblicazione: (2026)
Optimizer-Induced Mode Connectivity: From AdamW to Muon
di: Zhang, Fangzhao, et al.
Pubblicazione: (2026)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2026)
A Measure-Theoretic Finite-Sample Theory for Adaptive-Data Fitted Q-Iteration
di: Haussmann, Manuel, et al.
Pubblicazione: (2026)
di: Haussmann, Manuel, et al.
Pubblicazione: (2026)
Compressing Large Language Models using Low Rank and Low Precision Decomposition
di: Saha, Rajarshi, et al.
Pubblicazione: (2024)
di: Saha, Rajarshi, et al.
Pubblicazione: (2024)
Deep-Learning-Directed Preventive Dynamic Security Control via Coordinated Demand Response
di: Masoumi, Amin, et al.
Pubblicazione: (2025)
di: Masoumi, Amin, et al.
Pubblicazione: (2025)
Distributed Markov Chain Monte Carlo Sampling based on the Alternating Direction Method of Multipliers
di: Tzikas, Alexandros E., et al.
Pubblicazione: (2024)
di: Tzikas, Alexandros E., et al.
Pubblicazione: (2024)
Birefringence-free photoelastic modulator with centimeter-square aperture operating at 2.7 MHz with sub-watt drive power
di: Atalar, Okan, et al.
Pubblicazione: (2024)
di: Atalar, Okan, et al.
Pubblicazione: (2024)
Polarization-insensitive wide-angle resonant acousto-optic phase modulator
di: Atalar, Okan, et al.
Pubblicazione: (2024)
di: Atalar, Okan, et al.
Pubblicazione: (2024)
A Library of Mirrors: Deep Neural Nets in Low Dimensions are Convex Lasso Models with Reflection Features
di: Zeger, Emi, et al.
Pubblicazione: (2024)
di: Zeger, Emi, et al.
Pubblicazione: (2024)
ConvexECG: Lightweight and Explainable Neural Networks for Personalized, Continuous Cardiac Monitoring
di: Ansari, Rayan, et al.
Pubblicazione: (2024)
di: Ansari, Rayan, et al.
Pubblicazione: (2024)
Unexplored flaws in multiple-choice VQA evaluations
di: Rosenthal, Fabio, et al.
Pubblicazione: (2025)
di: Rosenthal, Fabio, et al.
Pubblicazione: (2025)
Gradient Coding with Iterative Block Leverage Score Sampling
di: Charalambides, Neophytos, et al.
Pubblicazione: (2023)
di: Charalambides, Neophytos, et al.
Pubblicazione: (2023)
Subtractive Training for Music Stem Insertion using Latent Diffusion Models
di: Villa-Renteria, Ivan, et al.
Pubblicazione: (2024)
di: Villa-Renteria, Ivan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Adaptive Large Language Models By Layerwise Attention Shortcuts
di: Verma, Prateek, et al.
Pubblicazione: (2024) -
From Complexity to Clarity: Analytical Expressions of Deep Neural Network Weights via Clifford's Geometric Algebra and Convexity
di: Pilanci, Mert
Pubblicazione: (2023) -
Convex Distillation: Efficient Compression of Deep Networks via Convex Optimization
di: Varshney, Prateek, et al.
Pubblicazione: (2024) -
Black Boxes and Looking Glasses: Multilevel Symmetries, Reflection Planes, and Convex Optimization in Deep Networks
di: Zeger, Emi, et al.
Pubblicazione: (2024) -
Optimal Sets and Solution Paths of ReLU Networks
di: Mishkin, Aaron, et al.
Pubblicazione: (2023)