LLM-Lasso: A Robust Framework for Domain-Informed Feature Selection and Regularization
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Erica, Goto, Ryunosuke, Sagan, Naomi, Mutter, Jurik, Phillips, Nick, Alizadeh, Ash, Lee, Kangwook, Blanchet, Jose, Pilanci, Mert, Tibshirani, Robert |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning When to Trust LLM Priors: A Validated Framework for Semantic Prior Integration
di: Zhang, Erica, et al.
Pubblicazione: (2026)
di: Zhang, Erica, et al.
Pubblicazione: (2026)
Pretraining and the Lasso
di: Craig, Erin, et al.
Pubblicazione: (2024)
di: Craig, Erin, et al.
Pubblicazione: (2024)
Compressing Large Language Models using Low Rank and Low Precision Decomposition
di: Saha, Rajarshi, et al.
Pubblicazione: (2024)
di: Saha, Rajarshi, et al.
Pubblicazione: (2024)
A Library of Mirrors: Deep Neural Nets in Low Dimensions are Convex Lasso Models with Reflection Features
di: Zeger, Emi, et al.
Pubblicazione: (2024)
di: Zeger, Emi, et al.
Pubblicazione: (2024)
From Complexity to Clarity: Analytical Expressions of Deep Neural Network Weights via Clifford's Geometric Algebra and Convexity
di: Pilanci, Mert
Pubblicazione: (2023)
di: Pilanci, Mert
Pubblicazione: (2023)
Active Learning of Deep Neural Networks via Gradient-Free Cutting Planes
di: Zhang, Erica, et al.
Pubblicazione: (2024)
di: Zhang, Erica, et al.
Pubblicazione: (2024)
When Should Humans Step In? Optimal Human Dispatching in AI-Assisted Decisions
di: Tan, Lezhi, et al.
Pubblicazione: (2026)
di: Tan, Lezhi, et al.
Pubblicazione: (2026)
Lassoed Forests: Random Forests with Adaptive Lasso Post-selection
di: Shang, Jing, et al.
Pubblicazione: (2025)
di: Shang, Jing, et al.
Pubblicazione: (2025)
Convex Low-resource Accent-Robust Language Detection in Speech Recognition
di: Feng, Miria, et al.
Pubblicazione: (2026)
di: Feng, Miria, et al.
Pubblicazione: (2026)
Analyzing Neural Network-Based Generative Diffusion Models through Convex Optimization
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
Optimal Sets and Solution Paths of ReLU Networks
di: Mishkin, Aaron, et al.
Pubblicazione: (2023)
di: Mishkin, Aaron, et al.
Pubblicazione: (2023)
Large Language Models Implicitly Learn to See and Hear Just By Reading
di: Verma, Prateek, et al.
Pubblicazione: (2025)
di: Verma, Prateek, et al.
Pubblicazione: (2025)
Thinking While Listening: Simple Test Time Scaling For Audio Classification
di: Verma, Prateek, et al.
Pubblicazione: (2025)
di: Verma, Prateek, et al.
Pubblicazione: (2025)
Adaptive Large Language Models By Layerwise Attention Shortcuts
di: Verma, Prateek, et al.
Pubblicazione: (2024)
di: Verma, Prateek, et al.
Pubblicazione: (2024)
Convex Distillation: Efficient Compression of Deep Networks via Convex Optimization
di: Varshney, Prateek, et al.
Pubblicazione: (2024)
di: Varshney, Prateek, et al.
Pubblicazione: (2024)
Convex Optimization for Alignment and Preference Learning on a Single GPU
di: Feng, Miria, et al.
Pubblicazione: (2026)
di: Feng, Miria, et al.
Pubblicazione: (2026)
Towards Signal Processing In Large Language Models
di: Verma, Prateek, et al.
Pubblicazione: (2024)
di: Verma, Prateek, et al.
Pubblicazione: (2024)
Convex Relaxations of ReLU Neural Networks Approximate Global Optima in Polynomial Time
di: Kim, Sungyoon, et al.
Pubblicazione: (2024)
di: Kim, Sungyoon, et al.
Pubblicazione: (2024)
Optimal Shrinkage for Distributed Second-Order Optimization
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
Unveiling Hidden Convexity in Deep Learning: a Sparse Signal Processing Perspective
di: Zeger, Emi, et al.
Pubblicazione: (2026)
di: Zeger, Emi, et al.
Pubblicazione: (2026)
Black Boxes and Looking Glasses: Multilevel Symmetries, Reflection Planes, and Convex Optimization in Deep Networks
di: Zeger, Emi, et al.
Pubblicazione: (2024)
di: Zeger, Emi, et al.
Pubblicazione: (2024)
Riemannian Preconditioned LoRA for Fine-Tuning Foundation Models
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
Spectral Adapter: Fine-Tuning in Spectral Space
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2024)
A Recovery Guarantee for Sparse Neural Networks
di: Fridovich-Keil, Sara, et al.
Pubblicazione: (2025)
di: Fridovich-Keil, Sara, et al.
Pubblicazione: (2025)
A Family of LZ78-based Universal Sequential Probability Assignments
di: Sagan, Naomi, et al.
Pubblicazione: (2024)
di: Sagan, Naomi, et al.
Pubblicazione: (2024)
Optimizer-Induced Mode Connectivity: From AdamW to Muon
di: Zhang, Fangzhao, et al.
Pubblicazione: (2026)
di: Zhang, Fangzhao, et al.
Pubblicazione: (2026)
An Information-Theoretic Perspective on LLM Tokenizers
di: Erdogan, Mete, et al.
Pubblicazione: (2026)
di: Erdogan, Mete, et al.
Pubblicazione: (2026)
Gradient Coding with Iterative Block Leverage Score Sampling
di: Charalambides, Neophytos, et al.
Pubblicazione: (2023)
di: Charalambides, Neophytos, et al.
Pubblicazione: (2023)
Exploring the loss landscape of regularized neural networks via convex duality
di: Kim, Sungyoon, et al.
Pubblicazione: (2024)
di: Kim, Sungyoon, et al.
Pubblicazione: (2024)
Fast Convex Optimization for Two-Layer ReLU Networks: Equivalent Model Classes and Cone Decompositions
di: Mishkin, Aaron, et al.
Pubblicazione: (2022)
di: Mishkin, Aaron, et al.
Pubblicazione: (2022)
Newton Meets Marchenko-Pastur: Massively Parallel Second-Order Optimization with Hessian Sketching and Debiasing
di: Romanov, Elad, et al.
Pubblicazione: (2024)
di: Romanov, Elad, et al.
Pubblicazione: (2024)
Faster Convergence of Stochastic Accelerated Gradient Descent under Interpolation
di: Mishkin, Aaron, et al.
Pubblicazione: (2024)
di: Mishkin, Aaron, et al.
Pubblicazione: (2024)
AdaPTwin: Low-Cost Adaptive Compression of Product Twins in Transformers
di: Biju, Emil, et al.
Pubblicazione: (2024)
di: Biju, Emil, et al.
Pubblicazione: (2024)
Optimal Scalar Quantization for Matrix Multiplication: Closed-Form Density and Phase Transition
di: Ang, Calvin, et al.
Pubblicazione: (2026)
di: Ang, Calvin, et al.
Pubblicazione: (2026)
CRONOS: Enhancing Deep Learning with Scalable GPU Accelerated Convex Neural Networks
di: Feng, Miria, et al.
Pubblicazione: (2024)
di: Feng, Miria, et al.
Pubblicazione: (2024)
Adversarial Training of Two-Layer Polynomial and ReLU Activation Networks via Convex Optimization
di: Kuelbs, Daniel, et al.
Pubblicazione: (2024)
di: Kuelbs, Daniel, et al.
Pubblicazione: (2024)
FlashSketch: Sketch-Kernel Co-Design for Fast Sparse Sketching on GPUs
di: Dwaraknath, Rajat Vadiraj, et al.
Pubblicazione: (2026)
di: Dwaraknath, Rajat Vadiraj, et al.
Pubblicazione: (2026)
HuBERT-VIC: Improving Noise-Robust Automatic Speech Recognition of Speech Foundation Model via Variance-Invariance-Covariance Regularization
di: Ahn, Hyebin, et al.
Pubblicazione: (2025)
di: Ahn, Hyebin, et al.
Pubblicazione: (2025)
Adaptive Inference: Theoretical Limits and Unexplored Opportunities
di: Hor, Soheil, et al.
Pubblicazione: (2024)
di: Hor, Soheil, et al.
Pubblicazione: (2024)
Optimal Ridge Regularization for Out-of-Distribution Prediction
di: Patil, Pratik, et al.
Pubblicazione: (2024)
di: Patil, Pratik, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Learning When to Trust LLM Priors: A Validated Framework for Semantic Prior Integration
di: Zhang, Erica, et al.
Pubblicazione: (2026) -
Pretraining and the Lasso
di: Craig, Erin, et al.
Pubblicazione: (2024) -
Compressing Large Language Models using Low Rank and Low Precision Decomposition
di: Saha, Rajarshi, et al.
Pubblicazione: (2024) -
A Library of Mirrors: Deep Neural Nets in Low Dimensions are Convex Lasso Models with Reflection Features
di: Zeger, Emi, et al.
Pubblicazione: (2024) -
From Complexity to Clarity: Analytical Expressions of Deep Neural Network Weights via Clifford's Geometric Algebra and Convexity
di: Pilanci, Mert
Pubblicazione: (2023)