GradMetaNet: An Equivariant Architecture for Learning on Gradients
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gelberg, Yoav, Eitan, Yam, Navon, Aviv, Shamsian, Aviv, Theo, Putterman, Bronstein, Michael, Maron, Haggai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Equivariant Deep Weight Space Alignment
von: Navon, Aviv, et al.
Veröffentlicht: (2023)
von: Navon, Aviv, et al.
Veröffentlicht: (2023)
Improved Generalization of Weight Space Networks via Augmentations
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024)
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024)
Training Transformers for KV Cache Compressibility
von: Gelberg, Yoav, et al.
Veröffentlicht: (2026)
von: Gelberg, Yoav, et al.
Veröffentlicht: (2026)
Learning on LoRAs: GL-Equivariant Processing of Low-Rank Weight Spaces for Large Finetuned Models
von: Putterman, Theo, et al.
Veröffentlicht: (2024)
von: Putterman, Theo, et al.
Veröffentlicht: (2024)
Go Beyond Your Means: Unlearning with Per-Sample Gradient Orthogonalization
von: Shamsian, Aviv, et al.
Veröffentlicht: (2025)
von: Shamsian, Aviv, et al.
Veröffentlicht: (2025)
Topological Blindspots: Understanding and Extending Topological Deep Learning Through the Lens of Expressivity
von: Eitan, Yam, et al.
Veröffentlicht: (2024)
von: Eitan, Yam, et al.
Veröffentlicht: (2024)
Whisper in Medusa's Ear: Multi-head Efficient Decoding for Transformer-based ASR
von: Segal-Feldman, Yael, et al.
Veröffentlicht: (2024)
von: Segal-Feldman, Yael, et al.
Veröffentlicht: (2024)
Multi Task Inverse Reinforcement Learning for Common Sense Reward
von: Glazer, Neta, et al.
Veröffentlicht: (2024)
von: Glazer, Neta, et al.
Veröffentlicht: (2024)
On the Expressive Power of Permutation-Equivariant Weight-Space Networks
von: Dayan, Adir, et al.
Veröffentlicht: (2026)
von: Dayan, Adir, et al.
Veröffentlicht: (2026)
PromptEvolver: Prompt Inversion through Evolutionary Optimization in Natural-Language Space
von: Buchnick, Asaf, et al.
Veröffentlicht: (2026)
von: Buchnick, Asaf, et al.
Veröffentlicht: (2026)
The Empirical Impact of Neural Parameter Symmetries, or Lack Thereof
von: Lim, Derek, et al.
Veröffentlicht: (2024)
von: Lim, Derek, et al.
Veröffentlicht: (2024)
On The Expressive Power of GNN Derivatives
von: Eitan, Yam, et al.
Veröffentlicht: (2025)
von: Eitan, Yam, et al.
Veröffentlicht: (2025)
Keyword-Guided Adaptation of Automatic Speech Recognition
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024)
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024)
A Flexible, Equivariant Framework for Subgraph GNNs via Graph Products and Graph Coarsening
von: Bar-Shalom, Guy, et al.
Veröffentlicht: (2024)
von: Bar-Shalom, Guy, et al.
Veröffentlicht: (2024)
FedSelect: Personalized Federated Learning with Customized Selection of Parameters for Fine-Tuning
von: Tamirisa, Rishub, et al.
Veröffentlicht: (2024)
von: Tamirisa, Rishub, et al.
Veröffentlicht: (2024)
WhisperNER: Unified Open Named Entity and Speech Recognition
von: Ayache, Gil, et al.
Veröffentlicht: (2024)
von: Ayache, Gil, et al.
Veröffentlicht: (2024)
Meta Reinforcement Learning with Finite Training Tasks -- a Density Estimation Approach
von: Rimon, Zohar, et al.
Veröffentlicht: (2022)
von: Rimon, Zohar, et al.
Veröffentlicht: (2022)
FS-KAN: Permutation Equivariant Kolmogorov-Arnold Networks via Function Sharing
von: Elbaz, Ran, et al.
Veröffentlicht: (2025)
von: Elbaz, Ran, et al.
Veröffentlicht: (2025)
Graph Metanetworks for Processing Diverse Neural Architectures
von: Lim, Derek, et al.
Veröffentlicht: (2023)
von: Lim, Derek, et al.
Veröffentlicht: (2023)
FlowTSE: Target Speaker Extraction with Flow Matching
von: Navon, Aviv, et al.
Veröffentlicht: (2025)
von: Navon, Aviv, et al.
Veröffentlicht: (2025)
Learning from Historical Activations in Graph Neural Networks
von: Galron, Yaniv, et al.
Veröffentlicht: (2026)
von: Galron, Yaniv, et al.
Veröffentlicht: (2026)
A Graph Meta-Network for Learning on Kolmogorov-Arnold Networks
von: Bar-Shalom, Guy, et al.
Veröffentlicht: (2026)
von: Bar-Shalom, Guy, et al.
Veröffentlicht: (2026)
A Comparison of Methods for Neural Network Aggregation
von: Pomerat, John, et al.
Veröffentlicht: (2023)
von: Pomerat, John, et al.
Veröffentlicht: (2023)
Balancing Efficiency and Expressiveness: Subgraph GNNs with Walk-Based Centrality
von: Southern, Joshua, et al.
Veröffentlicht: (2025)
von: Southern, Joshua, et al.
Veröffentlicht: (2025)
Drax: Speech Recognition with Discrete Flow Matching
von: Navon, Aviv, et al.
Veröffentlicht: (2025)
von: Navon, Aviv, et al.
Veröffentlicht: (2025)
Efficient GNN Training Through Structure-Aware Randomized Mini-Batching
von: Balaji, Vignesh, et al.
Veröffentlicht: (2025)
von: Balaji, Vignesh, et al.
Veröffentlicht: (2025)
GRANOLA: Adaptive Normalization for Graph Neural Networks
von: Eliasof, Moshe, et al.
Veröffentlicht: (2024)
von: Eliasof, Moshe, et al.
Veröffentlicht: (2024)
Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism
von: Bick, Aviv, et al.
Veröffentlicht: (2025)
von: Bick, Aviv, et al.
Veröffentlicht: (2025)
MoGU: Mixture-of-Gaussians with Uncertainty-based Gating for Time Series Forecasting
von: Aviv, Gilad, et al.
Veröffentlicht: (2025)
von: Aviv, Gilad, et al.
Veröffentlicht: (2025)
Understanding and Improving Laplacian Positional Encodings For Temporal GNNs
von: Galron, Yaniv, et al.
Veröffentlicht: (2025)
von: Galron, Yaniv, et al.
Veröffentlicht: (2025)
Retrieval-Aware Distillation for Transformer-SSM Hybrids
von: Bick, Aviv, et al.
Veröffentlicht: (2026)
von: Bick, Aviv, et al.
Veröffentlicht: (2026)
UmbraTTS: Adapting Text-to-Speech to Environmental Contexts with Flow Matching
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
GradINN: Gradient Informed Neural Network
von: Aglietti, Filippo, et al.
Veröffentlicht: (2024)
von: Aglietti, Filippo, et al.
Veröffentlicht: (2024)
GradTree: Learning Axis-Aligned Decision Trees with Gradient Descent
von: Marton, Sascha, et al.
Veröffentlicht: (2023)
von: Marton, Sascha, et al.
Veröffentlicht: (2023)
Future Directions in the Theory of Graph Machine Learning
von: Morris, Christopher, et al.
Veröffentlicht: (2024)
von: Morris, Christopher, et al.
Veröffentlicht: (2024)
GradMAP: Gradient-Based Multi-Agent Proximal Learning for Grid-Edge Flexibility
von: Zhou, Yihong, et al.
Veröffentlicht: (2026)
von: Zhou, Yihong, et al.
Veröffentlicht: (2026)
Beyond Transcription: Mechanistic Interpretability in ASR
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
von: Glazer, Neta, et al.
Veröffentlicht: (2025)
Llamba: Scaling Distilled Recurrent Models for Efficient Language Processing
von: Bick, Aviv, et al.
Veröffentlicht: (2025)
von: Bick, Aviv, et al.
Veröffentlicht: (2025)
Learning Probabilistic Symmetrization for Architecture Agnostic Equivariance
von: Kim, Jinwoo, et al.
Veröffentlicht: (2023)
von: Kim, Jinwoo, et al.
Veröffentlicht: (2023)
DenoGrad: A Gradient-Based Framework for Data Refinement in Tabular and Time-Series Learning
von: Alonso-Ramos, J. Javier, et al.
Veröffentlicht: (2025)
von: Alonso-Ramos, J. Javier, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Equivariant Deep Weight Space Alignment
von: Navon, Aviv, et al.
Veröffentlicht: (2023) -
Improved Generalization of Weight Space Networks via Augmentations
von: Shamsian, Aviv, et al.
Veröffentlicht: (2024) -
Training Transformers for KV Cache Compressibility
von: Gelberg, Yoav, et al.
Veröffentlicht: (2026) -
Learning on LoRAs: GL-Equivariant Processing of Low-Rank Weight Spaces for Large Finetuned Models
von: Putterman, Theo, et al.
Veröffentlicht: (2024) -
Go Beyond Your Means: Unlearning with Per-Sample Gradient Orthogonalization
von: Shamsian, Aviv, et al.
Veröffentlicht: (2025)