Saved in:
| Main Authors: | Zhu, Haoran, Majzoubi, Maryam, Jain, Arihant, Choromanska, Anna |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2210.03869 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Task-Level Contrastiveness for Cross-Domain Few-Shot Learning
by: Topollai, Kristi, et al.
Published: (2025)
by: Topollai, Kristi, et al.
Published: (2025)
Task-Agnostic Experts Composition for Continual Learning
by: Quarantiello, Luigi, et al.
Published: (2025)
by: Quarantiello, Luigi, et al.
Published: (2025)
Adaptive Memory Momentum via a Model-Based Framework for Deep Learning Optimization
by: Topollai, Kristi, et al.
Published: (2025)
by: Topollai, Kristi, et al.
Published: (2025)
Zero-Shot Cross-City Generalization in End-to-End Autonomous Driving: Self-Supervised versus Supervised Representations
by: Naeinian, Fatemeh, et al.
Published: (2026)
by: Naeinian, Fatemeh, et al.
Published: (2026)
Communication-Efficient Distributed Training for Collaborative Flat Optima Recovery in Deep Learning
by: Dimlioglu, Tolga, et al.
Published: (2025)
by: Dimlioglu, Tolga, et al.
Published: (2025)
Understanding Quantization of Optimizer States in LLM Pre-training: Dynamics of State Staleness and Effectiveness of State Resets
by: Topollai, Kristi, et al.
Published: (2026)
by: Topollai, Kristi, et al.
Published: (2026)
Split-on-Share: Mixture of Sparse Experts for Task-Agnostic Continual Learning
by: Siddika, Fatema, et al.
Published: (2026)
by: Siddika, Fatema, et al.
Published: (2026)
GRAWA: Gradient-based Weighted Averaging for Distributed Training of Deep Learning Models
by: Dimlioglu, Tolga, et al.
Published: (2024)
by: Dimlioglu, Tolga, et al.
Published: (2024)
A Survey of Optimization Methods for Training DL Models: Theoretical Perspective on Convergence and Generalization
by: Wang, Jing, et al.
Published: (2025)
by: Wang, Jing, et al.
Published: (2025)
Adjacent Leader Decentralized Stochastic Gradient Descent
by: He, Haoze, et al.
Published: (2024)
by: He, Haoze, et al.
Published: (2024)
$P^2$GNN: Two Prototype Sets to boost GNN Performance
by: Jain, Arihant, et al.
Published: (2026)
by: Jain, Arihant, et al.
Published: (2026)
Worker Disagreement Reveals Sharp Directions in Local SGD
by: Dimlioglu, Tolga, et al.
Published: (2026)
by: Dimlioglu, Tolga, et al.
Published: (2026)
Diversifying the Expert Knowledge for Task-Agnostic Pruning in Sparse Mixture-of-Experts
by: Zhang, Zeliang, et al.
Published: (2024)
by: Zhang, Zeliang, et al.
Published: (2024)
Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation
by: Hafez, Muhammad Burhan, et al.
Published: (2024)
by: Hafez, Muhammad Burhan, et al.
Published: (2024)
Self-Supervised JEPA-based World Models for LiDAR Occupancy Completion and Forecasting
by: Zhu, Haoran, et al.
Published: (2026)
by: Zhu, Haoran, et al.
Published: (2026)
Overcoming Growth-Induced Forgetting in Task-Agnostic Continual Learning
by: Zhao, Yuqing, et al.
Published: (2024)
by: Zhao, Yuqing, et al.
Published: (2024)
Optimal Task Order for Continual Learning of Multiple Tasks
by: Li, Ziyan, et al.
Published: (2025)
by: Li, Ziyan, et al.
Published: (2025)
Task-Agnostic Federated Continual Learning via Replay-Free Gradient Projection
by: Cha, Seohyeon, et al.
Published: (2025)
by: Cha, Seohyeon, et al.
Published: (2025)
Learning Task-Agnostic Motifs to Capture the Continuous Nature of Animal Behavior
by: Wang, Jiyi, et al.
Published: (2025)
by: Wang, Jiyi, et al.
Published: (2025)
Task-Agnostic Pre-training and Task-Guided Fine-tuning for Versatile Diffusion Planner
by: Fan, Chenyou, et al.
Published: (2024)
by: Fan, Chenyou, et al.
Published: (2024)
Outer-Momentum Restarting in High-Dimensional Two-Phase Optimization
by: Topollai, Kristi, et al.
Published: (2026)
by: Topollai, Kristi, et al.
Published: (2026)
TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking
by: Cheng, Yu, et al.
Published: (2026)
by: Cheng, Yu, et al.
Published: (2026)
TACOS: Task Agnostic Continual Learning in Spiking Neural Networks
by: Soures, Nicholas, et al.
Published: (2024)
by: Soures, Nicholas, et al.
Published: (2024)
GRID: Scalable Task-Agnostic Prompt-Based Continual Learning for Language Models
by: Tiwari, Anushka, et al.
Published: (2025)
by: Tiwari, Anushka, et al.
Published: (2025)
SLE-FNO: Single-Layer Extensions for Task-Agnostic Continual Learning in Fourier Neural Operators
by: Elhadidy, Mahmoud, et al.
Published: (2026)
by: Elhadidy, Mahmoud, et al.
Published: (2026)
FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning
by: Han, Xing, et al.
Published: (2026)
by: Han, Xing, et al.
Published: (2026)
OncoReason: Structuring Clinical Reasoning in LLMs for Robust and Interpretable Survival Prediction
by: Hemadri, Raghu Vamshi, et al.
Published: (2025)
by: Hemadri, Raghu Vamshi, et al.
Published: (2025)
Task-Agnostic Machine-Learning-Assisted Inference
by: Miao, Jiacheng, et al.
Published: (2024)
by: Miao, Jiacheng, et al.
Published: (2024)
Learning with Expert Abstractions for Efficient Multi-Task Continuous Control
by: Jewett, Jeff, et al.
Published: (2025)
by: Jewett, Jeff, et al.
Published: (2025)
T-TAME: Trainable Attention Mechanism for Explaining Convolutional Networks and Vision Transformers
by: Ntrougkas, Mariano V., et al.
Published: (2024)
by: Ntrougkas, Mariano V., et al.
Published: (2024)
Continual Learning at the Edge: An Agnostic IIoT Architecture
by: García-Santaclara, Pablo, et al.
Published: (2025)
by: García-Santaclara, Pablo, et al.
Published: (2025)
Task-Agnostic Contrastive Pretraining for Relational Deep Learning
by: Peleška, Jakub, et al.
Published: (2025)
by: Peleška, Jakub, et al.
Published: (2025)
Task-conditioned Ensemble of Expert Models for Continuous Learning
by: Sharma, Renu, et al.
Published: (2025)
by: Sharma, Renu, et al.
Published: (2025)
Learning Task-Agnostic Representations through Multi-Teacher Distillation
by: Formont, Philippe, et al.
Published: (2025)
by: Formont, Philippe, et al.
Published: (2025)
Principled Approaches for Learning to Defer with Multiple Experts
by: Mao, Anqi, et al.
Published: (2023)
by: Mao, Anqi, et al.
Published: (2023)
GRIP: Algorithm-Agnostic Machine Unlearning for Mixture-of-Experts via Geometric Router Constraints
by: Zhu, Andy, et al.
Published: (2026)
by: Zhu, Andy, et al.
Published: (2026)
A Unified Gradient-based Framework for Task-agnostic Continual Learning-Unlearning
by: Huang, Zhehao, et al.
Published: (2025)
by: Huang, Zhehao, et al.
Published: (2025)
DA-MoE: Towards Dynamic Expert Allocation for Mixture-of-Experts Models
by: Aghdam, Maryam Akhavan, et al.
Published: (2024)
by: Aghdam, Maryam Akhavan, et al.
Published: (2024)
Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts
by: Lou, Meng, et al.
Published: (2026)
by: Lou, Meng, et al.
Published: (2026)
U-TELL: Unsupervised Task Expert Lifelong Learning
by: Solomon, Indu, et al.
Published: (2024)
by: Solomon, Indu, et al.
Published: (2024)
Similar Items
-
Task-Level Contrastiveness for Cross-Domain Few-Shot Learning
by: Topollai, Kristi, et al.
Published: (2025) -
Task-Agnostic Experts Composition for Continual Learning
by: Quarantiello, Luigi, et al.
Published: (2025) -
Adaptive Memory Momentum via a Model-Based Framework for Deep Learning Optimization
by: Topollai, Kristi, et al.
Published: (2025) -
Zero-Shot Cross-City Generalization in End-to-End Autonomous Driving: Self-Supervised versus Supervised Representations
by: Naeinian, Fatemeh, et al.
Published: (2026) -
Communication-Efficient Distributed Training for Collaborative Flat Optima Recovery in Deep Learning
by: Dimlioglu, Tolga, et al.
Published: (2025)