MERMAIDE: Learning to Align Learners using Model-Based Meta-Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Banerjee, Arundhati, Phade, Soham, Ermon, Stefano, Zheng, Stephan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SequenceMatch: Imitation Learning for Autoregressive Sequence Modelling with Backtracking
por: Cundy, Chris, et al.
Publicado: (2023)
por: Cundy, Chris, et al.
Publicado: (2023)
Learning Neural PDE Solvers with Convergence Guarantees
por: Hsieh, Jun-Ting, et al.
Publicado: (2019)
por: Hsieh, Jun-Ting, et al.
Publicado: (2019)
Reviving Any-Subset Autoregressive Models with Principled Parallel Sampling and Speculative Decoding
por: Guo, Gabe, et al.
Publicado: (2025)
por: Guo, Gabe, et al.
Publicado: (2025)
BayPrAnoMeta: Bayesian Proto-MAML for Few-Shot Industrial Image Anomaly Detection
por: Sarkar, Soham, et al.
Publicado: (2026)
por: Sarkar, Soham, et al.
Publicado: (2026)
Bellman Diffusion: Generative Modeling as Learning a Linear Operator in the Distribution Space
por: Li, Yangming, et al.
Publicado: (2024)
por: Li, Yangming, et al.
Publicado: (2024)
Trust Region Continual Learning as an Implicit Meta-Learner
por: Wang, Zekun, et al.
Publicado: (2026)
por: Wang, Zekun, et al.
Publicado: (2026)
Decentralized Multi-Agent Active Search and Tracking when Targets Outnumber Agents
por: Banerjee, Arundhati, et al.
Publicado: (2024)
por: Banerjee, Arundhati, et al.
Publicado: (2024)
Discrete Diffusion Modeling by Estimating the Ratios of the Data Distribution
por: Lou, Aaron, et al.
Publicado: (2023)
por: Lou, Aaron, et al.
Publicado: (2023)
Align Your Structures: Generating Trajectories with Structure Pretraining for Molecular Dynamics
por: Iyengar, Aniketh, et al.
Publicado: (2026)
por: Iyengar, Aniketh, et al.
Publicado: (2026)
Cost-Aware Diffusion Active Search
por: Banerjee, Arundhati, et al.
Publicado: (2026)
por: Banerjee, Arundhati, et al.
Publicado: (2026)
Data Unlearning in Diffusion Models
por: Alberti, Silas, et al.
Publicado: (2025)
por: Alberti, Silas, et al.
Publicado: (2025)
Principled Fast and Meta Knowledge Learners for Continual Reinforcement Learning
por: Sun, Ke, et al.
Publicado: (2026)
por: Sun, Ke, et al.
Publicado: (2026)
Active Learning for Derivative-Based Global Sensitivity Analysis with Gaussian Processes
por: Belakaria, Syrine, et al.
Publicado: (2024)
por: Belakaria, Syrine, et al.
Publicado: (2024)
Probabilistic Graphical Models: A Concise Tutorial
por: Maasch, Jacqueline, et al.
Publicado: (2025)
por: Maasch, Jacqueline, et al.
Publicado: (2025)
Generative Modeling with Flux Matching
por: Pao-Huang, Peter, et al.
Publicado: (2026)
por: Pao-Huang, Peter, et al.
Publicado: (2026)
Calibrated Probabilistic Forecasts for Arbitrary Sequences
por: Marx, Charles, et al.
Publicado: (2024)
por: Marx, Charles, et al.
Publicado: (2024)
Learning Mamba as a Continual Learner: Meta-learning Selective State Space Models for Efficient Continual Learning
por: Zhao, Chongyang, et al.
Publicado: (2024)
por: Zhao, Chongyang, et al.
Publicado: (2024)
Newton Losses: Using Curvature Information for Learning with Differentiable Algorithms
por: Petersen, Felix, et al.
Publicado: (2024)
por: Petersen, Felix, et al.
Publicado: (2024)
CMT: Mid-Training for Efficient Learning of Consistency, Mean Flow, and Flow Map Models
por: Hu, Zheyuan, et al.
Publicado: (2025)
por: Hu, Zheyuan, et al.
Publicado: (2025)
Aligning Target-Aware Molecule Diffusion Models with Exact Energy Optimization
por: Gu, Siyi, et al.
Publicado: (2024)
por: Gu, Siyi, et al.
Publicado: (2024)
RelDiff: Relational Data Generative Modeling with Graph-Based Diffusion Models
por: Hudovernik, Valter, et al.
Publicado: (2025)
por: Hudovernik, Valter, et al.
Publicado: (2025)
MADiff: Offline Multi-agent Learning with Diffusion Models
por: Zhu, Zhengbang, et al.
Publicado: (2023)
por: Zhu, Zhengbang, et al.
Publicado: (2023)
Privacy-Constrained Policies via Mutual Information Regularized Policy Gradients
por: Cundy, Chris, et al.
Publicado: (2020)
por: Cundy, Chris, et al.
Publicado: (2020)
Inductive Moment Matching
por: Zhou, Linqi, et al.
Publicado: (2025)
por: Zhou, Linqi, et al.
Publicado: (2025)
Transferable SCF-Acceleration through Solver-Aligned Initialization Learning
por: Eberhard, Eike S., et al.
Publicado: (2026)
por: Eberhard, Eike S., et al.
Publicado: (2026)
RFG: Test-Time Scaling for Diffusion Large Language Model Reasoning with Reward-Free Guidance
por: Chen, Tianlang, et al.
Publicado: (2025)
por: Chen, Tianlang, et al.
Publicado: (2025)
LLMs as In-Context Meta-Learners for Model and Hyperparameter Selection
por: Hili, Youssef Attia El, et al.
Publicado: (2025)
por: Hili, Youssef Attia El, et al.
Publicado: (2025)
Generalizing Stochastic Smoothing for Differentiation and Gradient Estimation
por: Petersen, Felix, et al.
Publicado: (2024)
por: Petersen, Felix, et al.
Publicado: (2024)
Meta-Learning Approaches for Speaker-Dependent Voice Fatigue Models
por: Polle, Roseline, et al.
Publicado: (2025)
por: Polle, Roseline, et al.
Publicado: (2025)
MetaCD: A Meta Learning Framework for Cognitive Diagnosis based on Continual Learning
por: Wu, Jin, et al.
Publicado: (2025)
por: Wu, Jin, et al.
Publicado: (2025)
Collaborative Cognitive Diagnosis with Disentangled Representation Learning for Learner Modeling
por: Gao, Weibo, et al.
Publicado: (2024)
por: Gao, Weibo, et al.
Publicado: (2024)
Uncertainty Quantification for Forward and Inverse Problems of PDEs via Latent Global Evolution
por: Wu, Tailin, et al.
Publicado: (2024)
por: Wu, Tailin, et al.
Publicado: (2024)
Adaptive Meta-Learning for Identification of Rover-Terrain Dynamics
por: Banerjee, S., et al.
Publicado: (2020)
por: Banerjee, S., et al.
Publicado: (2020)
Self-Refining Diffusion Samplers: Enabling Parallelization via Parareal Iterations
por: Selvam, Nikil Roashan, et al.
Publicado: (2024)
por: Selvam, Nikil Roashan, et al.
Publicado: (2024)
TrAct: Making First-layer Pre-Activations Trainable
por: Petersen, Felix, et al.
Publicado: (2024)
por: Petersen, Felix, et al.
Publicado: (2024)
Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation
por: Manvi, Rohin, et al.
Publicado: (2024)
por: Manvi, Rohin, et al.
Publicado: (2024)
Beyond Pairwise Preferences: Listwise Reward-Aware Alignment for Diffusion Models
por: Wang, Austin, et al.
Publicado: (2026)
por: Wang, Austin, et al.
Publicado: (2026)
Stronger Baseline Models -- A Key Requirement for Aligning Machine Learning Research with Clinical Utility
por: Wolfrath, Nathan, et al.
Publicado: (2024)
por: Wolfrath, Nathan, et al.
Publicado: (2024)
Adversarial Attacks on Graph Neural Networks via Meta Learning
por: Zügner, Daniel, et al.
Publicado: (2019)
por: Zügner, Daniel, et al.
Publicado: (2019)
Mochi: Aligning Pre-training and Inference for Efficient Graph Foundation Models via Meta-Learning
por: Mattos, João, et al.
Publicado: (2026)
por: Mattos, João, et al.
Publicado: (2026)
Ejemplares similares
-
SequenceMatch: Imitation Learning for Autoregressive Sequence Modelling with Backtracking
por: Cundy, Chris, et al.
Publicado: (2023) -
Learning Neural PDE Solvers with Convergence Guarantees
por: Hsieh, Jun-Ting, et al.
Publicado: (2019) -
Reviving Any-Subset Autoregressive Models with Principled Parallel Sampling and Speculative Decoding
por: Guo, Gabe, et al.
Publicado: (2025) -
BayPrAnoMeta: Bayesian Proto-MAML for Few-Shot Industrial Image Anomaly Detection
por: Sarkar, Soham, et al.
Publicado: (2026) -
Bellman Diffusion: Generative Modeling as Learning a Linear Operator in the Distribution Space
por: Li, Yangming, et al.
Publicado: (2024)