Learning from Teaching Regularization: Generalizable Correlations Should be Easy to Imitate
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jin, Can, Che, Tong, Peng, Hongwu, Li, Yiyuan, Metaxas, Dimitris N., Pavone, Marco |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Your Reward Function for RL is Your Best PRM for Search: Unifying RL and Search-Based TTS
von: Jin, Can, et al.
Veröffentlicht: (2025)
von: Jin, Can, et al.
Veröffentlicht: (2025)
Two Heads are Better Than One: Test-time Scaling of Multi-agent Collaborative Reasoning
von: Jin, Can, et al.
Veröffentlicht: (2025)
von: Jin, Can, et al.
Veröffentlicht: (2025)
MILES: Making Imitation Learning Easy with Self-Supervision
von: Papagiannis, Georgios, et al.
Veröffentlicht: (2024)
von: Papagiannis, Georgios, et al.
Veröffentlicht: (2024)
DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
von: Zhou, Yang, et al.
Veröffentlicht: (2026)
RuleFuser: An Evidential Bayes Approach for Rule Injection in Imitation Learned Planners and Predictors for Robustness under Distribution Shifts
von: Patrikar, Jay, et al.
Veröffentlicht: (2024)
von: Patrikar, Jay, et al.
Veröffentlicht: (2024)
Test-Time Spectrum-Aware Latent Steering for Zero-Shot Generalization in Vision-Language Models
von: Dafnis, Konstantinos M., et al.
Veröffentlicht: (2025)
von: Dafnis, Konstantinos M., et al.
Veröffentlicht: (2025)
DIGIC: Domain Generalizable Imitation Learning by Causal Discovery
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
RIZE: Adaptive Regularization for Imitation Learning
von: Karimi, Adib, et al.
Veröffentlicht: (2025)
von: Karimi, Adib, et al.
Veröffentlicht: (2025)
Learning Multiple Initial Solutions to Optimization Problems
von: Sharony, Elad, et al.
Veröffentlicht: (2024)
von: Sharony, Elad, et al.
Veröffentlicht: (2024)
Constrained Sampling for Language Models Should Be Easy: An MCMC Perspective
von: Gonzalez, Emmanuel Anaya, et al.
Veröffentlicht: (2025)
von: Gonzalez, Emmanuel Anaya, et al.
Veröffentlicht: (2025)
Self-Improving Transformers Overcome Easy-to-Hard and Length Generalization Challenges
von: Lee, Nayoung, et al.
Veröffentlicht: (2025)
von: Lee, Nayoung, et al.
Veröffentlicht: (2025)
Deep Learning Warm Starts for Trajectory Optimization on the International Space Station
von: Banerjee, Somrita, et al.
Veröffentlicht: (2025)
von: Banerjee, Somrita, et al.
Veröffentlicht: (2025)
Denoising-based Contractive Imitation Learning
von: Shen, Macheng, et al.
Veröffentlicht: (2025)
von: Shen, Macheng, et al.
Veröffentlicht: (2025)
Diverse Policies Recovering via Pointwise Mutual Information Weighted Imitation Learning
von: Yang, Hanlin, et al.
Veröffentlicht: (2024)
von: Yang, Hanlin, et al.
Veröffentlicht: (2024)
Implicit In-context Learning
von: Li, Zhuowei, et al.
Veröffentlicht: (2024)
von: Li, Zhuowei, et al.
Veröffentlicht: (2024)
GABRIL: Gaze-Based Regularization for Mitigating Causal Confusion in Imitation Learning
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)
Imitating from auxiliary imperfect demonstrations via Adversarial Density Weighted Regression
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024)
von: Zhang, Ziqi, et al.
Veröffentlicht: (2024)
Unsupervisedly Learned Representations: Should the Quest be Over?
von: Nissani, Daniel N.
Veröffentlicht: (2020)
von: Nissani, Daniel N.
Veröffentlicht: (2020)
Shorter but not Worse: Frugal Reasoning via Easy Samples as Length Regularizers in Math RLVR
von: Bounhar, Abdelaziz, et al.
Veröffentlicht: (2025)
von: Bounhar, Abdelaziz, et al.
Veröffentlicht: (2025)
Confounded Causal Imitation Learning with Instrumental Variables
von: Zeng, Yan, et al.
Veröffentlicht: (2025)
von: Zeng, Yan, et al.
Veröffentlicht: (2025)
BLoB: Bayesian Low-Rank Adaptation by Backpropagation for Large Language Models
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
von: Wang, Yibin, et al.
Veröffentlicht: (2024)
Deep Reinforcement Learning for Artificial Upwelling Energy Management
von: Zhang, Yiyuan, et al.
Veröffentlicht: (2023)
von: Zhang, Yiyuan, et al.
Veröffentlicht: (2023)
Imitation Learning via Focused Satisficing
von: Shah, Rushit N., et al.
Veröffentlicht: (2025)
von: Shah, Rushit N., et al.
Veröffentlicht: (2025)
Anatomy-VLM: A Fine-grained Vision-Language Model for Medical Interpretation
von: Gu, Difei, et al.
Veröffentlicht: (2025)
von: Gu, Difei, et al.
Veröffentlicht: (2025)
SCL-GNN: Towards Generalizable Graph Neural Networks via Spurious Correlation Learning
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2026)
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2026)
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual Steering
von: Chatzoudis, Gerasimos, et al.
Veröffentlicht: (2025)
von: Chatzoudis, Gerasimos, et al.
Veröffentlicht: (2025)
PateGail: A Privacy-Preserving Mobility Trajectory Generator with Imitation Learning
von: Wang, Huandong, et al.
Veröffentlicht: (2024)
von: Wang, Huandong, et al.
Veröffentlicht: (2024)
Hybrid Imitation-Learning Motion Planner for Urban Driving
von: Gariboldi, Cristian, et al.
Veröffentlicht: (2024)
von: Gariboldi, Cristian, et al.
Veröffentlicht: (2024)
Imitation Bootstrapped Reinforcement Learning
von: Hu, Hengyuan, et al.
Veröffentlicht: (2023)
von: Hu, Hengyuan, et al.
Veröffentlicht: (2023)
Quantifying Generalisation in Imitation Learning
von: Gavenski, Nathan, et al.
Veröffentlicht: (2025)
von: Gavenski, Nathan, et al.
Veröffentlicht: (2025)
Robot-Gated Interactive Imitation Learning with Adaptive Intervention Mechanism
von: Cai, Haoyuan, et al.
Veröffentlicht: (2025)
von: Cai, Haoyuan, et al.
Veröffentlicht: (2025)
PR2: Predictive Routing Replay for MoE-Based LLM Reinforcement Learning
von: Dong, Daize, et al.
Veröffentlicht: (2026)
von: Dong, Daize, et al.
Veröffentlicht: (2026)
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
von: Xu, Haoran, et al.
Veröffentlicht: (2025)
von: Xu, Haoran, et al.
Veröffentlicht: (2025)
Offline Imitation Learning with Model-based Reverse Augmentation
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
von: Shao, Jie-Jing, et al.
Veröffentlicht: (2024)
Adversarial Imitation Learning via Boosting
von: Chang, Jonathan D., et al.
Veröffentlicht: (2024)
von: Chang, Jonathan D., et al.
Veröffentlicht: (2024)
Noise-Guided Transport for Imitation Learning
von: Blondé, Lionel, et al.
Veröffentlicht: (2025)
von: Blondé, Lionel, et al.
Veröffentlicht: (2025)
Sample-efficient Adversarial Imitation Learning
von: Jung, Dahuin, et al.
Veröffentlicht: (2023)
von: Jung, Dahuin, et al.
Veröffentlicht: (2023)
Boolean Satisfiability via Imitation Learning
von: Zhang, Zewei, et al.
Veröffentlicht: (2025)
von: Zhang, Zewei, et al.
Veröffentlicht: (2025)
Imitation Learning as Return Distribution Matching
von: Lazzati, Filippo, et al.
Veröffentlicht: (2025)
von: Lazzati, Filippo, et al.
Veröffentlicht: (2025)
RadAlign: Advancing Radiology Report Generation with Vision-Language Concept Alignment
von: Gu, Difei, et al.
Veröffentlicht: (2025)
von: Gu, Difei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Your Reward Function for RL is Your Best PRM for Search: Unifying RL and Search-Based TTS
von: Jin, Can, et al.
Veröffentlicht: (2025) -
Two Heads are Better Than One: Test-time Scaling of Multi-agent Collaborative Reasoning
von: Jin, Can, et al.
Veröffentlicht: (2025) -
MILES: Making Imitation Learning Easy with Self-Supervision
von: Papagiannis, Georgios, et al.
Veröffentlicht: (2024) -
DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation
von: Zhou, Yang, et al.
Veröffentlicht: (2026) -
RuleFuser: An Evidential Bayes Approach for Rule Injection in Imitation Learned Planners and Predictors for Robustness under Distribution Shifts
von: Patrikar, Jay, et al.
Veröffentlicht: (2024)