PreND: Enhancing Intrinsic Motivation in Reinforcement Learning through Pre-trained Network Distillation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Davoodabadi, Mohammadamin, Dijujin, Negin Hashemi, Baghshah, Mahdieh Soleymani |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CAREL: Instruction-guided reinforcement learning with cross-modal auxiliary objectives
von: Saghafian, Armin, et al.
Veröffentlicht: (2024)
von: Saghafian, Armin, et al.
Veröffentlicht: (2024)
Inductive Biases for Zero-shot Systematic Generalization in Language-informed Reinforcement Learning
von: Dijujin, Negin Hashemi, et al.
Veröffentlicht: (2025)
von: Dijujin, Negin Hashemi, et al.
Veröffentlicht: (2025)
SUSD: Structured Unsupervised Skill Discovery through State Factorization
von: Hosseini, Seyed Mohammad Hadi, et al.
Veröffentlicht: (2026)
von: Hosseini, Seyed Mohammad Hadi, et al.
Veröffentlicht: (2026)
CER: Confidence Enhanced Reasoning in LLMs
von: Razghandi, Ali, et al.
Veröffentlicht: (2025)
von: Razghandi, Ali, et al.
Veröffentlicht: (2025)
Dilated Balanced Cross Entropy Loss for Medical Image Segmentation
von: Hosseini, Seyed Mohsen, et al.
Veröffentlicht: (2024)
von: Hosseini, Seyed Mohsen, et al.
Veröffentlicht: (2024)
Classification of Breast Cancer Histopathology Images using a Modified Supervised Contrastive Learning Method
von: Sani, Matina Mahdizadeh, et al.
Veröffentlicht: (2024)
von: Sani, Matina Mahdizadeh, et al.
Veröffentlicht: (2024)
Efficient Adversarial Attacks on High-dimensional Offline Bandits
von: Hosseini, Seyed Mohammad Hadi, et al.
Veröffentlicht: (2026)
von: Hosseini, Seyed Mohammad Hadi, et al.
Veröffentlicht: (2026)
LibraGrad: Balancing Gradient Flow for Universally Better Vision Transformer Attributions
von: Mehri, Faridoun, et al.
Veröffentlicht: (2024)
von: Mehri, Faridoun, et al.
Veröffentlicht: (2024)
Language Plays a Pivotal Role in the Object-Attribute Compositional Generalization of CLIP
von: Abbasi, Reza, et al.
Veröffentlicht: (2024)
von: Abbasi, Reza, et al.
Veröffentlicht: (2024)
Limits and Gains of Test-Time Scaling in Vision-Language Reasoning
von: Ahmadpour, Mohammadjavad, et al.
Veröffentlicht: (2025)
von: Ahmadpour, Mohammadjavad, et al.
Veröffentlicht: (2025)
Improving 3D Few-Shot Segmentation with Inference-Time Pseudo-Labeling
von: Mozafari, Mohammad, et al.
Veröffentlicht: (2024)
von: Mozafari, Mohammad, et al.
Veröffentlicht: (2024)
PEAC: Unsupervised Pre-training for Cross-Embodiment Reinforcement Learning
von: Ying, Chengyang, et al.
Veröffentlicht: (2024)
von: Ying, Chengyang, et al.
Veröffentlicht: (2024)
Bridging Reasoning to Learning: Unmasking Illusions using Complexity Out of Distribution Generalization
von: Paqaleh, Mohammad Mahdi Samiei, et al.
Veröffentlicht: (2025)
von: Paqaleh, Mohammad Mahdi Samiei, et al.
Veröffentlicht: (2025)
HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2026)
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2026)
Can Distillation Mitigate Backdoor Attacks in Pre-trained Encoders?
von: Han, TIngxu, et al.
Veröffentlicht: (2024)
von: Han, TIngxu, et al.
Veröffentlicht: (2024)
Unleashing the Power of Pre-trained Language Models for Offline Reinforcement Learning
von: Shi, Ruizhe, et al.
Veröffentlicht: (2023)
von: Shi, Ruizhe, et al.
Veröffentlicht: (2023)
Pre-training with Synthetic Data Helps Offline Reinforcement Learning
von: Wang, Zecheng, et al.
Veröffentlicht: (2023)
von: Wang, Zecheng, et al.
Veröffentlicht: (2023)
Cross-domain Random Pre-training with Prototypes for Reinforcement Learning
von: Liu, Xin, et al.
Veröffentlicht: (2023)
von: Liu, Xin, et al.
Veröffentlicht: (2023)
Decompose-and-Compose: A Compositional Approach to Mitigating Spurious Correlation
von: Noohdani, Fahimeh Hosseini, et al.
Veröffentlicht: (2024)
von: Noohdani, Fahimeh Hosseini, et al.
Veröffentlicht: (2024)
Spurious-Aware Prototype Refinement for Reliable Out-of-Distribution Detection
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2025)
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2025)
RLeXplore: Accelerating Research in Intrinsically-Motivated Reinforcement Learning
von: Yuan, Mingqi, et al.
Veröffentlicht: (2024)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2024)
Enhanced Pre-training of Graph Neural Networks for Million-Scale Heterogeneous Graphs
von: Sun, Shengyin, et al.
Veröffentlicht: (2025)
von: Sun, Shengyin, et al.
Veröffentlicht: (2025)
Surprise-Adaptive Intrinsic Motivation for Unsupervised Reinforcement Learning
von: Hugessen, Adriana, et al.
Veröffentlicht: (2024)
von: Hugessen, Adriana, et al.
Veröffentlicht: (2024)
Stabilizing RNN Gradients through Pre-training
von: Herranz-Celotti, Luca, et al.
Veröffentlicht: (2023)
von: Herranz-Celotti, Luca, et al.
Veröffentlicht: (2023)
SOInter: A Novel Deep Energy Based Interpretation Method for Explaining Structured Output Models
von: Seyyedsalehi, S. Fatemeh, et al.
Veröffentlicht: (2022)
von: Seyyedsalehi, S. Fatemeh, et al.
Veröffentlicht: (2022)
Pre-trained Language Model and Knowledge Distillation for Lightweight Sequential Recommendation
von: Li, Li, et al.
Veröffentlicht: (2024)
von: Li, Li, et al.
Veröffentlicht: (2024)
Privacy Backdoors: Enhancing Membership Inference through Poisoning Pre-trained Models
von: Wen, Yuxin, et al.
Veröffentlicht: (2024)
von: Wen, Yuxin, et al.
Veröffentlicht: (2024)
Fostering Intrinsic Motivation in Reinforcement Learning with Pretrained Foundation Models
von: Andres, Alain, et al.
Veröffentlicht: (2024)
von: Andres, Alain, et al.
Veröffentlicht: (2024)
TrajGPT-R: Generating Urban Mobility Trajectory with Reinforcement Learning-Enhanced Generative Pre-trained Transformer
von: Wang, Jiawei, et al.
Veröffentlicht: (2026)
von: Wang, Jiawei, et al.
Veröffentlicht: (2026)
MLKD-BERT: Multi-level Knowledge Distillation for Pre-trained Language Models
von: Zhang, Ying, et al.
Veröffentlicht: (2024)
von: Zhang, Ying, et al.
Veröffentlicht: (2024)
LLM-Driven Intrinsic Motivation for Sparse Reward Reinforcement Learning
von: Quadros, André, et al.
Veröffentlicht: (2025)
von: Quadros, André, et al.
Veröffentlicht: (2025)
Fine-Grained Alignment and Noise Refinement for Compositional Text-to-Image Generation
von: Izadi, Amir Mohammad, et al.
Veröffentlicht: (2025)
von: Izadi, Amir Mohammad, et al.
Veröffentlicht: (2025)
Pre-training under infinite compute
von: Kim, Konwoo, et al.
Veröffentlicht: (2025)
von: Kim, Konwoo, et al.
Veröffentlicht: (2025)
PreMixer: MLP-Based Pre-training Enhanced MLP-Mixers for Large-scale Traffic Forecasting
von: Zhang, Tongtong, et al.
Veröffentlicht: (2024)
von: Zhang, Tongtong, et al.
Veröffentlicht: (2024)
Towards a General Framework for Continual Learning with Pre-training
von: Wang, Liyuan, et al.
Veröffentlicht: (2023)
von: Wang, Liyuan, et al.
Veröffentlicht: (2023)
Learning to Unlearn: Instance-wise Unlearning for Pre-trained Classifiers
von: Cha, Sungmin, et al.
Veröffentlicht: (2023)
von: Cha, Sungmin, et al.
Veröffentlicht: (2023)
Combining Pre-Trained Models for Enhanced Feature Representation in Reinforcement Learning
von: Piccoli, Elia, et al.
Veröffentlicht: (2025)
von: Piccoli, Elia, et al.
Veröffentlicht: (2025)
Deep Fusion: Efficient Network Training via Pre-trained Initializations
von: Mazzawi, Hanna, et al.
Veröffentlicht: (2023)
von: Mazzawi, Hanna, et al.
Veröffentlicht: (2023)
Text Classification: Neural Networks VS Machine Learning Models VS Pre-trained Models
von: Petridis, Christos
Veröffentlicht: (2024)
von: Petridis, Christos
Veröffentlicht: (2024)
Thinking Augmented Pre-training
von: Wang, Liang, et al.
Veröffentlicht: (2025)
von: Wang, Liang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
CAREL: Instruction-guided reinforcement learning with cross-modal auxiliary objectives
von: Saghafian, Armin, et al.
Veröffentlicht: (2024) -
Inductive Biases for Zero-shot Systematic Generalization in Language-informed Reinforcement Learning
von: Dijujin, Negin Hashemi, et al.
Veröffentlicht: (2025) -
SUSD: Structured Unsupervised Skill Discovery through State Factorization
von: Hosseini, Seyed Mohammad Hadi, et al.
Veröffentlicht: (2026) -
CER: Confidence Enhanced Reasoning in LLMs
von: Razghandi, Ali, et al.
Veröffentlicht: (2025) -
Dilated Balanced Cross Entropy Loss for Medical Image Segmentation
von: Hosseini, Seyed Mohsen, et al.
Veröffentlicht: (2024)