Predictive auxiliary objectives in deep RL mimic learning in the brain
Fuente:
arXiv
Saved in:
| Main Authors: | Fang, Ching, Stachenfeld, Kimberly L |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CAREL: Instruction-guided reinforcement learning with cross-modal auxiliary objectives
by: Saghafian, Armin, et al.
Published: (2024)
by: Saghafian, Armin, et al.
Published: (2024)
Scale-specific auxiliary multi-task contrastive learning for deep liver vessel segmentation
by: Sadikine, Amine, et al.
Published: (2024)
by: Sadikine, Amine, et al.
Published: (2024)
Permutative redundancy and uncertainty of the objective in deep learning
by: Glukhov, Vacslav
Published: (2024)
by: Glukhov, Vacslav
Published: (2024)
How does the primate brain combine generative and discriminative computations in vision?
by: Peters, Benjamin, et al.
Published: (2024)
by: Peters, Benjamin, et al.
Published: (2024)
Can AI mimic the human ability to define neologisms?
by: Georgiou, Georgios P.
Published: (2025)
by: Georgiou, Georgios P.
Published: (2025)
From Memories to Maps: Mechanisms of In-Context Reinforcement Learning in Transformers
by: Fang, Ching, et al.
Published: (2025)
by: Fang, Ching, et al.
Published: (2025)
Unsupervised decoding of encoded reasoning using language model interpretability
by: Fang, Ching, et al.
Published: (2025)
by: Fang, Ching, et al.
Published: (2025)
EARL: Entropy-Aware RL Alignment of LLMs for Reliable RTL Code Generation
by: Shi, Jiahe, et al.
Published: (2025)
by: Shi, Jiahe, et al.
Published: (2025)
Cross-Language Speaker Attribute Prediction Using MIL and RL
by: Shu, Sunny, et al.
Published: (2026)
by: Shu, Sunny, et al.
Published: (2026)
Agent^2 RL-Bench: Can LLM Agents Engineer Agentic RL Post-Training?
by: Chen, Wanyi, et al.
Published: (2026)
by: Chen, Wanyi, et al.
Published: (2026)
Yes, Q-learning Helps Offline In-Context RL
by: Tarasov, Denis, et al.
Published: (2025)
by: Tarasov, Denis, et al.
Published: (2025)
High-dimensional multiple imputation (HDMI) for partially observed confounders including natural language processing-derived auxiliary covariates
by: Weberpals, Janick, et al.
Published: (2024)
by: Weberpals, Janick, et al.
Published: (2024)
Improving action classification with brain-inspired deep networks
by: Aglinskas, Aidas, et al.
Published: (2025)
by: Aglinskas, Aidas, et al.
Published: (2025)
Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL
by: Sokar, Ghada, et al.
Published: (2024)
by: Sokar, Ghada, et al.
Published: (2024)
Explainable concept mappings of MRI: Revealing the mechanisms underlying deep learning-based brain disease classification
by: Tinauer, Christian, et al.
Published: (2024)
by: Tinauer, Christian, et al.
Published: (2024)
Memorization in deep learning: A survey
by: Wei, Jiaheng, et al.
Published: (2024)
by: Wei, Jiaheng, et al.
Published: (2024)
Contrastive learning-based agent modeling for deep reinforcement learning
by: Ma, Wenhao, et al.
Published: (2023)
by: Ma, Wenhao, et al.
Published: (2023)
NiceWebRL: a Python library for human subject experiments with reinforcement learning environments
by: Carvalho, Wilka, et al.
Published: (2025)
by: Carvalho, Wilka, et al.
Published: (2025)
Generalizability analysis of deep learning predictions of human brain responses to augmented and semantically novel visual stimuli
by: Piskovskyi, Valentyn, et al.
Published: (2024)
by: Piskovskyi, Valentyn, et al.
Published: (2024)
Imitating from auxiliary imperfect demonstrations via Adversarial Density Weighted Regression
by: Zhang, Ziqi, et al.
Published: (2024)
by: Zhang, Ziqi, et al.
Published: (2024)
Forager: a lightweight testbed for continual learning with partial observability in RL
by: Tang, Steven, et al.
Published: (2026)
by: Tang, Steven, et al.
Published: (2026)
Found-RL: foundation model-enhanced reinforcement learning for autonomous driving
by: Qu, Yansong, et al.
Published: (2026)
by: Qu, Yansong, et al.
Published: (2026)
Pediatric brain tumor classification using digital histopathology and deep learning: evaluation of SOTA methods on a multi-center Swedish cohort
by: Tampu, Iulian Emil, et al.
Published: (2024)
by: Tampu, Iulian Emil, et al.
Published: (2024)
Adaptive traffic signal safety and efficiency improvement by multi objective deep reinforcement learning approach
by: Mirbakhsh, Shahin, et al.
Published: (2024)
by: Mirbakhsh, Shahin, et al.
Published: (2024)
Bridging State and History Representations: Understanding Self-Predictive RL
by: Ni, Tianwei, et al.
Published: (2024)
by: Ni, Tianwei, et al.
Published: (2024)
Analyzing sequential activity and travel decisions with interpretable deep inverse reinforcement learning
by: Liang, Yuebing, et al.
Published: (2025)
by: Liang, Yuebing, et al.
Published: (2025)
Multi-objective hybrid knowledge distillation for efficient deep learning in smart agriculture
by: Hoang, Phi-Hung, et al.
Published: (2025)
by: Hoang, Phi-Hung, et al.
Published: (2025)
SkyRL-Agent: Efficient RL Training for Multi-turn LLM Agent
by: Cao, Shiyi, et al.
Published: (2025)
by: Cao, Shiyi, et al.
Published: (2025)
Dynamics-Predictive Sampling for Active RL Finetuning of Large Reasoning Models
by: Mao, Yixiu, et al.
Published: (2026)
by: Mao, Yixiu, et al.
Published: (2026)
DEAS: DEtached value learning with Action Sequence for Scalable Offline RL
by: Kim, Changyeon, et al.
Published: (2025)
by: Kim, Changyeon, et al.
Published: (2025)
Preventing overfitting in deep learning using differential privacy
by: Khatri, Alizishaan Anwar Hussein
Published: (2026)
by: Khatri, Alizishaan Anwar Hussein
Published: (2026)
Tenyidie Syllabification corpus creation and deep learning applications
by: Angami, Teisovi, et al.
Published: (2025)
by: Angami, Teisovi, et al.
Published: (2025)
DyPNIPP: Predicting Environment Dynamics for RL-based Robust Informative Path Planning
by: Deolasee, Srujan, et al.
Published: (2024)
by: Deolasee, Srujan, et al.
Published: (2024)
A comparative study of deep learning and ensemble learning to extend the horizon of traffic forecasting
by: Zheng, Xiao, et al.
Published: (2025)
by: Zheng, Xiao, et al.
Published: (2025)
Transformer-based deep imitation learning for dual-arm robot manipulation
by: Kim, Heecheol, et al.
Published: (2021)
by: Kim, Heecheol, et al.
Published: (2021)
RL$^3$: Boosting Meta Reinforcement Learning via RL inside RL$^2$
by: Bhatia, Abhinav, et al.
Published: (2023)
by: Bhatia, Abhinav, et al.
Published: (2023)
A comprehensive overview of deep learning models for object detection from videos/images
by: Zulfqar, Sukana, et al.
Published: (2026)
by: Zulfqar, Sukana, et al.
Published: (2026)
ProRL Agent: Rollout-as-a-Service for RL Training of Multi-Turn LLM Agents
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving
by: Mai, Xinji, et al.
Published: (2025)
by: Mai, Xinji, et al.
Published: (2025)
Can Prompt Difficulty be Online Predicted for Accelerating RL Finetuning of Reasoning Models?
by: Qu, Yun, et al.
Published: (2025)
by: Qu, Yun, et al.
Published: (2025)
Similar Items
-
CAREL: Instruction-guided reinforcement learning with cross-modal auxiliary objectives
by: Saghafian, Armin, et al.
Published: (2024) -
Scale-specific auxiliary multi-task contrastive learning for deep liver vessel segmentation
by: Sadikine, Amine, et al.
Published: (2024) -
Permutative redundancy and uncertainty of the objective in deep learning
by: Glukhov, Vacslav
Published: (2024) -
How does the primate brain combine generative and discriminative computations in vision?
by: Peters, Benjamin, et al.
Published: (2024) -
Can AI mimic the human ability to define neologisms?
by: Georgiou, Georgios P.
Published: (2025)