Learning from negative feedback, or positive feedback or both
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Abdolmaleki, Abbas, Piot, Bilal, Shahriari, Bobak, Springenberg, Jost Tobias, Hertweck, Tim, Joshi, Rishabh, Oh, Junhyuk, Bloesch, Michael, Lampe, Thomas, Heess, Nicolas, Buchli, Jonas, Riedmiller, Martin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Game On: Towards Language Models as RL Experimenters
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024)
Offline Actor-Critic Reinforcement Learning Scales to Large Models
von: Springenberg, Jost Tobias, et al.
Veröffentlicht: (2024)
von: Springenberg, Jost Tobias, et al.
Veröffentlicht: (2024)
Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement
von: Bloesch, Michael, et al.
Veröffentlicht: (2025)
von: Bloesch, Michael, et al.
Veröffentlicht: (2025)
Real-World Fluid Directed Rigid Body Control via Deep Reinforcement Learning
von: Bhardwaj, Mohak, et al.
Veröffentlicht: (2024)
von: Bhardwaj, Mohak, et al.
Veröffentlicht: (2024)
Less is more -- the Dispatcher/ Executor principle for multi-task Reinforcement Learning
von: Riedmiller, Martin, et al.
Veröffentlicht: (2023)
von: Riedmiller, Martin, et al.
Veröffentlicht: (2023)
Supervised Fine Tuning on Curated Data is Reinforcement Learning (and can be improved)
von: Qin, Chongli, et al.
Veröffentlicht: (2025)
von: Qin, Chongli, et al.
Veröffentlicht: (2025)
Grazing increases the positive feedback of legumes while decreasing the negative feedback of grass
von: Jiechao Chang, et al.
Veröffentlicht: (2024)
von: Jiechao Chang, et al.
Veröffentlicht: (2024)
The effects of supervisory negative feedback and coaching on employees' responses to feedback: The moderating role of mindset
von: Jetmir Zyberaj
Veröffentlicht: (2024)
von: Jetmir Zyberaj
Veröffentlicht: (2024)
Optogenetic feedback control of cardiac arrhythmias
von: Abidi, Syed Khizar Abbas
Veröffentlicht: (2025)
von: Abidi, Syed Khizar Abbas
Veröffentlicht: (2025)
Inshore-offshore sedimentation differences resulting from resuspension in the Eastern Basin of Lake Erie
von: Bloesch, J
Veröffentlicht: (1978)
von: Bloesch, J
Veröffentlicht: (1978)
Imitating Language via Scalable Inverse Reinforcement Learning
von: Wulfmeier, Markus, et al.
Veröffentlicht: (2024)
von: Wulfmeier, Markus, et al.
Veröffentlicht: (2024)
Structural stability for the bidirectional cyclic negative feedback systems
von: Cheng, Xu, et al.
Veröffentlicht: (2025)
von: Cheng, Xu, et al.
Veröffentlicht: (2025)
Private sector investment in Marine Protected Areas-Experiences of the Chumbe Island Coral Park in Zanzibar/Tanzania
von: Riedmiller, S.
Veröffentlicht: (2003)
von: Riedmiller, S.
Veröffentlicht: (2003)
Private Sector Management of Marine Protected Areas: The Chumbe Island Case
von: Riedmiller, S.
Veröffentlicht: (2000)
von: Riedmiller, S.
Veröffentlicht: (2000)
Local positive feedback in the overall negative: the impact of quasar winds on star formation in the FIRE cosmological simulations
von: Mercedes-Feliz, Jonathan, et al.
Veröffentlicht: (2023)
von: Mercedes-Feliz, Jonathan, et al.
Veröffentlicht: (2023)
Global convergence for time-periodic systems with negative feedback and applications
von: Wang, Yi, et al.
Veröffentlicht: (2024)
von: Wang, Yi, et al.
Veröffentlicht: (2024)
Time-dependent balls and bins model with positive feedback
von: Sidorova, Nadia
Veröffentlicht: (2018)
von: Sidorova, Nadia
Veröffentlicht: (2018)
A Perception-feedback position-tracking control for quadrotors
von: Espindola, Eduardo, et al.
Veröffentlicht: (2025)
von: Espindola, Eduardo, et al.
Veröffentlicht: (2025)
NMF-FFB: Non-negative matrix factorization with feedforward-feedback structure
von: Satoh, Kenichi
Veröffentlicht: (2025)
von: Satoh, Kenichi
Veröffentlicht: (2025)
Quality counts? Examining the role of feedback provider and feedback quality on students' feedback perceptions
von: Theresa Ruwe, et al.
Veröffentlicht: (2025)
von: Theresa Ruwe, et al.
Veröffentlicht: (2025)
An adenosinergic positive feedback loop extends pharmacological cardioprotection duration
von: Gerald Wölkart, et al.
Veröffentlicht: (2024)
von: Gerald Wölkart, et al.
Veröffentlicht: (2024)
Benefit of educational feedback for the use of positive expiratory pressure device
von: Gregory Reychler
Veröffentlicht: (2015)
von: Gregory Reychler
Veröffentlicht: (2015)
How negative feedback from filamentous actin affects cell shapes and motility
von: Hughes, Jack M., et al.
Veröffentlicht: (2026)
von: Hughes, Jack M., et al.
Veröffentlicht: (2026)
Vegetation-climate feedbacks across scales
von: Miralles, Diego G., et al.
Veröffentlicht: (2024)
von: Miralles, Diego G., et al.
Veröffentlicht: (2024)
Predicting oscillations in complex networks with delayed feedback
von: Liu, Shijie, et al.
Veröffentlicht: (2026)
von: Liu, Shijie, et al.
Veröffentlicht: (2026)
Negative feedback
von: James S. Huntley
Veröffentlicht: (2025)
von: James S. Huntley
Veröffentlicht: (2025)
A McKean--Vlasov equation with positive feedback and blow-ups
von: Hambly, Ben, et al.
Veröffentlicht: (2018)
von: Hambly, Ben, et al.
Veröffentlicht: (2018)
On the global dynamics of a forest model with monotone positive feedback and memory
von: Herrera, Franco, et al.
Veröffentlicht: (2024)
von: Herrera, Franco, et al.
Veröffentlicht: (2024)
A further investigation regarding the efficacy of and preference for positive and corrective feedback
von: Erik S. Godinez, et al.
Veröffentlicht: (2024)
von: Erik S. Godinez, et al.
Veröffentlicht: (2024)
Nonlinear systems and passivity: feedback control, model reduction, and time discretization
von: Breiten, Tobias, et al.
Veröffentlicht: (2025)
von: Breiten, Tobias, et al.
Veröffentlicht: (2025)
How negative feedback and the ambient environment limit the influence of recombination in common envelope evolution
von: Chamandy, Luke, et al.
Veröffentlicht: (2023)
von: Chamandy, Luke, et al.
Veröffentlicht: (2023)
The open dense conjecture on eventually slow oscillations of the differential equation with delayed negative feedback
von: Feng, Lirui
Veröffentlicht: (2025)
von: Feng, Lirui
Veröffentlicht: (2025)
Role of membrane estrogen receptor alpha on the negative feedback of estrogens on luteinizing hormone secretion
von: Mélanie C. Faure, et al.
Veröffentlicht: (2025)
von: Mélanie C. Faure, et al.
Veröffentlicht: (2025)
From critique to catalyst: How academic entrepreneurs transform negative feedback into pivots and performance
von: D. Carrington Motley, et al.
Veröffentlicht: (2025)
von: D. Carrington Motley, et al.
Veröffentlicht: (2025)
Can peer feedback substitute for faculty feedback in predoctoral dental education?
von: Brandon M. Veremis, et al.
Veröffentlicht: (2024)
von: Brandon M. Veremis, et al.
Veröffentlicht: (2024)
Creating feedback with ImPACT : Improving consistency in feedback through the design of principles of good feedback and creating common QuickMarks
von: Imogen Tijou, et al.
Veröffentlicht: (2026)
von: Imogen Tijou, et al.
Veröffentlicht: (2026)
Shrub encroachment promotes positive feedbacks from herbivores that reinforce ecosystem change
von: Kieran J. Andreoni, et al.
Veröffentlicht: (2025)
von: Kieran J. Andreoni, et al.
Veröffentlicht: (2025)
RL Token: Bootstrapping Online RL with Vision-Language-Action Models
von: Xu, Charles, et al.
Veröffentlicht: (2026)
von: Xu, Charles, et al.
Veröffentlicht: (2026)
To cascade feedback loops, or not?
von: Eduard Eitelberg
Veröffentlicht: (2024)
von: Eduard Eitelberg
Veröffentlicht: (2024)
Enhancing proactivity or encouraging cheating? The contrasting effects of leader negative feedback on employees' workplace behaviour
von: Niannian Dong, et al.
Veröffentlicht: (2026)
von: Niannian Dong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Game On: Towards Language Models as RL Experimenters
von: Zhang, Jingwei, et al.
Veröffentlicht: (2024) -
Offline Actor-Critic Reinforcement Learning Scales to Large Models
von: Springenberg, Jost Tobias, et al.
Veröffentlicht: (2024) -
Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement
von: Bloesch, Michael, et al.
Veröffentlicht: (2025) -
Real-World Fluid Directed Rigid Body Control via Deep Reinforcement Learning
von: Bhardwaj, Mohak, et al.
Veröffentlicht: (2024) -
Less is more -- the Dispatcher/ Executor principle for multi-task Reinforcement Learning
von: Riedmiller, Martin, et al.
Veröffentlicht: (2023)