Learning to Learn Faster from Human Feedback with Language Model Predictive Control
Fuente:
arXiv
Guardado en:
| Autores principales: | Liang, Jacky, Xia, Fei, Yu, Wenhao, Zeng, Andy, Arenas, Montserrat Gonzalez, Attarian, Maria, Bauza, Maria, Bennice, Matthew, Bewley, Alex, Dostmohamed, Adil, Fu, Chuyuan Kelly, Gileadi, Nimrod, Giustina, Marissa, Gopalakrishnan, Keerthana, Hasenclever, Leonard, Humplik, Jan, Hsu, Jasmine, Joshi, Nikhil, Jyenis, Ben, Kew, Chase, Kirmani, Sean, Lee, Tsang-Wei Edward, Lee, Kuang-Huei, Michaely, Assaf Hurwitz, Moore, Joss, Oslund, Ken, Rao, Dushyant, Ren, Allen, Tabanpour, Baruch, Vuong, Quan, Wahid, Ayzaan, Xiao, Ted, Xu, Ying, Zhuang, Vincent, Xu, Peng, Frey, Erik, Caluwaerts, Ken, Zhang, Tingnan, Ichter, Brian, Tompson, Jonathan, Takayama, Leila, Vanhoucke, Vincent, Shafran, Izhak, Mataric, Maja, Sadigh, Dorsa, Heess, Nicolas, Rao, Kanishka, Stewart, Nik, Tan, Jie, Parada, Carolina |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Proc4Gem: Foundation models for physical agency through procedural generation
por: Lin, Yixin, et al.
Publicado: (2025)
por: Lin, Yixin, et al.
Publicado: (2025)
Diffusion Augmented Agents: A Framework for Efficient Exploration and Transfer Learning
por: Di Palo, Norman, et al.
Publicado: (2024)
por: Di Palo, Norman, et al.
Publicado: (2024)
Self-Improving Embodied Foundation Models
por: Ghasemipour, Seyed Kamyar Seyed, et al.
Publicado: (2025)
por: Ghasemipour, Seyed Kamyar Seyed, et al.
Publicado: (2025)
Trade preferential agreements in Latin America : n ex-ante assessment / Michael Michaely
por: Michaely, Michael
Publicado: (1996)
por: Michaely, Michael
Publicado: (1996)
DB-KSVD: Scalable Alternating Optimization for Disentangling High-Dimensional Embedding Spaces
por: Valentin, Romeo, et al.
Publicado: (2025)
por: Valentin, Romeo, et al.
Publicado: (2025)
Perceptions of Wicked Policy Problems and Anti‐Privatization Attitudes: The Mediating Role of Political Leadership Preferences
por: Izhak Berkovich
Publicado: (2025)
por: Izhak Berkovich
Publicado: (2025)
DESAFIOS À EXPORTAÇÃO INDUSTRIAL DE PEQUENAS E MÉDIAS EMPRESAS BRASILEIRAS
por: Lia Hasenclever
Publicado: (2007)
por: Lia Hasenclever
Publicado: (2007)
Desempenho econômico do Rio de Janeiro: trajetórias passadas e perspectivas futuras
por: Lia Hasenclever
Publicado: (2012)
por: Lia Hasenclever
Publicado: (2012)
ALOHA Unleashed: A Simple Recipe for Robot Dexterity
por: Zhao, Tony Z., et al.
Publicado: (2024)
por: Zhao, Tony Z., et al.
Publicado: (2024)
Retrieval Augmented End-to-End Spoken Dialog Models
por: Wang, Mingqiu, et al.
Publicado: (2024)
por: Wang, Mingqiu, et al.
Publicado: (2024)
On the Response Entropy of APUFs
por: Dumoulin, Vincent, et al.
Publicado: (2024)
por: Dumoulin, Vincent, et al.
Publicado: (2024)
Source data for the graphs presented in "Synthesis, Anthelmintic Activity, and Mechanism of Action of 5-Aryl-1H-indoles"
por: Kadlecová, Alena, et al.
Publicado: (2026)
por: Kadlecová, Alena, et al.
Publicado: (2026)
Source data for the graphs presented in "Advanced screening methods for assessing motility and hatching in plant-parasitic nematodes"
por: Kadlecová, Alena, et al.
Publicado: (2026)
por: Kadlecová, Alena, et al.
Publicado: (2026)
Prosody for Intuitive Robotic Interface Design: It's Not What You Said, It's How You Said It
por: Sanoubari, Elaheh, et al.
Publicado: (2024)
por: Sanoubari, Elaheh, et al.
Publicado: (2024)
Raw VLINDER data August 2020
por: Vergauwen, Thomas, et al.
Publicado: (2026)
por: Vergauwen, Thomas, et al.
Publicado: (2026)
Learning Robot Soccer from Egocentric Vision with Deep Reinforcement Learning
por: Tirumala, Dhruva, et al.
Publicado: (2024)
por: Tirumala, Dhruva, et al.
Publicado: (2024)
Higher Education Professionals in the Age of NPM and Digital Knowledge: Distinction Strategies for Forming New Occupational Capital
por: Wasserman, Varda, et al.
Publicado: (2022)
por: Wasserman, Varda, et al.
Publicado: (2022)
Messenger RNA
por: Jerard Hurwitz
Publicado: (1960)
por: Jerard Hurwitz
Publicado: (1960)
Hydrothermal fluid flow modelling in homogeneous and stratified sedimentary basin
por: Galerne, Christophe, et al.
Publicado: (2019)
por: Galerne, Christophe, et al.
Publicado: (2019)
School Librarians as Co-Teachers of Literacy: Librarian Perceptions and Knowledge in the Context of the Literacy Instruction Role
por: Reed, Karen Nourse, et al.
Publicado: (2018)
por: Reed, Karen Nourse, et al.
Publicado: (2018)
Knowledge Graph Reasoning with Self-supervised Reinforcement Learning
por: Ma, Ying, et al.
Publicado: (2024)
por: Ma, Ying, et al.
Publicado: (2024)
Communicating Competencies and Collaboration.
por: Zipperer, Lorri, et al.
Publicado: (2002)
por: Zipperer, Lorri, et al.
Publicado: (2002)
PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs
por: Nasiriany, Soroush, et al.
Publicado: (2024)
por: Nasiriany, Soroush, et al.
Publicado: (2024)
Robot Data Curation with Mutual Information Estimators
por: Hejna, Joey, et al.
Publicado: (2025)
por: Hejna, Joey, et al.
Publicado: (2025)
Mapping Goffman’s Invisible College
por: Leeds-Hurwitz, Wendy
Publicado: (2025)
por: Leeds-Hurwitz, Wendy
Publicado: (2025)
Literacy Gains from Weekly Newsela ELA Use: A Quasi-Experimental Evaluation of Content-Rich Instruction
por: Lisa B. Hurwitz
Publicado: (2023)
por: Lisa B. Hurwitz
Publicado: (2023)
DemoStart: Demonstration-led auto-curriculum applied to sim-to-real with multi-fingered robots
por: Bauza, Maria, et al.
Publicado: (2024)
por: Bauza, Maria, et al.
Publicado: (2024)
Using CBT‐E in the Treatment of Anorexia Nervosa With Comorbid Obsessive‐Compulsive Personality Disorder and Clinical Perfectionism
por: Liv Sand, et al.
Publicado: (2025)
por: Liv Sand, et al.
Publicado: (2025)
Learned Neural Physics Simulation for Articulated 3D Human Pose Reconstruction
por: Andriluka, Mykhaylo, et al.
Publicado: (2024)
por: Andriluka, Mykhaylo, et al.
Publicado: (2024)
Vision Language Models are In-Context Value Learners
por: Ma, Yecheng Jason, et al.
Publicado: (2024)
por: Ma, Yecheng Jason, et al.
Publicado: (2024)
Leveraging Automatic CAD Annotations for Supervised Learning in 3D Scene Understanding
por: Rao, Yuchen, et al.
Publicado: (2025)
por: Rao, Yuchen, et al.
Publicado: (2025)
Unmet needs: Bringing physical rehabilitation to people experiencing homelessness
por: Max Hurwitz, et al.
Publicado: (2024)
por: Max Hurwitz, et al.
Publicado: (2024)
Financial regret at older ages and longevity awareness
por: Abigail Hurwitz, et al.
Publicado: (2025)
por: Abigail Hurwitz, et al.
Publicado: (2025)
Learning Analytics in Online Learning Environment: A Systematic Review on the Focuses and the Types of Student-Related Analytics Data
por: Kew, Si Na, et al.
Publicado: (2022)
por: Kew, Si Na, et al.
Publicado: (2022)
Constructing Interpretable Features from Compositional Neuron Groups
por: Shafran, Or, et al.
Publicado: (2025)
por: Shafran, Or, et al.
Publicado: (2025)
Relational Epipolar Graphs for Robust Relative Camera Pose Estimation
por: Rao, Prateeth, et al.
Publicado: (2026)
por: Rao, Prateeth, et al.
Publicado: (2026)
Detecting Gravitational Wave Bursts From Stellar-Mass Binaries in the Milli-hertz Band
por: Xuan, Zeyuan, et al.
Publicado: (2023)
por: Xuan, Zeyuan, et al.
Publicado: (2023)
Stochastic Gravitational Wave Background from Highly-Eccentric Stellar-Mass Binaries in the Milli-hertz Band
por: Xuan, Zeyuan, et al.
Publicado: (2024)
por: Xuan, Zeyuan, et al.
Publicado: (2024)
Heuristic Deep Reinforcement Learning for Phase Shift Optimization in RIS-assisted Secure Satellite Communication Systems with RSMA
por: Bao, Tingnan, et al.
Publicado: (2025)
por: Bao, Tingnan, et al.
Publicado: (2025)
Turning English-centric LLMs Into Polyglots: How Much Multilinguality Is Needed?
por: Kew, Tannon, et al.
Publicado: (2023)
por: Kew, Tannon, et al.
Publicado: (2023)
Ejemplares similares
-
Proc4Gem: Foundation models for physical agency through procedural generation
por: Lin, Yixin, et al.
Publicado: (2025) -
Diffusion Augmented Agents: A Framework for Efficient Exploration and Transfer Learning
por: Di Palo, Norman, et al.
Publicado: (2024) -
Self-Improving Embodied Foundation Models
por: Ghasemipour, Seyed Kamyar Seyed, et al.
Publicado: (2025) -
Trade preferential agreements in Latin America : n ex-ante assessment / Michael Michaely
por: Michaely, Michael
Publicado: (1996) -
DB-KSVD: Scalable Alternating Optimization for Disentangling High-Dimensional Embedding Spaces
por: Valentin, Romeo, et al.
Publicado: (2025)