Language Agents Mirror Human Causal Reasoning Biases. How Can We Help Them Think Like Scientists?
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | GX-Chen, Anthony, Lin, Dongyan, Samiei, Mandana, Precup, Doina, Richards, Blake A., Fergus, Rob, Marino, Kenneth |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Efficient Exploration and Discriminative World Model Learning with an Object-Centric Abstraction
par: GX-Chen, Anthony, et autres
Publié: (2024)
par: GX-Chen, Anthony, et autres
Publié: (2024)
Balancing Plasticity and Stability with Fast and Slow Successor Features
par: Chua, Raymond, et autres
Publié: (2026)
par: Chua, Raymond, et autres
Publié: (2026)
Functional Acceleration for Policy Mirror Descent
par: Chelu, Veronica, et autres
Publié: (2024)
par: Chelu, Veronica, et autres
Publié: (2024)
KL-Regularized Reinforcement Learning is Designed to Mode Collapse
par: GX-Chen, Anthony, et autres
Publié: (2025)
par: GX-Chen, Anthony, et autres
Publié: (2025)
Learning Successor Features the Simple Way
par: Chua, Raymond, et autres
Publié: (2024)
par: Chua, Raymond, et autres
Publié: (2024)
Diversity-Enriched Option-Critic
par: Kamat, Anand, et autres
Publié: (2020)
par: Kamat, Anand, et autres
Publié: (2020)
A Look at Value-Based Decision-Time vs. Background Planning Methods Across Different Settings
par: Alver, Safa, et autres
Publié: (2022)
par: Alver, Safa, et autres
Publié: (2022)
The New Technologies: What Are They, How Can We Get Them, and Why Don't We Have Them?
par: Lowenthal, Ralph A.
Publié: (1989)
par: Lowenthal, Ralph A.
Publié: (1989)
Thinking Like a Scientist: Can Interactive Simulations Foster Critical AI Literacy?
par: Zhao, Yiling, et autres
Publié: (2025)
par: Zhao, Yiling, et autres
Publié: (2025)
On the Privacy of Selection Mechanisms with Gaussian Noise
par: Lebensold, Jonathan, et autres
Publié: (2024)
par: Lebensold, Jonathan, et autres
Publié: (2024)
On Campus or out of Town: How Publishing Online Tutorials Can Help Your Patrons
par: Blake, Lindsay
Publié: (2009)
par: Blake, Lindsay
Publié: (2009)
Library Service to Pregnant Teens: How Can We Help?
par: Gross, Melissa
Publié: (1997)
par: Gross, Melissa
Publié: (1997)
Conditions on Preference Relations that Guarantee the Existence of Optimal Policies
par: Carr, Jonathan Colaço, et autres
Publié: (2023)
par: Carr, Jonathan Colaço, et autres
Publié: (2023)
Sparse-Reg: Improving Sample Complexity in Offline Reinforcement Learning using Sparsity
par: Arnob, Samin Yeasar, et autres
Publié: (2025)
par: Arnob, Samin Yeasar, et autres
Publié: (2025)
Adaptive Exploration for Data-Efficient General Value Function Evaluations
par: Jain, Arushi, et autres
Publié: (2024)
par: Jain, Arushi, et autres
Publié: (2024)
Fluid-Agent Reinforcement Learning
par: Sharma, Shishir, et autres
Publié: (2026)
par: Sharma, Shishir, et autres
Publié: (2026)
Partial Models for Building Adaptive Model-Based Reinforcement Learning Agents
par: Alver, Safa, et autres
Publié: (2024)
par: Alver, Safa, et autres
Publié: (2024)
Can We Change How We Think Food Will Make Us Feel?
par: Jenna R. Cummings, et autres
Publié: (2026)
par: Jenna R. Cummings, et autres
Publié: (2026)
How Far Can Fairness Constraints Help Recover From Biased Data?
par: Sharma, Mohit, et autres
Publié: (2023)
par: Sharma, Mohit, et autres
Publié: (2023)
Mind the Sim-to-Real Gap & Think Like a Scientist
par: Parikh, Harsh, et autres
Publié: (2026)
par: Parikh, Harsh, et autres
Publié: (2026)
What Can We Expect from Data Scientists?
par: Kurt Englmeier
Publié: (2017)
par: Kurt Englmeier
Publié: (2017)
The Walras-Bowley Lecture: Fragmentation of Matching Markets and How Economics Can Help Integrate Them
par: Kamada, Yuichiro, et autres
Publié: (2025)
par: Kamada, Yuichiro, et autres
Publié: (2025)
"How Can We Help?" The Contribution of University Libraries to Student Retention
par: Hagel, Pauline, et autres
Publié: (2012)
par: Hagel, Pauline, et autres
Publié: (2012)
LLMs Can Plan Only If We Tell Them
par: Sel, Bilgehan, et autres
Publié: (2025)
par: Sel, Bilgehan, et autres
Publié: (2025)
They Can't Hear Us Does Not Mean We Can't Serve Them.
par: McDaniel, Julie Ann
Publié: (1992)
par: McDaniel, Julie Ann
Publié: (1992)
Parseval Regularization for Continual Reinforcement Learning
par: Chung, Wesley, et autres
Publié: (2024)
par: Chung, Wesley, et autres
Publié: (2024)
Relative Trajectory Balance is equivalent to Trust-PCL
par: Deleu, Tristan, et autres
Publié: (2025)
par: Deleu, Tristan, et autres
Publié: (2025)
How Much Can RAG Help the Reasoning of LLM?
par: Liu, Jingyu, et autres
Publié: (2024)
par: Liu, Jingyu, et autres
Publié: (2024)
Distribution in an Electronic Environment, or Will There Be Libraries as We Know Them in the Internet World?
par: Dowlin, Kenneth E.
Publié: (1995)
par: Dowlin, Kenneth E.
Publié: (1995)
How We Can Help America Learn: A Summary of Major Activities.
Publié: (1995)
Publié: (1995)
Library Media Paraprofessionals--We Can't Live without Them!
par: Pawlowski, Connie, et autres
Publié: (1993)
par: Pawlowski, Connie, et autres
Publié: (1993)
Our Unruly Signs, and How We Cope with Them
par: Floyd Merrell
Publié: (2005)
par: Floyd Merrell
Publié: (2005)
Incorporating Spatial Information into Goal-Conditioned Hierarchical Reinforcement Learning via Graph Representations
par: Zhang, Shuyuan, et autres
Publié: (2025)
par: Zhang, Shuyuan, et autres
Publié: (2025)
SCAR: Shapley Credit Assignment for More Efficient RLHF
par: Cao, Meng, et autres
Publié: (2025)
par: Cao, Meng, et autres
Publié: (2025)
Capacity-Constrained Continual Learning
par: Wen, Zheng, et autres
Publié: (2025)
par: Wen, Zheng, et autres
Publié: (2025)
Finite time analysis of temporal difference learning with linear function approximation: Tail averaging and regularisation
par: Patil, Gandharv, et autres
Publié: (2022)
par: Patil, Gandharv, et autres
Publié: (2022)
Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning
par: Ishfaq, Haque, et autres
Publié: (2025)
par: Ishfaq, Haque, et autres
Publié: (2025)
Mirror, Mirror on the Wall -- Which is the Best Model of Them All?
par: Sayed, Dina, et autres
Publié: (2025)
par: Sayed, Dina, et autres
Publié: (2025)
Panic and the Lack of Moral Competence. How We Can Help to Prevent Panic Pandemics
par: Georg Lind
Publié: (2021)
par: Georg Lind
Publié: (2021)
Artificial Intelligence Generates Stereotypical Images of Scientists but Can Also Detect Them: A Pilot Study Using the Draw-A-Scientist Test
par: Lee, Gyeonggeon
Publié: (2025)
par: Lee, Gyeonggeon
Publié: (2025)
Documents similaires
-
Efficient Exploration and Discriminative World Model Learning with an Object-Centric Abstraction
par: GX-Chen, Anthony, et autres
Publié: (2024) -
Balancing Plasticity and Stability with Fast and Slow Successor Features
par: Chua, Raymond, et autres
Publié: (2026) -
Functional Acceleration for Policy Mirror Descent
par: Chelu, Veronica, et autres
Publié: (2024) -
KL-Regularized Reinforcement Learning is Designed to Mode Collapse
par: GX-Chen, Anthony, et autres
Publié: (2025) -
Learning Successor Features the Simple Way
par: Chua, Raymond, et autres
Publié: (2024)