Learning Human-like Representations to Enable Learning Human Values
Fuente:
arXiv
Saved in:
| Main Authors: | Wynn, Andrea, Sucholutsky, Ilia, Griffiths, Thomas L. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Revisiting Rogers' Paradox in the Context of Human-AI Interaction
by: Collins, Katherine M., et al.
Published: (2025)
by: Collins, Katherine M., et al.
Published: (2025)
Analyzing the Roles of Language and Vision in Learning from Limited Data
by: Chen, Allison, et al.
Published: (2024)
by: Chen, Allison, et al.
Published: (2024)
Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse
by: Liu, Ryan, et al.
Published: (2024)
by: Liu, Ryan, et al.
Published: (2024)
Learning with Language-Guided State Abstractions
by: Peng, Andi, et al.
Published: (2024)
by: Peng, Andi, et al.
Published: (2024)
Concept Alignment
by: Rane, Sunayana, et al.
Published: (2024)
by: Rane, Sunayana, et al.
Published: (2024)
Learning Human-Aligned Representations with Contrastive Learning and Generative Similarity
by: Marjieh, Raja, et al.
Published: (2024)
by: Marjieh, Raja, et al.
Published: (2024)
Large Language Models Assume People are More Rational than We Really are
by: Liu, Ryan, et al.
Published: (2024)
by: Liu, Ryan, et al.
Published: (2024)
Preference-Conditioned Language-Guided Abstraction
by: Peng, Andi, et al.
Published: (2024)
by: Peng, Andi, et al.
Published: (2024)
HER: Human-like Reasoning and Reinforcement Learning for LLM Role-playing
by: Du, Chengyu, et al.
Published: (2026)
by: Du, Chengyu, et al.
Published: (2026)
Distilling Symbolic Priors for Concept Learning into Neural Networks
by: Marinescu, Ioana, et al.
Published: (2024)
by: Marinescu, Ioana, et al.
Published: (2024)
Building Machines that Learn and Think with People
by: Collins, Katherine M., et al.
Published: (2024)
by: Collins, Katherine M., et al.
Published: (2024)
Human-Centered Learning Mechanics: A Dynamical Framework for Entropy-Regulated Representation Learning
by: Tran, Kim Phuc
Published: (2026)
by: Tran, Kim Phuc
Published: (2026)
Program-Based Strategy Induction for Reinforcement Learning
by: Correa, Carlos G., et al.
Published: (2024)
by: Correa, Carlos G., et al.
Published: (2024)
Enabling Self-Improving Agents to Learn at Test Time With Human-In-The-Loop Guidance
by: He, Yufei, et al.
Published: (2025)
by: He, Yufei, et al.
Published: (2025)
What is a Number, That a Large Language Model May Know It?
by: Marjieh, Raja, et al.
Published: (2025)
by: Marjieh, Raja, et al.
Published: (2025)
Can Differentiable Decision Trees Enable Interpretable Reward Learning from Human Feedback?
by: Kalra, Akansha, et al.
Published: (2023)
by: Kalra, Akansha, et al.
Published: (2023)
A Human-Inspired Decoupled Architecture for Efficient Audio Representation Learning
by: Kawano, Harunori, et al.
Published: (2026)
by: Kawano, Harunori, et al.
Published: (2026)
A collection of the accepted papers for the Human-Centric Representation Learning workshop at AAAI 2024
by: Spathis, Dimitris, et al.
Published: (2024)
by: Spathis, Dimitris, et al.
Published: (2024)
Human-like Category Learning by Injecting Ecological Priors from Large Language Models into Neural Networks
by: Jagadish, Akshay K., et al.
Published: (2024)
by: Jagadish, Akshay K., et al.
Published: (2024)
Learning Social Heuristics for Human-Aware Path Planning
by: Eirale, Andrea, et al.
Published: (2025)
by: Eirale, Andrea, et al.
Published: (2025)
Semi-Supervised Graph Representation Learning with Human-centric Explanation for Predicting Fatty Liver Disease
by: Kim, So Yeon, et al.
Published: (2024)
by: Kim, So Yeon, et al.
Published: (2024)
Representation Learning of Lab Values via Masked AutoEncoders
by: Restrepo, David, et al.
Published: (2025)
by: Restrepo, David, et al.
Published: (2025)
Stable Offline Value Function Learning with Bisimulation-based Representations
by: Pavse, Brahma S., et al.
Published: (2024)
by: Pavse, Brahma S., et al.
Published: (2024)
Meta-Learning at Scale for Large Language Models via Low-Rank Amortized Bayesian Meta-Learning
by: Zhang, Liyi, et al.
Published: (2025)
by: Zhang, Liyi, et al.
Published: (2025)
Human-like Forgetting Curves in Deep Neural Networks
by: Kline, Dylan
Published: (2025)
by: Kline, Dylan
Published: (2025)
Virtual Human Generative Model: Masked Modeling Approach for Learning Human Characteristics
by: Oono, Kenta, et al.
Published: (2023)
by: Oono, Kenta, et al.
Published: (2023)
Value Imprint: A Technique for Auditing the Human Values Embedded in RLHF Datasets
by: Obi, Ike, et al.
Published: (2024)
by: Obi, Ike, et al.
Published: (2024)
Toward Efficient Exploration by Large Language Model Agents
by: Arumugam, Dilip, et al.
Published: (2025)
by: Arumugam, Dilip, et al.
Published: (2025)
In-Context Learning can Perform Continual Learning Like Humans
by: Kang, Liuwang, et al.
Published: (2025)
by: Kang, Liuwang, et al.
Published: (2025)
On Benchmarking Human-Like Intelligence in Machines
by: Ying, Lance, et al.
Published: (2025)
by: Ying, Lance, et al.
Published: (2025)
Aligning Robot and Human Representations
by: Bobu, Andreea, et al.
Published: (2023)
by: Bobu, Andreea, et al.
Published: (2023)
Artificial Neural Nets and the Representation of Human Concepts
by: Freiesleben, Timo
Published: (2023)
by: Freiesleben, Timo
Published: (2023)
Human-like Working Memory Interference in Large Language Models
by: Xiong, Hua-Dong, et al.
Published: (2026)
by: Xiong, Hua-Dong, et al.
Published: (2026)
Human-Inspired Framework to Accelerate Reinforcement Learning
by: Beikmohammadi, Ali, et al.
Published: (2023)
by: Beikmohammadi, Ali, et al.
Published: (2023)
Are Human-generated Demonstrations Necessary for In-context Learning?
by: Li, Rui, et al.
Published: (2023)
by: Li, Rui, et al.
Published: (2023)
Understanding the Learning Dynamics of Alignment with Human Feedback
by: Im, Shawn, et al.
Published: (2024)
by: Im, Shawn, et al.
Published: (2024)
Human-Inspired Multi-Level Reinforcement Learning
by: Wu, Mingkang, et al.
Published: (2025)
by: Wu, Mingkang, et al.
Published: (2025)
Motion Tracks: A Unified Representation for Human-Robot Transfer in Few-Shot Imitation Learning
by: Ren, Juntao, et al.
Published: (2025)
by: Ren, Juntao, et al.
Published: (2025)
Contrastive Preference Learning: Learning from Human Feedback without RL
by: Hejna, Joey, et al.
Published: (2023)
by: Hejna, Joey, et al.
Published: (2023)
FANoise: Singular Value-Adaptive Noise Modulation for Robust Multimodal Representation Learning
by: Li, Jiaoyang, et al.
Published: (2025)
by: Li, Jiaoyang, et al.
Published: (2025)
Similar Items
-
Revisiting Rogers' Paradox in the Context of Human-AI Interaction
by: Collins, Katherine M., et al.
Published: (2025) -
Analyzing the Roles of Language and Vision in Learning from Limited Data
by: Chen, Allison, et al.
Published: (2024) -
Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse
by: Liu, Ryan, et al.
Published: (2024) -
Learning with Language-Guided State Abstractions
by: Peng, Andi, et al.
Published: (2024) -
Concept Alignment
by: Rane, Sunayana, et al.
Published: (2024)