Modeling Others' Minds as Code
Fuente:
arXiv
Salvato in:
| Autori principali: | Jha, Kunal, Huang, Aydan Yuenan, Ye, Eric, Jaques, Natasha, Kleiman-Weiner, Max |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
di: Jha, Kunal, et al.
Pubblicazione: (2025)
di: Jha, Kunal, et al.
Pubblicazione: (2025)
Task diversity produces systematic transfer but inhibits continual reinforcement learning
di: Seth, Purab, et al.
Pubblicazione: (2026)
di: Seth, Purab, et al.
Pubblicazione: (2026)
Estimating the Empowerment of Language Model Agents
di: Song, Jinyeop, et al.
Pubblicazione: (2025)
di: Song, Jinyeop, et al.
Pubblicazione: (2025)
Value Internalization: Learning and Generalizing from Social Reward
di: Rong, Frieda, et al.
Pubblicazione: (2024)
di: Rong, Frieda, et al.
Pubblicazione: (2024)
Evaluating LLMs in Open-Source Games
di: Sistla, Swadesh, et al.
Pubblicazione: (2025)
di: Sistla, Swadesh, et al.
Pubblicazione: (2025)
Generative Modeling for Robust Deep Reinforcement Learning on the Traveling Salesman Problem
di: Li, Michael, et al.
Pubblicazione: (2025)
di: Li, Michael, et al.
Pubblicazione: (2025)
Generative Value Conflicts Reveal LLM Priorities
di: Liu, Andy, et al.
Pubblicazione: (2025)
di: Liu, Andy, et al.
Pubblicazione: (2025)
Maximizing Mutual Information Between Prompt and Response Improves LLM Performance With No Additional Data
di: Nam, Hyunji, et al.
Pubblicazione: (2026)
di: Nam, Hyunji, et al.
Pubblicazione: (2026)
Impossibility Theorems for Feature Attribution
di: Bilodeau, Blair, et al.
Pubblicazione: (2022)
di: Bilodeau, Blair, et al.
Pubblicazione: (2022)
The Lock-in Hypothesis: Stagnation by Algorithm
di: Qiu, Tianyi Alex, et al.
Pubblicazione: (2025)
di: Qiu, Tianyi Alex, et al.
Pubblicazione: (2025)
How LLMs Distort Our Written Language
di: Abdulhai, Marwa, et al.
Pubblicazione: (2026)
di: Abdulhai, Marwa, et al.
Pubblicazione: (2026)
Learning to summarize user information for personalized reinforcement learning from human feedback
di: Nam, Hyunji, et al.
Pubblicazione: (2025)
di: Nam, Hyunji, et al.
Pubblicazione: (2025)
Evaluating & Reducing Deceptive Dialogue From Language Models with Multi-turn RL
di: Abdulhai, Marwa, et al.
Pubblicazione: (2025)
di: Abdulhai, Marwa, et al.
Pubblicazione: (2025)
CLadder: Assessing Causal Reasoning in Language Models
di: Jin, Zhijing, et al.
Pubblicazione: (2023)
di: Jin, Zhijing, et al.
Pubblicazione: (2023)
Infer Human's Intentions Before Following Natural Language Instructions
di: Wan, Yanming, et al.
Pubblicazione: (2024)
di: Wan, Yanming, et al.
Pubblicazione: (2024)
Learning to Cooperate with Humans using Generative Agents
di: Liang, Yancheng, et al.
Pubblicazione: (2024)
di: Liang, Yancheng, et al.
Pubblicazione: (2024)
Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
di: Poddar, Sriyash, et al.
Pubblicazione: (2024)
di: Poddar, Sriyash, et al.
Pubblicazione: (2024)
Towards Machine Theory of Mind with Large Language Model-Augmented Inverse Planning
di: Gelpí, Rebekah A., et al.
Pubblicazione: (2025)
di: Gelpí, Rebekah A., et al.
Pubblicazione: (2025)
MindCraft: How Concept Trees Take Shape In Deep Models
di: Tian, Bowei, et al.
Pubblicazione: (2025)
di: Tian, Bowei, et al.
Pubblicazione: (2025)
An Efficient Open World Environment for Multi-Agent Social Learning
di: Ye, Eric, et al.
Pubblicazione: (2025)
di: Ye, Eric, et al.
Pubblicazione: (2025)
When Empowerment Disempowers
di: Yang, Claire, et al.
Pubblicazione: (2025)
di: Yang, Claire, et al.
Pubblicazione: (2025)
Soup to go: mitigating forgetting during continual learning with model averaging
di: Kleiman, Anat, et al.
Pubblicazione: (2025)
di: Kleiman, Anat, et al.
Pubblicazione: (2025)
Deep Convolutional Neural Networks for predicting highest priority functional group in organic molecules
di: Khatri, Kunal, et al.
Pubblicazione: (2026)
di: Khatri, Kunal, et al.
Pubblicazione: (2026)
Transcendence: Generative Models Can Outperform The Experts That Train Them
di: Zhang, Edwin, et al.
Pubblicazione: (2024)
di: Zhang, Edwin, et al.
Pubblicazione: (2024)
ReLU's Revival: On the Entropic Overload in Normalization-Free Large Language Models
di: Jha, Nandan Kumar, et al.
Pubblicazione: (2024)
di: Jha, Nandan Kumar, et al.
Pubblicazione: (2024)
Mind the Model, Not the Agent: The Primacy Bias in Model-based RL
di: Qiao, Zhongjian, et al.
Pubblicazione: (2023)
di: Qiao, Zhongjian, et al.
Pubblicazione: (2023)
Co-Located Tests, Better AI Code: How Test Syntax Structure Affects Foundation Model Code Generation
di: Jacopin, Éric
Pubblicazione: (2026)
di: Jacopin, Éric
Pubblicazione: (2026)
UrbanMind: Towards Urban General Intelligence via Tool-Enhanced Retrieval-Augmented Generation and Multilevel Optimization
di: Yang, Kai, et al.
Pubblicazione: (2025)
di: Yang, Kai, et al.
Pubblicazione: (2025)
SBSC: Step-By-Step Coding for Improving Mathematical Olympiad Performance
di: Singh, Kunal, et al.
Pubblicazione: (2025)
di: Singh, Kunal, et al.
Pubblicazione: (2025)
Boundedly Rational Meta-Learning in Sequential Consumer Choice
di: Khosravi, Mehrzad, et al.
Pubblicazione: (2026)
di: Khosravi, Mehrzad, et al.
Pubblicazione: (2026)
CodeIt: Self-Improving Language Models with Prioritized Hindsight Replay
di: Butt, Natasha, et al.
Pubblicazione: (2024)
di: Butt, Natasha, et al.
Pubblicazione: (2024)
DMWM: Dual-Mind World Model with Long-Term Imagination
di: Wang, Lingyi, et al.
Pubblicazione: (2025)
di: Wang, Lingyi, et al.
Pubblicazione: (2025)
Fathom-DeepResearch: Unlocking Long Horizon Information Retrieval and Synthesis for SLMs
di: Singh, Shreyas, et al.
Pubblicazione: (2025)
di: Singh, Shreyas, et al.
Pubblicazione: (2025)
RAMAC: Multimodal Risk-Aware Offline Reinforcement Learning and the Role of Behavior Regularization
di: Fukazawa, Kai, et al.
Pubblicazione: (2025)
di: Fukazawa, Kai, et al.
Pubblicazione: (2025)
Exploring the Plausibility of Hate and Counter Speech Detectors with Explainable AI
di: Böck, Adrian Jaques, et al.
Pubblicazione: (2024)
di: Böck, Adrian Jaques, et al.
Pubblicazione: (2024)
AGENT: An Aerial Vehicle Generation and Design Tool Using Large Language Models
di: Samplawski, Colin, et al.
Pubblicazione: (2025)
di: Samplawski, Colin, et al.
Pubblicazione: (2025)
Throughput Optimization as a Strategic Lever in Large-Scale AI Systems: Evidence from Dataloader and Memory Profiling Innovations
di: Jha, Mayank
Pubblicazione: (2026)
di: Jha, Mayank
Pubblicazione: (2026)
Do Large Code Models Understand Programming Concepts? Counterfactual Analysis for Code Predicates
di: Hooda, Ashish, et al.
Pubblicazione: (2024)
di: Hooda, Ashish, et al.
Pubblicazione: (2024)
Deep Minds and Shallow Probes
di: Lee, Su Hyeong, et al.
Pubblicazione: (2026)
di: Lee, Su Hyeong, et al.
Pubblicazione: (2026)
A Notion of Complexity for Theory of Mind via Discrete World Models
di: Huang, X. Angelo, et al.
Pubblicazione: (2024)
di: Huang, X. Angelo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
di: Jha, Kunal, et al.
Pubblicazione: (2025) -
Task diversity produces systematic transfer but inhibits continual reinforcement learning
di: Seth, Purab, et al.
Pubblicazione: (2026) -
Estimating the Empowerment of Language Model Agents
di: Song, Jinyeop, et al.
Pubblicazione: (2025) -
Value Internalization: Learning and Generalizing from Social Reward
di: Rong, Frieda, et al.
Pubblicazione: (2024) -
Evaluating LLMs in Open-Source Games
di: Sistla, Swadesh, et al.
Pubblicazione: (2025)