HumanLM: Simulating Users with State Alignment Beats Response Imitation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Shirley, Choi, Evelyn, Khatua, Arpandeep, Wang, Zhanghan, He-Yueya, Joy, Weerasooriya, Tharindu Cyril, Wei, Wei, Yang, Diyi, Leskovec, Jure, Zou, James |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap
von: Chen, Tianlang, et al.
Veröffentlicht: (2026)
von: Chen, Tianlang, et al.
Veröffentlicht: (2026)
MultiGA: Leveraging Multi-Source Seeding in Genetic Algorithms
von: Ng, Isabelle Diana May-Xin, et al.
Veröffentlicht: (2025)
von: Ng, Isabelle Diana May-Xin, et al.
Veröffentlicht: (2025)
GraphMETRO: Mitigating Complex Graph Distribution Shifts via Mixture of Aligned Experts
von: Wu, Shirley, et al.
Veröffentlicht: (2023)
von: Wu, Shirley, et al.
Veröffentlicht: (2023)
Vicarious Offense and Noise Audit of Offensive Speech Classifiers: Unifying Human and Machine Disagreement on What is Offensive
von: Weerasooriya, Tharindu Cyril, et al.
Veröffentlicht: (2023)
von: Weerasooriya, Tharindu Cyril, et al.
Veröffentlicht: (2023)
Uncalibrated Reasoning: GRPO Induces Overconfidence for Stochastic Outcomes
von: Bereket, Michael, et al.
Veröffentlicht: (2025)
von: Bereket, Michael, et al.
Veröffentlicht: (2025)
Learning Efficient Positional Encodings with Graph Neural Networks
von: Kanatsoulis, Charilaos I., et al.
Veröffentlicht: (2025)
von: Kanatsoulis, Charilaos I., et al.
Veröffentlicht: (2025)
Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEM
von: Shetty, Samay U., et al.
Veröffentlicht: (2026)
von: Shetty, Samay U., et al.
Veröffentlicht: (2026)
Subasa - Adapting Language Models for Low-resourced Offensive Language Detection in Sinhala
von: Haturusinghe, Shanilka, et al.
Veröffentlicht: (2025)
von: Haturusinghe, Shanilka, et al.
Veröffentlicht: (2025)
LPI-RIT at LeWiDi-2025: Improving Distributional Predictions via Metadata and Loss Reweighting with DisCo
von: Sawkar, Mandira, et al.
Veröffentlicht: (2025)
von: Sawkar, Mandira, et al.
Veröffentlicht: (2025)
ProRefine: Inference-Time Prompt Refinement with Textual Feedback
von: Pandita, Deepak, et al.
Veröffentlicht: (2025)
von: Pandita, Deepak, et al.
Veröffentlicht: (2025)
RelGNN: Composite Message Passing for Relational Deep Learning
von: Chen, Tianlang, et al.
Veröffentlicht: (2025)
von: Chen, Tianlang, et al.
Veröffentlicht: (2025)
Reverse Image Retrieval Cues Parametric Memory in Multimodal LLMs
von: Xu, Jialiang, et al.
Veröffentlicht: (2024)
von: Xu, Jialiang, et al.
Veröffentlicht: (2024)
Psychometric Alignment: Capturing Human Knowledge Distributions via Language Models
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024)
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024)
Reinforcement Learning for Machine Learning Engineering Agents
von: Yang, Sherry, et al.
Veröffentlicht: (2025)
von: Yang, Sherry, et al.
Veröffentlicht: (2025)
Large Language Models are Good Relational Learners
von: Wu, Fang, et al.
Veröffentlicht: (2025)
von: Wu, Fang, et al.
Veröffentlicht: (2025)
Blind Spot Navigation in Large Language Model Reasoning with Thought Space Explorer
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024)
Detecting Corpus-Level Knowledge Inconsistencies in Wikipedia with Large Language Models
von: Semnani, Sina J., et al.
Veröffentlicht: (2025)
von: Semnani, Sina J., et al.
Veröffentlicht: (2025)
Rater Cohesion and Quality from a Vicarious Perspective
von: Pandita, Deepak, et al.
Veröffentlicht: (2024)
von: Pandita, Deepak, et al.
Veröffentlicht: (2024)
ARTICLE: Annotator Reliability Through In-Context Learning
von: Dutta, Sujan, et al.
Veröffentlicht: (2024)
von: Dutta, Sujan, et al.
Veröffentlicht: (2024)
Evaluating and Optimizing Educational Content with Large Language Model Judgments
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024)
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024)
RFG: Test-Time Scaling for Diffusion Large Language Model Reasoning with Reward-Free Guidance
von: Chen, Tianlang, et al.
Veröffentlicht: (2025)
von: Chen, Tianlang, et al.
Veröffentlicht: (2025)
MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation
von: Huang, Qian, et al.
Veröffentlicht: (2023)
von: Huang, Qian, et al.
Veröffentlicht: (2023)
Optimas: Optimizing Compound AI Systems with Globally Aligned Local Rewards
von: Wu, Shirley, et al.
Veröffentlicht: (2025)
von: Wu, Shirley, et al.
Veröffentlicht: (2025)
StaRPO: Stability-Augmented Reinforcement Policy Optimization
von: Zhang, Jinghan, et al.
Veröffentlicht: (2026)
von: Zhang, Jinghan, et al.
Veröffentlicht: (2026)
CooperBench: Why Coding Agents Cannot be Your Teammates Yet
von: Khatua, Arpandeep, et al.
Veröffentlicht: (2026)
von: Khatua, Arpandeep, et al.
Veröffentlicht: (2026)
Relational Deep Learning: Challenges, Foundations and Next-Generation Architectures
von: Dwivedi, Vijay Prakash, et al.
Veröffentlicht: (2025)
von: Dwivedi, Vijay Prakash, et al.
Veröffentlicht: (2025)
Spiking mode-based neural networks
von: Lin, Zhanghan, et al.
Veröffentlicht: (2023)
von: Lin, Zhanghan, et al.
Veröffentlicht: (2023)
CollabLLM: From Passive Responders to Active Collaborators
von: Wu, Shirley, et al.
Veröffentlicht: (2025)
von: Wu, Shirley, et al.
Veröffentlicht: (2025)
Uncertainty Quantification for Forward and Inverse Problems of PDEs via Latent Global Evolution
von: Wu, Tailin, et al.
Veröffentlicht: (2024)
von: Wu, Tailin, et al.
Veröffentlicht: (2024)
TimeGraphs: Graph-based Temporal Reasoning
von: Maheshwari, Paridhi, et al.
Veröffentlicht: (2024)
von: Maheshwari, Paridhi, et al.
Veröffentlicht: (2024)
Inferring Dynamic Networks from Marginals with Iterative Proportional Fitting
von: Chang, Serina, et al.
Veröffentlicht: (2024)
von: Chang, Serina, et al.
Veröffentlicht: (2024)
Learning over Positive and Negative Edges with Contrastive Message Passing
von: Pao-Huang, Peter, et al.
Veröffentlicht: (2026)
von: Pao-Huang, Peter, et al.
Veröffentlicht: (2026)
Compositional Generative Inverse Design
von: Wu, Tailin, et al.
Veröffentlicht: (2024)
von: Wu, Tailin, et al.
Veröffentlicht: (2024)
Novel Materials for the Removal of Microplastics and Nanoplastics in Drinking Water Treatment: A Comprehensive Review
von: Yueya Chang, et al.
Veröffentlicht: (2025)
von: Yueya Chang, et al.
Veröffentlicht: (2025)
Artificial Intelligence Applications in River Management: Challenges and Insights From a Bibliometric Review
von: Yueya Chang, et al.
Veröffentlicht: (2025)
von: Yueya Chang, et al.
Veröffentlicht: (2025)
Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment
von: Gan, Woody Haosheng, et al.
Veröffentlicht: (2026)
von: Gan, Woody Haosheng, et al.
Veröffentlicht: (2026)
AvaTaR: Optimizing LLM Agents for Tool Usage via Contrastive Reasoning
von: Wu, Shirley, et al.
Veröffentlicht: (2024)
von: Wu, Shirley, et al.
Veröffentlicht: (2024)
STaRK: Benchmarking LLM Retrieval on Textual and Relational Knowledge Bases
von: Wu, Shirley, et al.
Veröffentlicht: (2024)
von: Wu, Shirley, et al.
Veröffentlicht: (2024)
Frustration and chirality in three-dimensional trillium lattices: Insights and Perspectives
von: Khatua, J., et al.
Veröffentlicht: (2025)
von: Khatua, J., et al.
Veröffentlicht: (2025)
Interactive and Hybrid Imitation Learning: Provably Beating Behavior Cloning
von: Li, Yichen, et al.
Veröffentlicht: (2024)
von: Li, Yichen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap
von: Chen, Tianlang, et al.
Veröffentlicht: (2026) -
MultiGA: Leveraging Multi-Source Seeding in Genetic Algorithms
von: Ng, Isabelle Diana May-Xin, et al.
Veröffentlicht: (2025) -
GraphMETRO: Mitigating Complex Graph Distribution Shifts via Mixture of Aligned Experts
von: Wu, Shirley, et al.
Veröffentlicht: (2023) -
Vicarious Offense and Noise Audit of Offensive Speech Classifiers: Unifying Human and Machine Disagreement on What is Offensive
von: Weerasooriya, Tharindu Cyril, et al.
Veröffentlicht: (2023) -
Uncalibrated Reasoning: GRPO Induces Overconfidence for Stochastic Outcomes
von: Bereket, Michael, et al.
Veröffentlicht: (2025)