PolicySim: An LLM-Based Agent Social Simulation Sandbox for Proactive Policy Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Renhong, Tang, Ning, Xu, Jiarong, Cao, Yuxuan, Tu, Qingqian, Guo, Sheng, Zheng, Bo, Liu, Huiyuan, Yang, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RAPO: Expanding Exploration for LLM Agents via Retrieval-Augmented Policy Optimization
by: Zhang, Siwei, et al.
Published: (2026)
by: Zhang, Siwei, et al.
Published: (2026)
Proactive Constrained Policy Optimization with Preemptive Penalty
by: Yang, Ning, et al.
Published: (2025)
by: Yang, Ning, et al.
Published: (2025)
Can Modifying Data Address Graph Domain Adaptation?
by: Huang, Renhong, et al.
Published: (2024)
by: Huang, Renhong, et al.
Published: (2024)
Learning from the Irrecoverable: Error-Localized Policy Optimization for Tool-Integrated LLM Reasoning
by: Liang, Qiao, et al.
Published: (2026)
by: Liang, Qiao, et al.
Published: (2026)
WhatIf: Interactive Exploration of LLM-Powered Social Simulations for Policy Reasoning
by: Li, Yuxuan, et al.
Published: (2026)
by: Li, Yuxuan, et al.
Published: (2026)
SimRPD: Optimizing Recruitment Proactive Dialogue Agents through Simulator-Based Data Evaluation and Selection
by: Cao, Zhiyong, et al.
Published: (2026)
by: Cao, Zhiyong, et al.
Published: (2026)
SRAP-Agent: Simulating and Optimizing Scarce Resource Allocation Policy with LLM-based Agent
by: Ji, Jiarui, et al.
Published: (2024)
by: Ji, Jiarui, et al.
Published: (2024)
LitSim: A Conflict-aware Policy for Long-term Interactive Traffic Simulation
by: Xin, Haojie, et al.
Published: (2024)
by: Xin, Haojie, et al.
Published: (2024)
PolicySimEval: A Benchmark for Evaluating Policy Outcomes through Agent-Based Simulation
by: Kang, Jiaju, et al.
Published: (2025)
by: Kang, Jiaju, et al.
Published: (2025)
Simulation-Free Hierarchical Latent Policy Planning for Proactive Dialogues
by: He, Tao, et al.
Published: (2024)
by: He, Tao, et al.
Published: (2024)
ProAgentBench: Evaluating LLM Agents for Proactive Assistance with Real-World Data
by: Tang, Yuanbo, et al.
Published: (2026)
by: Tang, Yuanbo, et al.
Published: (2026)
Group-in-Group Policy Optimization for LLM Agent Training
by: Feng, Lang, et al.
Published: (2025)
by: Feng, Lang, et al.
Published: (2025)
Implicit Turn-Wise Policy Optimization for Proactive User-LLM Interaction
by: Wang, Haoyu, et al.
Published: (2026)
by: Wang, Haoyu, et al.
Published: (2026)
ForSim: Stepwise Forward Simulation for Traffic Policy Fine-Tuning
by: Chen, Keyu, et al.
Published: (2026)
by: Chen, Keyu, et al.
Published: (2026)
HCPO: Hierarchical Conductor-Based Policy Optimization in Multi-Agent Reinforcement Learning
by: Liu, Zejiao, et al.
Published: (2025)
by: Liu, Zejiao, et al.
Published: (2025)
What Makes LLM Agent Simulations Useful for Policy Practice? An Iterative Design Study in Emergency Preparedness
by: Li, Yuxuan, et al.
Published: (2025)
by: Li, Yuxuan, et al.
Published: (2025)
An Real-Sim-Real (RSR) Loop Framework for Generalizable Robotic Policy Transfer with Differentiable Simulation
by: Shi, Lu, et al.
Published: (2025)
by: Shi, Lu, et al.
Published: (2025)
Can Graph Neural Networks Expose Training Data Properties? An Efficient Risk Assessment Approach
by: Yuan, Hanyang, et al.
Published: (2024)
by: Yuan, Hanyang, et al.
Published: (2024)
Can Context Bridge the Reality Gap? Sim-to-Real Transfer of Context-Aware Policies
by: Iannotta, Marco, et al.
Published: (2025)
by: Iannotta, Marco, et al.
Published: (2025)
PTCG-Bench: Can LLM Agents Master Pokémon Trading Card Game?
by: Hua, Dongdong, et al.
Published: (2026)
by: Hua, Dongdong, et al.
Published: (2026)
TreeRPO: Tree Relative Policy Optimization
by: Yang, Zhicheng, et al.
Published: (2025)
by: Yang, Zhicheng, et al.
Published: (2025)
Think Outside the Policy: In-Context Steered Policy Optimization
by: Huang, Hsiu-Yuan, et al.
Published: (2025)
by: Huang, Hsiu-Yuan, et al.
Published: (2025)
DPEPO: Diverse Parallel Exploration Policy Optimization for LLM-based Agents
by: Zhang, Junshuo, et al.
Published: (2026)
by: Zhang, Junshuo, et al.
Published: (2026)
Adaptive Simulation Experiment for LLM Policy Optimization
by: Hu, Mingjie, et al.
Published: (2026)
by: Hu, Mingjie, et al.
Published: (2026)
Proactive Agent Research Environment: Simulating Active Users to Evaluate Proactive Assistants
by: Nathani, Deepak, et al.
Published: (2026)
by: Nathani, Deepak, et al.
Published: (2026)
Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization
by: Liu, Zeyuan, et al.
Published: (2026)
by: Liu, Zeyuan, et al.
Published: (2026)
The Dynamic Relationships Among Economic Policy Uncertainty, Bitcoin, and the Stock Market
by: Renhong Wu, et al.
Published: (2025)
by: Renhong Wu, et al.
Published: (2025)
SimKO: Simple Pass@K Policy Optimization
by: Peng, Ruotian, et al.
Published: (2025)
by: Peng, Ruotian, et al.
Published: (2025)
ComplexMCP: Evaluation of LLM Agents in Dynamic, Interdependent, and Large-Scale Tool Sandbox
by: Li, Yuanyang, et al.
Published: (2026)
by: Li, Yuanyang, et al.
Published: (2026)
Maestro: Learning to Collaborate via Conditional Listwise Policy Optimization for Multi-Agent LLMs
by: Yang, Wei, et al.
Published: (2025)
by: Yang, Wei, et al.
Published: (2025)
SetPO: Set-Level Policy Optimization for Diversity-Preserving LLM Reasoning
by: Li, Chenyi, et al.
Published: (2026)
by: Li, Chenyi, et al.
Published: (2026)
Training Proactive and Personalized LLM Agents
by: Sun, Weiwei, et al.
Published: (2025)
by: Sun, Weiwei, et al.
Published: (2025)
Graph-Enhanced Policy Optimization in LLM Agent Training
by: Yuan, Jiazhen, et al.
Published: (2025)
by: Yuan, Jiazhen, et al.
Published: (2025)
APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents
by: Li, Yibo, et al.
Published: (2026)
by: Li, Yibo, et al.
Published: (2026)
Real-to-Sim Robot Policy Evaluation with Gaussian Splatting Simulation of Soft-Body Interactions
by: Zhang, Kaifeng, et al.
Published: (2025)
by: Zhang, Kaifeng, et al.
Published: (2025)
SocialSim: Towards Socialized Simulation of Emotional Support Conversation
by: Chen, Zhuang, et al.
Published: (2025)
by: Chen, Zhuang, et al.
Published: (2025)
Analyzing and Internalizing Complex Policy Documents for LLM Agents
by: Liu, Jiateng, et al.
Published: (2025)
by: Liu, Jiateng, et al.
Published: (2025)
Adaptive Social Learning via Mode Policy Optimization for Language Agents
by: Wang, Minzheng, et al.
Published: (2025)
by: Wang, Minzheng, et al.
Published: (2025)
Dex4D: Task-Agnostic Point Track Policy for Sim-to-Real Dexterous Manipulation
by: Kuang, Yuxuan, et al.
Published: (2026)
by: Kuang, Yuxuan, et al.
Published: (2026)
Adaptive Collaboration with Humans: Metacognitive Policy Optimization for Multi-Agent LLMs with Continual Learning
by: Yang, Wei, et al.
Published: (2026)
by: Yang, Wei, et al.
Published: (2026)
Similar Items
-
RAPO: Expanding Exploration for LLM Agents via Retrieval-Augmented Policy Optimization
by: Zhang, Siwei, et al.
Published: (2026) -
Proactive Constrained Policy Optimization with Preemptive Penalty
by: Yang, Ning, et al.
Published: (2025) -
Can Modifying Data Address Graph Domain Adaptation?
by: Huang, Renhong, et al.
Published: (2024) -
Learning from the Irrecoverable: Error-Localized Policy Optimization for Tool-Integrated LLM Reasoning
by: Liang, Qiao, et al.
Published: (2026) -
WhatIf: Interactive Exploration of LLM-Powered Social Simulations for Policy Reasoning
by: Li, Yuxuan, et al.
Published: (2026)