Saved in:
| Main Authors: | Carbo, J., Sanchez, N., Molina, J. M. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2401.14153 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Attacks Meet Interpretability (AmI) Evaluation and Findings
by: Ma, Qian, et al.
Published: (2023)
by: Ma, Qian, et al.
Published: (2023)
Trust model of privacy-concerned, emotionally-aware agents in a cooperative logistics problem
by: Carbo, J., et al.
Published: (2024)
by: Carbo, J., et al.
Published: (2024)
Merging plans with incomplete knowledge about actions and goals through an agent-based reputation system
by: Carbo, Javier, et al.
Published: (2024)
by: Carbo, Javier, et al.
Published: (2024)
I Am Big, You Are Little; I Am Right, You Are Wrong
by: Kelly, David A., et al.
Published: (2025)
by: Kelly, David A., et al.
Published: (2025)
CARJAN: Agent-Based Generation and Simulation of Traffic Scenarios with AJAN
by: Neis, Leonard Frank, et al.
Published: (2025)
by: Neis, Leonard Frank, et al.
Published: (2025)
"Teammates, Am I Clear?": Analysing Legible Behaviours in Teams
by: Faria, Miguel, et al.
Published: (2025)
by: Faria, Miguel, et al.
Published: (2025)
"Who Am I, and Who Else Is Here?" Behavioral Differentiation Without Role Assignment in Multi-Agent LLM Systems
by: Kandoussi, Houssam EL
Published: (2026)
by: Kandoussi, Houssam EL
Published: (2026)
Beyond a Single Reference: Training and Evaluation with Paraphrases in Sign Language Translation
by: Javorek, Václav, et al.
Published: (2026)
by: Javorek, Václav, et al.
Published: (2026)
(A)I Am Not a Lawyer, But...: Engaging Legal Experts towards Responsible LLM Policies for Legal Advice
by: Cheong, Inyoung, et al.
Published: (2024)
by: Cheong, Inyoung, et al.
Published: (2024)
Optimal Decision Making Through Scenario Simulations Using Large Language Models
by: Rasal, Sumedh, et al.
Published: (2024)
by: Rasal, Sumedh, et al.
Published: (2024)
PIORS: Personalized Intelligent Outpatient Reception based on Large Language Model with Multi-Agents Medical Scenario Simulation
by: Bao, Zhijie, et al.
Published: (2024)
by: Bao, Zhijie, et al.
Published: (2024)
Exploring the Necessity of Reasoning in LLM-based Agent Scenarios
by: Zhou, Xueyang, et al.
Published: (2025)
by: Zhou, Xueyang, et al.
Published: (2025)
LLM-empowered Agents Simulation Framework for Scenario Generation in Service Ecosystem Governance
by: Zhou, Deyu, et al.
Published: (2025)
by: Zhou, Deyu, et al.
Published: (2025)
Generative Agents for Multi-Agent Autoformalization of Interaction Scenarios
by: Mensfelt, Agnieszka, et al.
Published: (2024)
by: Mensfelt, Agnieszka, et al.
Published: (2024)
Exploring Flexible Scenario Generation in Godot Simulator
by: Peraltai, Daniel, et al.
Published: (2024)
by: Peraltai, Daniel, et al.
Published: (2024)
MobilityBench: A Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios
by: Song, Zhiheng, et al.
Published: (2026)
by: Song, Zhiheng, et al.
Published: (2026)
MDGYM: Benchmarking AI Agents on Molecular Simulations
by: Kumar, Vinay, et al.
Published: (2026)
by: Kumar, Vinay, et al.
Published: (2026)
LLM-based Agent Simulation for Maternal Health Interventions: Uncertainty Estimation and Decision-focused Evaluation
by: Martinson, Sarah, et al.
Published: (2025)
by: Martinson, Sarah, et al.
Published: (2025)
Simulation-based Scenario Generation for Robust Hybrid AI for Autonomy
by: Keno, Hambisa, et al.
Published: (2024)
by: Keno, Hambisa, et al.
Published: (2024)
Multimodal Safety Evaluation in Generative Agent Social Simulations
by: Vera, Alhim, et al.
Published: (2025)
by: Vera, Alhim, et al.
Published: (2025)
AgentSimulator: An Agent-based Approach for Data-driven Business Process Simulation
by: Kirchdorfer, Lukas, et al.
Published: (2024)
by: Kirchdorfer, Lukas, et al.
Published: (2024)
Efficient Agent Evaluation via Diversity-Guided User Simulation
by: Nakash, Itay, et al.
Published: (2026)
by: Nakash, Itay, et al.
Published: (2026)
Simulating Policy Impacts: Developing a Generative Scenario Writing Method to Evaluate the Perceived Effects of Regulation
by: Barnett, Julia, et al.
Published: (2024)
by: Barnett, Julia, et al.
Published: (2024)
Evaluation of Agents under Simulated AI Marketplace Dynamics
by: Kim, To Eun, et al.
Published: (2026)
by: Kim, To Eun, et al.
Published: (2026)
Test Automation for Interactive Scenarios via Promptable Traffic Simulation
by: Mondelli, Augusto, et al.
Published: (2025)
by: Mondelli, Augusto, et al.
Published: (2025)
LLM-Agent-based Social Simulation for Attitude Diffusion
by: Reji, Deepak John
Published: (2026)
by: Reji, Deepak John
Published: (2026)
ShopSimulator: Evaluating and Exploring RL-Driven LLM Agent for Shopping Assistants
by: Wang, Pei, et al.
Published: (2026)
by: Wang, Pei, et al.
Published: (2026)
ClawsBench: Evaluating Capability and Safety of LLM Productivity Agents in Simulated Workspaces
by: Li, Xiangyi, et al.
Published: (2026)
by: Li, Xiangyi, et al.
Published: (2026)
The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios
by: Fu, Daocheng, et al.
Published: (2026)
by: Fu, Daocheng, et al.
Published: (2026)
TutorGym: A Testbed for Evaluating AI Agents as Tutors and Students
by: Weitekamp, Daniel, et al.
Published: (2025)
by: Weitekamp, Daniel, et al.
Published: (2025)
Functional Component Ablation Reveals Specialization Patterns in Hybrid Language Model Architectures
by: Borobia, Hector, et al.
Published: (2026)
by: Borobia, Hector, et al.
Published: (2026)
InfiAgent: Self-Evolving Pyramid Agent Framework for Infinite Scenarios
by: Yu, Chenglin, et al.
Published: (2025)
by: Yu, Chenglin, et al.
Published: (2025)
AgentSUMO: An Agentic Framework for Interactive Simulation Scenario Generation in SUMO via Large Language Models
by: Jeong, Minwoo, et al.
Published: (2025)
by: Jeong, Minwoo, et al.
Published: (2025)
Evaluating Generalization Capabilities of LLM-Based Agents in Mixed-Motive Scenarios Using Concordia
by: Smith, Chandler, et al.
Published: (2025)
by: Smith, Chandler, et al.
Published: (2025)
Improving Generalization of Speech Separation in Real-World Scenarios: Strategies in Simulation, Optimization, and Evaluation
by: Chen, Ke, et al.
Published: (2024)
by: Chen, Ke, et al.
Published: (2024)
Jenius Agent: Towards Experience-Driven Accuracy Optimization in Real-World Scenarios
by: Xia, Defei, et al.
Published: (2026)
by: Xia, Defei, et al.
Published: (2026)
CrashAgent: Crash Scenario Generation via Multi-modal Reasoning
by: Li, Miao, et al.
Published: (2025)
by: Li, Miao, et al.
Published: (2025)
LeHome: A Simulation Environment for Deformable Object Manipulation in Household Scenarios
by: Li, Zeyi, et al.
Published: (2026)
by: Li, Zeyi, et al.
Published: (2026)
Evaluating Software Development Agents: Patch Patterns, Code Quality, and Issue Complexity in Real-World GitHub Scenarios
by: Chen, Zhi, et al.
Published: (2024)
by: Chen, Zhi, et al.
Published: (2024)
Decision-aware User Simulation Agent for Evaluating Conversational Recommender Systems
by: Li, Yuan-Chi, et al.
Published: (2026)
by: Li, Yuan-Chi, et al.
Published: (2026)
Similar Items
-
Attacks Meet Interpretability (AmI) Evaluation and Findings
by: Ma, Qian, et al.
Published: (2023) -
Trust model of privacy-concerned, emotionally-aware agents in a cooperative logistics problem
by: Carbo, J., et al.
Published: (2024) -
Merging plans with incomplete knowledge about actions and goals through an agent-based reputation system
by: Carbo, Javier, et al.
Published: (2024) -
I Am Big, You Are Little; I Am Right, You Are Wrong
by: Kelly, David A., et al.
Published: (2025) -
CARJAN: Agent-Based Generation and Simulation of Traffic Scenarios with AJAN
by: Neis, Leonard Frank, et al.
Published: (2025)