Leveraging LLM Agents for Automated Video Game Testing
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Chengjia, Tang, Lanling, Yuan, Ming, Yu, Jiongchi, Xie, Xiaofei, Bu, Jiajun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
por: Yu, Jiongchi, et al.
Publicado: (2025)
por: Yu, Jiongchi, et al.
Publicado: (2025)
ARGUS: Defending LLM Agents Against Context-Aware Prompt Injection
por: Weng, Shihao, et al.
Publicado: (2026)
por: Weng, Shihao, et al.
Publicado: (2026)
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
por: Jiang, Weipeng, et al.
Publicado: (2025)
por: Jiang, Weipeng, et al.
Publicado: (2025)
Human in the Loop for Fuzz Testing: Literature Review and the Road Ahead
por: Yu, Jiongchi, et al.
Publicado: (2026)
por: Yu, Jiongchi, et al.
Publicado: (2026)
Chimera: Harnessing Multi-Agent LLMs for Automatic Insider Threat Simulation
por: Yu, Jiongchi, et al.
Publicado: (2025)
por: Yu, Jiongchi, et al.
Publicado: (2025)
Defects4C: Benchmarking Large Language Model Repair Capability with C/C++ Bugs
por: Wang, Jian, et al.
Publicado: (2025)
por: Wang, Jian, et al.
Publicado: (2025)
SpecGen: Automated Generation of Formal Program Specifications via Large Language Models
por: Ma, Lezhi, et al.
Publicado: (2024)
por: Ma, Lezhi, et al.
Publicado: (2024)
Bias Testing and Mitigation in LLM-based Code Generation
por: Huang, Dong, et al.
Publicado: (2023)
por: Huang, Dong, et al.
Publicado: (2023)
Understanding the Supply Chain and Risks of Large Language Model Applications
por: Ma, Yujie, et al.
Publicado: (2025)
por: Ma, Yujie, et al.
Publicado: (2025)
CXXCrafter: An LLM-Based Agent for Automated C/C++ Open Source Software Building
por: Yu, Zhengmin, et al.
Publicado: (2025)
por: Yu, Zhengmin, et al.
Publicado: (2025)
LogiAgent: Automated Logical Testing for REST Systems with LLM-Based Multi-Agents
por: Zhang, Ke, et al.
Publicado: (2025)
por: Zhang, Ke, et al.
Publicado: (2025)
MoDitector: Module-Directed Testing for Autonomous Driving Systems
por: Wang, Renzhi, et al.
Publicado: (2025)
por: Wang, Renzhi, et al.
Publicado: (2025)
From Exploration to Specification: LLM-Based Property Generation for Mobile App Testing
por: Xiong, Yiheng, et al.
Publicado: (2026)
por: Xiong, Yiheng, et al.
Publicado: (2026)
Beyond Accuracy: Policy Invariance as a Reliability Test for LLM Safety Judges
por: Weng, Shihao, et al.
Publicado: (2026)
por: Weng, Shihao, et al.
Publicado: (2026)
Intention is All You Need: Refining Your Code from Your Intention
por: Guo, Qi, et al.
Publicado: (2025)
por: Guo, Qi, et al.
Publicado: (2025)
CA2: Code-Aware Agent for Automated Game Testing
por: Adaikkappan, Valliappan Chidambaram, et al.
Publicado: (2026)
por: Adaikkappan, Valliappan Chidambaram, et al.
Publicado: (2026)
ContrastRepair: Enhancing Conversation-Based Automated Program Repair via Contrastive Test Case Pairs
por: Kong, Jiaolong, et al.
Publicado: (2024)
por: Kong, Jiaolong, et al.
Publicado: (2024)
Temac: Multi-Agent Collaboration for Automated Web GUI Testing
por: Liu, Chenxu, et al.
Publicado: (2025)
por: Liu, Chenxu, et al.
Publicado: (2025)
DriveTester: A Unified Platform for Simulation-Based Autonomous Driving Testing
por: Cheng, Mingfei, et al.
Publicado: (2024)
por: Cheng, Mingfei, et al.
Publicado: (2024)
What Makes a Good LLM Agent for Real-world Penetration Testing?
por: Deng, Gelei, et al.
Publicado: (2026)
por: Deng, Gelei, et al.
Publicado: (2026)
Towards Automated Crowdsourced Testing via Personified-LLM
por: Yu, Shengcheng, et al.
Publicado: (2026)
por: Yu, Shengcheng, et al.
Publicado: (2026)
LLM-Agents Driven Automated Simulation Testing and Analysis of small Uncrewed Aerial Systems
por: Duvvuru, Venkata Sai Aswath, et al.
Publicado: (2025)
por: Duvvuru, Venkata Sai Aswath, et al.
Publicado: (2025)
CAShift: Benchmarking Log-Based Cloud Attack Detection under Normality Shift
por: Yu, Jiongchi, et al.
Publicado: (2025)
por: Yu, Jiongchi, et al.
Publicado: (2025)
Enhancing Automated Program Repair via Faulty Token Localization and Quality-Aware Patch Refinement
por: Kong, Jiaolong, et al.
Publicado: (2025)
por: Kong, Jiaolong, et al.
Publicado: (2025)
KTester: Leveraging Domain and Testing Knowledge for More Effective LLM-based Test Generation
por: Li, Anji, et al.
Publicado: (2025)
por: Li, Anji, et al.
Publicado: (2025)
LLM Agents for Automated Dependency Upgrades
por: Tawosi, Vali, et al.
Publicado: (2025)
por: Tawosi, Vali, et al.
Publicado: (2025)
AgentRaft: Automated Detection of Data Over-Exposure in LLM Agents
por: Lin, Yixi, et al.
Publicado: (2026)
por: Lin, Yixi, et al.
Publicado: (2026)
Weaponizing the Commons: A Taxonomy and Detection Framework of Abuse on GitHub
por: Cheng, Yuli, et al.
Publicado: (2026)
por: Cheng, Yuli, et al.
Publicado: (2026)
AutoMT: A Multi-Agent LLM Framework for Automated Metamorphic Testing of Autonomous Driving Systems
por: Liang, Linfeng, et al.
Publicado: (2025)
por: Liang, Linfeng, et al.
Publicado: (2025)
UTFix: Change Aware Unit Test Repairing using LLM
por: Rahman, Shanto, et al.
Publicado: (2025)
por: Rahman, Shanto, et al.
Publicado: (2025)
Themis: Automatic and Efficient Deep Learning System Testing with Strong Fault Detection Capability
por: Huang, Dong, et al.
Publicado: (2024)
por: Huang, Dong, et al.
Publicado: (2024)
TrajAudit: Automated Failure Diagnosis for Agentic Coding Systems
por: Wang, Minxing, et al.
Publicado: (2026)
por: Wang, Minxing, et al.
Publicado: (2026)
Deploy-Master: Automating the Deployment of 50,000+ Agent-Ready Scientific Tools in One Day
por: Wang, Yi, et al.
Publicado: (2026)
por: Wang, Yi, et al.
Publicado: (2026)
Model-Enhanced LLM-Driven VUI Testing of VPA Apps
por: Li, Suwan, et al.
Publicado: (2024)
por: Li, Suwan, et al.
Publicado: (2024)
An Empirical Study on Leveraging Images in Automated Bug Report Reproduction
por: Wang, Dingbang, et al.
Publicado: (2025)
por: Wang, Dingbang, et al.
Publicado: (2025)
STCLocker: Deadlock Avoidance Testing for Autonomous Driving Systems
por: Cheng, Mingfei, et al.
Publicado: (2025)
por: Cheng, Mingfei, et al.
Publicado: (2025)
FT2Ra: A Fine-Tuning-Inspired Approach to Retrieval-Augmented Code Completion
por: Guo, Qi, et al.
Publicado: (2024)
por: Guo, Qi, et al.
Publicado: (2024)
Exploring and Lifting the Robustness of LLM-powered Automated Program Repair with Metamorphic Testing
por: Xue, Pengyu, et al.
Publicado: (2024)
por: Xue, Pengyu, et al.
Publicado: (2024)
Evaluating LLM Agents on Automated Software Analysis Tasks
por: Bouzenia, Islem, et al.
Publicado: (2026)
por: Bouzenia, Islem, et al.
Publicado: (2026)
A Comprehensive Study on Static Application Security Testing (SAST) Tools for Android
por: Zhu, Jingyun, et al.
Publicado: (2024)
por: Zhu, Jingyun, et al.
Publicado: (2024)
Ejemplares similares
-
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
por: Yu, Jiongchi, et al.
Publicado: (2025) -
ARGUS: Defending LLM Agents Against Context-Aware Prompt Injection
por: Weng, Shihao, et al.
Publicado: (2026) -
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
por: Jiang, Weipeng, et al.
Publicado: (2025) -
Human in the Loop for Fuzz Testing: Literature Review and the Road Ahead
por: Yu, Jiongchi, et al.
Publicado: (2026) -
Chimera: Harnessing Multi-Agent LLMs for Automatic Insider Threat Simulation
por: Yu, Jiongchi, et al.
Publicado: (2025)