Enabling Cost-Effective UI Automation Testing with Retrieval-Based LLMs: A Case Study in WeChat
Fuente:
arXiv
Guardado en:
| Autores principales: | Feng, Sidong, Lu, Haochuan, Jiang, Jianqin, Xiong, Ting, Huang, Likun, Liang, Yinglin, Li, Xiaoqin, Deng, Yuetang, Aleti, Aldeida |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Deep Dive into Retrieval-Augmented Generation for Code Completion: Experience on WeChat
por: Yang, Zezhou, et al.
Publicado: (2025)
por: Yang, Zezhou, et al.
Publicado: (2025)
Automated Prompt Generation for Code Intelligence: An Empirical study and Experience in WeChat
por: Ji, Kexing, et al.
Publicado: (2025)
por: Ji, Kexing, et al.
Publicado: (2025)
JSProtect: A Scalable Obfuscation Framework for Mini-Games in WeChat
por: Li, Zhihao, et al.
Publicado: (2025)
por: Li, Zhihao, et al.
Publicado: (2025)
ViBR: Automated Bug Replay from Video-based Reports using Vision-Language Models
por: Feng, Sidong, et al.
Publicado: (2026)
por: Feng, Sidong, et al.
Publicado: (2026)
The Role of Road Features and Vehicle Dynamics in Cost-Effective Autonomous Vehicles Safety Testing: Insights from Instance Space Analysis
por: Crespo-Rodriguez, Victor, et al.
Publicado: (2026)
por: Crespo-Rodriguez, Victor, et al.
Publicado: (2026)
PAFOT: A Position-Based Approach for Finding Optimal Tests of Autonomous Vehicles
por: Crespo-Rodriguez, Victor, et al.
Publicado: (2024)
por: Crespo-Rodriguez, Victor, et al.
Publicado: (2024)
Requirements-Driven Automated Software Testing: A Systematic Review
por: Wang, Fanyu, et al.
Publicado: (2025)
por: Wang, Fanyu, et al.
Publicado: (2025)
A Semantic-based Optimization Approach for Repairing LLMs: Case Study on Code Generation
por: Gu, Jian, et al.
Publicado: (2025)
por: Gu, Jian, et al.
Publicado: (2025)
Experimental evaluation of architectural software performance design patterns in microservices
por: Meijer, Willem, et al.
Publicado: (2024)
por: Meijer, Willem, et al.
Publicado: (2024)
UntrustVul: An Automated Approach for Identifying Untrustworthy Alerts in Vulnerability Detection Models
por: Tung, Lam Nguyen, et al.
Publicado: (2025)
por: Tung, Lam Nguyen, et al.
Publicado: (2025)
From Domain Documents to Requirements: Retrieval-Augmented Generation in the Space Industry
por: Arora, Chetan, et al.
Publicado: (2025)
por: Arora, Chetan, et al.
Publicado: (2025)
Trustworthy AI Software Engineers
por: Aleti, Aldeida, et al.
Publicado: (2026)
por: Aleti, Aldeida, et al.
Publicado: (2026)
Test-based Patch Clustering for Automatically-Generated Patches Assessment
por: Martinez, Matias, et al.
Publicado: (2022)
por: Martinez, Matias, et al.
Publicado: (2022)
Neuron Patching: Semantic-based Neuron-level Language Model Repair for Code Generation
por: Gu, Jian, et al.
Publicado: (2023)
por: Gu, Jian, et al.
Publicado: (2023)
Enhancing Large Language Models for Text-to-Testcase Generation
por: Alagarsamy, Saranya, et al.
Publicado: (2024)
por: Alagarsamy, Saranya, et al.
Publicado: (2024)
Beyond Pass or Fail: Multi-Dimensional Benchmarking of Foundation Models for Goal-based Mobile UI Navigation
por: Ran, Dezhi, et al.
Publicado: (2025)
por: Ran, Dezhi, et al.
Publicado: (2025)
MORTAR: Multi-turn Metamorphic Testing for LLM-based Dialogue Systems
por: Guo, Guoxiang, et al.
Publicado: (2024)
por: Guo, Guoxiang, et al.
Publicado: (2024)
Multi-Modal Requirements Data-based Acceptance Criteria Generation using LLMs
por: Wang, Fanyu, et al.
Publicado: (2025)
por: Wang, Fanyu, et al.
Publicado: (2025)
From User Interface to Agent Interface: Efficiency Optimization of UI Representations for LLM Agents
por: Ran, Dezhi, et al.
Publicado: (2025)
por: Ran, Dezhi, et al.
Publicado: (2025)
Empirical and Sustainability Aspects of Software Engineering Research in the Era of Large Language Models: A Reflection
por: Williams, David, et al.
Publicado: (2025)
por: Williams, David, et al.
Publicado: (2025)
Unveiling Practical Shortcomings of Patch Overfitting Detection Techniques
por: Williams, David, et al.
Publicado: (2026)
por: Williams, David, et al.
Publicado: (2026)
Automated Trustworthiness Oracle Generation for Machine Learning Text Classifiers
por: Tung, Lam Nguyen, et al.
Publicado: (2024)
por: Tung, Lam Nguyen, et al.
Publicado: (2024)
Improving Random Testing via LLM-powered UI Tarpit Escaping for Mobile Apps
por: Xu, Mengqian, et al.
Publicado: (2026)
por: Xu, Mengqian, et al.
Publicado: (2026)
Guiding ChatGPT to Fix Web UI Tests via Explanation-Consistency Checking
por: Xu, Zhuolin, et al.
Publicado: (2023)
por: Xu, Zhuolin, et al.
Publicado: (2023)
Prompting Is All You Need: Automated Android Bug Replay with Large Language Models
por: Feng, Sidong, et al.
Publicado: (2023)
por: Feng, Sidong, et al.
Publicado: (2023)
Bridging Design and Development with Automated Declarative UI Code Generation
por: Zhou, Ting, et al.
Publicado: (2024)
por: Zhou, Ting, et al.
Publicado: (2024)
JSidentify-V2: Leveraging Dynamic Memory Fingerprinting for Mini-Game Plagiarism Detection
por: Li, Zhihao, et al.
Publicado: (2025)
por: Li, Zhihao, et al.
Publicado: (2025)
Automated Tool Support for Category-Partition Testing: Design Decisions, UI and Examples of Use
por: Labiche, Yvan
Publicado: (2026)
por: Labiche, Yvan
Publicado: (2026)
Breaking Single-Tester Limits: Multi-Agent LLMs for Multi-User Feature Testing
por: Feng, Sidong, et al.
Publicado: (2025)
por: Feng, Sidong, et al.
Publicado: (2025)
Cascaded Code Editing: Large-Small Model Collaboration for Effective and Efficient Code Editing
por: Wang, Chaozheng, et al.
Publicado: (2026)
por: Wang, Chaozheng, et al.
Publicado: (2026)
MLLM-Based UI2Code Automation Guided by UI Layout Information
por: Wu, Fan, et al.
Publicado: (2025)
por: Wu, Fan, et al.
Publicado: (2025)
WeChat and the Chinese Diaspora
Publicado: (2023)
Publicado: (2023)
Skill-Adpative Imitation Learning for UI Test Reuse
por: Wu, Mengzhou, et al.
Publicado: (2024)
por: Wu, Mengzhou, et al.
Publicado: (2024)
Automated Testing of Task-based Chatbots: How Far Are We?
por: Clerissi, Diego, et al.
Publicado: (2026)
por: Clerissi, Diego, et al.
Publicado: (2026)
Test Oracle Automation in the era of LLMs
por: Molina, Facundo, et al.
Publicado: (2024)
por: Molina, Facundo, et al.
Publicado: (2024)
SmartLLMs Scheduler: A Framework for Cost-Effective LLMs Utilization
por: Liu, Yueyue, et al.
Publicado: (2025)
por: Liu, Yueyue, et al.
Publicado: (2025)
Advancing Mobile UI Testing by Learning Screen Usage Semantics
por: Khan, Safwat Ali
Publicado: (2025)
por: Khan, Safwat Ali
Publicado: (2025)
ChatUniTest: A Framework for LLM-Based Test Generation
por: Chen, Yinghao, et al.
Publicado: (2023)
por: Chen, Yinghao, et al.
Publicado: (2023)
Retrieval-Augmented Test Generation: How Far Are We?
por: Shin, Jiho, et al.
Publicado: (2024)
por: Shin, Jiho, et al.
Publicado: (2024)
MUCOCO: Automated Consistency Testing of Code LLMs
por: Chou, Chua Jin, et al.
Publicado: (2026)
por: Chou, Chua Jin, et al.
Publicado: (2026)
Ejemplares similares
-
A Deep Dive into Retrieval-Augmented Generation for Code Completion: Experience on WeChat
por: Yang, Zezhou, et al.
Publicado: (2025) -
Automated Prompt Generation for Code Intelligence: An Empirical study and Experience in WeChat
por: Ji, Kexing, et al.
Publicado: (2025) -
JSProtect: A Scalable Obfuscation Framework for Mini-Games in WeChat
por: Li, Zhihao, et al.
Publicado: (2025) -
ViBR: Automated Bug Replay from Video-based Reports using Vision-Language Models
por: Feng, Sidong, et al.
Publicado: (2026) -
The Role of Road Features and Vehicle Dynamics in Cost-Effective Autonomous Vehicles Safety Testing: Insights from Instance Space Analysis
por: Crespo-Rodriguez, Victor, et al.
Publicado: (2026)