AI Agents for Web Testing: A Case Study in the Wild
Fuente:
arXiv
Saved in:
| Main Authors: | Ye, Naimeng, Yu, Xiao, Xu, Ruize, Peng, Tianyi, Yu, Zhou |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AmbiBench: Benchmarking Mobile GUI Agents Beyond One-Shot Instructions in the Wild
by: Sun, Jiazheng, et al.
Published: (2026)
by: Sun, Jiazheng, et al.
Published: (2026)
Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild
by: van der Maden, Willem, et al.
Published: (2026)
by: van der Maden, Willem, et al.
Published: (2026)
A Case Study Investigating the Role of Generative AI in Quality Evaluations of Epics in Agile Software Development
by: Geyer, Werner, et al.
Published: (2025)
by: Geyer, Werner, et al.
Published: (2025)
How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignment in 20,574 Real-World Sessions
by: Tang, Ningzhi, et al.
Published: (2026)
by: Tang, Ningzhi, et al.
Published: (2026)
Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security
by: Li, Yuanchun, et al.
Published: (2024)
by: Li, Yuanchun, et al.
Published: (2024)
Professional Software Developers Don't Vibe, They Control: AI Agent Use for Coding in 2025
by: Huang, Ruanqianqian, et al.
Published: (2025)
by: Huang, Ruanqianqian, et al.
Published: (2025)
Generating Proto-Personas through Prompt Engineering: A Case Study on Efficiency, Effectiveness and Empathy
by: Ayach, Fernando, et al.
Published: (2025)
by: Ayach, Fernando, et al.
Published: (2025)
VTutor: An Open-Source SDK for Generative AI-Powered Animated Pedagogical Agents with Multi-Media Output
by: Chen, Eason, et al.
Published: (2025)
by: Chen, Eason, et al.
Published: (2025)
Agentic Metacognition: Designing a "Self-Aware" Low-Code Agent for Failure Prediction and Human Handoff
by: Xu, Jiexi
Published: (2025)
by: Xu, Jiexi
Published: (2025)
From Correctness to Collaboration: Toward a Human-Centered Framework for Evaluating AI Agent Behavior in Software Engineering
by: Dong, Tao, et al.
Published: (2025)
by: Dong, Tao, et al.
Published: (2025)
ACCESS: Prompt Engineering for Automated Web Accessibility Violation Corrections
by: Huang, Calista, et al.
Published: (2024)
by: Huang, Calista, et al.
Published: (2024)
Understanding the Weakness of Large Language Model Agents within a Complex Android Environment
by: Xing, Mingzhe, et al.
Published: (2024)
by: Xing, Mingzhe, et al.
Published: (2024)
Assistance or Disruption? Exploring and Evaluating the Design and Trade-offs of Proactive AI Programming Support
by: Pu, Kevin, et al.
Published: (2025)
by: Pu, Kevin, et al.
Published: (2025)
"My productivity is boosted, but ..." Demystifying Users' Perception on AI Coding Assistants
by: Lyu, Yunbo, et al.
Published: (2025)
by: Lyu, Yunbo, et al.
Published: (2025)
How Developers Interact with AI: A Taxonomy of Human-AI Collaboration in Software Engineering
by: Treude, Christoph, et al.
Published: (2025)
by: Treude, Christoph, et al.
Published: (2025)
Towards an Appropriate Level of Reliance on AI: A Preliminary Reliance-Control Framework for AI in Software Engineering
by: Ferino, Samuel, et al.
Published: (2026)
by: Ferino, Samuel, et al.
Published: (2026)
The SPACE of AI: Real-World Lessons on AI's Impact on Developers
by: Houck, Brian, et al.
Published: (2025)
by: Houck, Brian, et al.
Published: (2025)
Interacting with AI Reasoning Models: Harnessing "Thoughts" for AI-Driven Software Engineering
by: Treude, Christoph, et al.
Published: (2025)
by: Treude, Christoph, et al.
Published: (2025)
The Ann Arbor Architecture for Agent-Oriented Programming
by: Dong, Wei
Published: (2025)
by: Dong, Wei
Published: (2025)
On AI-Inspired UI-Design
by: Wei, Jialiang, et al.
Published: (2024)
by: Wei, Jialiang, et al.
Published: (2024)
Human-AI Experience in Integrated Development Environments: A Systematic Literature Review
by: Sergeyuk, Agnia, et al.
Published: (2025)
by: Sergeyuk, Agnia, et al.
Published: (2025)
Single Conversation Methodology: A Human-Centered Protocol for AI-Assisted Software Development
by: Escobedo, Salvador D.
Published: (2025)
by: Escobedo, Salvador D.
Published: (2025)
FhGenie: A Custom, Confidentiality-preserving Chat AI for Corporate and Scientific Use
by: Weber, Ingo, et al.
Published: (2024)
by: Weber, Ingo, et al.
Published: (2024)
TableTalk: Scaffolding Spreadsheet Development with a Language Agent
by: Liang, Jenny T., et al.
Published: (2025)
by: Liang, Jenny T., et al.
Published: (2025)
AI-Guided Exploration of Large-Scale Codebases
by: Alebachew, Yoseph Berhanu
Published: (2025)
by: Alebachew, Yoseph Berhanu
Published: (2025)
AI-Enhanced Operator Assistance for UNICOS Applications
by: Tam, Bernard, et al.
Published: (2025)
by: Tam, Bernard, et al.
Published: (2025)
AI for Requirements Engineering: Industry adoption and Practitioner perspectives
by: Rani, Lekshmi Murali, et al.
Published: (2025)
by: Rani, Lekshmi Murali, et al.
Published: (2025)
Interaction2Code: Benchmarking MLLM-based Interactive Webpage Code Generation from Interactive Prototyping
by: Xiao, Jingyu, et al.
Published: (2024)
by: Xiao, Jingyu, et al.
Published: (2024)
Why AI Agents Still Need You: Findings from Developer-Agent Collaborations in the Wild
by: Kumar, Aayush, et al.
Published: (2025)
by: Kumar, Aayush, et al.
Published: (2025)
Envisioning the Next-Generation AI Coding Assistants: Insights & Proposals
by: Nghiem, Khanh, et al.
Published: (2024)
by: Nghiem, Khanh, et al.
Published: (2024)
Intelligent Front-End Personalization: AI-Driven UI Adaptation
by: Rajhans, Mona
Published: (2026)
by: Rajhans, Mona
Published: (2026)
AI for Better UX in Computer-Aided Engineering: Is Academia Catching Up with Industry Demands? A Multivocal Literature Review
by: Uulu, Choro Ulan, et al.
Published: (2025)
by: Uulu, Choro Ulan, et al.
Published: (2025)
Generative AI for Self-Adaptive Systems: State of the Art and Research Roadmap
by: Li, Jialong, et al.
Published: (2025)
by: Li, Jialong, et al.
Published: (2025)
LLM Interactive Optimization of Open Source Python Libraries -- Case Studies and Generalization
by: Florath, Andreas
Published: (2023)
by: Florath, Andreas
Published: (2023)
Trust Calibration in IDEs: Paving the Way for Widespread Adoption of AI Refactoring
by: Borg, Markus
Published: (2024)
by: Borg, Markus
Published: (2024)
Exploring Interaction Patterns for Debugging: Enhancing Conversational Capabilities of AI-assistants
by: Chopra, Bhavya, et al.
Published: (2024)
by: Chopra, Bhavya, et al.
Published: (2024)
Exploring Human-AI Collaboration in Agile: Customised LLM Meeting Assistants
by: Cabrero-Daniel, Beatriz, et al.
Published: (2024)
by: Cabrero-Daniel, Beatriz, et al.
Published: (2024)
Stakeholder Participation for Responsible AI Development: Disconnects Between Guidance and Current Practice
by: Kallina, Emma, et al.
Published: (2025)
by: Kallina, Emma, et al.
Published: (2025)
Collaborative AI in Sentiment Analysis: System Architecture, Data Prediction and Deployment Strategies
by: Zhang, Chaofeng, et al.
Published: (2024)
by: Zhang, Chaofeng, et al.
Published: (2024)
How Do Hackathons Foster Creativity? Towards AI Collaborative Evaluation of Creativity at Scale
by: Falk, Jeanette, et al.
Published: (2025)
by: Falk, Jeanette, et al.
Published: (2025)
Similar Items
-
AmbiBench: Benchmarking Mobile GUI Agents Beyond One-Shot Instructions in the Wild
by: Sun, Jiazheng, et al.
Published: (2026) -
Results-Actionability Gap: Understanding How Practitioners Evaluate LLM Products in the Wild
by: van der Maden, Willem, et al.
Published: (2026) -
A Case Study Investigating the Role of Generative AI in Quality Evaluations of Epics in Agile Software Development
by: Geyer, Werner, et al.
Published: (2025) -
How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignment in 20,574 Real-World Sessions
by: Tang, Ningzhi, et al.
Published: (2026) -
Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security
by: Li, Yuanchun, et al.
Published: (2024)