XBIDetective: Leveraging Vision Language Models for Identifying Cross-Browser Visual Inconsistencies
Fuente:
arXiv
Saved in:
| Main Authors: | Grewal, Balreet, Graham, James, Muizelaar, Jeff, Odvarko, Jan Honza, Mujahid, Suhaib, Castelluccio, Marco, Bezemer, Cor-Paul |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Predicting the Impact of Crashes Across Release Channels
by: Mujahid, Suhaib, et al.
Published: (2024)
by: Mujahid, Suhaib, et al.
Published: (2024)
Exploring the Capabilities of Vision-Language Models to Detect Visual Bugs in HTML5 <canvas> Applications
by: Macklon, Finlay, et al.
Published: (2025)
by: Macklon, Finlay, et al.
Published: (2025)
Bridging the Language Gap: An Empirical Study of Bindings for Open Source Machine Learning Libraries Across Software Package Ecosystems
by: Li, Hao, et al.
Published: (2022)
by: Li, Hao, et al.
Published: (2022)
Automated Bug Frame Retrieval from Gameplay Videos Using Vision-Language Models
by: Lu, Wentao, et al.
Published: (2025)
by: Lu, Wentao, et al.
Published: (2025)
VideoGameBunny: Towards vision assistants for video games
by: Taesiri, Mohammad Reza, et al.
Published: (2024)
by: Taesiri, Mohammad Reza, et al.
Published: (2024)
Software Engineering and Foundation Models: Insights from Industry Blogs Using a Jury of Foundation Models
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Studying the Impact of TensorFlow and PyTorch Bindings on Machine Learning Software Quality
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Assessing the Feasibility of Selective Instrumentation for Runtime Code Coverage in Large C++ Game Engines
by: Gauk, Ian, et al.
Published: (2026)
by: Gauk, Ian, et al.
Published: (2026)
Impact of LLM-based Review Comment Generation in Practice: A Mixed Open-/Closed-source User Study
by: Olewicki, Doriane, et al.
Published: (2024)
by: Olewicki, Doriane, et al.
Published: (2024)
How Far Can VLMs Go for Visual Bug Detection? Studying 19,738 Keyframes from 41 Hours of Gameplay Videos
by: Lu, Wentao, et al.
Published: (2026)
by: Lu, Wentao, et al.
Published: (2026)
A Systematic Literature Review of Software Engineering Research on Jupyter Notebook
by: Siddik, Md Saeed, et al.
Published: (2025)
by: Siddik, Md Saeed, et al.
Published: (2025)
A Taxonomy of Testable HTML5 Canvas Issues
by: Macklon, Finlay, et al.
Published: (2022)
by: Macklon, Finlay, et al.
Published: (2022)
Keeping Deep Learning Models in Check: A History-Based Approach to Mitigate Overfitting
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Early Detection of Performance Regressions by Bridging Local Performance Data and Architectural Models
by: Liao, Lizhi, et al.
Published: (2024)
by: Liao, Lizhi, et al.
Published: (2024)
Understanding the Helpfulness of Stale Bot for Pull-based Development: An Empirical Study of 20 Large Open-Source Projects
by: Khatoonabadi, SayedHassan, et al.
Published: (2023)
by: Khatoonabadi, SayedHassan, et al.
Published: (2023)
Automated Generation of Issue-Reproducing Tests by Combining LLMs and Search-Based Testing
by: Kitsios, Konstantinos, et al.
Published: (2025)
by: Kitsios, Konstantinos, et al.
Published: (2025)
Group versus Individual Review Requests: Tradeoffs in Speed and Quality at Mozilla Firefox
by: Kucera, Matej, et al.
Published: (2026)
by: Kucera, Matej, et al.
Published: (2026)
Larger Is Not Always Better: Leveraging Structured Code Diffs for Comment Inconsistency Detection
by: Nguyen, Phong, et al.
Published: (2025)
by: Nguyen, Phong, et al.
Published: (2025)
A Dataset of Performance Measurements and Alerts from Mozilla (Data Artifact)
by: Besbes, Mohamed Bilel, et al.
Published: (2025)
by: Besbes, Mohamed Bilel, et al.
Published: (2025)
DocPrism: Local Categorization and External Filtering to Identify Relevant Code-Documentation Inconsistencies
by: Xu, Xiaomeng, et al.
Published: (2025)
by: Xu, Xiaomeng, et al.
Published: (2025)
"TODO: Fix the Mess Gemini Created": Towards Understanding GenAI-Induced Self-Admitted Technical Debt
by: Mujahid, Abdullah Al, et al.
Published: (2026)
by: Mujahid, Abdullah Al, et al.
Published: (2026)
The BrowserGym Ecosystem for Web Agent Research
by: De Chezelles, Thibault Le Sellier, et al.
Published: (2024)
by: De Chezelles, Thibault Le Sellier, et al.
Published: (2024)
Reverse Browser: Vector-Image-to-Code Generator
by: Toth-Czifra, Zoltan
Published: (2025)
by: Toth-Czifra, Zoltan
Published: (2025)
Mitigating Implicit Inconsistencies in Patch Porting
by: Pan, Shengyi, et al.
Published: (2026)
by: Pan, Shengyi, et al.
Published: (2026)
Inconsistencies in TeX-Produced Documents
by: Tan, Jovyn, et al.
Published: (2024)
by: Tan, Jovyn, et al.
Published: (2024)
ASSURE: Metamorphic Testing for AI-powered Browser Extensions
by: Gao, Xuanqi, et al.
Published: (2025)
by: Gao, Xuanqi, et al.
Published: (2025)
A Comparison of Conversational Models and Humans in Answering Technical Questions: the Firefox Case
by: Correia, Joao, et al.
Published: (2025)
by: Correia, Joao, et al.
Published: (2025)
Investigating the Impact of Code Comment Inconsistency on Bug Introducing
by: Radmanesh, Shiva, et al.
Published: (2024)
by: Radmanesh, Shiva, et al.
Published: (2024)
Understanding Inconsistent State Update Vulnerabilities in Smart Contracts
by: Li, Lantian, et al.
Published: (2025)
by: Li, Lantian, et al.
Published: (2025)
HTMLCure: Turning Browser Experience into State Guided Repair for Interactive HTML
by: Wu, Jiajun, et al.
Published: (2026)
by: Wu, Jiajun, et al.
Published: (2026)
An Empirical Study on Common Defects in Modern Web Browsers Using Knowledge Embedding in GPT-4o
by: Singh, Rahul, et al.
Published: (2025)
by: Singh, Rahul, et al.
Published: (2025)
CASCADE: Detecting Inconsistencies between Code and Documentation with Automatic Test Generation
by: Kiecker, Tobias, et al.
Published: (2026)
by: Kiecker, Tobias, et al.
Published: (2026)
BenchBrowser: Retrieving Evidence for Evaluating Benchmark Validity
by: Diddee, Harshita, et al.
Published: (2026)
by: Diddee, Harshita, et al.
Published: (2026)
CUJBench: Benchmarking LLM-Agent on Cross-Modal Failure Diagnosis from Browser to Backend
by: Meng, Haoming
Published: (2026)
by: Meng, Haoming
Published: (2026)
Predicting Safety Misbehaviours in Autonomous Driving Systems using Uncertainty Quantification
by: Grewal, Ruben, et al.
Published: (2024)
by: Grewal, Ruben, et al.
Published: (2024)
Event-Driven Inconsistency Detection Between UML Class and Sequence Diagrams
by: Lazzari, Luan, et al.
Published: (2025)
by: Lazzari, Luan, et al.
Published: (2025)
Domain-constrained Synthesis of Inconsistent Key Aspects in Textual Vulnerability Descriptions
by: Han, Linyi, et al.
Published: (2025)
by: Han, Linyi, et al.
Published: (2025)
Detecting Multi-Parameter Constraint Inconsistencies in Python Data Science Libraries
by: Xu, Xiufeng, et al.
Published: (2024)
by: Xu, Xiufeng, et al.
Published: (2024)
I3DE: An IDE for Inspecting Inconsistencies in PL/SQL Code
by: Liu, Jiangshan, et al.
Published: (2024)
by: Liu, Jiangshan, et al.
Published: (2024)
An Empirical Analysis of Git Commit Logs for Potential Inconsistency in Code Clones
by: Yokomori, Reishi, et al.
Published: (2024)
by: Yokomori, Reishi, et al.
Published: (2024)
Similar Items
-
Predicting the Impact of Crashes Across Release Channels
by: Mujahid, Suhaib, et al.
Published: (2024) -
Exploring the Capabilities of Vision-Language Models to Detect Visual Bugs in HTML5 <canvas> Applications
by: Macklon, Finlay, et al.
Published: (2025) -
Bridging the Language Gap: An Empirical Study of Bindings for Open Source Machine Learning Libraries Across Software Package Ecosystems
by: Li, Hao, et al.
Published: (2022) -
Automated Bug Frame Retrieval from Gameplay Videos Using Vision-Language Models
by: Lu, Wentao, et al.
Published: (2025) -
VideoGameBunny: Towards vision assistants for video games
by: Taesiri, Mohammad Reza, et al.
Published: (2024)