ScreenQA: Large-Scale Question-Answer Pairs over Mobile App Screenshots
Fuente:
arXiv
Saved in:
| Main Authors: | Hsiao, Yu-Chung, Zubach, Fedir, Baechler, Gilles, Sunkara, Srinivas, Carbune, Victor, Lin, Jason, Wang, Maria, Zhu, Yun, Chen, Jindong |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
by: Baechler, Gilles, et al.
Published: (2024)
by: Baechler, Gilles, et al.
Published: (2024)
WebQuest: A Benchmark for Multimodal QA on Web Page Sequences
by: Wang, Maria, et al.
Published: (2024)
by: Wang, Maria, et al.
Published: (2024)
Chart-based Reasoning: Transferring Capabilities from LLMs to VLMs
by: Carbune, Victor, et al.
Published: (2024)
by: Carbune, Victor, et al.
Published: (2024)
UISim: An Interactive Image-Based UI Simulator for Dynamic Mobile Environments
by: Xiang, Jiannan, et al.
Published: (2025)
by: Xiang, Jiannan, et al.
Published: (2025)
QA-Noun: Representing Nominal Semantics via Natural Language Question-Answer Pairs
by: Tseytlin, Maria, et al.
Published: (2025)
by: Tseytlin, Maria, et al.
Published: (2025)
EEE-QA: Exploring Effective and Efficient Question-Answer Representations
by: Hu, Zhanghao, et al.
Published: (2024)
by: Hu, Zhanghao, et al.
Published: (2024)
ExpertQA: Expert-Curated Questions and Attributed Answers
by: Malaviya, Chaitanya, et al.
Published: (2023)
by: Malaviya, Chaitanya, et al.
Published: (2023)
Rehearsing Answers to Probable Questions with Perspective-Taking
by: Shih, Yung-Yu, et al.
Published: (2024)
by: Shih, Yung-Yu, et al.
Published: (2024)
MentalQA: An Annotated Arabic Corpus for Questions and Answers of Mental Healthcare
by: Alhuzali, Hassan, et al.
Published: (2024)
by: Alhuzali, Hassan, et al.
Published: (2024)
Wrong Answers Can Also Be Useful: PlausibleQA -- A Large-Scale QA Dataset with Answer Plausibility Scores
by: Mozafari, Jamshid, et al.
Published: (2025)
by: Mozafari, Jamshid, et al.
Published: (2025)
TimelineKGQA: A Comprehensive Question-Answer Pair Generator for Temporal Knowledge Graphs
by: Sun, Qiang, et al.
Published: (2025)
by: Sun, Qiang, et al.
Published: (2025)
Scientific QA System with Verifiable Answers
by: Ljajić, Adela, et al.
Published: (2024)
by: Ljajić, Adela, et al.
Published: (2024)
Unlocking the conversion of Web Screenshots into HTML Code with the WebSight Dataset
by: Laurençon, Hugo, et al.
Published: (2024)
by: Laurençon, Hugo, et al.
Published: (2024)
FairytaleQA Translated: Enabling Educational Question and Answer Generation in Less-Resourced Languages
by: Leite, Bernardo, et al.
Published: (2024)
by: Leite, Bernardo, et al.
Published: (2024)
EduVidQA: Generating and Evaluating Long-form Answers to Student Questions based on Lecture Videos
by: Ray, Sourjyadip, et al.
Published: (2025)
by: Ray, Sourjyadip, et al.
Published: (2025)
Improving Language Understanding from Screenshots
by: Gao, Tianyu, et al.
Published: (2024)
by: Gao, Tianyu, et al.
Published: (2024)
SWE-QA: Can Language Models Answer Repository-level Code Questions?
by: Peng, Weihan, et al.
Published: (2025)
by: Peng, Weihan, et al.
Published: (2025)
Ego2Web: A Web Agent Benchmark Grounded in Egocentric Videos
by: Yu, Shoubin, et al.
Published: (2026)
by: Yu, Shoubin, et al.
Published: (2026)
Beyond Screenshots: Evaluating VLMs' Understanding of UI Animations
by: Liang, Chen, et al.
Published: (2026)
by: Liang, Chen, et al.
Published: (2026)
RealTime QA: What's the Answer Right Now?
by: Kasai, Jungo, et al.
Published: (2022)
by: Kasai, Jungo, et al.
Published: (2022)
LLMs Provide Unstable Answers to Legal Questions
by: Blair-Stanek, Andrew, et al.
Published: (2025)
by: Blair-Stanek, Andrew, et al.
Published: (2025)
FHIRPath-QA: Executable Question Answering over FHIR Electronic Health Records
by: Frew, Michael, et al.
Published: (2026)
by: Frew, Michael, et al.
Published: (2026)
Compound-QA: A Benchmark for Evaluating LLMs on Compound Questions
by: Hou, Yutao, et al.
Published: (2024)
by: Hou, Yutao, et al.
Published: (2024)
MeDiSumQA: Patient-Oriented Question-Answer Generation from Discharge Letters
by: Dada, Amin, et al.
Published: (2025)
by: Dada, Amin, et al.
Published: (2025)
AppGen: Mobility-aware App Usage Behavior Generation for Mobile Users
by: Huang, Zihan, et al.
Published: (2024)
by: Huang, Zihan, et al.
Published: (2024)
Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning
by: Chung, Ho-Lam, et al.
Published: (2025)
by: Chung, Ho-Lam, et al.
Published: (2025)
Is That Your Final Answer? Test-Time Scaling Improves Selective Question Answering
by: Jurayj, William, et al.
Published: (2025)
by: Jurayj, William, et al.
Published: (2025)
pdfQA: Diverse, Challenging, and Realistic Question Answering over PDFs
by: Schimanski, Tobias, et al.
Published: (2026)
by: Schimanski, Tobias, et al.
Published: (2026)
ASTRA-QA: A Benchmark for Abstract Question Answering over Documents
by: Wang, Shu, et al.
Published: (2026)
by: Wang, Shu, et al.
Published: (2026)
MobQA: A Benchmark Dataset for Semantic Understanding of Human Mobility Data through Question Answering
by: Asano, Hikaru, et al.
Published: (2025)
by: Asano, Hikaru, et al.
Published: (2025)
Return of EM: Entity-driven Answer Set Expansion for QA Evaluation
by: Lee, Dongryeol, et al.
Published: (2024)
by: Lee, Dongryeol, et al.
Published: (2024)
VQA Training Sets are Self-play Environments for Generating Few-shot Pools
by: Misiunas, Tautvydas, et al.
Published: (2024)
by: Misiunas, Tautvydas, et al.
Published: (2024)
Decomposed Prompting to Answer Questions on a Course Discussion Board
by: Jaipersaud, Brandon, et al.
Published: (2024)
by: Jaipersaud, Brandon, et al.
Published: (2024)
Automatic Feedback Generation for Short Answer Questions using Answer Diagnostic Graphs
by: Furuhashi, Momoka, et al.
Published: (2025)
by: Furuhashi, Momoka, et al.
Published: (2025)
PolQA: Polish Question Answering Dataset
by: Rybak, Piotr, et al.
Published: (2022)
by: Rybak, Piotr, et al.
Published: (2022)
Question Answering with LLMs and Learning from Answer Sets
by: Borroto, Manuel, et al.
Published: (2025)
by: Borroto, Manuel, et al.
Published: (2025)
Quantifying over Optimum Answer Sets
by: Mazzotta, Giuseppe, et al.
Published: (2024)
by: Mazzotta, Giuseppe, et al.
Published: (2024)
Answering Questions in Stages: Prompt Chaining for Contract QA
by: Roegiest, Adam, et al.
Published: (2024)
by: Roegiest, Adam, et al.
Published: (2024)
Knowledge-Augmented Question Error Correction for Chinese Question Answer System with QuestionRAG
by: Qiu, Longpeng, et al.
Published: (2025)
by: Qiu, Longpeng, et al.
Published: (2025)
Question-Answering (QA) Model for a Personalized Learning Assistant for Arabic Language
by: Sammoudi, Mohammad, et al.
Published: (2024)
by: Sammoudi, Mohammad, et al.
Published: (2024)
Similar Items
-
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
by: Baechler, Gilles, et al.
Published: (2024) -
WebQuest: A Benchmark for Multimodal QA on Web Page Sequences
by: Wang, Maria, et al.
Published: (2024) -
Chart-based Reasoning: Transferring Capabilities from LLMs to VLMs
by: Carbune, Victor, et al.
Published: (2024) -
UISim: An Interactive Image-Based UI Simulator for Dynamic Mobile Environments
by: Xiang, Jiannan, et al.
Published: (2025) -
QA-Noun: Representing Nominal Semantics via Natural Language Question-Answer Pairs
by: Tseytlin, Maria, et al.
Published: (2025)