Exploring Diagnostic Prompting Approach for Multimodal LLM-based Visual Complexity Assessment: A Case Study of Amazon Search Result Pages
Fuente:
arXiv
Saved in:
| Main Authors: | Murtadak, Divendar, Kim, Yoon, Akula, Trilokya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GenComUI: Exploring Generative Visual Aids as Medium to Support Task-Oriented Human-Robot Communication
by: Ge, Yate, et al.
Published: (2025)
by: Ge, Yate, et al.
Published: (2025)
MoXaRt: Audio-Visual Object-Guided Sound Interaction for XR
by: Xu, Tianyu, et al.
Published: (2026)
by: Xu, Tianyu, et al.
Published: (2026)
Widening the Role of Group Recommender Systems with CAJO
by: Ricci, Francesco, et al.
Published: (2025)
by: Ricci, Francesco, et al.
Published: (2025)
Emotions in the Loop: A Survey of Affective Computing for Emotional Support
by: Hegde, Karishma, et al.
Published: (2025)
by: Hegde, Karishma, et al.
Published: (2025)
MM-Conv: A Multimodal Dataset and Benchmark for Context-Aware Grounding in 3D Dialogue
by: Deichler, Anna, et al.
Published: (2026)
by: Deichler, Anna, et al.
Published: (2026)
FaceValue: Exploring Real-Time Self-View Overlays to Prompt Meaning-Oriented Self-Awareness in Remote Meetings
by: Park, Gun Woo Warren, et al.
Published: (2026)
by: Park, Gun Woo Warren, et al.
Published: (2026)
Semantic Reality: Interactive Context-Aware Visualization of Inter-Object Relationships in Augmented Reality
by: Liu, Xiaoan, et al.
Published: (2026)
by: Liu, Xiaoan, et al.
Published: (2026)
AiGet: Transforming Everyday Moments into Hidden Knowledge Discovery with AI Assistance on Smart Glasses
by: Cai, Runze, et al.
Published: (2025)
by: Cai, Runze, et al.
Published: (2025)
Distorted Perspectives of LLM-Simulated Preferences: Can AI Mislead Design?
by: Kuric, Eduard, et al.
Published: (2026)
by: Kuric, Eduard, et al.
Published: (2026)
OOPrompt: Reifying Intents into Structured Artifacts for Modular and Iterative Prompting
by: Xu, Tengyou, et al.
Published: (2026)
by: Xu, Tengyou, et al.
Published: (2026)
EvoCUA: Evolving Computer Use Agents via Learning from Scalable Synthetic Experience
by: Xue, Taofeng, et al.
Published: (2026)
by: Xue, Taofeng, et al.
Published: (2026)
OrganicHAR: Towards Activity Discovery in Organic Settings for Privacy Preserving Sensors Using Efficient Video Analysis
by: Patidar, Prasoon, et al.
Published: (2026)
by: Patidar, Prasoon, et al.
Published: (2026)
GroundLink: Exploring How Contextual Meeting Snippets Can Close Common Ground Gaps in Editing 3D Scenes for Virtual Production
by: Woo, Gun, et al.
Published: (2026)
by: Woo, Gun, et al.
Published: (2026)
Evaluating 5W3H Structured Prompting for Intent Alignment in Human-AI Interaction
by: Gang, Peng
Published: (2026)
by: Gang, Peng
Published: (2026)
The Silicon Mirror: Dynamic Behavioral Gating for Anti-Sycophancy in LLM Agents
by: Shah, Harshee Jignesh
Published: (2026)
by: Shah, Harshee Jignesh
Published: (2026)
Look and Tell: A Dataset for Multimodal Grounding Across Egocentric and Exocentric Views
by: Deichler, Anna, et al.
Published: (2025)
by: Deichler, Anna, et al.
Published: (2025)
A Reflective Storytelling Agent for Older Adults: Integrating Argumentation Schemes and Argument Mining in LLM-Based Personalised Narratives
by: Baskar, Jayalakshmi, et al.
Published: (2026)
by: Baskar, Jayalakshmi, et al.
Published: (2026)
FlyMeThrough: Human-AI Collaborative 3D Indoor Mapping with Commodity Drones
by: Su, Xia, et al.
Published: (2025)
by: Su, Xia, et al.
Published: (2025)
VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion
by: Guo, David, et al.
Published: (2025)
by: Guo, David, et al.
Published: (2025)
Conversational Forecasting Across Large Human Groups Using A Swarm of Surrogate AI Agents
by: Rosenberg, Louis, et al.
Published: (2026)
by: Rosenberg, Louis, et al.
Published: (2026)
Conversational Swarms of Humans and AI Agents enable Hybrid Collaborative Decision-making
by: Rosenberg, Louis, et al.
Published: (2024)
by: Rosenberg, Louis, et al.
Published: (2024)
ScaleMAP: Preserving Local Density and Neighborhood Structure in Low-Dimensional Embeddings
by: Poorna, Rajas, et al.
Published: (2026)
by: Poorna, Rajas, et al.
Published: (2026)
Mitigating Response Delays in Free-Form Conversations with LLM-powered Intelligent Virtual Agents
by: Maslych, Mykola, et al.
Published: (2025)
by: Maslych, Mykola, et al.
Published: (2025)
Functional Flexibility in Generative AI Interfaces: Text Editing with LLMs through Conversations, Toolbars, and Prompts
by: Lehmann, Florian, et al.
Published: (2024)
by: Lehmann, Florian, et al.
Published: (2024)
Yanyun-3: Enabling Cross-Platform Strategy Game Operation with Vision-Language Models
by: Wang, Guoyan, et al.
Published: (2025)
by: Wang, Guoyan, et al.
Published: (2025)
Exploring Hierarchical Classification Performance for Time Series Data: Dissimilarity Measures and Classifier Comparisons
by: Alagoz, Celal
Published: (2024)
by: Alagoz, Celal
Published: (2024)
Project Riley: Multimodal Multi-Agent LLM Collaboration with Emotional Reasoning and Voting
by: Ortigoso, Ana Rita, et al.
Published: (2025)
by: Ortigoso, Ana Rita, et al.
Published: (2025)
Sample-Efficient Language Model for Hinglish Conversational AI
by: Singh, Sakshi, et al.
Published: (2025)
by: Singh, Sakshi, et al.
Published: (2025)
Human Agency, Causality, and the Human Computer Interface in High-Stakes Artificial Intelligence
by: Hattab, Georges
Published: (2026)
by: Hattab, Georges
Published: (2026)
WatchHAR: Real-time On-device Human Activity Recognition System for Smartwatches
by: Yeon, Taeyoung, et al.
Published: (2025)
by: Yeon, Taeyoung, et al.
Published: (2025)
Distilled HuBERT for Mobile Speech Emotion Recognition: A Cross-Corpus Validation Study
by: Ismail, Saifelden M.
Published: (2025)
by: Ismail, Saifelden M.
Published: (2025)
MAESTRO: Adapting GUIs and Guiding Navigation with User Preferences in Conversational Agents with GUIs
by: Lee, Sangwook, et al.
Published: (2026)
by: Lee, Sangwook, et al.
Published: (2026)
Designing Transparent AI-Mediated Language Support for Intergenerational Family Communication
by: Kang, Sora, et al.
Published: (2026)
by: Kang, Sora, et al.
Published: (2026)
Documenting SME Processes with Conversational AI: From Tacit Knowledge to BPMN
by: Radhakrishnan, Unnikrishnan
Published: (2025)
by: Radhakrishnan, Unnikrishnan
Published: (2025)
Generative Confidants: How do People Experience Trust in Emotional Support from Generative AI?
by: Volpato, Riccardo, et al.
Published: (2026)
by: Volpato, Riccardo, et al.
Published: (2026)
Towards Human-AI Accessibility Mapping in India: VLM-Guided Annotations and POI-Centric Analysis in Chandigarh
by: Lalwani, Varchita, et al.
Published: (2026)
by: Lalwani, Varchita, et al.
Published: (2026)
QoSGMAA: A Robust Multi-Order Graph Attention and Adversarial Framework for Sparse QoS Prediction
by: Du, Guanchen, et al.
Published: (2025)
by: Du, Guanchen, et al.
Published: (2025)
Experimentation Accelerator: Interpretable Insights and Creative Recommendations for A/B Testing with Content-Aware ranking
by: Hu, Zhengmian, et al.
Published: (2026)
by: Hu, Zhengmian, et al.
Published: (2026)
What Would GPT Click: Practical Effects of Human-AI Behavioral Misalignment and the Cost of Synthetic Participants in User Experience
by: Kuric, Eduard, et al.
Published: (2026)
by: Kuric, Eduard, et al.
Published: (2026)
Human-Robot Dialogue Annotation for Multi-Modal Common Ground
by: Bonial, Claire, et al.
Published: (2024)
by: Bonial, Claire, et al.
Published: (2024)
Similar Items
-
GenComUI: Exploring Generative Visual Aids as Medium to Support Task-Oriented Human-Robot Communication
by: Ge, Yate, et al.
Published: (2025) -
MoXaRt: Audio-Visual Object-Guided Sound Interaction for XR
by: Xu, Tianyu, et al.
Published: (2026) -
Widening the Role of Group Recommender Systems with CAJO
by: Ricci, Francesco, et al.
Published: (2025) -
Emotions in the Loop: A Survey of Affective Computing for Emotional Support
by: Hegde, Karishma, et al.
Published: (2025) -
MM-Conv: A Multimodal Dataset and Benchmark for Context-Aware Grounding in 3D Dialogue
by: Deichler, Anna, et al.
Published: (2026)