PhotoScout: Synthesis-Powered Multi-Modal Image Search
Fuente:
arXiv
Saved in:
| Main Authors: | Barnaby, Celeste, Chen, Qiaochu, Wang, Chenglong, Dillig, Isil |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking Dataset Discovery with DataScout
by: Lin, Rachel, et al.
Published: (2025)
by: Lin, Rachel, et al.
Published: (2025)
CellScout: Visual Analytics for Mining Biomarkers in Cell State Discovery
by: Sheng, Rui, et al.
Published: (2025)
by: Sheng, Rui, et al.
Published: (2025)
Active Learning for Neurosymbolic Program Synthesis
by: Barnaby, Celeste, et al.
Published: (2025)
by: Barnaby, Celeste, et al.
Published: (2025)
Lightweight and Generalizable Multi-Sensor Human Activity Recognition via Cascaded Fusion and Style-Augmented Decomposition
by: Chenglong, Wang, et al.
Published: (2026)
by: Chenglong, Wang, et al.
Published: (2026)
Understanding Modality Preferences in Search Clarification
by: Tavakoli, Leila, et al.
Published: (2024)
by: Tavakoli, Leila, et al.
Published: (2024)
SceneScout: Towards AI Agent-driven Access to Street View Imagery for Blind Users
by: Jain, Gaurav, et al.
Published: (2025)
by: Jain, Gaurav, et al.
Published: (2025)
Envisioning Generative Artificial Intelligence in Cartography and Mapmaking
by: Kang, Yuhao, et al.
Published: (2025)
by: Kang, Yuhao, et al.
Published: (2025)
DataScout: Automatic Data Fact Retrieval for Statement Augmentation with an LLM-Based Agent
by: Chen, Chuer, et al.
Published: (2025)
by: Chen, Chuer, et al.
Published: (2025)
LegiScout: A Visual Tool for Understanding Complex Legislation
by: Patel, Aadarsh Rajiv, et al.
Published: (2025)
by: Patel, Aadarsh Rajiv, et al.
Published: (2025)
Understanding and Co-designing Photo-based Reminiscence with Older Adults
by: Zhang, Zhongyue, et al.
Published: (2024)
by: Zhang, Zhongyue, et al.
Published: (2024)
Hypergraph Multi-Modal Learning for EEG-based Emotion Recognition in Conversation
by: Kang, Zijian, et al.
Published: (2025)
by: Kang, Zijian, et al.
Published: (2025)
RASSAR: Room Accessibility and Safety Scanning in Augmented Reality
by: Su, Xia, et al.
Published: (2024)
by: Su, Xia, et al.
Published: (2024)
ReUseIt: Synthesizing Reusable AI Agent Workflows for Web Automation
by: Liu, Yimeng, et al.
Published: (2025)
by: Liu, Yimeng, et al.
Published: (2025)
Visualizing the Invisible: A Generative AR System for Intuitive Multi-Modal Sensor Data Presentation
by: Guo, Yunqi, et al.
Published: (2024)
by: Guo, Yunqi, et al.
Published: (2024)
VCEMO: Multi-Modal Emotion Recognition for Chinese Voiceprints
by: Tang, Jinghua, et al.
Published: (2024)
by: Tang, Jinghua, et al.
Published: (2024)
DynaVis: Dynamically Synthesized UI Widgets for Visualization Editing
by: Vaithilingam, Priyan, et al.
Published: (2024)
by: Vaithilingam, Priyan, et al.
Published: (2024)
Memory Reviver: Supporting Photo-Collection Reminiscence for People with Visual Impairment via a Proactive Chatbot
by: Xu, Shuchang, et al.
Published: (2025)
by: Xu, Shuchang, et al.
Published: (2025)
How Do Analysts Understand and Verify AI-Assisted Data Analyses?
by: Gu, Ken, et al.
Published: (2023)
by: Gu, Ken, et al.
Published: (2023)
FIRETWIN: Digital Twin Advancing Multi-Modal Sensing, Interactive Analytics for Wildfire Response
by: Raha, Mayamin Hamid, et al.
Published: (2025)
by: Raha, Mayamin Hamid, et al.
Published: (2025)
Raising Awareness of Location Information Vulnerabilities in Social Media Photos using LLMs
by: Ma, Ying, et al.
Published: (2025)
by: Ma, Ying, et al.
Published: (2025)
Emotion-Aware Interaction Design in Intelligent User Interface Using Multi-Modal Deep Learning
by: Duan, Shiyu, et al.
Published: (2024)
by: Duan, Shiyu, et al.
Published: (2024)
LaMI: Large Language Models for Multi-Modal Human-Robot Interaction
by: Wang, Chao, et al.
Published: (2024)
by: Wang, Chao, et al.
Published: (2024)
The Evolving Duet of Two Modalities: A Survey on Integrating Text and Visualization for Data Communication
by: Lan, Xingyu, et al.
Published: (2026)
by: Lan, Xingyu, et al.
Published: (2026)
HeartbeatCam: Self-Triggered Photo Elicitation of Stress Events Using Wearable Sensing
by: Zhou, Boyang, et al.
Published: (2026)
by: Zhou, Boyang, et al.
Published: (2026)
How to Make Your Multi-Image Posts Popular? An Approach to Enhanced Grid for Nine Images on Social Media
by: Xi, Qi, et al.
Published: (2025)
by: Xi, Qi, et al.
Published: (2025)
Steering Large Text-to-Image Model for Abstract Art Synthesis: Preference-based Prompt Optimization and Visualization
by: Zhou, Aven-Le, et al.
Published: (2024)
by: Zhou, Aven-Le, et al.
Published: (2024)
MambaGesture: Enhancing Co-Speech Gesture Generation with Mamba and Disentangled Multi-Modality Fusion
by: Fu, Chencan, et al.
Published: (2024)
by: Fu, Chencan, et al.
Published: (2024)
MemoryDiorama: Generating Dynamic 3D Diorama from Everyday Photos for Memory Recall
by: Ihara, Keiichi, et al.
Published: (2026)
by: Ihara, Keiichi, et al.
Published: (2026)
UniMind: Unleashing the Power of LLMs for Unified Multi-Task Brain Decoding
by: Lu, Weiheng, et al.
Published: (2025)
by: Lu, Weiheng, et al.
Published: (2025)
Intent Lenses: Inferring Capture-Time Intent to Transform Opportunistic Photo Captures into Structured Visual Notes
by: Ram, Ashwin, et al.
Published: (2026)
by: Ram, Ashwin, et al.
Published: (2026)
Challenges & Opportunities with LLM-Assisted Visualization Retargeting
by: Snyder, Luke S., et al.
Published: (2025)
by: Snyder, Luke S., et al.
Published: (2025)
LLM4DESIGN: An Automated Multi-Modal System for Architectural and Environmental Design
by: Chen, Ran, et al.
Published: (2024)
by: Chen, Ran, et al.
Published: (2024)
Choose, Don't Label: Multiple-Choice Query Synthesis for Program Disambiguation
by: Barnaby, Celeste, et al.
Published: (2026)
by: Barnaby, Celeste, et al.
Published: (2026)
MagicCopy: Bring my data along with me beyond boundaries of apps
by: Vaithilingam, Priyan, et al.
Published: (2026)
by: Vaithilingam, Priyan, et al.
Published: (2026)
Virtual Takeovers in the Metaverse: Interrogating Power in Our Past and Future(s) with Multi-Layered Narratives
by: Quinn, Heather Snyder, et al.
Published: (2024)
by: Quinn, Heather Snyder, et al.
Published: (2024)
More Modality, More AI: Exploring Design Opportunities of AI-Based Multi-modal Remote Monitoring Technologies for Early Detection of Mental Health Sequelae in Youth Concussion Patients
by: Yao, Bingsheng, et al.
Published: (2025)
by: Yao, Bingsheng, et al.
Published: (2025)
A Multi-Modal Interaction Framework for Efficient Human-Robot Collaborative Shelf Picking
by: Pathak, Abhinav, et al.
Published: (2025)
by: Pathak, Abhinav, et al.
Published: (2025)
To Search or To Gen? Exploring the Synergy between Generative AI and Web Search in Programming
by: Yen, Ryan, et al.
Published: (2024)
by: Yen, Ryan, et al.
Published: (2024)
Accessibility Scout: Personalized Accessibility Scans of Built Environments
by: Huang, William, et al.
Published: (2025)
by: Huang, William, et al.
Published: (2025)
The Value, Benefits, and Concerns of Generative AI-Powered Assistance in Writing
by: Li, Zhuoyan, et al.
Published: (2024)
by: Li, Zhuoyan, et al.
Published: (2024)
Similar Items
-
Rethinking Dataset Discovery with DataScout
by: Lin, Rachel, et al.
Published: (2025) -
CellScout: Visual Analytics for Mining Biomarkers in Cell State Discovery
by: Sheng, Rui, et al.
Published: (2025) -
Active Learning for Neurosymbolic Program Synthesis
by: Barnaby, Celeste, et al.
Published: (2025) -
Lightweight and Generalizable Multi-Sensor Human Activity Recognition via Cascaded Fusion and Style-Augmented Decomposition
by: Chenglong, Wang, et al.
Published: (2026) -
Understanding Modality Preferences in Search Clarification
by: Tavakoli, Leila, et al.
Published: (2024)