Salvato in:
| Autori principali: | Li, Jiahao Nick, Zhang, Zhuohao Jerry, Ma, Jiaju |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2409.08250 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
OmniActions: Predicting Digital Actions in Response to Real-World Multimodal Sensory Inputs with LLMs
di: Li, Jiahao Nick, et al.
Pubblicazione: (2024)
di: Li, Jiahao Nick, et al.
Pubblicazione: (2024)
SensorChat: Answering Qualitative and Quantitative Questions during Long-Term Multimodal Sensor Interactions
di: Yu, Xiaofan, et al.
Pubblicazione: (2025)
di: Yu, Xiaofan, et al.
Pubblicazione: (2025)
TreeHop: Generate and Filter Next Query Embeddings Efficiently for Multi-hop Question Answering
di: Li, Zhonghao, et al.
Pubblicazione: (2025)
di: Li, Zhonghao, et al.
Pubblicazione: (2025)
Question Answering for Decisionmaking in Green Building Design: A Multimodal Data Reasoning Method Driven by Large Language Models
di: Li, Yihui, et al.
Pubblicazione: (2024)
di: Li, Yihui, et al.
Pubblicazione: (2024)
OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environments
di: Henry, Felix, et al.
Pubblicazione: (2026)
di: Henry, Felix, et al.
Pubblicazione: (2026)
Understanding and Supporting Formal Email Exchange by Answering AI-Generated Questions
di: Miura, Yusuke, et al.
Pubblicazione: (2025)
di: Miura, Yusuke, et al.
Pubblicazione: (2025)
RAG-VR: Leveraging Retrieval-Augmented Generation for 3D Question Answering in VR Environments
di: Ding, Shiyi, et al.
Pubblicazione: (2025)
di: Ding, Shiyi, et al.
Pubblicazione: (2025)
Enabling Personalized Long-term Interactions in LLM-based Agents through Persistent Memory and User Profiles
di: Westhäußer, Rebecca, et al.
Pubblicazione: (2025)
di: Westhäußer, Rebecca, et al.
Pubblicazione: (2025)
Personalized to Persuade: The Effects of Contextualization and Warmth on Trust and Reliance in Conversational AI
di: Yazan, Mert, et al.
Pubblicazione: (2026)
di: Yazan, Mert, et al.
Pubblicazione: (2026)
OpenOmni: A Collaborative Open Source Tool for Building Future-Ready Multimodal Conversational Agents
di: Sun, Qiang, et al.
Pubblicazione: (2024)
di: Sun, Qiang, et al.
Pubblicazione: (2024)
Say It My Way: Exploring Control in Conversational Visual Question Answering with Blind Users
di: Zeraati, Farnaz Zamiri, et al.
Pubblicazione: (2026)
di: Zeraati, Farnaz Zamiri, et al.
Pubblicazione: (2026)
QACP: An Annotated Question Answering Dataset for Assisting Chinese Python Programming Learners
di: Xiao, Rui, et al.
Pubblicazione: (2024)
di: Xiao, Rui, et al.
Pubblicazione: (2024)
ClearFairy: Capturing Creative Workflows through Decision Structuring, In-Situ Questioning, and Rationale Inference
di: Son, Kihoon, et al.
Pubblicazione: (2025)
di: Son, Kihoon, et al.
Pubblicazione: (2025)
A Survey of Large Language Model Agents for Question Answering
di: Yue, Murong
Pubblicazione: (2025)
di: Yue, Murong
Pubblicazione: (2025)
MapAgent: Trajectory-Constructed Memory-Augmented Planning for Mobile Task Automation
di: Kong, Yi, et al.
Pubblicazione: (2025)
di: Kong, Yi, et al.
Pubblicazione: (2025)
Open-Ended Multi-Modal Relational Reasoning for Video Question Answering
di: Luo, Haozheng, et al.
Pubblicazione: (2020)
di: Luo, Haozheng, et al.
Pubblicazione: (2020)
Enabling On-Device LLMs Personalization with Smartphone Sensing
di: Zhang, Shiquan, et al.
Pubblicazione: (2024)
di: Zhang, Shiquan, et al.
Pubblicazione: (2024)
CUPID: Evaluating Personalized and Contextualized Alignment of LLMs from Interactions
di: Kim, Tae Soo, et al.
Pubblicazione: (2025)
di: Kim, Tae Soo, et al.
Pubblicazione: (2025)
Evaluating Contextually Personalized Programming Exercises Created with Generative AI
di: Logacheva, Evanfiya, et al.
Pubblicazione: (2024)
di: Logacheva, Evanfiya, et al.
Pubblicazione: (2024)
AgentEconomist: An End-to-end Agentic System Translating Economic Intuitions into Executable Computational Experiments
di: Chen, Jiaju, et al.
Pubblicazione: (2026)
di: Chen, Jiaju, et al.
Pubblicazione: (2026)
Can LLMs Address Mental Health Questions? A Comparison with Human Therapists
di: Wang, Synthia, et al.
Pubblicazione: (2025)
di: Wang, Synthia, et al.
Pubblicazione: (2025)
Comparing RAG and GraphRAG for Page-Level Retrieval Question Answering on Math Textbook
di: Chen, Eason, et al.
Pubblicazione: (2025)
di: Chen, Eason, et al.
Pubblicazione: (2025)
OmniResponse: Online Multimodal Conversational Response Generation in Dyadic Interactions
di: Luo, Cheng, et al.
Pubblicazione: (2025)
di: Luo, Cheng, et al.
Pubblicazione: (2025)
RealitySummary: Exploring On-Demand Mixed Reality Text Summarization and Question Answering using Large Language Models
di: Gunturu, Aditya, et al.
Pubblicazione: (2024)
di: Gunturu, Aditya, et al.
Pubblicazione: (2024)
AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?
di: Gor, Maharshi, et al.
Pubblicazione: (2026)
di: Gor, Maharshi, et al.
Pubblicazione: (2026)
OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web
di: Kapoor, Raghav, et al.
Pubblicazione: (2024)
di: Kapoor, Raghav, et al.
Pubblicazione: (2024)
See or Recall: A Sanity Check for the Role of Vision in Solving Visualization Question Answer Tasks with Multimodal LLMs
di: Li, Zhimin, et al.
Pubblicazione: (2025)
di: Li, Zhimin, et al.
Pubblicazione: (2025)
Gensors: Authoring Personalized Visual Sensors with Multimodal Foundation Models and Reasoning
di: Liu, Michael Xieyang, et al.
Pubblicazione: (2025)
di: Liu, Michael Xieyang, et al.
Pubblicazione: (2025)
Contextualized Counterspeech: Strategies for Adaptation, Personalization, and Evaluation
di: Cima, Lorenzo, et al.
Pubblicazione: (2024)
di: Cima, Lorenzo, et al.
Pubblicazione: (2024)
GhostWriter: Augmenting Collaborative Human-AI Writing Experiences Through Personalization and Agency
di: Yeh, Catherine, et al.
Pubblicazione: (2024)
di: Yeh, Catherine, et al.
Pubblicazione: (2024)
What Do People Want to Know About Artificial Intelligence (AI)? The Importance of Answering End-User Questions to Explain Autonomous Vehicle (AV) Decisions
di: Molaei, Somayeh, et al.
Pubblicazione: (2025)
di: Molaei, Somayeh, et al.
Pubblicazione: (2025)
AffectAI-Capture: A Reproducible Multimodal Protocol for Small-Group Meeting Research
di: Seikavandi, Meisam Jamshidi, et al.
Pubblicazione: (2026)
di: Seikavandi, Meisam Jamshidi, et al.
Pubblicazione: (2026)
Resonance: Drawing from Memories to Imagine Positive Futures through AI-Augmented Journaling
di: Zulfikar, Wazeer, et al.
Pubblicazione: (2025)
di: Zulfikar, Wazeer, et al.
Pubblicazione: (2025)
Towards Real-World Validity in Generative AI Benchmarks: Understanding and Designing Domain-Centered Evaluations for Journalism Practitioners
di: Li, Charlotte, et al.
Pubblicazione: (2025)
di: Li, Charlotte, et al.
Pubblicazione: (2025)
PersonaAI: Leveraging Retrieval-Augmented Generation and Personalized Context for AI-Driven Digital Avatars
di: Kimara, Elvis, et al.
Pubblicazione: (2025)
di: Kimara, Elvis, et al.
Pubblicazione: (2025)
Through the Lens of Human-Human Collaboration: A Configurable Research Platform for Exploring Human-Agent Collaboration
di: Yao, Bingsheng, et al.
Pubblicazione: (2025)
di: Yao, Bingsheng, et al.
Pubblicazione: (2025)
Cognitive Prosthetic: An AI-Enabled Multimodal System for Episodic Recall in Knowledge Work
di: Obiuwevwi, Lawrence, et al.
Pubblicazione: (2026)
di: Obiuwevwi, Lawrence, et al.
Pubblicazione: (2026)
NaviSense: A Multimodal Assistive Mobile application for Object Retrieval by Persons with Visual Impairment
di: Sridhar, Ajay Narayanan, et al.
Pubblicazione: (2025)
di: Sridhar, Ajay Narayanan, et al.
Pubblicazione: (2025)
Are We Asking the Right Questions? On Ambiguity in Natural Language Queries for Tabular Data Analysis
di: Gomm, Daniel, et al.
Pubblicazione: (2025)
di: Gomm, Daniel, et al.
Pubblicazione: (2025)
PILAR: Personalizing Augmented Reality Interactions with LLM-based Human-Centric and Trustworthy Explanations for Daily Use Cases
di: Kundu, Ripan Kumar, et al.
Pubblicazione: (2025)
di: Kundu, Ripan Kumar, et al.
Pubblicazione: (2025)
Documenti analoghi
-
OmniActions: Predicting Digital Actions in Response to Real-World Multimodal Sensory Inputs with LLMs
di: Li, Jiahao Nick, et al.
Pubblicazione: (2024) -
SensorChat: Answering Qualitative and Quantitative Questions during Long-Term Multimodal Sensor Interactions
di: Yu, Xiaofan, et al.
Pubblicazione: (2025) -
TreeHop: Generate and Filter Next Query Embeddings Efficiently for Multi-hop Question Answering
di: Li, Zhonghao, et al.
Pubblicazione: (2025) -
Question Answering for Decisionmaking in Green Building Design: A Multimodal Data Reasoning Method Driven by Large Language Models
di: Li, Yihui, et al.
Pubblicazione: (2024) -
OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environments
di: Henry, Felix, et al.
Pubblicazione: (2026)