Towards a Multimodal Document-grounded Conversational AI System for Education
Fuente:
arXiv
Saved in:
| Main Authors: | Taneja, Karan, Singh, Anjali, Goel, Ashok K. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MuDoC: An Interactive Multimodal Document-grounded Conversational AI System
by: Taneja, Karan, et al.
Published: (2025)
by: Taneja, Karan, et al.
Published: (2025)
Impact of Multimodal and Conversational AI on Learning Outcomes and Experience
by: Taneja, Karan, et al.
Published: (2026)
by: Taneja, Karan, et al.
Published: (2026)
OmniResponse: Online Multimodal Conversational Response Generation in Dyadic Interactions
by: Luo, Cheng, et al.
Published: (2025)
by: Luo, Cheng, et al.
Published: (2025)
GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents
by: Ouyang, Mingyu, et al.
Published: (2026)
by: Ouyang, Mingyu, et al.
Published: (2026)
MetaCues: Enabling Critical Engagement with Generative AI for Information Seeking and Sensemaking
by: Singh, Anjali, et al.
Published: (2026)
by: Singh, Anjali, et al.
Published: (2026)
Simulating Clinical AI Assistance using Multimodal LLMs: A Case Study in Diabetic Retinopathy
by: Barakat, Nadim, et al.
Published: (2025)
by: Barakat, Nadim, et al.
Published: (2025)
TraitSpaces: Towards Interpretable Visual Creativity for Human-AI Co-Creation
by: Luthra, Prerna
Published: (2025)
by: Luthra, Prerna
Published: (2025)
Towards user-centered interactive medical image segmentation in VR with an assistive AI agent
by: Spiegler, Pascal, et al.
Published: (2025)
by: Spiegler, Pascal, et al.
Published: (2025)
Towards a Humanized Social-Media Ecosystem: AI-Augmented HCI Design Patterns for Safety, Agency & Well-Being
by: Ameen, Mohd Ruhul, et al.
Published: (2025)
by: Ameen, Mohd Ruhul, et al.
Published: (2025)
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
by: Yang, Boyin, et al.
Published: (2025)
by: Yang, Boyin, et al.
Published: (2025)
Visual Evaluative AI: A Hypothesis-Driven Tool with Concept-Based Explanations and Weight of Evidence
by: Le, Thao, et al.
Published: (2024)
by: Le, Thao, et al.
Published: (2024)
Scene-Aware Conversational ADAS with Generative AI for Real-Time Driver Assistance
by: Han, Kyungtae, et al.
Published: (2025)
by: Han, Kyungtae, et al.
Published: (2025)
Toward a Universal Color Naming System: A Clustering-Based Approach using Multisource Data
by: Sabitkyzy, Aruzhan, et al.
Published: (2026)
by: Sabitkyzy, Aruzhan, et al.
Published: (2026)
Let's Go Real Talk: Spoken Dialogue Model for Face-to-Face Conversation
by: Park, Se Jin, et al.
Published: (2024)
by: Park, Se Jin, et al.
Published: (2024)
Design-o-meter: Towards Evaluating and Refining Graphic Designs
by: Goyal, Sahil, et al.
Published: (2024)
by: Goyal, Sahil, et al.
Published: (2024)
Milmer: a Framework for Multiple Instance Learning based Multimodal Emotion Recognition
by: Wang, Zaitian, et al.
Published: (2025)
by: Wang, Zaitian, et al.
Published: (2025)
Real-Time Intuitive AI Drawing System for Collaboration: Enhancing Human Creativity through Formal and Contextual Intent Integration
by: Song, Jookyung, et al.
Published: (2025)
by: Song, Jookyung, et al.
Published: (2025)
See-Control: A Multimodal Agent Framework for Smartphone Interaction with a Robotic Arm
by: Zhao, Haoyu, et al.
Published: (2025)
by: Zhao, Haoyu, et al.
Published: (2025)
EEG-based Multimodal Representation Learning for Emotion Recognition
by: Yin, Kang, et al.
Published: (2024)
by: Yin, Kang, et al.
Published: (2024)
Lightweight Structured Multimodal Reasoning for Clinical Scene Understanding in Robotics
by: Jha, Saurav, et al.
Published: (2025)
by: Jha, Saurav, et al.
Published: (2025)
GazeLLM: Multimodal LLMs incorporating Human Visual Attention
by: Rekimoto, Jun
Published: (2025)
by: Rekimoto, Jun
Published: (2025)
AI-based Multimodal Biometrics for Detecting Smartphone Distractions: Application to Online Learning
by: Becerra, Alvaro, et al.
Published: (2025)
by: Becerra, Alvaro, et al.
Published: (2025)
ViSIL: Unified Evaluation of Information Loss in Multimodal Video Captioning
by: Li, Po-han, et al.
Published: (2026)
by: Li, Po-han, et al.
Published: (2026)
CG-MER: A Card Game-based Multimodal dataset for Emotion Recognition
by: Farhat, Nessrine, et al.
Published: (2025)
by: Farhat, Nessrine, et al.
Published: (2025)
Toward a Machine Bertin: Why Visualization Needs Design Principles for Machine Cognition
by: Keith-Norambuena, Brian
Published: (2026)
by: Keith-Norambuena, Brian
Published: (2026)
Towards Safer and Understandable Driver Intention Prediction
by: Karuppasamy, Mukilan, et al.
Published: (2025)
by: Karuppasamy, Mukilan, et al.
Published: (2025)
Generative AI for Cel-Animation: A Survey
by: Tang, Yolo Y., et al.
Published: (2025)
by: Tang, Yolo Y., et al.
Published: (2025)
Enhancing Online Learning by Integrating Biosensors and Multimodal Learning Analytics for Detecting and Predicting Student Behavior: A Review
by: Becerra, Alvaro, et al.
Published: (2025)
by: Becerra, Alvaro, et al.
Published: (2025)
Classification Metrics for Image Explanations: Towards Building Reliable XAI-Evaluations
by: Fresz, Benjamin, et al.
Published: (2024)
by: Fresz, Benjamin, et al.
Published: (2024)
How to Distinguish AI-Generated Images from Authentic Photographs
by: Kamali, Negar, et al.
Published: (2024)
by: Kamali, Negar, et al.
Published: (2024)
Do Vision Language Models Understand Human Engagement in Games?
by: Wang, Ziyi, et al.
Published: (2026)
by: Wang, Ziyi, et al.
Published: (2026)
ASAP: Interpretable Analysis and Summarization of AI-generated Image Patterns at Scale
by: Huang, Jinbin, et al.
Published: (2024)
by: Huang, Jinbin, et al.
Published: (2024)
How good are humans at detecting AI-generated images? Learnings from an experiment
by: Roca, Thomas, et al.
Published: (2025)
by: Roca, Thomas, et al.
Published: (2025)
Sketch2Prototype: Rapid Conceptual Design Exploration and Prototyping with Generative AI
by: Edwards, Kristen M., et al.
Published: (2024)
by: Edwards, Kristen M., et al.
Published: (2024)
A User-Centric Analysis of Explainability in AI-Based Medical Image Diagnosis
by: Wagner, Julia, et al.
Published: (2026)
by: Wagner, Julia, et al.
Published: (2026)
AI-Based Facial Emotion Recognition Solutions for Education: A Study of Teacher-User and Other Categories
by: Ravenor, R. Yamamoto
Published: (2023)
by: Ravenor, R. Yamamoto
Published: (2023)
Pencils to Pixels: A Systematic Study of Creative Drawings across Children, Adults and AI
by: Nath, Surabhi S, et al.
Published: (2025)
by: Nath, Surabhi S, et al.
Published: (2025)
Dermatologist-like explainable AI enhances melanoma diagnosis accuracy: eye-tracking study
by: Chanda, Tirtha, et al.
Published: (2024)
by: Chanda, Tirtha, et al.
Published: (2024)
PedaCo-Gen: Scaffolding Pedagogical Agency in Human-AI Collaborative Video Authoring
by: Baek, Injun, et al.
Published: (2026)
by: Baek, Injun, et al.
Published: (2026)
Real-Time Feedback and Benchmark Dataset for Isometric Pose Evaluation
by: Jaiswal, Abhishek, et al.
Published: (2025)
by: Jaiswal, Abhishek, et al.
Published: (2025)
Similar Items
-
MuDoC: An Interactive Multimodal Document-grounded Conversational AI System
by: Taneja, Karan, et al.
Published: (2025) -
Impact of Multimodal and Conversational AI on Learning Outcomes and Experience
by: Taneja, Karan, et al.
Published: (2026) -
OmniResponse: Online Multimodal Conversational Response Generation in Dyadic Interactions
by: Luo, Cheng, et al.
Published: (2025) -
GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents
by: Ouyang, Mingyu, et al.
Published: (2026) -
MetaCues: Enabling Critical Engagement with Generative AI for Information Seeking and Sensemaking
by: Singh, Anjali, et al.
Published: (2026)