Toward a Dialogue System Using a Large Language Model to Recognize User Emotions with a Camera
Fuente:
arXiv
Salvato in:
| Autori principali: | Tanioka, Hiroki, Ueta, Tetsushi, Sano, Masahiko |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
di: Menon, Anjali R., et al.
Pubblicazione: (2025)
di: Menon, Anjali R., et al.
Pubblicazione: (2025)
Human-Robot Dialogue Annotation for Multi-Modal Common Ground
di: Bonial, Claire, et al.
Pubblicazione: (2024)
di: Bonial, Claire, et al.
Pubblicazione: (2024)
SCOUT: A Situated and Multi-Modal Human-Robot Dialogue Corpus
di: Lukin, Stephanie M., et al.
Pubblicazione: (2024)
di: Lukin, Stephanie M., et al.
Pubblicazione: (2024)
Inducing Causal World Models in LLMs for Zero-Shot Physical Reasoning
di: Sharma, Aditya, et al.
Pubblicazione: (2025)
di: Sharma, Aditya, et al.
Pubblicazione: (2025)
Autonomous Navigation and Collision Avoidance for Mobile Robots: Classification and Review
di: de Carvalho, Marcus Vinicius Leal, et al.
Pubblicazione: (2024)
di: de Carvalho, Marcus Vinicius Leal, et al.
Pubblicazione: (2024)
Growing Perspectives: Modelling Embodied Perspective Taking and Inner Narrative Development Using Large Language Models
di: Patania, Sabrina, et al.
Pubblicazione: (2025)
di: Patania, Sabrina, et al.
Pubblicazione: (2025)
Emotions in the Loop: A Survey of Affective Computing for Emotional Support
di: Hegde, Karishma, et al.
Pubblicazione: (2025)
di: Hegde, Karishma, et al.
Pubblicazione: (2025)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025)
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025)
ScreenSpot-Pro: GUI Grounding for Professional High-Resolution Computer Use
di: Li, Kaixin, et al.
Pubblicazione: (2025)
di: Li, Kaixin, et al.
Pubblicazione: (2025)
Domain-Specific Fine-Tuning of Large Language Models for Interactive Robot Programming
di: Alt, Benjamin, et al.
Pubblicazione: (2023)
di: Alt, Benjamin, et al.
Pubblicazione: (2023)
PerspAct: Enhancing LLM Situated Collaboration Skills through Perspective Taking and Active Vision
di: Patania, Sabrina, et al.
Pubblicazione: (2025)
di: Patania, Sabrina, et al.
Pubblicazione: (2025)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
di: Li, Danyang, et al.
Pubblicazione: (2025)
di: Li, Danyang, et al.
Pubblicazione: (2025)
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
MerNav: A Highly Generalizable Memory-Execute-Review Framework for Zero-Shot Object Goal Navigation
di: Qi, Dekang, et al.
Pubblicazione: (2026)
di: Qi, Dekang, et al.
Pubblicazione: (2026)
LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning
di: Luijkx, Jelle, et al.
Pubblicazione: (2025)
di: Luijkx, Jelle, et al.
Pubblicazione: (2025)
Incremental Bootstrapping and Classification of Structured Scenes in a Fuzzy Ontology
di: Buoncompagni, Luca, et al.
Pubblicazione: (2024)
di: Buoncompagni, Luca, et al.
Pubblicazione: (2024)
A Survey on Vision-Language-Action Models for Embodied AI
di: Ma, Yueen, et al.
Pubblicazione: (2024)
di: Ma, Yueen, et al.
Pubblicazione: (2024)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
di: Huo, Dongjie, et al.
Pubblicazione: (2026)
di: Huo, Dongjie, et al.
Pubblicazione: (2026)
A Replicable Robotics Awareness Method Using LLM-Enabled Robotics Interaction: Evidence from a Corporate Challenge
di: Prieto, S. A., et al.
Pubblicazione: (2026)
di: Prieto, S. A., et al.
Pubblicazione: (2026)
CognitiveArm: Enabling Real-Time EEG-Controlled Prosthetic Arm Using Embodied Machine Learning
di: Basit, Abdul, et al.
Pubblicazione: (2025)
di: Basit, Abdul, et al.
Pubblicazione: (2025)
Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests
di: Sáez, Arnau Igualde, et al.
Pubblicazione: (2025)
di: Sáez, Arnau Igualde, et al.
Pubblicazione: (2025)
MaP-AVR: A Meta-Action Planner for Agents Leveraging Vision Language Models and Retrieval-Augmented Generation
di: Guo, Zhenglong, et al.
Pubblicazione: (2025)
di: Guo, Zhenglong, et al.
Pubblicazione: (2025)
RACAS: Controlling Diverse Robots With a Single Agentic System
di: Ashley, Dylan R., et al.
Pubblicazione: (2026)
di: Ashley, Dylan R., et al.
Pubblicazione: (2026)
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
di: Alanazi, Ahmed, et al.
Pubblicazione: (2025)
di: Alanazi, Ahmed, et al.
Pubblicazione: (2025)
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
Who Sees What? Structured Thought-Action Sequences for Epistemic Reasoning in LLMs
di: Annese, Luca, et al.
Pubblicazione: (2025)
di: Annese, Luca, et al.
Pubblicazione: (2025)
Anonymization-Enhanced Privacy Protection for Mobile GUI Agents: Available but Invisible
di: Zhao, Lepeng, et al.
Pubblicazione: (2026)
di: Zhao, Lepeng, et al.
Pubblicazione: (2026)
Improving Knowledge Extraction from LLMs for Task Learning through Agent Analysis
di: Kirk, James R., et al.
Pubblicazione: (2023)
di: Kirk, James R., et al.
Pubblicazione: (2023)
BRAVE: Brain-Controlled Prosthetic Arm with Voice Integration and Embodied Learning for Enhanced Mobility
di: Basit, Abdul, et al.
Pubblicazione: (2025)
di: Basit, Abdul, et al.
Pubblicazione: (2025)
TensLoRA: Tensor Alternatives for Low-Rank Adaptation
di: Marmoret, Axel, et al.
Pubblicazione: (2025)
di: Marmoret, Axel, et al.
Pubblicazione: (2025)
Text-to-Events: Synthetic Event Camera Streams from Conditional Text Input
di: Ott, Joachim, et al.
Pubblicazione: (2024)
di: Ott, Joachim, et al.
Pubblicazione: (2024)
MoodBench 1.0: An Evaluation Benchmark for Emotional Companionship Dialogue Systems
di: Jing, Haifeng, et al.
Pubblicazione: (2025)
di: Jing, Haifeng, et al.
Pubblicazione: (2025)
To Whom are You Talking? A Deep Learning Model to Endow Social Robots with Addressee Estimation Skills
di: Mazzola, Carlo, et al.
Pubblicazione: (2023)
di: Mazzola, Carlo, et al.
Pubblicazione: (2023)
Visual Aesthetic Benchmark: Can Frontier Models Judge Beauty?
di: Feng, Yichen, et al.
Pubblicazione: (2026)
di: Feng, Yichen, et al.
Pubblicazione: (2026)
Talking Tennis: Language Feedback from 3D Biomechanical Action Recognition
di: Dashore, Arushi, et al.
Pubblicazione: (2025)
di: Dashore, Arushi, et al.
Pubblicazione: (2025)
MAP: Multi-user Personalization with Collaborative LLM-powered Agents
di: Lee, Christine, et al.
Pubblicazione: (2025)
di: Lee, Christine, et al.
Pubblicazione: (2025)
Comparing Apples to Oranges: LLM-powered Multimodal Intention Prediction in an Object Categorization Task
di: Ali, Hassan, et al.
Pubblicazione: (2024)
di: Ali, Hassan, et al.
Pubblicazione: (2024)
K-MetBench: A Multi-Dimensional Benchmark for Fine-Grained Evaluation of Expert Reasoning, Locality, and Multimodality in Meteorology
di: Kim, Soyeon, et al.
Pubblicazione: (2026)
di: Kim, Soyeon, et al.
Pubblicazione: (2026)
HuMoCon: Concept Discovery for Human Motion Understanding
di: Fang, Qihang, et al.
Pubblicazione: (2025)
di: Fang, Qihang, et al.
Pubblicazione: (2025)
Perception-Consistency Multimodal Large Language Models Reasoning via Caption-Regularized Policy Optimization
di: Tu, Songjun, et al.
Pubblicazione: (2025)
di: Tu, Songjun, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
di: Menon, Anjali R., et al.
Pubblicazione: (2025) -
Human-Robot Dialogue Annotation for Multi-Modal Common Ground
di: Bonial, Claire, et al.
Pubblicazione: (2024) -
SCOUT: A Situated and Multi-Modal Human-Robot Dialogue Corpus
di: Lukin, Stephanie M., et al.
Pubblicazione: (2024) -
Inducing Causal World Models in LLMs for Zero-Shot Physical Reasoning
di: Sharma, Aditya, et al.
Pubblicazione: (2025) -
Autonomous Navigation and Collision Avoidance for Mobile Robots: Classification and Review
di: de Carvalho, Marcus Vinicius Leal, et al.
Pubblicazione: (2024)