CARMA: Context-Aware Situational Grounding of Human-Robot Group Interactions by Combining Vision-Language Models with Object and Action Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Deigmoeller, Joerg, Hasler, Stephan, Agarwal, Nakul, Tanneberg, Daniel, Belardinelli, Anna, Ghoddoosian, Reza, Wang, Chao, Ocker, Felix, Zhang, Fan, Dariush, Behzad, Gienger, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MERGE: Guided Vision-Language Models for Multi-Actor Event Reasoning and Grounding in Human-Robot Interaction
by: Deigmoeller, Joerg, et al.
Published: (2026)
by: Deigmoeller, Joerg, et al.
Published: (2026)
To Help or Not to Help: LLM-based Attentive Support for Human-Robot Group Interactions
by: Tanneberg, Daniel, et al.
Published: (2024)
by: Tanneberg, Daniel, et al.
Published: (2024)
CoPAL: Corrective Planning of Robot Actions with Large Language Models
by: Joublin, Frank, et al.
Published: (2023)
by: Joublin, Frank, et al.
Published: (2023)
LaMI: Large Language Models for Multi-Modal Human-Robot Interaction
by: Wang, Chao, et al.
Published: (2024)
by: Wang, Chao, et al.
Published: (2024)
Pose-Aware Weakly-Supervised Action Segmentation
by: Zhao, Seth Z., et al.
Published: (2025)
by: Zhao, Seth Z., et al.
Published: (2025)
ACE: Action Concept Enhancement of Video-Language Models in Procedural Videos
by: Ghoddoosian, Reza, et al.
Published: (2024)
by: Ghoddoosian, Reza, et al.
Published: (2024)
Efficient Symbolic Planning with Views
by: Hasler, Stephan, et al.
Published: (2024)
by: Hasler, Stephan, et al.
Published: (2024)
Mirror Eyes: Explainable Human-Robot Interaction at a Glance
by: Krüger, Matti, et al.
Published: (2025)
by: Krüger, Matti, et al.
Published: (2025)
Tulip Agent -- Enabling LLM-Based Agents to Solve Tasks Using Large Tool Libraries
by: Ocker, Felix, et al.
Published: (2024)
by: Ocker, Felix, et al.
Published: (2024)
Learning Type-Generalized Actions for Symbolic Planning
by: Tanneberg, Daniel, et al.
Published: (2023)
by: Tanneberg, Daniel, et al.
Published: (2023)
SemanticScanpath: Combining Gaze and Speech for Situated Human-Robot Interaction Using LLMs
by: Menendez, Elisabeth, et al.
Published: (2025)
by: Menendez, Elisabeth, et al.
Published: (2025)
Towards Driver Behavior Understanding: Weakly-Supervised Risk Perception in Driving Scenes
by: Agarwal, Nakul, et al.
Published: (2026)
by: Agarwal, Nakul, et al.
Published: (2026)
XR$^3$: An Extended Reality Platform for Social-Physical Human-Robot Interaction
by: Wang, Chao, et al.
Published: (2026)
by: Wang, Chao, et al.
Published: (2026)
Local Pairwise Distance Matching for Backpropagation-Free Reinforcement Learning
by: Tanneberg, Daniel
Published: (2025)
by: Tanneberg, Daniel
Published: (2025)
Affordance-based Robot Manipulation with Flow Matching
by: Zhang, Fan, et al.
Published: (2024)
by: Zhang, Fan, et al.
Published: (2024)
Superpositions of CARMA processes
by: Grahovac, Danijel, et al.
Published: (2026)
by: Grahovac, Danijel, et al.
Published: (2026)
Learning Robot Manipulation from Audio World Models
by: Zhang, Fan, et al.
Published: (2025)
by: Zhang, Fan, et al.
Published: (2025)
A Grounded Memory System For Smart Personal Assistants
by: Ocker, Felix, et al.
Published: (2025)
by: Ocker, Felix, et al.
Published: (2025)
Measure-Valued CARMA Processes
by: Benth, Fred Espen, et al.
Published: (2025)
by: Benth, Fred Espen, et al.
Published: (2025)
Robustness Evaluation of Machine Learning Models for Robot Arm Action Recognition in Noisy Environments
by: Motamedi, Elaheh, et al.
Published: (2024)
by: Motamedi, Elaheh, et al.
Published: (2024)
On the interference of the scattered wave and the incident wave in light scattering problems with Gaussian beams
by: Gienger, Jonas
Published: (2025)
by: Gienger, Jonas
Published: (2025)
CARMA: Collocation-Aware Resource Manager
by: Yousefzadeh-Asl-Miandoab, Ehsan, et al.
Published: (2025)
by: Yousefzadeh-Asl-Miandoab, Ehsan, et al.
Published: (2025)
Stacked Confusion Reject Plots (SCORE)
by: Hasler, Stephan, et al.
Published: (2024)
by: Hasler, Stephan, et al.
Published: (2024)
Can't make an Omelette without Breaking some Eggs: Plausible Action Anticipation using Large Video-Language Models
by: Mittal, Himangi, et al.
Published: (2024)
by: Mittal, Himangi, et al.
Published: (2024)
Learning Action-Conditional and Object-Centric Gaussian Splatting World Models for Rigid Objects
by: Kreber, Jens U., et al.
Published: (2026)
by: Kreber, Jens U., et al.
Published: (2026)
Ground States for Infrared Renormalized Translation-Invariant Non-Relativistic QED
by: Hasler, David, et al.
Published: (2022)
by: Hasler, David, et al.
Published: (2022)
Ground States for translationally invariant Pauli-Fierz Models at zero Momentum
by: Hasler, David, et al.
Published: (2020)
by: Hasler, David, et al.
Published: (2020)
Generation of Real-time Robotic Emotional Expressions Learning from Human Demonstration in Mixed Reality
by: Wang, Chao, et al.
Published: (2025)
by: Wang, Chao, et al.
Published: (2025)
Vamos: Versatile Action Models for Video Understanding
by: Wang, Shijie, et al.
Published: (2023)
by: Wang, Shijie, et al.
Published: (2023)
Characterizing the Molecular Gas in Infrared Bright Galaxies with CARMA
by: Alatalo, Katherine, et al.
Published: (2024)
by: Alatalo, Katherine, et al.
Published: (2024)
ALGO: Object-Grounded Visual Commonsense Reasoning for Open-World Egocentric Action Recognition
by: Kundu, Sanjoy, et al.
Published: (2024)
by: Kundu, Sanjoy, et al.
Published: (2024)
Generalized Mission Planning for Heterogeneous Multi-Robot Teams via LLM-constructed Hierarchical Trees
by: Gupta, Piyush, et al.
Published: (2025)
by: Gupta, Piyush, et al.
Published: (2025)
Optimal Driver Warning Generation in Dynamic Driving Environment
by: Li, Chenran, et al.
Published: (2024)
by: Li, Chenran, et al.
Published: (2024)
Option Pricing with a Compound CARMA(p,q)-Hawkes
by: Mercuri, Lorenzo, et al.
Published: (2024)
by: Mercuri, Lorenzo, et al.
Published: (2024)
Estimation of Lévy-driven CARMA models under renewal sampling
by: Bosserhoff, Frank, et al.
Published: (2026)
by: Bosserhoff, Frank, et al.
Published: (2026)
Do You Need a Hand? -- a Bimanual Robotic Dressing Assistance Scheme
by: Zhu, Jihong, et al.
Published: (2023)
by: Zhu, Jihong, et al.
Published: (2023)
Bio-Inspired Event-Based Visual Servoing for Ground Robots
by: Mordad, Maral, et al.
Published: (2026)
by: Mordad, Maral, et al.
Published: (2026)
Protective Effects of Curcumin and Menthol in Rat Models of Global Cerebral Ischemia: A Comparative Study of Combined and Separate Treatments
by: Dariush Mehboodi, et al.
Published: (2025)
by: Dariush Mehboodi, et al.
Published: (2025)
Neuro-Symbolic Imitation Learning: Discovering Symbolic Abstractions for Skill Learning
by: Keller, Leon, et al.
Published: (2025)
by: Keller, Leon, et al.
Published: (2025)
Social History, Religious History, and the Case for Theology
by: Christopher Ocker
Published: (2026)
by: Christopher Ocker
Published: (2026)
Similar Items
-
MERGE: Guided Vision-Language Models for Multi-Actor Event Reasoning and Grounding in Human-Robot Interaction
by: Deigmoeller, Joerg, et al.
Published: (2026) -
To Help or Not to Help: LLM-based Attentive Support for Human-Robot Group Interactions
by: Tanneberg, Daniel, et al.
Published: (2024) -
CoPAL: Corrective Planning of Robot Actions with Large Language Models
by: Joublin, Frank, et al.
Published: (2023) -
LaMI: Large Language Models for Multi-Modal Human-Robot Interaction
by: Wang, Chao, et al.
Published: (2024) -
Pose-Aware Weakly-Supervised Action Segmentation
by: Zhao, Seth Z., et al.
Published: (2025)