OV-HHIR: Open Vocabulary Human Interaction Recognition Using Cross-modal Integration of Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Ray, Lala Shakti Swarup, Zhou, Bo, Suh, Sungho, Lukowicz, Paul |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Initial Findings on Sensor based Open Vocabulary Activity Recognition via Text Embedding Inversion
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025)
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025)
A Synthetic Benchmarking Pipeline to Compare Camera Calibration Algorithms
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2023)
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2023)
TxP: Reciprocal Generation of Ground Pressure Dynamics and Activity Descriptions for Improving Human Activity Recognition
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025)
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025)
SImpHAR: Advancing impedance-based human activity recognition using 3D simulation and text-to-motion models
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025)
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025)
OpenMarcie: Dataset for Multimodal Action Recognition in Industrial Environments
por: Bello, Hymalai, et al.
Publicado: (2026)
por: Bello, Hymalai, et al.
Publicado: (2026)
Sensor Data Augmentation from Skeleton Pose Sequences for Improving Human Activity Recognition
por: Zolfaghari, Parham, et al.
Publicado: (2024)
por: Zolfaghari, Parham, et al.
Publicado: (2024)
Enhancing Inertial Hand based HAR through Joint Representation of Language, Pose and Synthetic IMUs
por: Rey, Vitor Fortes, et al.
Publicado: (2024)
por: Rey, Vitor Fortes, et al.
Publicado: (2024)
W2W: A Simulated Exploration of IMU Placement Across the Human Body for Designing Smarter Wearable
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025)
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025)
ALS-HAR: Harnessing Wearable Ambient Light Sensors to Enhance IMU-based Human Activity Recogntion
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2024)
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2024)
ChairPose: Pressure-based Chair Morphology Grounded Sitting Pose Estimation through Simulation-Assisted Training
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025)
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025)
PIM: Physics-Informed Multi-task Pre-training for Improving Inertial Sensor-Based Human Activity Recognition
por: Nshimyimana, Dominique, et al.
Publicado: (2025)
por: Nshimyimana, Dominique, et al.
Publicado: (2025)
MeciFace: Mechanomyography and Inertial Fusion-based Glasses for Edge Real-Time Recognition of Facial and Eating Activities
por: Bello, Hymalai, et al.
Publicado: (2023)
por: Bello, Hymalai, et al.
Publicado: (2023)
Text me the data: Generating Ground Pressure Sequence from Textual Descriptions for HAR
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2024)
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2024)
From Neck to Head: Bio-Impedance Sensing for Head Pose Estimation
por: Liu, Mengxi, et al.
Publicado: (2025)
por: Liu, Mengxi, et al.
Publicado: (2025)
SPECTRA: An Efficient Spectral-Informed Neural Network for Sensor-Based Activity Recognition
por: Gurung, Deepika, et al.
Publicado: (2026)
por: Gurung, Deepika, et al.
Publicado: (2026)
OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion
por: Wang, Hao, et al.
Publicado: (2024)
por: Wang, Hao, et al.
Publicado: (2024)
CaptAinGlove: Capacitive and Inertial Fusion-Based Glove for Real-Time on Edge Hand Gesture Recognition for Drone Control
por: Bello, Hymalai, et al.
Publicado: (2023)
por: Bello, Hymalai, et al.
Publicado: (2023)
Origami Single-end Capacitive Sensing for Continuous Shape Estimation of Morphing Structures
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2023)
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2023)
MuJo: Multimodal Joint Feature Space Learning for Human Activity Recognition
por: Fritsch, Stefan Gerd, et al.
Publicado: (2024)
por: Fritsch, Stefan Gerd, et al.
Publicado: (2024)
Body-Area Capacitive or Electric Field Sensing for Human Activity Recognition and Human-Computer Interaction: A Comprehensive Survey
por: Bian, Sizhen, et al.
Publicado: (2024)
por: Bian, Sizhen, et al.
Publicado: (2024)
iMove: Exploring Bio-impedance Sensing for Fitness Activity Recognition
por: Liu, Mengxi, et al.
Publicado: (2024)
por: Liu, Mengxi, et al.
Publicado: (2024)
OV-COAST: Cost Aggregation with Optimal Transport for Open-Vocabulary Semantic Segmentation
por: Gandhamal, Aditya, et al.
Publicado: (2025)
por: Gandhamal, Aditya, et al.
Publicado: (2025)
SegEarth-OV: Towards Training-Free Open-Vocabulary Segmentation for Remote Sensing Images
por: Li, Kaiyu, et al.
Publicado: (2024)
por: Li, Kaiyu, et al.
Publicado: (2024)
OV9D: Open-Vocabulary Category-Level 9D Object Pose and Size Estimation
por: Cai, Junhao, et al.
Publicado: (2024)
por: Cai, Junhao, et al.
Publicado: (2024)
OV-DEIM: Real-time DETR-Style Open-Vocabulary Object Detection with GridSynthetic Augmentation
por: Wang, Leilei, et al.
Publicado: (2026)
por: Wang, Leilei, et al.
Publicado: (2026)
OV-MAP : Open-Vocabulary Zero-Shot 3D Instance Segmentation Map for Robots
por: Kim, Juno, et al.
Publicado: (2025)
por: Kim, Juno, et al.
Publicado: (2025)
OV-DQUO: Open-Vocabulary DETR with Denoising Text Query Training and Open-World Unknown Objects Supervision
por: Wang, Junjie, et al.
Publicado: (2024)
por: Wang, Junjie, et al.
Publicado: (2024)
Reconstructing In-the-Wild Open-Vocabulary Human-Object Interactions
por: Wen, Boran, et al.
Publicado: (2025)
por: Wen, Boran, et al.
Publicado: (2025)
SegEarth-OV3: Exploring SAM 3 for Open-Vocabulary Semantic Segmentation in Remote Sensing Images
por: Li, Kaiyu, et al.
Publicado: (2025)
por: Li, Kaiyu, et al.
Publicado: (2025)
OV-SCAN: Semantically Consistent Alignment for Novel Object Discovery in Open-Vocabulary 3D Object Detection
por: Chow, Adrian, et al.
Publicado: (2025)
por: Chow, Adrian, et al.
Publicado: (2025)
Bilateral Collaboration with Large Vision-Language Models for Open Vocabulary Human-Object Interaction Detection
por: Hu, Yupeng, et al.
Publicado: (2025)
por: Hu, Yupeng, et al.
Publicado: (2025)
BeSound: Bluetooth-Based Position Estimation Enhancing with Cross-Modality Distillation
por: Bello, Hymalai, et al.
Publicado: (2024)
por: Bello, Hymalai, et al.
Publicado: (2024)
Streamlined Open-Vocabulary Human-Object Interaction Detection
por: Sun, Chang, et al.
Publicado: (2026)
por: Sun, Chang, et al.
Publicado: (2026)
OV-Uni3DETR: Towards Unified Open-Vocabulary 3D Object Detection via Cycle-Modality Propagation
por: Wang, Zhenyu, et al.
Publicado: (2024)
por: Wang, Zhenyu, et al.
Publicado: (2024)
OV-Stitcher: A Global Context-Aware Framework for Training-Free Open-Vocabulary Semantic Segmentation
por: Moon, Seungjae, et al.
Publicado: (2026)
por: Moon, Seungjae, et al.
Publicado: (2026)
ActivityNarrated: An Open-Ended Narrative Paradigm for Wearable Human Activity Understanding
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2026)
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2026)
Rethinking CLIP-based Video Learners in Cross-Domain Open-Vocabulary Action Recognition
por: Lin, Kun-Yu, et al.
Publicado: (2024)
por: Lin, Kun-Yu, et al.
Publicado: (2024)
UniM-OV3D: Uni-Modality Open-Vocabulary 3D Scene Understanding with Fine-Grained Feature Representation
por: He, Qingdong, et al.
Publicado: (2024)
por: He, Qingdong, et al.
Publicado: (2024)
XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation
por: Wang, Ziyi, et al.
Publicado: (2024)
por: Wang, Ziyi, et al.
Publicado: (2024)
Explainable Parkinsons Disease Gait Recognition Using Multimodal RGB-D Fusion and Large Language Models
por: Alnaasan, Manar, et al.
Publicado: (2025)
por: Alnaasan, Manar, et al.
Publicado: (2025)
Ejemplares similares
-
Initial Findings on Sensor based Open Vocabulary Activity Recognition via Text Embedding Inversion
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025) -
A Synthetic Benchmarking Pipeline to Compare Camera Calibration Algorithms
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2023) -
TxP: Reciprocal Generation of Ground Pressure Dynamics and Activity Descriptions for Improving Human Activity Recognition
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025) -
SImpHAR: Advancing impedance-based human activity recognition using 3D simulation and text-to-motion models
por: Ray, Lala Shakti Swarup, et al.
Publicado: (2025) -
OpenMarcie: Dataset for Multimodal Action Recognition in Industrial Environments
por: Bello, Hymalai, et al.
Publicado: (2026)