Saved in:
| Main Authors: | Aparcedo, Alejandro, Lopez, Christian, Kotta, Abhinav, Li, Mengjie |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.00017 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dynamic Bandwidth Allocation for Hybrid Event-RGB Transmission
by: Yang, Pujing, et al.
Published: (2025)
by: Yang, Pujing, et al.
Published: (2025)
A PolSAR Scattering Power Factorization Framework and Novel Roll-Invariant Parameters Based Unsupervised Classification Scheme Using a Geodesic Distance
by: Ratha, Debanshu, et al.
Published: (2019)
by: Ratha, Debanshu, et al.
Published: (2019)
Power-LLaVA: Large Language and Vision Assistant for Power Transmission Line Inspection
by: Wang, Jiahao, et al.
Published: (2024)
by: Wang, Jiahao, et al.
Published: (2024)
Multimodal Indoor Localization Using Crowdsourced Radio Maps
by: Yi, Zhaoguang, et al.
Published: (2023)
by: Yi, Zhaoguang, et al.
Published: (2023)
Deep Learning Based Speckle Filtering for Polarimetric SAR Images. Application to Sentinel-1
by: Mestre-Quereda, Alejandro, et al.
Published: (2024)
by: Mestre-Quereda, Alejandro, et al.
Published: (2024)
Task-Oriented Feature Compression for Multimodal Understanding via Device-Edge Co-Inference
by: Yuan, Cheng, et al.
Published: (2025)
by: Yuan, Cheng, et al.
Published: (2025)
OpenMarcie: Dataset for Multimodal Action Recognition in Industrial Environments
by: Bello, Hymalai, et al.
Published: (2026)
by: Bello, Hymalai, et al.
Published: (2026)
A Renderer-Enabled Framework for Computing Parameter Estimation Lower Bounds in Plenoptic Imaging Systems
by: Sambasivan, Abhinav V., et al.
Published: (2026)
by: Sambasivan, Abhinav V., et al.
Published: (2026)
Generative AI Empowered LiDAR Point Cloud Generation with Multimodal Transformer
by: Farzanullah, Mohammad, et al.
Published: (2024)
by: Farzanullah, Mohammad, et al.
Published: (2024)
X-Fi: A Modality-Invariant Foundation Model for Multimodal Human Sensing
by: Chen, Xinyan, et al.
Published: (2024)
by: Chen, Xinyan, et al.
Published: (2024)
Col-OLHTR: A Novel Framework for Multimodal Online Handwritten Text Recognition
by: Liu, Chenyu, et al.
Published: (2025)
by: Liu, Chenyu, et al.
Published: (2025)
Foundation-Model-Boosted Multimodal Learning for fMRI-based Neuropathic Pain Drug Response Prediction
by: Fan, Wenrui, et al.
Published: (2025)
by: Fan, Wenrui, et al.
Published: (2025)
Introducing Multimodal Paradigm for Learning Sleep Staging PSG via General-Purpose Model
by: Zhou, Jianheng, et al.
Published: (2025)
by: Zhou, Jianheng, et al.
Published: (2025)
VidSole: A Multimodal Dataset for Joint Kinetics Quantification and Disease Detection with Deep Learning
by: Kambhamettu, Archit, et al.
Published: (2025)
by: Kambhamettu, Archit, et al.
Published: (2025)
Evaluation of Video-Based rPPG in Challenging Environments: Artifact Mitigation and Network Resilience
by: Nguyen, Nhi, et al.
Published: (2024)
by: Nguyen, Nhi, et al.
Published: (2024)
Fast Deep Predictive Coding Networks for Videos Feature Extraction without Labels
by: Xue, Wenqian, et al.
Published: (2024)
by: Xue, Wenqian, et al.
Published: (2024)
High-Dynamic Radar Sequence Prediction for Weather Nowcasting Using Spatiotemporal Coherent Gaussian Representation
by: Wang, Ziye, et al.
Published: (2025)
by: Wang, Ziye, et al.
Published: (2025)
Differentiable High-Performance Ray Tracing-Based Simulation of Radio Propagation with Point Clouds
by: Vaara, Niklas, et al.
Published: (2025)
by: Vaara, Niklas, et al.
Published: (2025)
Perspective-aware fusion of incomplete depth maps and surface normals for accurate 3D reconstruction
by: Hlinka, Ondrej, et al.
Published: (2026)
by: Hlinka, Ondrej, et al.
Published: (2026)
EMPD: An Event-based Multimodal Physiological Dataset for Remote Pulse Wave Detection
by: Feng, Qian, et al.
Published: (2026)
by: Feng, Qian, et al.
Published: (2026)
Quality-Aware Framework for Video-Derived Respiratory Signals
by: Nguyen, Nhi, et al.
Published: (2025)
by: Nguyen, Nhi, et al.
Published: (2025)
Radar-Based Recognition of Static Hand Gestures in American Sign Language
by: Schuessler, Christian, et al.
Published: (2024)
by: Schuessler, Christian, et al.
Published: (2024)
Deep, Deep Learning with BART
by: Blumenthal, Moritz, et al.
Published: (2022)
by: Blumenthal, Moritz, et al.
Published: (2022)
Tactile-based Multimodal Fusion in Embodied Intelligence: A Survey of Vision, Language, and Contact-Driven Paradigms
by: Cao, Zhixiang, et al.
Published: (2026)
by: Cao, Zhixiang, et al.
Published: (2026)
GAP9Shield: A 150GOPS AI-capable Ultra-low Power Module for Vision and Ranging Applications on Nano-drones
by: Müller, Hanna, et al.
Published: (2024)
by: Müller, Hanna, et al.
Published: (2024)
Hierarchical Attention Networks for Lossless Point Cloud Attribute Compression
by: Chen, Yueru, et al.
Published: (2025)
by: Chen, Yueru, et al.
Published: (2025)
Exploring Textual Semantics Diversity for Image Transmission in Semantic Communication Systems using Visual Language Model
by: Huang, Peishan, et al.
Published: (2025)
by: Huang, Peishan, et al.
Published: (2025)
LoFi: Vision-Aided Label Generator for Wi-Fi Localization and Tracking
by: Zhao, Zijian, et al.
Published: (2024)
by: Zhao, Zijian, et al.
Published: (2024)
Efficient Point Clouds Upsampling via Flow Matching
by: Liu, Zhi-Song, et al.
Published: (2025)
by: Liu, Zhi-Song, et al.
Published: (2025)
3D Photon Counting CT Image Super-Resolution Using Conditional Diffusion Model
by: Niu, Chuang, et al.
Published: (2024)
by: Niu, Chuang, et al.
Published: (2024)
Scene Understanding Enabled Semantic Communication with Open Channel Coding
by: Xiang, Zhe, et al.
Published: (2025)
by: Xiang, Zhe, et al.
Published: (2025)
MARS: Radio Map Super-resolution and Reconstruction Method under Sparse Channel Measurements
by: Deng, Chuyun, et al.
Published: (2025)
by: Deng, Chuyun, et al.
Published: (2025)
Multimodal Latent Fusion of ECG Leads for Early Assessment of Pulmonary Hypertension
by: Suvon, Mohammod N. I., et al.
Published: (2025)
by: Suvon, Mohammod N. I., et al.
Published: (2025)
Multi-Modal Self-Supervised Semantic Communication
by: Zhao, Hang, et al.
Published: (2025)
by: Zhao, Hang, et al.
Published: (2025)
Cross-Domain Multi-Person Human Activity Recognition via Near-Field Wi-Fi Sensing
by: Li, Xin, et al.
Published: (2025)
by: Li, Xin, et al.
Published: (2025)
Summit Vitals: Multi-Camera and Multi-Signal Biosensing at High Altitudes
by: Liu, Ke, et al.
Published: (2024)
by: Liu, Ke, et al.
Published: (2024)
Large Language Model-Driven Distributed Integrated Multimodal Sensing and Semantic Communications
by: Peng, Yubo, et al.
Published: (2025)
by: Peng, Yubo, et al.
Published: (2025)
Generative Video Semantic Communication via Multimodal Semantic Fusion with Large Model
by: Yin, Hang, et al.
Published: (2025)
by: Yin, Hang, et al.
Published: (2025)
CognitionCapturer: Decoding Visual Stimuli From Human EEG Signal With Multimodal Information
by: Zhang, Kaifan, et al.
Published: (2024)
by: Zhang, Kaifan, et al.
Published: (2024)
SegmentAnyMuscle: A universal muscle segmentation model across different locations in MRI
by: Colglazier, Roy, et al.
Published: (2025)
by: Colglazier, Roy, et al.
Published: (2025)
Similar Items
-
Dynamic Bandwidth Allocation for Hybrid Event-RGB Transmission
by: Yang, Pujing, et al.
Published: (2025) -
A PolSAR Scattering Power Factorization Framework and Novel Roll-Invariant Parameters Based Unsupervised Classification Scheme Using a Geodesic Distance
by: Ratha, Debanshu, et al.
Published: (2019) -
Power-LLaVA: Large Language and Vision Assistant for Power Transmission Line Inspection
by: Wang, Jiahao, et al.
Published: (2024) -
Multimodal Indoor Localization Using Crowdsourced Radio Maps
by: Yi, Zhaoguang, et al.
Published: (2023) -
Deep Learning Based Speckle Filtering for Polarimetric SAR Images. Application to Sentinel-1
by: Mestre-Quereda, Alejandro, et al.
Published: (2024)