Saved in:
| Main Authors: | Tang, Bin, Pan, Keqi, Zheng, Miao, Zhou, Ning, Sui, Jialu, Zhu, Dandan, Deng, Cheng-Long, Kuai, Shu-Guang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2503.12912 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SelfPose3d: Self-Supervised Multi-Person Multi-View 3d Pose Estimation
by: Srivastav, Vinkle, et al.
Published: (2024)
by: Srivastav, Vinkle, et al.
Published: (2024)
Label-Synchronous Neural Transducer for Adaptable Online E2E Speech Recognition
by: Deng, Keqi, et al.
Published: (2023)
by: Deng, Keqi, et al.
Published: (2023)
Human-Inspired Soft Anthropomorphic Hand System for Neuromorphic Object and Pose Recognition Using Multimodal Signals
by: Wang, Fengyi, et al.
Published: (2025)
by: Wang, Fengyi, et al.
Published: (2025)
SAP-CoPE: Social-Aware Planning using Cooperative Pose Estimation with Infrastructure Sensor Nodes
by: Ning, Minghao, et al.
Published: (2025)
by: Ning, Minghao, et al.
Published: (2025)
A Dataset and Benchmarks for Deep Learning-Based Optical Microrobot Pose and Depth Perception
by: Wei, Lan, et al.
Published: (2025)
by: Wei, Lan, et al.
Published: (2025)
Label-Synchronous Neural Transducer for E2E Simultaneous Speech Translation
by: Deng, Keqi, et al.
Published: (2024)
by: Deng, Keqi, et al.
Published: (2024)
Multi-head Temporal Latent Attention
by: Deng, Keqi, et al.
Published: (2025)
by: Deng, Keqi, et al.
Published: (2025)
Exploring the Necessity of Visual Modality in Multimodal Machine Translation using Authentic Datasets
by: Long, Zi, et al.
Published: (2024)
by: Long, Zi, et al.
Published: (2024)
Adaptive Multimodal Person Recognition: A Robust Framework for Handling Missing Modalities
by: Farhadipour, Aref, et al.
Published: (2025)
by: Farhadipour, Aref, et al.
Published: (2025)
Dark Side of Modalities: Reinforced Multimodal Distillation for Multimodal Knowledge Graph Reasoning
by: Zhao, Yu, et al.
Published: (2025)
by: Zhao, Yu, et al.
Published: (2025)
Learning to Learn from Multimodal Experience
by: Sui, Xingyu, et al.
Published: (2026)
by: Sui, Xingyu, et al.
Published: (2026)
OCCDiff: Occupancy Diffusion Model for High-Fidelity 3D Building Reconstruction from Noisy Point Clouds
by: Sui, Jialu, et al.
Published: (2025)
by: Sui, Jialu, et al.
Published: (2025)
Differential Mental Disorder Detection with Psychology-Inspired Multimodal Stimuli
by: Zhou, Zhiyuan, et al.
Published: (2026)
by: Zhou, Zhiyuan, et al.
Published: (2026)
Multi-Modal Experience Inspired AI Creation
by: Cao, Qian, et al.
Published: (2022)
by: Cao, Qian, et al.
Published: (2022)
A Theory-Inspired Framework for Few-Shot Cross-Modal Sketch Person Re-Identification
by: Gong, Yunpeng, et al.
Published: (2025)
by: Gong, Yunpeng, et al.
Published: (2025)
Revealing Personality Traits: A New Benchmark Dataset for Explainable Personality Recognition on Dialogues
by: Sun, Lei, et al.
Published: (2024)
by: Sun, Lei, et al.
Published: (2024)
Pedestrian Attribute Recognition via Hierarchical Cross-Modality HyperGraph Learning
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
Toward Intelligent and Personalized Skin Healing: Responsive Natural Hydrogels Bridging Sensing and Therapy
by: Xinheng Wang, et al.
Published: (2026)
by: Xinheng Wang, et al.
Published: (2026)
Towards Comprehensive Multimodal Perception: Introducing the Touch-Language-Vision Dataset
by: Cheng, Ning, et al.
Published: (2024)
by: Cheng, Ning, et al.
Published: (2024)
Range and Bird's Eye View Fused Cross-Modal Visual Place Recognition
by: Peng, Jianyi, et al.
Published: (2025)
by: Peng, Jianyi, et al.
Published: (2025)
Decision Making in Urban Traffic: A Game Theoretic Approach for Autonomous Vehicles Adhering to Traffic Rules
by: Shu, Keqi, et al.
Published: (2025)
by: Shu, Keqi, et al.
Published: (2025)
Real-World Deployment of Cloud-based Autonomous Mobility Systems for Outdoor and Indoor Environments
by: Yang, Yufeng, et al.
Published: (2025)
by: Yang, Yufeng, et al.
Published: (2025)
MDPE: A Multimodal Deception Dataset with Personality and Emotional Characteristics
by: Cai, Cong, et al.
Published: (2024)
by: Cai, Cong, et al.
Published: (2024)
iNews: A Multimodal Dataset for Modeling Personalized Affective Responses to News
by: Hu, Tiancheng, et al.
Published: (2025)
by: Hu, Tiancheng, et al.
Published: (2025)
Metronome: Efficient Scheduling for Periodic Traffic Jobs with Network and Priority Awareness
by: Jiang, Hao, et al.
Published: (2025)
by: Jiang, Hao, et al.
Published: (2025)
Wav2Prompt: End-to-End Speech Prompt Generation and Tuning For LLM in Zero and Few-shot Learning
by: Deng, Keqi, et al.
Published: (2024)
by: Deng, Keqi, et al.
Published: (2024)
AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition
by: Lian, Zheng, et al.
Published: (2024)
by: Lian, Zheng, et al.
Published: (2024)
Learning from Synchronization: Self-Supervised Uncalibrated Multi-View Person Association in Challenging Scenes
by: Chen, Keqi, et al.
Published: (2025)
by: Chen, Keqi, et al.
Published: (2025)
Transducer-Llama: Integrating LLMs into Streamable Transducer-based Speech Recognition
by: Deng, Keqi, et al.
Published: (2024)
by: Deng, Keqi, et al.
Published: (2024)
Dataset for Person Identification from Pose Estimates in Sign Language
by: Battisti, Alessia, et al.
Published: (2024)
by: Battisti, Alessia, et al.
Published: (2024)
SPACT18: Spiking Human Action Recognition Benchmark Dataset with Complementary RGB and Thermal Modalities
by: Ashraf, Yasser, et al.
Published: (2025)
by: Ashraf, Yasser, et al.
Published: (2025)
Elastic constant ratio for fatigue evaluation on rubber isolators
by: Robert Keqi Luo
Published: (2024)
by: Robert Keqi Luo
Published: (2024)
Temporal Interest-Driven Multimodal Personalized Content Generation
by: Miao, Tian
Published: (2025)
by: Miao, Tian
Published: (2025)
Incomplete Multimodal Industrial Anomaly Detection via Cross-Modal Distillation
by: Sui, Wenbo, et al.
Published: (2024)
by: Sui, Wenbo, et al.
Published: (2024)
SignAligner: Harmonizing Complementary Pose Modalities for Coherent Sign Language Generation
by: Wang, Xu, et al.
Published: (2025)
by: Wang, Xu, et al.
Published: (2025)
MultiDiffSense: Diffusion-Based Multi-Modal Visuo-Tactile Image Generation Conditioned on Object Shape and Contact Pose
by: Bhouri, Sirine, et al.
Published: (2026)
by: Bhouri, Sirine, et al.
Published: (2026)
DF4LCZ: A SAM-Empowered Data Fusion Framework for Scene-Level Local Climate Zone Classification
by: Wu, Qianqian, et al.
Published: (2024)
by: Wu, Qianqian, et al.
Published: (2024)
Adaptive Semantic-Enhanced Denoising Diffusion Probabilistic Model for Remote Sensing Image Super-Resolution
by: Sui, Jialu, et al.
Published: (2024)
by: Sui, Jialu, et al.
Published: (2024)
PPMamba: A Pyramid Pooling Local Auxiliary SSM-Based Model for Remote Sensing Image Semantic Segmentation
by: Hu, Yin, et al.
Published: (2024)
by: Hu, Yin, et al.
Published: (2024)
T2Vs Meet VLMs: A Scalable Multimodal Dataset for Visual Harmfulness Recognition
by: Yeh, Chen, et al.
Published: (2024)
by: Yeh, Chen, et al.
Published: (2024)
Similar Items
-
SelfPose3d: Self-Supervised Multi-Person Multi-View 3d Pose Estimation
by: Srivastav, Vinkle, et al.
Published: (2024) -
Label-Synchronous Neural Transducer for Adaptable Online E2E Speech Recognition
by: Deng, Keqi, et al.
Published: (2023) -
Human-Inspired Soft Anthropomorphic Hand System for Neuromorphic Object and Pose Recognition Using Multimodal Signals
by: Wang, Fengyi, et al.
Published: (2025) -
SAP-CoPE: Social-Aware Planning using Cooperative Pose Estimation with Infrastructure Sensor Nodes
by: Ning, Minghao, et al.
Published: (2025) -
A Dataset and Benchmarks for Deep Learning-Based Optical Microrobot Pose and Depth Perception
by: Wei, Lan, et al.
Published: (2025)