Gespeichert in:
| Hauptverfasser: | Huang, Shuokang, McCann, Julie A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2503.09537 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WiMANS: A Benchmark Dataset for WiFi-based Multi-user Activity Sensing
von: Huang, Shuokang, et al.
Veröffentlicht: (2024)
von: Huang, Shuokang, et al.
Veröffentlicht: (2024)
KeyNode-Driven Geometry Coding for Real-World Scanned Human Dynamic Mesh Compression
von: Hoang, Huong, et al.
Veröffentlicht: (2025)
von: Hoang, Huong, et al.
Veröffentlicht: (2025)
Full-reference Point Cloud Quality Assessment Using Spectral Graph Wavelets
von: Watanabe, Ryosuke, et al.
Veröffentlicht: (2024)
von: Watanabe, Ryosuke, et al.
Veröffentlicht: (2024)
Frequency-Spatial Interaction Driven Network for Low-Light Image Enhancement
von: Tao, Yunhong, et al.
Veröffentlicht: (2025)
von: Tao, Yunhong, et al.
Veröffentlicht: (2025)
Latency-Aware Generative Semantic Communications with Pre-Trained Diffusion Models
von: Qiao, Li, et al.
Veröffentlicht: (2024)
von: Qiao, Li, et al.
Veröffentlicht: (2024)
Communicate Less, Synthesize the Rest: Latency-aware Intent-based Generative Semantic Multicasting with Diffusion Models
von: Liu, Xinkai, et al.
Veröffentlicht: (2024)
von: Liu, Xinkai, et al.
Veröffentlicht: (2024)
Exploring Event-based Human Pose Estimation with 3D Event Representations
von: Yin, Xiaoting, et al.
Veröffentlicht: (2023)
von: Yin, Xiaoting, et al.
Veröffentlicht: (2023)
Enhanced Radar Perception via Multi-Task Learning: Towards Refined Data for Sensor Fusion Applications
von: Sun, Huawei, et al.
Veröffentlicht: (2024)
von: Sun, Huawei, et al.
Veröffentlicht: (2024)
WDMIR: Wavelet-Driven Multimodal Intent Recognition
von: Gong, Weiyin, et al.
Veröffentlicht: (2025)
von: Gong, Weiyin, et al.
Veröffentlicht: (2025)
Token Communications: A Large Model-Driven Framework for Cross-modal Context-aware Semantic Communications
von: Qiao, Li, et al.
Veröffentlicht: (2025)
von: Qiao, Li, et al.
Veröffentlicht: (2025)
Radio Frequency Signal based Human Silhouette Segmentation: A Sequential Diffusion Approach
von: Wen, Penghui, et al.
Veröffentlicht: (2024)
von: Wen, Penghui, et al.
Veröffentlicht: (2024)
Improving Noise Robust Audio-Visual Speech Recognition via Router-Gated Cross-Modal Feature Fusion
von: Lim, DongHoon, et al.
Veröffentlicht: (2025)
von: Lim, DongHoon, et al.
Veröffentlicht: (2025)
Tackle CSM in JPEG Steganalysis with Data Adaptation
von: Abecidan, Rony, et al.
Veröffentlicht: (2026)
von: Abecidan, Rony, et al.
Veröffentlicht: (2026)
HDiffTG: A Lightweight Hybrid Diffusion-Transformer-GCN Architecture for 3D Human Pose Estimation
von: Fu, Yajie, et al.
Veröffentlicht: (2025)
von: Fu, Yajie, et al.
Veröffentlicht: (2025)
RT-Pose: A 4D Radar Tensor-based 3D Human Pose Estimation and Localization Benchmark
von: Ho, Yuan-Hao, et al.
Veröffentlicht: (2024)
von: Ho, Yuan-Hao, et al.
Veröffentlicht: (2024)
Exploiting Frequency Correlation for Hyperspectral Image Reconstruction
von: Yan, Muge, et al.
Veröffentlicht: (2024)
von: Yan, Muge, et al.
Veröffentlicht: (2024)
AudioGen-Omni: A Unified Multimodal Diffusion Transformer for Video-Synchronized Audio, Speech, and Song Generation
von: Wang, Le, et al.
Veröffentlicht: (2025)
von: Wang, Le, et al.
Veröffentlicht: (2025)
PF-D2M: A Pose-free Diffusion Model for Universal Dance-to-Music Generation
von: Im, Jaekwon, et al.
Veröffentlicht: (2026)
von: Im, Jaekwon, et al.
Veröffentlicht: (2026)
RAPTR: Radar-based 3D Pose Estimation using Transformer
von: Kato, Sorachi, et al.
Veröffentlicht: (2025)
von: Kato, Sorachi, et al.
Veröffentlicht: (2025)
Frequency-Assisted Adaptive Sharpening Scheme Considering Bitrate and Quality Tradeoff
von: Pang, Yingxue, et al.
Veröffentlicht: (2025)
von: Pang, Yingxue, et al.
Veröffentlicht: (2025)
SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation
von: Xu, Xiangyu, et al.
Veröffentlicht: (2024)
von: Xu, Xiangyu, et al.
Veröffentlicht: (2024)
RadioDUN: A Physics-Inspired Deep Unfolding Network for Radio Map Estimation
von: Chen, Taiqin, et al.
Veröffentlicht: (2025)
von: Chen, Taiqin, et al.
Veröffentlicht: (2025)
Federated Multi-Agent DRL for Radio Resource Management in Industrial 6G in-X subnetworks
von: Madsen, Bjarke, et al.
Veröffentlicht: (2024)
von: Madsen, Bjarke, et al.
Veröffentlicht: (2024)
Automated Retinal Image Analysis and Medical Report Generation through Deep Learning
von: Huang, Jia-Hong
Veröffentlicht: (2024)
von: Huang, Jia-Hong
Veröffentlicht: (2024)
AdaMesh: Personalized Facial Expressions and Head Poses for Adaptive Speech-Driven 3D Facial Animation
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
Visible Light Positioning With Lamé Curve LEDs: A Generic Approach for Camera Pose Estimation
von: Pan, Wenxuan, et al.
Veröffentlicht: (2026)
von: Pan, Wenxuan, et al.
Veröffentlicht: (2026)
CounterFlow: A Two-Phase Inference-Time Sampling for Counterfactual Video Foley Generation
von: Lee, Gyubin, et al.
Veröffentlicht: (2026)
von: Lee, Gyubin, et al.
Veröffentlicht: (2026)
Learning Perceptual Representations for Gaming NR-VQA with Multi-Task FR Signals
von: Chen, Yu-Chih, et al.
Veröffentlicht: (2026)
von: Chen, Yu-Chih, et al.
Veröffentlicht: (2026)
Contrastive Multi-Modal Hypergraph Reasoning for 3D Crowd Mesh Recovery
von: Sun, Minghao, et al.
Veröffentlicht: (2026)
von: Sun, Minghao, et al.
Veröffentlicht: (2026)
R$^3$D: Regional-guided Residual Radar Diffusion
von: Li, Hao, et al.
Veröffentlicht: (2026)
von: Li, Hao, et al.
Veröffentlicht: (2026)
Identity-Preserving Text-to-Video Generation by Frequency Decomposition
von: Yuan, Shenghai, et al.
Veröffentlicht: (2024)
von: Yuan, Shenghai, et al.
Veröffentlicht: (2024)
Follow-Your-MultiPose: Tuning-Free Multi-Character Text-to-Video Generation via Pose Guidance
von: Zhang, Beiyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Beiyuan, et al.
Veröffentlicht: (2024)
Discriminative-Generative Synergy for Occlusion Robust 3D Human Mesh Recovery
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
ConcealGS: Concealing Invisible Copyright Information in 3D Gaussian Splatting
von: Yang, Yifeng, et al.
Veröffentlicht: (2025)
von: Yang, Yifeng, et al.
Veröffentlicht: (2025)
From Neck to Head: Bio-Impedance Sensing for Head Pose Estimation
von: Liu, Mengxi, et al.
Veröffentlicht: (2025)
von: Liu, Mengxi, et al.
Veröffentlicht: (2025)
Image Quality Assessment: From Human to Machine Preference
von: Li, Chunyi, et al.
Veröffentlicht: (2025)
von: Li, Chunyi, et al.
Veröffentlicht: (2025)
SUPER: Seated Upper Body Pose Estimation using mmWave Radars
von: Zhang, Bo, et al.
Veröffentlicht: (2024)
von: Zhang, Bo, et al.
Veröffentlicht: (2024)
Acoustic Neural 3D Reconstruction Under Pose Drift
von: Lin, Tianxiang, et al.
Veröffentlicht: (2025)
von: Lin, Tianxiang, et al.
Veröffentlicht: (2025)
Sphere-GAN: a GAN-based Approach for Saliency Estimation in 360° Videos
von: Wahba, Mahmoud Z. A., et al.
Veröffentlicht: (2025)
von: Wahba, Mahmoud Z. A., et al.
Veröffentlicht: (2025)
SoundLoc3D: Invisible 3D Sound Source Localization and Classification Using a Multimodal RGB-D Acoustic Camera
von: He, Yuhang, et al.
Veröffentlicht: (2024)
von: He, Yuhang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
WiMANS: A Benchmark Dataset for WiFi-based Multi-user Activity Sensing
von: Huang, Shuokang, et al.
Veröffentlicht: (2024) -
KeyNode-Driven Geometry Coding for Real-World Scanned Human Dynamic Mesh Compression
von: Hoang, Huong, et al.
Veröffentlicht: (2025) -
Full-reference Point Cloud Quality Assessment Using Spectral Graph Wavelets
von: Watanabe, Ryosuke, et al.
Veröffentlicht: (2024) -
Frequency-Spatial Interaction Driven Network for Low-Light Image Enhancement
von: Tao, Yunhong, et al.
Veröffentlicht: (2025) -
Latency-Aware Generative Semantic Communications with Pre-Trained Diffusion Models
von: Qiao, Li, et al.
Veröffentlicht: (2024)