Self-supervised 3D Patient Modeling with Multi-modal Attentive Fusion
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Meng, Planche, Benjamin, Gong, Xuan, Yang, Fan, Chen, Terrence, Wu, Ziyan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Few-Shot 3D Volumetric Segmentation with Multi-Surrogate Fusion
by: Zheng, Meng, et al.
Published: (2024)
by: Zheng, Meng, et al.
Published: (2024)
Automated Patient Positioning with Learned 3D Hand Gestures
by: Gao, Zhongpai, et al.
Published: (2024)
by: Gao, Zhongpai, et al.
Published: (2024)
Automating Catheterization Labs with Real-Time Perception
by: Yang, Fan, et al.
Published: (2024)
by: Yang, Fan, et al.
Published: (2024)
3D Vision-Language Gaussian Splatting
by: Peng, Qucheng, et al.
Published: (2024)
by: Peng, Qucheng, et al.
Published: (2024)
Anatomy-Aware Conditional Image-Text Retrieval
by: Zheng, Meng, et al.
Published: (2025)
by: Zheng, Meng, et al.
Published: (2025)
PolypSegTrack: Unified Foundation Model for Colonoscopy Video Analysis
by: Choudhuri, Anwesa, et al.
Published: (2025)
by: Choudhuri, Anwesa, et al.
Published: (2025)
Render-FM: A Foundation Model for Real-time Photorealistic Volumetric Rendering
by: Gao, Zhongpai, et al.
Published: (2025)
by: Gao, Zhongpai, et al.
Published: (2025)
Self-learning Canonical Space for Multi-view 3D Human Pose Estimation
by: Li, Xiaoben, et al.
Published: (2024)
by: Li, Xiaoben, et al.
Published: (2024)
7DGS: Unified Spatial-Temporal-Angular Gaussian Splatting
by: Gao, Zhongpai, et al.
Published: (2025)
by: Gao, Zhongpai, et al.
Published: (2025)
6DGS: Enhanced Direction-Aware Gaussian Splatting for Volumetric Rendering
by: Gao, Zhongpai, et al.
Published: (2024)
by: Gao, Zhongpai, et al.
Published: (2024)
DDGS-CT: Direction-Disentangled Gaussian Splatting for Realistic Volume Rendering
by: Gao, Zhongpai, et al.
Published: (2024)
by: Gao, Zhongpai, et al.
Published: (2024)
Exploring Cycle Consistency Learning in Interactive Volume Segmentation
by: Liu, Qin, et al.
Published: (2023)
by: Liu, Qin, et al.
Published: (2023)
PBADet: A One-Stage Anchor-Free Approach for Part-Body Association
by: Gao, Zhongpai, et al.
Published: (2024)
by: Gao, Zhongpai, et al.
Published: (2024)
Seq2Time: Sequential Knowledge Transfer for Video LLM Temporal Grounding
by: Deng, Andong, et al.
Published: (2024)
by: Deng, Andong, et al.
Published: (2024)
CHROME: Clothed Human Reconstruction with Occlusion-Resilience and Multiview-Consistency from a Single Image
by: Dutta, Arindam, et al.
Published: (2025)
by: Dutta, Arindam, et al.
Published: (2025)
S$^3$POT: Contrast-Driven Face Occlusion Segmentation via Self-Supervised Prompt Learning
by: Wang, Lingsong, et al.
Published: (2026)
by: Wang, Lingsong, et al.
Published: (2026)
Human Mesh Recovery from Arbitrary Multi-view Images
by: Li, Xiaoben, et al.
Published: (2024)
by: Li, Xiaoben, et al.
Published: (2024)
MedGRPO: Multi-Task Reinforcement Learning for Heterogeneous Medical Video Understanding
by: Su, Yuhao, et al.
Published: (2025)
by: Su, Yuhao, et al.
Published: (2025)
DaRePlane: Direction-aware Representations for Dynamic Scene Reconstruction
by: Lou, Ange, et al.
Published: (2024)
by: Lou, Ange, et al.
Published: (2024)
Divide and Fuse: Body Part Mesh Recovery from Partially Visible Human Images
by: Luan, Tianyu, et al.
Published: (2024)
by: Luan, Tianyu, et al.
Published: (2024)
Multi-View Attentive Contextualization for Multi-View 3D Object Detection
by: Liu, Xianpeng, et al.
Published: (2024)
by: Liu, Xianpeng, et al.
Published: (2024)
Order-aware Interactive Segmentation
by: Wang, Bin, et al.
Published: (2024)
by: Wang, Bin, et al.
Published: (2024)
Neural Finite-State Machines for Surgical Phase Recognition
by: Ding, Hao, et al.
Published: (2024)
by: Ding, Hao, et al.
Published: (2024)
DaReNeRF: Direction-aware Representation for Dynamic Scenes
by: Lou, Ange, et al.
Published: (2024)
by: Lou, Ange, et al.
Published: (2024)
An Interpretable Cross-Attentive Multi-modal MRI Fusion Framework for Schizophrenia Diagnosis
by: Zhou, Ziyu, et al.
Published: (2024)
by: Zhou, Ziyu, et al.
Published: (2024)
Dissecting RGB-D Learning for Improved Multi-modal Fusion
by: Chen, Hao, et al.
Published: (2023)
by: Chen, Hao, et al.
Published: (2023)
CmFNet: Cross-modal Fusion Network for Weakly-supervised Segmentation of Medical Images
by: Meng, Dongdong, et al.
Published: (2025)
by: Meng, Dongdong, et al.
Published: (2025)
Consistent Instance Field for Dynamic Scene Understanding
by: Wu, Junyi, et al.
Published: (2025)
by: Wu, Junyi, et al.
Published: (2025)
Universal Beta Splatting
by: Liu, Rong, et al.
Published: (2025)
by: Liu, Rong, et al.
Published: (2025)
Dual Diffusion Models for Multi-modal Guided 3D Avatar Generation
by: Li, Hong, et al.
Published: (2026)
by: Li, Hong, et al.
Published: (2026)
Mars Traversability Prediction: A Multi-modal Self-supervised Approach for Costmap Generation
by: Xie, Zongwu, et al.
Published: (2025)
by: Xie, Zongwu, et al.
Published: (2025)
MSSDF: Modality-Shared Self-supervised Distillation for High-Resolution Multi-modal Remote Sensing Image Learning
by: Wang, Tong, et al.
Published: (2025)
by: Wang, Tong, et al.
Published: (2025)
Semi-supervised Single-view 3D Reconstruction via Multi Shape Prior Fusion Strategy and Self-Attention
by: Zhoua, Wei, et al.
Published: (2024)
by: Zhoua, Wei, et al.
Published: (2024)
SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning
by: Zhang, Runmin, et al.
Published: (2024)
by: Zhang, Runmin, et al.
Published: (2024)
Multi-modal Video Representation Alignment for Robust Self-supervised Driver Distraction Detection
by: Lerch, David J., et al.
Published: (2026)
by: Lerch, David J., et al.
Published: (2026)
Self-supervised Multiplex Consensus Mamba for General Image Fusion
by: Wang, Yingying, et al.
Published: (2025)
by: Wang, Yingying, et al.
Published: (2025)
MambaFusion: Height-Fidelity Dense Global Fusion for Multi-modal 3D Object Detection
by: Wang, Hanshi, et al.
Published: (2025)
by: Wang, Hanshi, et al.
Published: (2025)
Textual Inversion and Self-supervised Refinement for Radiology Report Generation
by: Luo, Yuanjiang, et al.
Published: (2024)
by: Luo, Yuanjiang, et al.
Published: (2024)
Cross-Attentive Multiview Fusion of Vision-Language Embeddings
by: Martins, Tomas Berriel, et al.
Published: (2026)
by: Martins, Tomas Berriel, et al.
Published: (2026)
VideoFusion: A Spatio-Temporal Collaborative Network for Multi-modal Video Fusion
by: Tang, Linfeng, et al.
Published: (2025)
by: Tang, Linfeng, et al.
Published: (2025)
Similar Items
-
Few-Shot 3D Volumetric Segmentation with Multi-Surrogate Fusion
by: Zheng, Meng, et al.
Published: (2024) -
Automated Patient Positioning with Learned 3D Hand Gestures
by: Gao, Zhongpai, et al.
Published: (2024) -
Automating Catheterization Labs with Real-Time Perception
by: Yang, Fan, et al.
Published: (2024) -
3D Vision-Language Gaussian Splatting
by: Peng, Qucheng, et al.
Published: (2024) -
Anatomy-Aware Conditional Image-Text Retrieval
by: Zheng, Meng, et al.
Published: (2025)