Enregistré dans:
| Auteurs principaux: | Wan, Zishuo, Gao, Yu, Pang, Wanyuan, Ding, Dawei |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2501.03482 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
TARDis: Time Attenuated Representation Disentanglement for Incomplete Multi-Modal Tumor Segmentation and Classification
par: Wan, Zishuo, et autres
Publié: (2025)
par: Wan, Zishuo, et autres
Publié: (2025)
VOILA: Evaluation of MLLMs For Perceptual Understanding and Analogical Reasoning
par: Yilmaz, Nilay, et autres
Publié: (2025)
par: Yilmaz, Nilay, et autres
Publié: (2025)
VOILA: Value-of-Information Guided Fidelity Selection for Cost-Aware Multimodal Question Answering
par: Bhope, Rahul Atul, et autres
Publié: (2026)
par: Bhope, Rahul Atul, et autres
Publié: (2026)
Sparsity-Aware Voxel Attention and Foreground Modulation for 3D Semantic Scene Completion
par: Xue, Yu, et autres
Publié: (2026)
par: Xue, Yu, et autres
Publié: (2026)
Interactive Test-Time Adaptation with Reliable Spatial-Temporal Voxels for Multi-Modal Segmentation
par: Cao, Haozhi, et autres
Publié: (2024)
par: Cao, Haozhi, et autres
Publié: (2024)
DynamicTree: Interactive Real Tree Animation via Sparse Voxel Spectrum
par: Li, Yaokun, et autres
Publié: (2025)
par: Li, Yaokun, et autres
Publié: (2025)
Context and Geometry Aware Voxel Transformer for Semantic Scene Completion
par: Yu, Zhu, et autres
Publié: (2024)
par: Yu, Zhu, et autres
Publié: (2024)
DivAS: Interactive 3D Segmentation of NeRFs via Depth-Weighted Voxel Aggregation
par: Pande, Ayush
Publié: (2026)
par: Pande, Ayush
Publié: (2026)
Adapting Vision-Language Model with Fine-grained Semantics for Open-Vocabulary Segmentation
par: Chng, Yong Xien, et autres
Publié: (2024)
par: Chng, Yong Xien, et autres
Publié: (2024)
Towards Universal Text-driven CT Image Segmentation
par: Li, Yuheng, et autres
Publié: (2025)
par: Li, Yuheng, et autres
Publié: (2025)
TiFRe: Text-guided Video Frame Reduction for Efficient Video Multi-modal Large Language Models
par: Zheng, Xiangtian, et autres
Publié: (2026)
par: Zheng, Xiangtian, et autres
Publié: (2026)
OpenVoxel: Training-Free Grouping and Captioning Voxels for Open-Vocabulary 3D Scene Understanding
par: Huang, Sheng-Yu, et autres
Publié: (2026)
par: Huang, Sheng-Yu, et autres
Publié: (2026)
Taming Mambas for Voxel Level 3D Medical Image Segmentation
par: Lumetti, Luca, et autres
Publié: (2024)
par: Lumetti, Luca, et autres
Publié: (2024)
Not All Voxels Are Equal: Hardness-Aware Semantic Scene Completion with Self-Distillation
par: Wang, Song, et autres
Publié: (2024)
par: Wang, Song, et autres
Publié: (2024)
LiteVoxel: Low-memory Intelligent Thresholding for Efficient Voxel Rasterization
par: Lee, Jee Won, et autres
Publié: (2025)
par: Lee, Jee Won, et autres
Publié: (2025)
UniVoxel: Fast Inverse Rendering by Unified Voxelization of Scene Representation
par: Wu, Shuang, et autres
Publié: (2024)
par: Wu, Shuang, et autres
Publié: (2024)
Learning Trajectory-Aware Multimodal Large Language Models for Video Reasoning Segmentation
par: Luo, Jingnan, et autres
Publié: (2026)
par: Luo, Jingnan, et autres
Publié: (2026)
Towards Interactive Lesion Segmentation in Whole-Body PET/CT with Promptable Models
par: Rokuss, Maximilian, et autres
Publié: (2025)
par: Rokuss, Maximilian, et autres
Publié: (2025)
Interactive Segmentation and Report Generation for CT Images
par: Gu, Yannian, et autres
Publié: (2025)
par: Gu, Yannian, et autres
Publié: (2025)
Global Position Aware Group Choreography using Large Language Model
par: Pang, Haozhou, et autres
Publié: (2025)
par: Pang, Haozhou, et autres
Publié: (2025)
Geometrical Cross-Attention and Nonvoid Voxelization for Efficient 3D Medical Image Segmentation
par: Yuan, Chenxin, et autres
Publié: (2026)
par: Yuan, Chenxin, et autres
Publié: (2026)
GLUS: Global-Local Reasoning Unified into A Single Large Language Model for Video Segmentation
par: Lin, Lang, et autres
Publié: (2025)
par: Lin, Lang, et autres
Publié: (2025)
Devil is in Details: Locality-Aware 3D Abdominal CT Volume Generation for Self-Supervised Organ Segmentation
par: Wang, Yuran, et autres
Publié: (2024)
par: Wang, Yuran, et autres
Publié: (2024)
Voxel Densification for Serialized 3D Object Detection: Mitigating Sparsity via Pre-serialization Expansion
par: Liu, Qifeng, et autres
Publié: (2025)
par: Liu, Qifeng, et autres
Publié: (2025)
Open-set Anomaly Segmentation in Complex Scenarios
par: Xia, Song, et autres
Publié: (2025)
par: Xia, Song, et autres
Publié: (2025)
4D Neural Voxel Splatting: Dynamic Scene Rendering with Voxelized Guassian Splatting
par: Wu, Chun-Tin, et autres
Publié: (2025)
par: Wu, Chun-Tin, et autres
Publié: (2025)
VoxelTrack: Exploring Voxel Representation for 3D Point Cloud Object Tracking
par: Lu, Yuxuan, et autres
Publié: (2024)
par: Lu, Yuxuan, et autres
Publié: (2024)
SVRecon: Sparse Voxel Rasterization for Surface Reconstruction
par: Oh, Seunghun, et autres
Publié: (2025)
par: Oh, Seunghun, et autres
Publié: (2025)
PVNet: Point-Voxel Interaction LiDAR Scene Upsampling Via Diffusion Models
par: Cheng, Xianjing, et autres
Publié: (2025)
par: Cheng, Xianjing, et autres
Publié: (2025)
ASSR-NeRF: Arbitrary-Scale Super-Resolution on Voxel Grid for High-Quality Radiance Fields Reconstruction
par: Huang, Ding-Jiun, et autres
Publié: (2024)
par: Huang, Ding-Jiun, et autres
Publié: (2024)
Context-Aware Interaction Network for RGB-T Semantic Segmentation
par: Lv, Ying, et autres
Publié: (2024)
par: Lv, Ying, et autres
Publié: (2024)
Advancing Structured Priors for Sparse-Voxel Surface Reconstruction
par: Chi, Ting-Hsun, et autres
Publié: (2026)
par: Chi, Ting-Hsun, et autres
Publié: (2026)
PointVoxelFormer -- Reviving point cloud networks for 3D medical imaging
par: Heinrich, Mattias Paul
Publié: (2024)
par: Heinrich, Mattias Paul
Publié: (2024)
Language and Geometry Grounded Sparse Voxel Representations for Holistic Scene Understanding
par: Wu, Guile, et autres
Publié: (2026)
par: Wu, Guile, et autres
Publié: (2026)
Dynamic Scene Understanding through Object-Centric Voxelization and Neural Rendering
par: Zhao, Yanpeng, et autres
Publié: (2024)
par: Zhao, Yanpeng, et autres
Publié: (2024)
Interactive Segmentation Model for Placenta Segmentation from 3D Ultrasound images
par: Li, Hao, et autres
Publié: (2024)
par: Li, Hao, et autres
Publié: (2024)
VoxelOpt: Voxel-Adaptive Message Passing for Discrete Optimization in Deformable Abdominal CT Registration
par: Zhang, Hang, et autres
Publié: (2025)
par: Zhang, Hang, et autres
Publié: (2025)
Irregularity Inspection using Neural Radiance Field
par: Ding, Tianqi, et autres
Publié: (2024)
par: Ding, Tianqi, et autres
Publié: (2024)
PIVOT-Net: Heterogeneous Point-Voxel-Tree-based Framework for Point Cloud Compression
par: Pang, Jiahao, et autres
Publié: (2024)
par: Pang, Jiahao, et autres
Publié: (2024)
CAR-SAM: Cross-Attention Reconstruction for Post-Training Quantization of the Segment Anything Model
par: Wen, Houji, et autres
Publié: (2026)
par: Wen, Houji, et autres
Publié: (2026)
Documents similaires
-
TARDis: Time Attenuated Representation Disentanglement for Incomplete Multi-Modal Tumor Segmentation and Classification
par: Wan, Zishuo, et autres
Publié: (2025) -
VOILA: Evaluation of MLLMs For Perceptual Understanding and Analogical Reasoning
par: Yilmaz, Nilay, et autres
Publié: (2025) -
VOILA: Value-of-Information Guided Fidelity Selection for Cost-Aware Multimodal Question Answering
par: Bhope, Rahul Atul, et autres
Publié: (2026) -
Sparsity-Aware Voxel Attention and Foreground Modulation for 3D Semantic Scene Completion
par: Xue, Yu, et autres
Publié: (2026) -
Interactive Test-Time Adaptation with Reliable Spatial-Temporal Voxels for Multi-Modal Segmentation
par: Cao, Haozhi, et autres
Publié: (2024)