CUS3D :CLIP-based Unsupervised 3D Segmentation via Object-level Denoise
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Fuyang, Tian, Runze, Wang, Zhen, Wang, Xiaochuan, Liang, Xiaohui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Ego3DT: Tracking Every 3D Object in Ego-centric Videos
von: Hao, Shengyu, et al.
Veröffentlicht: (2024)
von: Hao, Shengyu, et al.
Veröffentlicht: (2024)
MTFusion: Reconstructing Any 3D Object from Single Image Using Multi-word Textual Inversion
von: Liu, Yu, et al.
Veröffentlicht: (2024)
von: Liu, Yu, et al.
Veröffentlicht: (2024)
Enhancing 3D Gaussian Splatting Compression via Spatial Condition-based Prediction
von: Ma, Jingui, et al.
Veröffentlicht: (2025)
von: Ma, Jingui, et al.
Veröffentlicht: (2025)
REArtGS: Reconstructing and Generating Articulated Objects via 3D Gaussian Splatting with Geometric and Motion Constraints
von: Wu, Di, et al.
Veröffentlicht: (2025)
von: Wu, Di, et al.
Veröffentlicht: (2025)
Querying Autonomous Vehicle Point Clouds: Enhanced by 3D Object Counting with CounterNet
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)
3D Gaussian Editing with A Single Image
von: Luo, Guan, et al.
Veröffentlicht: (2024)
von: Luo, Guan, et al.
Veröffentlicht: (2024)
DanceCamera3D: 3D Camera Movement Synthesis with Music and Dance
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
Zero-Shot Visual Grounding in 3D Gaussians via View Retrieval
von: Liao, Liwei, et al.
Veröffentlicht: (2025)
von: Liao, Liwei, et al.
Veröffentlicht: (2025)
Learning Efficient Unsupervised Satellite Image-based Building Damage Detection
von: Zhang, Yiyun, et al.
Veröffentlicht: (2023)
von: Zhang, Yiyun, et al.
Veröffentlicht: (2023)
Pre-training CLIP against Data Poisoning with Optimal Transport-based Matching and Alignment
von: Zhang, Tong, et al.
Veröffentlicht: (2025)
von: Zhang, Tong, et al.
Veröffentlicht: (2025)
Riemann-based Multi-scale Attention Reasoning Network for Text-3D Retrieval
von: Li, Wenrui, et al.
Veröffentlicht: (2024)
von: Li, Wenrui, et al.
Veröffentlicht: (2024)
GaussianTalker: Speaker-specific Talking Head Synthesis via 3D Gaussian Splatting
von: Yu, Hongyun, et al.
Veröffentlicht: (2024)
von: Yu, Hongyun, et al.
Veröffentlicht: (2024)
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding
von: Li, Zeju, et al.
Veröffentlicht: (2024)
von: Li, Zeju, et al.
Veröffentlicht: (2024)
MOC-3D: Manifold-Order Consistency for Text-to-3D Generation
von: Fan, Chenyang, et al.
Veröffentlicht: (2026)
von: Fan, Chenyang, et al.
Veröffentlicht: (2026)
PixCLIP: Achieving Fine-grained Visual Language Understanding via Any-granularity Pixel-Text Alignment Learning
von: Xiao, Yicheng, et al.
Veröffentlicht: (2025)
von: Xiao, Yicheng, et al.
Veröffentlicht: (2025)
FlashSplat: 2D to 3D Gaussian Splatting Segmentation Solved Optimally
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
von: Shen, Qiuhong, et al.
Veröffentlicht: (2024)
SizeGS: Size-aware Compression of 3D Gaussian Splatting via Mixed Integer Programming
von: Xie, Shuzhao, et al.
Veröffentlicht: (2024)
von: Xie, Shuzhao, et al.
Veröffentlicht: (2024)
Unlearning the Noisy Correspondence Makes CLIP More Robust
von: Han, Haochen, et al.
Veröffentlicht: (2025)
von: Han, Haochen, et al.
Veröffentlicht: (2025)
VideoCLIP-XL: Advancing Long Description Understanding for Video CLIP Models
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024)
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024)
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
von: Zeng, Kang, et al.
Veröffentlicht: (2024)
von: Zeng, Kang, et al.
Veröffentlicht: (2024)
HCNQA: Enhancing 3D VQA with Hierarchical Concentration Narrowing Supervision
von: Zhou, Shengli, et al.
Veröffentlicht: (2025)
von: Zhou, Shengli, et al.
Veröffentlicht: (2025)
Mono3DVG-EnSD: Enhanced Spatial-aware and Dimension-decoupled Text Encoding for Monocular 3D Visual Grounding
von: Li, Yuzhen, et al.
Veröffentlicht: (2025)
von: Li, Yuzhen, et al.
Veröffentlicht: (2025)
Hi3D: Pursuing High-Resolution Image-to-3D Generation with Video Diffusion Models
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
CLIP Brings Better Features to Visual Aesthetics Learners
von: Xu, Liwu, et al.
Veröffentlicht: (2023)
von: Xu, Liwu, et al.
Veröffentlicht: (2023)
Selective Vision-Language Subspace Projection for Few-shot CLIP
von: Zhu, Xingyu, et al.
Veröffentlicht: (2024)
von: Zhu, Xingyu, et al.
Veröffentlicht: (2024)
RealX3D: A Physically-Degraded 3D Benchmark for Multi-view Visual Restoration and Reconstruction
von: Liu, Shuhong, et al.
Veröffentlicht: (2025)
von: Liu, Shuhong, et al.
Veröffentlicht: (2025)
Perceive-Sample-Compress: Towards Real-Time 3D Gaussian Splatting
von: Wang, Zijian, et al.
Veröffentlicht: (2025)
von: Wang, Zijian, et al.
Veröffentlicht: (2025)
Local Neighborhood Features for 3D Classification
von: Sheshappanavar, Shivanand Venkanna, et al.
Veröffentlicht: (2022)
von: Sheshappanavar, Shivanand Venkanna, et al.
Veröffentlicht: (2022)
DanceCamAnimator: Keyframe-Based Controllable 3D Dance Camera Synthesis
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
von: Wang, Zixuan, et al.
Veröffentlicht: (2024)
Magic3DSketch: Create Colorful 3D Models From Sketch-Based 3D Modeling Guided by Text and Language-Image Pre-Training
von: Zang, Ying, et al.
Veröffentlicht: (2024)
von: Zang, Ying, et al.
Veröffentlicht: (2024)
FreeEnhance: Tuning-Free Image Enhancement via Content-Consistent Noising-and-Denoising Process
von: Luo, Yang, et al.
Veröffentlicht: (2024)
von: Luo, Yang, et al.
Veröffentlicht: (2024)
HDiffTG: A Lightweight Hybrid Diffusion-Transformer-GCN Architecture for 3D Human Pose Estimation
von: Fu, Yajie, et al.
Veröffentlicht: (2025)
von: Fu, Yajie, et al.
Veröffentlicht: (2025)
Test-Time Adaptation with CLIP Reward for Zero-Shot Generalization in Vision-Language Models
von: Zhao, Shuai, et al.
Veröffentlicht: (2023)
von: Zhao, Shuai, et al.
Veröffentlicht: (2023)
UniScene: Multi-Camera Unified Pre-training via 3D Scene Reconstruction for Autonomous Driving
von: Min, Chen, et al.
Veröffentlicht: (2023)
von: Min, Chen, et al.
Veröffentlicht: (2023)
Rendering-Oriented 3D Point Cloud Attribute Compression using Sparse Tensor-based Transformer
von: Huo, Xiao, et al.
Veröffentlicht: (2024)
von: Huo, Xiao, et al.
Veröffentlicht: (2024)
Unified Coding for Both Human Perception and Generalized Machine Analytics with CLIP Supervision
von: Yin, Kangsheng, et al.
Veröffentlicht: (2025)
von: Yin, Kangsheng, et al.
Veröffentlicht: (2025)
Adaptive 3D Gaussian Splatting Video Streaming
von: Gong, Han, et al.
Veröffentlicht: (2025)
von: Gong, Han, et al.
Veröffentlicht: (2025)
SMC++: Masked Learning of Unsupervised Video Semantic Compression
von: Tian, Yuan, et al.
Veröffentlicht: (2024)
von: Tian, Yuan, et al.
Veröffentlicht: (2024)
3D2M Dataset: A 3-Dimension diverse Mesh Dataset
von: Dasgupta, Sankarshan
Veröffentlicht: (2024)
von: Dasgupta, Sankarshan
Veröffentlicht: (2024)
Ähnliche Einträge
-
Ego3DT: Tracking Every 3D Object in Ego-centric Videos
von: Hao, Shengyu, et al.
Veröffentlicht: (2024) -
MTFusion: Reconstructing Any 3D Object from Single Image Using Multi-word Textual Inversion
von: Liu, Yu, et al.
Veröffentlicht: (2024) -
Enhancing 3D Gaussian Splatting Compression via Spatial Condition-based Prediction
von: Ma, Jingui, et al.
Veröffentlicht: (2025) -
REArtGS: Reconstructing and Generating Articulated Objects via 3D Gaussian Splatting with Geometric and Motion Constraints
von: Wu, Di, et al.
Veröffentlicht: (2025) -
Querying Autonomous Vehicle Point Clouds: Enhanced by 3D Object Counting with CounterNet
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2025)