Saved in:
| Main Authors: | Hao, Yu, Huang, Hao, Yuan, Shuaihang, Fang, Yi |
|---|---|
| Format: | Preprint |
| Published: |
2021
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2110.03854 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Chain-of-Thought Subspace Meta-Learning for Few-shot Image Captioning with Large Vision and Language Models
by: Huang, Hao, et al.
Published: (2025)
by: Huang, Hao, et al.
Published: (2025)
AnyImageNav: Any-View Geometry for Precise Last-Meter Image-Goal Navigation
by: Deng, Yijie, et al.
Published: (2026)
by: Deng, Yijie, et al.
Published: (2026)
A Multi-Modal Foundation Model to Assist People with Blindness and Low Vision in Environmental Interaction
by: Hao, Yu, et al.
Published: (2023)
by: Hao, Yu, et al.
Published: (2023)
Hierarchical Scoring with 3D Gaussian Splatting for Instance Image-Goal Navigation
by: Deng, Yijie, et al.
Published: (2025)
by: Deng, Yijie, et al.
Published: (2025)
3D Unsupervised Region-Aware Registration Transformer
by: Hao, Yu, et al.
Published: (2021)
by: Hao, Yu, et al.
Published: (2021)
MapBERT: Bitwise Masked Modeling for Real-Time Semantic Mapping Generation
by: Deng, Yijie, et al.
Published: (2025)
by: Deng, Yijie, et al.
Published: (2025)
FairCLIP: Harnessing Fairness in Vision-Language Learning
by: Luo, Yan, et al.
Published: (2024)
by: Luo, Yan, et al.
Published: (2024)
SCA3D: Enhancing Cross-modal 3D Retrieval via 3D Shape and Caption Paired Data Augmentation
by: Ren, Junlong, et al.
Published: (2025)
by: Ren, Junlong, et al.
Published: (2025)
Tex3D: Objects as Attack Surfaces via Adversarial 3D Textures for Vision-Language-Action Models
by: Chen, Jiawei, et al.
Published: (2026)
by: Chen, Jiawei, et al.
Published: (2026)
D-Convexity: A Unified Differentiable Convex Shape Prior via Quasi-Concavity for Data-driven Image Segmentation
by: Chen, Shengzhe, et al.
Published: (2026)
by: Chen, Shengzhe, et al.
Published: (2026)
Deep Convolutional Neural Networks Meet Variational Shape Compactness Priors for Image Segmentation
by: Zhang, Kehui, et al.
Published: (2024)
by: Zhang, Kehui, et al.
Published: (2024)
Hierarchical Transformers for Unsupervised 3D Shape Abstraction
by: Vora, Aditya, et al.
Published: (2025)
by: Vora, Aditya, et al.
Published: (2025)
Shape My Moves: Text-Driven Shape-Aware Synthesis of Human Motions
by: Liao, Ting-Hsuan, et al.
Published: (2025)
by: Liao, Ting-Hsuan, et al.
Published: (2025)
CSGaussian: Progressive Rate-Distortion Compression and Segmentation for 3D Gaussian Splatting
by: Tseng, Yu-Jen, et al.
Published: (2026)
by: Tseng, Yu-Jen, et al.
Published: (2026)
Voxify3D: Pixel Art Meets Volumetric Rendering
by: Huang, Yi-Chuan, et al.
Published: (2025)
by: Huang, Yi-Chuan, et al.
Published: (2025)
SparseFlex: High-Resolution and Arbitrary-Topology 3D Shape Modeling
by: He, Xianglong, et al.
Published: (2025)
by: He, Xianglong, et al.
Published: (2025)
Hierarchical Action Learning for Weakly-Supervised Action Segmentation
by: Huang, Junxian, et al.
Published: (2026)
by: Huang, Junxian, et al.
Published: (2026)
FAST3DIS: Feed-forward Anchored Scene Transformer for 3D Instance Segmentation
by: Li, Changyang, et al.
Published: (2026)
by: Li, Changyang, et al.
Published: (2026)
PARTFIELD: Learning 3D Feature Fields for Part Segmentation and Beyond
by: Liu, Minghua, et al.
Published: (2025)
by: Liu, Minghua, et al.
Published: (2025)
3rd Place Solution for MeViS Track in CVPR 2024 PVUW workshop: Motion Expression guided Video Segmentation
by: Pan, Feiyu, et al.
Published: (2024)
by: Pan, Feiyu, et al.
Published: (2024)
PinPoint3D: Fine-Grained 3D Part Segmentation from a Few Clicks
by: Zhang, Bojun, et al.
Published: (2025)
by: Zhang, Bojun, et al.
Published: (2025)
Revisiting the Scale Loss Function and Gaussian-Shape Convolution for Infrared Small Target Detection
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
3D Weakly Supervised Semantic Segmentation with 2D Vision-Language Guidance
by: Xu, Xiaoxu, et al.
Published: (2024)
by: Xu, Xiaoxu, et al.
Published: (2024)
S3O: A Dual-Phase Approach for Reconstructing Dynamic Shape and Skeleton of Articulated Objects from Single Monocular Video
by: Zhang, Hao, et al.
Published: (2024)
by: Zhang, Hao, et al.
Published: (2024)
SegGraph: Leveraging Graphs of SAM Segments for Few-Shot 3D Part Segmentation
by: Hu, Yueyang, et al.
Published: (2025)
by: Hu, Yueyang, et al.
Published: (2025)
CNS-Edit: 3D Shape Editing via Coupled Neural Shape Optimization
by: Hu, Jingyu, et al.
Published: (2024)
by: Hu, Jingyu, et al.
Published: (2024)
CSL: Class-Agnostic Structure-Constrained Learning for Segmentation Including the Unseen
by: Zhang, Hao, et al.
Published: (2023)
by: Zhang, Hao, et al.
Published: (2023)
Graph-Guided Dual-Level Augmentation for 3D Scene Segmentation
by: Lin, Hongbin, et al.
Published: (2025)
by: Lin, Hongbin, et al.
Published: (2025)
Learning Spectral-Decomposed Tokens for Domain Generalized Semantic Segmentation
by: Yi, Jingjun, et al.
Published: (2024)
by: Yi, Jingjun, et al.
Published: (2024)
PartDistill: 3D Shape Part Segmentation by Vision-Language Model Distillation
by: Umam, Ardian, et al.
Published: (2023)
by: Umam, Ardian, et al.
Published: (2023)
Open-Vocabulary 3D Semantic Segmentation with Text-to-Image Diffusion Models
by: Zhu, Xiaoyu, et al.
Published: (2024)
by: Zhu, Xiaoyu, et al.
Published: (2024)
ShapeMamba-EM: Fine-Tuning Foundation Model with Local Shape Descriptors and Mamba Blocks for 3D EM Image Segmentation
by: Shi, Ruohua, et al.
Published: (2024)
by: Shi, Ruohua, et al.
Published: (2024)
PASS:Test-Time Prompting to Adapt Styles and Semantic Shapes in Medical Image Segmentation
by: Zhang, Chuyan, et al.
Published: (2024)
by: Zhang, Chuyan, et al.
Published: (2024)
PDF: A Probability-Driven Framework for Open World 3D Point Cloud Semantic Segmentation
by: Xu, Jinfeng, et al.
Published: (2024)
by: Xu, Jinfeng, et al.
Published: (2024)
3D-R1: Enhancing Reasoning in 3D VLMs for Unified Scene Understanding
by: Huang, Ting, et al.
Published: (2025)
by: Huang, Ting, et al.
Published: (2025)
MASH: Masked Anchored SpHerical Distances for 3D Shape Representation and Generation
by: Li, Changhao, et al.
Published: (2025)
by: Li, Changhao, et al.
Published: (2025)
Unify3D: An Augmented Holistic End-to-end Monocular 3D Human Reconstruction via Anatomy Shaping and Twins Negotiating
by: Yao, Nanjie, et al.
Published: (2025)
by: Yao, Nanjie, et al.
Published: (2025)
A Light and Smart Wearable Platform with Multimodal Foundation Model for Enhanced Spatial Reasoning in People with Blindness and Low Vision
by: Magay, Alexey, et al.
Published: (2025)
by: Magay, Alexey, et al.
Published: (2025)
3x2: 3D Object Part Segmentation by 2D Semantic Correspondences
by: Thai, Anh, et al.
Published: (2024)
by: Thai, Anh, et al.
Published: (2024)
3D Dental Model Segmentation with Geometrical Boundary Preserving
by: Xi, Shufan, et al.
Published: (2025)
by: Xi, Shufan, et al.
Published: (2025)
Similar Items
-
A Chain-of-Thought Subspace Meta-Learning for Few-shot Image Captioning with Large Vision and Language Models
by: Huang, Hao, et al.
Published: (2025) -
AnyImageNav: Any-View Geometry for Precise Last-Meter Image-Goal Navigation
by: Deng, Yijie, et al.
Published: (2026) -
A Multi-Modal Foundation Model to Assist People with Blindness and Low Vision in Environmental Interaction
by: Hao, Yu, et al.
Published: (2023) -
Hierarchical Scoring with 3D Gaussian Splatting for Instance Image-Goal Navigation
by: Deng, Yijie, et al.
Published: (2025) -
3D Unsupervised Region-Aware Registration Transformer
by: Hao, Yu, et al.
Published: (2021)