Saved in:
| Main Authors: | Liu, Jing, Wang, Duanchu, Gong, Haoran, Wang, Chongyu, Zhu, Jihua, Wang, Di |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.03637 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CompetitorFormer: Competitor Transformer for 3D Instance Segmentation
by: Wang, Duanchu, et al.
Published: (2024)
by: Wang, Duanchu, et al.
Published: (2024)
OpenUrban3D: Annotation-Free Open-Vocabulary Semantic Segmentation of Large-Scale Urban Point Clouds
by: Wang, Chongyu, et al.
Published: (2025)
by: Wang, Chongyu, et al.
Published: (2025)
ForestHG-Trace: Traceable Long-Horizon Ecological Reasoning over Large-Scale Forest Scenes
by: Cheng, Zihang, et al.
Published: (2026)
by: Cheng, Zihang, et al.
Published: (2026)
PointQ-Bench: Benchmarking Diagnostic and Interpretable Point Cloud Quality Assessment
by: Wang, Duanchu, et al.
Published: (2026)
by: Wang, Duanchu, et al.
Published: (2026)
Efficiently Expanding Receptive Fields: Local Split Attention and Parallel Aggregation for Enhanced Large-scale Point Cloud Semantic Segmentation
by: Wang, Haodong, et al.
Published: (2024)
by: Wang, Haodong, et al.
Published: (2024)
PartNeXt: A Next-Generation Dataset for Fine-Grained and Hierarchical 3D Part Understanding
by: Wang, Penghao, et al.
Published: (2025)
by: Wang, Penghao, et al.
Published: (2025)
BuildAnyPoint: 3D Building Structured Abstraction from Diverse Point Clouds
by: Hua, Tongyan, et al.
Published: (2026)
by: Hua, Tongyan, et al.
Published: (2026)
Multilateral Cascading Network for Semantic Segmentation of Large-Scale Outdoor Point Clouds
by: Gong, Haoran, et al.
Published: (2024)
by: Gong, Haoran, et al.
Published: (2024)
Advancing Fine-Grained Visual Understanding with Multi-Scale Alignment in Multi-Modal Models
by: Wang, Wei, et al.
Published: (2024)
by: Wang, Wei, et al.
Published: (2024)
Understanding the Role of Pathways in a Deep Neural Network
by: Lyu, Lei, et al.
Published: (2024)
by: Lyu, Lei, et al.
Published: (2024)
New Dataset and Methods for Fine-Grained Compositional Referring Expression Comprehension via Specialist-MLLM Collaboration
by: Yang, Xuzheng, et al.
Published: (2025)
by: Yang, Xuzheng, et al.
Published: (2025)
UniM-OV3D: Uni-Modality Open-Vocabulary 3D Scene Understanding with Fine-Grained Feature Representation
by: He, Qingdong, et al.
Published: (2024)
by: He, Qingdong, et al.
Published: (2024)
Text-guided Fine-Grained Video Anomaly Understanding
by: Gu, Jihao, et al.
Published: (2025)
by: Gu, Jihao, et al.
Published: (2025)
Parameter-efficient Prompt Learning for 3D Point Cloud Understanding
by: Sun, Hongyu, et al.
Published: (2024)
by: Sun, Hongyu, et al.
Published: (2024)
CMHANet: A Cross-Modal Hybrid Attention Network for Point Cloud Registration
by: Zhang, Dongxu, et al.
Published: (2026)
by: Zhang, Dongxu, et al.
Published: (2026)
SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding
by: Luo, Junwei, et al.
Published: (2024)
by: Luo, Junwei, et al.
Published: (2024)
Matching Distance and Geometric Distribution Aided Learning Multiview Point Cloud Registration
by: Li, Shiqi, et al.
Published: (2025)
by: Li, Shiqi, et al.
Published: (2025)
Prior-Constrained Association Learning for Fine-Grained Generalized Category Discovery
by: Wang, Menglin, et al.
Published: (2025)
by: Wang, Menglin, et al.
Published: (2025)
Fine-Grained Open-Vocabulary Object Detection with Fined-Grained Prompts: Task, Dataset and Benchmark
by: Liu, Ying, et al.
Published: (2025)
by: Liu, Ying, et al.
Published: (2025)
SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes
by: Huang, Jiaxin, et al.
Published: (2025)
by: Huang, Jiaxin, et al.
Published: (2025)
Let Geometry GUIDE: Layer-wise Unrolling of Geometric Priors in Multimodal LLMs
by: Wang, Chongyu, et al.
Published: (2026)
by: Wang, Chongyu, et al.
Published: (2026)
Structure-Aware Fine-Grained Gaussian Splatting for Expressive Avatar Reconstruction
by: Su, Yuze, et al.
Published: (2026)
by: Su, Yuze, et al.
Published: (2026)
Advancing Fine-Grained Classification by Structure and Subject Preserving Augmentation
by: Michaeli, Eyal, et al.
Published: (2024)
by: Michaeli, Eyal, et al.
Published: (2024)
Fine-Grained Domain Generalization with Feature Structuralization
by: Yu, Wenlong, et al.
Published: (2024)
by: Yu, Wenlong, et al.
Published: (2024)
FD$^2$: A Dedicated Framework for Fine-Grained Dataset Distillation
by: Ma, Hongxu, et al.
Published: (2026)
by: Ma, Hongxu, et al.
Published: (2026)
KPM-Bench: A Kinematic Parsing Motion Benchmark for Fine-grained Motion-centric Video Understanding
by: Lin, Boda, et al.
Published: (2026)
by: Lin, Boda, et al.
Published: (2026)
3D-JEPA: A Joint Embedding Predictive Architecture for 3D Self-Supervised Representation Learning
by: Hu, Naiwen, et al.
Published: (2024)
by: Hu, Naiwen, et al.
Published: (2024)
MMDocBench: Benchmarking Large Vision-Language Models for Fine-Grained Visual Document Understanding
by: Zhu, Fengbin, et al.
Published: (2024)
by: Zhu, Fengbin, et al.
Published: (2024)
FineCops-Ref: A new Dataset and Task for Fine-Grained Compositional Referring Expression Comprehension
by: Liu, Junzhuo, et al.
Published: (2024)
by: Liu, Junzhuo, et al.
Published: (2024)
VSFormer: Mining Correlations in Flexible View Set for Multi-view 3D Shape Understanding
by: Sun, Hongyu, et al.
Published: (2024)
by: Sun, Hongyu, et al.
Published: (2024)
VERIFIED: A Video Corpus Moment Retrieval Benchmark for Fine-Grained Video Understanding
by: Chen, Houlun, et al.
Published: (2024)
by: Chen, Houlun, et al.
Published: (2024)
FiLA-Video: Spatio-Temporal Compression for Fine-Grained Long Video Understanding
by: Guo, Yanan, et al.
Published: (2025)
by: Guo, Yanan, et al.
Published: (2025)
DenseScan: Advancing 3D Scene Understanding with 2D Dense Annotation
by: Wang, Zirui, et al.
Published: (2025)
by: Wang, Zirui, et al.
Published: (2025)
Mantis: Mamba-native Tuning is Efficient for 3D Point Cloud Foundation Models
by: Guo, Zihao, et al.
Published: (2026)
by: Guo, Zihao, et al.
Published: (2026)
Towards Fine-Grained Human Motion Video Captioning
by: Song, Guorui, et al.
Published: (2025)
by: Song, Guorui, et al.
Published: (2025)
Hyperbolic Image-and-Pointcloud Contrastive Learning for 3D Classification
by: Hu, Naiwen, et al.
Published: (2024)
by: Hu, Naiwen, et al.
Published: (2024)
Noisy Ostracods: A Fine-Grained, Imbalanced Real-World Dataset for Benchmarking Robust Machine Learning and Label Correction Methods
by: Hu, Jiamian, et al.
Published: (2024)
by: Hu, Jiamian, et al.
Published: (2024)
ArtiMuse: Fine-Grained Image Aesthetics Assessment with Joint Scoring and Expert-Level Understanding
by: Cao, Shuo, et al.
Published: (2025)
by: Cao, Shuo, et al.
Published: (2025)
StreamForest: Efficient Online Video Understanding with Persistent Event Memory
by: Zeng, Xiangyu, et al.
Published: (2025)
by: Zeng, Xiangyu, et al.
Published: (2025)
Fine-Grained Generalization via Structuralizing Concept and Feature Space into Commonality, Specificity and Confounding
by: Wang, Zhen, et al.
Published: (2026)
by: Wang, Zhen, et al.
Published: (2026)
Similar Items
-
CompetitorFormer: Competitor Transformer for 3D Instance Segmentation
by: Wang, Duanchu, et al.
Published: (2024) -
OpenUrban3D: Annotation-Free Open-Vocabulary Semantic Segmentation of Large-Scale Urban Point Clouds
by: Wang, Chongyu, et al.
Published: (2025) -
ForestHG-Trace: Traceable Long-Horizon Ecological Reasoning over Large-Scale Forest Scenes
by: Cheng, Zihang, et al.
Published: (2026) -
PointQ-Bench: Benchmarking Diagnostic and Interpretable Point Cloud Quality Assessment
by: Wang, Duanchu, et al.
Published: (2026) -
Efficiently Expanding Receptive Fields: Local Split Attention and Parallel Aggregation for Enhanced Large-scale Point Cloud Semantic Segmentation
by: Wang, Haodong, et al.
Published: (2024)