Understanding Multi-Granularity for Open-Vocabulary Part Segmentation
Fuente:
arXiv
Guardado en:
| Autores principales: | Choi, Jiho, Lee, Seonho, Lee, Seungho, Lee, Minhyun, Shim, Hyunjung |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Fine-Grained Image-Text Correspondence with Cost Aggregation for Open-Vocabulary Part Segmentation
por: Choi, Jiho, et al.
Publicado: (2025)
por: Choi, Jiho, et al.
Publicado: (2025)
MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation
por: Lee, Minhyun, et al.
Publicado: (2024)
por: Lee, Minhyun, et al.
Publicado: (2024)
Learning from Spatio-temporal Correlation for Semi-Supervised LiDAR Semantic Segmentation
por: Lee, Seungho, et al.
Publicado: (2024)
por: Lee, Seungho, et al.
Publicado: (2024)
Scribble-Guided Diffusion for Training-free Text-to-Image Generation
por: Lee, Seonho, et al.
Publicado: (2024)
por: Lee, Seonho, et al.
Publicado: (2024)
Self-Supervised Vision Transformers Are Efficient Segmentation Learners for Imperfect Labels
por: Lee, Seungho, et al.
Publicado: (2024)
por: Lee, Seungho, et al.
Publicado: (2024)
DreamCatalyst: Fast and High-Quality 3D Editing via Controlling Editability and Identity Preservation
por: Kim, Jiwook, et al.
Publicado: (2024)
por: Kim, Jiwook, et al.
Publicado: (2024)
Weakly Supervised Semantic Segmentation for Driving Scenes
por: Kim, Dongseob, et al.
Publicado: (2023)
por: Kim, Dongseob, et al.
Publicado: (2023)
3D-Aware Vision-Language Models Fine-Tuning with Geometric Distillation
por: Lee, Seonho, et al.
Publicado: (2025)
por: Lee, Seonho, et al.
Publicado: (2025)
What "Not" to Detect: Negation-Aware VLMs via Structured Reasoning and Token Merging
por: Kang, Inha, et al.
Publicado: (2025)
por: Kang, Inha, et al.
Publicado: (2025)
Blind to Position, Biased in Language: Probing Mid-Layer Representational Bias in Vision-Language Encoders for Zero-Shot Language-Grounded Spatial Understanding
por: An, Na Min, et al.
Publicado: (2025)
por: An, Na Min, et al.
Publicado: (2025)
Sampling Bag of Views for Open-Vocabulary Object Detection
por: Choi, Hojun, et al.
Publicado: (2024)
por: Choi, Hojun, et al.
Publicado: (2024)
SeiT++: Masked Token Modeling Improves Storage-efficient Training
por: Lee, Minhyun, et al.
Publicado: (2023)
por: Lee, Minhyun, et al.
Publicado: (2023)
CoT-PL: Chain-of-Thought Pseudo-Labeling for Open-Vocabulary Object Detection
por: Choi, Hojun, et al.
Publicado: (2025)
por: Choi, Hojun, et al.
Publicado: (2025)
Effective SAM Combination for Open-Vocabulary Semantic Segmentation
por: Lee, Minhyeok, et al.
Publicado: (2024)
por: Lee, Minhyeok, et al.
Publicado: (2024)
MomentMix Augmentation with Length-Aware DETR for Temporally Robust Moment Retrieval
por: Park, Seojeong, et al.
Publicado: (2024)
por: Park, Seojeong, et al.
Publicado: (2024)
Personalized OVSS: Understanding Personal Concept in Open-Vocabulary Semantic Segmentation
por: Park, Sunghyun, et al.
Publicado: (2025)
por: Park, Sunghyun, et al.
Publicado: (2025)
WaymoQA: A Multi-View Visual Question Answering Dataset for Safety-Critical Reasoning in Autonomous Driving
por: Yu, Seungjun, et al.
Publicado: (2025)
por: Yu, Seungjun, et al.
Publicado: (2025)
No Thing, Nothing: Highlighting Safety-Critical Classes for Robust LiDAR Semantic Segmentation in Adverse Weather
por: Park, Junsung, et al.
Publicado: (2025)
por: Park, Junsung, et al.
Publicado: (2025)
Beyond-Labels: Advancing Open-Vocabulary Segmentation With Vision-Language Models
por: Rahman, Muhammad Atta ur, et al.
Publicado: (2025)
por: Rahman, Muhammad Atta ur, et al.
Publicado: (2025)
Periodic-MAE: Periodic Video Masked Autoencoder for rPPG Estimation
por: Choi, Jiho, et al.
Publicado: (2025)
por: Choi, Jiho, et al.
Publicado: (2025)
Classifier-guided CLIP Distillation for Unsupervised Multi-label Classification
por: Kim, Dongseob, et al.
Publicado: (2025)
por: Kim, Dongseob, et al.
Publicado: (2025)
EmbodiedSplat: Online Feed-Forward Semantic 3DGS for Open-Vocabulary 3D Scene Understanding
por: Lee, Seungjun, et al.
Publicado: (2026)
por: Lee, Seungjun, et al.
Publicado: (2026)
Rethinking Open-Vocabulary Segmentation of Radiance Fields in 3D Space
por: Lee, Hyunjee, et al.
Publicado: (2024)
por: Lee, Hyunjee, et al.
Publicado: (2024)
OVS Meets Continual Learning: Towards Sustainable Open-Vocabulary Segmentation
por: Hwang, Dongjun, et al.
Publicado: (2024)
por: Hwang, Dongjun, et al.
Publicado: (2024)
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
por: Lim, Youngsun, et al.
Publicado: (2024)
por: Lim, Youngsun, et al.
Publicado: (2024)
Rethinking Data Augmentation for Robust LiDAR Semantic Segmentation in Adverse Weather
por: Park, Junsung, et al.
Publicado: (2024)
por: Park, Junsung, et al.
Publicado: (2024)
Exploring Hierarchical Consistency and Unbiased Objectness for Open-Vocabulary Object Detection
por: Lee, Sanghoon, et al.
Publicado: (2026)
por: Lee, Sanghoon, et al.
Publicado: (2026)
Latent Inversion with Timestep-aware Sampling for Training-free Non-rigid Editing
por: Jung, Yunji, et al.
Publicado: (2024)
por: Jung, Yunji, et al.
Publicado: (2024)
Open-Vocabulary Semantic Part Segmentation of 3D Human
por: Suzuki, Keito, et al.
Publicado: (2025)
por: Suzuki, Keito, et al.
Publicado: (2025)
econSG: Efficient and Multi-view Consistent Open-Vocabulary 3D Semantic Gaussians
por: Zhang, Can, et al.
Publicado: (2025)
por: Zhang, Can, et al.
Publicado: (2025)
AnyDepth-DETR/-YOLO: Any-depth object detection with a single network
por: Kang, Woochul, et al.
Publicado: (2026)
por: Kang, Woochul, et al.
Publicado: (2026)
LangHOPS: Language Grounded Hierarchical Open-Vocabulary Part Segmentation
por: Miao, Yang, et al.
Publicado: (2025)
por: Miao, Yang, et al.
Publicado: (2025)
Precision matters: Precision-aware ensemble for weakly supervised semantic segmentation
por: Park, Junsung, et al.
Publicado: (2024)
por: Park, Junsung, et al.
Publicado: (2024)
Addressing Image Hallucination in Text-to-Image Generation through Factual Image Retrieval
por: Lim, Youngsun, et al.
Publicado: (2024)
por: Lim, Youngsun, et al.
Publicado: (2024)
GUIDED: Granular Understanding via Identification, Detection, and Discrimination for Fine-Grained Open-Vocabulary Object Detection
por: Li, Jiaming, et al.
Publicado: (2026)
por: Li, Jiaming, et al.
Publicado: (2026)
Video-Oasis: Rethinking Evaluation of Video Understanding
por: Lim, Geuntaek, et al.
Publicado: (2026)
por: Lim, Geuntaek, et al.
Publicado: (2026)
Prompt the Unseen: Evaluating Visual-Language Alignment Beyond Supervision
por: Jung, Raehyuk, et al.
Publicado: (2025)
por: Jung, Raehyuk, et al.
Publicado: (2025)
PCA-Seg: Revisiting Cost Aggregation for Open-Vocabulary Semantic and Part Segmentation
por: Yin, Jianjian, et al.
Publicado: (2026)
por: Yin, Jianjian, et al.
Publicado: (2026)
Mosaic3D: Foundation Dataset and Model for Open-Vocabulary 3D Segmentation
por: Lee, Junha, et al.
Publicado: (2025)
por: Lee, Junha, et al.
Publicado: (2025)
From Open-Vocabulary to Vocabulary-Free Semantic Segmentation
por: Reichard, Klara, et al.
Publicado: (2025)
por: Reichard, Klara, et al.
Publicado: (2025)
Ejemplares similares
-
Fine-Grained Image-Text Correspondence with Cost Aggregation for Open-Vocabulary Part Segmentation
por: Choi, Jiho, et al.
Publicado: (2025) -
MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation
por: Lee, Minhyun, et al.
Publicado: (2024) -
Learning from Spatio-temporal Correlation for Semi-Supervised LiDAR Semantic Segmentation
por: Lee, Seungho, et al.
Publicado: (2024) -
Scribble-Guided Diffusion for Training-free Text-to-Image Generation
por: Lee, Seonho, et al.
Publicado: (2024) -
Self-Supervised Vision Transformers Are Efficient Segmentation Learners for Imperfect Labels
por: Lee, Seungho, et al.
Publicado: (2024)