Task-aligned Part-aware Panoptic Segmentation through Joint Object-Part Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | de Geus, Daan, Dubbelman, Gijs |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How to Benchmark Vision Foundation Models for Semantic Segmentation?
von: Kerssies, Tommie, et al.
Veröffentlicht: (2024)
von: Kerssies, Tommie, et al.
Veröffentlicht: (2024)
PMT: Plain Mask Transformer for Image and Video Segmentation with Frozen Vision Encoders
von: Cavagnero, Niccolò, et al.
Veröffentlicht: (2026)
von: Cavagnero, Niccolò, et al.
Veröffentlicht: (2026)
ALGM: Adaptive Local-then-Global Token Merging for Efficient Semantic Segmentation with Plain Vision Transformers
von: Norouzi, Narges, et al.
Veröffentlicht: (2024)
von: Norouzi, Narges, et al.
Veröffentlicht: (2024)
First Place Solution to the ECCV 2024 BRAVO Challenge: Evaluating Robustness of Vision Foundation Models for Semantic Segmentation
von: Kerssies, Tommie, et al.
Veröffentlicht: (2024)
von: Kerssies, Tommie, et al.
Veröffentlicht: (2024)
VidEoMT: Your ViT is Secretly Also a Video Segmentation Model
von: Norouzi, Narges, et al.
Veröffentlicht: (2026)
von: Norouzi, Narges, et al.
Veröffentlicht: (2026)
Exploring the Benefits of Vision Foundation Models for Unsupervised Domain Adaptation
von: Englert, Brunó B., et al.
Veröffentlicht: (2024)
von: Englert, Brunó B., et al.
Veröffentlicht: (2024)
Your ViT is Secretly an Image Segmentation Model
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
PanopticPartFormer++: A Unified and Decoupled View for Panoptic Part Segmentation
von: Li, Xiangtai, et al.
Veröffentlicht: (2023)
von: Li, Xiangtai, et al.
Veröffentlicht: (2023)
A Frame is Worth One Token: Efficient Generative World Modeling with Delta Tokens
von: Kerssies, Tommie, et al.
Veröffentlicht: (2026)
von: Kerssies, Tommie, et al.
Veröffentlicht: (2026)
VFM-UDA++: Improving Network Architectures and Data Strategies for Unsupervised Domain Adaptive Semantic Segmentation
von: Englert, Brunó B., et al.
Veröffentlicht: (2025)
von: Englert, Brunó B., et al.
Veröffentlicht: (2025)
Depth-aware Panoptic Segmentation
von: Nguyen, Tuan, et al.
Veröffentlicht: (2024)
von: Nguyen, Tuan, et al.
Veröffentlicht: (2024)
Towards Data-Efficient Video Pre-training with Frozen Image Foundation Models
von: Orlova, Svetlana, et al.
Veröffentlicht: (2026)
von: Orlova, Svetlana, et al.
Veröffentlicht: (2026)
Orion-Lite: Distilling LLM Reasoning into Efficient Vision-Only Driving Models
von: Gu, Jing, et al.
Veröffentlicht: (2026)
von: Gu, Jing, et al.
Veröffentlicht: (2026)
REFNet++: Multi-Task Efficient Fusion of Camera and Radar Sensor Data in Bird's-Eye Polar View
von: Chandrasekaran, Kavin, et al.
Veröffentlicht: (2026)
von: Chandrasekaran, Kavin, et al.
Veröffentlicht: (2026)
DONUT: A Decoder-Only Model for Trajectory Prediction
von: Knoche, Markus, et al.
Veröffentlicht: (2025)
von: Knoche, Markus, et al.
Veröffentlicht: (2025)
A Resource Efficient Fusion Network for Object Detection in Bird's-Eye View using Camera and Raw Radar Data
von: Chandrasekaran, Kavin, et al.
Veröffentlicht: (2024)
von: Chandrasekaran, Kavin, et al.
Veröffentlicht: (2024)
What is the Added Value of UDA in the VFM Era?
von: Englert, Brunó B., et al.
Veröffentlicht: (2025)
von: Englert, Brunó B., et al.
Veröffentlicht: (2025)
BPJDet: Extended Object Representation for Generic Body-Part Joint Detection
von: Zhou, Huayi, et al.
Veröffentlicht: (2023)
von: Zhou, Huayi, et al.
Veröffentlicht: (2023)
PanDepth: Joint Panoptic Segmentation and Depth Completion
von: Lagos, Juan, et al.
Veröffentlicht: (2022)
von: Lagos, Juan, et al.
Veröffentlicht: (2022)
Part-aware Prompted Segment Anything Model for Adaptive Segmentation
von: Zhao, Chenhui, et al.
Veröffentlicht: (2024)
von: Zhao, Chenhui, et al.
Veröffentlicht: (2024)
Beyond Viewpoint: Robust 3D Object Recognition under Arbitrary Views through Joint Multi-Part Representation
von: Fan, Linlong, et al.
Veröffentlicht: (2024)
von: Fan, Linlong, et al.
Veröffentlicht: (2024)
Simplifying Traffic Anomaly Detection with Video Foundation Models
von: Orlova, Svetlana, et al.
Veröffentlicht: (2025)
von: Orlova, Svetlana, et al.
Veröffentlicht: (2025)
PanSR: An Object-Centric Mask Transformer for Panoptic Segmentation
von: Žust, Lojze, et al.
Veröffentlicht: (2024)
von: Žust, Lojze, et al.
Veröffentlicht: (2024)
PartSTAD: 2D-to-3D Part Segmentation Task Adaptation
von: Kim, Hyunjin, et al.
Veröffentlicht: (2024)
von: Kim, Hyunjin, et al.
Veröffentlicht: (2024)
Balancing Shared and Task-Specific Representations: A Hybrid Approach to Depth-Aware Video Panoptic Segmentation
von: Stolle, Kurt H. W.
Veröffentlicht: (2024)
von: Stolle, Kurt H. W.
Veröffentlicht: (2024)
Multi-Part Object Representations via Graph Structures and Co-Part Discovery
von: Foo, Alex, et al.
Veröffentlicht: (2025)
von: Foo, Alex, et al.
Veröffentlicht: (2025)
DINO in the Room: Leveraging 2D Foundation Models for 3D Segmentation
von: Knaebel, Karim, et al.
Veröffentlicht: (2025)
von: Knaebel, Karim, et al.
Veröffentlicht: (2025)
InstructPart: Task-Oriented Part Segmentation with Instruction Reasoning
von: Wan, Zifu, et al.
Veröffentlicht: (2025)
von: Wan, Zifu, et al.
Veröffentlicht: (2025)
Redefining Instance Matching: A Unified Framework for Part-Aware Matching in Panoptic Segmentation Evaluation
von: Großkopf, Erik, et al.
Veröffentlicht: (2026)
von: Großkopf, Erik, et al.
Veröffentlicht: (2026)
Revisiting Radar Perception With Spectral Point Clouds
von: Alsharif, Hamza, et al.
Veröffentlicht: (2026)
von: Alsharif, Hamza, et al.
Veröffentlicht: (2026)
Sa2VA-i: Improving Sa2VA Results with Consistent Training and Inference
von: Nekrasov, Alexey, et al.
Veröffentlicht: (2025)
von: Nekrasov, Alexey, et al.
Veröffentlicht: (2025)
How Important are Videos for Training Video LLMs?
von: Lydakis, George, et al.
Veröffentlicht: (2025)
von: Lydakis, George, et al.
Veröffentlicht: (2025)
Volume Transformer: Revisiting Vanilla Transformers for 3D Scene Understanding
von: Yilmaz, Kadir, et al.
Veröffentlicht: (2026)
von: Yilmaz, Kadir, et al.
Veröffentlicht: (2026)
Mitigating Objectness Bias and Region-to-Text Misalignment for Open-Vocabulary Panoptic Segmentation
von: Kormushev, Nikolay, et al.
Veröffentlicht: (2026)
von: Kormushev, Nikolay, et al.
Veröffentlicht: (2026)
SurgPIS: Surgical-instrument-level Instances and Part-level Semantics for Weakly-supervised Part-aware Instance Segmentation
von: Wei, Meng, et al.
Veröffentlicht: (2025)
von: Wei, Meng, et al.
Veröffentlicht: (2025)
PartCraft: Crafting Creative Objects by Parts
von: Ng, Kam Woh, et al.
Veröffentlicht: (2024)
von: Ng, Kam Woh, et al.
Veröffentlicht: (2024)
InstanceBEV: Unifying Instance and BEV Representation for 3D Panoptic Segmentation
von: Li, Feng, et al.
Veröffentlicht: (2025)
von: Li, Feng, et al.
Veröffentlicht: (2025)
Part-aware Unified Representation of Language and Skeleton for Zero-shot Action Recognition
von: Zhu, Anqi, et al.
Veröffentlicht: (2024)
von: Zhu, Anqi, et al.
Veröffentlicht: (2024)
COCONut-PanCap: Joint Panoptic Segmentation and Grounded Captions for Fine-Grained Understanding and Generation
von: Deng, Xueqing, et al.
Veröffentlicht: (2025)
von: Deng, Xueqing, et al.
Veröffentlicht: (2025)
Lidar Panoptic Segmentation in an Open World
von: Chakravarthy, Anirudh S, et al.
Veröffentlicht: (2024)
von: Chakravarthy, Anirudh S, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
How to Benchmark Vision Foundation Models for Semantic Segmentation?
von: Kerssies, Tommie, et al.
Veröffentlicht: (2024) -
PMT: Plain Mask Transformer for Image and Video Segmentation with Frozen Vision Encoders
von: Cavagnero, Niccolò, et al.
Veröffentlicht: (2026) -
ALGM: Adaptive Local-then-Global Token Merging for Efficient Semantic Segmentation with Plain Vision Transformers
von: Norouzi, Narges, et al.
Veröffentlicht: (2024) -
First Place Solution to the ECCV 2024 BRAVO Challenge: Evaluating Robustness of Vision Foundation Models for Semantic Segmentation
von: Kerssies, Tommie, et al.
Veröffentlicht: (2024) -
VidEoMT: Your ViT is Secretly Also a Video Segmentation Model
von: Norouzi, Narges, et al.
Veröffentlicht: (2026)