Point-VOS: Pointing Up Video Object Segmentation
Fuente:
arXiv
Guardado en:
| Autores principales: | Zulfikar, Idil Esen, Mahadevan, Sabarinath, Voigtlaender, Paul, Leibe, Bastian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Interactive4D: Interactive 4D LiDAR Segmentation
por: Fradlin, Ilya, et al.
Publicado: (2024)
por: Fradlin, Ilya, et al.
Publicado: (2024)
VidEoMT: Your ViT is Secretly Also a Video Segmentation Model
por: Norouzi, Narges, et al.
Publicado: (2026)
por: Norouzi, Narges, et al.
Publicado: (2026)
AGILE3D: Attention Guided Interactive Multi-object 3D Segmentation
por: Yue, Yuanwen, et al.
Publicado: (2023)
por: Yue, Yuanwen, et al.
Publicado: (2023)
Point2Vec for Self-Supervised Representation Learning on Point Clouds
por: Knaebel, Karim, et al.
Publicado: (2023)
por: Knaebel, Karim, et al.
Publicado: (2023)
ClickVOS: Click Video Object Segmentation
por: Guo, Pinxue, et al.
Publicado: (2024)
por: Guo, Pinxue, et al.
Publicado: (2024)
ActionVOS: Actions as Prompts for Video Object Segmentation
por: Ouyang, Liangyang, et al.
Publicado: (2024)
por: Ouyang, Liangyang, et al.
Publicado: (2024)
DeVOS: Flow-Guided Deformable Transformer for Video Object Segmentation
por: Fedynyak, Volodymyr, et al.
Publicado: (2024)
por: Fedynyak, Volodymyr, et al.
Publicado: (2024)
LiVOS: Light Video Object Segmentation with Gated Linear Matching
por: Liu, Qin, et al.
Publicado: (2024)
por: Liu, Qin, et al.
Publicado: (2024)
Mask4Former: Mask Transformer for 4D Panoptic Segmentation
por: Yilmaz, Kadir, et al.
Publicado: (2023)
por: Yilmaz, Kadir, et al.
Publicado: (2023)
SurGe: Improved Surface Geometry in Point Maps
por: Knaebel, Karim, et al.
Publicado: (2026)
por: Knaebel, Karim, et al.
Publicado: (2026)
OneVOS: Unifying Video Object Segmentation with All-in-One Transformer Framework
por: Li, Wanyun, et al.
Publicado: (2024)
por: Li, Wanyun, et al.
Publicado: (2024)
UW-VOS: A Large-Scale Dataset for Underwater Video Object Segmentation
por: Zhao, Hongshen, et al.
Publicado: (2026)
por: Zhao, Hongshen, et al.
Publicado: (2026)
How Important are Videos for Training Video LLMs?
por: Lydakis, George, et al.
Publicado: (2025)
por: Lydakis, George, et al.
Publicado: (2025)
Look Gauss, No Pose: Novel View Synthesis using Gaussian Splatting without Accurate Pose Initialization
por: Schmidt, Christian, et al.
Publicado: (2024)
por: Schmidt, Christian, et al.
Publicado: (2024)
Panoptic-CUDAL: Rural Australia Point Cloud Dataset in Rainy Conditions
por: Tseng, Tzu-Yun, et al.
Publicado: (2025)
por: Tseng, Tzu-Yun, et al.
Publicado: (2025)
OoDIS: Anomaly Instance Segmentation and Detection Benchmark
por: Nekrasov, Alexey, et al.
Publicado: (2024)
por: Nekrasov, Alexey, et al.
Publicado: (2024)
Video Object Segmentation via SAM 2: The 4th Solution for LSVOS Challenge VOS Track
por: Pan, Feiyu, et al.
Publicado: (2024)
por: Pan, Feiyu, et al.
Publicado: (2024)
DONUT: A Decoder-Only Model for Trajectory Prediction
por: Knoche, Markus, et al.
Publicado: (2025)
por: Knoche, Markus, et al.
Publicado: (2025)
PointGauss: Point Cloud-Guided Multi-Object Segmentation for Gaussian Splatting
por: Sun, Wentao, et al.
Publicado: (2025)
por: Sun, Wentao, et al.
Publicado: (2025)
OpenSplat3D: Open-Vocabulary 3D Instance Segmentation using Gaussian Splatting
por: Piekenbrinck, Jens, et al.
Publicado: (2025)
por: Piekenbrinck, Jens, et al.
Publicado: (2025)
Spotting the Unexpected (STU): A 3D LiDAR Dataset for Anomaly Segmentation in Autonomous Driving
por: Nekrasov, Alexey, et al.
Publicado: (2025)
por: Nekrasov, Alexey, et al.
Publicado: (2025)
Point2Insert: Video Object Insertion via Sparse Point Guidance
por: Zhou, Yu, et al.
Publicado: (2026)
por: Zhou, Yu, et al.
Publicado: (2026)
Generative Data Augmentation for Object Point Cloud Segmentation
por: Zhu, Dekai, et al.
Publicado: (2025)
por: Zhu, Dekai, et al.
Publicado: (2025)
M$^3$-VOS: Multi-Phase, Multi-Transition, and Multi-Scenery Video Object Segmentation
por: Chen, Zixuan, et al.
Publicado: (2024)
por: Chen, Zixuan, et al.
Publicado: (2024)
PanoVOS: Bridging Non-panoramic and Panoramic Views with Transformer for Video Segmentation
por: Yan, Shilin, et al.
Publicado: (2023)
por: Yan, Shilin, et al.
Publicado: (2023)
P2Object: Single Point Supervised Object Detection and Instance Segmentation
por: Chen, Pengfei, et al.
Publicado: (2025)
por: Chen, Pengfei, et al.
Publicado: (2025)
What is Point Supervision Worth in Video Instance Segmentation?
por: Huang, Shuaiyi, et al.
Publicado: (2024)
por: Huang, Shuaiyi, et al.
Publicado: (2024)
Point, Segment and Count: A Generalized Framework for Object Counting
por: Huang, Zhizhong, et al.
Publicado: (2023)
por: Huang, Zhizhong, et al.
Publicado: (2023)
DINO in the Room: Leveraging 2D Foundation Models for 3D Segmentation
por: Knaebel, Karim, et al.
Publicado: (2025)
por: Knaebel, Karim, et al.
Publicado: (2025)
Your ViT is Secretly an Image Segmentation Model
por: Kerssies, Tommie, et al.
Publicado: (2025)
por: Kerssies, Tommie, et al.
Publicado: (2025)
Text Prompting for Multi-Concept Video Customization by Autoregressive Generation
por: Kothandaraman, Divya, et al.
Publicado: (2024)
por: Kothandaraman, Divya, et al.
Publicado: (2024)
Vanishing-Point-Guided Video Semantic Segmentation of Driving Scenes
por: Guo, Diandian, et al.
Publicado: (2024)
por: Guo, Diandian, et al.
Publicado: (2024)
OCCUQ: Exploring Efficient Uncertainty Quantification for 3D Occupancy Prediction
por: Heidrich, Severin, et al.
Publicado: (2025)
por: Heidrich, Severin, et al.
Publicado: (2025)
Block-Sparse Global Attention for Efficient Multi-View Geometry Transformers
por: Wang, Chung-Shien Brian, et al.
Publicado: (2025)
por: Wang, Chung-Shien Brian, et al.
Publicado: (2025)
Acquisition of high-quality images for camera calibration in robotics applications via speech prompts
por: Linder, Timm, et al.
Publicado: (2025)
por: Linder, Timm, et al.
Publicado: (2025)
Query2Uncertainty: Robust Uncertainty Quantification and Calibration for 3D Object Detection under Distribution Shift
por: Beemelmanns, Till, et al.
Publicado: (2026)
por: Beemelmanns, Till, et al.
Publicado: (2026)
FreePoint: Unsupervised Point Cloud Instance Segmentation
por: Zhang, Zhikai, et al.
Publicado: (2023)
por: Zhang, Zhikai, et al.
Publicado: (2023)
Multi-Granularity Video Object Segmentation
por: Lim, Sangbeom, et al.
Publicado: (2024)
por: Lim, Sangbeom, et al.
Publicado: (2024)
Sa2VA-i: Improving Sa2VA Results with Consistent Training and Inference
por: Nekrasov, Alexey, et al.
Publicado: (2025)
por: Nekrasov, Alexey, et al.
Publicado: (2025)
Volume Transformer: Revisiting Vanilla Transformers for 3D Scene Understanding
por: Yilmaz, Kadir, et al.
Publicado: (2026)
por: Yilmaz, Kadir, et al.
Publicado: (2026)
Ejemplares similares
-
Interactive4D: Interactive 4D LiDAR Segmentation
por: Fradlin, Ilya, et al.
Publicado: (2024) -
VidEoMT: Your ViT is Secretly Also a Video Segmentation Model
por: Norouzi, Narges, et al.
Publicado: (2026) -
AGILE3D: Attention Guided Interactive Multi-object 3D Segmentation
por: Yue, Yuanwen, et al.
Publicado: (2023) -
Point2Vec for Self-Supervised Representation Learning on Point Clouds
por: Knaebel, Karim, et al.
Publicado: (2023) -
ClickVOS: Click Video Object Segmentation
por: Guo, Pinxue, et al.
Publicado: (2024)