Learning segmentation from point trajectories
Fuente:
arXiv
Saved in:
| Main Authors: | Karazija, Laurynas, Laina, Iro, Rupprecht, Christian, Vedaldi, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
What Happens Next? Anticipating Future Motion by Generating Point Trajectories
by: Boduljak, Gabrijel, et al.
Published: (2025)
by: Boduljak, Gabrijel, et al.
Published: (2025)
Diffusion Models for Open-Vocabulary Segmentation
by: Karazija, Laurynas, et al.
Published: (2023)
by: Karazija, Laurynas, et al.
Published: (2023)
IM-3D: Iterative Multiview Diffusion and Reconstruction for High-Quality 3D Generation
by: Melas-Kyriazi, Luke, et al.
Published: (2024)
by: Melas-Kyriazi, Luke, et al.
Published: (2024)
N2F2: Hierarchical Scene Understanding with Nested Neural Feature Fields
by: Bhalgat, Yash, et al.
Published: (2024)
by: Bhalgat, Yash, et al.
Published: (2024)
3D-Aware Instance Segmentation and Tracking in Egocentric Videos
by: Bhalgat, Yash, et al.
Published: (2024)
by: Bhalgat, Yash, et al.
Published: (2024)
DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness
by: Li, Ruining, et al.
Published: (2025)
by: Li, Ruining, et al.
Published: (2025)
VFMF: World Modeling by Forecasting Vision Foundation Model Features
by: Boduljak, Gabrijel, et al.
Published: (2025)
by: Boduljak, Gabrijel, et al.
Published: (2025)
Invisible Stitch: Generating Smooth 3D Scenes with Depth Inpainting
by: Engstler, Paul, et al.
Published: (2024)
by: Engstler, Paul, et al.
Published: (2024)
SynCity: Training-Free Generation of 3D Worlds
by: Engstler, Paul, et al.
Published: (2025)
by: Engstler, Paul, et al.
Published: (2025)
DGE: Direct Gaussian 3D Editing by Consistent Multi-view Editing
by: Chen, Minghao, et al.
Published: (2024)
by: Chen, Minghao, et al.
Published: (2024)
Puppet-Master: Scaling Interactive Video Generation as a Motion Prior for Part-Level Dynamics
by: Li, Ruining, et al.
Published: (2024)
by: Li, Ruining, et al.
Published: (2024)
Layered Motion Fusion: Lifting Motion Segmentation to 3D in Egocentric Videos
by: Tschernezki, Vadim, et al.
Published: (2025)
by: Tschernezki, Vadim, et al.
Published: (2025)
TRAVL: A Recipe for Making Video-Language Models Better Judges of Physics Implausibility
by: Motamed, Saman, et al.
Published: (2025)
by: Motamed, Saman, et al.
Published: (2025)
Splatt3R: Zero-shot Gaussian Splatting from Uncalibrated Image Pairs
by: Smart, Brandon, et al.
Published: (2024)
by: Smart, Brandon, et al.
Published: (2024)
Mesh4D: 4D Mesh Reconstruction and Tracking from Monocular Video
by: Jiang, Zeren, et al.
Published: (2026)
by: Jiang, Zeren, et al.
Published: (2026)
When Do Diffusion Models learn to Generate Multiple Objects?
by: Jeong, Yujin, et al.
Published: (2026)
by: Jeong, Yujin, et al.
Published: (2026)
AnimalClue: Recognizing Animals by their Traces
by: Shinoda, Risa, et al.
Published: (2025)
by: Shinoda, Risa, et al.
Published: (2025)
Meta 3D TextureGen: Fast and Consistent Texture Generation for 3D Objects
by: Bensadoun, Raphael, et al.
Published: (2024)
by: Bensadoun, Raphael, et al.
Published: (2024)
Particulate: Feed-Forward 3D Object Articulation
by: Li, Ruining, et al.
Published: (2025)
by: Li, Ruining, et al.
Published: (2025)
Semantic segmentation with coarse annotations
by: de Jong, Jort, et al.
Published: (2025)
by: de Jong, Jort, et al.
Published: (2025)
Ensemble architecture in polyp segmentation
by: Hsu, Hao-Yun, et al.
Published: (2024)
by: Hsu, Hao-Yun, et al.
Published: (2024)
Advanced computer vision for extracting georeferenced vehicle trajectories from drone imagery
by: Fonod, Robert, et al.
Published: (2024)
by: Fonod, Robert, et al.
Published: (2024)
Pitfalls of topology-aware image segmentation
by: Berger, Alexander H., et al.
Published: (2024)
by: Berger, Alexander H., et al.
Published: (2024)
A methodology for clinically driven interactive segmentation evaluation
by: Esmaeili, Parhom, et al.
Published: (2025)
by: Esmaeili, Parhom, et al.
Published: (2025)
Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction
by: Jiang, Zeren, et al.
Published: (2025)
by: Jiang, Zeren, et al.
Published: (2025)
DOSE3 : Diffusion-based Out-of-distribution detection on SE(3) trajectories
by: Cheng, Hongzhe, et al.
Published: (2025)
by: Cheng, Hongzhe, et al.
Published: (2025)
SHIC: Shape-Image Correspondences with no Keypoint Supervision
by: Shtedritski, Aleksandar, et al.
Published: (2024)
by: Shtedritski, Aleksandar, et al.
Published: (2024)
Splatter Image: Ultra-Fast Single-View 3D Reconstruction
by: Szymanowicz, Stanislaw, et al.
Published: (2023)
by: Szymanowicz, Stanislaw, et al.
Published: (2023)
Tracktention: Leveraging Point Tracking to Attend Videos Faster and Better
by: Lai, Zihang, et al.
Published: (2025)
by: Lai, Zihang, et al.
Published: (2025)
Universal representations:The missing link between faces, text, planktons, and cat breeds
by: Bilen, Hakan, et al.
Published: (2017)
by: Bilen, Hakan, et al.
Published: (2017)
Autonomous state-space segmentation for Deep-RL sparse reward scenarios
by: Maselli, Gianluca, et al.
Published: (2025)
by: Maselli, Gianluca, et al.
Published: (2025)
Revisiting MAE pre-training for 3D medical image segmentation
by: Wald, Tassilo, et al.
Published: (2024)
by: Wald, Tassilo, et al.
Published: (2024)
COIN: Counterfactual inpainting for weakly supervised semantic segmentation for medical images
by: Shvetsov, Dmytro, et al.
Published: (2024)
by: Shvetsov, Dmytro, et al.
Published: (2024)
Towards Multi-Modal Animal Pose Estimation: A Survey and In-Depth Analysis
by: Deng, Qianyi, et al.
Published: (2024)
by: Deng, Qianyi, et al.
Published: (2024)
MultiSolSegment: Multi-channel segmentation of overlapping features in electroluminescence images of photovoltaic cells
by: Sanghi, Ojas, et al.
Published: (2026)
by: Sanghi, Ojas, et al.
Published: (2026)
DragAPart: Learning a Part-Level Motion Prior for Articulated Objects
by: Li, Ruining, et al.
Published: (2024)
by: Li, Ruining, et al.
Published: (2024)
WCEbleedGen: A wireless capsule endoscopy dataset and its benchmarking for automatic bleeding classification, detection, and segmentation
by: Handa, Palak, et al.
Published: (2024)
by: Handa, Palak, et al.
Published: (2024)
How to select slices for annotation to train best-performing deep learning segmentation models for cross-sectional medical images?
by: Zhang, Yixin, et al.
Published: (2024)
by: Zhang, Yixin, et al.
Published: (2024)
Syn4D: A Multiview Synthetic 4D Dataset
by: Jiang, Zeren, et al.
Published: (2026)
by: Jiang, Zeren, et al.
Published: (2026)
Universal Algorithm-Implicit Learning
by: Woerner, Stefano, et al.
Published: (2026)
by: Woerner, Stefano, et al.
Published: (2026)
Similar Items
-
What Happens Next? Anticipating Future Motion by Generating Point Trajectories
by: Boduljak, Gabrijel, et al.
Published: (2025) -
Diffusion Models for Open-Vocabulary Segmentation
by: Karazija, Laurynas, et al.
Published: (2023) -
IM-3D: Iterative Multiview Diffusion and Reconstruction for High-Quality 3D Generation
by: Melas-Kyriazi, Luke, et al.
Published: (2024) -
N2F2: Hierarchical Scene Understanding with Nested Neural Feature Fields
by: Bhalgat, Yash, et al.
Published: (2024) -
3D-Aware Instance Segmentation and Tracking in Egocentric Videos
by: Bhalgat, Yash, et al.
Published: (2024)