Multi Activity Sequence Alignment via Implicit Clustering
Fuente:
arXiv
Saved in:
| Main Authors: | Kwon, Taein, Pataki, Zador, Rad, Mahdi, Pollefeys, Marc |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MP-SfM: Monocular Surface Priors for Robust Structure-from-Motion
by: Pataki, Zador, et al.
Published: (2025)
by: Pataki, Zador, et al.
Published: (2025)
Recognising BSL Fingerspelling in Continuous Signing Sequences
by: Chan, Alyssa, et al.
Published: (2026)
by: Chan, Alyssa, et al.
Published: (2026)
EgoPressure: A Dataset for Hand Pressure and Pose Estimation in Egocentric Vision
by: Zhao, Yiming, et al.
Published: (2024)
by: Zhao, Yiming, et al.
Published: (2024)
Space3D-Bench: Spatial 3D Question Answering Benchmark
by: Szymanska, Emilia, et al.
Published: (2024)
by: Szymanska, Emilia, et al.
Published: (2024)
AdaptToken: Entropy-based Adaptive Token Selection for MLLM Long Video Understanding
by: Qi, Haozhe, et al.
Published: (2026)
by: Qi, Haozhe, et al.
Published: (2026)
Loc3R-VLM: Language-based Localization and 3D Reasoning with Vision-Language Models
by: Qu, Kevin, et al.
Published: (2026)
by: Qu, Kevin, et al.
Published: (2026)
EgoWorld: Translating Exocentric View to Egocentric View using Rich Exocentric Observations
by: Park, Junho, et al.
Published: (2025)
by: Park, Junho, et al.
Published: (2025)
Understanding Co-speech Gestures in-the-wild
by: Hegde, Sindhu B, et al.
Published: (2025)
by: Hegde, Sindhu B, et al.
Published: (2025)
CrossOver: 3D Scene Cross-Modal Alignment
by: Sarkar, Sayan Deb, et al.
Published: (2025)
by: Sarkar, Sayan Deb, et al.
Published: (2025)
TORA: Topological Representation Alignment for 3D Shape Assembly
by: Lee, Nahyuk, et al.
Published: (2026)
by: Lee, Nahyuk, et al.
Published: (2026)
CoPE-VideoLM: Leveraging Codec Primitives For Efficient Video Language Modeling
by: Sarkar, Sayan Deb, et al.
Published: (2026)
by: Sarkar, Sayan Deb, et al.
Published: (2026)
VisualChef: Generating Visual Aids in Cooking via Mask Inpainting
by: Kuzyk, Oleh, et al.
Published: (2025)
by: Kuzyk, Oleh, et al.
Published: (2025)
Lightweight and Accurate Multi-View Stereo with Confidence-Aware Diffusion Model
by: Wang, Fangjinhua, et al.
Published: (2025)
by: Wang, Fangjinhua, et al.
Published: (2025)
Object-X: Learning to Reconstruct Multi-Modal 3D Object Representations
by: Di Lorenzo, Gaia, et al.
Published: (2025)
by: Di Lorenzo, Gaia, et al.
Published: (2025)
Fixing the RANSAC Stopping Criterion
by: Schönberger, Johannes, et al.
Published: (2025)
by: Schönberger, Johannes, et al.
Published: (2025)
Gravity-aligned Rotation Averaging with Circular Regression
by: Pan, Linfei, et al.
Published: (2024)
by: Pan, Linfei, et al.
Published: (2024)
Learning to Make Keypoints Sub-Pixel Accurate
by: Kim, Shinjeong, et al.
Published: (2024)
by: Kim, Shinjeong, et al.
Published: (2024)
Retrieval Robust to Object Motion Blur
by: Zou, Rong, et al.
Published: (2024)
by: Zou, Rong, et al.
Published: (2024)
Active Visual Localization for Multi-Agent Collaboration: A Data-Driven Approach
by: Hanlon, Matthew, et al.
Published: (2023)
by: Hanlon, Matthew, et al.
Published: (2023)
Multi-Level Neural Scene Graphs for Dynamic Urban Environments
by: Fischer, Tobias, et al.
Published: (2024)
by: Fischer, Tobias, et al.
Published: (2024)
EgoXtreme: A Dataset for Robust Object Pose Estimation in Egocentric Views under Extreme Conditions
by: Yoon, Taegyoon, et al.
Published: (2026)
by: Yoon, Taegyoon, et al.
Published: (2026)
LSA: Localized Semantic Alignment for Enhancing Temporal Consistency in Traffic Video Generation
by: Karimov, Mirlan, et al.
Published: (2026)
by: Karimov, Mirlan, et al.
Published: (2026)
Synthesizing Consistent Novel Views via 3D Epipolar Attention without Re-Training
by: Ye, Botao, et al.
Published: (2025)
by: Ye, Botao, et al.
Published: (2025)
Enhancing Video Super-Resolution via Implicit Resampling-based Alignment
by: Xu, Kai, et al.
Published: (2023)
by: Xu, Kai, et al.
Published: (2023)
ReSplat: Learning Recurrent Gaussian Splatting
by: Xu, Haofei, et al.
Published: (2025)
by: Xu, Haofei, et al.
Published: (2025)
SegSplat: Feed-forward Gaussian Splatting and Open-Set Semantic Segmentation
by: Siegel, Peter, et al.
Published: (2025)
by: Siegel, Peter, et al.
Published: (2025)
MegaFlow: Zero-Shot Large Displacement Optical Flow
by: Zhang, Dingxi, et al.
Published: (2026)
by: Zhang, Dingxi, et al.
Published: (2026)
Multiway Point Cloud Mosaicking with Diffusion and Global Optimization
by: Jin, Shengze, et al.
Published: (2024)
by: Jin, Shengze, et al.
Published: (2024)
LEAP-VO: Long-term Effective Any Point Tracking for Visual Odometry
by: Chen, Weirong, et al.
Published: (2024)
by: Chen, Weirong, et al.
Published: (2024)
ResFields: Residual Neural Fields for Spatiotemporal Signals
by: Mihajlovic, Marko, et al.
Published: (2023)
by: Mihajlovic, Marko, et al.
Published: (2023)
No Pose, No Problem in 4D: Feed-Forward Dynamic Gaussians from Unposed Multi-View Videos
by: Balice, Matteo, et al.
Published: (2026)
by: Balice, Matteo, et al.
Published: (2026)
UMA: Ultra-detailed Human Avatars via Multi-level Surface Alignment
by: Zhu, Heming, et al.
Published: (2025)
by: Zhu, Heming, et al.
Published: (2025)
From Synchrony to Sequence: Exo-to-Ego Generation via Interpolation
by: Mahdi, Mohammad, et al.
Published: (2026)
by: Mahdi, Mohammad, et al.
Published: (2026)
Global Structure-from-Motion Revisited
by: Pan, Linfei, et al.
Published: (2024)
by: Pan, Linfei, et al.
Published: (2024)
GeoCalib: Learning Single-image Calibration with Geometric Optimization
by: Veicht, Alexander, et al.
Published: (2024)
by: Veicht, Alexander, et al.
Published: (2024)
Know Your Neighbors: Improving Single-View Reconstruction via Spatial Vision-Language Reasoning
by: Li, Rui, et al.
Published: (2024)
by: Li, Rui, et al.
Published: (2024)
INST-Align: Implicit Neural Alignment for Spatial Transcriptomics via Canonical Expression Fields
by: Han, Bonian, et al.
Published: (2026)
by: Han, Bonian, et al.
Published: (2026)
HouseTour: A Virtual Real Estate A(I)gent
by: Çelen, Ata, et al.
Published: (2025)
by: Çelen, Ata, et al.
Published: (2025)
MuRF: Multi-Baseline Radiance Fields
by: Xu, Haofei, et al.
Published: (2023)
by: Xu, Haofei, et al.
Published: (2023)
Learning-based Multi-View Stereo: A Survey
by: Wang, Fangjinhua, et al.
Published: (2024)
by: Wang, Fangjinhua, et al.
Published: (2024)
Similar Items
-
MP-SfM: Monocular Surface Priors for Robust Structure-from-Motion
by: Pataki, Zador, et al.
Published: (2025) -
Recognising BSL Fingerspelling in Continuous Signing Sequences
by: Chan, Alyssa, et al.
Published: (2026) -
EgoPressure: A Dataset for Hand Pressure and Pose Estimation in Egocentric Vision
by: Zhao, Yiming, et al.
Published: (2024) -
Space3D-Bench: Spatial 3D Question Answering Benchmark
by: Szymanska, Emilia, et al.
Published: (2024) -
AdaptToken: Entropy-based Adaptive Token Selection for MLLM Long Video Understanding
by: Qi, Haozhe, et al.
Published: (2026)