Learning Local and Global Temporal Contexts for Video Semantic Segmentation
Fuente:
arXiv
Salvato in:
| Autori principali: | Sun, Guolei, Liu, Yun, Ding, Henghui, Wu, Min, Van Gool, Luc |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Self-Explainable Affordance Learning with Embodied Caption
di: Zhang, Zhipeng, et al.
Pubblicazione: (2024)
di: Zhang, Zhipeng, et al.
Pubblicazione: (2024)
Rethinking Global Context in Crowd Counting
di: Sun, Guolei, et al.
Pubblicazione: (2021)
di: Sun, Guolei, et al.
Pubblicazione: (2021)
Evaluating SAM2 for Video Semantic Segmentation
di: Ariff, Syed Hesham Syed, et al.
Pubblicazione: (2025)
di: Ariff, Syed Hesham Syed, et al.
Pubblicazione: (2025)
Rethinking Few-shot 3D Point Cloud Semantic Segmentation
di: An, Zhaochong, et al.
Pubblicazione: (2024)
di: An, Zhaochong, et al.
Pubblicazione: (2024)
TRAVL: A Recipe for Making Video-Language Models Better Judges of Physics Implausibility
di: Motamed, Saman, et al.
Pubblicazione: (2025)
di: Motamed, Saman, et al.
Pubblicazione: (2025)
CamSAM2: Segment Anything Accurately in Camouflaged Videos
di: Zhou, Yuli, et al.
Pubblicazione: (2025)
di: Zhou, Yuli, et al.
Pubblicazione: (2025)
ObjectRelator: Enabling Cross-View Object Relation Understanding Across Ego-Centric and Exo-Centric Perspectives
di: Fu, Yuqian, et al.
Pubblicazione: (2024)
di: Fu, Yuqian, et al.
Pubblicazione: (2024)
Learning Generative Interactive Environments By Trained Agent Exploration
di: Kazemi, Naser, et al.
Pubblicazione: (2024)
di: Kazemi, Naser, et al.
Pubblicazione: (2024)
Language-Guided Instance-Aware Domain-Adaptive Panoptic Segmentation
di: Mansour, Elham Amin, et al.
Pubblicazione: (2024)
di: Mansour, Elham Amin, et al.
Pubblicazione: (2024)
Cross-View Multi-Modal Segmentation @ Ego-Exo4D Challenges 2025
di: Fu, Yuqian, et al.
Pubblicazione: (2025)
di: Fu, Yuqian, et al.
Pubblicazione: (2025)
VOID: Video Object and Interaction Deletion
di: Motamed, Saman, et al.
Pubblicazione: (2026)
di: Motamed, Saman, et al.
Pubblicazione: (2026)
When SAM2 Meets Video Camouflaged Object Segmentation: A Comprehensive Evaluation and Adaptation
di: Zhou, Yuli, et al.
Pubblicazione: (2024)
di: Zhou, Yuli, et al.
Pubblicazione: (2024)
MedVSR: Medical Video Super-Resolution with Cross State-Space Propagation
di: Liu, Xinyu, et al.
Pubblicazione: (2025)
di: Liu, Xinyu, et al.
Pubblicazione: (2025)
IGL-DT: Iterative Global-Local Feature Learning with Dual-Teacher Semantic Segmentation Framework under Limited Annotation Scheme
di: Tran, Dinh Dai Quan, et al.
Pubblicazione: (2025)
di: Tran, Dinh Dai Quan, et al.
Pubblicazione: (2025)
Vision Transformers with Hierarchical Attention
di: Liu, Yun, et al.
Pubblicazione: (2021)
di: Liu, Yun, et al.
Pubblicazione: (2021)
MUSTAN: Multi-scale Temporal Context as Attention for Robust Video Foreground Segmentation
di: Pokala, Praveen Kumar, et al.
Pubblicazione: (2024)
di: Pokala, Praveen Kumar, et al.
Pubblicazione: (2024)
Leveraging Swin Transformer for Local-to-Global Weakly Supervised Semantic Segmentation
di: Ahmadi, Rozhan, et al.
Pubblicazione: (2024)
di: Ahmadi, Rozhan, et al.
Pubblicazione: (2024)
EgoNight: Towards Egocentric Vision Understanding at Night with a Challenging Benchmark
di: Zhang, Deheng, et al.
Pubblicazione: (2025)
di: Zhang, Deheng, et al.
Pubblicazione: (2025)
LSA: Localized Semantic Alignment for Enhancing Temporal Consistency in Traffic Video Generation
di: Karimov, Mirlan, et al.
Pubblicazione: (2026)
di: Karimov, Mirlan, et al.
Pubblicazione: (2026)
Few-Shot Segmentation with Global and Local Contrastive Learning
di: Liu, Weide, et al.
Pubblicazione: (2021)
di: Liu, Weide, et al.
Pubblicazione: (2021)
MetaSeg: MetaFormer-based Global Contexts-aware Network for Efficient Semantic Segmentation
di: Kang, Beoungwoo, et al.
Pubblicazione: (2024)
di: Kang, Beoungwoo, et al.
Pubblicazione: (2024)
FedSaaS: Class-Consistency Federated Semantic Segmentation via Global Prototype Supervision and Local Adversarial Harmonization
di: Yu, Xiaoyang, et al.
Pubblicazione: (2025)
di: Yu, Xiaoyang, et al.
Pubblicazione: (2025)
Semantic Localization Guiding Segment Anything Model For Reference Remote Sensing Image Segmentation
di: Li, Shuyang, et al.
Pubblicazione: (2025)
di: Li, Shuyang, et al.
Pubblicazione: (2025)
Learning Spatial-Semantic Features for Robust Video Object Segmentation
di: Li, Xin, et al.
Pubblicazione: (2024)
di: Li, Xin, et al.
Pubblicazione: (2024)
Condition-Invariant Semantic Segmentation
di: Sakaridis, Christos, et al.
Pubblicazione: (2023)
di: Sakaridis, Christos, et al.
Pubblicazione: (2023)
Vanishing-Point-Guided Video Semantic Segmentation of Driving Scenes
di: Guo, Diandian, et al.
Pubblicazione: (2024)
di: Guo, Diandian, et al.
Pubblicazione: (2024)
EgoCross: Benchmarking Multimodal Large Language Models for Cross-Domain Egocentric Video Question Answering
di: Li, Yanjun, et al.
Pubblicazione: (2025)
di: Li, Yanjun, et al.
Pubblicazione: (2025)
SweetTok: Semantic-Aware Spatial-Temporal Tokenizer for Compact Video Discretization
di: Tan, Zhentao, et al.
Pubblicazione: (2024)
di: Tan, Zhentao, et al.
Pubblicazione: (2024)
Mask-RadarNet: Enhancing Transformer With Spatial-Temporal Semantic Context for Radar Object Detection in Autonomous Driving
di: Wu, Yuzhi, et al.
Pubblicazione: (2024)
di: Wu, Yuzhi, et al.
Pubblicazione: (2024)
Boosting Audio Visual Question Answering via Key Semantic-Aware Cues
di: Li, Guangyao, et al.
Pubblicazione: (2024)
di: Li, Guangyao, et al.
Pubblicazione: (2024)
Hunting Attributes: Context Prototype-Aware Learning for Weakly Supervised Semantic Segmentation
di: Tang, Feilong, et al.
Pubblicazione: (2024)
di: Tang, Feilong, et al.
Pubblicazione: (2024)
VideoZoomer: Reinforcement-Learned Temporal Focusing for Long Video Reasoning
di: Ding, Yang, et al.
Pubblicazione: (2025)
di: Ding, Yang, et al.
Pubblicazione: (2025)
Optimizing against Infeasible Inclusions from Data for Semantic Segmentation through Morphology
di: Basu, Shamik, et al.
Pubblicazione: (2024)
di: Basu, Shamik, et al.
Pubblicazione: (2024)
Video Understanding: From Geometry and Semantics to Unified Models
di: An, Zhaochong, et al.
Pubblicazione: (2026)
di: An, Zhaochong, et al.
Pubblicazione: (2026)
Semantic Segmentation of Video Sequences with Convolutional LSTMs
di: Pfeuffer, Andreas, et al.
Pubblicazione: (2019)
di: Pfeuffer, Andreas, et al.
Pubblicazione: (2019)
Looking Beyond the Window: Global-Local Aligned CLIP for Training-free Open-Vocabulary Semantic Segmentation
di: Lee, ByeongCheol, et al.
Pubblicazione: (2026)
di: Lee, ByeongCheol, et al.
Pubblicazione: (2026)
Visual Prompt Selection for In-Context Learning Segmentation
di: Suo, Wei, et al.
Pubblicazione: (2024)
di: Suo, Wei, et al.
Pubblicazione: (2024)
C^2DA: Contrastive and Context-aware Domain Adaptive Semantic Segmentation
di: Khan, Md. Al-Masrur, et al.
Pubblicazione: (2024)
di: Khan, Md. Al-Masrur, et al.
Pubblicazione: (2024)
Probabilistic Sampling of Balanced K-Means using Adiabatic Quantum Computing
di: Zaech, Jan-Nico, et al.
Pubblicazione: (2023)
di: Zaech, Jan-Nico, et al.
Pubblicazione: (2023)
Context-Aware Temporal Embedding of Objects in Video Data
di: Farhan, Ahnaf, et al.
Pubblicazione: (2024)
di: Farhan, Ahnaf, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Self-Explainable Affordance Learning with Embodied Caption
di: Zhang, Zhipeng, et al.
Pubblicazione: (2024) -
Rethinking Global Context in Crowd Counting
di: Sun, Guolei, et al.
Pubblicazione: (2021) -
Evaluating SAM2 for Video Semantic Segmentation
di: Ariff, Syed Hesham Syed, et al.
Pubblicazione: (2025) -
Rethinking Few-shot 3D Point Cloud Semantic Segmentation
di: An, Zhaochong, et al.
Pubblicazione: (2024) -
TRAVL: A Recipe for Making Video-Language Models Better Judges of Physics Implausibility
di: Motamed, Saman, et al.
Pubblicazione: (2025)