3rd Place at CVPR 2026 CASTLE Challenge: Agentic Multi-View Long-Context Video Understanding via Hierarchical Knowledge Graph Retrieval
Fuente:
arXiv
Salvato in:
| Autori principali: | Albusayes, Raghad, Alyahya, Munirah |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
3rd Place Solution for MOSE Track in CVPR 2024 PVUW workshop: Complex Video Object Segmentation
di: Liu, Xinyu, et al.
Pubblicazione: (2024)
di: Liu, Xinyu, et al.
Pubblicazione: (2024)
CuriosAI Submission to the CASTLE Challenge at EgoVis 2026
di: Kanda, Yuto, et al.
Pubblicazione: (2026)
di: Kanda, Yuto, et al.
Pubblicazione: (2026)
3rd Place Solution for MeViS Track in CVPR 2024 PVUW workshop: Motion Expression guided Video Segmentation
di: Pan, Feiyu, et al.
Pubblicazione: (2024)
di: Pan, Feiyu, et al.
Pubblicazione: (2024)
MARS: Technical Report for the CASTLE Challenge at EgoVis 2026
di: Zhang, Haoyu, et al.
Pubblicazione: (2026)
di: Zhang, Haoyu, et al.
Pubblicazione: (2026)
CASTLE2026 Team WDL Technical Report
di: Li, Zhengyang, et al.
Pubblicazione: (2026)
di: Li, Zhengyang, et al.
Pubblicazione: (2026)
3rd Place Solution for PVUW Challenge 2024: Video Panoptic Segmentation
di: Wu, Ruipu, et al.
Pubblicazione: (2024)
di: Wu, Ruipu, et al.
Pubblicazione: (2024)
1st Place Winner of the 2024 Pixel-level Video Understanding in the Wild (CVPR'24 PVUW) Challenge in Video Panoptic Segmentation and Best Long Video Consistency of Video Semantic Segmentation
di: Liu, Qingfeng, et al.
Pubblicazione: (2024)
di: Liu, Qingfeng, et al.
Pubblicazione: (2024)
Visual Agentic Memory: Enabling Online Long Video Understanding via Online Indexing, Hierarchical Memory, and Agentic Retrieval
di: Li, Aiden Yiliu, et al.
Pubblicazione: (2026)
di: Li, Aiden Yiliu, et al.
Pubblicazione: (2026)
X-Restormer++: 1st Place Solution for the UG2+ CVPR 2026 All-Weather Restoration Challenge
di: Pan, Youwei, et al.
Pubblicazione: (2026)
di: Pan, Youwei, et al.
Pubblicazione: (2026)
TempRet: Temporal Enhancement and Two-Stage Reranking for CVPR 2026 EPIC-KITCHENS-100 Multi-Instance Retrieval Challenge
di: Li, Zixu, et al.
Pubblicazione: (2026)
di: Li, Zixu, et al.
Pubblicazione: (2026)
SAMSON: 3rd Place Solution of LSVOS 2025 VOS Challenge
di: Xie, Yujie, et al.
Pubblicazione: (2025)
di: Xie, Yujie, et al.
Pubblicazione: (2025)
3rd Place Solution to ICCV LargeFineFoodAI Retrieval
di: Zhong, Yang, et al.
Pubblicazione: (2025)
di: Zhong, Yang, et al.
Pubblicazione: (2025)
EgoAdapt: A Multi-Scene Egocentric Adaptation Method for CVPR 2026 HD-EPIC VQA Challenge
di: Chen, Zhiwei, et al.
Pubblicazione: (2026)
di: Chen, Zhiwei, et al.
Pubblicazione: (2026)
The Instance-centric Transformer for the RVOS Track of LSVOS Challenge: 3rd Place Solution
di: Cao, Bin, et al.
Pubblicazione: (2024)
di: Cao, Bin, et al.
Pubblicazione: (2024)
FVOS for MOSE Track of 4th PVUW Challenge: 3rd Place Solution
di: Wang, Mengjiao, et al.
Pubblicazione: (2025)
di: Wang, Mengjiao, et al.
Pubblicazione: (2025)
VideoARM: Agentic Reasoning over Hierarchical Memory for Long-Form Video Understanding
di: Yin, Yufei, et al.
Pubblicazione: (2025)
di: Yin, Yufei, et al.
Pubblicazione: (2025)
The CASTLE 2024 Dataset: Advancing the Art of Multimodal Understanding
di: Rossetto, Luca, et al.
Pubblicazione: (2025)
di: Rossetto, Luca, et al.
Pubblicazione: (2025)
1st Place Solution for MOSE Track in CVPR 2024 PVUW Workshop: Complex Video Object Segmentation
di: Miao, Deshui, et al.
Pubblicazione: (2024)
di: Miao, Deshui, et al.
Pubblicazione: (2024)
2nd Place Solution for MOSE Track in CVPR 2024 PVUW workshop: Complex Video Object Segmentation
di: Xu, Zhensong, et al.
Pubblicazione: (2024)
di: Xu, Zhensong, et al.
Pubblicazione: (2024)
Hierarchical Long Video Understanding with Audiovisual Entity Cohesion and Agentic Search
di: Yin, Xinlei, et al.
Pubblicazione: (2026)
di: Yin, Xinlei, et al.
Pubblicazione: (2026)
An Effective Solution for the CVPR 2026 8th UG2+ Challenge Track 3: Dynamic Object Segmentation in Turbulence
di: Li, Hongzhen, et al.
Pubblicazione: (2026)
di: Li, Hongzhen, et al.
Pubblicazione: (2026)
Multi-View Hierarchical Graph Neural Network for Sketch-Based 3D Shape Retrieval
di: Cheng, Hang, et al.
Pubblicazione: (2026)
di: Cheng, Hang, et al.
Pubblicazione: (2026)
LSVOS Challenge 3rd Place Report: SAM2 and Cutie based VOS
di: Liu, Xinyu, et al.
Pubblicazione: (2024)
di: Liu, Xinyu, et al.
Pubblicazione: (2024)
Winner of CVPR2026 NTIRE Challenge on Image Shadow Removal: Semantic and Geometric Guidance for Shadow Removal via Cascaded Refinement
di: Beltrame, Lorenzo, et al.
Pubblicazione: (2026)
di: Beltrame, Lorenzo, et al.
Pubblicazione: (2026)
Technique Report of CVPR 2024 PBDL Challenges
di: Fu, Ying, et al.
Pubblicazione: (2024)
di: Fu, Ying, et al.
Pubblicazione: (2024)
3rd Place Solution for VisDA 2021 Challenge -- Universally Domain Adaptive Image Recognition
di: Liao, Haojin, et al.
Pubblicazione: (2021)
di: Liao, Haojin, et al.
Pubblicazione: (2021)
Agentic Very Long Video Understanding
di: Rege, Aniket, et al.
Pubblicazione: (2026)
di: Rege, Aniket, et al.
Pubblicazione: (2026)
A Robust Semantic Segmentation Pipeline for the CVPR 2026 8th UG2+ Challenge Track 2
di: Chai, Jinming, et al.
Pubblicazione: (2026)
di: Chai, Jinming, et al.
Pubblicazione: (2026)
2nd Place Solution for MeViS Track in CVPR 2024 PVUW Workshop: Motion Expression guided Video Segmentation
di: Cao, Bin, et al.
Pubblicazione: (2024)
di: Cao, Bin, et al.
Pubblicazione: (2024)
1st Place Solution for MeViS Track in CVPR 2024 PVUW Workshop: Motion Expression guided Video Segmentation
di: Gao, Mingqi, et al.
Pubblicazione: (2024)
di: Gao, Mingqi, et al.
Pubblicazione: (2024)
Vgent: Graph-based Retrieval-Reasoning-Augmented Generation For Long Video Understanding
di: Shen, Xiaoqian, et al.
Pubblicazione: (2025)
di: Shen, Xiaoqian, et al.
Pubblicazione: (2025)
The Solution for the CVPR2024 NICE Image Captioning Challenge
di: Huang, Longfei, et al.
Pubblicazione: (2024)
di: Huang, Longfei, et al.
Pubblicazione: (2024)
LongVidSearch: An Agentic Benchmark for Multi-hop Evidence Retrieval Planning in Long Videos
di: Yu, Rongyi, et al.
Pubblicazione: (2026)
di: Yu, Rongyi, et al.
Pubblicazione: (2026)
DIVE: Deep-search Iterative Video Exploration A Technical Report for the CVRR Challenge at CVPR 2025
di: Kamoto, Umihiro, et al.
Pubblicazione: (2025)
di: Kamoto, Umihiro, et al.
Pubblicazione: (2025)
Resolving Evidence Sparsity: Agentic Context Engineering for Long-Document Understanding
di: Liu, Keliang, et al.
Pubblicazione: (2025)
di: Liu, Keliang, et al.
Pubblicazione: (2025)
Technical Report for the 5th CLVision Challenge at CVPR: Addressing the Class-Incremental with Repetition using Unlabeled Data -- 4th Place Solution
di: Moraiti, Panagiota, et al.
Pubblicazione: (2025)
di: Moraiti, Panagiota, et al.
Pubblicazione: (2025)
Long Video Understanding with Learnable Retrieval in Video-Language Models
di: Xu, Jiaqi, et al.
Pubblicazione: (2023)
di: Xu, Jiaqi, et al.
Pubblicazione: (2023)
DrVideo: Document Retrieval Based Long Video Understanding
di: Ma, Ziyu, et al.
Pubblicazione: (2024)
di: Ma, Ziyu, et al.
Pubblicazione: (2024)
EgoAction: Egocentric Action Composition with Reliability-Aware Temporal Fusion for the EPIC-KITCHENS Action Detection Challenge at CVPR 2026
di: Fu, Zhiheng, et al.
Pubblicazione: (2026)
di: Fu, Zhiheng, et al.
Pubblicazione: (2026)
3rd Place Solution to Large-scale Fine-grained Food Recognition
di: Zhong, Yang, et al.
Pubblicazione: (2025)
di: Zhong, Yang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
3rd Place Solution for MOSE Track in CVPR 2024 PVUW workshop: Complex Video Object Segmentation
di: Liu, Xinyu, et al.
Pubblicazione: (2024) -
CuriosAI Submission to the CASTLE Challenge at EgoVis 2026
di: Kanda, Yuto, et al.
Pubblicazione: (2026) -
3rd Place Solution for MeViS Track in CVPR 2024 PVUW workshop: Motion Expression guided Video Segmentation
di: Pan, Feiyu, et al.
Pubblicazione: (2024) -
MARS: Technical Report for the CASTLE Challenge at EgoVis 2026
di: Zhang, Haoyu, et al.
Pubblicazione: (2026) -
CASTLE2026 Team WDL Technical Report
di: Li, Zhengyang, et al.
Pubblicazione: (2026)