3rd Place at CVPR 2026 CASTLE Challenge: Agentic Multi-View Long-Context Video Understanding via Hierarchical Knowledge Graph Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Albusayes, Raghad, Alyahya, Munirah |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
3rd Place Solution for MOSE Track in CVPR 2024 PVUW workshop: Complex Video Object Segmentation
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
CuriosAI Submission to the CASTLE Challenge at EgoVis 2026
von: Kanda, Yuto, et al.
Veröffentlicht: (2026)
von: Kanda, Yuto, et al.
Veröffentlicht: (2026)
3rd Place Solution for MeViS Track in CVPR 2024 PVUW workshop: Motion Expression guided Video Segmentation
von: Pan, Feiyu, et al.
Veröffentlicht: (2024)
von: Pan, Feiyu, et al.
Veröffentlicht: (2024)
MARS: Technical Report for the CASTLE Challenge at EgoVis 2026
von: Zhang, Haoyu, et al.
Veröffentlicht: (2026)
von: Zhang, Haoyu, et al.
Veröffentlicht: (2026)
CASTLE2026 Team WDL Technical Report
von: Li, Zhengyang, et al.
Veröffentlicht: (2026)
von: Li, Zhengyang, et al.
Veröffentlicht: (2026)
3rd Place Solution for PVUW Challenge 2024: Video Panoptic Segmentation
von: Wu, Ruipu, et al.
Veröffentlicht: (2024)
von: Wu, Ruipu, et al.
Veröffentlicht: (2024)
1st Place Winner of the 2024 Pixel-level Video Understanding in the Wild (CVPR'24 PVUW) Challenge in Video Panoptic Segmentation and Best Long Video Consistency of Video Semantic Segmentation
von: Liu, Qingfeng, et al.
Veröffentlicht: (2024)
von: Liu, Qingfeng, et al.
Veröffentlicht: (2024)
Visual Agentic Memory: Enabling Online Long Video Understanding via Online Indexing, Hierarchical Memory, and Agentic Retrieval
von: Li, Aiden Yiliu, et al.
Veröffentlicht: (2026)
von: Li, Aiden Yiliu, et al.
Veröffentlicht: (2026)
X-Restormer++: 1st Place Solution for the UG2+ CVPR 2026 All-Weather Restoration Challenge
von: Pan, Youwei, et al.
Veröffentlicht: (2026)
von: Pan, Youwei, et al.
Veröffentlicht: (2026)
TempRet: Temporal Enhancement and Two-Stage Reranking for CVPR 2026 EPIC-KITCHENS-100 Multi-Instance Retrieval Challenge
von: Li, Zixu, et al.
Veröffentlicht: (2026)
von: Li, Zixu, et al.
Veröffentlicht: (2026)
SAMSON: 3rd Place Solution of LSVOS 2025 VOS Challenge
von: Xie, Yujie, et al.
Veröffentlicht: (2025)
von: Xie, Yujie, et al.
Veröffentlicht: (2025)
3rd Place Solution to ICCV LargeFineFoodAI Retrieval
von: Zhong, Yang, et al.
Veröffentlicht: (2025)
von: Zhong, Yang, et al.
Veröffentlicht: (2025)
EgoAdapt: A Multi-Scene Egocentric Adaptation Method for CVPR 2026 HD-EPIC VQA Challenge
von: Chen, Zhiwei, et al.
Veröffentlicht: (2026)
von: Chen, Zhiwei, et al.
Veröffentlicht: (2026)
The Instance-centric Transformer for the RVOS Track of LSVOS Challenge: 3rd Place Solution
von: Cao, Bin, et al.
Veröffentlicht: (2024)
von: Cao, Bin, et al.
Veröffentlicht: (2024)
FVOS for MOSE Track of 4th PVUW Challenge: 3rd Place Solution
von: Wang, Mengjiao, et al.
Veröffentlicht: (2025)
von: Wang, Mengjiao, et al.
Veröffentlicht: (2025)
VideoARM: Agentic Reasoning over Hierarchical Memory for Long-Form Video Understanding
von: Yin, Yufei, et al.
Veröffentlicht: (2025)
von: Yin, Yufei, et al.
Veröffentlicht: (2025)
The CASTLE 2024 Dataset: Advancing the Art of Multimodal Understanding
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)
von: Rossetto, Luca, et al.
Veröffentlicht: (2025)
1st Place Solution for MOSE Track in CVPR 2024 PVUW Workshop: Complex Video Object Segmentation
von: Miao, Deshui, et al.
Veröffentlicht: (2024)
von: Miao, Deshui, et al.
Veröffentlicht: (2024)
2nd Place Solution for MOSE Track in CVPR 2024 PVUW workshop: Complex Video Object Segmentation
von: Xu, Zhensong, et al.
Veröffentlicht: (2024)
von: Xu, Zhensong, et al.
Veröffentlicht: (2024)
Hierarchical Long Video Understanding with Audiovisual Entity Cohesion and Agentic Search
von: Yin, Xinlei, et al.
Veröffentlicht: (2026)
von: Yin, Xinlei, et al.
Veröffentlicht: (2026)
An Effective Solution for the CVPR 2026 8th UG2+ Challenge Track 3: Dynamic Object Segmentation in Turbulence
von: Li, Hongzhen, et al.
Veröffentlicht: (2026)
von: Li, Hongzhen, et al.
Veröffentlicht: (2026)
Multi-View Hierarchical Graph Neural Network for Sketch-Based 3D Shape Retrieval
von: Cheng, Hang, et al.
Veröffentlicht: (2026)
von: Cheng, Hang, et al.
Veröffentlicht: (2026)
LSVOS Challenge 3rd Place Report: SAM2 and Cutie based VOS
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
Winner of CVPR2026 NTIRE Challenge on Image Shadow Removal: Semantic and Geometric Guidance for Shadow Removal via Cascaded Refinement
von: Beltrame, Lorenzo, et al.
Veröffentlicht: (2026)
von: Beltrame, Lorenzo, et al.
Veröffentlicht: (2026)
Technique Report of CVPR 2024 PBDL Challenges
von: Fu, Ying, et al.
Veröffentlicht: (2024)
von: Fu, Ying, et al.
Veröffentlicht: (2024)
3rd Place Solution for VisDA 2021 Challenge -- Universally Domain Adaptive Image Recognition
von: Liao, Haojin, et al.
Veröffentlicht: (2021)
von: Liao, Haojin, et al.
Veröffentlicht: (2021)
Agentic Very Long Video Understanding
von: Rege, Aniket, et al.
Veröffentlicht: (2026)
von: Rege, Aniket, et al.
Veröffentlicht: (2026)
A Robust Semantic Segmentation Pipeline for the CVPR 2026 8th UG2+ Challenge Track 2
von: Chai, Jinming, et al.
Veröffentlicht: (2026)
von: Chai, Jinming, et al.
Veröffentlicht: (2026)
2nd Place Solution for MeViS Track in CVPR 2024 PVUW Workshop: Motion Expression guided Video Segmentation
von: Cao, Bin, et al.
Veröffentlicht: (2024)
von: Cao, Bin, et al.
Veröffentlicht: (2024)
1st Place Solution for MeViS Track in CVPR 2024 PVUW Workshop: Motion Expression guided Video Segmentation
von: Gao, Mingqi, et al.
Veröffentlicht: (2024)
von: Gao, Mingqi, et al.
Veröffentlicht: (2024)
Vgent: Graph-based Retrieval-Reasoning-Augmented Generation For Long Video Understanding
von: Shen, Xiaoqian, et al.
Veröffentlicht: (2025)
von: Shen, Xiaoqian, et al.
Veröffentlicht: (2025)
The Solution for the CVPR2024 NICE Image Captioning Challenge
von: Huang, Longfei, et al.
Veröffentlicht: (2024)
von: Huang, Longfei, et al.
Veröffentlicht: (2024)
LongVidSearch: An Agentic Benchmark for Multi-hop Evidence Retrieval Planning in Long Videos
von: Yu, Rongyi, et al.
Veröffentlicht: (2026)
von: Yu, Rongyi, et al.
Veröffentlicht: (2026)
DIVE: Deep-search Iterative Video Exploration A Technical Report for the CVRR Challenge at CVPR 2025
von: Kamoto, Umihiro, et al.
Veröffentlicht: (2025)
von: Kamoto, Umihiro, et al.
Veröffentlicht: (2025)
Resolving Evidence Sparsity: Agentic Context Engineering for Long-Document Understanding
von: Liu, Keliang, et al.
Veröffentlicht: (2025)
von: Liu, Keliang, et al.
Veröffentlicht: (2025)
Technical Report for the 5th CLVision Challenge at CVPR: Addressing the Class-Incremental with Repetition using Unlabeled Data -- 4th Place Solution
von: Moraiti, Panagiota, et al.
Veröffentlicht: (2025)
von: Moraiti, Panagiota, et al.
Veröffentlicht: (2025)
Long Video Understanding with Learnable Retrieval in Video-Language Models
von: Xu, Jiaqi, et al.
Veröffentlicht: (2023)
von: Xu, Jiaqi, et al.
Veröffentlicht: (2023)
DrVideo: Document Retrieval Based Long Video Understanding
von: Ma, Ziyu, et al.
Veröffentlicht: (2024)
von: Ma, Ziyu, et al.
Veröffentlicht: (2024)
EgoAction: Egocentric Action Composition with Reliability-Aware Temporal Fusion for the EPIC-KITCHENS Action Detection Challenge at CVPR 2026
von: Fu, Zhiheng, et al.
Veröffentlicht: (2026)
von: Fu, Zhiheng, et al.
Veröffentlicht: (2026)
3rd Place Solution to Large-scale Fine-grained Food Recognition
von: Zhong, Yang, et al.
Veröffentlicht: (2025)
von: Zhong, Yang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
3rd Place Solution for MOSE Track in CVPR 2024 PVUW workshop: Complex Video Object Segmentation
von: Liu, Xinyu, et al.
Veröffentlicht: (2024) -
CuriosAI Submission to the CASTLE Challenge at EgoVis 2026
von: Kanda, Yuto, et al.
Veröffentlicht: (2026) -
3rd Place Solution for MeViS Track in CVPR 2024 PVUW workshop: Motion Expression guided Video Segmentation
von: Pan, Feiyu, et al.
Veröffentlicht: (2024) -
MARS: Technical Report for the CASTLE Challenge at EgoVis 2026
von: Zhang, Haoyu, et al.
Veröffentlicht: (2026) -
CASTLE2026 Team WDL Technical Report
von: Li, Zhengyang, et al.
Veröffentlicht: (2026)