Saved in:
| Main Authors: | Chen, Shimin, Li, Wei, Chu, Jiaming, Chen, Chen, Zhang, Chen, Guo, Yandong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2411.00881 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Technical Report for ActivityNet Challenge 2022 -- Temporal Action Localization
by: Chen, Shimin, et al.
Published: (2024)
by: Chen, Shimin, et al.
Published: (2024)
SoccerNet 2023 Challenges Results
by: Cioppa, Anthony, et al.
Published: (2023)
by: Cioppa, Anthony, et al.
Published: (2023)
SoccerNet 2024 Challenges Results
by: Cioppa, Anthony, et al.
Published: (2024)
by: Cioppa, Anthony, et al.
Published: (2024)
SoccerNet 2025 Challenges Results
by: Giancola, Silvio, et al.
Published: (2025)
by: Giancola, Silvio, et al.
Published: (2025)
SoccerNet-Caption: Dense Video Captioning for Soccer Broadcasts Commentaries
by: Mkhallati, Hassan, et al.
Published: (2023)
by: Mkhallati, Hassan, et al.
Published: (2023)
SoccerNet-Tracking: Multiple Object Tracking Dataset and Benchmark in Soccer Videos
by: Cioppa, Anthony, et al.
Published: (2022)
by: Cioppa, Anthony, et al.
Published: (2022)
MM-SEAL: A Large-scale Video Dataset of Multi-person Multi-grained Spatio-temporally Action Localization
by: Chen, Shimin, et al.
Published: (2022)
by: Chen, Shimin, et al.
Published: (2022)
SoccerNet-v3D: Leveraging Sports Broadcast Replays for 3D Scene Understanding
by: Gutiérrez-Pérez, Marc, et al.
Published: (2025)
by: Gutiérrez-Pérez, Marc, et al.
Published: (2025)
Action Anticipation from SoccerNet Football Video Broadcasts
by: Dalal, Mohamad, et al.
Published: (2025)
by: Dalal, Mohamad, et al.
Published: (2025)
From Broadcast to Minimap: Achieving State-of-the-Art SoccerNet Game State Reconstruction
by: Golovkin, Vladimir, et al.
Published: (2025)
by: Golovkin, Vladimir, et al.
Published: (2025)
Technical Report for Soccernet 2023 -- Dense Video Captioning
by: Ruan, Zheng, et al.
Published: (2024)
by: Ruan, Zheng, et al.
Published: (2024)
SoccerNet Game State Reconstruction: End-to-End Athlete Tracking and Identification on a Minimap
by: Somers, Vladimir, et al.
Published: (2024)
by: Somers, Vladimir, et al.
Published: (2024)
Deep Understanding of Soccer Match Videos
by: Xu, Shikun, et al.
Published: (2024)
by: Xu, Shikun, et al.
Published: (2024)
A Vanilla Multi-Task Framework for Dense Visual Prediction Solution to 1st VCL Challenge -- Multi-Task Robustness Track
by: Chen, Zehui, et al.
Published: (2024)
by: Chen, Zehui, et al.
Published: (2024)
3DWG: 3D Weakly Supervised Visual Grounding via Category and Instance-Level Alignment
by: Li, Xiaoqi, et al.
Published: (2025)
by: Li, Xiaoqi, et al.
Published: (2025)
BEVUDA: Multi-geometric Space Alignments for Domain Adaptive BEV 3D Object Detection
by: Liu, Jiaming, et al.
Published: (2022)
by: Liu, Jiaming, et al.
Published: (2022)
An Efficient and Effective Transformer Decoder-Based Framework for Multi-Task Visual Grounding
by: Chen, Wei, et al.
Published: (2024)
by: Chen, Wei, et al.
Published: (2024)
Logics-Parsing Technical Report
by: Chen, Xiangyang, et al.
Published: (2025)
by: Chen, Xiangyang, et al.
Published: (2025)
Octopus v3: Technical Report for On-device Sub-billion Multimodal AI Agent
by: Chen, Wei, et al.
Published: (2024)
by: Chen, Wei, et al.
Published: (2024)
Continual-MAE: Adaptive Distribution Masked Autoencoders for Continual Test-Time Adaptation
by: Liu, Jiaming, et al.
Published: (2023)
by: Liu, Jiaming, et al.
Published: (2023)
Enhancing Soccer Camera Calibration Through Keypoint Exploitation
by: Falaleev, Nikolay S., et al.
Published: (2024)
by: Falaleev, Nikolay S., et al.
Published: (2024)
Kwai Keye-VL Technical Report
by: Kwai Keye Team, et al.
Published: (2025)
by: Kwai Keye Team, et al.
Published: (2025)
Ovis-Image Technical Report
by: Wang, Guo-Hua, et al.
Published: (2025)
by: Wang, Guo-Hua, et al.
Published: (2025)
Ovis-U1 Technical Report
by: Wang, Guo-Hua, et al.
Published: (2025)
by: Wang, Guo-Hua, et al.
Published: (2025)
Technical Report: Competition Solution For Modelscope-Sora
by: Chen, Shengfu, et al.
Published: (2024)
by: Chen, Shengfu, et al.
Published: (2024)
Replay-Free Continual Low-Rank Adaptation with Dynamic Memory
by: Chen, Huancheng, et al.
Published: (2024)
by: Chen, Huancheng, et al.
Published: (2024)
Qwen-Image Technical Report
by: Wu, Chenfei, et al.
Published: (2025)
by: Wu, Chenfei, et al.
Published: (2025)
NTO3D: Neural Target Object 3D Reconstruction with Segment Anything
by: Wei, Xiaobao, et al.
Published: (2023)
by: Wei, Xiaobao, et al.
Published: (2023)
ReplayCAD: Generative Diffusion Replay for Continual Anomaly Detection
by: Hu, Lei, et al.
Published: (2025)
by: Hu, Lei, et al.
Published: (2025)
SoccerLens: Grounded Soccer Video Understanding Beyond Accuracy
by: Elsharkawi, Ismael, et al.
Published: (2026)
by: Elsharkawi, Ismael, et al.
Published: (2026)
StreamingClaw Technical Report
by: Chen, Jiawei, et al.
Published: (2026)
by: Chen, Jiawei, et al.
Published: (2026)
Kwai Keye-VL 1.5 Technical Report
by: Yang, Biao, et al.
Published: (2025)
by: Yang, Biao, et al.
Published: (2025)
Kimi-VL Technical Report
by: Kimi Team, et al.
Published: (2025)
by: Kimi Team, et al.
Published: (2025)
Technical Report for Argoverse2 Scenario Mining Challenges on Iterative Error Correction and Spatially-Aware Prompting
by: Chen, Yifei, et al.
Published: (2025)
by: Chen, Yifei, et al.
Published: (2025)
MARS: Technical Report for the CASTLE Challenge at EgoVis 2026
by: Zhang, Haoyu, et al.
Published: (2026)
by: Zhang, Haoyu, et al.
Published: (2026)
RefChartQA: Grounding Visual Answer on Chart Images through Instruction Tuning
by: Vogel, Alexander, et al.
Published: (2025)
by: Vogel, Alexander, et al.
Published: (2025)
Seeing Beyond Classes: Zero-Shot Grounded Situation Recognition via Language Explainer
by: Lei, Jiaming, et al.
Published: (2024)
by: Lei, Jiaming, et al.
Published: (2024)
JFAA: Technical Report for the EPIC-KITCHENS-100 Action Anticipation Challenge at EgoVis 2026
by: Chu, Qiaohui, et al.
Published: (2026)
by: Chu, Qiaohui, et al.
Published: (2026)
Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025
by: Ying, Zonghao, et al.
Published: (2025)
by: Ying, Zonghao, et al.
Published: (2025)
Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model
by: Ma, Guoqing, et al.
Published: (2025)
by: Ma, Guoqing, et al.
Published: (2025)
Similar Items
-
Technical Report for ActivityNet Challenge 2022 -- Temporal Action Localization
by: Chen, Shimin, et al.
Published: (2024) -
SoccerNet 2023 Challenges Results
by: Cioppa, Anthony, et al.
Published: (2023) -
SoccerNet 2024 Challenges Results
by: Cioppa, Anthony, et al.
Published: (2024) -
SoccerNet 2025 Challenges Results
by: Giancola, Silvio, et al.
Published: (2025) -
SoccerNet-Caption: Dense Video Captioning for Soccer Broadcasts Commentaries
by: Mkhallati, Hassan, et al.
Published: (2023)