CoSTL: Comprehensive Spatial-Temporal Representation Learning for Moment Retrieval and Highlight Detection
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Dong, Xin, Geng, Wenjia, Deng, Wenfeng, Tang, Yansong |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Localization-Aware Multi-Scale Representation Learning for Repetitive Action Counting
par: Wang, Sujia, et autres
Publié: (2025)
par: Wang, Sujia, et autres
Publié: (2025)
Boosting Zero-Shot 3D Style Transfer with 2D Pre-trained Priors
par: Dong, Xin, et autres
Publié: (2026)
par: Dong, Xin, et autres
Publié: (2026)
Saliency-Guided DETR for Moment Retrieval and Highlight Detection
par: Gordeev, Aleksandr, et autres
Publié: (2024)
par: Gordeev, Aleksandr, et autres
Publié: (2024)
DiffusionVMR: Diffusion Model for Joint Video Moment Retrieval and Highlight Detection
par: Zhao, Henghao, et autres
Publié: (2023)
par: Zhao, Henghao, et autres
Publié: (2023)
MS-DETR: Towards Effective Video Moment Retrieval and Highlight Detection by Joint Motion-Semantic Learning
par: Ma, Hongxu, et autres
Publié: (2025)
par: Ma, Hongxu, et autres
Publié: (2025)
Multi-modal Fusion and Query Refinement Network for Video Moment Retrieval and Highlight Detection
par: Xu, Yifang, et autres
Publié: (2025)
par: Xu, Yifang, et autres
Publié: (2025)
TR-DETR: Task-Reciprocal Transformer for Joint Moment Retrieval and Highlight Detection
par: Sun, Hao, et autres
Publié: (2024)
par: Sun, Hao, et autres
Publié: (2024)
Moment and Highlight Detection via MLLM Frame Segmentation
par: Jiwanta, I Putu Andika Bagas, et autres
Publié: (2025)
par: Jiwanta, I Putu Andika Bagas, et autres
Publié: (2025)
Lighthouse: A User-Friendly Library for Reproducible Video Moment Retrieval and Highlight Detection
par: Nishimura, Taichi, et autres
Publié: (2024)
par: Nishimura, Taichi, et autres
Publié: (2024)
GPTSee: Enhancing Moment Retrieval and Highlight Detection via Description-Based Similarity Features
par: Sun, Yunzhuo, et autres
Publié: (2024)
par: Sun, Yunzhuo, et autres
Publié: (2024)
UniMD: Towards Unifying Moment Retrieval and Temporal Action Detection
par: Zeng, Yingsen, et autres
Publié: (2024)
par: Zeng, Yingsen, et autres
Publié: (2024)
Adaptive Evidential Learning for Temporal-Semantic Robustness in Moment Retrieval
par: Huang, Haojian, et autres
Publié: (2025)
par: Huang, Haojian, et autres
Publié: (2025)
See, Rank, and Filter: Important Word-Aware Clip Filtering via Scene Understanding for Moment Retrieval and Highlight Detection
par: Lee, YuEun, et autres
Publié: (2025)
par: Lee, YuEun, et autres
Publié: (2025)
Task-Driven Exploration: Decoupling and Inter-Task Feedback for Joint Moment Retrieval and Highlight Detection
par: Yang, Jin, et autres
Publié: (2024)
par: Yang, Jin, et autres
Publié: (2024)
Watch Video, Catch Keyword: Context-aware Keyword Attention for Moment Retrieval and Highlight Detection
par: Um, Sung Jin, et autres
Publié: (2025)
par: Um, Sung Jin, et autres
Publié: (2025)
Background-aware Moment Detection for Video Moment Retrieval
par: Jung, Minjoon, et autres
Publié: (2023)
par: Jung, Minjoon, et autres
Publié: (2023)
Skeleton-Guided Spatial-Temporal Feature Learning for Video-Based Visible-Infrared Person Re-Identification
par: Jiang, Wenjia, et autres
Publié: (2024)
par: Jiang, Wenjia, et autres
Publié: (2024)
Augmenting Moment Retrieval: Zero-Dependency Two-Stage Learning
par: Wei, Zhengxuan, et autres
Publié: (2025)
par: Wei, Zhengxuan, et autres
Publié: (2025)
LD-DETR: Loop Decoder DEtection TRansformer for Video Moment Retrieval and Highlight Detection
par: Zhao, Pengcheng, et autres
Publié: (2025)
par: Zhao, Pengcheng, et autres
Publié: (2025)
VideoLights: Feature Refinement and Cross-Task Alignment Transformer for Joint Video Highlight Detection and Moment Retrieval
par: Paul, Dhiman, et autres
Publié: (2024)
par: Paul, Dhiman, et autres
Publié: (2024)
Unsupervised Modality-Transferable Video Highlight Detection with Representation Activation Sequence Learning
par: Li, Tingtian, et autres
Publié: (2024)
par: Li, Tingtian, et autres
Publié: (2024)
Coarse Correspondences Boost Spatial-Temporal Reasoning in Multimodal Language Model
par: Liu, Benlin, et autres
Publié: (2024)
par: Liu, Benlin, et autres
Publié: (2024)
Described Spatial-Temporal Video Detection
par: Ji, Wei, et autres
Publié: (2024)
par: Ji, Wei, et autres
Publié: (2024)
Neural Spatial-Temporal Tensor Representation for Infrared Small Target Detection
par: Wu, Fengyi, et autres
Publié: (2024)
par: Wu, Fengyi, et autres
Publié: (2024)
MomentMix Augmentation with Length-Aware DETR for Temporally Robust Moment Retrieval
par: Park, Seojeong, et autres
Publié: (2024)
par: Park, Seojeong, et autres
Publié: (2024)
3SHNet: Boosting Image-Sentence Retrieval via Visual Semantic-Spatial Self-Highlighting
par: Ge, Xuri, et autres
Publié: (2024)
par: Ge, Xuri, et autres
Publié: (2024)
FDDet: Achieving Data-Efficient Food Defect Detection Under Real-World Scenarios
par: Xu, Ruihao, et autres
Publié: (2026)
par: Xu, Ruihao, et autres
Publié: (2026)
When One Moment Isn't Enough: Multi-Moment Retrieval with Cross-Moment Interactions
par: Cao, Zhuo, et autres
Publié: (2025)
par: Cao, Zhuo, et autres
Publié: (2025)
HighlightMe: Detecting Highlights from Human-Centric Videos
par: Bhattacharya, Uttaran, et autres
Publié: (2021)
par: Bhattacharya, Uttaran, et autres
Publié: (2021)
Moment of Untruth: Dealing with Negative Queries in Video Moment Retrieval
par: Flanagan, Kevin, et autres
Publié: (2025)
par: Flanagan, Kevin, et autres
Publié: (2025)
SRC-Net: Bi-Temporal Spatial Relationship Concerned Network for Change Detection
par: Chen, Hongjia, et autres
Publié: (2024)
par: Chen, Hongjia, et autres
Publié: (2024)
Retrieving Any Relevant Moments: Benchmark and Models for Generalized Moment Retrieval
par: Ding, Yiming, et autres
Publié: (2026)
par: Ding, Yiming, et autres
Publié: (2026)
Self-Supervised Representation Learning with Spatial-Temporal Consistency for Sign Language Recognition
par: Zhao, Weichao, et autres
Publié: (2024)
par: Zhao, Weichao, et autres
Publié: (2024)
The Devil is in the Spurious Correlations: Boosting Moment Retrieval with Dynamic Learning
par: Zhou, Xinyang, et autres
Publié: (2025)
par: Zhou, Xinyang, et autres
Publié: (2025)
A Lightweight Moment Retrieval System with Global Re-Ranking and Robust Adaptive Bidirectional Temporal Search
par: Nguyen-Nhu, Tinh-Anh, et autres
Publié: (2025)
par: Nguyen-Nhu, Tinh-Anh, et autres
Publié: (2025)
CREM: Compression-Driven Representation Enhancement for Multimodal Retrieval and Comprehension
par: Liu, Lihao, et autres
Publié: (2026)
par: Liu, Lihao, et autres
Publié: (2026)
Jointly Learning Spatial, Angular, and Temporal Information for Enhanced Lane Detection
par: Alam, Muhammad Zeshan
Publié: (2024)
par: Alam, Muhammad Zeshan
Publié: (2024)
Moment Quantization for Video Temporal Grounding
par: Sun, Xiaolong, et autres
Publié: (2025)
par: Sun, Xiaolong, et autres
Publié: (2025)
Learning Dual-Level Deformable Implicit Representation for Real-World Scale Arbitrary Super-Resolution
par: Li, Zhiheng, et autres
Publié: (2024)
par: Li, Zhiheng, et autres
Publié: (2024)
Hybrid-Learning Video Moment Retrieval across Multi-Domain Labels
par: Cai, Weitong, et autres
Publié: (2024)
par: Cai, Weitong, et autres
Publié: (2024)
Documents similaires
-
Localization-Aware Multi-Scale Representation Learning for Repetitive Action Counting
par: Wang, Sujia, et autres
Publié: (2025) -
Boosting Zero-Shot 3D Style Transfer with 2D Pre-trained Priors
par: Dong, Xin, et autres
Publié: (2026) -
Saliency-Guided DETR for Moment Retrieval and Highlight Detection
par: Gordeev, Aleksandr, et autres
Publié: (2024) -
DiffusionVMR: Diffusion Model for Joint Video Moment Retrieval and Highlight Detection
par: Zhao, Henghao, et autres
Publié: (2023) -
MS-DETR: Towards Effective Video Moment Retrieval and Highlight Detection by Joint Motion-Semantic Learning
par: Ma, Hongxu, et autres
Publié: (2025)