Hybrid-Learning Video Moment Retrieval across Multi-Domain Labels
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cai, Weitong, Huang, Jiabo, Gong, Shaogang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MLLM as Video Narrator: Mitigating Modality Imbalance in Video Moment Retrieval
von: Cai, Weitong, et al.
Veröffentlicht: (2024)
von: Cai, Weitong, et al.
Veröffentlicht: (2024)
Generative Video Diffusion for Unseen Novel Semantic Video Moment Retrieval
von: Luo, Dezhao, et al.
Veröffentlicht: (2024)
von: Luo, Dezhao, et al.
Veröffentlicht: (2024)
Neuro-Symbolic Spatial Reasoning in Segmentation
von: Lin, Jiayi, et al.
Veröffentlicht: (2025)
von: Lin, Jiayi, et al.
Veröffentlicht: (2025)
InvSeg: Test-Time Prompt Inversion for Semantic Segmentation
von: Lin, Jiayi, et al.
Veröffentlicht: (2024)
von: Lin, Jiayi, et al.
Veröffentlicht: (2024)
Grounding Video Reasoning in Physical Signals
von: Osmanli, Alibay, et al.
Veröffentlicht: (2026)
von: Osmanli, Alibay, et al.
Veröffentlicht: (2026)
Training-free Zero-shot Composed Image Retrieval with Local Concept Reranking
von: Sun, Shitong, et al.
Veröffentlicht: (2023)
von: Sun, Shitong, et al.
Veröffentlicht: (2023)
Background-aware Moment Detection for Video Moment Retrieval
von: Jung, Minjoon, et al.
Veröffentlicht: (2023)
von: Jung, Minjoon, et al.
Veröffentlicht: (2023)
Few-Shot Image Generation by Conditional Relaxing Diffusion Inversion
von: Cao, Yu, et al.
Veröffentlicht: (2024)
von: Cao, Yu, et al.
Veröffentlicht: (2024)
Moment of Untruth: Dealing with Negative Queries in Video Moment Retrieval
von: Flanagan, Kevin, et al.
Veröffentlicht: (2025)
von: Flanagan, Kevin, et al.
Veröffentlicht: (2025)
Denoise-then-Retrieve: Text-Conditioned Video Denoising for Video Moment Retrieval
von: Liu, Weijia, et al.
Veröffentlicht: (2025)
von: Liu, Weijia, et al.
Veröffentlicht: (2025)
ViSMaP: Unsupervised Hour-long Video Summarisation by Meta-Prompting
von: Hu, Jian, et al.
Veröffentlicht: (2025)
von: Hu, Jian, et al.
Veröffentlicht: (2025)
CoS: Chain-of-Shot Prompting for Long Video Understanding
von: Hu, Jian, et al.
Veröffentlicht: (2025)
von: Hu, Jian, et al.
Veröffentlicht: (2025)
EgoGraph: Temporal Knowledge Graph for Egocentric Video Understanding
von: Sun, Shitong, et al.
Veröffentlicht: (2026)
von: Sun, Shitong, et al.
Veröffentlicht: (2026)
Multi-Modal Cross-Domain Alignment Network for Video Moment Retrieval
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
von: Fang, Xiang, et al.
Veröffentlicht: (2022)
Object-Centric Framework for Video Moment Retrieval
von: Li, Zongyao, et al.
Veröffentlicht: (2025)
von: Li, Zongyao, et al.
Veröffentlicht: (2025)
V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
von: Cheng, Zixu, et al.
Veröffentlicht: (2025)
von: Cheng, Zixu, et al.
Veröffentlicht: (2025)
Adaptive Domain Shift in Diffusion Models for Cross-Modality Image Translation
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
von: Wang, Zihao, et al.
Veröffentlicht: (2026)
MaRI: Material Retrieval Integration across Domains
von: Wang, Jianhui, et al.
Veröffentlicht: (2025)
von: Wang, Jianhui, et al.
Veröffentlicht: (2025)
Multi-modal Fusion and Query Refinement Network for Video Moment Retrieval and Highlight Detection
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
von: Xu, Yifang, et al.
Veröffentlicht: (2025)
Beyond Caption-Based Queries for Video Moment Retrieval
von: Pujol-Perich, David, et al.
Veröffentlicht: (2026)
von: Pujol-Perich, David, et al.
Veröffentlicht: (2026)
INT: Instance-Specific Negative Mining for Task-Generic Promptable Segmentation
von: Hu, Jian, et al.
Veröffentlicht: (2025)
von: Hu, Jian, et al.
Veröffentlicht: (2025)
Multi-proposal Collaboration and Multi-task Training for Weakly-supervised Video Moment Retrieval
von: Zhang, Bolin, et al.
Veröffentlicht: (2026)
von: Zhang, Bolin, et al.
Veröffentlicht: (2026)
Vid-Morp: Video Moment Retrieval Pretraining from Unlabeled Videos in the Wild
von: Bao, Peijun, et al.
Veröffentlicht: (2024)
von: Bao, Peijun, et al.
Veröffentlicht: (2024)
XFMamba: Cross-Fusion Mamba for Multi-View Medical Image Classification
von: Zheng, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Zheng, Xiaoyu, et al.
Veröffentlicht: (2025)
Selective Query-guided Debiasing for Video Corpus Moment Retrieval
von: Yoon, Sunjae, et al.
Veröffentlicht: (2022)
von: Yoon, Sunjae, et al.
Veröffentlicht: (2022)
HybridMamba: A Dual-domain Mamba for 3D Medical Image Segmentation
von: Wu, Weitong, et al.
Veröffentlicht: (2025)
von: Wu, Weitong, et al.
Veröffentlicht: (2025)
Source-free Video Domain Adaptation by Learning from Noisy Labels
von: Dasgupta, Avijit, et al.
Veröffentlicht: (2023)
von: Dasgupta, Avijit, et al.
Veröffentlicht: (2023)
Towards Efficient Partially Relevant Video Retrieval with Active Moment Discovering
von: Song, Peipei, et al.
Veröffentlicht: (2025)
von: Song, Peipei, et al.
Veröffentlicht: (2025)
Adaptive Evidential Learning for Temporal-Semantic Robustness in Moment Retrieval
von: Huang, Haojian, et al.
Veröffentlicht: (2025)
von: Huang, Haojian, et al.
Veröffentlicht: (2025)
Text-Video Multi-Grained Integration for Video Moment Montage
von: Yin, Zhihui, et al.
Veröffentlicht: (2024)
von: Yin, Zhihui, et al.
Veröffentlicht: (2024)
Event-aware Video Corpus Moment Retrieval
von: Hou, Danyang, et al.
Veröffentlicht: (2024)
von: Hou, Danyang, et al.
Veröffentlicht: (2024)
GIRL-DETR: Gradient-Isolated Reinforcement Learning for Video Moment Retrieval
von: Zhang, Shihang, et al.
Veröffentlicht: (2026)
von: Zhang, Shihang, et al.
Veröffentlicht: (2026)
MomentSeeker: A Task-Oriented Benchmark For Long-Video Moment Retrieval
von: Yuan, Huaying, et al.
Veröffentlicht: (2025)
von: Yuan, Huaying, et al.
Veröffentlicht: (2025)
MS-DETR: Towards Effective Video Moment Retrieval and Highlight Detection by Joint Motion-Semantic Learning
von: Ma, Hongxu, et al.
Veröffentlicht: (2025)
von: Ma, Hongxu, et al.
Veröffentlicht: (2025)
Context-Enhanced Video Moment Retrieval with Large Language Models
von: Liu, Weijia, et al.
Veröffentlicht: (2024)
von: Liu, Weijia, et al.
Veröffentlicht: (2024)
Leveraging Hallucinations to Reduce Manual Prompt Dependency in Promptable Segmentation
von: Hu, Jian, et al.
Veröffentlicht: (2024)
von: Hu, Jian, et al.
Veröffentlicht: (2024)
Enhancing Zero-Shot Facial Expression Recognition by LLM Knowledge Transfer
von: Zhao, Zengqun, et al.
Veröffentlicht: (2024)
von: Zhao, Zengqun, et al.
Veröffentlicht: (2024)
GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking
von: Cheng, Zixu, et al.
Veröffentlicht: (2026)
von: Cheng, Zixu, et al.
Veröffentlicht: (2026)
Temporal Score Analysis for Understanding and Correcting Diffusion Artifacts
von: Cao, Yu, et al.
Veröffentlicht: (2025)
von: Cao, Yu, et al.
Veröffentlicht: (2025)
LLaVA-MR: Large Language-and-Vision Assistant for Video Moment Retrieval
von: Lu, Weiheng, et al.
Veröffentlicht: (2024)
von: Lu, Weiheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MLLM as Video Narrator: Mitigating Modality Imbalance in Video Moment Retrieval
von: Cai, Weitong, et al.
Veröffentlicht: (2024) -
Generative Video Diffusion for Unseen Novel Semantic Video Moment Retrieval
von: Luo, Dezhao, et al.
Veröffentlicht: (2024) -
Neuro-Symbolic Spatial Reasoning in Segmentation
von: Lin, Jiayi, et al.
Veröffentlicht: (2025) -
InvSeg: Test-Time Prompt Inversion for Semantic Segmentation
von: Lin, Jiayi, et al.
Veröffentlicht: (2024) -
Grounding Video Reasoning in Physical Signals
von: Osmanli, Alibay, et al.
Veröffentlicht: (2026)