Uneven Event Modeling for Partially Relevant Video Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Sa, Chen, Huashan, Zhang, Wanqian, Zhang, Jinchao, Yang, Zexian, Hao, Xiaoshuai, Li, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mitigating Semantic Collapse in Partially Relevant Video Retrieval
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
Knowledge-Refined Dual Context-Aware Network for Partially Relevant Video Retrieval
von: Yang, Junkai, et al.
Veröffentlicht: (2026)
von: Yang, Junkai, et al.
Veröffentlicht: (2026)
Ambiguity-Restrained Text-Video Representation Learning for Partially Relevant Video Retrieval
von: Cho, CH, et al.
Veröffentlicht: (2025)
von: Cho, CH, et al.
Veröffentlicht: (2025)
Prototypes are Balanced Units for Efficient and Effective Partially Relevant Video Retrieval
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
GMMFormer: Gaussian-Mixture-Model Based Transformer for Efficient Partially Relevant Video Retrieval
von: Wang, Yuting, et al.
Veröffentlicht: (2023)
von: Wang, Yuting, et al.
Veröffentlicht: (2023)
Denoise and Align: Diffusion-Driven Foreground Knowledge Prompting for Open-Vocabulary Temporal Action Detection
von: Zhu, Sa, et al.
Veröffentlicht: (2026)
von: Zhu, Sa, et al.
Veröffentlicht: (2026)
Decompose and Transfer: CoT-Prompting Enhanced Alignment for Open-Vocabulary Temporal Action Detection
von: Zhu, Sa, et al.
Veröffentlicht: (2026)
von: Zhu, Sa, et al.
Veröffentlicht: (2026)
ILDiff: Generate Transparent Animated Stickers by Implicit Layout Distillation
von: Zhang, Ting, et al.
Veröffentlicht: (2024)
von: Zhang, Ting, et al.
Veröffentlicht: (2024)
Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events
von: Liu, Xiaolin, et al.
Veröffentlicht: (2026)
von: Liu, Xiaolin, et al.
Veröffentlicht: (2026)
Scaling the Long Video Understanding of Multimodal Large Language Models via Visual Memory Mechanism
von: Chen, Tao, et al.
Veröffentlicht: (2026)
von: Chen, Tao, et al.
Veröffentlicht: (2026)
Event-VStream: Event-Driven Real-Time Understanding for Long Video Streams
von: Guo, Zhenghui, et al.
Veröffentlicht: (2026)
von: Guo, Zhenghui, et al.
Veröffentlicht: (2026)
Uneven Evolution of Cognition Across Generations of Generative AI Models
von: Galatzer-Levy, Isaac, et al.
Veröffentlicht: (2026)
von: Galatzer-Levy, Isaac, et al.
Veröffentlicht: (2026)
VERHallu: Evaluating and Mitigating Event Relation Hallucination in Video Large Language Models
von: Zhang, Zefan, et al.
Veröffentlicht: (2026)
von: Zhang, Zefan, et al.
Veröffentlicht: (2026)
VERIFIED: A Video Corpus Moment Retrieval Benchmark for Fine-Grained Video Understanding
von: Chen, Houlun, et al.
Veröffentlicht: (2024)
von: Chen, Houlun, et al.
Veröffentlicht: (2024)
PRVR: Partially Relevant Video Retrieval
von: Chen, Xianke, et al.
Veröffentlicht: (2022)
von: Chen, Xianke, et al.
Veröffentlicht: (2022)
Compressed Deepfake Video Detection Based on 3D Spatiotemporal Trajectories
von: Chen, Zongmei, et al.
Veröffentlicht: (2024)
von: Chen, Zongmei, et al.
Veröffentlicht: (2024)
YOLO-RD: Introducing Relevant and Compact Explicit Knowledge to YOLO by Retriever-Dictionary
von: Tsui, Hao-Tang, et al.
Veröffentlicht: (2024)
von: Tsui, Hao-Tang, et al.
Veröffentlicht: (2024)
Reason-RFT: Reinforcement Fine-Tuning for Visual Reasoning of Vision Language Models
von: Tan, Huajie, et al.
Veröffentlicht: (2025)
von: Tan, Huajie, et al.
Veröffentlicht: (2025)
Enhancing Visual Reliance in Text Generation: A Bayesian Perspective on Mitigating Hallucination in Large Vision-Language Models
von: Hu, Nanxing, et al.
Veröffentlicht: (2025)
von: Hu, Nanxing, et al.
Veröffentlicht: (2025)
WalkVLM:Aid Visually Impaired People Walking by Vision Language Model
von: Yuan, Zhiqiang, et al.
Veröffentlicht: (2024)
von: Yuan, Zhiqiang, et al.
Veröffentlicht: (2024)
MADTempo: An Interactive System for Multi-Event Temporal Video Retrieval with Query Augmentation
von: Vu, Huu-An, et al.
Veröffentlicht: (2025)
von: Vu, Huu-An, et al.
Veröffentlicht: (2025)
Video-EM: Event-Centric Episodic Memory for Long-Form Video Understanding
von: Wang, Yun, et al.
Veröffentlicht: (2025)
von: Wang, Yun, et al.
Veröffentlicht: (2025)
Enhancing Long Video Understanding via Hierarchical Event-Based Memory
von: Cheng, Dingxin, et al.
Veröffentlicht: (2024)
von: Cheng, Dingxin, et al.
Veröffentlicht: (2024)
Towards Video Anomaly Retrieval from Video Anomaly Detection: New Benchmarks and Model
von: Wu, Peng, et al.
Veröffentlicht: (2023)
von: Wu, Peng, et al.
Veröffentlicht: (2023)
ShotFinder: Imagination-Driven Open-Domain Video Shot Retrieval via Web Search
von: Yu, Tao, et al.
Veröffentlicht: (2026)
von: Yu, Tao, et al.
Veröffentlicht: (2026)
Enhancing Partially Relevant Video Retrieval with Robust Alignment Learning
von: Zhang, Long, et al.
Veröffentlicht: (2025)
von: Zhang, Long, et al.
Veröffentlicht: (2025)
Any-to-Bokeh: Arbitrary-Subject Video Refocusing with Video Diffusion Model
von: Yang, Yang, et al.
Veröffentlicht: (2025)
von: Yang, Yang, et al.
Veröffentlicht: (2025)
Enhancing Adversarial Robustness of Vision-Language Models through Low-Rank Adaptation
von: Ji, Yuheng, et al.
Veröffentlicht: (2024)
von: Ji, Yuheng, et al.
Veröffentlicht: (2024)
VidEvent: A Large Dataset for Understanding Dynamic Evolution of Events in Videos
von: Liang, Baoyu, et al.
Veröffentlicht: (2025)
von: Liang, Baoyu, et al.
Veröffentlicht: (2025)
Localizing Events in Videos with Multimodal Queries
von: Zhang, Gengyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Gengyuan, et al.
Veröffentlicht: (2024)
Unleashing the Power of CNN and Transformer for Balanced RGB-Event Video Recognition
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
Event-Enhanced Blurry Video Super-Resolution
von: Kai, Dachun, et al.
Veröffentlicht: (2025)
von: Kai, Dachun, et al.
Veröffentlicht: (2025)
AdvDMD: Adversarial Reward Meets DMD For High-Quality Few-Step Generation
von: Wang, Xu, et al.
Veröffentlicht: (2026)
von: Wang, Xu, et al.
Veröffentlicht: (2026)
Towards Efficient Partially Relevant Video Retrieval with Active Moment Discovering
von: Song, Peipei, et al.
Veröffentlicht: (2025)
von: Song, Peipei, et al.
Veröffentlicht: (2025)
QuRe: Query-Relevant Retrieval through Hard Negative Sampling in Composed Image Retrieval
von: Kwak, Jaehyun, et al.
Veröffentlicht: (2025)
von: Kwak, Jaehyun, et al.
Veröffentlicht: (2025)
Adversarial Attack for RGB-Event based Visual Object Tracking
von: Chen, Qiang, et al.
Veröffentlicht: (2025)
von: Chen, Qiang, et al.
Veröffentlicht: (2025)
GIRL-DETR: Gradient-Isolated Reinforcement Learning for Video Moment Retrieval
von: Zhang, Shihang, et al.
Veröffentlicht: (2026)
von: Zhang, Shihang, et al.
Veröffentlicht: (2026)
Event-Causal RAG: A Retrieval-Augmented Generation Framework for Long Video Reasoning in Complex Scenarios
von: Yan, Peizheng, et al.
Veröffentlicht: (2026)
von: Yan, Peizheng, et al.
Veröffentlicht: (2026)
REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing
von: Xu, Weihan, et al.
Veröffentlicht: (2025)
von: Xu, Weihan, et al.
Veröffentlicht: (2025)
AdaFocus: Adaptive Relevance-Diversity Sampling with Zero-Cache Look-back for Efficient Long Video Understanding
von: Yang, Xiao, et al.
Veröffentlicht: (2026)
von: Yang, Xiao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Mitigating Semantic Collapse in Partially Relevant Video Retrieval
von: Moon, WonJun, et al.
Veröffentlicht: (2025) -
Knowledge-Refined Dual Context-Aware Network for Partially Relevant Video Retrieval
von: Yang, Junkai, et al.
Veröffentlicht: (2026) -
Ambiguity-Restrained Text-Video Representation Learning for Partially Relevant Video Retrieval
von: Cho, CH, et al.
Veröffentlicht: (2025) -
Prototypes are Balanced Units for Efficient and Effective Partially Relevant Video Retrieval
von: Moon, WonJun, et al.
Veröffentlicht: (2025) -
GMMFormer: Gaussian-Mixture-Model Based Transformer for Efficient Partially Relevant Video Retrieval
von: Wang, Yuting, et al.
Veröffentlicht: (2023)