Ambiguity-Restrained Text-Video Representation Learning for Partially Relevant Video Retrieval
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Cho, CH, Moon, WJ, Jun, W, Jung, MS, Heo, JP |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Mitigating Semantic Collapse in Partially Relevant Video Retrieval
par: Moon, WonJun, et autres
Publié: (2025)
par: Moon, WonJun, et autres
Publié: (2025)
Prototypes are Balanced Units for Efficient and Effective Partially Relevant Video Retrieval
par: Moon, WonJun, et autres
Publié: (2025)
par: Moon, WonJun, et autres
Publié: (2025)
Uneven Event Modeling for Partially Relevant Video Retrieval
par: Zhu, Sa, et autres
Publié: (2025)
par: Zhu, Sa, et autres
Publié: (2025)
Knowledge-Refined Dual Context-Aware Network for Partially Relevant Video Retrieval
par: Yang, Junkai, et autres
Publié: (2026)
par: Yang, Junkai, et autres
Publié: (2026)
ReSpec: Relevance and Specificity Grounded Online Filtering for Learning on Video-Text Data Streams
par: Kim, Chris Dongjoo, et autres
Publié: (2025)
par: Kim, Chris Dongjoo, et autres
Publié: (2025)
Temporal Alignment-Free Video Matching for Few-shot Action Recognition
par: Lee, SuBeen, et autres
Publié: (2025)
par: Lee, SuBeen, et autres
Publié: (2025)
VidVec: Unlocking Video MLLM Embeddings for Video-Text Retrieval
par: Tzachor, Issar, et autres
Publié: (2026)
par: Tzachor, Issar, et autres
Publié: (2026)
GMMFormer: Gaussian-Mixture-Model Based Transformer for Efficient Partially Relevant Video Retrieval
par: Wang, Yuting, et autres
Publié: (2023)
par: Wang, Yuting, et autres
Publié: (2023)
Dual-Modal Attention-Enhanced Text-Video Retrieval with Triplet Partial Margin Contrastive Learning
par: Jiang, Chen, et autres
Publié: (2023)
par: Jiang, Chen, et autres
Publié: (2023)
BlobGEN-Vid: Compositional Text-to-Video Generation with Blob Video Representations
par: Feng, Weixi, et autres
Publié: (2025)
par: Feng, Weixi, et autres
Publié: (2025)
xGen-VideoSyn-1: High-fidelity Text-to-Video Synthesis with Compressed Representations
par: Qin, Can, et autres
Publié: (2024)
par: Qin, Can, et autres
Publié: (2024)
VideoSAGE: Video Summarization with Graph Representation Learning
par: Chaves, Jose M. Rojas, et autres
Publié: (2024)
par: Chaves, Jose M. Rojas, et autres
Publié: (2024)
Multi-Scale Temporal Difference Transformer for Video-Text Retrieval
par: Wang, Ni, et autres
Publié: (2024)
par: Wang, Ni, et autres
Publié: (2024)
Selective Contrastive Learning for Weakly Supervised Affordance Grounding
par: Moon, WonJun, et autres
Publié: (2025)
par: Moon, WonJun, et autres
Publié: (2025)
Video Text Preservation with Synthetic Text-Rich Videos
par: Liu, Ziyang, et autres
Publié: (2025)
par: Liu, Ziyang, et autres
Publié: (2025)
Multi-Group Proportional Representation for Text-to-Image Models
par: Jung, Sangwon, et autres
Publié: (2025)
par: Jung, Sangwon, et autres
Publié: (2025)
AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
par: Wang, Zun, et autres
Publié: (2026)
par: Wang, Zun, et autres
Publié: (2026)
PRVR: Partially Relevant Video Retrieval
par: Chen, Xianke, et autres
Publié: (2022)
par: Chen, Xianke, et autres
Publié: (2022)
Frame-Difference Guided Dynamic Region Perception for CLIP Adaptation in Text-Video Retrieval
par: Yu, Jiaao, et autres
Publié: (2025)
par: Yu, Jiaao, et autres
Publié: (2025)
AURA: Development and Validation of an Augmented Unplanned Removal Alert System using Synthetic ICU Videos
par: Seo, Junhyuk, et autres
Publié: (2025)
par: Seo, Junhyuk, et autres
Publié: (2025)
Video Representation Learning with Joint-Embedding Predictive Architectures
par: Drozdov, Katrina, et autres
Publié: (2024)
par: Drozdov, Katrina, et autres
Publié: (2024)
EPiC: Efficient Video Camera Control Learning with Precise Anchor-Video Guidance
par: Wang, Zun, et autres
Publié: (2025)
par: Wang, Zun, et autres
Publié: (2025)
Snap Video: Scaled Spatiotemporal Transformers for Text-to-Video Synthesis
par: Menapace, Willi, et autres
Publié: (2024)
par: Menapace, Willi, et autres
Publié: (2024)
Video-RAG: Visually-aligned Retrieval-Augmented Long Video Comprehension
par: Luo, Yongdong, et autres
Publié: (2024)
par: Luo, Yongdong, et autres
Publié: (2024)
Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement
par: Lee, Daeun, et autres
Publié: (2024)
par: Lee, Daeun, et autres
Publié: (2024)
Modeling Depth Ambiguity: A Mixture-Density Representation for Flying-Point-Free Depth Estimation
par: Bian, Siyuan, et autres
Publié: (2026)
par: Bian, Siyuan, et autres
Publié: (2026)
Retrieval, Refinement, and Ranking for Text-to-Video Generation via Prompt Optimization and Test-Time Scaling
par: Rahman, Zillur, et autres
Publié: (2026)
par: Rahman, Zillur, et autres
Publié: (2026)
Watch Video, Catch Keyword: Context-aware Keyword Attention for Moment Retrieval and Highlight Detection
par: Um, Sung Jin, et autres
Publié: (2025)
par: Um, Sung Jin, et autres
Publié: (2025)
REMAP: Regularized Matching and Partial Alignment of Video Embeddings
par: Chandra, Soumyadeep, et autres
Publié: (2025)
par: Chandra, Soumyadeep, et autres
Publié: (2025)
BadVideo: Stealthy Backdoor Attack against Text-to-Video Generation
par: Wang, Ruotong, et autres
Publié: (2025)
par: Wang, Ruotong, et autres
Publié: (2025)
Embedded Representation Learning Network for Animating Styled Video Portrait
par: Wang, Tianyong, et autres
Publié: (2024)
par: Wang, Tianyong, et autres
Publié: (2024)
SLVMEval: Synthetic Meta Evaluation Benchmark for Text-to-Long Video Generation
par: Matsuda, Ryosuke, et autres
Publié: (2026)
par: Matsuda, Ryosuke, et autres
Publié: (2026)
GIRL-DETR: Gradient-Isolated Reinforcement Learning for Video Moment Retrieval
par: Zhang, Shihang, et autres
Publié: (2026)
par: Zhang, Shihang, et autres
Publié: (2026)
HawkEye: Training Video-Text LLMs for Grounding Text in Videos
par: Wang, Yueqian, et autres
Publié: (2024)
par: Wang, Yueqian, et autres
Publié: (2024)
Masking Matters: Unlocking the Spatial Reasoning Capabilities of LLMs for 3D Scene-Language Understanding
par: Jeon, Yerim, et autres
Publié: (2025)
par: Jeon, Yerim, et autres
Publié: (2025)
SViTT-Ego: A Sparse Video-Text Transformer for Egocentric Video
par: Valdez, Hector A., et autres
Publié: (2024)
par: Valdez, Hector A., et autres
Publié: (2024)
Towards Video Anomaly Retrieval from Video Anomaly Detection: New Benchmarks and Model
par: Wu, Peng, et autres
Publié: (2023)
par: Wu, Peng, et autres
Publié: (2023)
VERIFIED: A Video Corpus Moment Retrieval Benchmark for Fine-Grained Video Understanding
par: Chen, Houlun, et autres
Publié: (2024)
par: Chen, Houlun, et autres
Publié: (2024)
Enhancing Partially Relevant Video Retrieval with Robust Alignment Learning
par: Zhang, Long, et autres
Publié: (2025)
par: Zhang, Long, et autres
Publié: (2025)
Learning Multi-modal Representations by Watching Hundreds of Surgical Video Lectures
par: Yuan, Kun, et autres
Publié: (2023)
par: Yuan, Kun, et autres
Publié: (2023)
Documents similaires
-
Mitigating Semantic Collapse in Partially Relevant Video Retrieval
par: Moon, WonJun, et autres
Publié: (2025) -
Prototypes are Balanced Units for Efficient and Effective Partially Relevant Video Retrieval
par: Moon, WonJun, et autres
Publié: (2025) -
Uneven Event Modeling for Partially Relevant Video Retrieval
par: Zhu, Sa, et autres
Publié: (2025) -
Knowledge-Refined Dual Context-Aware Network for Partially Relevant Video Retrieval
par: Yang, Junkai, et autres
Publié: (2026) -
ReSpec: Relevance and Specificity Grounded Online Filtering for Learning on Video-Text Data Streams
par: Kim, Chris Dongjoo, et autres
Publié: (2025)