Selective Query-guided Debiasing for Video Corpus Moment Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yoon, Sunjae, Hong, Ji Woo, Yoon, Eunseop, Kim, Dahyun, Kim, Junyeong, Yoon, Hee Suk, Yoo, Chang D. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SCANet: Scene Complexity Aware Network for Weakly-Supervised Video Moment Retrieval
von: Yoon, Sunjae, et al.
Veröffentlicht: (2023)
von: Yoon, Sunjae, et al.
Veröffentlicht: (2023)
HEAR: Hearing Enhanced Audio Response for Video-grounded Dialogue
von: Yoon, Sunjae, et al.
Veröffentlicht: (2023)
von: Yoon, Sunjae, et al.
Veröffentlicht: (2023)
Can Video LLMs Refuse to Answer? Alignment for Answerability in Video Large Language Models
von: Yoon, Eunseop, et al.
Veröffentlicht: (2025)
von: Yoon, Eunseop, et al.
Veröffentlicht: (2025)
ESD: Expected Squared Difference as a Tuning-Free Trainable Calibration Measure
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2023)
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2023)
GranAlign: Granularity-Aware Alignment Framework for Zero-Shot Video Moment Retrieval
von: Jeon, Mingyu, et al.
Veröffentlicht: (2026)
von: Jeon, Mingyu, et al.
Veröffentlicht: (2026)
DNI: Dilutional Noise Initialization for Diffusion Video Editing
von: Yoon, Sunjae, et al.
Veröffentlicht: (2024)
von: Yoon, Sunjae, et al.
Veröffentlicht: (2024)
FRAG: Frequency Adapting Group for Diffusion Video Editing
von: Yoon, Sunjae, et al.
Veröffentlicht: (2024)
von: Yoon, Sunjae, et al.
Veröffentlicht: (2024)
FlowDrag: 3D-aware Drag-based Image Editing with Mesh-guided Deformation Vector Flow Fields
von: Koo, Gwanhyeong, et al.
Veröffentlicht: (2025)
von: Koo, Gwanhyeong, et al.
Veröffentlicht: (2025)
SimPSI: A Simple Strategy to Preserve Spectral Information in Time Series Data Augmentation
von: Ryu, Hyun, et al.
Veröffentlicht: (2023)
von: Ryu, Hyun, et al.
Veröffentlicht: (2023)
FlexiEdit: Frequency-Aware Latent Refinement for Enhanced Non-Rigid Editing
von: Koo, Gwanhyeong, et al.
Veröffentlicht: (2024)
von: Koo, Gwanhyeong, et al.
Veröffentlicht: (2024)
High-Fidelity Text-to-Image Generation from Pre-Trained Vision-Language Models via Distribution-Conditioned Diffusion Decoding
von: Hong, Ji Woo, et al.
Veröffentlicht: (2026)
von: Hong, Ji Woo, et al.
Veröffentlicht: (2026)
Point to Span: Zero-Shot Moment Retrieval for Navigating Unseen Hour-Long Videos
von: Jeon, Mingyu, et al.
Veröffentlicht: (2025)
von: Jeon, Mingyu, et al.
Veröffentlicht: (2025)
Decomposed On-Policy Distillation for Vision-Language Reasoning: Steering Gradients for Visual Grounding
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2026)
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2026)
Occlusion-robust Stylization for Drawing-based 3D Animation
von: Yoon, Sunjae, et al.
Veröffentlicht: (2025)
von: Yoon, Sunjae, et al.
Veröffentlicht: (2025)
Wavelet-Guided Acceleration of Text Inversion in Diffusion-Based Image Editing
von: Koo, Gwanhyeong, et al.
Veröffentlicht: (2024)
von: Koo, Gwanhyeong, et al.
Veröffentlicht: (2024)
C-TPT: Calibrated Test-Time Prompt Tuning for Vision-Language Models via Text Feature Dispersion
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2024)
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2024)
Visual Funnel: Resolving Contextual Blindness in Multimodal Large Language Models
von: Jung, Woojun, et al.
Veröffentlicht: (2025)
von: Jung, Woojun, et al.
Veröffentlicht: (2025)
ConfPO: Exploiting Policy Model Confidence for Critical Token Selection in Preference Optimization
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2025)
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2025)
ITA-MDT: Image-Timestep-Adaptive Masked Diffusion Transformer Framework for Image-Based Virtual Try-On
von: Hong, Ji Woo, et al.
Veröffentlicht: (2025)
von: Hong, Ji Woo, et al.
Veröffentlicht: (2025)
TPC: Test-time Procrustes Calibration for Diffusion-based Human Image Animation
von: Yoon, Sunjae, et al.
Veröffentlicht: (2024)
von: Yoon, Sunjae, et al.
Veröffentlicht: (2024)
Language-Grounded Multi-Domain Image Translation via Semantic Difference Guidance
von: Ryu, Jongwon, et al.
Veröffentlicht: (2026)
von: Ryu, Jongwon, et al.
Veröffentlicht: (2026)
See More, Store Less: Memory-Efficient Resolution for Video Moment Retrieval
von: Jeon, Mingyu, et al.
Veröffentlicht: (2026)
von: Jeon, Mingyu, et al.
Veröffentlicht: (2026)
CMTA: Cross-Modal Temporal Alignment for Event-guided Video Deblurring
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
MVMR: A New Framework for Evaluating Faithfulness of Video Moment Retrieval against Multiple Distractors
von: Yang, Nakyeong, et al.
Veröffentlicht: (2023)
von: Yang, Nakyeong, et al.
Veröffentlicht: (2023)
EGLOCE: Training-Free Energy-Guided Latent Optimization for Concept Erasure
von: Ahn, Junyeong, et al.
Veröffentlicht: (2026)
von: Ahn, Junyeong, et al.
Veröffentlicht: (2026)
LI-TTA: Language Informed Test-Time Adaptation for Automatic Speech Recognition
von: Yoon, Eunseop, et al.
Veröffentlicht: (2024)
von: Yoon, Eunseop, et al.
Veröffentlicht: (2024)
AdaMER-CTC: Connectionist Temporal Classification with Adaptive Maximum Entropy Regularization for Automatic Speech Recognition
von: Eom, SooHwan, et al.
Veröffentlicht: (2024)
von: Eom, SooHwan, et al.
Veröffentlicht: (2024)
Learning Event-guided Exposure-agnostic Video Frame Interpolation via Adaptive Feature Blending
von: Jung, Junsik, et al.
Veröffentlicht: (2025)
von: Jung, Junsik, et al.
Veröffentlicht: (2025)
Towards Test-time Efficient Visual Place Recognition via Asymmetric Query Processing
von: Kim, Jaeyoon, et al.
Veröffentlicht: (2025)
von: Kim, Jaeyoon, et al.
Veröffentlicht: (2025)
Large Language Models Facilitate Vision Reflection in Image Classification
von: An, Guoyuan, et al.
Veröffentlicht: (2025)
von: An, Guoyuan, et al.
Veröffentlicht: (2025)
Zero-Shot Dual-Path Integration Framework for Open-Vocabulary 3D Instance Segmentation
von: Ton, Tri, et al.
Veröffentlicht: (2024)
von: Ton, Tri, et al.
Veröffentlicht: (2024)
Towards Real-world Event-guided Low-light Video Enhancement and Deblurring
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
Referring Video Object Segmentation via Language-aligned Track Selection
von: Kim, Seongchan, et al.
Veröffentlicht: (2024)
von: Kim, Seongchan, et al.
Veröffentlicht: (2024)
Progressive Fourier Neural Representation for Sequential Video Compilation
von: Kang, Haeyong, et al.
Veröffentlicht: (2023)
von: Kang, Haeyong, et al.
Veröffentlicht: (2023)
Your Super Resolution Model is not Enough for Tackling Real-World Scenarios
von: Yoon, Dongsik, et al.
Veröffentlicht: (2025)
von: Yoon, Dongsik, et al.
Veröffentlicht: (2025)
From Prompts to Deployment: Auto-Curated Domain-Specific Dataset Generation via Diffusion Models
von: Yoon, Dongsik, et al.
Veröffentlicht: (2026)
von: Yoon, Dongsik, et al.
Veröffentlicht: (2026)
Moment of Untruth: Dealing with Negative Queries in Video Moment Retrieval
von: Flanagan, Kevin, et al.
Veröffentlicht: (2025)
von: Flanagan, Kevin, et al.
Veröffentlicht: (2025)
Event-aware Video Corpus Moment Retrieval
von: Hou, Danyang, et al.
Veröffentlicht: (2024)
von: Hou, Danyang, et al.
Veröffentlicht: (2024)
Background-aware Moment Detection for Video Moment Retrieval
von: Jung, Minjoon, et al.
Veröffentlicht: (2023)
von: Jung, Minjoon, et al.
Veröffentlicht: (2023)
Continual Learning: Forget-free Winning Subnetworks for Video Representations
von: Kang, Haeyong, et al.
Veröffentlicht: (2023)
von: Kang, Haeyong, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
SCANet: Scene Complexity Aware Network for Weakly-Supervised Video Moment Retrieval
von: Yoon, Sunjae, et al.
Veröffentlicht: (2023) -
HEAR: Hearing Enhanced Audio Response for Video-grounded Dialogue
von: Yoon, Sunjae, et al.
Veröffentlicht: (2023) -
Can Video LLMs Refuse to Answer? Alignment for Answerability in Video Large Language Models
von: Yoon, Eunseop, et al.
Veröffentlicht: (2025) -
ESD: Expected Squared Difference as a Tuning-Free Trainable Calibration Measure
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2023) -
GranAlign: Granularity-Aware Alignment Framework for Zero-Shot Video Moment Retrieval
von: Jeon, Mingyu, et al.
Veröffentlicht: (2026)