Correlation-Guided Query-Dependency Calibration for Video Temporal Grounding
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Moon, WonJun, Hyun, Sangeek, Lee, SuBeen, Heo, Jae-Pil |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Temporal Alignment-Free Video Matching for Few-shot Action Recognition
par: Lee, SuBeen, et autres
Publié: (2025)
par: Lee, SuBeen, et autres
Publié: (2025)
Progressive Proxy Anchor Propagation for Unsupervised Semantic Segmentation
par: Seong, Hyun Seok, et autres
Publié: (2024)
par: Seong, Hyun Seok, et autres
Publié: (2024)
Auxiliary Descriptive Knowledge for Few-Shot Adaptation of Vision-Language Model
par: Lee, SuBeen, et autres
Publié: (2025)
par: Lee, SuBeen, et autres
Publié: (2025)
Mitigating Background Shift in Class-Incremental Semantic Segmentation
par: Park, Gilhan, et autres
Publié: (2024)
par: Park, Gilhan, et autres
Publié: (2024)
Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric Learning
par: Moon, WonJun, et autres
Publié: (2026)
par: Moon, WonJun, et autres
Publié: (2026)
Selective Contrastive Learning for Weakly Supervised Affordance Grounding
par: Moon, WonJun, et autres
Publié: (2025)
par: Moon, WonJun, et autres
Publié: (2025)
From Vicious to Virtuous Cycles: Synergistic Representation Learning for Unsupervised Video Object-Centric Learning
par: Seong, Hyun Seok, et autres
Publié: (2026)
par: Seong, Hyun Seok, et autres
Publié: (2026)
Looking Beyond the Window: Global-Local Aligned CLIP for Training-free Open-Vocabulary Semantic Segmentation
par: Lee, ByeongCheol, et autres
Publié: (2026)
par: Lee, ByeongCheol, et autres
Publié: (2026)
Temporally Consistent Long-Term Memory for 3D Single Object Tracking
par: Yoo, Jaejoon, et autres
Publié: (2026)
par: Yoo, Jaejoon, et autres
Publié: (2026)
White Aggregation and Restoration for Few-shot 3D Point Cloud Semantic Segmentation
par: Im, Jiyun, et autres
Publié: (2025)
par: Im, Jiyun, et autres
Publié: (2025)
GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats
par: Hyun, Sangeek, et autres
Publié: (2024)
par: Hyun, Sangeek, et autres
Publié: (2024)
Foreground-Covering Prototype Generation and Matching for SAM-Aided Few-Shot Segmentation
par: Park, Suho, et autres
Publié: (2025)
par: Park, Suho, et autres
Publié: (2025)
Masking Matters: Unlocking the Spatial Reasoning Capabilities of LLMs for 3D Scene-Language Understanding
par: Jeon, Yerim, et autres
Publié: (2025)
par: Jeon, Yerim, et autres
Publié: (2025)
VLCounter: Text-aware Visual Representation for Zero-Shot Object Counting
par: Kang, Seunggu, et autres
Publié: (2023)
par: Kang, Seunggu, et autres
Publié: (2023)
Style Injection in Diffusion: A Training-free Approach for Adapting Large-scale Diffusion Models for Style Transfer
par: Chung, Jiwoo, et autres
Publié: (2023)
par: Chung, Jiwoo, et autres
Publié: (2023)
Cross-scale Aligned Supervision for Training GANs
par: Hyun, Sangeek, et autres
Publié: (2026)
par: Hyun, Sangeek, et autres
Publié: (2026)
Disambiguating 2D-3D Correspondences in Gaussian Splatting-based Feature Fields for Visual Localization
par: Lee, Miso, et autres
Publié: (2026)
par: Lee, Miso, et autres
Publié: (2026)
Scalable GANs with Transformers
par: Hyun, Sangeek, et autres
Publié: (2025)
par: Hyun, Sangeek, et autres
Publié: (2025)
Auto-Encoded Supervision for Perceptual Image Super-Resolution
par: Lee, MinKyu, et autres
Publié: (2024)
par: Lee, MinKyu, et autres
Publié: (2024)
Diversity-aware Channel Pruning for StyleGAN Compression
par: Chung, Jiwoo, et autres
Publié: (2024)
par: Chung, Jiwoo, et autres
Publié: (2024)
Prototypes are Balanced Units for Efficient and Effective Partially Relevant Video Retrieval
par: Moon, WonJun, et autres
Publié: (2025)
par: Moon, WonJun, et autres
Publié: (2025)
Analyzing the Training Dynamics of Image Restoration Transformers: A Revisit to Layer Normalization
par: Lee, MinKyu, et autres
Publié: (2025)
par: Lee, MinKyu, et autres
Publié: (2025)
Mitigating Semantic Collapse in Partially Relevant Video Retrieval
par: Moon, WonJun, et autres
Publié: (2025)
par: Moon, WonJun, et autres
Publié: (2025)
Fine-Tuning Visual Autoregressive Models for Subject-Driven Generation
par: Chung, Jiwoo, et autres
Publié: (2025)
par: Chung, Jiwoo, et autres
Publié: (2025)
Long-term Pre-training for Temporal Action Detection with Transformers
par: Kim, Jihwan, et autres
Publié: (2024)
par: Kim, Jihwan, et autres
Publié: (2024)
SeaCache: Spectral-Evolution-Aware Cache for Accelerating Diffusion Models
par: Chung, Jiwoo, et autres
Publié: (2026)
par: Chung, Jiwoo, et autres
Publié: (2026)
Noise-free Optimization in Early Training Steps for Image Super-Resolution
par: Lee, MinKyu, et autres
Publié: (2023)
par: Lee, MinKyu, et autres
Publié: (2023)
Prediction-Feedback DETR for Temporal Action Detection
par: Kim, Jihwan, et autres
Publié: (2024)
par: Kim, Jihwan, et autres
Publié: (2024)
Boundary-Recovering Network for Temporal Action Detection
par: Kim, Jihwan, et autres
Publié: (2024)
par: Kim, Jihwan, et autres
Publié: (2024)
Activating Self-Attention for Multi-Scene Absolute Pose Regression
par: Lee, Miso, et autres
Publié: (2024)
par: Lee, Miso, et autres
Publié: (2024)
Learning to Refuse: Refusal-Aware Reinforcement Fine-Tuning for Hard-Irrelevant Queries in Video Temporal Grounding
par: Lee, Jin-Seop, et autres
Publié: (2025)
par: Lee, Jin-Seop, et autres
Publié: (2025)
Mutually-Aware Feature Learning for Few-Shot Object Counting
par: Jeon, Yerim, et autres
Publié: (2024)
par: Jeon, Yerim, et autres
Publié: (2024)
Stay in your Lane: Role Specific Queries with Overlap Suppression Loss for Dense Video Captioning
par: Baek, Seung Hyup, et autres
Publié: (2026)
par: Baek, Seung Hyup, et autres
Publié: (2026)
Let Me Finish My Sentence: Video Temporal Grounding with Holistic Text Understanding
par: Woo, Jongbhin, et autres
Publié: (2024)
par: Woo, Jongbhin, et autres
Publié: (2024)
Diversifying Query: Region-Guided Transformer for Temporal Sentence Grounding
par: Sun, Xiaolong, et autres
Publié: (2024)
par: Sun, Xiaolong, et autres
Publié: (2024)
Temporal Grounding as a Learning Signal for Referring Video Object Segmentation
par: Lee, Seunghun, et autres
Publié: (2025)
par: Lee, Seunghun, et autres
Publié: (2025)
Perceive, Query & Reason: Enhancing Video QA with Question-Guided Temporal Queries
par: Amoroso, Roberto, et autres
Publié: (2024)
par: Amoroso, Roberto, et autres
Publié: (2024)
Context-Guided Spatio-Temporal Video Grounding
par: Gu, Xin, et autres
Publié: (2024)
par: Gu, Xin, et autres
Publié: (2024)
CVA: Context-aware Video-text Alignment for Video Temporal Grounding
par: Moon, Sungho, et autres
Publié: (2026)
par: Moon, Sungho, et autres
Publié: (2026)
TimeExpert: An Expert-Guided Video LLM for Video Temporal Grounding
par: Yang, Zuhao, et autres
Publié: (2025)
par: Yang, Zuhao, et autres
Publié: (2025)
Documents similaires
-
Temporal Alignment-Free Video Matching for Few-shot Action Recognition
par: Lee, SuBeen, et autres
Publié: (2025) -
Progressive Proxy Anchor Propagation for Unsupervised Semantic Segmentation
par: Seong, Hyun Seok, et autres
Publié: (2024) -
Auxiliary Descriptive Knowledge for Few-Shot Adaptation of Vision-Language Model
par: Lee, SuBeen, et autres
Publié: (2025) -
Mitigating Background Shift in Class-Incremental Semantic Segmentation
par: Park, Gilhan, et autres
Publié: (2024) -
Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric Learning
par: Moon, WonJun, et autres
Publié: (2026)