TAG: A Simple Yet Effective Temporal-Aware Approach for Zero-Shot Video Temporal Grounding
Fuente:
arXiv
Guardado en:
| Autores principales: | Lee, Jin-Seop, Lee, SungJoon, Ahn, Jaehan, Choi, YunSeok, Lee, Jee-Hyong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning to Refuse: Refusal-Aware Reinforcement Fine-Tuning for Hard-Irrelevant Queries in Video Temporal Grounding
por: Lee, Jin-Seop, et al.
Publicado: (2025)
por: Lee, Jin-Seop, et al.
Publicado: (2025)
DeCAP: Context-Adaptive Prompt Generation for Debiasing Zero-shot Question Answering in Large Language Models
por: Bae, Suyoung, et al.
Publicado: (2025)
por: Bae, Suyoung, et al.
Publicado: (2025)
DCG-SQL: Enhancing In-Context Learning for Text-to-SQL with Deep Contextual Schema Link Graph
por: Lee, Jihyung, et al.
Publicado: (2025)
por: Lee, Jihyung, et al.
Publicado: (2025)
SALAD: Improving Robustness and Generalization through Contrastive Learning with Structure-Aware and LLM-Driven Augmented Data
por: Bae, Suyoung, et al.
Publicado: (2025)
por: Bae, Suyoung, et al.
Publicado: (2025)
BD-Net: Has Depth-Wise Convolution Ever Been Applied in Binary Neural Networks?
por: Kim, DoYoung, et al.
Publicado: (2025)
por: Kim, DoYoung, et al.
Publicado: (2025)
Q-FAKER: Query-free Hard Black-box Attack via Controlled Generation
por: Na, CheolWon, et al.
Publicado: (2025)
por: Na, CheolWon, et al.
Publicado: (2025)
Remember Your Trace: Memory-Guided Long-Horizon Agentic Framework for Consistent and Hierarchical Repository-Level Code Documentation
por: Bae, Suyoung, et al.
Publicado: (2026)
por: Bae, Suyoung, et al.
Publicado: (2026)
Stabilizing Open-Set Test-Time Adaptation via Primary-Auxiliary Filtering and Knowledge-Integrated Prediction
por: Lee, Byung-Joon, et al.
Publicado: (2025)
por: Lee, Byung-Joon, et al.
Publicado: (2025)
ReFEree: Reference-Free and Fine-Grained Method for Evaluating Factual Consistency in Real-World Code Summarization
por: Bae, Suyoung, et al.
Publicado: (2026)
por: Bae, Suyoung, et al.
Publicado: (2026)
CountCluster: Training-Free Object Quantity Guidance with Cross-Attention Map Clustering for Text-to-Image Generation
por: Lee, Joohyeon, et al.
Publicado: (2025)
por: Lee, Joohyeon, et al.
Publicado: (2025)
DomCLP: Domain-wise Contrastive Learning with Prototype Mixup for Unsupervised Domain Generalization
por: Lee, Jin-Seop, et al.
Publicado: (2024)
por: Lee, Jin-Seop, et al.
Publicado: (2024)
G-MIXER: Geodesic Mixup-based Implicit Semantic Expansion and Explicit Semantic Re-ranking for Zero-Shot Composed Image Retrieval
por: Lim, Jiyoung, et al.
Publicado: (2026)
por: Lim, Jiyoung, et al.
Publicado: (2026)
EVIDENT: Routing MLLM Adaptation through Entity-Grounded Visual Evidence for Cross-Domain Video Temporal Grounding
por: Ahn, Geo, et al.
Publicado: (2026)
por: Ahn, Geo, et al.
Publicado: (2026)
ORIGEN: Zero-Shot 3D Orientation Grounding in Text-to-Image Generation
por: Min, Yunhong, et al.
Publicado: (2025)
por: Min, Yunhong, et al.
Publicado: (2025)
Information Density Enhancement Using Lossy Compression in DNA Data Storage
por: Seongjun Seo, et al.
Publicado: (2024)
por: Seongjun Seo, et al.
Publicado: (2024)
A Sensitivity Analysis of Multi-Event Audio Grounding in Audio LLMs
por: Lee, Taehan, et al.
Publicado: (2026)
por: Lee, Taehan, et al.
Publicado: (2026)
Are Video Models Zero-Shot Learners and Reasoners in Education? EduVideoBench, A Knowledge-Skills-Attitude Benchmark for Educational Video Generation
por: Lee, Unggi, et al.
Publicado: (2026)
por: Lee, Unggi, et al.
Publicado: (2026)
SlotVTG: Object-Centric Adapter for Generalizable Video Temporal Grounding
por: Han, Jiwook, et al.
Publicado: (2026)
por: Han, Jiwook, et al.
Publicado: (2026)
Unleashing the Potential of Multimodal LLMs for Zero-Shot Spatio-Temporal Video Grounding
por: Yang, Zaiquan, et al.
Publicado: (2025)
por: Yang, Zaiquan, et al.
Publicado: (2025)
VTG-GPT: Tuning-Free Zero-Shot Video Temporal Grounding with GPT
por: Xu, Yifang, et al.
Publicado: (2024)
por: Xu, Yifang, et al.
Publicado: (2024)
VoicePrompter: Robust Zero-Shot Voice Conversion with Voice Prompt and Conditional Flow Matching
por: Choi, Ha-Yeong, et al.
Publicado: (2025)
por: Choi, Ha-Yeong, et al.
Publicado: (2025)
Learning Uncertainty-Aware Temporally-Extended Actions
por: Lee, Joongkyu, et al.
Publicado: (2024)
por: Lee, Joongkyu, et al.
Publicado: (2024)
Self-distillation Regularized Connectionist Temporal Classification Loss for Text Recognition: A Simple Yet Effective Approach
por: Zhang, Ziyin, et al.
Publicado: (2023)
por: Zhang, Ziyin, et al.
Publicado: (2023)
CVA: Context-aware Video-text Alignment for Video Temporal Grounding
por: Moon, Sungho, et al.
Publicado: (2026)
por: Moon, Sungho, et al.
Publicado: (2026)
Dual‐Hydrogen Bond Donor‐Functionalized Carbon Nanotube Fibers: Enhancing Anion‐Sensing Performance Through Functionalization Approaches
por: Seung‐Ho Choi, et al.
Publicado: (2024)
por: Seung‐Ho Choi, et al.
Publicado: (2024)
BAM-DETR: Boundary-Aligned Moment Detection Transformer for Temporal Sentence Grounding in Videos
por: Lee, Pilhyeon, et al.
Publicado: (2023)
por: Lee, Pilhyeon, et al.
Publicado: (2023)
Temporal Grounding as a Learning Signal for Referring Video Object Segmentation
por: Lee, Seunghun, et al.
Publicado: (2025)
por: Lee, Seunghun, et al.
Publicado: (2025)
TAG: Tangential Amplifying Guidance for Hallucination-Resistant Sampling
por: Cho, Hyunmin, et al.
Publicado: (2025)
por: Cho, Hyunmin, et al.
Publicado: (2025)
Tsanet: Temporal and Scale Alignment for Unsupervised Video Object Segmentation
por: Lee, Seunghoon, et al.
Publicado: (2023)
por: Lee, Seunghoon, et al.
Publicado: (2023)
Two-Pass Zero-Shot Temporal-Spatial Grounding of Rare Traffic Events in Surveillance Video
por: Huang, Jiantang
Publicado: (2026)
por: Huang, Jiantang
Publicado: (2026)
Infusing Environmental Captions for Long-Form Video Language Grounding
por: Lee, Hyogun, et al.
Publicado: (2024)
por: Lee, Hyogun, et al.
Publicado: (2024)
Zero-Shot Temporal Interaction Localization for Egocentric Videos
por: Zhang, Erhang, et al.
Publicado: (2025)
por: Zhang, Erhang, et al.
Publicado: (2025)
Gaussian Mixture Proposals with Pull-Push Learning Scheme to Capture Diverse Events for Weakly Supervised Temporal Video Grounding
por: Kim, Sunoh, et al.
Publicado: (2023)
por: Kim, Sunoh, et al.
Publicado: (2023)
Zero-TIG: Temporal Consistency-Aware Zero-Shot Illumination-Guided Low-light Video Enhancement
por: Li, Yini, et al.
Publicado: (2025)
por: Li, Yini, et al.
Publicado: (2025)
Correlation-Guided Query-Dependency Calibration for Video Temporal Grounding
por: Moon, WonJun, et al.
Publicado: (2023)
por: Moon, WonJun, et al.
Publicado: (2023)
SimBase: A Simple Baseline for Temporal Video Grounding
por: Bao, Peijun, et al.
Publicado: (2024)
por: Bao, Peijun, et al.
Publicado: (2024)
Exploring Temporally-Aware Features for Point Tracking
por: Kim, Inès Hyeonsu, et al.
Publicado: (2025)
por: Kim, Inès Hyeonsu, et al.
Publicado: (2025)
AnyAnomaly: Zero-Shot Customizable Video Anomaly Detection with LVLM
por: Ahn, Sunghyun, et al.
Publicado: (2025)
por: Ahn, Sunghyun, et al.
Publicado: (2025)
TempCore: Are Video QA Benchmarks Temporally Grounded? A Frame Selection Sensitivity Analysis and Benchmark
por: Ok, Hyunjong, et al.
Publicado: (2025)
por: Ok, Hyunjong, et al.
Publicado: (2025)
Empower Words: DualGround for Structured Phrase and Sentence-Level Temporal Grounding
por: Kang, Minseok, et al.
Publicado: (2025)
por: Kang, Minseok, et al.
Publicado: (2025)
Ejemplares similares
-
Learning to Refuse: Refusal-Aware Reinforcement Fine-Tuning for Hard-Irrelevant Queries in Video Temporal Grounding
por: Lee, Jin-Seop, et al.
Publicado: (2025) -
DeCAP: Context-Adaptive Prompt Generation for Debiasing Zero-shot Question Answering in Large Language Models
por: Bae, Suyoung, et al.
Publicado: (2025) -
DCG-SQL: Enhancing In-Context Learning for Text-to-SQL with Deep Contextual Schema Link Graph
por: Lee, Jihyung, et al.
Publicado: (2025) -
SALAD: Improving Robustness and Generalization through Contrastive Learning with Structure-Aware and LLM-Driven Augmented Data
por: Bae, Suyoung, et al.
Publicado: (2025) -
BD-Net: Has Depth-Wise Convolution Ever Been Applied in Binary Neural Networks?
por: Kim, DoYoung, et al.
Publicado: (2025)