Temporal Grounding as a Learning Signal for Referring Video Object Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Seunghun, Seo, Jiwan, Kim, Jeonghoon, Moon, Sungho, Kim, Siwon, Yun, Haeun, Jeon, Hyogyeong, Choi, Wonhyeok, Jeong, Jaehoon, Durante, Zane, Park, Sang Hyun, Im, Sunghoon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CVA: Context-aware Video-text Alignment for Video Temporal Grounding
by: Moon, Sungho, et al.
Published: (2026)
by: Moon, Sungho, et al.
Published: (2026)
Latest Object Memory Management for Temporally Consistent Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2025)
by: Lee, Seunghun, et al.
Published: (2025)
Depth-discriminative Metric Learning for Monocular 3D Object Detection
by: Choi, Wonhyeok, et al.
Published: (2024)
by: Choi, Wonhyeok, et al.
Published: (2024)
CAVIS: Context-Aware Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2024)
by: Lee, Seunghun, et al.
Published: (2024)
Bridging Geometric and Semantic Foundation Models for Generalized Monocular Depth Estimation
by: Ma, Sanggyun, et al.
Published: (2025)
by: Ma, Sanggyun, et al.
Published: (2025)
Rethinking LiDAR Domain Generalization: Single Source as Multiple Density Domains
by: Kim, Jaeyeul, et al.
Published: (2023)
by: Kim, Jaeyeul, et al.
Published: (2023)
Multi-task Learning for Real-time Autonomous Driving Leveraging Task-adaptive Attention Generator
by: Choi, Wonhyeok, et al.
Published: (2024)
by: Choi, Wonhyeok, et al.
Published: (2024)
Self-supervised Monocular Depth Estimation Robust to Reflective Surface Leveraged by Triplet Mining
by: Choi, Wonhyeok, et al.
Published: (2025)
by: Choi, Wonhyeok, et al.
Published: (2025)
A Training-Free Style-Personalization via SVD-Based Feature Decomposition
by: Lee, Kyoungmin, et al.
Published: (2025)
by: Lee, Kyoungmin, et al.
Published: (2025)
Objectomaly: Objectness-Aware Refinement for OoD Segmentation with Structural Consistency and Boundary Precision
by: Song, Jeonghoon, et al.
Published: (2025)
by: Song, Jeonghoon, et al.
Published: (2025)
On Rainbow Turán Densities of Trees
by: Seonghyuk Im, et al.
Published: (2025)
by: Seonghyuk Im, et al.
Published: (2025)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
by: Choi, Wonhyeok, et al.
Published: (2026)
by: Choi, Wonhyeok, et al.
Published: (2026)
Style-Editor: Text-driven object-centric style editing
by: Park, Jihun, et al.
Published: (2024)
by: Park, Jihun, et al.
Published: (2024)
Development of a Validation and Inspection Tool for Armband-based Lifelog Data (VITAL) to Facilitate the Clinical Use of Wearable Data: A Prototype and Usability Evaluation
by: Eunyoung, Im, et al.
Published: (2025)
by: Eunyoung, Im, et al.
Published: (2025)
Intrinsic Image Decomposition for Robust Self-supervised Monocular Depth Estimation on Reflective Surfaces
by: Choi, Wonhyeok, et al.
Published: (2025)
by: Choi, Wonhyeok, et al.
Published: (2025)
Sufficient conditions for additivity of the zero-error classical capacity of quantum channels
by: Park, Jeonghoon, et al.
Published: (2026)
by: Park, Jeonghoon, et al.
Published: (2026)
Ion-Trap Chip Architecture Optimized for Implementation of Quantum Error-Correcting Code
by: Lee, Jeonghoon, et al.
Published: (2025)
by: Lee, Jeonghoon, et al.
Published: (2025)
Multi-Rate Task-Oriented Communication for Multi-Edge Cooperative Inference
by: Kim, Dongwon, et al.
Published: (2025)
by: Kim, Dongwon, et al.
Published: (2025)
Battling the Non-stationarity in Time Series Forecasting via Test-time Adaptation
by: Kim, HyunGi, et al.
Published: (2025)
by: Kim, HyunGi, et al.
Published: (2025)
Adaptive Graph Rewiring to Mitigate Over-Squashing in Mesh-Based GNNs for Fluid Dynamics Simulations
by: Seo, Sangwoo, et al.
Published: (2025)
by: Seo, Sangwoo, et al.
Published: (2025)
CompoDistill: Attention Distillation for Compositional Reasoning in Multimodal LLMs
by: Kim, Jiwan, et al.
Published: (2025)
by: Kim, Jiwan, et al.
Published: (2025)
Beat Tracking as Object Detection
by: Ahn, Jaehoon, et al.
Published: (2025)
by: Ahn, Jaehoon, et al.
Published: (2025)
Infinite-Story: A Training-Free Consistent Text-to-Image Generation
by: Park, Jihun, et al.
Published: (2025)
by: Park, Jihun, et al.
Published: (2025)
Scale-invariant and View-relational Representation Learning for Full Surround Monocular Depth
by: Hwang, Kyumin, et al.
Published: (2025)
by: Hwang, Kyumin, et al.
Published: (2025)
SU(4) Kondo Lattice in Semiconductor Moiré Materials
by: Kim, Sunghoon
Published: (2025)
by: Kim, Sunghoon
Published: (2025)
Korean Books and FRBR: An Investigation
by: Kim, Jeong-Hyen, et al.
Published: (2010)
by: Kim, Jeong-Hyen, et al.
Published: (2010)
Cross, Dwell, or Pinch: Designing and Evaluating Around-Device Selection Methods for Unmodified Smartwatches
by: Kim, Jiwan, et al.
Published: (2025)
by: Kim, Jiwan, et al.
Published: (2025)
REP: Resource-Efficient Prompting for Rehearsal-Free Continual Learning
by: Jeon, Sungho, et al.
Published: (2024)
by: Jeon, Sungho, et al.
Published: (2024)
In Their Own Words: Reasoning Traces Tailored for Small Models Make Them Better Reasoners
by: Kim, Jaehoon, et al.
Published: (2025)
by: Kim, Jaehoon, et al.
Published: (2025)
Ramsey--Dirac theory for bounded degree hypertrees
by: Han, Jie, et al.
Published: (2024)
by: Han, Jie, et al.
Published: (2024)
Application of Preoperative Transarterial Chemoembolization Before Hilar Liver Tumour Resection in a Dog
by: Hyunglak Son, et al.
Published: (2025)
by: Hyunglak Son, et al.
Published: (2025)
Locally Convex Global Loss Network for Decision-Focused Learning
by: Jeon, Haeun, et al.
Published: (2024)
by: Jeon, Haeun, et al.
Published: (2024)
Prediction Loss Guided Decision-Focused Learning
by: Jeon, Haeun, et al.
Published: (2025)
by: Jeon, Haeun, et al.
Published: (2025)
Quantum Support Vector Machine-Based Classification of GPS Signal Reception Conditions
by: Jeong, Suhui, et al.
Published: (2024)
by: Jeong, Suhui, et al.
Published: (2024)
Multi-stream deep learning framework to predict mild cognitive impairment with Rey Complex Figure Test
by: Park, Junyoung, et al.
Published: (2024)
by: Park, Junyoung, et al.
Published: (2024)
Computed Tomographic Analysis of the Anatomical Characteristics of the Canine Prostatic Artery and Development of a Three‐Dimensional Canine Prostate Cancer Model for Simulation of Prostatic Artery Embolization
by: Jaepung Han, et al.
Published: (2025)
by: Jaepung Han, et al.
Published: (2025)
Investigating Long-term Training for Remote Sensing Object Detection
by: Park, JongHyun, et al.
Published: (2024)
by: Park, JongHyun, et al.
Published: (2024)
Inlier-Centric Post-Training Quantization for Object Detection Models
by: Kim, Minsu, et al.
Published: (2026)
by: Kim, Minsu, et al.
Published: (2026)
Designing and Evaluating In-Vehicle Temporal Decoupling Pointing System for Selecting External Object
by: Pyun, Jaehoon, et al.
Published: (2022)
by: Pyun, Jaehoon, et al.
Published: (2022)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
by: Chung, Jiwan, et al.
Published: (2025)
by: Chung, Jiwan, et al.
Published: (2025)
Similar Items
-
CVA: Context-aware Video-text Alignment for Video Temporal Grounding
by: Moon, Sungho, et al.
Published: (2026) -
Latest Object Memory Management for Temporally Consistent Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2025) -
Depth-discriminative Metric Learning for Monocular 3D Object Detection
by: Choi, Wonhyeok, et al.
Published: (2024) -
CAVIS: Context-Aware Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2024) -
Bridging Geometric and Semantic Foundation Models for Generalized Monocular Depth Estimation
by: Ma, Sanggyun, et al.
Published: (2025)