V-CORE: Temporally Consistent Video Understanding for Video-LLM
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kang, Zhengjian, Chen, Qi, Liu, Rui, Mo, Kangtong, Zhang, Xingyu, Deng, Xiaoyu, Zhang, Ye |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PaQ-DETR: Learning Pattern and Quality-Aware Dynamic Queries for Object Detection
von: Kang, Zhengjian, et al.
Veröffentlicht: (2026)
von: Kang, Zhengjian, et al.
Veröffentlicht: (2026)
Dual-R-DETR: Resolving Query Competition with Pairwise Routing in Transformer Decoders
von: Zhang, Ye, et al.
Veröffentlicht: (2025)
von: Zhang, Ye, et al.
Veröffentlicht: (2025)
StableV2V: Stablizing Shape Consistency in Video-to-Video Editing
von: Liu, Chang, et al.
Veröffentlicht: (2024)
von: Liu, Chang, et al.
Veröffentlicht: (2024)
VideoExpert: Augmented LLM for Temporal-Sensitive Video Understanding
von: Zhao, Henghao, et al.
Veröffentlicht: (2025)
von: Zhao, Henghao, et al.
Veröffentlicht: (2025)
CoVis: A Collaborative Framework for Fine-grained Graphic Visual Understanding
von: Deng, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Deng, Xiaoyu, et al.
Veröffentlicht: (2024)
LP-DETR: Layer-wise Progressive Relations for Object Detection
von: Kang, Zhengjian, et al.
Veröffentlicht: (2025)
von: Kang, Zhengjian, et al.
Veröffentlicht: (2025)
Video-QTR: Query-Driven Temporal Reasoning Framework for Lightweight Video Understanding
von: Zhao, Xinkui, et al.
Veröffentlicht: (2025)
von: Zhao, Xinkui, et al.
Veröffentlicht: (2025)
Stereo Any Video: Temporally Consistent Stereo Matching
von: Jing, Junpeng, et al.
Veröffentlicht: (2025)
von: Jing, Junpeng, et al.
Veröffentlicht: (2025)
Edit Temporal-Consistent Videos with Image Diffusion Model
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2023)
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2023)
SlowFocus: Enhancing Fine-grained Temporal Understanding in Video LLM
von: Nie, Ming, et al.
Veröffentlicht: (2026)
von: Nie, Ming, et al.
Veröffentlicht: (2026)
Learning Temporally Consistent Video Depth from Video Diffusion Priors
von: Shao, Jiahao, et al.
Veröffentlicht: (2024)
von: Shao, Jiahao, et al.
Veröffentlicht: (2024)
InstaVSR: Taming Diffusion for Efficient and Temporally Consistent Video Super-Resolution
von: Hu, Jintong, et al.
Veröffentlicht: (2026)
von: Hu, Jintong, et al.
Veröffentlicht: (2026)
When and What: Diffusion-Grounded VideoLLM with Entity Aware Segmentation for Long Video Understanding
von: Fang, Pengcheng, et al.
Veröffentlicht: (2025)
von: Fang, Pengcheng, et al.
Veröffentlicht: (2025)
On the Consistency of Video Large Language Models in Temporal Comprehension
von: Jung, Minjoon, et al.
Veröffentlicht: (2024)
von: Jung, Minjoon, et al.
Veröffentlicht: (2024)
VideoLucy: Deep Memory Backtracking for Long Video Understanding
von: Zuo, Jialong, et al.
Veröffentlicht: (2025)
von: Zuo, Jialong, et al.
Veröffentlicht: (2025)
Training-Free Motion-Guided Video Generation with Enhanced Temporal Consistency Using Motion Consistency Loss
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
Collaborative Temporal Consistency Learning for Point-supervised Natural Language Video Localization
von: Tao, Zhuo, et al.
Veröffentlicht: (2025)
von: Tao, Zhuo, et al.
Veröffentlicht: (2025)
ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation
von: Ren, Weiming, et al.
Veröffentlicht: (2024)
von: Ren, Weiming, et al.
Veröffentlicht: (2024)
VideoITG: Multimodal Video Understanding with Instructed Temporal Grounding
von: Wang, Shihao, et al.
Veröffentlicht: (2025)
von: Wang, Shihao, et al.
Veröffentlicht: (2025)
RelightVid: Temporal-Consistent Diffusion Model for Video Relighting
von: Fang, Ye, et al.
Veröffentlicht: (2025)
von: Fang, Ye, et al.
Veröffentlicht: (2025)
Mixup Helps Understanding Multimodal Video Better
von: Ma, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Ma, Xiaoyu, et al.
Veröffentlicht: (2025)
VideoRefer Suite: Advancing Spatial-Temporal Object Understanding with Video LLM
von: Yuan, Yuqian, et al.
Veröffentlicht: (2024)
von: Yuan, Yuqian, et al.
Veröffentlicht: (2024)
FastInit: Fast Noise Initialization for Temporally Consistent Video Generation
von: Bai, Chengyu, et al.
Veröffentlicht: (2025)
von: Bai, Chengyu, et al.
Veröffentlicht: (2025)
FiLA-Video: Spatio-Temporal Compression for Fine-Grained Long Video Understanding
von: Guo, Yanan, et al.
Veröffentlicht: (2025)
von: Guo, Yanan, et al.
Veröffentlicht: (2025)
Temporal2Seq: A Unified Framework for Temporal Video Understanding Tasks
von: Yang, Min, et al.
Veröffentlicht: (2024)
von: Yang, Min, et al.
Veröffentlicht: (2024)
Enhancing Temporal Consistency in Video Editing by Reconstructing Videos with 3D Gaussian Splatting
von: Shin, Inkyu, et al.
Veröffentlicht: (2024)
von: Shin, Inkyu, et al.
Veröffentlicht: (2024)
TSPO: Temporal Sampling Policy Optimization for Long-form Video Language Understanding
von: Tang, Canhui, et al.
Veröffentlicht: (2025)
von: Tang, Canhui, et al.
Veröffentlicht: (2025)
DRFusion: Drift-Resilient Temporally Consistent Infrared-Visible Video Fusion
von: Li, Xingyuan, et al.
Veröffentlicht: (2026)
von: Li, Xingyuan, et al.
Veröffentlicht: (2026)
Vectorized Video Representation with Easy Editing via Hierarchical Spatio-Temporally Consistent Proxy Embedding
von: Chen, Ye, et al.
Veröffentlicht: (2025)
von: Chen, Ye, et al.
Veröffentlicht: (2025)
Temporal-Consistent Video Restoration with Pre-trained Diffusion Models
von: Wang, Hengkang, et al.
Veröffentlicht: (2025)
von: Wang, Hengkang, et al.
Veröffentlicht: (2025)
WaterWave: Bridging Underwater Image Enhancement into Video Streams via Wavelet-based Temporal Consistency Field
von: Zhu, Qi, et al.
Veröffentlicht: (2025)
von: Zhu, Qi, et al.
Veröffentlicht: (2025)
VideoCompressa: Data-Efficient Video Understanding via Joint Temporal Compression and Spatial Reconstruction
von: Wang, Shaobo, et al.
Veröffentlicht: (2025)
von: Wang, Shaobo, et al.
Veröffentlicht: (2025)
MotionLLM: Understanding Human Behaviors from Human Motions and Videos
von: Chen, Ling-Hao, et al.
Veröffentlicht: (2024)
von: Chen, Ling-Hao, et al.
Veröffentlicht: (2024)
Spatial Degradation-Aware and Temporal Consistent Diffusion Model for Compressed Video Super-Resolution
von: An, Hongyu, et al.
Veröffentlicht: (2025)
von: An, Hongyu, et al.
Veröffentlicht: (2025)
TROPHIES: Temporal Reconstruction of Places, Humans, and Cameras from Multi-view Videos
von: Liu, Jinpeng, et al.
Veröffentlicht: (2026)
von: Liu, Jinpeng, et al.
Veröffentlicht: (2026)
DropletVideo: A Dataset and Approach to Explore Integral Spatio-Temporal Consistent Video Generation
von: Zhang, Runze, et al.
Veröffentlicht: (2025)
von: Zhang, Runze, et al.
Veröffentlicht: (2025)
Incentivizing Temporal-Awareness in Egocentric Video Understanding Models
von: Xu, Zhiyang, et al.
Veröffentlicht: (2026)
von: Xu, Zhiyang, et al.
Veröffentlicht: (2026)
VTG-LLM: Integrating Timestamp Knowledge into Video LLMs for Enhanced Video Temporal Grounding
von: Guo, Yongxin, et al.
Veröffentlicht: (2024)
von: Guo, Yongxin, et al.
Veröffentlicht: (2024)
STAR: Spatial-Temporal Augmentation with Text-to-Video Models for Real-World Video Super-Resolution
von: Xie, Rui, et al.
Veröffentlicht: (2025)
von: Xie, Rui, et al.
Veröffentlicht: (2025)
TimeSearch-R: Adaptive Temporal Search for Long-Form Video Understanding via Self-Verification Reinforcement Learning
von: Pan, Junwen, et al.
Veröffentlicht: (2025)
von: Pan, Junwen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PaQ-DETR: Learning Pattern and Quality-Aware Dynamic Queries for Object Detection
von: Kang, Zhengjian, et al.
Veröffentlicht: (2026) -
Dual-R-DETR: Resolving Query Competition with Pairwise Routing in Transformer Decoders
von: Zhang, Ye, et al.
Veröffentlicht: (2025) -
StableV2V: Stablizing Shape Consistency in Video-to-Video Editing
von: Liu, Chang, et al.
Veröffentlicht: (2024) -
VideoExpert: Augmented LLM for Temporal-Sensitive Video Understanding
von: Zhao, Henghao, et al.
Veröffentlicht: (2025) -
CoVis: A Collaborative Framework for Fine-grained Graphic Visual Understanding
von: Deng, Xiaoyu, et al.
Veröffentlicht: (2024)