Saved in:
| Main Authors: | Zhang, Zheyuan, Tang, Weihao, Chen, Hong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2508.06640 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Koala: Key frame-conditioned long video-LLM
by: Tan, Reuben, et al.
Published: (2024)
by: Tan, Reuben, et al.
Published: (2024)
UCVC: A Unified Contextual Video Compression Framework with Joint P-frame and B-frame Coding
by: Yang, Jiayu, et al.
Published: (2024)
by: Yang, Jiayu, et al.
Published: (2024)
DFME: A New Benchmark for Dynamic Facial Micro-expression Recognition
by: Zhao, Sirui, et al.
Published: (2023)
by: Zhao, Sirui, et al.
Published: (2023)
A Benchmark for Incremental Micro-expression Recognition
by: Lai, Zhengqin, et al.
Published: (2025)
by: Lai, Zhengqin, et al.
Published: (2025)
D$^{2}$-VPR: A Parameter-efficient Visual-foundation-model-based Visual Place Recognition Method via Knowledge Distillation and Deformable Aggregation
by: Zhang, Zheyuan, et al.
Published: (2025)
by: Zhang, Zheyuan, et al.
Published: (2025)
Mono-ViFI: A Unified Learning Framework for Self-supervised Single- and Multi-frame Monocular Depth Estimation
by: Liu, Jinfeng, et al.
Published: (2024)
by: Liu, Jinfeng, et al.
Published: (2024)
EPIR: An Efficient Patch Tokenization, Integration and Representation Framework for Micro-expression Recognition
by: Wang, Junbo, et al.
Published: (2026)
by: Wang, Junbo, et al.
Published: (2026)
Low-Level Matters: An Efficient Hybrid Architecture for Robust Multi-frame Infrared Small Target Detection
by: Shen, Zhihua, et al.
Published: (2025)
by: Shen, Zhihua, et al.
Published: (2025)
Micro-expression Recognition Based on Dual-branch Feature Extraction and Fusion
by: Zhang, Mingjie, et al.
Published: (2026)
by: Zhang, Mingjie, et al.
Published: (2026)
Continuous Sign Language Recognition Based on Motor attention mechanism and frame-level Self-distillation
by: Zhu, Qidan, et al.
Published: (2024)
by: Zhu, Qidan, et al.
Published: (2024)
Adaptive Temporal Motion Guided Graph Convolution Network for Micro-expression Recognition
by: Zhang, Fengyuan, et al.
Published: (2024)
by: Zhang, Fengyuan, et al.
Published: (2024)
Rethinking the Key Factors for the Generalization of Remote Sensing Stereo Matching Networks
by: Jiang, Liting, et al.
Published: (2024)
by: Jiang, Liting, et al.
Published: (2024)
Extending Video Masked Autoencoders to 128 frames
by: Gundavarapu, Nitesh Bharadwaj, et al.
Published: (2024)
by: Gundavarapu, Nitesh Bharadwaj, et al.
Published: (2024)
DENOISER: Rethinking the Robustness for Open-Vocabulary Action Recognition
by: Cheng, Haozhe, et al.
Published: (2024)
by: Cheng, Haozhe, et al.
Published: (2024)
VS3R: Robust Full-frame Video Stabilization via Deep 3D Reconstruction
by: Zhu, Muhua, et al.
Published: (2026)
by: Zhu, Muhua, et al.
Published: (2026)
RDTF: Resource-efficient Dual-mask Training Framework for Multi-frame Animated Sticker Generation
by: Yuan, Zhiqiang, et al.
Published: (2025)
by: Yuan, Zhiqiang, et al.
Published: (2025)
OmniParser: A Unified Framework for Text Spotting, Key Information Extraction and Table Recognition
by: Wan, Jianqiang, et al.
Published: (2024)
by: Wan, Jianqiang, et al.
Published: (2024)
Efficient motion-based metrics for video frame interpolation
by: Daly, Conall, et al.
Published: (2025)
by: Daly, Conall, et al.
Published: (2025)
FMANet: A Novel Dual-Phase Optical Flow Approach with Fusion Motion Attention Network for Robust Micro-expression Recognition
by: Nguyen, Luu Tu, et al.
Published: (2025)
by: Nguyen, Luu Tu, et al.
Published: (2025)
KeyPoint Relative Position Encoding for Face Recognition
by: Kim, Minchul, et al.
Published: (2024)
by: Kim, Minchul, et al.
Published: (2024)
DESSERT: Diffusion-based Event-driven Single-frame Synthesis via Residual Training
by: Kong, Jiyun, et al.
Published: (2025)
by: Kong, Jiyun, et al.
Published: (2025)
SplatFlow: Learning Multi-frame Optical Flow via Splatting
by: Wang, Bo, et al.
Published: (2023)
by: Wang, Bo, et al.
Published: (2023)
Space Object Detection using Multi-frame Temporal Trajectory Completion Method
by: Lan, Xiaoqing, et al.
Published: (2025)
by: Lan, Xiaoqing, et al.
Published: (2025)
DIFEM: Key-points Interaction based Feature Extraction Module for Violence Recognition in Videos
by: Mittal, Himanshu, et al.
Published: (2024)
by: Mittal, Himanshu, et al.
Published: (2024)
WVSC: Wireless Video Semantic Communication with Multi-frame Compensation
by: Xie, Bingyan, et al.
Published: (2025)
by: Xie, Bingyan, et al.
Published: (2025)
MFSeg: Efficient Multi-frame 3D Semantic Segmentation
by: Huang, Chengjie, et al.
Published: (2025)
by: Huang, Chengjie, et al.
Published: (2025)
Hierarchical B-frame Video Coding for Long Group of Pictures
by: Kirillov, Ivan, et al.
Published: (2024)
by: Kirillov, Ivan, et al.
Published: (2024)
PEEK: Picking Essential frames via Efficient Knowledge distillation
by: Steunou, Killian, et al.
Published: (2026)
by: Steunou, Killian, et al.
Published: (2026)
DeltaFlow: An Efficient Multi-frame Scene Flow Estimation Method
by: Zhang, Qingwen, et al.
Published: (2025)
by: Zhang, Qingwen, et al.
Published: (2025)
Improving Video Question Answering through query-based frame selection
by: Patil, Himanshu, et al.
Published: (2026)
by: Patil, Himanshu, et al.
Published: (2026)
RingID: Rethinking Tree-Ring Watermarking for Enhanced Multi-Key Identification
by: Ci, Hai, et al.
Published: (2024)
by: Ci, Hai, et al.
Published: (2024)
eWand: A calibration framework for wide baseline frame-based and event-based camera systems
by: Gossard, Thomas, et al.
Published: (2023)
by: Gossard, Thomas, et al.
Published: (2023)
Revealing Key Details to See Differences: A Novel Prototypical Perspective for Skeleton-based Action Recognition
by: Liu, Hongda, et al.
Published: (2024)
by: Liu, Hongda, et al.
Published: (2024)
A Robust Low-Rank Prior Model for Structured Cartoon-Texture Image Decomposition with Heavy-Tailed Noise
by: Tang, Weihao, et al.
Published: (2026)
by: Tang, Weihao, et al.
Published: (2026)
Mesh deformation-based single-view 3D reconstruction of thin eyeglasses frames with differentiable rendering
by: Zhang, Fan, et al.
Published: (2024)
by: Zhang, Fan, et al.
Published: (2024)
Dual-frame Fluid Motion Estimation with Test-time Optimization and Zero-divergence Loss
by: Zhang, Yifei, et al.
Published: (2024)
by: Zhang, Yifei, et al.
Published: (2024)
Key Patches Are All You Need: A Multiple Instance Learning Framework For Robust Medical Diagnosis
by: Araújo, Diogo J., et al.
Published: (2024)
by: Araújo, Diogo J., et al.
Published: (2024)
Neural B-frame Video Compression with Bi-directional Reference Harmonization
by: Liu, Yuxi, et al.
Published: (2025)
by: Liu, Yuxi, et al.
Published: (2025)
Learning Multi-frame and Monocular Prior for Estimating Geometry in Dynamic Scenes
by: Park, Seong Hyeon, et al.
Published: (2025)
by: Park, Seong Hyeon, et al.
Published: (2025)
Beyond sparse denoising in frames: minimax estimation with a scattering transform
by: Cuvelle--Magar, Nathanaël, et al.
Published: (2025)
by: Cuvelle--Magar, Nathanaël, et al.
Published: (2025)
Similar Items
-
Koala: Key frame-conditioned long video-LLM
by: Tan, Reuben, et al.
Published: (2024) -
UCVC: A Unified Contextual Video Compression Framework with Joint P-frame and B-frame Coding
by: Yang, Jiayu, et al.
Published: (2024) -
DFME: A New Benchmark for Dynamic Facial Micro-expression Recognition
by: Zhao, Sirui, et al.
Published: (2023) -
A Benchmark for Incremental Micro-expression Recognition
by: Lai, Zhengqin, et al.
Published: (2025) -
D$^{2}$-VPR: A Parameter-efficient Visual-foundation-model-based Visual Place Recognition Method via Knowledge Distillation and Deformable Aggregation
by: Zhang, Zheyuan, et al.
Published: (2025)