Gespeichert in:
| Hauptverfasser: | Wang, Yuan, Liao, Borui, Huang, Huijuan, Lu, Jinda, Li, Ouxiang, Liu, Kuien, Wang, Meng, Wang, Xiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2601.04033 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling
von: Wang, Yuan, et al.
Veröffentlicht: (2026)
von: Wang, Yuan, et al.
Veröffentlicht: (2026)
Precise, Fast, and Low-cost Concept Erasure in Value Space: Orthogonal Complement Matters
von: Wang, Yuan, et al.
Veröffentlicht: (2024)
von: Wang, Yuan, et al.
Veröffentlicht: (2024)
Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models
von: Liu, Zijian, et al.
Veröffentlicht: (2026)
von: Liu, Zijian, et al.
Veröffentlicht: (2026)
FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting
von: He, Zefeng, et al.
Veröffentlicht: (2025)
von: He, Zefeng, et al.
Veröffentlicht: (2025)
Easier Painting Than Thinking: Can Text-to-Image Models Set the Stage, but Not Direct the Play?
von: Li, Ouxiang, et al.
Veröffentlicht: (2025)
von: Li, Ouxiang, et al.
Veröffentlicht: (2025)
Beyond Where to Look: Trajectory-Guided Reinforcement Learning for Multimodal RLVR
von: Lu, Jinda, et al.
Veröffentlicht: (2026)
von: Lu, Jinda, et al.
Veröffentlicht: (2026)
When Thinking Hurts: Mitigating Visual Forgetting in Video Reasoning via Frame Repetition
von: Sun, Xiaokun, et al.
Veröffentlicht: (2026)
von: Sun, Xiaokun, et al.
Veröffentlicht: (2026)
Frame-Level Captions for Long Video Generation with Complex Multi Scenes
von: Zheng, Guangcong, et al.
Veröffentlicht: (2025)
von: Zheng, Guangcong, et al.
Veröffentlicht: (2025)
Self-supervised Learning of Event-guided Video Frame Interpolation for Rolling Shutter Frames
von: Lu, Yunfan, et al.
Veröffentlicht: (2023)
von: Lu, Yunfan, et al.
Veröffentlicht: (2023)
VideoAR: Autoregressive Video Generation via Next-Frame & Scale Prediction
von: Ji, Longbin, et al.
Veröffentlicht: (2026)
von: Ji, Longbin, et al.
Veröffentlicht: (2026)
TempoMaster: Efficient Long Video Generation via Next-Frame-Rate Prediction
von: Ma, Yukuo, et al.
Veröffentlicht: (2025)
von: Ma, Yukuo, et al.
Veröffentlicht: (2025)
Detecting AI-Generated Video via Frame Consistency
von: Ma, Long, et al.
Veröffentlicht: (2024)
von: Ma, Long, et al.
Veröffentlicht: (2024)
Autoregressive Video Generation beyond Next Frames Prediction
von: Ren, Sucheng, et al.
Veröffentlicht: (2025)
von: Ren, Sucheng, et al.
Veröffentlicht: (2025)
FrameMind: Frame-Interleaved Video Reasoning via Reinforcement Learning
von: Ge, Haonan, et al.
Veröffentlicht: (2025)
von: Ge, Haonan, et al.
Veröffentlicht: (2025)
Video Frame Interpolation for Polarization via Swin-Transformer
von: Huang, Feng, et al.
Veröffentlicht: (2024)
von: Huang, Feng, et al.
Veröffentlicht: (2024)
VGDFR: Diffusion-based Video Generation with Dynamic Latent Frame Rate
von: Yuan, Zhihang, et al.
Veröffentlicht: (2025)
von: Yuan, Zhihang, et al.
Veröffentlicht: (2025)
Rethinking Visual Content Refinement in Low-Shot CLIP Adaptation
von: Lu, Jinda, et al.
Veröffentlicht: (2024)
von: Lu, Jinda, et al.
Veröffentlicht: (2024)
360VFI: A Dataset and Benchmark for Omnidirectional Video Frame Interpolation
von: Lu, Wenxuan, et al.
Veröffentlicht: (2024)
von: Lu, Wenxuan, et al.
Veröffentlicht: (2024)
Perception-Oriented Video Frame Interpolation via Asymmetric Blending
von: Wu, Guangyang, et al.
Veröffentlicht: (2024)
von: Wu, Guangyang, et al.
Veröffentlicht: (2024)
Velocity Disambiguation for Video Frame Interpolation
von: Zhong, Zhihang, et al.
Veröffentlicht: (2023)
von: Zhong, Zhihang, et al.
Veröffentlicht: (2023)
Frame by Familiar Frame: Understanding Replication in Video Diffusion Models
von: Rahman, Aimon, et al.
Veröffentlicht: (2024)
von: Rahman, Aimon, et al.
Veröffentlicht: (2024)
Frame-Voyager: Learning to Query Frames for Video Large Language Models
von: Yu, Sicheng, et al.
Veröffentlicht: (2024)
von: Yu, Sicheng, et al.
Veröffentlicht: (2024)
Frame In-N-Out: Unbounded Controllable Image-to-Video Generation
von: Wang, Boyang, et al.
Veröffentlicht: (2025)
von: Wang, Boyang, et al.
Veröffentlicht: (2025)
Generative Inbetweening through Frame-wise Conditions-Driven Video Generation
von: Zhu, Tianyi, et al.
Veröffentlicht: (2024)
von: Zhu, Tianyi, et al.
Veröffentlicht: (2024)
DLFR-VAE: Dynamic Latent Frame Rate VAE for Video Generation
von: Yuan, Zhihang, et al.
Veröffentlicht: (2025)
von: Yuan, Zhihang, et al.
Veröffentlicht: (2025)
Motion-aware Latent Diffusion Models for Video Frame Interpolation
von: Huang, Zhilin, et al.
Veröffentlicht: (2024)
von: Huang, Zhilin, et al.
Veröffentlicht: (2024)
Beyond the Last Frame: Process-aware Evaluation for Generative Video Reasoning
von: Li, Yifan, et al.
Veröffentlicht: (2025)
von: Li, Yifan, et al.
Veröffentlicht: (2025)
VFIMamba: Video Frame Interpolation with State Space Models
von: Zhang, Guozhen, et al.
Veröffentlicht: (2024)
von: Zhang, Guozhen, et al.
Veröffentlicht: (2024)
STORYANCHORS: Generating Consistent Multi-Scene Story Frames for Long-Form Narratives
von: Wang, Bo, et al.
Veröffentlicht: (2025)
von: Wang, Bo, et al.
Veröffentlicht: (2025)
InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing
von: Yang, Shaoshu, et al.
Veröffentlicht: (2025)
von: Yang, Shaoshu, et al.
Veröffentlicht: (2025)
Sparse Global Matching for Video Frame Interpolation with Large Motion
von: Liu, Chunxu, et al.
Veröffentlicht: (2024)
von: Liu, Chunxu, et al.
Veröffentlicht: (2024)
Benchmarking Video Frame Interpolation
von: Kiefhaber, Simon, et al.
Veröffentlicht: (2024)
von: Kiefhaber, Simon, et al.
Veröffentlicht: (2024)
Chain-of-Frames: Advancing Video Understanding in Multimodal LLMs via Frame-Aware Reasoning
von: Ghazanfari, Sara, et al.
Veröffentlicht: (2025)
von: Ghazanfari, Sara, et al.
Veröffentlicht: (2025)
Mamba-FETrack: Frame-Event Tracking via State Space Model
von: Huang, Ju, et al.
Veröffentlicht: (2024)
von: Huang, Ju, et al.
Veröffentlicht: (2024)
End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling
von: Liang, Jianxin, et al.
Veröffentlicht: (2024)
von: Liang, Jianxin, et al.
Veröffentlicht: (2024)
FrameBridge: Improving Image-to-Video Generation with Bridge Models
von: Wang, Yuji, et al.
Veröffentlicht: (2024)
von: Wang, Yuji, et al.
Veröffentlicht: (2024)
Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models
von: Zhang, Lvmin, et al.
Veröffentlicht: (2025)
von: Zhang, Lvmin, et al.
Veröffentlicht: (2025)
DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes
von: Song, Zhende, et al.
Veröffentlicht: (2024)
von: Song, Zhende, et al.
Veröffentlicht: (2024)
DreaMontage: Arbitrary Frame-Guided One-Shot Video Generation
von: Liu, Jiawei, et al.
Veröffentlicht: (2025)
von: Liu, Jiawei, et al.
Veröffentlicht: (2025)
Think-Clip-Sample: Slow-Fast Frame Selection for Video Understanding
von: Tan, Wenhui, et al.
Veröffentlicht: (2026)
von: Tan, Wenhui, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Think, then Score: Decoupled Reasoning and Scoring for Video Reward Modeling
von: Wang, Yuan, et al.
Veröffentlicht: (2026) -
Precise, Fast, and Low-cost Concept Erasure in Value Space: Orthogonal Complement Matters
von: Wang, Yuan, et al.
Veröffentlicht: (2024) -
Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models
von: Liu, Zijian, et al.
Veröffentlicht: (2026) -
FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting
von: He, Zefeng, et al.
Veröffentlicht: (2025) -
Easier Painting Than Thinking: Can Text-to-Image Models Set the Stage, but Not Direct the Play?
von: Li, Ouxiang, et al.
Veröffentlicht: (2025)