Memory-Anchored Multimodal Reasoning for Explainable Video Forensics
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Chen, Li, Runze, Zhang, Zejun, Zhao, Pukun, Zhou, Fanqing, Wang, Longxiang, Huang, Haojian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HippoMM: Hippocampal-inspired Multimodal Memory for Long Audiovisual Event Understanding
by: Lin, Yueqian, et al.
Published: (2025)
by: Lin, Yueqian, et al.
Published: (2025)
Audio-Visual Cross-Modal Compression for Generative Face Video Coding
by: Xu, Youmin, et al.
Published: (2025)
by: Xu, Youmin, et al.
Published: (2025)
Progressive Frame Patching for FoV-based Point Cloud Video Streaming
by: Zong, Tongyu, et al.
Published: (2023)
by: Zong, Tongyu, et al.
Published: (2023)
QoE Optimization for Semantic Self-Correcting Video Transmission in Multi-UAV Networks
by: Chen, Xuyang, et al.
Published: (2025)
by: Chen, Xuyang, et al.
Published: (2025)
Perception-Aware Video Semantic Communication
by: Huang, Yinhuan, et al.
Published: (2026)
by: Huang, Yinhuan, et al.
Published: (2026)
Tube-Structured Incremental Semantic HARQ for Generative Video Receivers
by: Wang, Xuesong, et al.
Published: (2026)
by: Wang, Xuesong, et al.
Published: (2026)
Towards User-level QoE: Large-scale Practice in Personalized Optimization of Adaptive Video Streaming
by: Jia, Lianchen, et al.
Published: (2025)
by: Jia, Lianchen, et al.
Published: (2025)
EvEnhancer: Empowering Effectiveness, Efficiency and Generalizability for Continuous Space-Time Video Super-Resolution with Events
by: Wei, Shuoyan, et al.
Published: (2025)
by: Wei, Shuoyan, et al.
Published: (2025)
H.265/HEVC Video Steganalysis Based on CU Block Structure Gradients and IPM Mapping
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
Camel: Frame-Level Bandwidth Estimation for Low-Latency Live Streaming under Video Bitrate Undershooting
by: Liu, Liming, et al.
Published: (2026)
by: Liu, Liming, et al.
Published: (2026)
Prompt-based Multimodal Semantic Communication for Multi-spectral Image Segmentation
by: Zhang, Haoshuo, et al.
Published: (2025)
by: Zhang, Haoshuo, et al.
Published: (2025)
Symmetric Entropy-Constrained Video Coding for Machines
by: Sun, Yuxiao, et al.
Published: (2025)
by: Sun, Yuxiao, et al.
Published: (2025)
Smaller is Better: Generative Models Can Power Short Video Preloading
by: Liu, Liming, et al.
Published: (2026)
by: Liu, Liming, et al.
Published: (2026)
Rate-Quality or Energy-Quality Pareto Fronts for Adaptive Video Streaming?
by: Katsenou, Angeliki, et al.
Published: (2024)
by: Katsenou, Angeliki, et al.
Published: (2024)
A H.265/HEVC Fine-Grained ROI Video Encryption Algorithm Based on Coding Unit and Prompt Segmentation
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
Optimizing Mobile-Friendly Viewport Prediction for Live 360-Degree Video Streaming
by: Zhang, Lei, et al.
Published: (2024)
by: Zhang, Lei, et al.
Published: (2024)
Interactive $360^{\circ}$ Video Streaming Using FoV-Adaptive Coding with Temporal Prediction
by: Mao, Yixiang, et al.
Published: (2024)
by: Mao, Yixiang, et al.
Published: (2024)
A Versatile Depth Video Encoding Scheme Based on Low-rank Tensor Modeling for Free Viewpoint Video
by: Sharma, Mansi, et al.
Published: (2021)
by: Sharma, Mansi, et al.
Published: (2021)
Generative Flow Networks for Personalized Multimedia Systems: A Case Study on Short Video Feeds
by: Jin, Yili, et al.
Published: (2025)
by: Jin, Yili, et al.
Published: (2025)
Generative Video Compression: Towards 0.01% Compression Rate for Video Transmission
by: Chen, Xiangyu, et al.
Published: (2025)
by: Chen, Xiangyu, et al.
Published: (2025)
Compact Visual Data Representation for Green Multimedia -- A Human Visual System Perspective
by: Chen, Peilin, et al.
Published: (2024)
by: Chen, Peilin, et al.
Published: (2024)
Efficient Sub-pixel Motion Compensation in Learned Video Codecs
by: Ladune, Théo, et al.
Published: (2025)
by: Ladune, Théo, et al.
Published: (2025)
Fast Multirate Encoding for 360° Video in OMAF Streaming Workflows
by: Premkumar, Amritha, et al.
Published: (2026)
by: Premkumar, Amritha, et al.
Published: (2026)
Adaptive Resolution and Chroma Subsampling for Energy-Efficient Video Coding
by: Premkumar, Amritha, et al.
Published: (2026)
by: Premkumar, Amritha, et al.
Published: (2026)
Convex-hull Estimation using XPSNR for Versatile Video Coding
by: Menon, Vignesh V, et al.
Published: (2024)
by: Menon, Vignesh V, et al.
Published: (2024)
Rethinking Security of Diffusion-based Generative Steganography
by: Zhu, Jihao, et al.
Published: (2026)
by: Zhu, Jihao, et al.
Published: (2026)
Neural Compression of 360-Degree Equirectangular Videos using Quality Parameter Adaptation
by: Arai, Daichi, et al.
Published: (2025)
by: Arai, Daichi, et al.
Published: (2025)
Encoding Time and Energy Model for SVT-AV1 based on Video Complexity
by: Eichermüller, Lena, et al.
Published: (2024)
by: Eichermüller, Lena, et al.
Published: (2024)
Content-Driven Frame-Level Bit Prediction for Rate Control in Versatile Video Coding
by: Premkumar, Amritha, et al.
Published: (2026)
by: Premkumar, Amritha, et al.
Published: (2026)
DiV-INR: Extreme Low-Bitrate Diffusion Video Compression with INR Conditioning
by: Çetin, Eren, et al.
Published: (2026)
by: Çetin, Eren, et al.
Published: (2026)
DQ-Ladder: A Deep Reinforcement Learning-based Bitrate Ladder for Adaptive Video Streaming
by: Farahani, Reza, et al.
Published: (2026)
by: Farahani, Reza, et al.
Published: (2026)
Video Compression Beyond VVC: Quantitative Analysis of Intra Coding Tools in Enhanced Compression Model (ECM)
by: Abdoli, Mohsen, et al.
Published: (2024)
by: Abdoli, Mohsen, et al.
Published: (2024)
Beyond Interpretability: Exploring the Comprehensibility of Adaptive Video Streaming through Large Language Models
by: Jia, Lianchen, et al.
Published: (2025)
by: Jia, Lianchen, et al.
Published: (2025)
Robust Live Streaming over LEO Satellite Constellations: Measurement, Analysis, and Handover-Aware Adaptation
by: Fang, Hao, et al.
Published: (2025)
by: Fang, Hao, et al.
Published: (2025)
TVMC: Time-Varying Mesh Compression via Multi-Stage Anchor Mesh Generation
by: Huang, He, et al.
Published: (2025)
by: Huang, He, et al.
Published: (2025)
Dynamic resolution switching for live streaming
by: Xiong, Xin, et al.
Published: (2026)
by: Xiong, Xin, et al.
Published: (2026)
HybridPrompt: Bridging Generative Priors and Traditional Codecs for Mobile Streaming
by: Liu, Liming, et al.
Published: (2026)
by: Liu, Liming, et al.
Published: (2026)
Enhanced Quality Aware-Scalable Underwater Image Compression
by: Zhu, Linwei, et al.
Published: (2025)
by: Zhu, Linwei, et al.
Published: (2025)
Dual Inverse Degradation Network for Real-World SDRTV-to-HDRTV Conversion
by: Xu, Kepeng, et al.
Published: (2023)
by: Xu, Kepeng, et al.
Published: (2023)
Cross-Layer Encrypted Semantic Communication Framework for Panoramic Video Transmission
by: Gao, Haixiao, et al.
Published: (2024)
by: Gao, Haixiao, et al.
Published: (2024)
Similar Items
-
HippoMM: Hippocampal-inspired Multimodal Memory for Long Audiovisual Event Understanding
by: Lin, Yueqian, et al.
Published: (2025) -
Audio-Visual Cross-Modal Compression for Generative Face Video Coding
by: Xu, Youmin, et al.
Published: (2025) -
Progressive Frame Patching for FoV-based Point Cloud Video Streaming
by: Zong, Tongyu, et al.
Published: (2023) -
QoE Optimization for Semantic Self-Correcting Video Transmission in Multi-UAV Networks
by: Chen, Xuyang, et al.
Published: (2025) -
Perception-Aware Video Semantic Communication
by: Huang, Yinhuan, et al.
Published: (2026)