HippoMM: Hippocampal-inspired Multimodal Memory for Long Audiovisual Event Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Yueqian, Zhang, Jingyang, Wang, Qinsi, Ye, Hancheng, Fu, Yuzhe, Liu, Yudong, Li, Hai "Helen", Chen, Yiran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Memory-Anchored Multimodal Reasoning for Explainable Video Forensics
von: Chen, Chen, et al.
Veröffentlicht: (2025)
von: Chen, Chen, et al.
Veröffentlicht: (2025)
EvEnhancer: Empowering Effectiveness, Efficiency and Generalizability for Continuous Space-Time Video Super-Resolution with Events
von: Wei, Shuoyan, et al.
Veröffentlicht: (2025)
von: Wei, Shuoyan, et al.
Veröffentlicht: (2025)
Prompt-based Multimodal Semantic Communication for Multi-spectral Image Segmentation
von: Zhang, Haoshuo, et al.
Veröffentlicht: (2025)
von: Zhang, Haoshuo, et al.
Veröffentlicht: (2025)
Dynamic resolution switching for live streaming
von: Xiong, Xin, et al.
Veröffentlicht: (2026)
von: Xiong, Xin, et al.
Veröffentlicht: (2026)
H.265/HEVC Video Steganalysis Based on CU Block Structure Gradients and IPM Mapping
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
A H.265/HEVC Fine-Grained ROI Video Encryption Algorithm Based on Coding Unit and Prompt Segmentation
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
Enhanced Quality Aware-Scalable Underwater Image Compression
von: Zhu, Linwei, et al.
Veröffentlicht: (2025)
von: Zhu, Linwei, et al.
Veröffentlicht: (2025)
Enhanced Template-based Intra Mode Derivation with Adaptive Block Vector Replacement
von: Zhang, Jiaqi, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaqi, et al.
Veröffentlicht: (2025)
Robust Multi-modal Task-oriented Communications with Redundancy-aware Representations
von: Fu, Jingwen, et al.
Veröffentlicht: (2025)
von: Fu, Jingwen, et al.
Veröffentlicht: (2025)
Robust Live Streaming over LEO Satellite Constellations: Measurement, Analysis, and Handover-Aware Adaptation
von: Fang, Hao, et al.
Veröffentlicht: (2025)
von: Fang, Hao, et al.
Veröffentlicht: (2025)
Transform and Entropy Coding in AV2
von: Nalci, Alican, et al.
Veröffentlicht: (2026)
von: Nalci, Alican, et al.
Veröffentlicht: (2026)
Tube-Structured Incremental Semantic HARQ for Generative Video Receivers
von: Wang, Xuesong, et al.
Veröffentlicht: (2026)
von: Wang, Xuesong, et al.
Veröffentlicht: (2026)
Rate-Quality or Energy-Quality Pareto Fronts for Adaptive Video Streaming?
von: Katsenou, Angeliki, et al.
Veröffentlicht: (2024)
von: Katsenou, Angeliki, et al.
Veröffentlicht: (2024)
Smaller is Better: Generative Models Can Power Short Video Preloading
von: Liu, Liming, et al.
Veröffentlicht: (2026)
von: Liu, Liming, et al.
Veröffentlicht: (2026)
Camel: Frame-Level Bandwidth Estimation for Low-Latency Live Streaming under Video Bitrate Undershooting
von: Liu, Liming, et al.
Veröffentlicht: (2026)
von: Liu, Liming, et al.
Veröffentlicht: (2026)
NiMark: A Non-intrusive Watermarking Framework against Screen-shooting Attacks
von: Wu, Yufeng, et al.
Veröffentlicht: (2026)
von: Wu, Yufeng, et al.
Veröffentlicht: (2026)
Fast Multirate Encoding for 360° Video in OMAF Streaming Workflows
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
Decoding Complexity-Rate-Quality Pareto-Front for Adaptive VVC Streaming
von: Katsenou, Angeliki, et al.
Veröffentlicht: (2024)
von: Katsenou, Angeliki, et al.
Veröffentlicht: (2024)
GScomp-QA: A Subjective Dataset for Quality Assessment of Compressed Gaussian Splatting
von: Martin, Pedro, et al.
Veröffentlicht: (2026)
von: Martin, Pedro, et al.
Veröffentlicht: (2026)
Unravelling the Power of Single-Pass Look-Ahead in Modern Codecs for Optimized Transcoding Deployment
von: Vibhoothi, Vibhoothi, et al.
Veröffentlicht: (2024)
von: Vibhoothi, Vibhoothi, et al.
Veröffentlicht: (2024)
Perception-Aware Video Semantic Communication
von: Huang, Yinhuan, et al.
Veröffentlicht: (2026)
von: Huang, Yinhuan, et al.
Veröffentlicht: (2026)
Video Compression Beyond VVC: Quantitative Analysis of Intra Coding Tools in Enhanced Compression Model (ECM)
von: Abdoli, Mohsen, et al.
Veröffentlicht: (2024)
von: Abdoli, Mohsen, et al.
Veröffentlicht: (2024)
Generative Flow Networks for Personalized Multimedia Systems: A Case Study on Short Video Feeds
von: Jin, Yili, et al.
Veröffentlicht: (2025)
von: Jin, Yili, et al.
Veröffentlicht: (2025)
HybridPrompt: Bridging Generative Priors and Traditional Codecs for Mobile Streaming
von: Liu, Liming, et al.
Veröffentlicht: (2026)
von: Liu, Liming, et al.
Veröffentlicht: (2026)
Partition Tree Search Acceleration for VVC: Survey and Evaluation with VTM Evolution
von: Kherchouche, M. E. A., et al.
Veröffentlicht: (2026)
von: Kherchouche, M. E. A., et al.
Veröffentlicht: (2026)
Sandwiched Compression: Repurposing Standard Codecs with Neural Network Wrappers
von: Guleryuz, Onur G., et al.
Veröffentlicht: (2024)
von: Guleryuz, Onur G., et al.
Veröffentlicht: (2024)
Encoding Time and Energy Model for SVT-AV1 based on Video Complexity
von: Eichermüller, Lena, et al.
Veröffentlicht: (2024)
von: Eichermüller, Lena, et al.
Veröffentlicht: (2024)
Interactive $360^{\circ}$ Video Streaming Using FoV-Adaptive Coding with Temporal Prediction
von: Mao, Yixiang, et al.
Veröffentlicht: (2024)
von: Mao, Yixiang, et al.
Veröffentlicht: (2024)
Maximum entropy and quantized metric models for absolute category ratings
von: Saupe, Dietmar, et al.
Veröffentlicht: (2024)
von: Saupe, Dietmar, et al.
Veröffentlicht: (2024)
Avoiding Quality Saturation in UGC Compression Using Denoised References
von: Xiong, Xin, et al.
Veröffentlicht: (2025)
von: Xiong, Xin, et al.
Veröffentlicht: (2025)
A Versatile Depth Video Encoding Scheme Based on Low-rank Tensor Modeling for Free Viewpoint Video
von: Sharma, Mansi, et al.
Veröffentlicht: (2021)
von: Sharma, Mansi, et al.
Veröffentlicht: (2021)
Adaptive Resolution and Chroma Subsampling for Energy-Efficient Video Coding
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
ABC: Adaptive BayesNet Structure Learning for Computational Scalable Multi-task Image Compression
von: Zhang, Yufeng, et al.
Veröffentlicht: (2025)
von: Zhang, Yufeng, et al.
Veröffentlicht: (2025)
Foveated Compression for Immersive Telepresence Visualization
von: Schwarz, Max, et al.
Veröffentlicht: (2025)
von: Schwarz, Max, et al.
Veröffentlicht: (2025)
Is there a relationship between Mean Opinion Score (MOS) and Just Noticeable Difference (JND)?
von: Zhu, Jingwen, et al.
Veröffentlicht: (2026)
von: Zhu, Jingwen, et al.
Veröffentlicht: (2026)
QoE Optimization for Semantic Self-Correcting Video Transmission in Multi-UAV Networks
von: Chen, Xuyang, et al.
Veröffentlicht: (2025)
von: Chen, Xuyang, et al.
Veröffentlicht: (2025)
Generative AI for Multimedia Communication: Recent Advances, An Information-Theoretic Framework, and Future Opportunities
von: Jin, Yili, et al.
Veröffentlicht: (2025)
von: Jin, Yili, et al.
Veröffentlicht: (2025)
Symmetric Entropy-Constrained Video Coding for Machines
von: Sun, Yuxiao, et al.
Veröffentlicht: (2025)
von: Sun, Yuxiao, et al.
Veröffentlicht: (2025)
Content-Driven Frame-Level Bit Prediction for Rate Control in Versatile Video Coding
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
von: Premkumar, Amritha, et al.
Veröffentlicht: (2026)
Convex-hull Estimation using XPSNR for Versatile Video Coding
von: Menon, Vignesh V, et al.
Veröffentlicht: (2024)
von: Menon, Vignesh V, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Memory-Anchored Multimodal Reasoning for Explainable Video Forensics
von: Chen, Chen, et al.
Veröffentlicht: (2025) -
EvEnhancer: Empowering Effectiveness, Efficiency and Generalizability for Continuous Space-Time Video Super-Resolution with Events
von: Wei, Shuoyan, et al.
Veröffentlicht: (2025) -
Prompt-based Multimodal Semantic Communication for Multi-spectral Image Segmentation
von: Zhang, Haoshuo, et al.
Veröffentlicht: (2025) -
Dynamic resolution switching for live streaming
von: Xiong, Xin, et al.
Veröffentlicht: (2026) -
H.265/HEVC Video Steganalysis Based on CU Block Structure Gradients and IPM Mapping
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)