CFMW: Cross-modality Fusion Mamba for Robust Object Detection under Adverse Weather
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Haoyuan, Hu, Qi, Zhou, Binjia, Yao, You, Lin, Jiacheng, Yang, Kailun, Chen, Peng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
by: Zeng, Kang, et al.
Published: (2024)
by: Zeng, Kang, et al.
Published: (2024)
RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
by: Peng, Kunyu, et al.
Published: (2025)
by: Peng, Kunyu, et al.
Published: (2025)
E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes
by: Zhai, Jiajun, et al.
Published: (2026)
by: Zhai, Jiajun, et al.
Published: (2026)
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
by: Wei, Yiping, et al.
Published: (2023)
by: Wei, Yiping, et al.
Published: (2023)
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
by: Chen, Yifei, et al.
Published: (2023)
by: Chen, Yifei, et al.
Published: (2023)
Exploring Event-based Human Pose Estimation with 3D Event Representations
by: Yin, Xiaoting, et al.
Published: (2023)
by: Yin, Xiaoting, et al.
Published: (2023)
Robust Live Streaming over LEO Satellite Constellations: Measurement, Analysis, and Handover-Aware Adaptation
by: Fang, Hao, et al.
Published: (2025)
by: Fang, Hao, et al.
Published: (2025)
Robust Multi-modal Task-oriented Communications with Redundancy-aware Representations
by: Fu, Jingwen, et al.
Published: (2025)
by: Fu, Jingwen, et al.
Published: (2025)
Symmetric Entropy-Constrained Video Coding for Machines
by: Sun, Yuxiao, et al.
Published: (2025)
by: Sun, Yuxiao, et al.
Published: (2025)
HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios
by: Peng, Kunyu, et al.
Published: (2025)
by: Peng, Kunyu, et al.
Published: (2025)
EdgeNavMamba: Mamba Optimized Object Detection for Energy Efficient Edge Devices
by: Aalishah, Romina, et al.
Published: (2025)
by: Aalishah, Romina, et al.
Published: (2025)
Audio-Visual Cross-Modal Compression for Generative Face Video Coding
by: Xu, Youmin, et al.
Published: (2025)
by: Xu, Youmin, et al.
Published: (2025)
MarsSQE: Stereo Quality Enhancement for Martian Images Using Bi-level Cross-view Attention
by: Xu, Mai, et al.
Published: (2024)
by: Xu, Mai, et al.
Published: (2024)
Camel: Frame-Level Bandwidth Estimation for Low-Latency Live Streaming under Video Bitrate Undershooting
by: Liu, Liming, et al.
Published: (2026)
by: Liu, Liming, et al.
Published: (2026)
Interactive $360^{\circ}$ Video Streaming Using FoV-Adaptive Coding with Temporal Prediction
by: Mao, Yixiang, et al.
Published: (2024)
by: Mao, Yixiang, et al.
Published: (2024)
End-to-End RGB-IR Joint Image Compression With Channel-wise Cross-modality Entropy Model
by: Wang, Haofeng, et al.
Published: (2025)
by: Wang, Haofeng, et al.
Published: (2025)
EvEnhancer: Empowering Effectiveness, Efficiency and Generalizability for Continuous Space-Time Video Super-Resolution with Events
by: Wei, Shuoyan, et al.
Published: (2025)
by: Wei, Shuoyan, et al.
Published: (2025)
Progressive Frame Patching for FoV-based Point Cloud Video Streaming
by: Zong, Tongyu, et al.
Published: (2023)
by: Zong, Tongyu, et al.
Published: (2023)
TVMC: Time-Varying Mesh Compression via Multi-Stage Anchor Mesh Generation
by: Huang, He, et al.
Published: (2025)
by: Huang, He, et al.
Published: (2025)
QoE Optimization for Semantic Self-Correcting Video Transmission in Multi-UAV Networks
by: Chen, Xuyang, et al.
Published: (2025)
by: Chen, Xuyang, et al.
Published: (2025)
H.265/HEVC Video Steganalysis Based on CU Block Structure Gradients and IPM Mapping
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
Unified ROI-based Image Compression Paradigm with Generalized Gaussian Model
by: Hu, Kai, et al.
Published: (2026)
by: Hu, Kai, et al.
Published: (2026)
Region-Adaptive Learned Hierarchical Encoding for 3D Gaussian Splatting Data
by: Sridhara, Shashank N., et al.
Published: (2025)
by: Sridhara, Shashank N., et al.
Published: (2025)
A H.265/HEVC Fine-Grained ROI Video Encryption Algorithm Based on Coding Unit and Prompt Segmentation
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
Editing Away the Evidence: Diffusion-Based Image Manipulation and the Failure Modes of Robust Watermarking
by: Qi, Qian, et al.
Published: (2026)
by: Qi, Qian, et al.
Published: (2026)
ABC: Adaptive BayesNet Structure Learning for Computational Scalable Multi-task Image Compression
by: Zhang, Yufeng, et al.
Published: (2025)
by: Zhang, Yufeng, et al.
Published: (2025)
HippoMM: Hippocampal-inspired Multimodal Memory for Long Audiovisual Event Understanding
by: Lin, Yueqian, et al.
Published: (2025)
by: Lin, Yueqian, et al.
Published: (2025)
CoBEVMoE: Heterogeneity-aware Feature Fusion with Dynamic Mixture-of-Experts for Collaborative Perception
by: Kong, Lingzhao, et al.
Published: (2025)
by: Kong, Lingzhao, et al.
Published: (2025)
Transform and Entropy Coding in AV2
by: Nalci, Alican, et al.
Published: (2026)
by: Nalci, Alican, et al.
Published: (2026)
CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection
by: Wang, Hang, et al.
Published: (2026)
by: Wang, Hang, et al.
Published: (2026)
Rip Current Detection in Nearshore Areas through UAV Video Analysis with Almost Local-Isometric Embedding Techniques on Sphere
by: Sun, Anchen, et al.
Published: (2023)
by: Sun, Anchen, et al.
Published: (2023)
Enhanced Template-based Intra Mode Derivation with Adaptive Block Vector Replacement
by: Zhang, Jiaqi, et al.
Published: (2025)
by: Zhang, Jiaqi, et al.
Published: (2025)
Tube-Structured Incremental Semantic HARQ for Generative Video Receivers
by: Wang, Xuesong, et al.
Published: (2026)
by: Wang, Xuesong, et al.
Published: (2026)
Rate-Quality or Energy-Quality Pareto Fronts for Adaptive Video Streaming?
by: Katsenou, Angeliki, et al.
Published: (2024)
by: Katsenou, Angeliki, et al.
Published: (2024)
Smaller is Better: Generative Models Can Power Short Video Preloading
by: Liu, Liming, et al.
Published: (2026)
by: Liu, Liming, et al.
Published: (2026)
NiMark: A Non-intrusive Watermarking Framework against Screen-shooting Attacks
by: Wu, Yufeng, et al.
Published: (2026)
by: Wu, Yufeng, et al.
Published: (2026)
Fast Multirate Encoding for 360° Video in OMAF Streaming Workflows
by: Premkumar, Amritha, et al.
Published: (2026)
by: Premkumar, Amritha, et al.
Published: (2026)
Decoding Complexity-Rate-Quality Pareto-Front for Adaptive VVC Streaming
by: Katsenou, Angeliki, et al.
Published: (2024)
by: Katsenou, Angeliki, et al.
Published: (2024)
Memory-Anchored Multimodal Reasoning for Explainable Video Forensics
by: Chen, Chen, et al.
Published: (2025)
by: Chen, Chen, et al.
Published: (2025)
GScomp-QA: A Subjective Dataset for Quality Assessment of Compressed Gaussian Splatting
by: Martin, Pedro, et al.
Published: (2026)
by: Martin, Pedro, et al.
Published: (2026)
Similar Items
-
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
by: Zeng, Kang, et al.
Published: (2024) -
RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
by: Peng, Kunyu, et al.
Published: (2025) -
E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes
by: Zhai, Jiajun, et al.
Published: (2026) -
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
by: Wei, Yiping, et al.
Published: (2023) -
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
by: Chen, Yifei, et al.
Published: (2023)