SUM: Saliency Unification through Mamba for Visual Attention Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hosseini, Alireza, Kazerouni, Amirhossein, Akhavan, Saeed, Brudno, Michael, Taati, Babak |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LIFT: Latent Implicit Functions for Task- and Data-Agnostic Encoding
von: Kazerouni, Amirhossein, et al.
Veröffentlicht: (2025)
von: Kazerouni, Amirhossein, et al.
Veröffentlicht: (2025)
STAF: Sinusoidal Trainable Activation Functions for Implicit Neural Representation
von: Morsali, Alireza, et al.
Veröffentlicht: (2025)
von: Morsali, Alireza, et al.
Veröffentlicht: (2025)
SalM$^{2}$: An Extremely Lightweight Saliency Mamba Model for Real-Time Cognitive Awareness of Driver Attention
von: Zhao, Chunyu, et al.
Veröffentlicht: (2025)
von: Zhao, Chunyu, et al.
Veröffentlicht: (2025)
DTFSal: Audio-Visual Dynamic Token Fusion for Video Saliency Prediction
von: Hooshanfar, Kiana, et al.
Veröffentlicht: (2025)
von: Hooshanfar, Kiana, et al.
Veröffentlicht: (2025)
SynPAIN: A Synthetic Dataset of Pain and Non-Pain Facial Expressions
von: Taati, Babak, et al.
Veröffentlicht: (2025)
von: Taati, Babak, et al.
Veröffentlicht: (2025)
Brand Visibility in Packaging: A Deep Learning Approach for Logo Detection, Saliency-Map Prediction, and Logo Placement Analysis
von: Hosseini, Alireza, et al.
Veröffentlicht: (2024)
von: Hosseini, Alireza, et al.
Veröffentlicht: (2024)
Saliency Guided Longitudinal Medical Visual Question Answering
von: Wu, Jialin, et al.
Veröffentlicht: (2025)
von: Wu, Jialin, et al.
Veröffentlicht: (2025)
MambaU-Lite: A Lightweight Model based on Mamba and Integrated Channel-Spatial Attention for Skin Lesion Segmentation
von: Nguyen, Thi-Nhu-Quynh, et al.
Veröffentlicht: (2024)
von: Nguyen, Thi-Nhu-Quynh, et al.
Veröffentlicht: (2024)
Saliency Suppressed, Semantics Surfaced: Visual Transformations in Neural Networks and the Brain
von: Opiełka, Gustaw, et al.
Veröffentlicht: (2024)
von: Opiełka, Gustaw, et al.
Veröffentlicht: (2024)
SRMA-Mamba: Spatial Reverse Mamba Attention Network for Pathological Liver Segmentation in MRI Volumes
von: Zeng, Jun, et al.
Veröffentlicht: (2025)
von: Zeng, Jun, et al.
Veröffentlicht: (2025)
Selective Visual Prompting in Vision Mamba
von: Yao, Yifeng, et al.
Veröffentlicht: (2024)
von: Yao, Yifeng, et al.
Veröffentlicht: (2024)
Jodi: Unification of Visual Generation and Understanding via Joint Modeling
von: Xu, Yifeng, et al.
Veröffentlicht: (2025)
von: Xu, Yifeng, et al.
Veröffentlicht: (2025)
Face2Scene: Using Facial Degradation as an Oracle for Diffusion-Based Scene Restoration
von: Kazerouni, Amirhossein, et al.
Veröffentlicht: (2026)
von: Kazerouni, Amirhossein, et al.
Veröffentlicht: (2026)
Saliency-Guided Representation with Consistency Policy Learning for Visual Unsupervised Reinforcement Learning
von: Sun, Jingbo, et al.
Veröffentlicht: (2026)
von: Sun, Jingbo, et al.
Veröffentlicht: (2026)
LocalMamba: Visual State Space Model with Windowed Selective Scan
von: Huang, Tao, et al.
Veröffentlicht: (2024)
von: Huang, Tao, et al.
Veröffentlicht: (2024)
Revealing the Gap in Human and VLM Scene Perception through Counterfactual Semantic Saliency
von: Wen, Ziqi, et al.
Veröffentlicht: (2026)
von: Wen, Ziqi, et al.
Veröffentlicht: (2026)
Serial Over Parallel: Learning Continual Unification for Multi-Modal Visual Object Tracking and Benchmarking
von: Tang, Zhangyong, et al.
Veröffentlicht: (2025)
von: Tang, Zhangyong, et al.
Veröffentlicht: (2025)
Skip Mamba Diffusion for Monocular 3D Semantic Scene Completion
von: Liang, Li, et al.
Veröffentlicht: (2025)
von: Liang, Li, et al.
Veröffentlicht: (2025)
Towards A Comprehensive Visual Saliency Explanation Framework for AI-based Face Recognition Systems
von: Lu, Yuhang, et al.
Veröffentlicht: (2024)
von: Lu, Yuhang, et al.
Veröffentlicht: (2024)
Evaluating the Performance of Deep Learning Models in Whole-body Dynamic 3D Posture Prediction During Load-reaching Activities
von: Hosseini, Seyede Niloofar, et al.
Veröffentlicht: (2025)
von: Hosseini, Seyede Niloofar, et al.
Veröffentlicht: (2025)
Continuous Sign Language Recognition Using Intra-inter Gloss Attention
von: Ranjbar, Hossein, et al.
Veröffentlicht: (2024)
von: Ranjbar, Hossein, et al.
Veröffentlicht: (2024)
FaceSaliencyAug: Mitigating Geographic, Gender and Stereotypical Biases via Saliency-Based Data Augmentation
von: Kumar, Teerath, et al.
Veröffentlicht: (2024)
von: Kumar, Teerath, et al.
Veröffentlicht: (2024)
MambaEVT: Event Stream based Visual Object Tracking using State Space Model
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
Modeling Saliency Dataset Bias
von: Kümmerer, Matthias, et al.
Veröffentlicht: (2025)
von: Kümmerer, Matthias, et al.
Veröffentlicht: (2025)
Correlation of Object Detection Performance with Visual Saliency and Depth Estimation
von: Bartolo, Matthias, et al.
Veröffentlicht: (2024)
von: Bartolo, Matthias, et al.
Veröffentlicht: (2024)
EfficientVMamba: Atrous Selective Scan for Light Weight Visual Mamba
von: Pei, Xiaohuan, et al.
Veröffentlicht: (2024)
von: Pei, Xiaohuan, et al.
Veröffentlicht: (2024)
MixerCSeg: An Efficient Mixer Architecture for Crack Segmentation via Decoupled Mamba Attention
von: Zhao, Zilong, et al.
Veröffentlicht: (2026)
von: Zhao, Zilong, et al.
Veröffentlicht: (2026)
FMNet: Frequency-Assisted Mamba-Like Linear Attention Network for Camouflaged Object Detection
von: Deng, Ming, et al.
Veröffentlicht: (2025)
von: Deng, Ming, et al.
Veröffentlicht: (2025)
SEMA: a Scalable and Efficient Mamba like Attention via Token Localization and Averaging
von: Tran, Nhat Thanh, et al.
Veröffentlicht: (2025)
von: Tran, Nhat Thanh, et al.
Veröffentlicht: (2025)
Salience Adjustment for Context-Based Emotion Recognition
von: Han, Bin, et al.
Veröffentlicht: (2025)
von: Han, Bin, et al.
Veröffentlicht: (2025)
Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models
von: Yang, Songlin, et al.
Veröffentlicht: (2026)
von: Yang, Songlin, et al.
Veröffentlicht: (2026)
MambaEye: A Size-Agnostic Visual Encoder with Causal Sequential Processing
von: Choi, Changho, et al.
Veröffentlicht: (2025)
von: Choi, Changho, et al.
Veröffentlicht: (2025)
FAMSeg: Fetal Femur and Cranial Ultrasound Segmentation Using Feature-Aware Attention and Mamba Enhancement
von: He, Jie, et al.
Veröffentlicht: (2025)
von: He, Jie, et al.
Veröffentlicht: (2025)
Do Sparse Subnetworks Exhibit Cognitively Aligned Attention? Effects of Pruning on Saliency Map Fidelity, Sparsity, and Concept Coherence
von: Suwal, Sanish, et al.
Veröffentlicht: (2025)
von: Suwal, Sanish, et al.
Veröffentlicht: (2025)
TextMamba: Scene Text Detector with Mamba
von: Zhao, Qiyan, et al.
Veröffentlicht: (2025)
von: Zhao, Qiyan, et al.
Veröffentlicht: (2025)
Surgical-MambaLLM: Mamba2-enhanced Multimodal Large Language Model for VQLA in Robotic Surgery
von: Hao, Pengfei, et al.
Veröffentlicht: (2025)
von: Hao, Pengfei, et al.
Veröffentlicht: (2025)
TrajMamba: An Ego-Motion-Guided Mamba Model for Pedestrian Trajectory Prediction from an Egocentric Perspective
von: Peng, Yusheng, et al.
Veröffentlicht: (2026)
von: Peng, Yusheng, et al.
Veröffentlicht: (2026)
Empathetic Response in Audio-Visual Conversations Using Emotion Preference Optimization and MambaCompressor
von: Kim, Yeonju, et al.
Veröffentlicht: (2024)
von: Kim, Yeonju, et al.
Veröffentlicht: (2024)
Pan-FM: A Pan-Organ Foundation Model with Saliency-Guided Masking for Missing Robustness
von: Wu, Qiangqiang, et al.
Veröffentlicht: (2026)
von: Wu, Qiangqiang, et al.
Veröffentlicht: (2026)
Hierarchical Modeling for Medical Visual Question Answering with Cross-Attention Fusion
von: Zhang, Junkai, et al.
Veröffentlicht: (2025)
von: Zhang, Junkai, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LIFT: Latent Implicit Functions for Task- and Data-Agnostic Encoding
von: Kazerouni, Amirhossein, et al.
Veröffentlicht: (2025) -
STAF: Sinusoidal Trainable Activation Functions for Implicit Neural Representation
von: Morsali, Alireza, et al.
Veröffentlicht: (2025) -
SalM$^{2}$: An Extremely Lightweight Saliency Mamba Model for Real-Time Cognitive Awareness of Driver Attention
von: Zhao, Chunyu, et al.
Veröffentlicht: (2025) -
DTFSal: Audio-Visual Dynamic Token Fusion for Video Saliency Prediction
von: Hooshanfar, Kiana, et al.
Veröffentlicht: (2025) -
SynPAIN: A Synthetic Dataset of Pain and Non-Pain Facial Expressions
von: Taati, Babak, et al.
Veröffentlicht: (2025)