DGSSM: Diffusion guided state-space models for multimodal salient object detection
Fuente:
arXiv
Saved in:
| Main Authors: | Ghosh, Suklav, Sur, Arijit, Mitra, Pinaki |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
C-LEAD: Contrastive Learning for Enhanced Adversarial Defense
by: Ghosh, Suklav, et al.
Published: (2025)
by: Ghosh, Suklav, et al.
Published: (2025)
Trans-defense: Transformer-based Denoiser for Adversarial Defense with Spatial-Frequency Domain Representation
by: Pramanick, Alik, et al.
Published: (2025)
by: Pramanick, Alik, et al.
Published: (2025)
Do multimodal models imagine electric sheep?
by: Ramakrishnan, Santhosh Kumar, et al.
Published: (2026)
by: Ramakrishnan, Santhosh Kumar, et al.
Published: (2026)
D4: Text-guided diffusion model-based domain adaptive data augmentation for vineyard shoot detection
by: Hirahara, Kentaro, et al.
Published: (2024)
by: Hirahara, Kentaro, et al.
Published: (2024)
SADA: Stability-guided Adaptive Diffusion Acceleration
by: Jiang, Ting, et al.
Published: (2025)
by: Jiang, Ting, et al.
Published: (2025)
Sketch-guided Image Inpainting with Partial Discrete Diffusion Process
by: Sharma, Nakul, et al.
Published: (2024)
by: Sharma, Nakul, et al.
Published: (2024)
Explaining latent representations of generative models with large multimodal models
by: Zhu, Mengdan, et al.
Published: (2024)
by: Zhu, Mengdan, et al.
Published: (2024)
Contrastive Denoising Score for Text-guided Latent Diffusion Image Editing
by: Nam, Hyelin, et al.
Published: (2023)
by: Nam, Hyelin, et al.
Published: (2023)
From classical techniques to convolution-based models: A review of object detection algorithms
by: Neha, Fnu, et al.
Published: (2024)
by: Neha, Fnu, et al.
Published: (2024)
Review of multimodal machine learning approaches in healthcare
by: Krones, Felix, et al.
Published: (2024)
by: Krones, Felix, et al.
Published: (2024)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
by: Jeong, Hyeonho, et al.
Published: (2024)
by: Jeong, Hyeonho, et al.
Published: (2024)
DOSE3 : Diffusion-based Out-of-distribution detection on SE(3) trajectories
by: Cheng, Hongzhe, et al.
Published: (2025)
by: Cheng, Hongzhe, et al.
Published: (2025)
Autonomous state-space segmentation for Deep-RL sparse reward scenarios
by: Maselli, Gianluca, et al.
Published: (2025)
by: Maselli, Gianluca, et al.
Published: (2025)
Towards Physics-informed Diffusion for Anomaly Detection in Trajectories
by: Sharma, Arun, et al.
Published: (2025)
by: Sharma, Arun, et al.
Published: (2025)
Hybrid deep convolution model for lung cancer detection with transfer learning
by: Saxena, Sugandha, et al.
Published: (2025)
by: Saxena, Sugandha, et al.
Published: (2025)
MAN TruckScenes: A multimodal dataset for autonomous trucking in diverse conditions
by: Fent, Felix, et al.
Published: (2024)
by: Fent, Felix, et al.
Published: (2024)
Open-set object detection: towards unified problem formulation and benchmarking
by: Ammar, Hejer, et al.
Published: (2024)
by: Ammar, Hejer, et al.
Published: (2024)
Generating crossmodal gene expression from cancer histopathology improves multimodal AI predictions
by: Dey, Samiran, et al.
Published: (2025)
by: Dey, Samiran, et al.
Published: (2025)
What to align in multimodal contrastive learning?
by: Dufumier, Benoit, et al.
Published: (2024)
by: Dufumier, Benoit, et al.
Published: (2024)
Moving object detection from multi-depth images with an attention-enhanced CNN
by: Shibukawa, Masato, et al.
Published: (2025)
by: Shibukawa, Masato, et al.
Published: (2025)
FeatInv: Spatially resolved mapping from feature space to input space using conditional diffusion models
by: Neukirch, Nils, et al.
Published: (2025)
by: Neukirch, Nils, et al.
Published: (2025)
SELECTOR: Heterogeneous graph network with convolutional masked autoencoder for multimodal robust prediction of cancer survival
by: Pan, Liangrui, et al.
Published: (2024)
by: Pan, Liangrui, et al.
Published: (2024)
Deep Attention-guided Adaptive Subsampling
by: Shankaranarayana, Sharath M, et al.
Published: (2025)
by: Shankaranarayana, Sharath M, et al.
Published: (2025)
Self-learned representation-guided latent diffusion model for breast cancer classification in deep ultraviolet whole surface images
by: Afshin, Pouya, et al.
Published: (2026)
by: Afshin, Pouya, et al.
Published: (2026)
PBCAT: Patch-based composite adversarial training against physically realizable attacks on object detection
by: Li, Xiao, et al.
Published: (2025)
by: Li, Xiao, et al.
Published: (2025)
RQR3D: Reparametrizing the regression targets for BEV-based 3D object detection
by: Kilinc, Ozsel, et al.
Published: (2025)
by: Kilinc, Ozsel, et al.
Published: (2025)
Feedback-guided Data Synthesis for Imbalanced Classification
by: Hemmat, Reyhane Askari, et al.
Published: (2023)
by: Hemmat, Reyhane Askari, et al.
Published: (2023)
Enhancing multimodal cooperation via sample-level modality valuation
by: Wei, Yake, et al.
Published: (2023)
by: Wei, Yake, et al.
Published: (2023)
Human-like object concept representations emerge naturally in multimodal large language models
by: Du, Changde, et al.
Published: (2024)
by: Du, Changde, et al.
Published: (2024)
An attempt to generate new bridge types from latent space of denoising diffusion Implicit model
by: Zhang, Hongjun
Published: (2024)
by: Zhang, Hongjun
Published: (2024)
Learning Hyperspectral Images with Curated Text Prompts for Efficient Multimodal Alignment
by: Chatterjee, Abhiroop, et al.
Published: (2025)
by: Chatterjee, Abhiroop, et al.
Published: (2025)
Structured Unrestricted-Rank Matrices for Parameter Efficient Fine-tuning
by: Sehanobish, Arijit, et al.
Published: (2024)
by: Sehanobish, Arijit, et al.
Published: (2024)
CDAN: Convolutional dense attention-guided network for low-light image enhancement
by: Shakibania, Hossein, et al.
Published: (2023)
by: Shakibania, Hossein, et al.
Published: (2023)
Saliency-guided Emotion Modeling: Predicting Viewer Reactions from Video Stimuli
by: Yaragoppa, Akhila, et al.
Published: (2025)
by: Yaragoppa, Akhila, et al.
Published: (2025)
Enhancing kelp forest detection in remote sensing images using crowdsourced labels with Mixed Vision Transformers and ConvNeXt segmentation models
by: Nasios, Ioannis
Published: (2025)
by: Nasios, Ioannis
Published: (2025)
PIF: Anomaly detection via preference embedding
by: Leveni, Filippo, et al.
Published: (2025)
by: Leveni, Filippo, et al.
Published: (2025)
Defect detection using weakly supervised learning
by: Sevetlidis, Vasileios, et al.
Published: (2023)
by: Sevetlidis, Vasileios, et al.
Published: (2023)
Few-Shot Classification and Anatomical Localization of Tissues in SPECT Imaging
by: Khan, Mohammed Abdul Hafeez, et al.
Published: (2025)
by: Khan, Mohammed Abdul Hafeez, et al.
Published: (2025)
UPCMR: A Universal Prompt-guided Model for Random Sampling Cardiac MRI Reconstruction
by: Lyu, Donghang, et al.
Published: (2025)
by: Lyu, Donghang, et al.
Published: (2025)
DiffusionNFT: Online Diffusion Reinforcement with Forward Process
by: Zheng, Kaiwen, et al.
Published: (2025)
by: Zheng, Kaiwen, et al.
Published: (2025)
Similar Items
-
C-LEAD: Contrastive Learning for Enhanced Adversarial Defense
by: Ghosh, Suklav, et al.
Published: (2025) -
Trans-defense: Transformer-based Denoiser for Adversarial Defense with Spatial-Frequency Domain Representation
by: Pramanick, Alik, et al.
Published: (2025) -
Do multimodal models imagine electric sheep?
by: Ramakrishnan, Santhosh Kumar, et al.
Published: (2026) -
D4: Text-guided diffusion model-based domain adaptive data augmentation for vineyard shoot detection
by: Hirahara, Kentaro, et al.
Published: (2024) -
SADA: Stability-guided Adaptive Diffusion Acceleration
by: Jiang, Ting, et al.
Published: (2025)