SLIM-Diff: Shared Latent Image-Mask Diffusion with Lp loss for Data-Scarce Epilepsy FLAIR MRI
Fuente:
arXiv
Guardado en:
| Autores principales: | Pascual-González, Mario, Jiménez-Partinen, Ariadna, Luque-Baena, R. M., Nagib-Raya, Fátima, López-Rubio, Ezequiel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
por: Raoufi, Behnam, et al.
Publicado: (2025)
por: Raoufi, Behnam, et al.
Publicado: (2025)
U-Net-Like Spiking Neural Networks for Single Image Dehazing
por: Li, Huibin, et al.
Publicado: (2025)
por: Li, Huibin, et al.
Publicado: (2025)
PRISM: Differentiable Analysis-by-Synthesis for Fixel Recovery in Diffusion MRI
por: Abouagour, Mohamed, et al.
Publicado: (2026)
por: Abouagour, Mohamed, et al.
Publicado: (2026)
Interpretable Tau-PET Synthesis from Multimodal T1-Weighted and FLAIR MRI Using Partial Information Decomposition Guided Disentangled Quantized Half-UNet
por: Chopra, Agamdeep S., et al.
Publicado: (2026)
por: Chopra, Agamdeep S., et al.
Publicado: (2026)
Exploring the Capabilities of Large Language Model Encoders for Image-Text Retrieval in Chest X-rays
por: Ko, Hanbin, et al.
Publicado: (2025)
por: Ko, Hanbin, et al.
Publicado: (2025)
Breast Cell Segmentation Under Extreme Data Constraints: Quantum Enhancement Meets Adaptive Loss Stabilization
por: Dasoju, Varun Kumar, et al.
Publicado: (2025)
por: Dasoju, Varun Kumar, et al.
Publicado: (2025)
Learning Association via Track-Detection Matching for Multi-Object Tracking
por: Adžemović, Momir
Publicado: (2025)
por: Adžemović, Momir
Publicado: (2025)
BreastDCEDL: A Comprehensive Breast Cancer DCE-MRI Dataset and Transformer Implementation for Treatment Response Prediction
por: Fridman, Naomi, et al.
Publicado: (2025)
por: Fridman, Naomi, et al.
Publicado: (2025)
When in Doubt, Think Slow: Iterative Reasoning with Latent Imagination
por: Benfeghoul, Martin, et al.
Publicado: (2024)
por: Benfeghoul, Martin, et al.
Publicado: (2024)
SpinalSAM-R1: A Vision-Language Multimodal Interactive System for Spine CT Segmentation
por: Liu, Jiaming, et al.
Publicado: (2025)
por: Liu, Jiaming, et al.
Publicado: (2025)
Mask-Conditioned Voxel Diffusion for Joint Geometry and Color Inpainting
por: Sumuk, Aarya
Publicado: (2026)
por: Sumuk, Aarya
Publicado: (2026)
Efficient Image Pre-Training with Siamese Cropped Masked Autoencoders
por: Eymaël, Alexandre, et al.
Publicado: (2024)
por: Eymaël, Alexandre, et al.
Publicado: (2024)
Motion Perceiver: Real-Time Occupancy Forecasting for Embedded Systems
por: Ferenczi, Bryce, et al.
Publicado: (2023)
por: Ferenczi, Bryce, et al.
Publicado: (2023)
Learning Continuous Receive Apodization Weights via Implicit Neural Representation for Ultrafast ICE Ultrasound Imaging
por: Delaunay, Rémi, et al.
Publicado: (2025)
por: Delaunay, Rémi, et al.
Publicado: (2025)
Modulated INR with Prior Embeddings for Ultrasound Imaging Reconstruction
por: Delaunay, Rémi, et al.
Publicado: (2025)
por: Delaunay, Rémi, et al.
Publicado: (2025)
Heart Failure Prediction using Modal Decomposition and Masked Autoencoders for Scarce Echocardiography Databases
por: Bell-Navas, Andrés, et al.
Publicado: (2025)
por: Bell-Navas, Andrés, et al.
Publicado: (2025)
Preserving instance continuity and length in segmentation through connectivity-aware loss computation
por: Szustakowski, Karol, et al.
Publicado: (2025)
por: Szustakowski, Karol, et al.
Publicado: (2025)
Topology-Aware Latent Diffusion for 3D Shape Generation
por: Hu, Jiangbei, et al.
Publicado: (2024)
por: Hu, Jiangbei, et al.
Publicado: (2024)
Semi-supervised Latent Disentangled Diffusion Model for Textile Pattern Generation
por: Hu, Chenggong, et al.
Publicado: (2026)
por: Hu, Chenggong, et al.
Publicado: (2026)
Quantized Vision-Language Models for Damage Assessment: A Comparative Study of LLaVA-1.5-7B Quantization Levels
por: Yasuno, Takato
Publicado: (2026)
por: Yasuno, Takato
Publicado: (2026)
Cora: Correspondence-aware image editing using few step diffusion
por: Alimohammadi, Amirhossein, et al.
Publicado: (2025)
por: Alimohammadi, Amirhossein, et al.
Publicado: (2025)
Pointing-Based Object Recognition
por: Hajdúch, Lukáš, et al.
Publicado: (2026)
por: Hajdúch, Lukáš, et al.
Publicado: (2026)
Autoregressive Medical Image Segmentation via Next-Scale Mask Prediction
por: Chen, Tao, et al.
Publicado: (2025)
por: Chen, Tao, et al.
Publicado: (2025)
PhysVideoGenerator: Towards Physically Aware Video Generation via Latent Physics Guidance
por: Satish, Siddarth Nilol Kundur, et al.
Publicado: (2026)
por: Satish, Siddarth Nilol Kundur, et al.
Publicado: (2026)
SITransformer: Shared Information-Guided Transformer for Extreme Multimodal Summarization
por: Liu, Sicheng, et al.
Publicado: (2024)
por: Liu, Sicheng, et al.
Publicado: (2024)
Visual Enhanced Depth Scaling for Multimodal Latent Reasoning
por: Han, Yudong, et al.
Publicado: (2026)
por: Han, Yudong, et al.
Publicado: (2026)
Medico 2025: Visual Question Answering for Gastrointestinal Imaging
por: Gautam, Sushant, et al.
Publicado: (2025)
por: Gautam, Sushant, et al.
Publicado: (2025)
Can Local Vision-Language Models improve Activity Recognition over Vision Transformers? -- Case Study on Newborn Resuscitation
por: Guerriero, Enrico, et al.
Publicado: (2026)
por: Guerriero, Enrico, et al.
Publicado: (2026)
Low Dose CT for Stroke Diagnosis: A Dual Pipeline Deep Learning Framework for Portable Neuroimaging
por: Ghosal, Rhea, et al.
Publicado: (2026)
por: Ghosal, Rhea, et al.
Publicado: (2026)
Bayesian Deep Learning Approaches for Uncertainty-Aware Retinal OCT Image Segmentation for Multiple Sclerosis
por: Ball, Samuel T. M.
Publicado: (2025)
por: Ball, Samuel T. M.
Publicado: (2025)
An Empirical Study for Representations of Videos in Video Question Answering via MLLMs
por: Li, Zhi, et al.
Publicado: (2025)
por: Li, Zhi, et al.
Publicado: (2025)
A Comparative Analysis of Recurrent and Attention Architectures for Isolated Sign Language Recognition
por: Alishzade, Nigar, et al.
Publicado: (2025)
por: Alishzade, Nigar, et al.
Publicado: (2025)
Extracting Explanations, Justification, and Uncertainty from Black-Box Deep Neural Networks
por: Ardis, Paul, et al.
Publicado: (2024)
por: Ardis, Paul, et al.
Publicado: (2024)
Technology prediction of a 3D model using Neural Network
por: Miebs, Grzegorz, et al.
Publicado: (2025)
por: Miebs, Grzegorz, et al.
Publicado: (2025)
Sat-JEPA-Diff: Bridging Self-Supervised Learning and Generative Diffusion for Remote Sensing
por: Komurcu, Kursat, et al.
Publicado: (2026)
por: Komurcu, Kursat, et al.
Publicado: (2026)
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
por: Dai, Song, et al.
Publicado: (2025)
por: Dai, Song, et al.
Publicado: (2025)
Universal Adversarial Attack on Aligned Multimodal LLMs
por: Rahmatullaev, Temurbek, et al.
Publicado: (2025)
por: Rahmatullaev, Temurbek, et al.
Publicado: (2025)
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
por: Tong, Jingqi, et al.
Publicado: (2025)
por: Tong, Jingqi, et al.
Publicado: (2025)
SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing
por: Meng, Zi, et al.
Publicado: (2026)
por: Meng, Zi, et al.
Publicado: (2026)
AGOP as Explanation: From Feature Learning to Per-Sample Attribution in Image Classifiers
por: Katakam, Raj Kiran Gupta
Publicado: (2026)
por: Katakam, Raj Kiran Gupta
Publicado: (2026)
Ejemplares similares
-
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
por: Raoufi, Behnam, et al.
Publicado: (2025) -
U-Net-Like Spiking Neural Networks for Single Image Dehazing
por: Li, Huibin, et al.
Publicado: (2025) -
PRISM: Differentiable Analysis-by-Synthesis for Fixel Recovery in Diffusion MRI
por: Abouagour, Mohamed, et al.
Publicado: (2026) -
Interpretable Tau-PET Synthesis from Multimodal T1-Weighted and FLAIR MRI Using Partial Information Decomposition Guided Disentangled Quantized Half-UNet
por: Chopra, Agamdeep S., et al.
Publicado: (2026) -
Exploring the Capabilities of Large Language Model Encoders for Image-Text Retrieval in Chest X-rays
por: Ko, Hanbin, et al.
Publicado: (2025)