Versatile Framework with Semantic and Structural guidance for Image Reconstruction from Brain Activity
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Yizhuo, Du, Changde, Zhou, Qiongyi, Jiang, Liuyun, He, Huiguang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Animate Your Thoughts: Decoupled Reconstruction of Dynamic Natural Vision from Slow Brain Activity
by: Lu, Yizhuo, et al.
Published: (2024)
by: Lu, Yizhuo, et al.
Published: (2024)
CLIP-MUSED: CLIP-Guided Multi-Subject Visual Neural Information Semantic Decoding
by: Zhou, Qiongyi, et al.
Published: (2024)
by: Zhou, Qiongyi, et al.
Published: (2024)
NeuralOOD: Improving Out-of-Distribution Generalization Performance with Brain-machine Fusion Learning Framework
by: Zhao, Shuangchen, et al.
Published: (2024)
by: Zhao, Shuangchen, et al.
Published: (2024)
Mind-Omni: A Unified Multi-Task Framework for Brain-Vision-Language Modeling via Discrete Diffusion
by: Lu, Yizhuo, et al.
Published: (2026)
by: Lu, Yizhuo, et al.
Published: (2026)
SemVideo: Reconstructs What You Watch from Brain Activity via Hierarchical Semantic Guidance
by: Yang, Minghan, et al.
Published: (2026)
by: Yang, Minghan, et al.
Published: (2026)
HAVIR: HierArchical Vision to Image Reconstruction using CLIP-Guided Versatile Diffusion
by: Zhang, Shiyi, et al.
Published: (2025)
by: Zhang, Shiyi, et al.
Published: (2025)
TGC-Net: A Structure-Aware and Semantically-Aligned Framework for Text-Guided Medical Image Segmentation
by: Lin, Gaoren, et al.
Published: (2025)
by: Lin, Gaoren, et al.
Published: (2025)
Reverse the auditory processing pathway: Coarse-to-fine audio reconstruction from fMRI
by: Liu, Che, et al.
Published: (2024)
by: Liu, Che, et al.
Published: (2024)
BrainVis: Exploring the Bridge between Brain and Visual Signals via Image Reconstruction
by: Fu, Honghao, et al.
Published: (2023)
by: Fu, Honghao, et al.
Published: (2023)
NeuroMamba: Multi-Perspective Feature Interaction with Visual Mamba for Neuron Segmentation
by: Jiang, Liuyun, et al.
Published: (2026)
by: Jiang, Liuyun, et al.
Published: (2026)
Medical Visual Prompting (MVP): A Unified Framework for Versatile and High-Quality Medical Image Segmentation
by: Chen, Yulin, et al.
Published: (2024)
by: Chen, Yulin, et al.
Published: (2024)
HAVIR: HierArchical Vision to Image Reconstruction using CLIP-Guided Versatile Diffusion
by: Zhang, Shiyi, et al.
Published: (2025)
by: Zhang, Shiyi, et al.
Published: (2025)
Pixel3DMM: Versatile Screen-Space Priors for Single-Image 3D Face Reconstruction
by: Giebenhain, Simon, et al.
Published: (2025)
by: Giebenhain, Simon, et al.
Published: (2025)
SpikeVAEDiff: Neural Spike-based Natural Visual Scene Reconstruction via VD-VAE and Versatile Diffusion
by: Li, Jialu, et al.
Published: (2026)
by: Li, Jialu, et al.
Published: (2026)
Diffusion Model-Based Data Augmentation for Enhanced Neuron Segmentation
by: Jiang, Liuyun, et al.
Published: (2026)
by: Jiang, Liuyun, et al.
Published: (2026)
Brain-CLIPLM: Decoding Compressed Semantic Representations in EEG for Language Reconstruction
by: Yang, Xiaoli, et al.
Published: (2026)
by: Yang, Xiaoli, et al.
Published: (2026)
Zero-shot Text-guided Infinite Image Synthesis with LLM guidance
by: Kwon, Soyeong, et al.
Published: (2024)
by: Kwon, Soyeong, et al.
Published: (2024)
SAT3D: Image-driven Semantic Attribute Transfer in 3D
by: Zhai, Zhijun, et al.
Published: (2024)
by: Zhai, Zhijun, et al.
Published: (2024)
Enabling Versatile Controls for Video Diffusion Models
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
A Lightweight Transformer for Pain Recognition from Brain Activity
by: Gkikas, Stefanos, et al.
Published: (2026)
by: Gkikas, Stefanos, et al.
Published: (2026)
ChromouVQA: Benchmarking Vision-Language Models under Chromatic Camouflaged Images
by: Zhang, Yunfei, et al.
Published: (2025)
by: Zhang, Yunfei, et al.
Published: (2025)
End-to-End Deep Learning for Structural Brain Imaging: A Unified Framework
by: Su, Yao, et al.
Published: (2025)
by: Su, Yao, et al.
Published: (2025)
RadMamba: Efficient Human Activity Recognition through Radar-based Micro-Doppler-Oriented Mamba State-Space Model
by: Wu, Yizhuo, et al.
Published: (2025)
by: Wu, Yizhuo, et al.
Published: (2025)
Adaptive Clinical-Aware Latent Diffusion for Multimodal Brain Image Generation and Missing Modality Imputation
by: Zhou, Rong, et al.
Published: (2026)
by: Zhou, Rong, et al.
Published: (2026)
Brain-IT: Image Reconstruction from fMRI via Brain-Interaction Transformer
by: Beliy, Roman, et al.
Published: (2025)
by: Beliy, Roman, et al.
Published: (2025)
Knowledge-Base based Semantic Image Transmission Using CLIP
by: Li, Chongyang, et al.
Published: (2025)
by: Li, Chongyang, et al.
Published: (2025)
Semantic Is Enough: Only Semantic Information For NeRF Reconstruction
by: Wang, Ruibo, et al.
Published: (2024)
by: Wang, Ruibo, et al.
Published: (2024)
AugGS: Self-augmented Gaussians with Structural Masks for Sparse-view 3D Reconstruction
by: Du, Bi'an, et al.
Published: (2024)
by: Du, Bi'an, et al.
Published: (2024)
SPIE: Semantic and Structural Post-Training of Image Editing Diffusion Models with AI feedback
by: Benarous, Elior, et al.
Published: (2025)
by: Benarous, Elior, et al.
Published: (2025)
An Efficient Deep Learning Framework for Brain Stroke Diagnosis Using Computed Tomography Images
by: Hossen, Md. Sabbir, et al.
Published: (2025)
by: Hossen, Md. Sabbir, et al.
Published: (2025)
FlexMUSE: Multimodal Unification and Semantics Enhancement Framework with Flexible interaction for Creative Writing
by: Chen, Jiahao, et al.
Published: (2025)
by: Chen, Jiahao, et al.
Published: (2025)
SwinSF: Image Reconstruction from Spatial-Temporal Spike Streams
by: Jiang, Liangyan, et al.
Published: (2024)
by: Jiang, Liangyan, et al.
Published: (2024)
Joint Imaging-ROI Representation Learning via Cross-View Contrastive Alignment for Brain Disorder Classification
by: Liang, Wei, et al.
Published: (2026)
by: Liang, Wei, et al.
Published: (2026)
2D Gaussian Splatting with Semantic Alignment for Image Inpainting
by: Li, Hongyu, et al.
Published: (2025)
by: Li, Hongyu, et al.
Published: (2025)
MOSE: Monocular Semantic Reconstruction Using NeRF-Lifted Noisy Priors
by: Du, Zhenhua, et al.
Published: (2024)
by: Du, Zhenhua, et al.
Published: (2024)
Weakly-Supervised Image Forgery Localization via Vision-Language Collaborative Reasoning Framework
by: Sheng, Ziqi, et al.
Published: (2025)
by: Sheng, Ziqi, et al.
Published: (2025)
ReBrain: Brain MRI Reconstruction from Sparse CT Slice via Retrieval-Augmented Diffusion
by: Liu, Junming, et al.
Published: (2025)
by: Liu, Junming, et al.
Published: (2025)
Towards Unified Semantic and Controllable Image Fusion: A Diffusion Transformer Approach
by: Li, Jiayang, et al.
Published: (2025)
by: Li, Jiayang, et al.
Published: (2025)
RealignDiff: Boosting Text-to-Image Diffusion Model with Coarse-to-fine Semantic Re-alignment
by: Jiang, Zutao, et al.
Published: (2023)
by: Jiang, Zutao, et al.
Published: (2023)
LaRE$^2$: Latent Reconstruction Error Based Method for Diffusion-Generated Image Detection
by: Luo, Yunpeng, et al.
Published: (2024)
by: Luo, Yunpeng, et al.
Published: (2024)
Similar Items
-
Animate Your Thoughts: Decoupled Reconstruction of Dynamic Natural Vision from Slow Brain Activity
by: Lu, Yizhuo, et al.
Published: (2024) -
CLIP-MUSED: CLIP-Guided Multi-Subject Visual Neural Information Semantic Decoding
by: Zhou, Qiongyi, et al.
Published: (2024) -
NeuralOOD: Improving Out-of-Distribution Generalization Performance with Brain-machine Fusion Learning Framework
by: Zhao, Shuangchen, et al.
Published: (2024) -
Mind-Omni: A Unified Multi-Task Framework for Brain-Vision-Language Modeling via Discrete Diffusion
by: Lu, Yizhuo, et al.
Published: (2026) -
SemVideo: Reconstructs What You Watch from Brain Activity via Hierarchical Semantic Guidance
by: Yang, Minghan, et al.
Published: (2026)