Optical Flow Representation Alignment Mamba Diffusion Model for Medical Video Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Zhenbin, Zhang, Lei, Wang, Lituan, Zhu, Minjuan, Zhang, Zhenwei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Soft Masked Mamba Diffusion Model for CT to MRI Conversion
by: Wang, Zhenbin, et al.
Published: (2024)
by: Wang, Zhenbin, et al.
Published: (2024)
LanDA: Language-Guided Multi-Source Domain Adaptation
by: Wang, Zhenbin, et al.
Published: (2024)
by: Wang, Zhenbin, et al.
Published: (2024)
EAUWSeg: Eliminating annotation uncertainty in weakly-supervised medical image segmentation
by: Lituan, Wang, et al.
Published: (2025)
by: Lituan, Wang, et al.
Published: (2025)
PCLMix: Weakly Supervised Medical Image Segmentation via Pixel-Level Contrastive Learning and Dynamic Mix Augmentation
by: Lei, Yu, et al.
Published: (2024)
by: Lei, Yu, et al.
Published: (2024)
Training-Free Representation Guidance for Diffusion Models with a Representation Alignment Projector
by: Zu, Wenqiang, et al.
Published: (2026)
by: Zu, Wenqiang, et al.
Published: (2026)
A Physical Coherence Benchmark for Evaluating Video Generation Models via Optical Flow-guided Frame Prediction
by: Chen, Yongfan, et al.
Published: (2025)
by: Chen, Yongfan, et al.
Published: (2025)
FaRMamba: Frequency-based learning and Reconstruction aided Mamba for Medical Segmentation
by: Rong, Ze, et al.
Published: (2025)
by: Rong, Ze, et al.
Published: (2025)
MLVTG: Mamba-Based Feature Alignment and LLM-Driven Purification for Multi-Modal Video Temporal Grounding
by: Zhu, Zhiyi, et al.
Published: (2025)
by: Zhu, Zhiyi, et al.
Published: (2025)
MagDiff: Multi-Alignment Diffusion for High-Fidelity Video Generation and Editing
by: Zhao, Haoyu, et al.
Published: (2023)
by: Zhao, Haoyu, et al.
Published: (2023)
Surgical-MambaLLM: Mamba2-enhanced Multimodal Large Language Model for VQLA in Robotic Surgery
by: Hao, Pengfei, et al.
Published: (2025)
by: Hao, Pengfei, et al.
Published: (2025)
MambaFlow: A Novel and Flow-guided State Space Model for Scene Flow Estimation
by: Luo, Jiehao, et al.
Published: (2025)
by: Luo, Jiehao, et al.
Published: (2025)
MedGen: Unlocking Medical Video Generation by Scaling Granularly-annotated Medical Videos
by: Wang, Rongsheng, et al.
Published: (2025)
by: Wang, Rongsheng, et al.
Published: (2025)
FlowAct-R1: Towards Interactive Humanoid Video Generation
by: Wang, Lizhen, et al.
Published: (2026)
by: Wang, Lizhen, et al.
Published: (2026)
EMMA: Empowering Multi-modal Mamba with Structural and Hierarchical Alignment
by: Xing, Yifei, et al.
Published: (2024)
by: Xing, Yifei, et al.
Published: (2024)
Personalized Safety Alignment for Text-to-Image Diffusion Models
by: Lei, Yu, et al.
Published: (2025)
by: Lei, Yu, et al.
Published: (2025)
FADE: Frequency-Aware Diffusion Model Factorization for Video Editing
by: Zhu, Yixuan, et al.
Published: (2025)
by: Zhu, Yixuan, et al.
Published: (2025)
Surface Vision Mamba: Leveraging Bidirectional State Space Model for Efficient Spherical Manifold Representation
by: He, Rongzhao, et al.
Published: (2025)
by: He, Rongzhao, et al.
Published: (2025)
Attention-Mamba: A Mamba-Enhanced Multi-Scale Parallel Inference Network for Medical Image Segmentation
by: Zhang, Yanhua, et al.
Published: (2024)
by: Zhang, Yanhua, et al.
Published: (2024)
HMAFlow: Learning More Accurate Optical Flow via Hierarchical Motion Field Alignment
by: Ma, Dianbo, et al.
Published: (2024)
by: Ma, Dianbo, et al.
Published: (2024)
Consistent Video Editing as Flow-Driven Image-to-Video Generation
by: Wang, Ge, et al.
Published: (2025)
by: Wang, Ge, et al.
Published: (2025)
Not Like Transformers: Drop the Beat Representation for Dance Generation with Mamba-Based Diffusion Model
by: Park, Sangjune, et al.
Published: (2026)
by: Park, Sangjune, et al.
Published: (2026)
MOFA-Video: Controllable Image Animation via Generative Motion Field Adaptions in Frozen Image-to-Video Diffusion Model
by: Niu, Muyao, et al.
Published: (2024)
by: Niu, Muyao, et al.
Published: (2024)
BIVDiff: A Training-Free Framework for General-Purpose Video Synthesis via Bridging Image and Video Diffusion Models
by: Shi, Fengyuan, et al.
Published: (2023)
by: Shi, Fengyuan, et al.
Published: (2023)
Discrete Diffusion Models with MLLMs for Unified Medical Multimodal Generation
by: Mao, Jiawei, et al.
Published: (2025)
by: Mao, Jiawei, et al.
Published: (2025)
ELBO-T2IAlign: A Generic ELBO-Based Method for Calibrating Pixel-level Text-Image Alignment in Diffusion Models
by: Zhou, Qin, et al.
Published: (2025)
by: Zhou, Qin, et al.
Published: (2025)
NOVA3D: Normal Aligned Video Diffusion Model for Single Image to 3D Generation
by: Yang, Yuxiao, et al.
Published: (2025)
by: Yang, Yuxiao, et al.
Published: (2025)
AlignMamba: Enhancing Multimodal Mamba with Local and Global Cross-modal Alignment
by: Li, Yan, et al.
Published: (2024)
by: Li, Yan, et al.
Published: (2024)
Mamba-VMR: Multimodal Query Augmentation via Generated Videos for Precise Temporal Grounding
by: Sun, Yunzhuo, et al.
Published: (2026)
by: Sun, Yunzhuo, et al.
Published: (2026)
Scaling Diffusion Mamba with Bidirectional SSMs for Efficient Image and Video Generation
by: Mo, Shentong, et al.
Published: (2024)
by: Mo, Shentong, et al.
Published: (2024)
MAVIN: Multi-Action Video Generation with Diffusion Models via Transition Video Infilling
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
Online Iterative Self-Alignment for Radiology Report Generation
by: Xiao, Ting, et al.
Published: (2025)
by: Xiao, Ting, et al.
Published: (2025)
DIVD: Deblurring with Improved Video Diffusion Model
by: Long, Haoyang, et al.
Published: (2024)
by: Long, Haoyang, et al.
Published: (2024)
Timeline and Boundary Guided Diffusion Network for Video Shadow Detection
by: Zhou, Haipeng, et al.
Published: (2024)
by: Zhou, Haipeng, et al.
Published: (2024)
No Alignment Needed for Generation: Learning Linearly Separable Representations in Diffusion Models
by: Yun, Junno, et al.
Published: (2025)
by: Yun, Junno, et al.
Published: (2025)
Enhancing Medical Large Vision-Language Models via Alignment Distillation
by: Chang, Aofei, et al.
Published: (2025)
by: Chang, Aofei, et al.
Published: (2025)
Enabling Versatile Controls for Video Diffusion Models
by: Zhang, Xu, et al.
Published: (2025)
by: Zhang, Xu, et al.
Published: (2025)
Sat2Flow: A Structure-Aware Diffusion Framework for Human Flow Generation from Satellite Imagery
by: Wang, Xiangxu, et al.
Published: (2025)
by: Wang, Xiangxu, et al.
Published: (2025)
VideoDPO: Omni-Preference Alignment for Video Diffusion Generation
by: Liu, Runtao, et al.
Published: (2024)
by: Liu, Runtao, et al.
Published: (2024)
MSCrackMamba: Leveraging Vision Mamba for Crack Detection in Fused Multispectral Imagery
by: Zhu, Qinfeng, et al.
Published: (2024)
by: Zhu, Qinfeng, et al.
Published: (2024)
AnyFlow: Any-Step Video Diffusion Model with On-Policy Flow Map Distillation
by: Gu, Yuchao, et al.
Published: (2026)
by: Gu, Yuchao, et al.
Published: (2026)
Similar Items
-
Soft Masked Mamba Diffusion Model for CT to MRI Conversion
by: Wang, Zhenbin, et al.
Published: (2024) -
LanDA: Language-Guided Multi-Source Domain Adaptation
by: Wang, Zhenbin, et al.
Published: (2024) -
EAUWSeg: Eliminating annotation uncertainty in weakly-supervised medical image segmentation
by: Lituan, Wang, et al.
Published: (2025) -
PCLMix: Weakly Supervised Medical Image Segmentation via Pixel-Level Contrastive Learning and Dynamic Mix Augmentation
by: Lei, Yu, et al.
Published: (2024) -
Training-Free Representation Guidance for Diffusion Models with a Representation Alignment Projector
by: Zu, Wenqiang, et al.
Published: (2026)