Diffusion-Guided Semantic Consistency for Multimodal Heterogeneity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Jing, Guo, Zhengliang, Wang, Yan, Zhu, Xiaoguang, Du, Yao, Wang, Zehua, Leung, Victor C. M. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation
von: Zheng, Wenjie, et al.
Veröffentlicht: (2026)
von: Zheng, Wenjie, et al.
Veröffentlicht: (2026)
Harnessing Group-Oriented Consistency Constraints for Semi-Supervised Semantic Segmentation in CdZnTe Semiconductors
von: Li, Peihao, et al.
Veröffentlicht: (2025)
von: Li, Peihao, et al.
Veröffentlicht: (2025)
GarmentDiffusion: 3D Garment Sewing Pattern Generation with Multimodal Diffusion Transformers
von: Li, Xinyu, et al.
Veröffentlicht: (2025)
von: Li, Xinyu, et al.
Veröffentlicht: (2025)
Task-Oriented Real-time Visual Inference for IoVT Systems: A Co-design Framework of Neural Networks and Edge Deployment
von: Wu, Jiaqi, et al.
Veröffentlicht: (2024)
von: Wu, Jiaqi, et al.
Veröffentlicht: (2024)
POCI-Diff: Position Objects Consistently and Interactively with 3D-Layout Guided Diffusion
von: Rigo, Andrea, et al.
Veröffentlicht: (2026)
von: Rigo, Andrea, et al.
Veröffentlicht: (2026)
A Dual-way Enhanced Framework from Text Matching Point of View for Multimodal Entity Linking
von: Song, Shezheng, et al.
Veröffentlicht: (2023)
von: Song, Shezheng, et al.
Veröffentlicht: (2023)
CLIP-MUSED: CLIP-Guided Multi-Subject Visual Neural Information Semantic Decoding
von: Zhou, Qiongyi, et al.
Veröffentlicht: (2024)
von: Zhou, Qiongyi, et al.
Veröffentlicht: (2024)
Reward Guided Latent Consistency Distillation
von: Li, Jiachen, et al.
Veröffentlicht: (2024)
von: Li, Jiachen, et al.
Veröffentlicht: (2024)
SSMG: Spatial-Semantic Map Guided Diffusion Model for Free-form Layout-to-Image Generation
von: Jia, Chengyou, et al.
Veröffentlicht: (2023)
von: Jia, Chengyou, et al.
Veröffentlicht: (2023)
The Thinking Pixel: Recursive Sparse Reasoning in Multimodal Diffusion Latents
von: Sun, Yuwei, et al.
Veröffentlicht: (2026)
von: Sun, Yuwei, et al.
Veröffentlicht: (2026)
Homogeneous and Heterogeneous Consistency progressive Re-ranking for Visible-Infrared Person Re-identification
von: Wang, Yiming
Veröffentlicht: (2026)
von: Wang, Yiming
Veröffentlicht: (2026)
CLIPin: A Non-contrastive Plug-in to CLIP for Multimodal Semantic Alignment
von: Yang, Shengzhu, et al.
Veröffentlicht: (2025)
von: Yang, Shengzhu, et al.
Veröffentlicht: (2025)
Semantic Localization Guiding Segment Anything Model For Reference Remote Sensing Image Segmentation
von: Li, Shuyang, et al.
Veröffentlicht: (2025)
von: Li, Shuyang, et al.
Veröffentlicht: (2025)
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
von: Lv, Zheqi, et al.
Veröffentlicht: (2025)
von: Lv, Zheqi, et al.
Veröffentlicht: (2025)
PointGS: Semantic-Consistent Unsupervised 3D Point Cloud Segmentation with 3D Gaussian Splatting
von: Song, Yixiao, et al.
Veröffentlicht: (2026)
von: Song, Yixiao, et al.
Veröffentlicht: (2026)
Reference-Guided Diffusion Inpainting For Multimodal Counterfactual Generation
von: Buburuzan, Alexandru
Veröffentlicht: (2025)
von: Buburuzan, Alexandru
Veröffentlicht: (2025)
SVGDreamer: Text Guided SVG Generation with Diffusion Model
von: Xing, Ximing, et al.
Veröffentlicht: (2023)
von: Xing, Ximing, et al.
Veröffentlicht: (2023)
Affinity-Graph-Guided Contractive Learning for Pretext-Free Medical Image Segmentation with Minimal Annotation
von: Cheng, Zehua, et al.
Veröffentlicht: (2024)
von: Cheng, Zehua, et al.
Veröffentlicht: (2024)
DreamStory: Open-Domain Story Visualization by LLM-Guided Multi-Subject Consistent Diffusion
von: He, Huiguo, et al.
Veröffentlicht: (2024)
von: He, Huiguo, et al.
Veröffentlicht: (2024)
Semantic Surgery: Zero-Shot Concept Erasure in Diffusion Models
von: Xiong, Lexiang, et al.
Veröffentlicht: (2025)
von: Xiong, Lexiang, et al.
Veröffentlicht: (2025)
Sam-Guided Enhanced Fine-Grained Encoding with Mixed Semantic Learning for Medical Image Captioning
von: Zhang, Zhenyu, et al.
Veröffentlicht: (2023)
von: Zhang, Zhenyu, et al.
Veröffentlicht: (2023)
Robust Polyp Detection and Diagnosis through Compositional Prompt-Guided Diffusion Models
von: Yu, Jia, et al.
Veröffentlicht: (2025)
von: Yu, Jia, et al.
Veröffentlicht: (2025)
Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models
von: Kim, Keuntae, et al.
Veröffentlicht: (2026)
von: Kim, Keuntae, et al.
Veröffentlicht: (2026)
Morphable Diffusion: 3D-Consistent Diffusion for Single-image Avatar Creation
von: Chen, Xiyi, et al.
Veröffentlicht: (2024)
von: Chen, Xiyi, et al.
Veröffentlicht: (2024)
Timeline and Boundary Guided Diffusion Network for Video Shadow Detection
von: Zhou, Haipeng, et al.
Veröffentlicht: (2024)
von: Zhou, Haipeng, et al.
Veröffentlicht: (2024)
GS-ID: Illumination Decomposition on Gaussian Splatting via Adaptive Light Aggregation and Diffusion-Guided Material Priors
von: Du, Kang, et al.
Veröffentlicht: (2024)
von: Du, Kang, et al.
Veröffentlicht: (2024)
Text-Guided Layer Fusion Mitigates Hallucination in Multimodal LLMs
von: Lin, Chenchen, et al.
Veröffentlicht: (2026)
von: Lin, Chenchen, et al.
Veröffentlicht: (2026)
MCITlib: Multimodal Continual Instruction Tuning Library and Benchmark
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
MuDD: A Multimodal Deception Detection Dataset and GSR-Guided Progressive Distillation for Non-Contact Deception Detection
von: Jiang, Peiyuan, et al.
Veröffentlicht: (2026)
von: Jiang, Peiyuan, et al.
Veröffentlicht: (2026)
TGC-Net: A Structure-Aware and Semantically-Aligned Framework for Text-Guided Medical Image Segmentation
von: Lin, Gaoren, et al.
Veröffentlicht: (2025)
von: Lin, Gaoren, et al.
Veröffentlicht: (2025)
AnchorDS: Anchoring Dynamic Sources for Semantically Consistent Text-to-3D Generation
von: Zhu, Jiayin, et al.
Veröffentlicht: (2025)
von: Zhu, Jiayin, et al.
Veröffentlicht: (2025)
Task Consistent Prototype Learning for Incremental Few-shot Semantic Segmentation
von: Xu, Wenbo, et al.
Veröffentlicht: (2024)
von: Xu, Wenbo, et al.
Veröffentlicht: (2024)
Exploring Homogeneous and Heterogeneous Consistent Label Associations for Unsupervised Visible-Infrared Person ReID
von: He, Lingfeng, et al.
Veröffentlicht: (2024)
von: He, Lingfeng, et al.
Veröffentlicht: (2024)
Mitigating Hallucinations in Large Vision-Language Models via Entity-Centric Multimodal Preference Optimization
von: Wu, Jiulong, et al.
Veröffentlicht: (2025)
von: Wu, Jiulong, et al.
Veröffentlicht: (2025)
SAM Guided Semantic and Motion Changed Region Mining for Remote Sensing Change Captioning
von: Wang, Futian, et al.
Veröffentlicht: (2025)
von: Wang, Futian, et al.
Veröffentlicht: (2025)
Phys4D: Fine-Grained Physics-Consistent 4D Modeling from Video Diffusion
von: Lu, Haoran, et al.
Veröffentlicht: (2026)
von: Lu, Haoran, et al.
Veröffentlicht: (2026)
SemHiTok: A Unified Image Tokenizer via Semantic-Guided Hierarchical Codebook for Multimodal Understanding and Generation
von: Chen, Zisheng, et al.
Veröffentlicht: (2025)
von: Chen, Zisheng, et al.
Veröffentlicht: (2025)
Bringing Diversity from Diffusion Models to Semantic-Guided Face Asset Generation
von: Cai, Yunxuan, et al.
Veröffentlicht: (2025)
von: Cai, Yunxuan, et al.
Veröffentlicht: (2025)
DECADE: A Temporally-Consistent Unsupervised Diffusion Model for Enhanced Rb-82 Dynamic Cardiac PET Image Denoising
von: Zhou, Yinchi, et al.
Veröffentlicht: (2026)
von: Zhou, Yinchi, et al.
Veröffentlicht: (2026)
DiffSketcher: Text Guided Vector Sketch Synthesis through Latent Diffusion Models
von: Xing, Ximing, et al.
Veröffentlicht: (2023)
von: Xing, Ximing, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation
von: Zheng, Wenjie, et al.
Veröffentlicht: (2026) -
Harnessing Group-Oriented Consistency Constraints for Semi-Supervised Semantic Segmentation in CdZnTe Semiconductors
von: Li, Peihao, et al.
Veröffentlicht: (2025) -
GarmentDiffusion: 3D Garment Sewing Pattern Generation with Multimodal Diffusion Transformers
von: Li, Xinyu, et al.
Veröffentlicht: (2025) -
Task-Oriented Real-time Visual Inference for IoVT Systems: A Co-design Framework of Neural Networks and Edge Deployment
von: Wu, Jiaqi, et al.
Veröffentlicht: (2024) -
POCI-Diff: Position Objects Consistently and Interactively with 3D-Layout Guided Diffusion
von: Rigo, Andrea, et al.
Veröffentlicht: (2026)