Foodfusion: A Novel Approach for Food Image Composition via Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Shi, Chaohua, Wang, Xuan, Shi, Si, Wang, Xule, Zhu, Mingrui, Wang, Nannan, Gao, Xinbo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bridging Generative and Discriminative Models for Unified Visual Perception with Diffusion Priors
by: Dong, Shiyin, et al.
Published: (2024)
by: Dong, Shiyin, et al.
Published: (2024)
CSHNet: A Novel Information Asymmetric Image Translation Method
by: Yang, Xi, et al.
Published: (2025)
by: Yang, Xi, et al.
Published: (2025)
TeSG: Textual Semantic Guidance for Infrared and Visible Image Fusion
by: Zhu, Mingrui, et al.
Published: (2025)
by: Zhu, Mingrui, et al.
Published: (2025)
Disentangle Before Anonymize: A Two-stage Framework for Attribute-preserved and Occlusion-robust De-identification
by: Zhu, Mingrui, et al.
Published: (2023)
by: Zhu, Mingrui, et al.
Published: (2023)
Mixture of Ranks with Degradation-Aware Routing for One-Step Real-World Image Super-Resolution
by: He, Xiao, et al.
Published: (2025)
by: He, Xiao, et al.
Published: (2025)
Effective Diffusion Transformer Architecture for Image Super-Resolution
by: Cheng, Kun, et al.
Published: (2024)
by: Cheng, Kun, et al.
Published: (2024)
One-Step Diffusion-based Real-World Image Super-Resolution with Visual Perception Distillation
by: Wu, Xue, et al.
Published: (2025)
by: Wu, Xue, et al.
Published: (2025)
Imperceptible Face Forgery Attack via Adversarial Semantic Mask
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization
by: Yang, Xi, et al.
Published: (2025)
by: Yang, Xi, et al.
Published: (2025)
Lightweight RGB-D Salient Object Detection from a Speed-Accuracy Tradeoff Perspective
by: Duan, Songsong, et al.
Published: (2025)
by: Duan, Songsong, et al.
Published: (2025)
One Step Diffusion-based Super-Resolution with Time-Aware Distillation
by: He, Xiao, et al.
Published: (2024)
by: He, Xiao, et al.
Published: (2024)
Hierarchical Identity Learning for Unsupervised Visible-Infrared Person Re-Identification
by: Shi, Haonan, et al.
Published: (2025)
by: Shi, Haonan, et al.
Published: (2025)
InstructBrush: Learning Attention-based Instruction Optimization for Image Editing
by: Zhao, Ruoyu, et al.
Published: (2024)
by: Zhao, Ruoyu, et al.
Published: (2024)
3D Test-time Adaptation via Graph Spectral Driven Point Shift
by: Wei, Xin, et al.
Published: (2025)
by: Wei, Xin, et al.
Published: (2025)
Generalizable Prompt Learning of CLIP: A Brief Overview
by: Cui, Fangming, et al.
Published: (2025)
by: Cui, Fangming, et al.
Published: (2025)
Thinking Racial Bias in Fair Forgery Detection: Models, Datasets and Evaluations
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
Improving Adversarial Robustness via Decoupled Visual Representation Masking
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
iFADIT: Invertible Face Anonymization via Disentangled Identity Transform
by: Yuan, Lin, et al.
Published: (2025)
by: Yuan, Lin, et al.
Published: (2025)
Exploring Homogeneous and Heterogeneous Consistent Label Associations for Unsupervised Visible-Infrared Person ReID
by: He, Lingfeng, et al.
Published: (2024)
by: He, Lingfeng, et al.
Published: (2024)
Semantic-Aligned Learning with Collaborative Refinement for Unsupervised VI-ReID
by: Cheng, De, et al.
Published: (2025)
by: Cheng, De, et al.
Published: (2025)
Improving White-box Robustness of Pre-processing Defenses via Joint Adversarial Training
by: Zhou, Dawei, et al.
Published: (2021)
by: Zhou, Dawei, et al.
Published: (2021)
Structure-Accurate Medical Image Translation via Dynamic Frequency Balance and Knowledge Guidance
by: Xu, Jiahua, et al.
Published: (2025)
by: Xu, Jiahua, et al.
Published: (2025)
EtC: Temporal Boundary Expand then Clarify for Weakly Supervised Video Grounding with Multimodal Large Language Model
by: Li, Guozhang, et al.
Published: (2023)
by: Li, Guozhang, et al.
Published: (2023)
Masked Attribute Description Embedding for Cloth-Changing Person Re-identification
by: Peng, Chunlei, et al.
Published: (2024)
by: Peng, Chunlei, et al.
Published: (2024)
Motion Artifact Removal in Pixel-Frequency Domain via Alternate Masks and Diffusion Model
by: Xu, Jiahua, et al.
Published: (2024)
by: Xu, Jiahua, et al.
Published: (2024)
Mamba-CL: Optimizing Selective State Space Model in Null Space for Continual Learning
by: Cheng, De, et al.
Published: (2024)
by: Cheng, De, et al.
Published: (2024)
Federated Face Forgery Detection Learning with Personalized Representation
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
Knowledge-Enhanced Facial Expression Recognition with Emotional-to-Neutral Transformation
by: Li, Hangyu, et al.
Published: (2024)
by: Li, Hangyu, et al.
Published: (2024)
Prompt Disentanglement via Language Guidance and Representation Alignment for Domain Generalization
by: Cheng, De, et al.
Published: (2025)
by: Cheng, De, et al.
Published: (2025)
CPDM: Content-Preserving Diffusion Model for Underwater Image Enhancement
by: Shi, Xiaowen, et al.
Published: (2024)
by: Shi, Xiaowen, et al.
Published: (2024)
Free-Mask: A Novel Paradigm of Integration Between the Segmentation Diffusion Model and Image Editing
by: Gao, Bo, et al.
Published: (2024)
by: Gao, Bo, et al.
Published: (2024)
CKAA: Cross-subspace Knowledge Alignment and Aggregation for Robust Continual Learning
by: He, Lingfeng, et al.
Published: (2025)
by: He, Lingfeng, et al.
Published: (2025)
EKPC: Elastic Knowledge Preservation and Compensation for Class-Incremental Learning
by: Wang, Huaijie, et al.
Published: (2025)
by: Wang, Huaijie, et al.
Published: (2025)
Versatile Diffusion: Text, Images and Variations All in One Diffusion Model
by: Xu, Xingqian, et al.
Published: (2022)
by: Xu, Xingqian, et al.
Published: (2022)
Bridging Data Trials and Task Barriers: A Unified Framework for Sketch Biometric Identification
by: Liu, Decheng, et al.
Published: (2026)
by: Liu, Decheng, et al.
Published: (2026)
Task-Driven Subspace Decomposition for Knowledge Sharing and Isolation in LoRA-based Continual Learning
by: He, Lingfeng, et al.
Published: (2026)
by: He, Lingfeng, et al.
Published: (2026)
Efficient Bilateral Cross-Modality Cluster Matching for Unsupervised Visible-Infrared Person ReID
by: Cheng, De, et al.
Published: (2023)
by: Cheng, De, et al.
Published: (2023)
Consistent-1-to-3: Consistent Image to 3D View Synthesis via Geometry-aware Diffusion Models
by: Ye, Jianglong, et al.
Published: (2023)
by: Ye, Jianglong, et al.
Published: (2023)
MooD: Perception-Enhanced Efficient Affective Image Editing via Continuous Valence-Arousal Modeling
by: Yin, Xinyi, et al.
Published: (2026)
by: Yin, Xinyi, et al.
Published: (2026)
IMG: Calibrating Diffusion Models via Implicit Multimodal Guidance
by: Guo, Jiayi, et al.
Published: (2025)
by: Guo, Jiayi, et al.
Published: (2025)
Similar Items
-
Bridging Generative and Discriminative Models for Unified Visual Perception with Diffusion Priors
by: Dong, Shiyin, et al.
Published: (2024) -
CSHNet: A Novel Information Asymmetric Image Translation Method
by: Yang, Xi, et al.
Published: (2025) -
TeSG: Textual Semantic Guidance for Infrared and Visible Image Fusion
by: Zhu, Mingrui, et al.
Published: (2025) -
Disentangle Before Anonymize: A Two-stage Framework for Attribute-preserved and Occlusion-robust De-identification
by: Zhu, Mingrui, et al.
Published: (2023) -
Mixture of Ranks with Degradation-Aware Routing for One-Step Real-World Image Super-Resolution
by: He, Xiao, et al.
Published: (2025)