Bridging Generative and Discriminative Models for Unified Visual Perception with Diffusion Priors
Fuente:
arXiv
Saved in:
| Main Authors: | Dong, Shiyin, Zhu, Mingrui, Cheng, Kun, Wang, Nannan, Gao, Xinbo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Foodfusion: A Novel Approach for Food Image Composition via Diffusion Models
by: Shi, Chaohua, et al.
Published: (2024)
by: Shi, Chaohua, et al.
Published: (2024)
Mixture of Ranks with Degradation-Aware Routing for One-Step Real-World Image Super-Resolution
by: He, Xiao, et al.
Published: (2025)
by: He, Xiao, et al.
Published: (2025)
Disentangle Before Anonymize: A Two-stage Framework for Attribute-preserved and Occlusion-robust De-identification
by: Zhu, Mingrui, et al.
Published: (2023)
by: Zhu, Mingrui, et al.
Published: (2023)
TeSG: Textual Semantic Guidance for Infrared and Visible Image Fusion
by: Zhu, Mingrui, et al.
Published: (2025)
by: Zhu, Mingrui, et al.
Published: (2025)
Effective Diffusion Transformer Architecture for Image Super-Resolution
by: Cheng, Kun, et al.
Published: (2024)
by: Cheng, Kun, et al.
Published: (2024)
One-Step Diffusion-based Real-World Image Super-Resolution with Visual Perception Distillation
by: Wu, Xue, et al.
Published: (2025)
by: Wu, Xue, et al.
Published: (2025)
One Step Diffusion-based Super-Resolution with Time-Aware Distillation
by: He, Xiao, et al.
Published: (2024)
by: He, Xiao, et al.
Published: (2024)
Bridging Data Trials and Task Barriers: A Unified Framework for Sketch Biometric Identification
by: Liu, Decheng, et al.
Published: (2026)
by: Liu, Decheng, et al.
Published: (2026)
Aligning Generative Denoising with Discriminative Objectives Unleashes Diffusion for Visual Perception
by: Pang, Ziqi, et al.
Published: (2025)
by: Pang, Ziqi, et al.
Published: (2025)
Exploring Homogeneous and Heterogeneous Consistent Label Associations for Unsupervised Visible-Infrared Person ReID
by: He, Lingfeng, et al.
Published: (2024)
by: He, Lingfeng, et al.
Published: (2024)
Semantic-Aligned Learning with Collaborative Refinement for Unsupervised VI-ReID
by: Cheng, De, et al.
Published: (2025)
by: Cheng, De, et al.
Published: (2025)
Improving Adversarial Robustness via Decoupled Visual Representation Masking
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization
by: Yang, Xi, et al.
Published: (2025)
by: Yang, Xi, et al.
Published: (2025)
Lightweight RGB-D Salient Object Detection from a Speed-Accuracy Tradeoff Perspective
by: Duan, Songsong, et al.
Published: (2025)
by: Duan, Songsong, et al.
Published: (2025)
EtC: Temporal Boundary Expand then Clarify for Weakly Supervised Video Grounding with Multimodal Large Language Model
by: Li, Guozhang, et al.
Published: (2023)
by: Li, Guozhang, et al.
Published: (2023)
Visual Bridge: Universal Visual Perception Representations Generating
by: Gao, Yilin, et al.
Published: (2025)
by: Gao, Yilin, et al.
Published: (2025)
DDAE++: Enhancing Diffusion Models Towards Unified Generative and Discriminative Learning
by: Xiang, Weilai, et al.
Published: (2025)
by: Xiang, Weilai, et al.
Published: (2025)
Prior Does Matter: Visual Navigation via Denoising Diffusion Bridge Models
by: Ren, Hao, et al.
Published: (2025)
by: Ren, Hao, et al.
Published: (2025)
Diffusion Models Need Visual Priors for Image Generation
by: Yue, Xiaoyu, et al.
Published: (2024)
by: Yue, Xiaoyu, et al.
Published: (2024)
Prompt Disentanglement via Language Guidance and Representation Alignment for Domain Generalization
by: Cheng, De, et al.
Published: (2025)
by: Cheng, De, et al.
Published: (2025)
Thinking Racial Bias in Fair Forgery Detection: Models, Datasets and Evaluations
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
PriorFusion: Unified Integration of Priors for Robust Road Perception in Autonomous Driving
by: Tang, Xuewei, et al.
Published: (2025)
by: Tang, Xuewei, et al.
Published: (2025)
CSHNet: A Novel Information Asymmetric Image Translation Method
by: Yang, Xi, et al.
Published: (2025)
by: Yang, Xi, et al.
Published: (2025)
Mamba-CL: Optimizing Selective State Space Model in Null Space for Continual Learning
by: Cheng, De, et al.
Published: (2024)
by: Cheng, De, et al.
Published: (2024)
Imperceptible Face Forgery Attack via Adversarial Semantic Mask
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
A Knowledge-guided Adversarial Defense for Resisting Malicious Visual Manipulation
by: Zhou, Dawei, et al.
Published: (2025)
by: Zhou, Dawei, et al.
Published: (2025)
CKAA: Cross-subspace Knowledge Alignment and Aggregation for Robust Continual Learning
by: He, Lingfeng, et al.
Published: (2025)
by: He, Lingfeng, et al.
Published: (2025)
EKPC: Elastic Knowledge Preservation and Compensation for Class-Incremental Learning
by: Wang, Huaijie, et al.
Published: (2025)
by: Wang, Huaijie, et al.
Published: (2025)
3D Test-time Adaptation via Graph Spectral Driven Point Shift
by: Wei, Xin, et al.
Published: (2025)
by: Wei, Xin, et al.
Published: (2025)
Towards Generalized Proactive Defense against Face Swapping with Contour-Hybrid Watermark
by: Xia, Ruiyang, et al.
Published: (2025)
by: Xia, Ruiyang, et al.
Published: (2025)
InstructBrush: Learning Attention-based Instruction Optimization for Image Editing
by: Zhao, Ruoyu, et al.
Published: (2024)
by: Zhao, Ruoyu, et al.
Published: (2024)
Masked Attribute Description Embedding for Cloth-Changing Person Re-identification
by: Peng, Chunlei, et al.
Published: (2024)
by: Peng, Chunlei, et al.
Published: (2024)
Task-Driven Subspace Decomposition for Knowledge Sharing and Isolation in LoRA-based Continual Learning
by: He, Lingfeng, et al.
Published: (2026)
by: He, Lingfeng, et al.
Published: (2026)
Hierarchical Identity Learning for Unsupervised Visible-Infrared Person Re-Identification
by: Shi, Haonan, et al.
Published: (2025)
by: Shi, Haonan, et al.
Published: (2025)
Efficient Bilateral Cross-Modality Cluster Matching for Unsupervised Visible-Infrared Person ReID
by: Cheng, De, et al.
Published: (2023)
by: Cheng, De, et al.
Published: (2023)
Exploiting Discriminative Codebook Prior for Autoregressive Image Generation
by: Tang, Longxiang, et al.
Published: (2025)
by: Tang, Longxiang, et al.
Published: (2025)
Federated Face Forgery Detection Learning with Personalized Representation
by: Liu, Decheng, et al.
Published: (2024)
by: Liu, Decheng, et al.
Published: (2024)
Knowledge-Enhanced Facial Expression Recognition with Emotional-to-Neutral Transformation
by: Li, Hangyu, et al.
Published: (2024)
by: Li, Hangyu, et al.
Published: (2024)
iFADIT: Invertible Face Anonymization via Disentangled Identity Transform
by: Yuan, Lin, et al.
Published: (2025)
by: Yuan, Lin, et al.
Published: (2025)
Steering Visual Generation in Unified Multimodal Models with Understanding Supervision
by: Liu, Zeyu, et al.
Published: (2026)
by: Liu, Zeyu, et al.
Published: (2026)
Similar Items
-
Foodfusion: A Novel Approach for Food Image Composition via Diffusion Models
by: Shi, Chaohua, et al.
Published: (2024) -
Mixture of Ranks with Degradation-Aware Routing for One-Step Real-World Image Super-Resolution
by: He, Xiao, et al.
Published: (2025) -
Disentangle Before Anonymize: A Two-stage Framework for Attribute-preserved and Occlusion-robust De-identification
by: Zhu, Mingrui, et al.
Published: (2023) -
TeSG: Textual Semantic Guidance for Infrared and Visible Image Fusion
by: Zhu, Mingrui, et al.
Published: (2025) -
Effective Diffusion Transformer Architecture for Image Super-Resolution
by: Cheng, Kun, et al.
Published: (2024)