CtrLoRA: An Extensible and Efficient Framework for Controllable Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Yifeng, He, Zhenliang, Shan, Shiguang, Chen, Xilin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FullLoRA: Efficiently Boosting the Robustness of Pretrained Vision Transformers
von: Yuan, Zheng, et al.
Veröffentlicht: (2024)
von: Yuan, Zheng, et al.
Veröffentlicht: (2024)
Jodi: Unification of Visual Generation and Understanding via Joint Modeling
von: Xu, Yifeng, et al.
Veröffentlicht: (2025)
von: Xu, Yifeng, et al.
Veröffentlicht: (2025)
T2IShield: Defending Against Backdoors on Text-to-Image Diffusion Models
von: Wang, Zhongqi, et al.
Veröffentlicht: (2024)
von: Wang, Zhongqi, et al.
Veröffentlicht: (2024)
Dynamic Attention Analysis for Backdoor Detection in Text-to-Image Diffusion Models
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025)
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025)
GLip: A Global-Local Integrated Progressive Framework for Robust Visual Speech Recognition
von: Wang, Tianyue, et al.
Veröffentlicht: (2025)
von: Wang, Tianyue, et al.
Veröffentlicht: (2025)
OSI: One-step Inversion Excels in Extracting Diffusion Watermarks
von: Chen, Yuwei, et al.
Veröffentlicht: (2026)
von: Chen, Yuwei, et al.
Veröffentlicht: (2026)
EfficientMT: Efficient Temporal Adaptation for Motion Transfer in Text-to-Video Diffusion Models
von: Cai, Yufei, et al.
Veröffentlicht: (2025)
von: Cai, Yufei, et al.
Veröffentlicht: (2025)
JoPano: Unified Panorama Generation via Joint Modeling
von: Feng, Wancheng, et al.
Veröffentlicht: (2025)
von: Feng, Wancheng, et al.
Veröffentlicht: (2025)
Trigger without Trace: Towards Stealthy Backdoor Attack on Text-to-Image Diffusion Models
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
V-Attack: Targeting Disentangled Value Features for Controllable Adversarial Attacks on LVLMs
von: Nie, Sen, et al.
Veröffentlicht: (2025)
von: Nie, Sen, et al.
Veröffentlicht: (2025)
UniPose: A Unified Multimodal Framework for Human Pose Comprehension, Generation and Editing
von: Li, Yiheng, et al.
Veröffentlicht: (2024)
von: Li, Yiheng, et al.
Veröffentlicht: (2024)
Semantic or Covariate? A Study on the Intractable Case of Out-of-Distribution Detection
von: Long, Xingming, et al.
Veröffentlicht: (2024)
von: Long, Xingming, et al.
Veröffentlicht: (2024)
Rethinking the Evaluation of Out-of-Distribution Detection: A Sorites Paradox
von: Long, Xingming, et al.
Veröffentlicht: (2024)
von: Long, Xingming, et al.
Veröffentlicht: (2024)
VOPE: Revisiting Hallucination of Vision-Language Models in Voluntary Imagination Task
von: Long, Xingming, et al.
Veröffentlicht: (2025)
von: Long, Xingming, et al.
Veröffentlicht: (2025)
Assimilation Matters: Model-level Backdoor Detection in Vision-Language Pretrained Models
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025)
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025)
Task-adaptive Q-Face
von: Sun, Haomiao, et al.
Veröffentlicht: (2024)
von: Sun, Haomiao, et al.
Veröffentlicht: (2024)
StylizedGS: Controllable Stylization for 3D Gaussian Splatting
von: Zhang, Dingxi, et al.
Veröffentlicht: (2024)
von: Zhang, Dingxi, et al.
Veröffentlicht: (2024)
Towards Transferable Defense Against Malicious Image Edits
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
von: Zhang, Jie, et al.
Veröffentlicht: (2025)
AutoLoRA: Automatic LoRA Retrieval and Fine-Grained Gated Fusion for Text-to-Image Generation
von: Li, Zhiwen, et al.
Veröffentlicht: (2025)
von: Li, Zhiwen, et al.
Veröffentlicht: (2025)
Towards Robust Semantic Segmentation against Patch-based Attack via Attention Refinement
von: Yuan, Zheng, et al.
Veröffentlicht: (2024)
von: Yuan, Zheng, et al.
Veröffentlicht: (2024)
Contrastive Spectral Rectification: Test-Time Defense towards Zero-shot Adversarial Robustness of CLIP
von: Nie, Sen, et al.
Veröffentlicht: (2026)
von: Nie, Sen, et al.
Veröffentlicht: (2026)
Component-Based Out-of-Distribution Detection
von: Liu, Wenrui, et al.
Veröffentlicht: (2026)
von: Liu, Wenrui, et al.
Veröffentlicht: (2026)
ACT Now: Preempting LVLM Hallucinations via Adaptive Context Integration
von: Yan, Bei, et al.
Veröffentlicht: (2026)
von: Yan, Bei, et al.
Veröffentlicht: (2026)
Neural Gate: Mitigating Privacy Risks in LVLMs via Neuron-Level Gradient Gating
von: Cao, Xiangkui, et al.
Veröffentlicht: (2026)
von: Cao, Xiangkui, et al.
Veröffentlicht: (2026)
EntropyScan: Towards Model-level Backdoor Detection in LVLMs via Visual Attention Entropy
von: Ge, Xuanyu, et al.
Veröffentlicht: (2026)
von: Ge, Xuanyu, et al.
Veröffentlicht: (2026)
HyperLoRA: Parameter-Efficient Adaptive Generation for Portrait Synthesis
von: Li, Mengtian, et al.
Veröffentlicht: (2025)
von: Li, Mengtian, et al.
Veröffentlicht: (2025)
Video2LoRA: Unified Semantic-Controlled Video Generation via Per-Reference-Video LoRA
von: Wu, Zexi, et al.
Veröffentlicht: (2026)
von: Wu, Zexi, et al.
Veröffentlicht: (2026)
Anonymization Prompt Learning for Facial Privacy-Preserving Text-to-Image Generation
von: Shi, Liang, et al.
Veröffentlicht: (2024)
von: Shi, Liang, et al.
Veröffentlicht: (2024)
DIVE: Inverting Conditional Diffusion Models for Discriminative Tasks
von: Li, Yinqi, et al.
Veröffentlicht: (2025)
von: Li, Yinqi, et al.
Veröffentlicht: (2025)
TC-LoRA: Temporally Modulated Conditional LoRA for Adaptive Diffusion Control
von: Cho, Minkyoung, et al.
Veröffentlicht: (2025)
von: Cho, Minkyoung, et al.
Veröffentlicht: (2025)
MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models
von: Yan, Bei, et al.
Veröffentlicht: (2024)
von: Yan, Bei, et al.
Veröffentlicht: (2024)
Semantic Mismatch and Perceptual Degradation: A New Perspective on Image Editing Immunity
von: Dong, Shuai, et al.
Veröffentlicht: (2025)
von: Dong, Shuai, et al.
Veröffentlicht: (2025)
LoRA of Change: Learning to Generate LoRA for the Editing Instruction from A Single Before-After Image Pair
von: Song, Xue, et al.
Veröffentlicht: (2024)
von: Song, Xue, et al.
Veröffentlicht: (2024)
Generalized Semi-Supervised Learning via Self-Supervised Feature Adaptation
von: Liang, Jiachen, et al.
Veröffentlicht: (2024)
von: Liang, Jiachen, et al.
Veröffentlicht: (2024)
LiON-LoRA: Rethinking LoRA Fusion to Unify Controllable Spatial and Temporal Generation for Video Diffusion
von: Zhang, Yisu, et al.
Veröffentlicht: (2025)
von: Zhang, Yisu, et al.
Veröffentlicht: (2025)
LoRA-IR: Taming Low-Rank Experts for Efficient All-in-One Image Restoration
von: Ai, Yuang, et al.
Veröffentlicht: (2024)
von: Ai, Yuang, et al.
Veröffentlicht: (2024)
HPNet: Dynamic Trajectory Forecasting with Historical Prediction Attention
von: Tang, Xiaolong, et al.
Veröffentlicht: (2024)
von: Tang, Xiaolong, et al.
Veröffentlicht: (2024)
UMFC: Unsupervised Multi-Domain Feature Calibration for Vision-Language Models
von: Liang, Jiachen, et al.
Veröffentlicht: (2024)
von: Liang, Jiachen, et al.
Veröffentlicht: (2024)
un$^2$CLIP: Improving CLIP's Visual Detail Capturing Ability via Inverting unCLIP
von: Li, Yinqi, et al.
Veröffentlicht: (2025)
von: Li, Yinqi, et al.
Veröffentlicht: (2025)
Revisiting Logit Distributions for Reliable Out-of-Distribution Detection
von: Liang, Jiachen, et al.
Veröffentlicht: (2025)
von: Liang, Jiachen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FullLoRA: Efficiently Boosting the Robustness of Pretrained Vision Transformers
von: Yuan, Zheng, et al.
Veröffentlicht: (2024) -
Jodi: Unification of Visual Generation and Understanding via Joint Modeling
von: Xu, Yifeng, et al.
Veröffentlicht: (2025) -
T2IShield: Defending Against Backdoors on Text-to-Image Diffusion Models
von: Wang, Zhongqi, et al.
Veröffentlicht: (2024) -
Dynamic Attention Analysis for Backdoor Detection in Text-to-Image Diffusion Models
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025) -
GLip: A Global-Local Integrated Progressive Framework for Robust Visual Speech Recognition
von: Wang, Tianyue, et al.
Veröffentlicht: (2025)