Gespeichert in:
| Hauptverfasser: | Xie, Kangyang, Yang, Binbin, Chen, Hao, Wang, Meng, Zou, Cheng, Xue, Hui, Yang, Ming, Shen, Chunhua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2403.11077 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models
von: Wang, Wen, et al.
Veröffentlicht: (2023)
von: Wang, Wen, et al.
Veröffentlicht: (2023)
What Matters When Repurposing Diffusion Models for General Dense Perception Tasks?
von: Xu, Guangkai, et al.
Veröffentlicht: (2024)
von: Xu, Guangkai, et al.
Veröffentlicht: (2024)
Generative Video Matting
von: Ge, Yongtao, et al.
Veröffentlicht: (2025)
von: Ge, Yongtao, et al.
Veröffentlicht: (2025)
ZipGait: Bridging Skeleton and Silhouette with Diffusion Model for Advancing Gait Recognition
von: Min, Fanxu, et al.
Veröffentlicht: (2024)
von: Min, Fanxu, et al.
Veröffentlicht: (2024)
StyleTokenizer: Defining Image Style by a Single Instance for Controlling Diffusion Models
von: Li, Wen, et al.
Veröffentlicht: (2024)
von: Li, Wen, et al.
Veröffentlicht: (2024)
Diffusion Models are Efficient Data Generators for Human Mesh Recovery
von: Ge, Yongtao, et al.
Veröffentlicht: (2024)
von: Ge, Yongtao, et al.
Veröffentlicht: (2024)
FreeCompose: Generic Zero-Shot Image Composition with Diffusion Prior
von: Chen, Zhekai, et al.
Veröffentlicht: (2024)
von: Chen, Zhekai, et al.
Veröffentlicht: (2024)
Unleashing the Potential of the Diffusion Model in Few-shot Semantic Segmentation
von: Zhu, Muzhi, et al.
Veröffentlicht: (2024)
von: Zhu, Muzhi, et al.
Veröffentlicht: (2024)
Active-O3: Empowering Multimodal Large Language Models with Active Perception via GRPO
von: Zhu, Muzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Muzhi, et al.
Veröffentlicht: (2025)
Traffic Scene Parsing through the TSP6K Dataset
von: Jiang, Peng-Tao, et al.
Veröffentlicht: (2023)
von: Jiang, Peng-Tao, et al.
Veröffentlicht: (2023)
DiverGen: Improving Instance Segmentation by Learning Wider Data Distribution with More Diverse Generative Data
von: Fan, Chengxiang, et al.
Veröffentlicht: (2024)
von: Fan, Chengxiang, et al.
Veröffentlicht: (2024)
SegAgent: Exploring Pixel Understanding Capabilities in MLLMs by Imitating Human Annotator Trajectories
von: Zhu, Muzhi, et al.
Veröffentlicht: (2025)
von: Zhu, Muzhi, et al.
Veröffentlicht: (2025)
ForensicZip: More Tokens are Better but Not Necessary in Forensic Vision-Language Models
von: Lai, Yingxin, et al.
Veröffentlicht: (2026)
von: Lai, Yingxin, et al.
Veröffentlicht: (2026)
Video Virtual Try-on with Conditional Diffusion Transformer Inpainter
von: Zou, Cheng, et al.
Veröffentlicht: (2025)
von: Zou, Cheng, et al.
Veröffentlicht: (2025)
RGM: A Robust Generalizable Matching Model
von: Zhang, Songyan, et al.
Veröffentlicht: (2023)
von: Zhang, Songyan, et al.
Veröffentlicht: (2023)
Diffusion Knows Transparency: Repurposing Video Diffusion for Transparent Object Depth and Normal Estimation
von: Xu, Shaocong, et al.
Veröffentlicht: (2025)
von: Xu, Shaocong, et al.
Veröffentlicht: (2025)
ZipIR: Latent Pyramid Diffusion Transformer for High-Resolution Image Restoration
von: Yu, Yongsheng, et al.
Veröffentlicht: (2025)
von: Yu, Yongsheng, et al.
Veröffentlicht: (2025)
DICEPTION: A Generalist Diffusion Model for Visual Perceptual Tasks
von: Zhao, Canyu, et al.
Veröffentlicht: (2025)
von: Zhao, Canyu, et al.
Veröffentlicht: (2025)
Tinker: Diffusion's Gift to 3D--Multi-View Consistent Editing From Sparse Inputs without Per-Scene Optimization
von: Zhao, Canyu, et al.
Veröffentlicht: (2025)
von: Zhao, Canyu, et al.
Veröffentlicht: (2025)
Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching
von: Liu, Yang, et al.
Veröffentlicht: (2023)
von: Liu, Yang, et al.
Veröffentlicht: (2023)
Object-aware Inversion and Reassembly for Image Editing
von: Yang, Zhen, et al.
Veröffentlicht: (2023)
von: Yang, Zhen, et al.
Veröffentlicht: (2023)
UnZipLoRA: Separating Content and Style from a Single Image
von: Liu, Chang, et al.
Veröffentlicht: (2024)
von: Liu, Chang, et al.
Veröffentlicht: (2024)
HQ-DM: Single Hadamard Transformation-Based Quantization-Aware Training for Low-Bit Diffusion Models
von: Mao, Shizhuo, et al.
Veröffentlicht: (2025)
von: Mao, Shizhuo, et al.
Veröffentlicht: (2025)
Unpaired Deblurring via Decoupled Diffusion Model
von: Cheng, Junhao, et al.
Veröffentlicht: (2025)
von: Cheng, Junhao, et al.
Veröffentlicht: (2025)
A Geometric Perspective on Diffusion Models
von: Chen, Defang, et al.
Veröffentlicht: (2023)
von: Chen, Defang, et al.
Veröffentlicht: (2023)
MARBLE: Multi-Aspect Reward Balance for Diffusion RL
von: Zhao, Canyu, et al.
Veröffentlicht: (2026)
von: Zhao, Canyu, et al.
Veröffentlicht: (2026)
A Single Neuron Works: Precise Concept Erasure in Text-to-Image Diffusion Models
von: He, Qinqin, et al.
Veröffentlicht: (2025)
von: He, Qinqin, et al.
Veröffentlicht: (2025)
VisionZip: Longer is Better but Not Necessary in Vision Language Models
von: Yang, Senqiao, et al.
Veröffentlicht: (2024)
von: Yang, Senqiao, et al.
Veröffentlicht: (2024)
Paragraph-to-Image Generation with Information-Enriched Diffusion Model
von: Wu, Weijia, et al.
Veröffentlicht: (2023)
von: Wu, Weijia, et al.
Veröffentlicht: (2023)
Towards a Transparent and Interpretable AI Model for Medical Image Classifications
von: Wen, Binbin, et al.
Veröffentlicht: (2025)
von: Wen, Binbin, et al.
Veröffentlicht: (2025)
StaMo: Unsupervised Learning of Generalizable Robot Motion from Compact State Representation
von: Liu, Mingyu, et al.
Veröffentlicht: (2025)
von: Liu, Mingyu, et al.
Veröffentlicht: (2025)
DiffuMask: Synthesizing Images with Pixel-level Annotations for Semantic Segmentation Using Diffusion Models
von: Wu, Weijia, et al.
Veröffentlicht: (2023)
von: Wu, Weijia, et al.
Veröffentlicht: (2023)
Generative Active Learning for Long-tailed Instance Segmentation
von: Zhu, Muzhi, et al.
Veröffentlicht: (2024)
von: Zhu, Muzhi, et al.
Veröffentlicht: (2024)
FreeCustom: Tuning-Free Customized Image Generation for Multi-Concept Composition
von: Ding, Ganggui, et al.
Veröffentlicht: (2024)
von: Ding, Ganggui, et al.
Veröffentlicht: (2024)
A Simple Image Segmentation Framework via In-Context Examples
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
PromptRR: Diffusion Models as Prompt Generators for Single Image Reflection Removal
von: Wang, Tao, et al.
Veröffentlicht: (2024)
von: Wang, Tao, et al.
Veröffentlicht: (2024)
Distribution-Aware Data Expansion with Diffusion Models
von: Zhu, Haowei, et al.
Veröffentlicht: (2024)
von: Zhu, Haowei, et al.
Veröffentlicht: (2024)
On the Trajectory Regularity of ODE-based Diffusion Sampling
von: Chen, Defang, et al.
Veröffentlicht: (2024)
von: Chen, Defang, et al.
Veröffentlicht: (2024)
FIND: Fine-tuning Initial Noise Distribution with Policy Optimization for Diffusion Models
von: Chen, Changgu, et al.
Veröffentlicht: (2024)
von: Chen, Changgu, et al.
Veröffentlicht: (2024)
HieraTok: Multi-Scale Visual Tokenizer Improves Image Reconstruction and Generation
von: Chen, Cong, et al.
Veröffentlicht: (2025)
von: Chen, Cong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models
von: Wang, Wen, et al.
Veröffentlicht: (2023) -
What Matters When Repurposing Diffusion Models for General Dense Perception Tasks?
von: Xu, Guangkai, et al.
Veröffentlicht: (2024) -
Generative Video Matting
von: Ge, Yongtao, et al.
Veröffentlicht: (2025) -
ZipGait: Bridging Skeleton and Silhouette with Diffusion Model for Advancing Gait Recognition
von: Min, Fanxu, et al.
Veröffentlicht: (2024) -
StyleTokenizer: Defining Image Style by a Single Instance for Controlling Diffusion Models
von: Li, Wen, et al.
Veröffentlicht: (2024)