Condition-Aware Neural Network for Controlled Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cai, Han, Li, Muyang, Zhang, Zhuoyang, Zhang, Qinsheng, Liu, Ming-Yu, Han, Song |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EfficientViT-SAM: Accelerated Segment Anything Model Without Accuracy Loss
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2024)
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2024)
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
von: Wu, Yecheng, et al.
Veröffentlicht: (2025)
Locality-aware Parallel Decoding for Efficient Autoregressive Image Generation
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2025)
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2025)
Spatial-Aware Latent Initialization for Controllable Image Generation
von: Sun, Wenqiang, et al.
Veröffentlicht: (2024)
von: Sun, Wenqiang, et al.
Veröffentlicht: (2024)
CPO: Condition Preference Optimization for Controllable Image Generation
von: Lyu, Zonglin, et al.
Veröffentlicht: (2025)
von: Lyu, Zonglin, et al.
Veröffentlicht: (2025)
Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models
von: Chen, Junyu, et al.
Veröffentlicht: (2024)
von: Chen, Junyu, et al.
Veröffentlicht: (2024)
CtrlNeRF: The Generative Neural Radiation Fields for the Controllable Synthesis of High-fidelity 3D-Aware Images
von: Liu, Jian, et al.
Veröffentlicht: (2024)
von: Liu, Jian, et al.
Veröffentlicht: (2024)
DistriFusion: Distributed Parallel Inference for High-Resolution Diffusion Models
von: Li, Muyang, et al.
Veröffentlicht: (2024)
von: Li, Muyang, et al.
Veröffentlicht: (2024)
JetViT: Efficient High-Resolution Vision Transformer with Post-Training Attention Search
von: Zou, Dongyun, et al.
Veröffentlicht: (2026)
von: Zou, Dongyun, et al.
Veröffentlicht: (2026)
DC-VideoGen: Efficient Video Generation with Deep Compression Video Autoencoder
von: Chen, Junyu, et al.
Veröffentlicht: (2025)
von: Chen, Junyu, et al.
Veröffentlicht: (2025)
Graph Conditioned Diffusion for Controllable Histopathology Image Generation
von: Cechnicka, Sarah, et al.
Veröffentlicht: (2025)
von: Cechnicka, Sarah, et al.
Veröffentlicht: (2025)
HART: Efficient Visual Generation with Hybrid Autoregressive Transformer
von: Tang, Haotian, et al.
Veröffentlicht: (2024)
von: Tang, Haotian, et al.
Veröffentlicht: (2024)
Spatial-Frequency Aware for Object Detection in RAW Image
von: Ye, Zhuohua, et al.
Veröffentlicht: (2025)
von: Ye, Zhuohua, et al.
Veröffentlicht: (2025)
Deformation-Invariant Neural Network and Its Applications in Distorted Image Restoration and Analysis
von: Zhang, Han, et al.
Veröffentlicht: (2023)
von: Zhang, Han, et al.
Veröffentlicht: (2023)
PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
Condition Weaving Meets Expert Modulation: Towards Universal and Controllable Image Generation
von: Zhang, Guoqing, et al.
Veröffentlicht: (2025)
von: Zhang, Guoqing, et al.
Veröffentlicht: (2025)
SmartDirector: Keyframe-Conditioned Cinematic Video Generation with Narrative Pacing Control
von: Zhang, Zhida, et al.
Veröffentlicht: (2026)
von: Zhang, Zhida, et al.
Veröffentlicht: (2026)
DC-Gen: Post-Training Diffusion Acceleration with Deeply Compressed Latent Space
von: He, Wenkun, et al.
Veröffentlicht: (2025)
von: He, Wenkun, et al.
Veröffentlicht: (2025)
LANS: A Layout-Aware Neural Solver for Plane Geometry Problem
von: Li, Zhong-Zhi, et al.
Veröffentlicht: (2023)
von: Li, Zhong-Zhi, et al.
Veröffentlicht: (2023)
Memory-based Cross-modal Semantic Alignment Network for Radiology Report Generation
von: Tao, Yitian, et al.
Veröffentlicht: (2024)
von: Tao, Yitian, et al.
Veröffentlicht: (2024)
SANA-Video: Efficient Video Generation with Block Linear Diffusion Transformer
von: Chen, Junsong, et al.
Veröffentlicht: (2025)
von: Chen, Junsong, et al.
Veröffentlicht: (2025)
ICM-SR: Image-Conditioned Manifold Regularization for Image Super-Resolution
von: Kang, Junoh, et al.
Veröffentlicht: (2025)
von: Kang, Junoh, et al.
Veröffentlicht: (2025)
Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model
von: Zhang, Tianqiu, et al.
Veröffentlicht: (2026)
von: Zhang, Tianqiu, et al.
Veröffentlicht: (2026)
Coordinating Multiple Conditions for Trajectory-Controlled Human Motion Generation
von: Cai, Deli, et al.
Veröffentlicht: (2026)
von: Cai, Deli, et al.
Veröffentlicht: (2026)
Epistemic Uncertainty for Generated Image Detection
von: Nie, Jun, et al.
Veröffentlicht: (2024)
von: Nie, Jun, et al.
Veröffentlicht: (2024)
Relation-Aware Meta-Learning for Zero-shot Sketch-Based Image Retrieval
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
MetaFind: Scene-Aware 3D Asset Retrieval for Coherent Metaverse Scene Generation
von: Pan, Zhenyu, et al.
Veröffentlicht: (2025)
von: Pan, Zhenyu, et al.
Veröffentlicht: (2025)
CGI: Identifying Conditional Generative Models with Example Images
von: Zhou, Zhi, et al.
Veröffentlicht: (2025)
von: Zhou, Zhi, et al.
Veröffentlicht: (2025)
ScaleNet: Scaling up Pretrained Neural Networks with Incremental Parameters
von: Hao, Zhiwei, et al.
Veröffentlicht: (2025)
von: Hao, Zhiwei, et al.
Veröffentlicht: (2025)
Exploring Annotation-free Image Captioning with Retrieval-augmented Pseudo Sentence Generation
von: Li, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Li, Zhiyuan, et al.
Veröffentlicht: (2023)
Zero-Shot Industrial Anomaly Segmentation with Image-Aware Prompt Generation
von: Park, SoYoung, et al.
Veröffentlicht: (2025)
von: Park, SoYoung, et al.
Veröffentlicht: (2025)
CLIP Model for Images to Textual Prompts Based on Top-k Neighbors
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
The Geometry of Compromise: Unlocking Generative Capabilities via Controllable Modality Alignment
von: Liu, Hongyuan, et al.
Veröffentlicht: (2026)
von: Liu, Hongyuan, et al.
Veröffentlicht: (2026)
CannyEdit: Selective Canny Control and Dual-Prompt Guidance for Training-Free Image Editing
von: Xie, Weiyan, et al.
Veröffentlicht: (2025)
von: Xie, Weiyan, et al.
Veröffentlicht: (2025)
Controllable Contextualized Image Captioning: Directing the Visual Narrative through User-Defined Highlights
von: Mao, Shunqi, et al.
Veröffentlicht: (2024)
von: Mao, Shunqi, et al.
Veröffentlicht: (2024)
Pack-PTQ: Advancing Post-training Quantization of Neural Networks by Pack-wise Reconstruction
von: Li, Changjun, et al.
Veröffentlicht: (2025)
von: Li, Changjun, et al.
Veröffentlicht: (2025)
Gradient-Direction-Aware Density Control for 3D Gaussian Splatting
von: Zhou, Zheng, et al.
Veröffentlicht: (2025)
von: Zhou, Zheng, et al.
Veröffentlicht: (2025)
CA-Edit: Causality-Aware Condition Adapter for High-Fidelity Local Facial Attribute Editing
von: Xian, Xiaole, et al.
Veröffentlicht: (2024)
von: Xian, Xiaole, et al.
Veröffentlicht: (2024)
VM-BHINet:Vision Mamba Bimanual Hand Interaction Network for 3D Interacting Hand Mesh Recovery From a Single RGB Image
von: Bi, Han, et al.
Veröffentlicht: (2025)
von: Bi, Han, et al.
Veröffentlicht: (2025)
Non-Markov Multi-Round Conversational Image Generation with History-Conditioned MLLMs
von: Zhang, Haochen, et al.
Veröffentlicht: (2026)
von: Zhang, Haochen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
EfficientViT-SAM: Accelerated Segment Anything Model Without Accuracy Loss
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2024) -
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer
von: Wu, Yecheng, et al.
Veröffentlicht: (2025) -
Locality-aware Parallel Decoding for Efficient Autoregressive Image Generation
von: Zhang, Zhuoyang, et al.
Veröffentlicht: (2025) -
Spatial-Aware Latent Initialization for Controllable Image Generation
von: Sun, Wenqiang, et al.
Veröffentlicht: (2024) -
CPO: Condition Preference Optimization for Controllable Image Generation
von: Lyu, Zonglin, et al.
Veröffentlicht: (2025)