Gespeichert in:
| Hauptverfasser: | Chen, Changjian, Lv, Fei, Guan, Yalong, Wang, Pengcheng, Yu, Shengjie, Zhang, Yifan, Tang, Zhuo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2412.16839 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Small-Scale Dataset Expansion with Triplet-Connection-based Sample Re-Weighting
von: Xiang, Ting, et al.
Veröffentlicht: (2025)
von: Xiang, Ting, et al.
Veröffentlicht: (2025)
MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale
von: Tang, Zhicong, et al.
Veröffentlicht: (2026)
von: Tang, Zhicong, et al.
Veröffentlicht: (2026)
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
Data Extrapolation for Text-to-image Generation on Small Datasets
von: Ye, Senmao, et al.
Veröffentlicht: (2024)
von: Ye, Senmao, et al.
Veröffentlicht: (2024)
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
SkeleGuide: Explicit Skeleton Reasoning for Context-Aware Human-in-Place Image Synthesis
von: Wu, Chuqiao, et al.
Veröffentlicht: (2026)
von: Wu, Chuqiao, et al.
Veröffentlicht: (2026)
Invisible Clean-Label Backdoor Attacks for Generative Data Augmentation
von: Xiang, Ting, et al.
Veröffentlicht: (2026)
von: Xiang, Ting, et al.
Veröffentlicht: (2026)
Diffusion Cocktail: Mixing Domain-Specific Diffusion Models for Diversified Image Generations
von: Liu, Haoming, et al.
Veröffentlicht: (2023)
von: Liu, Haoming, et al.
Veröffentlicht: (2023)
GPT-NAS: Evolutionary Neural Architecture Search with the Generative Pre-Trained Model
von: Yu, Caiyang, et al.
Veröffentlicht: (2023)
von: Yu, Caiyang, et al.
Veröffentlicht: (2023)
A Comprehensive Dataset for Human vs. AI Generated Image Detection
von: Roy, Rajarshi, et al.
Veröffentlicht: (2026)
von: Roy, Rajarshi, et al.
Veröffentlicht: (2026)
Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion
von: Lv, Zheqi, et al.
Veröffentlicht: (2025)
von: Lv, Zheqi, et al.
Veröffentlicht: (2025)
A High-Quality Dataset and Reliable Evaluation for Interleaved Image-Text Generation
von: Feng, Yukang, et al.
Veröffentlicht: (2025)
von: Feng, Yukang, et al.
Veröffentlicht: (2025)
OnlineHOI: Towards Online Human-Object Interaction Generation and Perception
von: Ji, Yihong, et al.
Veröffentlicht: (2025)
von: Ji, Yihong, et al.
Veröffentlicht: (2025)
Using Synthetic Images to Augment Small Medical Image Datasets
von: Vu, Minh H., et al.
Veröffentlicht: (2025)
von: Vu, Minh H., et al.
Veröffentlicht: (2025)
ImageGem: In-the-wild Generative Image Interaction Dataset for Generative Model Personalization
von: Guo, Yuanhe, et al.
Veröffentlicht: (2025)
von: Guo, Yuanhe, et al.
Veröffentlicht: (2025)
CCIS-Diff: A Generative Model with Stable Diffusion Prior for Controlled Colonoscopy Image Synthesis
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
von: Xie, Yifan, et al.
Veröffentlicht: (2024)
Boosting Semi-Supervised Medical Image Segmentation via Masked Image Consistency and Discrepancy Learning
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)
MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing
von: Zhang, Kai, et al.
Veröffentlicht: (2023)
von: Zhang, Kai, et al.
Veröffentlicht: (2023)
Guiding Perception-Reasoning Closer to Human in Blind Image Quality Assessment
von: Li, Yuan, et al.
Veröffentlicht: (2025)
von: Li, Yuan, et al.
Veröffentlicht: (2025)
Image Forgery Localization via Guided Noise and Multi-Scale Feature Aggregation
von: Niu, Yakun, et al.
Veröffentlicht: (2024)
von: Niu, Yakun, et al.
Veröffentlicht: (2024)
RealGen: Photorealistic Text-to-Image Generation via Detector-Guided Rewards
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
von: Ye, Junyan, et al.
Veröffentlicht: (2025)
15M Multimodal Facial Image-Text Dataset
von: Dai, Dawei, et al.
Veröffentlicht: (2024)
von: Dai, Dawei, et al.
Veröffentlicht: (2024)
Guided and Variance-Corrected Fusion with One-shot Style Alignment for Large-Content Image Generation
von: Sun, Shoukun, et al.
Veröffentlicht: (2024)
von: Sun, Shoukun, et al.
Veröffentlicht: (2024)
Scaling Image Tokenizers with Grouped Spherical Quantization
von: Wang, Jiangtao, et al.
Veröffentlicht: (2024)
von: Wang, Jiangtao, et al.
Veröffentlicht: (2024)
Camyla: Scaling Autonomous Research in Medical Image Segmentation
von: Gao, Yifan, et al.
Veröffentlicht: (2026)
von: Gao, Yifan, et al.
Veröffentlicht: (2026)
Audio-Infused Automatic Image Colorization by Exploiting Audio Scene Semantics
von: Zhao, Pengcheng, et al.
Veröffentlicht: (2024)
von: Zhao, Pengcheng, et al.
Veröffentlicht: (2024)
Evaluating Text-to-Image Generative Models: An Empirical Study on Human Image Synthesis
von: Chen, Muxi, et al.
Veröffentlicht: (2024)
von: Chen, Muxi, et al.
Veröffentlicht: (2024)
LSSGen: Leveraging Latent Space Scaling in Flow and Diffusion for Efficient Text to Image Generation
von: Tang, Jyun-Ze, et al.
Veröffentlicht: (2025)
von: Tang, Jyun-Ze, et al.
Veröffentlicht: (2025)
Language-Guided Image Tokenization for Generation
von: Zha, Kaiwen, et al.
Veröffentlicht: (2024)
von: Zha, Kaiwen, et al.
Veröffentlicht: (2024)
Learning Patient-Specific Disease Dynamics with Latent Flow Matching for Longitudinal Imaging Generation
von: Chen, Hao, et al.
Veröffentlicht: (2025)
von: Chen, Hao, et al.
Veröffentlicht: (2025)
Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation
von: Chen, Junying, et al.
Veröffentlicht: (2025)
von: Chen, Junying, et al.
Veröffentlicht: (2025)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
OpenGPT-4o-Image: A Comprehensive Dataset for Advanced Image Generation and Editing
von: Chen, Zhihong, et al.
Veröffentlicht: (2025)
von: Chen, Zhihong, et al.
Veröffentlicht: (2025)
FlowSteer: Guiding Few-Step Image Synthesis with Authentic Trajectories
von: Ke, Lei, et al.
Veröffentlicht: (2025)
von: Ke, Lei, et al.
Veröffentlicht: (2025)
Training-Free Watermarking for Autoregressive Image Generation
von: Tong, Yu, et al.
Veröffentlicht: (2025)
von: Tong, Yu, et al.
Veröffentlicht: (2025)
Latent Action Control for Reasoning-Guided Unified Image Generation
von: Zhai, Fuxiang, et al.
Veröffentlicht: (2026)
von: Zhai, Fuxiang, et al.
Veröffentlicht: (2026)
Training-Free Text-Guided Image Editing with Visual Autoregressive Model
von: Wang, Yufei, et al.
Veröffentlicht: (2025)
von: Wang, Yufei, et al.
Veröffentlicht: (2025)
HP-Edit: A Human-Preference Post-Training Framework for Image Editing
von: Li, Fan, et al.
Veröffentlicht: (2026)
von: Li, Fan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Enhancing Small-Scale Dataset Expansion with Triplet-Connection-based Sample Re-Weighting
von: Xiang, Ting, et al.
Veröffentlicht: (2025) -
MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale
von: Tang, Zhicong, et al.
Veröffentlicht: (2026) -
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024) -
Data Extrapolation for Text-to-image Generation on Small Datasets
von: Ye, Senmao, et al.
Veröffentlicht: (2024) -
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)