Salvato in:
| Autori principali: | Jiang, Jing, Ling, Yiran, Li, Binzhu, Li, Pengxiang, Piao, Junming, Zhang, Yu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2407.06196 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Pixels, Patterns, but No Poetry: To See The World like Humans
di: Gao, Hongcheng, et al.
Pubblicazione: (2025)
di: Gao, Hongcheng, et al.
Pubblicazione: (2025)
Single Image Iterative Subject-driven Generation and Editing
di: Shpitzer, Yair, et al.
Pubblicazione: (2025)
di: Shpitzer, Yair, et al.
Pubblicazione: (2025)
Seeing the Poem: Image-Semantic Detection of AI-Generated Modern Chinese Poetry with MLLMs
di: Wang, Shanshan, et al.
Pubblicazione: (2026)
di: Wang, Shanshan, et al.
Pubblicazione: (2026)
Iterative Refinement Improves Compositional Image Generation
di: Jaiswal, Shantanu, et al.
Pubblicazione: (2026)
di: Jaiswal, Shantanu, et al.
Pubblicazione: (2026)
Iterative Adversarial Attack on Image-guided Story Ending Generation
di: Wang, Youze, et al.
Pubblicazione: (2023)
di: Wang, Youze, et al.
Pubblicazione: (2023)
Autoregressive Image Generation with Linear Complexity: A Spatial-Aware Decay Perspective
di: Mao, Yuxin, et al.
Pubblicazione: (2025)
di: Mao, Yuxin, et al.
Pubblicazione: (2025)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
di: Zhao, Yu, et al.
Pubblicazione: (2024)
di: Zhao, Yu, et al.
Pubblicazione: (2024)
DCMM-Transformer: Degree-Corrected Mixed-Membership Attention for Medical Imaging
di: Cheng, Huimin, et al.
Pubblicazione: (2025)
di: Cheng, Huimin, et al.
Pubblicazione: (2025)
Regeneration Based Training-free Attribution of Fake Images Generated by Text-to-Image Generative Models
di: Li, Meiling, et al.
Pubblicazione: (2024)
di: Li, Meiling, et al.
Pubblicazione: (2024)
ELBO-T2IAlign: A Generic ELBO-Based Method for Calibrating Pixel-level Text-Image Alignment in Diffusion Models
di: Zhou, Qin, et al.
Pubblicazione: (2025)
di: Zhou, Qin, et al.
Pubblicazione: (2025)
Loupe: A Generalizable and Adaptive Framework for Image Forgery Detection
di: Jiang, Yuchu, et al.
Pubblicazione: (2025)
di: Jiang, Yuchu, et al.
Pubblicazione: (2025)
Knowledge Completes the Vision: A Multimodal Entity-aware Retrieval-Augmented Generation Framework for News Image Captioning
di: You, Xiaoxing, et al.
Pubblicazione: (2025)
di: You, Xiaoxing, et al.
Pubblicazione: (2025)
StableI2I: Spotting Unintended Changes in Image-to-Image Transition
di: Li, Jiayang, et al.
Pubblicazione: (2026)
di: Li, Jiayang, et al.
Pubblicazione: (2026)
GenShield: Unified Detection and Artifact Correction for AI-Generated Images
di: Xu, Zhipei, et al.
Pubblicazione: (2026)
di: Xu, Zhipei, et al.
Pubblicazione: (2026)
A Framework For Image Synthesis Using Supervised Contrastive Learning
di: Liu, Yibin, et al.
Pubblicazione: (2024)
di: Liu, Yibin, et al.
Pubblicazione: (2024)
CookAnything: A Framework for Flexible and Consistent Multi-Step Recipe Image Generation
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
di: Zhang, Ruoxuan, et al.
Pubblicazione: (2025)
Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement
di: Jeong, Suchae, et al.
Pubblicazione: (2025)
di: Jeong, Suchae, et al.
Pubblicazione: (2025)
Condition-Aware Neural Network for Controlled Image Generation
di: Cai, Han, et al.
Pubblicazione: (2024)
di: Cai, Han, et al.
Pubblicazione: (2024)
Self-Corrected Image Generation with Explainable Latent Rewards
di: Luo, Yinyi, et al.
Pubblicazione: (2026)
di: Luo, Yinyi, et al.
Pubblicazione: (2026)
Relative-Absolute Fusion: Rethinking Feature Extraction in Image-Based Iterative Method Selection for Solving Sparse Linear Systems
di: Zhang, Kaiqi, et al.
Pubblicazione: (2025)
di: Zhang, Kaiqi, et al.
Pubblicazione: (2025)
Multi-Agent Image Restoration
di: Jiang, Xu, et al.
Pubblicazione: (2025)
di: Jiang, Xu, et al.
Pubblicazione: (2025)
ImageEdit-R1: Boosting Multi-Agent Image Editing via Reinforcement Learning
di: Zhao, Yiran, et al.
Pubblicazione: (2026)
di: Zhao, Yiran, et al.
Pubblicazione: (2026)
Improving Generalization of Medical Image Registration Foundation Model
di: Hu, Jing, et al.
Pubblicazione: (2025)
di: Hu, Jing, et al.
Pubblicazione: (2025)
Incentivizing Tool-augmented Thinking with Images for Medical Image Analysis
di: Jiang, Yankai, et al.
Pubblicazione: (2025)
di: Jiang, Yankai, et al.
Pubblicazione: (2025)
GEBench: Benchmarking Image Generation Models as GUI Environments
di: Li, Haodong, et al.
Pubblicazione: (2026)
di: Li, Haodong, et al.
Pubblicazione: (2026)
REVEAL: Reasoning-Enhanced Forensic Evidence Analysis for Explainable AI-Generated Image Detection
di: Cao, Huangsen, et al.
Pubblicazione: (2025)
di: Cao, Huangsen, et al.
Pubblicazione: (2025)
Cross Modality Image Translation In Medical Imaging Using Generative Frameworks
di: Romoli, Giulia, et al.
Pubblicazione: (2026)
di: Romoli, Giulia, et al.
Pubblicazione: (2026)
IA-T2I: Internet-Augmented Text-to-Image Generation
di: Li, Chuanhao, et al.
Pubblicazione: (2025)
di: Li, Chuanhao, et al.
Pubblicazione: (2025)
Poetry in Pixels: Prompt Tuning for Poem Image Generation via Diffusion Models
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
di: Jamil, Sofia, et al.
Pubblicazione: (2025)
Culture-inspired Multi-modal Color Palette Generation and Colorization: A Chinese Youth Subculture Case
di: Li, Yufan, et al.
Pubblicazione: (2021)
di: Li, Yufan, et al.
Pubblicazione: (2021)
Diagnostic Benchmark and Iterative Inpainting for Layout-Guided Image Generation
di: Cho, Jaemin, et al.
Pubblicazione: (2023)
di: Cho, Jaemin, et al.
Pubblicazione: (2023)
RL-I2IT: Image-to-Image Translation with Deep Reinforcement Learning
di: Hu, Jing, et al.
Pubblicazione: (2023)
di: Hu, Jing, et al.
Pubblicazione: (2023)
VLM-Guided Iterative Refinement for Surgical Image Segmentation with Foundation Models
di: Lou, Ange, et al.
Pubblicazione: (2026)
di: Lou, Ange, et al.
Pubblicazione: (2026)
A High-Quality Dataset and Reliable Evaluation for Interleaved Image-Text Generation
di: Feng, Yukang, et al.
Pubblicazione: (2025)
di: Feng, Yukang, et al.
Pubblicazione: (2025)
MS-UMamba: An Improved Vision Mamba Unet for Fetal Abdominal Medical Image Segmentation
di: Xu, Caixu, et al.
Pubblicazione: (2025)
di: Xu, Caixu, et al.
Pubblicazione: (2025)
Spatial-Aware Latent Initialization for Controllable Image Generation
di: Sun, Wenqiang, et al.
Pubblicazione: (2024)
di: Sun, Wenqiang, et al.
Pubblicazione: (2024)
VModA: An Effective Framework for Adaptive NSFW Image Moderation
di: Bao, Han, et al.
Pubblicazione: (2025)
di: Bao, Han, et al.
Pubblicazione: (2025)
Heterogeneous Generative Knowledge Distillation with Masked Image Modeling
di: Wang, Ziming, et al.
Pubblicazione: (2023)
di: Wang, Ziming, et al.
Pubblicazione: (2023)
WithAnyone: Towards Controllable and ID Consistent Image Generation
di: Xu, Hengyuan, et al.
Pubblicazione: (2025)
di: Xu, Hengyuan, et al.
Pubblicazione: (2025)
Compress3D: a Compressed Latent Space for 3D Generation from a Single Image
di: Zhang, Bowen, et al.
Pubblicazione: (2024)
di: Zhang, Bowen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Pixels, Patterns, but No Poetry: To See The World like Humans
di: Gao, Hongcheng, et al.
Pubblicazione: (2025) -
Single Image Iterative Subject-driven Generation and Editing
di: Shpitzer, Yair, et al.
Pubblicazione: (2025) -
Seeing the Poem: Image-Semantic Detection of AI-Generated Modern Chinese Poetry with MLLMs
di: Wang, Shanshan, et al.
Pubblicazione: (2026) -
Iterative Refinement Improves Compositional Image Generation
di: Jaiswal, Shantanu, et al.
Pubblicazione: (2026) -
Iterative Adversarial Attack on Image-guided Story Ending Generation
di: Wang, Youze, et al.
Pubblicazione: (2023)