You Only Submit One Image to Find the Most Suitable Generative Model
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Zhi, Guo, Lan-Zhe, Song, Peng-Xiao, Li, Yu-Feng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CGI: Identifying Conditional Generative Models with Example Images
by: Zhou, Zhi, et al.
Published: (2025)
by: Zhou, Zhi, et al.
Published: (2025)
You Only Need One Color Space: An Efficient Network for Low-light Image Enhancement
by: Yan, Qingsen, et al.
Published: (2024)
by: Yan, Qingsen, et al.
Published: (2024)
OMGSR: You Only Need One Mid-timestep Guidance for Real-World Image Super-Resolution
by: Wu, Zhiqiang, et al.
Published: (2025)
by: Wu, Zhiqiang, et al.
Published: (2025)
You Only Need One Stage: Novel-View Synthesis From A Single Blind Face Image
by: Wang, Taoyue, et al.
Published: (2026)
by: Wang, Taoyue, et al.
Published: (2026)
Image is All You Need to Empower Large-scale Diffusion Models for In-Domain Generation
by: Cao, Pu, et al.
Published: (2023)
by: Cao, Pu, et al.
Published: (2023)
Learning from One and Only One Shot
by: Yu, Haizi, et al.
Published: (2022)
by: Yu, Haizi, et al.
Published: (2022)
LAST: Leveraging Tools as Hints to Enhance Spatial Reasoning for Multimodal Large Language Models
by: Tian, Shi-Yu, et al.
Published: (2026)
by: Tian, Shi-Yu, et al.
Published: (2026)
VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning
by: Jia, Zi-Yi, et al.
Published: (2026)
by: Jia, Zi-Yi, et al.
Published: (2026)
You Only Judge Once: Multi-response Reward Modeling in a Single Forward Pass
by: Yang, Yinuo, et al.
Published: (2026)
by: Yang, Yinuo, et al.
Published: (2026)
Intelligent Communication Mixture-of-Experts Boosted-Medical Image Segmentation Foundation Model
by: Zhang, Xinwei, et al.
Published: (2025)
by: Zhang, Xinwei, et al.
Published: (2025)
LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token
by: Zhang, Shaolei, et al.
Published: (2025)
by: Zhang, Shaolei, et al.
Published: (2025)
Oasis: One Image is All You Need for Multimodal Instruction Data Synthesis
by: Zhang, Letian, et al.
Published: (2025)
by: Zhang, Letian, et al.
Published: (2025)
Off-The-Shelf Image-to-Image Models Are All You Need To Defeat Image Protection Schemes
by: Pleimling, Xavier, et al.
Published: (2026)
by: Pleimling, Xavier, et al.
Published: (2026)
From Noise to Nuance: Advances in Deep Generative Image Models
by: Peng, Benji, et al.
Published: (2024)
by: Peng, Benji, et al.
Published: (2024)
SUSTechGAN: Image Generation for Object Detection in Adverse Conditions of Autonomous Driving
by: Lan, Gongjin, et al.
Published: (2024)
by: Lan, Gongjin, et al.
Published: (2024)
StyleMamba : State Space Model for Efficient Text-driven Image Style Transfer
by: Wang, Zijia, et al.
Published: (2024)
by: Wang, Zijia, et al.
Published: (2024)
RewardFlow: Generate Images by Optimizing What You Reward
by: Susladkar, Onkar, et al.
Published: (2026)
by: Susladkar, Onkar, et al.
Published: (2026)
Wired Perspectives: Multi-View Wire Art Embraces Generative AI
by: Qu, Zhiyu, et al.
Published: (2023)
by: Qu, Zhiyu, et al.
Published: (2023)
Three Creates All: You Only Sample 3 Steps
by: Cai, Yuren, et al.
Published: (2026)
by: Cai, Yuren, et al.
Published: (2026)
Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
by: Zhou, Chunting, et al.
Published: (2024)
by: Zhou, Chunting, et al.
Published: (2024)
Memory augment is All You Need for image restoration
by: Zhang, Xiao Feng, et al.
Published: (2023)
by: Zhang, Xiao Feng, et al.
Published: (2023)
Is One GPU Enough? Pushing Image Generation at Higher-Resolutions with Foundation Models
by: Tragakis, Athanasios, et al.
Published: (2024)
by: Tragakis, Athanasios, et al.
Published: (2024)
Bayesian Exploration of Pre-trained Models for Low-shot Image Classification
by: Miao, Yibo, et al.
Published: (2024)
by: Miao, Yibo, et al.
Published: (2024)
How Long Can Unified Multimodal Models Generate Images Reliably? Taming Long-Horizon Interleaved Image Generation via Context Curation
by: Chen, Haoyu, et al.
Published: (2026)
by: Chen, Haoyu, et al.
Published: (2026)
Not All Attention Heads Are What You Need: Refining CLIP's Image Representation with Attention Ablation
by: Lin, Feng, et al.
Published: (2025)
by: Lin, Feng, et al.
Published: (2025)
CTS: A Consistency-Based Medical Image Segmentation Model
by: Zhang, Kejia, et al.
Published: (2024)
by: Zhang, Kejia, et al.
Published: (2024)
Img2Loc: Revisiting Image Geolocalization using Multi-modality Foundation Models and Image-based Retrieval-Augmented Generation
by: Zhou, Zhongliang, et al.
Published: (2024)
by: Zhou, Zhongliang, et al.
Published: (2024)
One-Forcing: Towards Stable One-Step Autoregressive Video Generation
by: Feng, Jiaqi, et al.
Published: (2026)
by: Feng, Jiaqi, et al.
Published: (2026)
DPDEdit: Detail-Preserved Diffusion Models for Multimodal Fashion Image Editing
by: Wang, Xiaolong, et al.
Published: (2024)
by: Wang, Xiaolong, et al.
Published: (2024)
SemVideo: Reconstructs What You Watch from Brain Activity via Hierarchical Semantic Guidance
by: Yang, Minghan, et al.
Published: (2026)
by: Yang, Minghan, et al.
Published: (2026)
Mamba-CAD: State Space Model For 3D Computer-Aided Design Generative Modeling
by: Li, Xueyang, et al.
Published: (2026)
by: Li, Xueyang, et al.
Published: (2026)
FGM-HD: Boosting Generation Diversity of Fractal Generative Models through Hausdorff Dimension Induction
by: Zhang, Haowei, et al.
Published: (2025)
by: Zhang, Haowei, et al.
Published: (2025)
Beyond Matching to Tiles: Bridging Unaligned Aerial and Satellite Views for Vision-Only UAV Navigation
by: Liu, Kejia, et al.
Published: (2026)
by: Liu, Kejia, et al.
Published: (2026)
Unified Thinker: A General Reasoning Modular Core for Image Generation
by: Zhou, Sashuai, et al.
Published: (2026)
by: Zhou, Sashuai, et al.
Published: (2026)
SpatialReward: Verifiable Spatial Reward Modeling for Fine-Grained Spatial Consistency in Text-to-Image Generation
by: Zhou, Sashuai, et al.
Published: (2026)
by: Zhou, Sashuai, et al.
Published: (2026)
One Head Eight Arms: Block Matrix based Low Rank Adaptation for CLIP-based Few-Shot Learning
by: Zhou, Chunpeng, et al.
Published: (2025)
by: Zhou, Chunpeng, et al.
Published: (2025)
EfficientFSL: Enhancing Few-Shot Classification via Query-Only Tuning in Vision Transformers
by: Liao, Wenwen, et al.
Published: (2026)
by: Liao, Wenwen, et al.
Published: (2026)
Progressive Image Restoration via Text-Conditioned Video Generation
by: Kang, Peng, et al.
Published: (2025)
by: Kang, Peng, et al.
Published: (2025)
VAP-Diffusion: Enriching Descriptions with MLLMs for Enhanced Medical Image Generation
by: Huang, Peng, et al.
Published: (2025)
by: Huang, Peng, et al.
Published: (2025)
HARIVO: Harnessing Text-to-Image Models for Video Generation
by: Kwon, Mingi, et al.
Published: (2024)
by: Kwon, Mingi, et al.
Published: (2024)
Similar Items
-
CGI: Identifying Conditional Generative Models with Example Images
by: Zhou, Zhi, et al.
Published: (2025) -
You Only Need One Color Space: An Efficient Network for Low-light Image Enhancement
by: Yan, Qingsen, et al.
Published: (2024) -
OMGSR: You Only Need One Mid-timestep Guidance for Real-World Image Super-Resolution
by: Wu, Zhiqiang, et al.
Published: (2025) -
You Only Need One Stage: Novel-View Synthesis From A Single Blind Face Image
by: Wang, Taoyue, et al.
Published: (2026) -
Image is All You Need to Empower Large-scale Diffusion Models for In-Domain Generation
by: Cao, Pu, et al.
Published: (2023)