Why Settle for One? Text-to-ImageSet Generation and Evaluation
Fuente:
arXiv
Salvato in:
| Autori principali: | Jia, Chengyou, Shen, Xin, Dang, Zhuohang, Xia, Changliang, Wu, Weijia, Zhang, Xinyu, Qian, Hangwei, Tsang, Ivor W., Luo, Minnan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ChatGen: Automatic Text-to-Image Generation From FreeStyle Chatting
di: Jia, Chengyou, et al.
Pubblicazione: (2024)
di: Jia, Chengyou, et al.
Pubblicazione: (2024)
PaCo-RL: Advancing Reinforcement Learning for Consistent Image Generation with Pairwise Reward Modeling
di: Ping, Bowen, et al.
Pubblicazione: (2025)
di: Ping, Bowen, et al.
Pubblicazione: (2025)
Multi-Modal Dataset Distillation in the Wild
di: Dang, Zhuohang, et al.
Pubblicazione: (2025)
di: Dang, Zhuohang, et al.
Pubblicazione: (2025)
$\mathrm{D}^\mathrm{3}$-Predictor: Noise-Free Deterministic Diffusion for Dense Prediction
di: Xia, Changliang, et al.
Pubblicazione: (2025)
di: Xia, Changliang, et al.
Pubblicazione: (2025)
From Ideal to Real: Unified and Data-Efficient Dense Prediction for Real-World Scenarios
di: Xia, Changliang, et al.
Pubblicazione: (2025)
di: Xia, Changliang, et al.
Pubblicazione: (2025)
Flow-Factory: A Unified Framework for Reinforcement Learning in Flow-Matching Models
di: Ping, Bowen, et al.
Pubblicazione: (2026)
di: Ping, Bowen, et al.
Pubblicazione: (2026)
AutoGPS: Automated Geometry Problem Solving via Multimodal Formalization and Deductive Reasoning
di: Ping, Bowen, et al.
Pubblicazione: (2025)
di: Ping, Bowen, et al.
Pubblicazione: (2025)
SSMG: Spatial-Semantic Map Guided Diffusion Model for Free-form Layout-to-Image Generation
di: Jia, Chengyou, et al.
Pubblicazione: (2023)
di: Jia, Chengyou, et al.
Pubblicazione: (2023)
PSDiff: Diffusion Model for Person Search with Iterative and Collaborative Refinement
di: Jia, Chengyou, et al.
Pubblicazione: (2023)
di: Jia, Chengyou, et al.
Pubblicazione: (2023)
Disentangled Representation Learning with Transmitted Information Bottleneck
di: Dang, Zhuohang, et al.
Pubblicazione: (2023)
di: Dang, Zhuohang, et al.
Pubblicazione: (2023)
AgentStore: Scalable Integration of Heterogeneous Agents As Specialized Generalist Computer Assistant
di: Jia, Chengyou, et al.
Pubblicazione: (2024)
di: Jia, Chengyou, et al.
Pubblicazione: (2024)
Exploring the Effectiveness and Interpretability of Texts in LLM-based Time Series Models
di: Sun, Zhengke, et al.
Pubblicazione: (2025)
di: Sun, Zhengke, et al.
Pubblicazione: (2025)
Disentangled Noisy Correspondence Learning
di: Dang, Zhuohang, et al.
Pubblicazione: (2024)
di: Dang, Zhuohang, et al.
Pubblicazione: (2024)
ImageSet2Text: Describing Sets of Images through Text
di: Riccio, Piera, et al.
Pubblicazione: (2025)
di: Riccio, Piera, et al.
Pubblicazione: (2025)
CoFFT: Chain of Foresight-Focus Thought for Visual Language Models
di: Zhang, Xinyu, et al.
Pubblicazione: (2025)
di: Zhang, Xinyu, et al.
Pubblicazione: (2025)
The Role of Reactive Oxygen Species in Alzheimer's Disease: From Mechanism to Biomaterials Therapy
di: Zhuohang Yu, et al.
Pubblicazione: (2024)
di: Zhuohang Yu, et al.
Pubblicazione: (2024)
The Role of Reactive Oxygen Species in Alzheimer's Disease: From Mechanism to Biomaterials Therapy (Adv. Healthcare Mater. 29/2024)
di: Zhuohang Yu, et al.
Pubblicazione: (2024)
di: Zhuohang Yu, et al.
Pubblicazione: (2024)
Collected environmental change and nitrogen removal data
di: Dong, Liang, et al.
Pubblicazione: (2025)
di: Dong, Liang, et al.
Pubblicazione: (2025)
Learning ORDER-Aware Multimodal Representations for Composite Materials Design
di: Li, Xinyao, et al.
Pubblicazione: (2026)
di: Li, Xinyao, et al.
Pubblicazione: (2026)
Uncover and Unlearn Nuisances: Agnostic Fully Test-Time Adaptation
di: Srey, Ponhvoan, et al.
Pubblicazione: (2025)
di: Srey, Ponhvoan, et al.
Pubblicazione: (2025)
Cross-Context Backdoor Attacks against Graph Prompt Learning
di: Lyu, Xiaoting, et al.
Pubblicazione: (2024)
di: Lyu, Xiaoting, et al.
Pubblicazione: (2024)
Effective Bethe Ansatz for Spin-1 Non-integrable Models
di: Wang, Zhuohang, et al.
Pubblicazione: (2026)
di: Wang, Zhuohang, et al.
Pubblicazione: (2026)
SCOUT-RAG: Scalable and Cost-Efficient Unifying Traversal for Agentic Graph-RAG over Distributed Domains
di: Li, Longkun, et al.
Pubblicazione: (2026)
di: Li, Longkun, et al.
Pubblicazione: (2026)
Hierarchical Mask-Enhanced Dual Reconstruction Network for Few-Shot Fine-Grained Image Classification
di: Luo, Ning, et al.
Pubblicazione: (2025)
di: Luo, Ning, et al.
Pubblicazione: (2025)
Second-Order Fine-Tuning without Pain for LLMs:A Hessian Informed Zeroth-Order Optimizer
di: Zhao, Yanjun, et al.
Pubblicazione: (2024)
di: Zhao, Yanjun, et al.
Pubblicazione: (2024)
Why Settle for Equity?
di: Timothy A. Carey
Pubblicazione: (2025)
di: Timothy A. Carey
Pubblicazione: (2025)
Improved Turbo Message Passing for Compressive Robust Principal Component Analysis: Algorithm Design and Asymptotic Analysis
di: He, Zhuohang, et al.
Pubblicazione: (2024)
di: He, Zhuohang, et al.
Pubblicazione: (2024)
A Generic and Efficient Python Runtime Verification System and its Large-scale Evaluation
di: Shen, Zhuohang, et al.
Pubblicazione: (2025)
di: Shen, Zhuohang, et al.
Pubblicazione: (2025)
CogBlender: Towards Continuous Cognitive Intervention in Text-to-Image Generation
di: Dang, Shengqi, et al.
Pubblicazione: (2026)
di: Dang, Shengqi, et al.
Pubblicazione: (2026)
SHAN: Object-Level Privacy Detection via Inference on Scene Heterogeneous Graph
di: Jiang, Zhuohang, et al.
Pubblicazione: (2024)
di: Jiang, Zhuohang, et al.
Pubblicazione: (2024)
Beyond Visual Appearances: Privacy-sensitive Objects Identification via Hybrid Graph Reasoning
di: Jiang, Zhuohang, et al.
Pubblicazione: (2024)
di: Jiang, Zhuohang, et al.
Pubblicazione: (2024)
Evaluating LLMs Without Oracle Feedback: Agentic Annotation Evaluation Through Unsupervised Consistency Signals
di: Chen, Cheng, et al.
Pubblicazione: (2025)
di: Chen, Cheng, et al.
Pubblicazione: (2025)
Why Does Differential Privacy with Large Epsilon Defend Against Practical Membership Inference Attacks?
di: Lowy, Andrew, et al.
Pubblicazione: (2024)
di: Lowy, Andrew, et al.
Pubblicazione: (2024)
Prompt-based Ingredient-Oriented All-in-One Image Restoration
di: Gao, Hu, et al.
Pubblicazione: (2023)
di: Gao, Hu, et al.
Pubblicazione: (2023)
IML-ViT: Benchmarking Image Manipulation Localization by Vision Transformer
di: Ma, Xiaochen, et al.
Pubblicazione: (2023)
di: Ma, Xiaochen, et al.
Pubblicazione: (2023)
Matching-Free Depth Recovery from Structured Light
di: Yu, Zhuohang, et al.
Pubblicazione: (2025)
di: Yu, Zhuohang, et al.
Pubblicazione: (2025)
FinCast: A Foundation Model for Financial Time-Series Forecasting
di: Zhu, Zhuohang, et al.
Pubblicazione: (2025)
di: Zhu, Zhuohang, et al.
Pubblicazione: (2025)
Mid-infrared edge-enhanced imaging via angle-selective nonlinear filtering
di: Wei, Zhuohang, et al.
Pubblicazione: (2026)
di: Wei, Zhuohang, et al.
Pubblicazione: (2026)
Conversion of Farmland to Apple Orchards Modifies Water–Carbon–Nitrogen Trade‐Offs in Deep Loess Deposits
di: Zhuohang Jin, et al.
Pubblicazione: (2025)
di: Zhuohang Jin, et al.
Pubblicazione: (2025)
EmotiCrafter: Text-to-Emotional-Image Generation based on Valence-Arousal Model
di: Dang, Shengqi, et al.
Pubblicazione: (2025)
di: Dang, Shengqi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
ChatGen: Automatic Text-to-Image Generation From FreeStyle Chatting
di: Jia, Chengyou, et al.
Pubblicazione: (2024) -
PaCo-RL: Advancing Reinforcement Learning for Consistent Image Generation with Pairwise Reward Modeling
di: Ping, Bowen, et al.
Pubblicazione: (2025) -
Multi-Modal Dataset Distillation in the Wild
di: Dang, Zhuohang, et al.
Pubblicazione: (2025) -
$\mathrm{D}^\mathrm{3}$-Predictor: Noise-Free Deterministic Diffusion for Dense Prediction
di: Xia, Changliang, et al.
Pubblicazione: (2025) -
From Ideal to Real: Unified and Data-Efficient Dense Prediction for Real-World Scenarios
di: Xia, Changliang, et al.
Pubblicazione: (2025)