MULAN: A Multi Layer Annotated Dataset for Controllable Text-to-Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Tudosiu, Petru-Daniel, Yang, Yongxin, Zhang, Shifeng, Chen, Fei, McDonagh, Steven, Lampouras, Gerasimos, Iacobacci, Ignacio, Parisot, Sarah |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Generating Compositional Scenes via Text-to-image RGBA Instance Generation
di: Fontanella, Alessandro, et al.
Pubblicazione: (2024)
di: Fontanella, Alessandro, et al.
Pubblicazione: (2024)
A Practical Investigation of Spatially-Controlled Image Generation with Transformers
di: Xia, Guoxuan, et al.
Pubblicazione: (2025)
di: Xia, Guoxuan, et al.
Pubblicazione: (2025)
Exploiting Mixture-of-Experts Redundancy Unlocks Multimodal Generative Abilities
di: Dutt, Raman, et al.
Pubblicazione: (2025)
di: Dutt, Raman, et al.
Pubblicazione: (2025)
HumanRankEval: Automatic Evaluation of LMs as Conversational Assistants
di: Gritta, Milan, et al.
Pubblicazione: (2024)
di: Gritta, Milan, et al.
Pubblicazione: (2024)
Code-Optimise: Self-Generated Preference Data for Correctness and Efficiency
di: Gee, Leonidas, et al.
Pubblicazione: (2024)
di: Gee, Leonidas, et al.
Pubblicazione: (2024)
Improving Object Detection via Local-global Contrastive Learning
di: Triantafyllidou, Danai, et al.
Pubblicazione: (2024)
di: Triantafyllidou, Danai, et al.
Pubblicazione: (2024)
Text-to-Code Generation with Modality-relative Pre-training
di: Christopoulou, Fenia, et al.
Pubblicazione: (2024)
di: Christopoulou, Fenia, et al.
Pubblicazione: (2024)
Chapter 4 Improving the estate
di: McDonagh, Briony
Pubblicazione: (2019)
di: McDonagh, Briony
Pubblicazione: (2019)
Chapter 3 Managing the estate
di: McDonagh, Briony
Pubblicazione: (2019)
di: McDonagh, Briony
Pubblicazione: (2019)
Chapter 2 Women, land and property
di: McDonagh, Briony
Pubblicazione: (2019)
di: McDonagh, Briony
Pubblicazione: (2019)
Elite Women and the Agricultural Landscape, 1700–1830
di: McDonagh, Briony
Pubblicazione: (2019)
di: McDonagh, Briony
Pubblicazione: (2019)
Microfinance strategies for HIV/AIDS mitigation and prevention in Sub-Saharan Africa
di: Amy McDonagh
Pubblicazione: (2001)
di: Amy McDonagh
Pubblicazione: (2001)
JoseBellido and KathyBowrey, Adventures in Childhood: Intellectual Property, Imagination and the Business of Play, Cambridge, Cambridge University Press, 2022, 250pp, hb, £85.00
di: Luke McDonagh
Pubblicazione: (2024)
di: Luke McDonagh
Pubblicazione: (2024)
Findings of the First Workshop on Simulating Conversational Intelligence in Chat
di: Graham, Yvette, et al.
Pubblicazione: (2024)
di: Graham, Yvette, et al.
Pubblicazione: (2024)
Label-Efficient Object Detection via Region Proposal Network Pre-Training
di: Dong, Nanqing, et al.
Pubblicazione: (2022)
di: Dong, Nanqing, et al.
Pubblicazione: (2022)
Benchmarking Self-Supervised Learning Methods for Accelerated MRI Reconstruction
di: Wang, Andrew, et al.
Pubblicazione: (2025)
di: Wang, Andrew, et al.
Pubblicazione: (2025)
CSEval: A Framework for Evaluating Clinical Semantics in Text-to-Image Generation
di: Cronshaw, Robert, et al.
Pubblicazione: (2026)
di: Cronshaw, Robert, et al.
Pubblicazione: (2026)
CRCE: Coreference-Retention Concept Erasure in Text-to-Image Diffusion Models
di: Xue, Yuyang, et al.
Pubblicazione: (2025)
di: Xue, Yuyang, et al.
Pubblicazione: (2025)
SceneForge: Structured World Supervision from 3D Interventions
di: Li, Jizhizi, et al.
Pubblicazione: (2026)
di: Li, Jizhizi, et al.
Pubblicazione: (2026)
Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models
di: Stogiannidis, Ilias, et al.
Pubblicazione: (2025)
di: Stogiannidis, Ilias, et al.
Pubblicazione: (2025)
TopoAlign: A Framework for Aligning Code to Math via Topological Decomposition
di: Li, Yupei, et al.
Pubblicazione: (2025)
di: Li, Yupei, et al.
Pubblicazione: (2025)
DReSD: Dense Retrieval for Speculative Decoding
di: Gritta, Milan, et al.
Pubblicazione: (2025)
di: Gritta, Milan, et al.
Pubblicazione: (2025)
Concept-based Adversarial Attack: a Probabilistic Perspective
di: Zhang, Andi, et al.
Pubblicazione: (2025)
di: Zhang, Andi, et al.
Pubblicazione: (2025)
Youth‐led theatre for climate resilience and action at COP26
di: Kate Smith, et al.
Pubblicazione: (2025)
di: Kate Smith, et al.
Pubblicazione: (2025)
EofE Manifesto: A Behavioral Ethics Layer for Human-AI Interaction
di: Iacobacci, Nicoletta
Pubblicazione: (2026)
di: Iacobacci, Nicoletta
Pubblicazione: (2026)
GaussianHeadTalk: Wobble-Free 3D Talking Heads with Audio Driven Gaussian Splatting
di: Agarwal, Madhav, et al.
Pubblicazione: (2025)
di: Agarwal, Madhav, et al.
Pubblicazione: (2025)
Ambient Physics: Training Neural PDE Solvers with Partial Observations
di: Majid, Harris Abdul, et al.
Pubblicazione: (2026)
di: Majid, Harris Abdul, et al.
Pubblicazione: (2026)
Erase to Enhance: Data-Efficient Machine Unlearning in MRI Reconstruction
di: Xue, Yuyang, et al.
Pubblicazione: (2024)
di: Xue, Yuyang, et al.
Pubblicazione: (2024)
An Extended Evaluation Split for DeepSpaceYoloDataset
di: Parisot, Olivier
Pubblicazione: (2026)
di: Parisot, Olivier
Pubblicazione: (2026)
DRIFT: Decompose, Retrieve, Illustrate, then Formalize Theorems
di: Zhang, Meiru, et al.
Pubblicazione: (2025)
di: Zhang, Meiru, et al.
Pubblicazione: (2025)
Why Do Vision Language Models Struggle To Recognize Human Emotions?
di: Agarwal, Madhav, et al.
Pubblicazione: (2026)
di: Agarwal, Madhav, et al.
Pubblicazione: (2026)
Rethinking Inter-LoRA Orthogonality in Adapter Merging: Insights from Orthogonal Monte Carlo Dropout
di: Zhang, Andi, et al.
Pubblicazione: (2025)
di: Zhang, Andi, et al.
Pubblicazione: (2025)
Causal Ordering for Structure Learning from Time Series
di: Sanchez, Pedro P., et al.
Pubblicazione: (2025)
di: Sanchez, Pedro P., et al.
Pubblicazione: (2025)
SparsePO: Controlling Preference Alignment of LLMs via Sparse Token Masks
di: Christopoulou, Fenia, et al.
Pubblicazione: (2024)
di: Christopoulou, Fenia, et al.
Pubblicazione: (2024)
View-Consistent Diffusion Representations for 3D-Consistent Video Generation
di: Danier, Duolikun, et al.
Pubblicazione: (2025)
di: Danier, Duolikun, et al.
Pubblicazione: (2025)
The Impact of Due Diligence Legislation on International Trade and Business: Analysis of Potential Trade‐Offs
di: Peter Draper, et al.
Pubblicazione: (2025)
di: Peter Draper, et al.
Pubblicazione: (2025)
No time to train! Training-Free Reference-Based Instance Segmentation
di: Espinosa, Miguel, et al.
Pubblicazione: (2025)
di: Espinosa, Miguel, et al.
Pubblicazione: (2025)
There is no SAMantics! Exploring SAM as a Backbone for Visual Understanding Tasks
di: Espinosa, Miguel, et al.
Pubblicazione: (2024)
di: Espinosa, Miguel, et al.
Pubblicazione: (2024)
Conjecturing: An Overlooked Step in Formal Mathematical Reasoning
di: Sivakumar, Jasivan Alex, et al.
Pubblicazione: (2025)
di: Sivakumar, Jasivan Alex, et al.
Pubblicazione: (2025)
On the Significance of Religion for Global Diplomacy
di: McDonagh, Philip, et al.
Pubblicazione: (2020)
di: McDonagh, Philip, et al.
Pubblicazione: (2020)
Documenti analoghi
-
Generating Compositional Scenes via Text-to-image RGBA Instance Generation
di: Fontanella, Alessandro, et al.
Pubblicazione: (2024) -
A Practical Investigation of Spatially-Controlled Image Generation with Transformers
di: Xia, Guoxuan, et al.
Pubblicazione: (2025) -
Exploiting Mixture-of-Experts Redundancy Unlocks Multimodal Generative Abilities
di: Dutt, Raman, et al.
Pubblicazione: (2025) -
HumanRankEval: Automatic Evaluation of LMs as Conversational Assistants
di: Gritta, Milan, et al.
Pubblicazione: (2024) -
Code-Optimise: Self-Generated Preference Data for Correctness and Efficiency
di: Gee, Leonidas, et al.
Pubblicazione: (2024)