Improving Explicit Spatial Relationships in Text-to-Image Generation through an Automatically Derived Dataset
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Salaberria, Ander, Azkune, Gorka, de Lacalle, Oier Lopez, Soroa, Aitor, Agirre, Eneko, Keller, Frank |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Grounding Spatial Relations in Text-Only Language Models
von: Azkune, Gorka, et al.
Veröffentlicht: (2024)
von: Azkune, Gorka, et al.
Veröffentlicht: (2024)
BiVLC: Extending Vision-Language Compositionality Evaluation with Text-to-Image Retrieval
von: Miranda, Imanol, et al.
Veröffentlicht: (2024)
von: Miranda, Imanol, et al.
Veröffentlicht: (2024)
Adding simple structure at inference improves Vision-Language Compositionality
von: Miranda, Imanol, et al.
Veröffentlicht: (2025)
von: Miranda, Imanol, et al.
Veröffentlicht: (2025)
Revisiting Compositionality in Dual-Encoder Vision-Language Models: The Role of Inference
von: Miranda, Imanol, et al.
Veröffentlicht: (2026)
von: Miranda, Imanol, et al.
Veröffentlicht: (2026)
Vision-Language Models Struggle to Align Entities across Modalities
von: Alonso, Iñigo, et al.
Veröffentlicht: (2025)
von: Alonso, Iñigo, et al.
Veröffentlicht: (2025)
BertaQA: How Much Do Language Models Know About Local Culture?
von: Etxaniz, Julen, et al.
Veröffentlicht: (2024)
von: Etxaniz, Julen, et al.
Veröffentlicht: (2024)
Multimodal Large Language Models for Low-Resource Languages: A Case Study for Basque
von: Arana, Lukas, et al.
Veröffentlicht: (2025)
von: Arana, Lukas, et al.
Veröffentlicht: (2025)
Do not be greedy, Think Twice: Sampling and Selection for Document-level Information Extraction
von: Zubillaga, Mikel, et al.
Veröffentlicht: (2026)
von: Zubillaga, Mikel, et al.
Veröffentlicht: (2026)
Improving the Efficiency of Visually Augmented Language Models
von: Ontalvilla, Paula, et al.
Veröffentlicht: (2024)
von: Ontalvilla, Paula, et al.
Veröffentlicht: (2024)
Event Extraction in Basque: Typologically motivated Cross-Lingual Transfer-Learning Analysis
von: Zubillaga, Mikel, et al.
Veröffentlicht: (2024)
von: Zubillaga, Mikel, et al.
Veröffentlicht: (2024)
Deriving MOND from f(T) teleparallel gravity: a selected interpolation function and acceleration scale
von: Ander, Azkune
Veröffentlicht: (2026)
von: Ander, Azkune
Veröffentlicht: (2026)
Conditioning LLMs to Generate Code-Switched Text
von: Heredia, Maite, et al.
Veröffentlicht: (2025)
von: Heredia, Maite, et al.
Veröffentlicht: (2025)
Automatic Logical Forms improve fidelity in Table-to-Text generation
von: Alonso, Iñigo, et al.
Veröffentlicht: (2023)
von: Alonso, Iñigo, et al.
Veröffentlicht: (2023)
Reasoning over Object Descriptions Improves Coreference Resolution in Task-Based Dialogue Systems
von: Ijurco, Oier, et al.
Veröffentlicht: (2026)
von: Ijurco, Oier, et al.
Veröffentlicht: (2026)
Multimodal LLMs Do Not Compose Skills Optimally Across Modalities
von: Ontalvilla, Paula, et al.
Veröffentlicht: (2025)
von: Ontalvilla, Paula, et al.
Veröffentlicht: (2025)
When to Retrieve: Teaching LLMs to Utilize Information Retrieval Effectively
von: Labruna, Tiziano, et al.
Veröffentlicht: (2024)
von: Labruna, Tiziano, et al.
Veröffentlicht: (2024)
GoLLIE: Annotation Guidelines improve Zero-Shot Information-Extraction
von: Sainz, Oscar, et al.
Veröffentlicht: (2023)
von: Sainz, Oscar, et al.
Veröffentlicht: (2023)
Automatic Essay Scoring and Feedback Generation in Basque Language Learning
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
TABLET: A Large-Scale Dataset for Robust Visual Table Understanding
von: Alonso, Iñigo, et al.
Veröffentlicht: (2025)
von: Alonso, Iñigo, et al.
Veröffentlicht: (2025)
A LLM-Based Ranking Method for the Evaluation of Automatic Counter-Narrative Generation
von: Zubiaga, Irune, et al.
Veröffentlicht: (2024)
von: Zubiaga, Irune, et al.
Veröffentlicht: (2024)
KNOW: Developing large-scale multilingual technologies for language understanding
von: Eneko Agirre
Veröffentlicht: (2009)
von: Eneko Agirre
Veröffentlicht: (2009)
Exploring feature set combinations for WSD
von: Eneko Agirre
Veröffentlicht: (2006)
von: Eneko Agirre
Veröffentlicht: (2006)
Lexical semantics, Basque and Spanish in QTLeap: Quality Translation by Deep Language Engineering Approaches
von: Eneko Agirre
Veröffentlicht: (2015)
von: Eneko Agirre
Veröffentlicht: (2015)
MCR for CLIR
von: Eneko Agirre
Veröffentlicht: (2007)
von: Eneko Agirre
Veröffentlicht: (2007)
KNOW2: Language understanding technologies for multilingual domain-oriented information access
von: Eneko Agirre
Veröffentlicht: (2010)
von: Eneko Agirre
Veröffentlicht: (2010)
PixT3: Pixel-based Table-To-Text Generation
von: Alonso, Iñigo, et al.
Veröffentlicht: (2023)
von: Alonso, Iñigo, et al.
Veröffentlicht: (2023)
Latxa: An Open Language Model and Evaluation Suite for Basque
von: Etxaniz, Julen, et al.
Veröffentlicht: (2024)
von: Etxaniz, Julen, et al.
Veröffentlicht: (2024)
Una arquitectura de integración de recursos léxicos de naturaleza heterogénea. Una aportación desde la perspectiva de la integración de datos.
von: Aitor Soroa Etxabe
Veröffentlicht: (2005)
von: Aitor Soroa Etxabe
Veröffentlicht: (2005)
Euskarazko lehen C1 ebaluatzaile automatikoa
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
von: Azurmendi, Ekhi, et al.
Veröffentlicht: (2025)
Time-Distributed Backdoor Attacks on Federated Spiking Learning
von: Abad, Gorka, et al.
Veröffentlicht: (2024)
von: Abad, Gorka, et al.
Veröffentlicht: (2024)
Test-time Controllable Image Generation by Explicit Spatial Constraint Enforcement
von: Zhang, Z., et al.
Veröffentlicht: (2025)
von: Zhang, Z., et al.
Veröffentlicht: (2025)
Sneaky Spikes: Uncovering Stealthy Backdoor Attacks in Spiking Neural Networks with Neuromorphic Data
von: Abad, Gorka, et al.
Veröffentlicht: (2023)
von: Abad, Gorka, et al.
Veröffentlicht: (2023)
Improving Text-to-Image Consistency via Automatic Prompt Optimization
von: Mañas, Oscar, et al.
Veröffentlicht: (2024)
von: Mañas, Oscar, et al.
Veröffentlicht: (2024)
Getting it Right: Improving Spatial Consistency in Text-to-Image Models
von: Chatterjee, Agneet, et al.
Veröffentlicht: (2024)
von: Chatterjee, Agneet, et al.
Veröffentlicht: (2024)
Bringing The Consistency Gap: Explicit Structured Memory for Interleaved Image-Text Generation
von: Lin, Zeteng, et al.
Veröffentlicht: (2025)
von: Lin, Zeteng, et al.
Veröffentlicht: (2025)
Type-R: Automatically Retouching Typos for Text-to-Image Generation
von: Shimoda, Wataru, et al.
Veröffentlicht: (2024)
von: Shimoda, Wataru, et al.
Veröffentlicht: (2024)
Synergistic Dual Spatial-aware Generation of Image-to-Text and Text-to-Image
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
von: Zhao, Yu, et al.
Veröffentlicht: (2024)
Evaluating the Generation of Spatial Relations in Text and Image Generative Models
von: Sim, Shang Hong, et al.
Veröffentlicht: (2024)
von: Sim, Shang Hong, et al.
Veröffentlicht: (2024)
Continual Learning for Image Captioning through Improved Image-Text Alignment
von: Taetz, Bertram, et al.
Veröffentlicht: (2025)
von: Taetz, Bertram, et al.
Veröffentlicht: (2025)
EuskañolDS: A Naturally Sourced Corpus for Basque-Spanish Code-Switching
von: Heredia, Maite, et al.
Veröffentlicht: (2025)
von: Heredia, Maite, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Grounding Spatial Relations in Text-Only Language Models
von: Azkune, Gorka, et al.
Veröffentlicht: (2024) -
BiVLC: Extending Vision-Language Compositionality Evaluation with Text-to-Image Retrieval
von: Miranda, Imanol, et al.
Veröffentlicht: (2024) -
Adding simple structure at inference improves Vision-Language Compositionality
von: Miranda, Imanol, et al.
Veröffentlicht: (2025) -
Revisiting Compositionality in Dual-Encoder Vision-Language Models: The Role of Inference
von: Miranda, Imanol, et al.
Veröffentlicht: (2026) -
Vision-Language Models Struggle to Align Entities across Modalities
von: Alonso, Iñigo, et al.
Veröffentlicht: (2025)