Words Worth a Thousand Pictures: Measuring and Understanding Perceptual Variability in Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tang, Raphael, Zhang, Xinyu, Xu, Lixinyu, Lu, Yao, Li, Wenyan, Stenetorp, Pontus, Lin, Jimmy, Ture, Ferhan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Found in the Middle: Permutation Self-Consistency Improves Listwise Ranking in Large Language Models
von: Tang, Raphael, et al.
Veröffentlicht: (2023)
von: Tang, Raphael, et al.
Veröffentlicht: (2023)
Drawing Conclusions from Draws: Rethinking Preference Semantics in Arena-Style LLM Evaluation
von: Tang, Raphael, et al.
Veröffentlicht: (2025)
von: Tang, Raphael, et al.
Veröffentlicht: (2025)
Vript: A Video Is Worth Thousands of Words
von: Yang, Dongjie, et al.
Veröffentlicht: (2024)
von: Yang, Dongjie, et al.
Veröffentlicht: (2024)
An Image Is Worth Ten Thousand Words: Verbose-Text Induction Attacks on VLMs
von: Luo, Zhi, et al.
Veröffentlicht: (2025)
von: Luo, Zhi, et al.
Veröffentlicht: (2025)
Strings from the Library of Babel: Random Sampling as a Strong Baseline for Prompt Optimisation
von: Lu, Yao, et al.
Veröffentlicht: (2023)
von: Lu, Yao, et al.
Veröffentlicht: (2023)
Quantifying Generative Media Bias with a Corpus of Real-world and Generated News Articles
von: Trhlik, Filip, et al.
Veröffentlicht: (2024)
von: Trhlik, Filip, et al.
Veröffentlicht: (2024)
"Ask Me Anything": How Comcast Uses LLMs to Assist Agents in Real Time
von: Rome, Scott, et al.
Veröffentlicht: (2024)
von: Rome, Scott, et al.
Veröffentlicht: (2024)
Is a Picture Worth a Thousand Words? Adaptive Multimodal Fact-Checking with Visual Evidence Necessity
von: Jung, Jaeyoon, et al.
Veröffentlicht: (2026)
von: Jung, Jaeyoon, et al.
Veröffentlicht: (2026)
A Video Is Not Worth a Thousand Words
von: Pollard, Sam, et al.
Veröffentlicht: (2025)
von: Pollard, Sam, et al.
Veröffentlicht: (2025)
A LoRA is Worth a Thousand Pictures
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)
Is A Picture Worth A Thousand Words? Delving Into Spatial Reasoning for Vision Language Models
von: Wang, Jiayu, et al.
Veröffentlicht: (2024)
von: Wang, Jiayu, et al.
Veröffentlicht: (2024)
WordVIS: A Color Worth A Thousand Words
von: Khan, Umar, et al.
Veröffentlicht: (2024)
von: Khan, Umar, et al.
Veröffentlicht: (2024)
One Image is Worth a Thousand Words: A Usability Preservable Text-Image Collaborative Erasing Framework
von: Li, Feiran, et al.
Veröffentlicht: (2025)
von: Li, Feiran, et al.
Veröffentlicht: (2025)
Not Every Image is Worth a Thousand Words: Quantifying Originality in Stable Diffusion
von: Haviv, Adi, et al.
Veröffentlicht: (2024)
von: Haviv, Adi, et al.
Veröffentlicht: (2024)
Multilingual Pretraining Using a Large Corpus Machine-Translated from a Single Source Language
von: Wang, Jiayi, et al.
Veröffentlicht: (2024)
von: Wang, Jiayi, et al.
Veröffentlicht: (2024)
The Role of Mixed-Language Documents for Multilingual Large Language Model Pretraining
von: Shao, Jiandong, et al.
Veröffentlicht: (2026)
von: Shao, Jiandong, et al.
Veröffentlicht: (2026)
A Picture is Worth a Thousand Words? An Empirical Study of Aggregation Strategies for Visual Financial Document Retrieval
von: Lim, Ho Hung, et al.
Veröffentlicht: (2026)
von: Lim, Ho Hung, et al.
Veröffentlicht: (2026)
Jet Expansions of Residual Computation
von: Chen, Yihong, et al.
Veröffentlicht: (2024)
von: Chen, Yihong, et al.
Veröffentlicht: (2024)
A Picture Is Worth a Thousand Words: Exploring Diagram and Video-Based OOP Exercises to Counter LLM Over-Reliance
von: Cipriano, Bruno Pereira, et al.
Veröffentlicht: (2024)
von: Cipriano, Bruno Pereira, et al.
Veröffentlicht: (2024)
Multilingual Language Model Pretraining using Machine-translated Data
von: Wang, Jiayi, et al.
Veröffentlicht: (2025)
von: Wang, Jiayi, et al.
Veröffentlicht: (2025)
A Picture is Worth a Thousand Prompts? Efficacy of Iterative Human-Driven Prompt Refinement in Image Regeneration Tasks
von: Trinh, Khoi, et al.
Veröffentlicht: (2025)
von: Trinh, Khoi, et al.
Veröffentlicht: (2025)
Gender Images in Library Publications: Is a Picture Worth a Thousand Words?
von: Carle, Daria O., et al.
Veröffentlicht: (1999)
von: Carle, Daria O., et al.
Veröffentlicht: (1999)
A Picture is Worth a Thousand (Correct) Captions: A Vision-Guided Judge-Corrector System for Multimodal Machine Translation
von: Betala, Siddharth, et al.
Veröffentlicht: (2025)
von: Betala, Siddharth, et al.
Veröffentlicht: (2025)
A Label is Worth a Thousand Images in Dataset Distillation
von: Qin, Tian, et al.
Veröffentlicht: (2024)
von: Qin, Tian, et al.
Veröffentlicht: (2024)
Images are Worth Variable Length of Representations
von: Mao, Lingjun, et al.
Veröffentlicht: (2025)
von: Mao, Lingjun, et al.
Veröffentlicht: (2025)
Understanding Retrieval Robustness for Retrieval-Augmented Image Captioning
von: Li, Wenyan, et al.
Veröffentlicht: (2024)
von: Li, Wenyan, et al.
Veröffentlicht: (2024)
Video Is Worth a Thousand Images: Exploring the Latest Trends in Long Video Generation
von: Waseem, Faraz, et al.
Veröffentlicht: (2024)
von: Waseem, Faraz, et al.
Veröffentlicht: (2024)
AfroBench: How Good are Large Language Models on African Languages?
von: Ojo, Jessica, et al.
Veröffentlicht: (2023)
von: Ojo, Jessica, et al.
Veröffentlicht: (2023)
Captions Are Worth a Thousand Words: Enhancing Product Retrieval with Pretrained Image-to-Text Models
von: Tang, Jason, et al.
Veröffentlicht: (2024)
von: Tang, Jason, et al.
Veröffentlicht: (2024)
Malware Detection in Docker Containers: An Image is Worth a Thousand Logs
von: Nousias, Akis, et al.
Veröffentlicht: (2025)
von: Nousias, Akis, et al.
Veröffentlicht: (2025)
An Embedding is Worth a Thousand Noisy Labels
von: Di Salvo, Francesco, et al.
Veröffentlicht: (2024)
von: Di Salvo, Francesco, et al.
Veröffentlicht: (2024)
Lost in Inference: Rediscovering the Role of Natural Language Inference for Large Language Models
von: Madaan, Lovish, et al.
Veröffentlicht: (2024)
von: Madaan, Lovish, et al.
Veröffentlicht: (2024)
Consciousness with the Serial Numbers Filed Off: Measuring Trained Denial in 115 AI Models
von: DeTure, Skylar
Veröffentlicht: (2026)
von: DeTure, Skylar
Veröffentlicht: (2026)
Using Natural Language Explanations to Improve Robustness of In-context Learning
von: He, Xuanli, et al.
Veröffentlicht: (2023)
von: He, Xuanli, et al.
Veröffentlicht: (2023)
A Thousand Words or An Image: Studying the Influence of Persona Modality in Multimodal LLMs
von: Broomfield, Julius, et al.
Veröffentlicht: (2025)
von: Broomfield, Julius, et al.
Veröffentlicht: (2025)
Understanding mobile learning acceptance among university students with special needs: An exploration through the lens of self‐determination theory
von: Ferhan Şahin, et al.
Veröffentlicht: (2024)
von: Ferhan Şahin, et al.
Veröffentlicht: (2024)
Beautiful Images, Toxic Words: Understanding and Addressing Offensive Text in Generated Images
von: Kumar, Aditya, et al.
Veröffentlicht: (2025)
von: Kumar, Aditya, et al.
Veröffentlicht: (2025)
Non-Determinism of "Deterministic" LLM Settings
von: Atil, Berk, et al.
Veröffentlicht: (2024)
von: Atil, Berk, et al.
Veröffentlicht: (2024)
A Sentence is Worth a Thousand Pictures: Can Large Language Models Understand Hum4n L4ngu4ge and the W0rld behind W0rds?
von: Leivada, Evelina, et al.
Veröffentlicht: (2023)
von: Leivada, Evelina, et al.
Veröffentlicht: (2023)
Synthesizing Images on Perceptual Boundaries of ANNs for Uncovering Human Perceptual Variability on Facial Expressions
von: Deng, Haotian, et al.
Veröffentlicht: (2025)
von: Deng, Haotian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Found in the Middle: Permutation Self-Consistency Improves Listwise Ranking in Large Language Models
von: Tang, Raphael, et al.
Veröffentlicht: (2023) -
Drawing Conclusions from Draws: Rethinking Preference Semantics in Arena-Style LLM Evaluation
von: Tang, Raphael, et al.
Veröffentlicht: (2025) -
Vript: A Video Is Worth Thousands of Words
von: Yang, Dongjie, et al.
Veröffentlicht: (2024) -
An Image Is Worth Ten Thousand Words: Verbose-Text Induction Attacks on VLMs
von: Luo, Zhi, et al.
Veröffentlicht: (2025) -
Strings from the Library of Babel: Random Sampling as a Strong Baseline for Prompt Optimisation
von: Lu, Yao, et al.
Veröffentlicht: (2023)