ANCHOR: LLM-driven Subject Conditioning for Text-to-Image Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Ramakrishnan, Aashish Anantha, Huang, Sharon X., Lee, Dongwon |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RONA: Pragmatically Diverse Image Captioning with Coherence Relations
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
CORDIAL: Can Multimodal Large Language Models Effectively Understand Coherence Relationships?
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
IRONIC: Coherence-Aware Reasoning Chains for Multi-Modal Sarcasm Detection
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025)
ANNA: Abstractive Text-to-Image Synthesis with Filtered News Captions
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2023)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2023)
Generative Active Testing: Efficient LLM Evaluation via Proxy Task Adaptation
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2026)
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2026)
Diffusion Lens: Interpreting Text Encoders in Text-to-Image Pipelines
by: Toker, Michael, et al.
Published: (2024)
by: Toker, Michael, et al.
Published: (2024)
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
by: Banthia, Saumya, et al.
Published: (2020)
by: Banthia, Saumya, et al.
Published: (2020)
State of the Practice for Medical Imaging Software
by: Smith, W. Spencer, et al.
Published: (2024)
by: Smith, W. Spencer, et al.
Published: (2024)
LLMs and the Human Condition
by: Wallis, Peter
Published: (2024)
by: Wallis, Peter
Published: (2024)
Adversarial Feature Alignment: Balancing Robustness and Accuracy in Deep Learning via Adversarial Training
by: Park, Leo Hyun, et al.
Published: (2024)
by: Park, Leo Hyun, et al.
Published: (2024)
STRICT: Stress Test of Rendering Images Containing Text
by: Zhang, Tianyu, et al.
Published: (2025)
by: Zhang, Tianyu, et al.
Published: (2025)
Quaternion Convolutional Neural Networks: Current Advances and Future Directions
by: Altamirano-Gomez, Gerardo, et al.
Published: (2023)
by: Altamirano-Gomez, Gerardo, et al.
Published: (2023)
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
by: Patel, Hitesh Laxmichand, et al.
Published: (2025)
by: Patel, Hitesh Laxmichand, et al.
Published: (2025)
ALISON: Fast and Effective Stylometric Authorship Obfuscation
by: Xing, Eric, et al.
Published: (2024)
by: Xing, Eric, et al.
Published: (2024)
A simple and effective approach for body part recognition on CT scans based on projection estimation
by: Hrzic, Franko, et al.
Published: (2025)
by: Hrzic, Franko, et al.
Published: (2025)
ViFiCon: Vision and Wireless Association Via Self-Supervised Contrastive Learning
by: Meegan, Nicholas, et al.
Published: (2022)
by: Meegan, Nicholas, et al.
Published: (2022)
MemeCraft: Contextual and Stance-Driven Multimodal Meme Generation
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
Is Our Chatbot Telling Lies? Assessing Correctness of an LLM-based Dutch Support Chatbot
by: Lassche, Herman, et al.
Published: (2024)
by: Lassche, Herman, et al.
Published: (2024)
Breaking the HISCO Barrier: Automatic Occupational Standardization with OccCANINE
by: Dahl, Christian Møller, et al.
Published: (2024)
by: Dahl, Christian Møller, et al.
Published: (2024)
Advances and Frontiers of LLM-based Issue Resolution in Software Engineering: A Comprehensive Survey
by: Li, Caihua, et al.
Published: (2026)
by: Li, Caihua, et al.
Published: (2026)
ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration
by: Deng, Minghang, et al.
Published: (2025)
by: Deng, Minghang, et al.
Published: (2025)
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
by: Du, Bangde, et al.
Published: (2025)
by: Du, Bangde, et al.
Published: (2025)
Syntactic Blind Spots: How Misalignment Leads to LLMs Mathematical Errors
by: Williamson, Dane, et al.
Published: (2025)
by: Williamson, Dane, et al.
Published: (2025)
Semantic Decomposition and Selective Context Filtering -- Text Processing Techniques for Context-Aware NLP-Based Systems
by: Villardar, Karl John
Published: (2025)
by: Villardar, Karl John
Published: (2025)
Neural Machine Translation for Malayalam Paraphrase Generation
by: Varghese, Christeena, et al.
Published: (2024)
by: Varghese, Christeena, et al.
Published: (2024)
Predictive Simultaneous Interpretation: Harnessing Large Language Models for Democratizing Real-Time Multilingual Communication
by: Iida, Kurando, et al.
Published: (2024)
by: Iida, Kurando, et al.
Published: (2024)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
by: Ovcharov, Volodymyr
Published: (2026)
by: Ovcharov, Volodymyr
Published: (2026)
NRR-Phi: Text-to-State Mapping for Ambiguity Preservation in LLM Inference
by: Saito, Kei
Published: (2026)
by: Saito, Kei
Published: (2026)
Seeing The Words: Evaluating AI-generated Biblical Art
by: Makimei, Hidde, et al.
Published: (2025)
by: Makimei, Hidde, et al.
Published: (2025)
Holistic Audit Dataset Generation for LLM Unlearning via Knowledge Graph Traversal and Redundancy Removal
by: Jiang, Weipeng, et al.
Published: (2025)
by: Jiang, Weipeng, et al.
Published: (2025)
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
by: Khanna, Danush, et al.
Published: (2025)
by: Khanna, Danush, et al.
Published: (2025)
Predicting Known Vulnerabilities from Attack Descriptions Using Sentence Transformers
by: Othman, Refat
Published: (2026)
by: Othman, Refat
Published: (2026)
Grandes modelos de lenguaje: de la predicción de palabras a la comprensión?
by: Gómez-Rodríguez, Carlos
Published: (2025)
by: Gómez-Rodríguez, Carlos
Published: (2025)
ActAlign: Zero-Shot Fine-Grained Video Classification via Language-Guided Sequence Alignment
by: Aghdam, Amir, et al.
Published: (2025)
by: Aghdam, Amir, et al.
Published: (2025)
A Survey of Text and Speech Resources for Hausa and Fongbe: Availability, Quality, and Gaps for NLP Development
by: Adjovi, Mahounan Pericles, et al.
Published: (2026)
by: Adjovi, Mahounan Pericles, et al.
Published: (2026)
Semantic Commit: Helping Users Update Intent Specifications for AI Memory at Scale
by: Vaithilingam, Priyan, et al.
Published: (2025)
by: Vaithilingam, Priyan, et al.
Published: (2025)
On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance
by: Casanova, Etienne, et al.
Published: (2026)
by: Casanova, Etienne, et al.
Published: (2026)
Model-Driven Legacy System Modernization at Scale
by: Böhm, Tobias, et al.
Published: (2026)
by: Böhm, Tobias, et al.
Published: (2026)
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
by: Chen, Renmiao, et al.
Published: (2025)
by: Chen, Renmiao, et al.
Published: (2025)
Towards Conditioning Clinical Text Generation for User Control
by: Koraş, Osman Alperen, et al.
Published: (2025)
by: Koraş, Osman Alperen, et al.
Published: (2025)
Similar Items
-
RONA: Pragmatically Diverse Image Captioning with Coherence Relations
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025) -
CORDIAL: Can Multimodal Large Language Models Effectively Understand Coherence Relationships?
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025) -
IRONIC: Coherence-Aware Reasoning Chains for Multi-Modal Sarcasm Detection
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2025) -
ANNA: Abstractive Text-to-Image Synthesis with Filtered News Captions
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2023) -
Generative Active Testing: Efficient LLM Evaluation via Proxy Task Adaptation
by: Ramakrishnan, Aashish Anantha, et al.
Published: (2026)