Relations, Negations, and Numbers: Looking for Logic in Generative Text-to-Image Models
Fuente:
arXiv
Saved in:
| Main Authors: | Conwell, Colin, Tawiah-Quashie, Rupert, Ullman, Tomer |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A New Hybrid Intelligent Approach for Multimodal Detection of Suspected Disinformation on TikTok
by: Guerrero-Sosa, Jared D. T., et al.
Published: (2025)
by: Guerrero-Sosa, Jared D. T., et al.
Published: (2025)
NePTune: A Neuro-Pythonic Framework for Tunable Compositional Reasoning on Vision-Language
by: Kamali, Danial, et al.
Published: (2025)
by: Kamali, Danial, et al.
Published: (2025)
Vector-Symbolic Architecture for Event-Based Optical Flow
by: You, Hongzhi, et al.
Published: (2024)
by: You, Hongzhi, et al.
Published: (2024)
Algebraic Representations for Faster Predictions in Convolutional Neural Networks
by: Joyce, Johnny, et al.
Published: (2024)
by: Joyce, Johnny, et al.
Published: (2024)
The Illusion-Illusion: Vision Language Models See Illusions Where There are None
by: Ullman, Tomer
Published: (2024)
by: Ullman, Tomer
Published: (2024)
We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
by: Qiao, Runqi, et al.
Published: (2024)
by: Qiao, Runqi, et al.
Published: (2024)
Hyperion -- A fast, versatile symbolic Gaussian Belief Propagation framework for Continuous-Time SLAM
by: Hug, David, et al.
Published: (2024)
by: Hug, David, et al.
Published: (2024)
A Neurosymbolic Approach to Adaptive Feature Extraction in SLAM
by: Chandio, Yasra, et al.
Published: (2024)
by: Chandio, Yasra, et al.
Published: (2024)
FactorHD: A Hyperdimensional Computing Model for Multi-Object Multi-Class Representation and Factorization
by: Zhou, Yifei, et al.
Published: (2025)
by: Zhou, Yifei, et al.
Published: (2025)
Large Language Models are Interpretable Learners
by: Wang, Ruochen, et al.
Published: (2024)
by: Wang, Ruochen, et al.
Published: (2024)
Self-Attention Based Semantic Decomposition in Vector Symbolic Architectures
by: Yeung, Calvin, et al.
Published: (2024)
by: Yeung, Calvin, et al.
Published: (2024)
Reasoning over the Behaviour of Objects in Video-Clips for Adverb-Type Recognition
by: Seshadri, Amrit Diggavi, et al.
Published: (2023)
by: Seshadri, Amrit Diggavi, et al.
Published: (2023)
Dynamic Training-Free Fusion of Subject and Style LoRAs
by: Cao, Qinglong, et al.
Published: (2026)
by: Cao, Qinglong, et al.
Published: (2026)
Distilling Formal Logic into Neural Spaces: A Kernel Alignment Approach for Signal Temporal Logic
by: Candussio, Sara, et al.
Published: (2026)
by: Candussio, Sara, et al.
Published: (2026)
Structural-Ambiguity-Aware Translation from Natural Language to Signal Temporal Logic
by: Fushimi, Kosei, et al.
Published: (2026)
by: Fushimi, Kosei, et al.
Published: (2026)
Bio-inspired AI: Integrating Biological Complexity into Artificial Intelligence
by: Dehghani, Nima, et al.
Published: (2024)
by: Dehghani, Nima, et al.
Published: (2024)
Rule-Based Spatial Mixture-of-Experts U-Net for Explainable Edge Detection
by: Dogga, Bharadwaj, et al.
Published: (2026)
by: Dogga, Bharadwaj, et al.
Published: (2026)
Amortized Equation Discovery in Hybrid Dynamical Systems
by: Liu, Yongtuo, et al.
Published: (2024)
by: Liu, Yongtuo, et al.
Published: (2024)
Hierarchical NeuroSymbolic Approach for Comprehensive and Explainable Action Quality Assessment
by: Okamoto, Lauren, et al.
Published: (2024)
by: Okamoto, Lauren, et al.
Published: (2024)
ParSEL: Parameterized Shape Editing with Language
by: Ganeshan, Aditya, et al.
Published: (2024)
by: Ganeshan, Aditya, et al.
Published: (2024)
Speaking in Words, Thinking in Logic: A Dual-Process Framework in QA Systems
by: Bui, Tuan, et al.
Published: (2025)
by: Bui, Tuan, et al.
Published: (2025)
Numerically Computing Galois Groups of Minimal Problems
by: Duff, Timothy
Published: (2025)
by: Duff, Timothy
Published: (2025)
Using Multimodal Deep Neural Networks to Disentangle Language from Visual Aesthetics
by: Conwell, Colin, et al.
Published: (2024)
by: Conwell, Colin, et al.
Published: (2024)
LogicOCR: Do Your Large Multimodal Models Excel at Logical Reasoning on Text-Rich Images?
by: Ye, Maoyuan, et al.
Published: (2025)
by: Ye, Maoyuan, et al.
Published: (2025)
Evaluating Task-Oriented Dialogue Consistency through Constraint Satisfaction
by: Labruna, Tiziano, et al.
Published: (2024)
by: Labruna, Tiziano, et al.
Published: (2024)
A Neuro-Symbolic Approach to Monitoring Salt Content in Food
by: Tayal, Anuja, et al.
Published: (2024)
by: Tayal, Anuja, et al.
Published: (2024)
Verified Language Processing with Hybrid Explainability: A Technical Report
by: Fox, Oliver Robert, et al.
Published: (2025)
by: Fox, Oliver Robert, et al.
Published: (2025)
GOFAI meets Generative AI: Development of Expert Systems by means of Large Language Models
by: Garrido-Merchán, Eduardo C., et al.
Published: (2025)
by: Garrido-Merchán, Eduardo C., et al.
Published: (2025)
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through Logic
by: Zhao, Xufeng, et al.
Published: (2023)
by: Zhao, Xufeng, et al.
Published: (2023)
Image2Text2Image: A Novel Framework for Label-Free Evaluation of Image-to-Text Generation with Text-to-Image Diffusion Models
by: Huang, Jia-Hong, et al.
Published: (2024)
by: Huang, Jia-Hong, et al.
Published: (2024)
DreamArtist++: Controllable One-Shot Text-to-Image Generation via Positive-Negative Adapter
by: Dong, Ziyi, et al.
Published: (2022)
by: Dong, Ziyi, et al.
Published: (2022)
A Comparative Study of Neurosymbolic AI Approaches to Interpretable Logical Reasoning
by: Chen, Michael K.
Published: (2025)
by: Chen, Michael K.
Published: (2025)
Visual Set Program Synthesizer
by: Cheng, Zehua, et al.
Published: (2026)
by: Cheng, Zehua, et al.
Published: (2026)
Large Language Models as Mirrors of Societal Moral Standards
by: Papadopoulou, Evi, et al.
Published: (2024)
by: Papadopoulou, Evi, et al.
Published: (2024)
Looking Beyond Text: Reducing Language bias in Large Vision-Language Models via Multimodal Dual-Attention and Soft-Image Guidance
by: Zhao, Haozhe, et al.
Published: (2024)
by: Zhao, Haozhe, et al.
Published: (2024)
Parameter-Efficient CT Reconstruction via Deep Graph Laplacian Regularization
by: Radhakrishnan, Veera Varuni, et al.
Published: (2026)
by: Radhakrishnan, Veera Varuni, et al.
Published: (2026)
Quantum Knowledge Graph: Modeling Context-Dependent Triplet Validity
by: Wang, Yao, et al.
Published: (2026)
by: Wang, Yao, et al.
Published: (2026)
Precision or Recall? An Analysis of Image Captions for Training Text-to-Image Generation Model
by: Cheng, Sheng, et al.
Published: (2024)
by: Cheng, Sheng, et al.
Published: (2024)
Chain of Time: In-Context Physical Simulation with Image Generation Models
by: Wang, YingQiao, et al.
Published: (2025)
by: Wang, YingQiao, et al.
Published: (2025)
Cognitive LLMs: Towards Integrating Cognitive Architectures and Large Language Models for Manufacturing Decision-making
by: Wu, Siyu, et al.
Published: (2024)
by: Wu, Siyu, et al.
Published: (2024)
Similar Items
-
A New Hybrid Intelligent Approach for Multimodal Detection of Suspected Disinformation on TikTok
by: Guerrero-Sosa, Jared D. T., et al.
Published: (2025) -
NePTune: A Neuro-Pythonic Framework for Tunable Compositional Reasoning on Vision-Language
by: Kamali, Danial, et al.
Published: (2025) -
Vector-Symbolic Architecture for Event-Based Optical Flow
by: You, Hongzhi, et al.
Published: (2024) -
Algebraic Representations for Faster Predictions in Convolutional Neural Networks
by: Joyce, Johnny, et al.
Published: (2024) -
The Illusion-Illusion: Vision Language Models See Illusions Where There are None
by: Ullman, Tomer
Published: (2024)