How and where does CLIP process negation?
Fuente:
arXiv
Guardado en:
| Autores principales: | Quantmeyer, Vincent, Mosteiro, Pablo, Gatt, Albert |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Grounded Misunderstandings in Asymmetric Dialogue: A Perspectivist Annotation Scheme for MapTask
por: Li, Nan, et al.
Publicado: (2025)
por: Li, Nan, et al.
Publicado: (2025)
Evaluating LLM-Generated Versus Human-Authored Responses in Role-Play Dialogues
por: Lu, Dongxu, et al.
Publicado: (2025)
por: Lu, Dongxu, et al.
Publicado: (2025)
Contrast Is All You Need
por: Kilic, Burak, et al.
Publicado: (2023)
por: Kilic, Burak, et al.
Publicado: (2023)
How does a Multilingual LM Handle Multiple Languages?
por: Kakarla, Santhosh, et al.
Publicado: (2025)
por: Kakarla, Santhosh, et al.
Publicado: (2025)
How does Misinformation Affect Large Language Model Behaviors and Preferences?
por: Peng, Miao, et al.
Publicado: (2025)
por: Peng, Miao, et al.
Publicado: (2025)
Evaluation of Multilingual Image Captioning: How far can we get with CLIP models?
por: Gomes, Gonçalo, et al.
Publicado: (2025)
por: Gomes, Gonçalo, et al.
Publicado: (2025)
How does fine-tuning improve sensorimotor representations in large language models?
por: Wu, Minghua, et al.
Publicado: (2026)
por: Wu, Minghua, et al.
Publicado: (2026)
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding
por: Chen, Xi, et al.
Publicado: (2025)
por: Chen, Xi, et al.
Publicado: (2025)
Functional Subspace, where language models can use vector algebra to solve problems
por: Lee, Jung H., et al.
Publicado: (2026)
por: Lee, Jung H., et al.
Publicado: (2026)
The Idola Tribus of AI: Large Language Models tend to perceive order where none exists
por: Ishikawa, Shin-nosuke, et al.
Publicado: (2025)
por: Ishikawa, Shin-nosuke, et al.
Publicado: (2025)
Where does In-context Translation Happen in Large Language Models
por: Sia, Suzanna, et al.
Publicado: (2024)
por: Sia, Suzanna, et al.
Publicado: (2024)
LatteCLIP: Unsupervised CLIP Fine-Tuning via LMM-Synthetic Texts
por: Cao, Anh-Quan, et al.
Publicado: (2024)
por: Cao, Anh-Quan, et al.
Publicado: (2024)
DisCoCLIP: A Distributional Compositional Tensor Network Encoder for Vision-Language Understanding
por: Lo, Kin Ian, et al.
Publicado: (2025)
por: Lo, Kin Ian, et al.
Publicado: (2025)
InterCLIP-MEP: Interactive CLIP and Memory-Enhanced Predictor for Multi-modal Sarcasm Detection
por: Chen, Junjie, et al.
Publicado: (2024)
por: Chen, Junjie, et al.
Publicado: (2024)
UrbanCLIP: Learning Text-enhanced Urban Region Profiling with Contrastive Language-Image Pretraining from the Web
por: Yan, Yibo, et al.
Publicado: (2023)
por: Yan, Yibo, et al.
Publicado: (2023)
Why mask diffusion does not work
por: Sun, Haocheng, et al.
Publicado: (2025)
por: Sun, Haocheng, et al.
Publicado: (2025)
The Confidence Shortcut: A Reasoning Failure Mode of Masked Diffusion Models
por: Kim, Dueun, et al.
Publicado: (2026)
por: Kim, Dueun, et al.
Publicado: (2026)
Updating CLIP to Prefer Descriptions Over Captions
por: Zur, Amir, et al.
Publicado: (2024)
por: Zur, Amir, et al.
Publicado: (2024)
Temperature-scaling surprisal estimates improve fit to human reading times -- but does it do so for the "right reasons"?
por: Liu, Tong, et al.
Publicado: (2023)
por: Liu, Tong, et al.
Publicado: (2023)
$How^{2}$: How to learn from procedural How-to questions
por: Dagan, Gautier, et al.
Publicado: (2025)
por: Dagan, Gautier, et al.
Publicado: (2025)
Clinical knowledge in LLMs does not translate to human interactions
por: Bean, Andrew M., et al.
Publicado: (2025)
por: Bean, Andrew M., et al.
Publicado: (2025)
UoR-NCL at SemEval-2025 Task 1: Using Generative LLMs and CLIP Models for Multilingual Multimodal Idiomaticity Representation
por: Markchom, Thanet, et al.
Publicado: (2025)
por: Markchom, Thanet, et al.
Publicado: (2025)
Which bird does not have wings: Negative-constrained KGQA with Schema-guided Semantic Matching and Self-directed Refinement
por: Shim, Midan, et al.
Publicado: (2026)
por: Shim, Midan, et al.
Publicado: (2026)
Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning
por: Schaffelder, Max, et al.
Publicado: (2025)
por: Schaffelder, Max, et al.
Publicado: (2025)
Which symbol grounding problem should we try to solve?
por: Müller, Vincent C.
Publicado: (2025)
por: Müller, Vincent C.
Publicado: (2025)
When a language model is optimized for reasoning, does it still show embers of autoregression? An analysis of OpenAI o1
por: McCoy, R. Thomas, et al.
Publicado: (2024)
por: McCoy, R. Thomas, et al.
Publicado: (2024)
Where does output diversity collapse in post-training?
por: Karouzos, Constantinos, et al.
Publicado: (2026)
por: Karouzos, Constantinos, et al.
Publicado: (2026)
ComCLIP: Training-Free Compositional Image and Text Matching
por: Jiang, Kenan, et al.
Publicado: (2022)
por: Jiang, Kenan, et al.
Publicado: (2022)
AllSummedUp: un framework open-source pour comparer les metriques d'evaluation de resume
por: Herserant, Tanguy, et al.
Publicado: (2025)
por: Herserant, Tanguy, et al.
Publicado: (2025)
SEval-Ex: A Statement-Level Framework for Explainable Summarization Evaluation
por: Herserant, Tanguy, et al.
Publicado: (2025)
por: Herserant, Tanguy, et al.
Publicado: (2025)
Explainable Rule Application via Structured Prompting: A Neural-Symbolic Approach
por: Sadowski, Albert, et al.
Publicado: (2025)
por: Sadowski, Albert, et al.
Publicado: (2025)
On Verifiable Legal Reasoning: A Multi-Agent Framework with Formalized Knowledge Representations
por: Sadowski, Albert, et al.
Publicado: (2025)
por: Sadowski, Albert, et al.
Publicado: (2025)
Bridging Legal Knowledge and AI: Retrieval-Augmented Generation with Vector Stores, Knowledge Graphs, and Hierarchical Non-negative Matrix Factorization
por: Barron, Ryan C., et al.
Publicado: (2025)
por: Barron, Ryan C., et al.
Publicado: (2025)
Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse
por: Liu, Ryan, et al.
Publicado: (2024)
por: Liu, Ryan, et al.
Publicado: (2024)
Explaining Caption-Image Interactions in CLIP Models with Second-Order Attributions
por: Möller, Lucas, et al.
Publicado: (2024)
por: Möller, Lucas, et al.
Publicado: (2024)
Does CLIP Bind Concepts? Probing Compositionality in Large Image Models
por: Lewis, Martha, et al.
Publicado: (2022)
por: Lewis, Martha, et al.
Publicado: (2022)
TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation
por: Lin, Haokun, et al.
Publicado: (2025)
por: Lin, Haokun, et al.
Publicado: (2025)
Natural language processing for African languages
por: Adelani, David Ifeoluwa
Publicado: (2025)
por: Adelani, David Ifeoluwa
Publicado: (2025)
How LLMs Might Think
por: Gottlieb, Joseph, et al.
Publicado: (2026)
por: Gottlieb, Joseph, et al.
Publicado: (2026)
How Persuasive is Your Context?
por: Nguyen, Tu, et al.
Publicado: (2025)
por: Nguyen, Tu, et al.
Publicado: (2025)
Ejemplares similares
-
Grounded Misunderstandings in Asymmetric Dialogue: A Perspectivist Annotation Scheme for MapTask
por: Li, Nan, et al.
Publicado: (2025) -
Evaluating LLM-Generated Versus Human-Authored Responses in Role-Play Dialogues
por: Lu, Dongxu, et al.
Publicado: (2025) -
Contrast Is All You Need
por: Kilic, Burak, et al.
Publicado: (2023) -
How does a Multilingual LM Handle Multiple Languages?
por: Kakarla, Santhosh, et al.
Publicado: (2025) -
How does Misinformation Affect Large Language Model Behaviors and Preferences?
por: Peng, Miao, et al.
Publicado: (2025)