Visual Word Sense Disambiguation with CLIP through Dual-Channel Text Prompting and Image Augmentations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhattacharya, Shamik, Perkins, Daniel, Dogan, Yaren, Konjeti, Vineeth, Srinivasan, Sudarshan, Begoli, Edmon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Prompt Engineering and the Effectiveness of Large Language Models in Enhancing Human Productivity
von: Anam, Rizal Khoirul
Veröffentlicht: (2025)
von: Anam, Rizal Khoirul
Veröffentlicht: (2025)
Disambiguation of Emotion Annotations by Contextualizing Events in Plausible Narratives
von: Schäfer, Johannes, et al.
Veröffentlicht: (2025)
von: Schäfer, Johannes, et al.
Veröffentlicht: (2025)
Extracting and Validating Explanatory Word Archipelagoes using Dual Entropy
von: Ohsawa, Yukio
Veröffentlicht: (2020)
von: Ohsawa, Yukio
Veröffentlicht: (2020)
RHealthTwin: Towards Responsible and Multimodal Digital Twins for Personalized Well-being
von: Ferdousi, Rahatara, et al.
Veröffentlicht: (2025)
von: Ferdousi, Rahatara, et al.
Veröffentlicht: (2025)
The Compression Paradox in LLM Inference: Provider-Dependent Energy Effects of Prompt Compression
von: Johnson, Warren
Veröffentlicht: (2026)
von: Johnson, Warren
Veröffentlicht: (2026)
GPT-2 as a Compression Preprocessor: Improving Gzip for Structured Text Domains
von: Ojha, Anurag Kumar
Veröffentlicht: (2025)
von: Ojha, Anurag Kumar
Veröffentlicht: (2025)
BayesRAG: Probabilistic Mutual Evidence Corroboration for Multimodal Retrieval-Augmented Generation
von: Li, Xuan, et al.
Veröffentlicht: (2026)
von: Li, Xuan, et al.
Veröffentlicht: (2026)
Jina CLIP: Your CLIP Model Is Also Your Text Retriever
von: Koukounas, Andreas, et al.
Veröffentlicht: (2024)
von: Koukounas, Andreas, et al.
Veröffentlicht: (2024)
Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs
von: Balter, Samuel G., et al.
Veröffentlicht: (2026)
von: Balter, Samuel G., et al.
Veröffentlicht: (2026)
Monetizing Currency Pair Sentiments through LLM Explainability
von: Limonad, Lior, et al.
Veröffentlicht: (2024)
von: Limonad, Lior, et al.
Veröffentlicht: (2024)
LLMs as Deceptive Agents: How Role-Based Prompting Induces Semantic Ambiguity in Puzzle Tasks
von: Yoo, Seunghyun
Veröffentlicht: (2025)
von: Yoo, Seunghyun
Veröffentlicht: (2025)
Bridging the Language Gap: Enhancing Multilingual Prompt-Based Code Generation in LLMs via Zero-Shot Cross-Lingual Transfer
von: Li, Mingda, et al.
Veröffentlicht: (2024)
von: Li, Mingda, et al.
Veröffentlicht: (2024)
Multicultural Spyfall: Assessing LLMs through Dynamic Multilingual Social Deduction Game
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2026)
von: Wibowo, Haryo Akbarianto, et al.
Veröffentlicht: (2026)
mHC-SSM: Manifold-Constrained Hyper-Connections for State Space Language Models with Stream-Specialized Adapters
von: Mutlu, Abdulvahap, et al.
Veröffentlicht: (2026)
von: Mutlu, Abdulvahap, et al.
Veröffentlicht: (2026)
Relating Word Embedding Gender Biases to Gender Gaps: A Cross-Cultural Analysis
von: Friedman, Scott, et al.
Veröffentlicht: (2026)
von: Friedman, Scott, et al.
Veröffentlicht: (2026)
LLM-Assisted Crisis Management: Building Advanced LLM Platforms for Effective Emergency Response and Public Collaboration
von: Otal, Hakan T., et al.
Veröffentlicht: (2024)
von: Otal, Hakan T., et al.
Veröffentlicht: (2024)
From MTEB to MTOB: Retrieval-Augmented Classification for Descriptive Grammars
von: Kornilov, Albert, et al.
Veröffentlicht: (2024)
von: Kornilov, Albert, et al.
Veröffentlicht: (2024)
Forging GEMs: Advancing Greek NLP through Quality-Based Corpus Curation
von: Apostolopoulou, Alexandra, et al.
Veröffentlicht: (2025)
von: Apostolopoulou, Alexandra, et al.
Veröffentlicht: (2025)
Does CLIP perceive art the same way we do?
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
Does It Make Sense to Explain a Black Box With Another Black Box?
von: Delaunay, Julien, et al.
Veröffentlicht: (2024)
von: Delaunay, Julien, et al.
Veröffentlicht: (2024)
From Scarcity to Efficiency: Investigating the Effects of Data Augmentation on African Machine Translation
von: Oduwole, Mardiyyah, et al.
Veröffentlicht: (2025)
von: Oduwole, Mardiyyah, et al.
Veröffentlicht: (2025)
Generic Embedding-Based Lexicons for Transparent and Reproducible Text Scoring
von: Moez, Catherine
Veröffentlicht: (2024)
von: Moez, Catherine
Veröffentlicht: (2024)
AI Assistants for Spaceflight Procedures: Combining Generative Pre-Trained Transformer and Retrieval-Augmented Generation on Knowledge Graphs With Augmented Reality Cues
von: Bensch, Oliver, et al.
Veröffentlicht: (2024)
von: Bensch, Oliver, et al.
Veröffentlicht: (2024)
Improving Recursive Transformers with Mixture of LoRAs
von: Nouriborji, Mohammadmahdi, et al.
Veröffentlicht: (2025)
von: Nouriborji, Mohammadmahdi, et al.
Veröffentlicht: (2025)
Tatarstan Toponyms: A Bilingual Dataset and Hybrid RAG System for Geospatial Question Answering
von: Arabov, Mullosharaf K.
Veröffentlicht: (2026)
von: Arabov, Mullosharaf K.
Veröffentlicht: (2026)
Survey of Swarm Intelligence Approaches to Search Documents Based On Semantic Similarity
von: Muniyappa, Chandrashekar, et al.
Veröffentlicht: (2025)
von: Muniyappa, Chandrashekar, et al.
Veröffentlicht: (2025)
LangMARL: Natural Language Multi-Agent Reinforcement Learning
von: Yao, Huaiyuan, et al.
Veröffentlicht: (2026)
von: Yao, Huaiyuan, et al.
Veröffentlicht: (2026)
Prompting Encoder Models for Zero-Shot Classification: A Cross-Domain Study in Italian
von: Auriemma, Serena, et al.
Veröffentlicht: (2024)
von: Auriemma, Serena, et al.
Veröffentlicht: (2024)
The Meta-Prompting Protocol: Orchestrating LLMs via Adversarial Feedback Loops
von: Fu, Fanzhe
Veröffentlicht: (2025)
von: Fu, Fanzhe
Veröffentlicht: (2025)
PairCFR: Enhancing Model Training on Paired Counterfactually Augmented Data through Contrastive Learning
von: Qiu, Xiaoqi, et al.
Veröffentlicht: (2024)
von: Qiu, Xiaoqi, et al.
Veröffentlicht: (2024)
Compression Method Matters: Benchmark-Dependent Output Dynamics in LLM Prompt Compression
von: Johnson, Warren
Veröffentlicht: (2026)
von: Johnson, Warren
Veröffentlicht: (2026)
Prompt Compression in Production Task Orchestration: A Pre-Registered Randomized Trial
von: Johnson, Warren, et al.
Veröffentlicht: (2026)
von: Johnson, Warren, et al.
Veröffentlicht: (2026)
Evaluating LLM Prompts for Data Augmentation in Multi-label Classification of Ecological Texts
von: Glazkova, Anna, et al.
Veröffentlicht: (2024)
von: Glazkova, Anna, et al.
Veröffentlicht: (2024)
Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
ReFACT: Updating Text-to-Image Models by Editing the Text Encoder
von: Arad, Dana, et al.
Veröffentlicht: (2023)
von: Arad, Dana, et al.
Veröffentlicht: (2023)
Profiling German Text Simplification with Interpretable Model-Fingerprints
von: Klöser, Lars, et al.
Veröffentlicht: (2026)
von: Klöser, Lars, et al.
Veröffentlicht: (2026)
When Retrieval Succeeds and Fails: Rethinking Retrieval-Augmented Generation for LLMs
von: Wang, Yongjie, et al.
Veröffentlicht: (2025)
von: Wang, Yongjie, et al.
Veröffentlicht: (2025)
A Survey of Text Watermarking in the Era of Large Language Models
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
von: Liu, Aiwei, et al.
Veröffentlicht: (2023)
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2023)
von: Muñoz-Ortiz, Alberto, et al.
Veröffentlicht: (2023)
Evaluating AI Grading on Real-World Handwritten College Mathematics: A Large-Scale Study Toward a Benchmark
von: Yu, Zhiqi, et al.
Veröffentlicht: (2026)
von: Yu, Zhiqi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Prompt Engineering and the Effectiveness of Large Language Models in Enhancing Human Productivity
von: Anam, Rizal Khoirul
Veröffentlicht: (2025) -
Disambiguation of Emotion Annotations by Contextualizing Events in Plausible Narratives
von: Schäfer, Johannes, et al.
Veröffentlicht: (2025) -
Extracting and Validating Explanatory Word Archipelagoes using Dual Entropy
von: Ohsawa, Yukio
Veröffentlicht: (2020) -
RHealthTwin: Towards Responsible and Multimodal Digital Twins for Personalized Well-being
von: Ferdousi, Rahatara, et al.
Veröffentlicht: (2025) -
The Compression Paradox in LLM Inference: Provider-Dependent Energy Effects of Prompt Compression
von: Johnson, Warren
Veröffentlicht: (2026)