"I See What You Did There": Can Large Vision-Language Models Understand Multimodal Puns?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Naen, Sheng, Jiayi, Li, Changjiang, Zhou, Chunyi, Li, Yuyuan, Du, Tianyu, Wang, Jun, Fu, Zhihui, Li, Jinbao, Ji, Shouling |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bridging the Copyright Gap: Do Large Vision-Language Models Recognize and Respect Copyrighted Content?
von: Xu, Naen, et al.
Veröffentlicht: (2025)
von: Xu, Naen, et al.
Veröffentlicht: (2025)
Compiling Activation Steering into Weights via Null-Space Constraints for Stealthy Backdoors
von: Yin, Rui, et al.
Veröffentlicht: (2026)
von: Yin, Rui, et al.
Veröffentlicht: (2026)
When Agents "Misremember" Collectively: Exploring the Mandela Effect in LLM-based Multi-Agent Systems
von: Xu, Naen, et al.
Veröffentlicht: (2026)
von: Xu, Naen, et al.
Veröffentlicht: (2026)
VideoEraser: Concept Erasure in Text-to-Video Diffusion Models
von: Xu, Naen, et al.
Veröffentlicht: (2025)
von: Xu, Naen, et al.
Veröffentlicht: (2025)
ACIArena: Toward Unified Evaluation for Agent Cascading Injection
von: An, Hengyu, et al.
Veröffentlicht: (2026)
von: An, Hengyu, et al.
Veröffentlicht: (2026)
FraudShield: Knowledge Graph Empowered Defense for LLMs against Fraud Attacks
von: Xu, Naen, et al.
Veröffentlicht: (2026)
von: Xu, Naen, et al.
Veröffentlicht: (2026)
Automatic Red Teaming LLM-based Agents with Model Context Protocol Tools
von: He, Ping, et al.
Veröffentlicht: (2025)
von: He, Ping, et al.
Veröffentlicht: (2025)
The Eminence in Shadow: Exploiting Feature Boundary Ambiguity for Robust Backdoor Attacks
von: Feng, Zhou, et al.
Veröffentlicht: (2025)
von: Feng, Zhou, et al.
Veröffentlicht: (2025)
Vision Language Models See What You Want but not What You See
von: Gao, Qingying, et al.
Veröffentlicht: (2024)
von: Gao, Qingying, et al.
Veröffentlicht: (2024)
CopyrightMeter: Revisiting Copyright Protection in Text-to-image Models
von: Xu, Naen, et al.
Veröffentlicht: (2024)
von: Xu, Naen, et al.
Veröffentlicht: (2024)
"Hey, Did You See This?"
von: Nelson, Cathy Jo
Veröffentlicht: (2011)
von: Nelson, Cathy Jo
Veröffentlicht: (2011)
But Did You See the Gorilla?
von: Lynne C. Messer, et al.
Veröffentlicht: (2025)
von: Lynne C. Messer, et al.
Veröffentlicht: (2025)
IPIGuard: A Novel Tool Dependency Graph-Based Defense Against Indirect Prompt Injection in LLM Agents
von: An, Hengyu, et al.
Veröffentlicht: (2025)
von: An, Hengyu, et al.
Veröffentlicht: (2025)
CAMH: Advancing Model Hijacking Attack in Machine Learning
von: He, Xing, et al.
Veröffentlicht: (2024)
von: He, Xing, et al.
Veröffentlicht: (2024)
Enkidu: Universal Frequential Perturbation for Real-Time Audio Privacy Protection against Voice Deepfakes
von: Feng, Zhou, et al.
Veröffentlicht: (2025)
von: Feng, Zhou, et al.
Veröffentlicht: (2025)
Rethinking the Vulnerabilities of Face Recognition Systems:From a Practical Perspective
von: Chen, Jiahao, et al.
Veröffentlicht: (2024)
von: Chen, Jiahao, et al.
Veröffentlicht: (2024)
"No Matter What You Do": Purifying GNN Models via Backdoor Unlearning
von: Zhang, Jiale, et al.
Veröffentlicht: (2024)
von: Zhang, Jiale, et al.
Veröffentlicht: (2024)
See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding
von: Sun, Boyuan, et al.
Veröffentlicht: (2026)
von: Sun, Boyuan, et al.
Veröffentlicht: (2026)
Enhancing Adversarial Transferability with Adversarial Weight Tuning
von: Chen, Jiahao, et al.
Veröffentlicht: (2024)
von: Chen, Jiahao, et al.
Veröffentlicht: (2024)
Pun Unintended: LLMs and the Illusion of Humor Understanding
von: Zangari, Alessandro, et al.
Veröffentlicht: (2025)
von: Zangari, Alessandro, et al.
Veröffentlicht: (2025)
Attributing and Exploiting Safety Vectors through Global Optimization in Large Language Models
von: Chu, Fengheng, et al.
Veröffentlicht: (2026)
von: Chu, Fengheng, et al.
Veröffentlicht: (2026)
LoRAShield: Data-Free Editing Alignment for Secure Personalized LoRA Sharing
von: Chen, Jiahao, et al.
Veröffentlicht: (2025)
von: Chen, Jiahao, et al.
Veröffentlicht: (2025)
"A good pun is its own reword": Can Large Language Models Understand Puns?
von: Xu, Zhijun, et al.
Veröffentlicht: (2024)
von: Xu, Zhijun, et al.
Veröffentlicht: (2024)
Andrew Pun
Veröffentlicht: (2024)
Veröffentlicht: (2024)
Andrew Pun
Veröffentlicht: (2024)
Veröffentlicht: (2024)
Profiling for Pennies: Unveiling the Privacy Iceberg of LLM Agents
von: Chen, Jiahao, et al.
Veröffentlicht: (2026)
von: Chen, Jiahao, et al.
Veröffentlicht: (2026)
What Did You Learn at the Museum Today?
von: Jairo Luis Romero‐Acosta, et al.
Veröffentlicht: (2025)
von: Jairo Luis Romero‐Acosta, et al.
Veröffentlicht: (2025)
UNIDOOR: A Universal Framework for Action-Level Backdoor Attacks in Deep Reinforcement Learning
von: Ma, Oubo, et al.
Veröffentlicht: (2025)
von: Ma, Oubo, et al.
Veröffentlicht: (2025)
Teacher: Can You See What I'm Saying? A Research Experience with Deaf Learners
von: Olga Lucía Ávila Caica
Veröffentlicht: (2011)
von: Olga Lucía Ávila Caica
Veröffentlicht: (2011)
The Struggle You Can’t See
von: Lierman, Ash
Veröffentlicht: (2024)
von: Lierman, Ash
Veröffentlicht: (2024)
On the Security Risks of ML-based Malware Detection Systems: A Survey
von: He, Ping, et al.
Veröffentlicht: (2025)
von: He, Ping, et al.
Veröffentlicht: (2025)
Can Retrieval Heads See Images? Multimodal Retrieval Heads in Long-Context Vision-Language Models
von: Li, Aaron Branson Cigres, et al.
Veröffentlicht: (2026)
von: Li, Aaron Branson Cigres, et al.
Veröffentlicht: (2026)
DP-GENG : Differentially Private Dataset Distillation Guided by DP-Generated Data
von: Shi, Shuo, et al.
Veröffentlicht: (2025)
von: Shi, Shuo, et al.
Veröffentlicht: (2025)
Do You See What I See?: Observed Race and the Ascription of American Identity
von: Raul S. Casarez
Veröffentlicht: (2025)
von: Raul S. Casarez
Veröffentlicht: (2025)
What You See is (Usually) What You Get: Multimodal Prototype Networks that Abstain from Expensive Modalities
von: Bahng, Muchang, et al.
Veröffentlicht: (2025)
von: Bahng, Muchang, et al.
Veröffentlicht: (2025)
Demo: TOSense -- What Did You Just Agree to?
von: Chen, Xinzhang, et al.
Veröffentlicht: (2025)
von: Chen, Xinzhang, et al.
Veröffentlicht: (2025)
What Did You Write about the War, Daddy?
von: Welch, Elizabeth H.
Veröffentlicht: (1972)
von: Welch, Elizabeth H.
Veröffentlicht: (1972)
Poison in the Well: Feature Embedding Disruption in Backdoor Attacks
von: Feng, Zhou, et al.
Veröffentlicht: (2025)
von: Feng, Zhou, et al.
Veröffentlicht: (2025)
"What Did It Actually Do?": Understanding Risk Awareness and Traceability for Computer-Use Agents
von: Peng, Zifan, et al.
Veröffentlicht: (2026)
von: Peng, Zifan, et al.
Veröffentlicht: (2026)
Seeing with You: Perception-Reasoning Coevolution for Multimodal Reasoning
von: Miao, Ziqi, et al.
Veröffentlicht: (2026)
von: Miao, Ziqi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Bridging the Copyright Gap: Do Large Vision-Language Models Recognize and Respect Copyrighted Content?
von: Xu, Naen, et al.
Veröffentlicht: (2025) -
Compiling Activation Steering into Weights via Null-Space Constraints for Stealthy Backdoors
von: Yin, Rui, et al.
Veröffentlicht: (2026) -
When Agents "Misremember" Collectively: Exploring the Mandela Effect in LLM-based Multi-Agent Systems
von: Xu, Naen, et al.
Veröffentlicht: (2026) -
VideoEraser: Concept Erasure in Text-to-Video Diffusion Models
von: Xu, Naen, et al.
Veröffentlicht: (2025) -
ACIArena: Toward Unified Evaluation for Agent Cascading Injection
von: An, Hengyu, et al.
Veröffentlicht: (2026)