VLSlice: Interactive Vision-and-Language Slice Discovery
Fuente:
arXiv
Saved in:
| Main Authors: | Slyman, Eric, Kahng, Minsuk, Lee, Stefan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FairDeDup: Detecting and Mitigating Vision-Language Fairness Disparities in Semantic Dataset Deduplication
by: Slyman, Eric, et al.
Published: (2024)
by: Slyman, Eric, et al.
Published: (2024)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
by: Shahin, Nada, et al.
Published: (2025)
by: Shahin, Nada, et al.
Published: (2025)
Therapy as an NLP Task: Psychologists' Comparison of LLMs and Human Peers in CBT
by: Iftikhar, Zainab, et al.
Published: (2024)
by: Iftikhar, Zainab, et al.
Published: (2024)
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
by: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Published: (2025)
by: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Published: (2025)
Topological Structure Description for Artcode Detection Using the Shape of Orientation Histogram
by: Xu, Liming, et al.
Published: (2025)
by: Xu, Liming, et al.
Published: (2025)
ADAT: Time-Series-Aware Adaptive Transformer Architecture for Sign Language Translation
by: Shahin, Nada, et al.
Published: (2025)
by: Shahin, Nada, et al.
Published: (2025)
Growing Perspectives: Modelling Embodied Perspective Taking and Inner Narrative Development Using Large Language Models
by: Patania, Sabrina, et al.
Published: (2025)
by: Patania, Sabrina, et al.
Published: (2025)
Detecting Effects of AI-Mediated Communication on Language Complexity and Sentiment
by: Sussman, Kristen, et al.
Published: (2025)
by: Sussman, Kristen, et al.
Published: (2025)
Towards a Robust Framework for Multimodal Hate Detection: A Study on Video vs. Image-based Content
by: Koushik, Girish A., et al.
Published: (2025)
by: Koushik, Girish A., et al.
Published: (2025)
PerspAct: Enhancing LLM Situated Collaboration Skills through Perspective Taking and Active Vision
by: Patania, Sabrina, et al.
Published: (2025)
by: Patania, Sabrina, et al.
Published: (2025)
Who Sees What? Structured Thought-Action Sequences for Epistemic Reasoning in LLMs
by: Annese, Luca, et al.
Published: (2025)
by: Annese, Luca, et al.
Published: (2025)
Talking Tennis: Language Feedback from 3D Biomechanical Action Recognition
by: Dashore, Arushi, et al.
Published: (2025)
by: Dashore, Arushi, et al.
Published: (2025)
HATL: Hierarchical Adaptive-Transfer Learning Framework for Sign Language Machine Translation
by: Shahin, Nada, et al.
Published: (2026)
by: Shahin, Nada, et al.
Published: (2026)
Visual Aesthetic Benchmark: Can Frontier Models Judge Beauty?
by: Feng, Yichen, et al.
Published: (2026)
by: Feng, Yichen, et al.
Published: (2026)
The Wisdom of Agent Crowds: A Human-AI Interaction Innovation Ignition Framework
by: Yang, Senhao, et al.
Published: (2025)
by: Yang, Senhao, et al.
Published: (2025)
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
by: Patel, Hitesh Laxmichand, et al.
Published: (2025)
by: Patel, Hitesh Laxmichand, et al.
Published: (2025)
Generalizable Multiscale Segmentation of Heterogeneous Map Collections
by: Petitpierre, Remi
Published: (2026)
by: Petitpierre, Remi
Published: (2026)
Analyzing Quality, Bias, and Performance in Text-to-Image Generative Models
by: Masrourisaadat, Nila, et al.
Published: (2024)
by: Masrourisaadat, Nila, et al.
Published: (2024)
On the Limitations of Vision-Language Models in Understanding Image Transforms
by: Anis, Ahmad Mustafa, et al.
Published: (2025)
by: Anis, Ahmad Mustafa, et al.
Published: (2025)
Enhancing Cross-Modal Contextual Congruence for Crowdfunding Success using Knowledge-infused Learning
by: Padhi, Trilok, et al.
Published: (2024)
by: Padhi, Trilok, et al.
Published: (2024)
Studying Maps at Scale: A Digital Investigation of Cartography and the Evolution of Figuration
by: Petitpierre, Remi
Published: (2025)
by: Petitpierre, Remi
Published: (2025)
Revisiting [CLS] and Patch Token Interaction in Vision Transformers
by: Marouani, Alexis, et al.
Published: (2026)
by: Marouani, Alexis, et al.
Published: (2026)
Cora: Correspondence-aware image editing using few step diffusion
by: Alimohammadi, Amirhossein, et al.
Published: (2025)
by: Alimohammadi, Amirhossein, et al.
Published: (2025)
Exploiting Causality Signals in Medical Images: A Pilot Study with Empirical Results
by: Carloni, Gianluca, et al.
Published: (2023)
by: Carloni, Gianluca, et al.
Published: (2023)
Quantized Vision-Language Models for Damage Assessment: A Comparative Study of LLaVA-1.5-7B Quantization Levels
by: Yasuno, Takato
Published: (2026)
by: Yasuno, Takato
Published: (2026)
Persona-aware and Explainable Bikeability Assessment: A Vision-Language Model Approach
by: Dai, Yilong, et al.
Published: (2026)
by: Dai, Yilong, et al.
Published: (2026)
Language Predicts Identity Fusion Across Cultures and Reveals Divergent Pathways to Violence
by: Wright, Devin R., et al.
Published: (2026)
by: Wright, Devin R., et al.
Published: (2026)
Dynamic Personality Adaptation in Large Language Models via State Machines
by: Pielage, Leon, et al.
Published: (2026)
by: Pielage, Leon, et al.
Published: (2026)
StyleID: A Perception-Aware Dataset and Metric for Stylization-Agnostic Facial Identity Recognition
by: Yun, Kwan, et al.
Published: (2026)
by: Yun, Kwan, et al.
Published: (2026)
LLM-empowered Dynamic Prompt Routing for Vision-Language Models Tuning under Long-Tailed Distributions
by: Jia, Yongju, et al.
Published: (2025)
by: Jia, Yongju, et al.
Published: (2025)
Unpacking the Eye of the Beholder: Social Location, Identity, and the Moving Target of Political Perspectives
by: Sirotkina, Elena
Published: (2026)
by: Sirotkina, Elena
Published: (2026)
Tversky Neural Networks: Psychologically Plausible Deep Learning with Differentiable Tversky Similarity
by: Doumbouya, Moussa Koulako Bala, et al.
Published: (2025)
by: Doumbouya, Moussa Koulako Bala, et al.
Published: (2025)
Discovering Differences in Strategic Behavior Between Humans and LLMs
by: Wang, Caroline, et al.
Published: (2026)
by: Wang, Caroline, et al.
Published: (2026)
Fine-Tuning Vision-Language Models for Understanding Current Damage and Scoring Priority with Quality Guard Agent
by: Yasuno, Takato
Published: (2026)
by: Yasuno, Takato
Published: (2026)
GeoVision Labeler: Zero-Shot Geospatial Classification with Vision and Language Models
by: Hacheme, Gilles Quentin, et al.
Published: (2025)
by: Hacheme, Gilles Quentin, et al.
Published: (2025)
Frequency-Decomposed INR for NIR-Assisted Low-Light RGB Image Denoising
by: Shi, Ligen, et al.
Published: (2026)
by: Shi, Ligen, et al.
Published: (2026)
Mechanisms of Prompt-Induced Hallucination in Vision-Language Models
by: Rudman, William, et al.
Published: (2026)
by: Rudman, William, et al.
Published: (2026)
STimage-1K4M: A histopathology image-gene expression dataset for spatial transcriptomics
by: Chen, Jiawen, et al.
Published: (2024)
by: Chen, Jiawen, et al.
Published: (2024)
The Representational Alignment between Humans and Language Models is implicitly driven by a Concreteness Effect
by: Iaia, Cosimo, et al.
Published: (2025)
by: Iaia, Cosimo, et al.
Published: (2025)
Cheap Learning: Maximising Performance of Language Models for Social Data Science Using Minimal Data
by: Castro-Gonzalez, Leonardo, et al.
Published: (2024)
by: Castro-Gonzalez, Leonardo, et al.
Published: (2024)
Similar Items
-
FairDeDup: Detecting and Mitigating Vision-Language Fairness Disparities in Semantic Dataset Deduplication
by: Slyman, Eric, et al.
Published: (2024) -
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
by: Shahin, Nada, et al.
Published: (2025) -
Therapy as an NLP Task: Psychologists' Comparison of LLMs and Human Peers in CBT
by: Iftikhar, Zainab, et al.
Published: (2024) -
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
by: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Published: (2025) -
Topological Structure Description for Artcode Detection Using the Shape of Orientation Histogram
by: Xu, Liming, et al.
Published: (2025)