NL-Eye: Abductive NLI for Images
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ventura, Mor, Toker, Michael, Calderon, Nitay, Gekhman, Zorik, Bitton, Yonatan, Reichart, Roi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DeLeaker: Dynamic Inference-Time Reweighting For Semantic Leakage Mitigation in Text-to-Image Models
von: Ventura, Mor, et al.
Veröffentlicht: (2025)
von: Ventura, Mor, et al.
Veröffentlicht: (2025)
Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis
von: Ventura, Mor, et al.
Veröffentlicht: (2026)
von: Ventura, Mor, et al.
Veröffentlicht: (2026)
LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
von: Toker, Gilat, et al.
Veröffentlicht: (2026)
von: Toker, Gilat, et al.
Veröffentlicht: (2026)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
von: Orgad, Hadas, et al.
Veröffentlicht: (2024)
von: Orgad, Hadas, et al.
Veröffentlicht: (2024)
On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
von: Calderon, Nitay, et al.
Veröffentlicht: (2024)
von: Calderon, Nitay, et al.
Veröffentlicht: (2024)
Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factuality
von: Calderon, Nitay, et al.
Veröffentlicht: (2026)
von: Calderon, Nitay, et al.
Veröffentlicht: (2026)
Diffusion Lens: Interpreting Text Encoders in Text-to-Image Pipelines
von: Toker, Michael, et al.
Veröffentlicht: (2024)
von: Toker, Michael, et al.
Veröffentlicht: (2024)
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
von: Calderon, Nitay, et al.
Veröffentlicht: (2025)
von: Calderon, Nitay, et al.
Veröffentlicht: (2025)
Seeing Isn't Knowing: Do VLMs Know When Not to Answer Spatial Questions (and Why)?
von: Zhang, Yue, et al.
Veröffentlicht: (2026)
von: Zhang, Yue, et al.
Veröffentlicht: (2026)
Error-Driven Scene Editing for 3D Grounding in Large Language Models
von: Zhang, Yue, et al.
Veröffentlicht: (2025)
von: Zhang, Yue, et al.
Veröffentlicht: (2025)
Don't Fight Hallucinations, Use Them: Estimating Image Realism using NLI over Atomic Facts
von: Rykov, Elisei, et al.
Veröffentlicht: (2025)
von: Rykov, Elisei, et al.
Veröffentlicht: (2025)
Leveraging NTPs for Efficient Hallucination Detection in VLMs
von: Azachi, Ofir, et al.
Veröffentlicht: (2025)
von: Azachi, Ofir, et al.
Veröffentlicht: (2025)
EditInspector: A Benchmark for Evaluation of Text-Guided Image Edits
von: Yosef, Ron, et al.
Veröffentlicht: (2025)
von: Yosef, Ron, et al.
Veröffentlicht: (2025)
DOCCI: Descriptions of Connected and Contrasting Images
von: Onoe, Yasumasa, et al.
Veröffentlicht: (2024)
von: Onoe, Yasumasa, et al.
Veröffentlicht: (2024)
Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs
von: Gekhman, Zorik, et al.
Veröffentlicht: (2026)
von: Gekhman, Zorik, et al.
Veröffentlicht: (2026)
Navigating Cultural Chasms: Exploring and Unlocking the Cultural POV of Text-To-Image Models
von: Ventura, Mor, et al.
Veröffentlicht: (2023)
von: Ventura, Mor, et al.
Veröffentlicht: (2023)
Video-STaR: Self-Training Enables Video Instruction Tuning with Any Supervision
von: Zohar, Orr, et al.
Veröffentlicht: (2024)
von: Zohar, Orr, et al.
Veröffentlicht: (2024)
Mismatch Quest: Visual and Textual Feedback for Image-Text Misalignment
von: Gordon, Brian, et al.
Veröffentlicht: (2023)
von: Gordon, Brian, et al.
Veröffentlicht: (2023)
3DLLM-Mem: Long-Term Spatial-Temporal Memory for Embodied 3D Large Language Model
von: Hu, Wenbo, et al.
Veröffentlicht: (2025)
von: Hu, Wenbo, et al.
Veröffentlicht: (2025)
Measuring the Robustness of NLP Models to Domain Shifts
von: Calderon, Nitay, et al.
Veröffentlicht: (2023)
von: Calderon, Nitay, et al.
Veröffentlicht: (2023)
Towards Conversational Medical AI with Eyes, Ears and a Voice
von: Shah, Meet, et al.
Veröffentlicht: (2026)
von: Shah, Meet, et al.
Veröffentlicht: (2026)
MulTaBench: Benchmarking Multimodal Tabular Learning with Text and Image
von: Arazi, Alan, et al.
Veröffentlicht: (2026)
von: Arazi, Alan, et al.
Veröffentlicht: (2026)
HawkEye: Training Video-Text LLMs for Grounding Text in Videos
von: Wang, Yueqian, et al.
Veröffentlicht: (2024)
von: Wang, Yueqian, et al.
Veröffentlicht: (2024)
Enabling Chatbots with Eyes and Ears: An Immersive Multimodal Conversation System for Dynamic Interactions
von: Jang, Jihyoung, et al.
Veröffentlicht: (2025)
von: Jang, Jihyoung, et al.
Veröffentlicht: (2025)
Assessing Cognitive Effort in L2 Idiomatic Processing: An Eye-Tracking Dataset
von: Santos, Eduardo, et al.
Veröffentlicht: (2026)
von: Santos, Eduardo, et al.
Veröffentlicht: (2026)
Can LLMs Learn Macroeconomic Narratives from Social Media?
von: Gueta, Almog, et al.
Veröffentlicht: (2024)
von: Gueta, Almog, et al.
Veröffentlicht: (2024)
Enhancing Human-Computer Interaction in Chest X-ray Analysis using Vision and Language Model with Eye Gaze Patterns
von: Kim, Yunsoo, et al.
Veröffentlicht: (2024)
von: Kim, Yunsoo, et al.
Veröffentlicht: (2024)
Don't Blame the Annotator: Bias Already Starts in the Annotation Instructions
von: Parmar, Mihir, et al.
Veröffentlicht: (2022)
von: Parmar, Mihir, et al.
Veröffentlicht: (2022)
Padding Tone: A Mechanistic Analysis of Padding Tokens in T2I Models
von: Toker, Michael, et al.
Veröffentlicht: (2025)
von: Toker, Michael, et al.
Veröffentlicht: (2025)
Black Swan: Abductive and Defeasible Video Reasoning in Unpredictable Events
von: Chinchure, Aditya, et al.
Veröffentlicht: (2024)
von: Chinchure, Aditya, et al.
Veröffentlicht: (2024)
Visual Riddles: a Commonsense and World Knowledge Challenge for Large Vision and Language Models
von: Bitton-Guetta, Nitzan, et al.
Veröffentlicht: (2024)
von: Bitton-Guetta, Nitzan, et al.
Veröffentlicht: (2024)
Fine-Grained Detection of Context-Grounded Hallucinations Using LLMs
von: Peisakhovsky, Yehonatan, et al.
Veröffentlicht: (2025)
von: Peisakhovsky, Yehonatan, et al.
Veröffentlicht: (2025)
Abductive Ego-View Accident Video Understanding for Safe Driving Perception
von: Fang, Jianwu, et al.
Veröffentlicht: (2024)
von: Fang, Jianwu, et al.
Veröffentlicht: (2024)
Losing Visual Needles in Image Haystacks: Vision Language Models are Easily Distracted in Short and Long Contexts
von: Sharma, Aditya, et al.
Veröffentlicht: (2024)
von: Sharma, Aditya, et al.
Veröffentlicht: (2024)
All in an Aggregated Image for In-Image Learning
von: Wang, Lei, et al.
Veröffentlicht: (2024)
von: Wang, Lei, et al.
Veröffentlicht: (2024)
The Colorful Future of LLMs: Evaluating and Improving LLMs as Emotional Supporters for Queer Youth
von: Lissak, Shir, et al.
Veröffentlicht: (2024)
von: Lissak, Shir, et al.
Veröffentlicht: (2024)
TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation
von: Feng, Weixi, et al.
Veröffentlicht: (2024)
von: Feng, Weixi, et al.
Veröffentlicht: (2024)
ImageInWords: Unlocking Hyper-Detailed Image Descriptions
von: Garg, Roopal, et al.
Veröffentlicht: (2024)
von: Garg, Roopal, et al.
Veröffentlicht: (2024)
Fine-Grained Image-Text Alignment in Medical Imaging Enables Explainable Cyclic Image-Report Generation
von: Chen, Wenting, et al.
Veröffentlicht: (2023)
von: Chen, Wenting, et al.
Veröffentlicht: (2023)
TALC: Time-Aligned Captions for Multi-Scene Text-to-Video Generation
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
von: Bansal, Hritik, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DeLeaker: Dynamic Inference-Time Reweighting For Semantic Leakage Mitigation in Text-to-Image Models
von: Ventura, Mor, et al.
Veröffentlicht: (2025) -
Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis
von: Ventura, Mor, et al.
Veröffentlicht: (2026) -
LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
von: Toker, Gilat, et al.
Veröffentlicht: (2026) -
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
von: Orgad, Hadas, et al.
Veröffentlicht: (2024) -
On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
von: Calderon, Nitay, et al.
Veröffentlicht: (2024)