MVTamperBench: Evaluating Robustness of Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Agarwal, Amit, Panda, Srikant, Charles, Angeline, Kumar, Bhargava, Patel, Hitesh, Pattnayak, Priyaranjan, Rafi, Taki Hasan, Kumar, Tejaswini, Meghwani, Hansa, Gupta, Karan, Chae, Dong-Kyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Aligning LLMs for Multilingual Consistency in Enterprise Applications
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
A Landmark-Aware Visual Navigation Dataset
von: Johnson, Faith, et al.
Veröffentlicht: (2024)
von: Johnson, Faith, et al.
Veröffentlicht: (2024)
Named entity recognition for Serbian legal documents: Design, methodology and dataset development
von: Kalušev, Vladimir, et al.
Veröffentlicht: (2025)
von: Kalušev, Vladimir, et al.
Veröffentlicht: (2025)
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
von: Cao, Bryan Bo, et al.
Veröffentlicht: (2024)
von: Cao, Bryan Bo, et al.
Veröffentlicht: (2024)
Semantic Modeling for World-Centered Architectures
von: Mantsivoda, Andrei, et al.
Veröffentlicht: (2026)
von: Mantsivoda, Andrei, et al.
Veröffentlicht: (2026)
Structured Basis Function Networks: Loss-Centric Multi-Hypothesis Ensembles with Controllable Diversity
von: Dominguez, Alejandro Rodriguez, et al.
Veröffentlicht: (2025)
von: Dominguez, Alejandro Rodriguez, et al.
Veröffentlicht: (2025)
RCI: A Score for Evaluating Global and Local Reasoning in Multimodal Benchmarks
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
von: Agarwal, Amit, et al.
Veröffentlicht: (2025)
LRCP: Low-Rank Compressibility Guided Visual Token Pruning for Efficient LVLMs
von: Lu, Hongyu, et al.
Veröffentlicht: (2026)
von: Lu, Hongyu, et al.
Veröffentlicht: (2026)
On Woolhouse's Cotton-Spinning Problem
von: Groote, Jan Friso, et al.
Veröffentlicht: (2024)
von: Groote, Jan Friso, et al.
Veröffentlicht: (2024)
A Cost-Effective Eye-Tracker for Early Detection of Mild Cognitive Impairment
von: Greco, Danilo, et al.
Veröffentlicht: (2024)
von: Greco, Danilo, et al.
Veröffentlicht: (2024)
Abductive explanations of classifiers under constraints: Complexity and properties
von: Cooper, Martin, et al.
Veröffentlicht: (2024)
von: Cooper, Martin, et al.
Veröffentlicht: (2024)
Deep Spectral Meshes: Multi-Frequency Facial Mesh Processing with Graph Neural Networks
von: Kosk, Robert, et al.
Veröffentlicht: (2024)
von: Kosk, Robert, et al.
Veröffentlicht: (2024)
On measuring grounding and generalizing grounding problems
von: Quigley, Daniel, et al.
Veröffentlicht: (2025)
von: Quigley, Daniel, et al.
Veröffentlicht: (2025)
ROI-GS: Interest-based Local Quality 3D Gaussian Splatting
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
ROI-NeRFs: Hi-Fi Visualization of Objects of Interest within a Scene by NeRFs Composition
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
Hyperbox Mixture Regression for Process Performance Prediction in Antibody Production
von: Nik-Khorasani, Ali, et al.
Veröffentlicht: (2024)
von: Nik-Khorasani, Ali, et al.
Veröffentlicht: (2024)
TowerVision: Understanding and Improving Multilinguality in Vision-Language Models
von: Viveiros, André G., et al.
Veröffentlicht: (2025)
von: Viveiros, André G., et al.
Veröffentlicht: (2025)
Swish-T : Enhancing Swish Activation with Tanh Bias for Improved Neural Network Performance
von: Seo, Youngmin, et al.
Veröffentlicht: (2024)
von: Seo, Youngmin, et al.
Veröffentlicht: (2024)
StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
von: Merugu, Ranjith, et al.
Veröffentlicht: (2025)
von: Merugu, Ranjith, et al.
Veröffentlicht: (2025)
Understanding the Uncertainty of LLM Explanations: A Perspective Based on Reasoning Topology
von: Da, Longchao, et al.
Veröffentlicht: (2025)
von: Da, Longchao, et al.
Veröffentlicht: (2025)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
YOLO Ensemble for UAV-based Multispectral Defect Detection in Wind Turbine Components
von: Svystun, Serhii, et al.
Veröffentlicht: (2025)
von: Svystun, Serhii, et al.
Veröffentlicht: (2025)
Smelly, dense, and spreaded: The Object Detection for Olfactory References (ODOR) dataset
von: Zinnen, Mathias, et al.
Veröffentlicht: (2025)
von: Zinnen, Mathias, et al.
Veröffentlicht: (2025)
Who's Asking? Investigating Bias Through the Lens of Disability Framed Queries in LLMs
von: Hari, Vishnu, et al.
Veröffentlicht: (2025)
von: Hari, Vishnu, et al.
Veröffentlicht: (2025)
Polarization-Based Eye Tracking with Personalized Siamese Architectures
von: Kalkanli, Beyza, et al.
Veröffentlicht: (2026)
von: Kalkanli, Beyza, et al.
Veröffentlicht: (2026)
TerraSeg: Self-Supervised Ground Segmentation for Any LiDAR
von: Lentsch, Ted, et al.
Veröffentlicht: (2026)
von: Lentsch, Ted, et al.
Veröffentlicht: (2026)
UNION: Unsupervised 3D Object Detection using Object Appearance-based Pseudo-Classes
von: Lentsch, Ted, et al.
Veröffentlicht: (2024)
von: Lentsch, Ted, et al.
Veröffentlicht: (2024)
AI Model for Predicting Binding Affinity of Antidiabetic Compounds Targeting PPAR
von: Aman, La Ode, et al.
Veröffentlicht: (2024)
von: Aman, La Ode, et al.
Veröffentlicht: (2024)
Semantic2Graph: Graph-based Multi-modal Feature Fusion for Action Segmentation in Videos
von: Zhang, Junbin, et al.
Veröffentlicht: (2022)
von: Zhang, Junbin, et al.
Veröffentlicht: (2022)
CVCM Track Circuits Pre-emptive Failure Diagnostics for Predictive Maintenance Using Deep Neural Networks
von: Mukherjee, Debdeep, et al.
Veröffentlicht: (2025)
von: Mukherjee, Debdeep, et al.
Veröffentlicht: (2025)
Contract-Driven QoE Auditing for Speech and Singing Services: From MOS Regression to Service Graphs
von: Du, Wenzhang
Veröffentlicht: (2025)
von: Du, Wenzhang
Veröffentlicht: (2025)
Hierarchical Point-Patch Fusion with Adaptive Patch Codebook for 3D Shape Anomaly Detection
von: Kang, Xueyang, et al.
Veröffentlicht: (2026)
von: Kang, Xueyang, et al.
Veröffentlicht: (2026)
Transforming faces into video stories -- VideoFace2.0
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
von: Brkljač, Branko, et al.
Veröffentlicht: (2025)
Advancing Brain Tumor Segmentation via Attention-based 3D U-Net Architecture and Digital Image Processing
von: Gad, Eyad, et al.
Veröffentlicht: (2025)
von: Gad, Eyad, et al.
Veröffentlicht: (2025)
Probabilistic Approach for Detection of High-Frequency Periodic Signals using an Event Camera
von: Ben-Ezra, David El-Chai, et al.
Veröffentlicht: (2022)
von: Ben-Ezra, David El-Chai, et al.
Veröffentlicht: (2022)
Heart Failure Prediction using Modal Decomposition and Masked Autoencoders for Scarce Echocardiography Databases
von: Bell-Navas, Andrés, et al.
Veröffentlicht: (2025)
von: Bell-Navas, Andrés, et al.
Veröffentlicht: (2025)
Mitigating Catastrophic Forgetting in Streaming Generative and Predictive Learning via Stateful Replay
von: Du, Wenzhang
Veröffentlicht: (2025)
von: Du, Wenzhang
Veröffentlicht: (2025)
Transfer learning with generative models for object detection on limited datasets
von: Paiano, Matteo, et al.
Veröffentlicht: (2024)
von: Paiano, Matteo, et al.
Veröffentlicht: (2024)
Decoder Generates Manufacturable Structures: A Framework for 3D-Printable Object Synthesis
von: Kumar, Abhishek
Veröffentlicht: (2026)
von: Kumar, Abhishek
Veröffentlicht: (2026)
Ähnliche Einträge
-
Aligning LLMs for Multilingual Consistency in Enterprise Applications
von: Agarwal, Amit, et al.
Veröffentlicht: (2025) -
A Landmark-Aware Visual Navigation Dataset
von: Johnson, Faith, et al.
Veröffentlicht: (2024) -
Named entity recognition for Serbian legal documents: Design, methodology and dataset development
von: Kalušev, Vladimir, et al.
Veröffentlicht: (2025) -
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025) -
Few-Class Arena: A Benchmark for Efficient Selection of Vision Models and Dataset Difficulty Measurement
von: Cao, Bryan Bo, et al.
Veröffentlicht: (2024)