VALUED -- Vision and Logical Understanding Evaluation Dataset
Fuente:
arXiv
Saved in:
| Main Authors: | Saha, Soumadeep, Saha, Saptarshi, Garain, Utpal |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Region Mixup
by: Saha, Saptarshi, et al.
Published: (2024)
by: Saha, Saptarshi, et al.
Published: (2024)
Language Models are Crossword Solvers
by: Saha, Soumadeep, et al.
Published: (2024)
by: Saha, Soumadeep, et al.
Published: (2024)
KisMATH: Do LLMs Have Knowledge of Implicit Structures in Mathematical Reasoning?
by: Saha, Soumadeep, et al.
Published: (2025)
by: Saha, Soumadeep, et al.
Published: (2025)
Evaluation of State-of-the-Art Deep Learning Techniques for Plant Disease and Pest Detection
by: Banerjee, Saptarshi, et al.
Published: (2025)
by: Banerjee, Saptarshi, et al.
Published: (2025)
Can Deep Learning Trigger Alerts from Mobile-Captured Images?
by: Sarkar, Pritisha, et al.
Published: (2025)
by: Sarkar, Pritisha, et al.
Published: (2025)
Cyclic Counterfactuals under Shift-Scale Interventions
by: Saha, Saptarshi, et al.
Published: (2025)
by: Saha, Saptarshi, et al.
Published: (2025)
Pause and Think: A Dataset and Benchmark for Video-Grounded Assistive Action Suggestion
by: Singh, Shivam, et al.
Published: (2026)
by: Singh, Shivam, et al.
Published: (2026)
Externally Validated Multi-Task Learning via Consistency Regularization Using Differentiable BI-RADS Features for Breast Ultrasound Tumor Segmentation
by: Zhang, Jingru, et al.
Published: (2025)
by: Zhang, Jingru, et al.
Published: (2025)
Deep Reinforcement Learning-driven Edge Offloading for Latency-constrained XR pipelines
by: Saha, Sourya, et al.
Published: (2026)
by: Saha, Sourya, et al.
Published: (2026)
Exploring the Frontier of Vision-Language Models: A Survey of Current Methodologies and Future Directions
by: Ghosh, Akash, et al.
Published: (2024)
by: Ghosh, Akash, et al.
Published: (2024)
Evaluating Vision Language Models (VLMs) for Radiology: A Comprehensive Analysis
by: Li, Frank, et al.
Published: (2025)
by: Li, Frank, et al.
Published: (2025)
Enhancing Adverse Drug Event Detection with Multimodal Dataset: Corpus Creation and Model Development
by: Sahoo, Pranab, et al.
Published: (2024)
by: Sahoo, Pranab, et al.
Published: (2024)
LogicQA: Logical Anomaly Detection with Vision Language Model Generated Questions
by: Kwon, Yejin, et al.
Published: (2025)
by: Kwon, Yejin, et al.
Published: (2025)
When Words Can't Capture It All: Towards Video-Based User Complaint Text Generation with Multimodal Video Complaint Dataset
by: Das, Sarmistha, et al.
Published: (2025)
by: Das, Sarmistha, et al.
Published: (2025)
FlowLearn: Evaluating Large Vision-Language Models on Flowchart Understanding
by: Pan, Huitong, et al.
Published: (2024)
by: Pan, Huitong, et al.
Published: (2024)
Investigating Deep Learning Models for Ejection Fraction Estimation from Echocardiography Videos
by: Saranyan, Shravan, et al.
Published: (2025)
by: Saranyan, Shravan, et al.
Published: (2025)
The Promise of Analog Deep Learning: Recent Advances, Challenges and Opportunities
by: Datar, Aditya, et al.
Published: (2024)
by: Datar, Aditya, et al.
Published: (2024)
Comparing and Integrating Different Notions of Representational Correspondence in Neural Systems
by: Wu, Jialin, et al.
Published: (2025)
by: Wu, Jialin, et al.
Published: (2025)
Predicting Road Crossing Behaviour using Pose Detection and Sequence Modelling
by: Dasgupta, Subhasis, et al.
Published: (2025)
by: Dasgupta, Subhasis, et al.
Published: (2025)
Local-to-Global Logical Explanations for Deep Vision Models
by: Vasu, Bhavan, et al.
Published: (2026)
by: Vasu, Bhavan, et al.
Published: (2026)
LeafNet: A Large-Scale Dataset and Comprehensive Benchmark for Foundational Vision-Language Understanding of Plant Diseases
by: Quoc, Khang Nguyen, et al.
Published: (2026)
by: Quoc, Khang Nguyen, et al.
Published: (2026)
Rice-VL: Evaluating Vision-Language Models for Cultural Understanding Across ASEAN Countries
by: Pranav, Tushar, et al.
Published: (2025)
by: Pranav, Tushar, et al.
Published: (2025)
Exploring Curriculum Learning for Vision-Language Tasks: A Study on Small-Scale Multimodal Training
by: Saha, Rohan, et al.
Published: (2024)
by: Saha, Rohan, et al.
Published: (2024)
SkySenseGPT: A Fine-Grained Instruction Tuning Dataset and Model for Remote Sensing Vision-Language Understanding
by: Luo, Junwei, et al.
Published: (2024)
by: Luo, Junwei, et al.
Published: (2024)
MetaLogic: Robustness Evaluation of Text-to-Image Models via Logically Equivalent Prompts
by: Shen, Yifan, et al.
Published: (2025)
by: Shen, Yifan, et al.
Published: (2025)
Compositional Attribute Imbalance in Vision Datasets
by: Chen, Jiayi, et al.
Published: (2025)
by: Chen, Jiayi, et al.
Published: (2025)
Improved Anomaly Detection in Medical Images via Mean Shift Density Enhancement
by: Kar, Pritam, et al.
Published: (2026)
by: Kar, Pritam, et al.
Published: (2026)
Benchmarking Large Vision-Language Models on CFMME: A Comprehensive Chinese Financial Multimodal Evaluation Dataset
by: Chen, Qian, et al.
Published: (2026)
by: Chen, Qian, et al.
Published: (2026)
Multi-Attention Stacked Ensemble for Lung Cancer Detection in CT Scans
by: Saha, Uzzal, et al.
Published: (2025)
by: Saha, Uzzal, et al.
Published: (2025)
ADP-FL-MedSeg: Adaptive Differential Privacy for Federated Medical Segmentation Across Diverse Modalities
by: Saha, Puja, et al.
Published: (2026)
by: Saha, Puja, et al.
Published: (2026)
Unlocking Financial Insights: An advanced Multimodal Summarization with Multimodal Output Framework for Financial Advisory Videos
by: Das, Sarmistha, et al.
Published: (2025)
by: Das, Sarmistha, et al.
Published: (2025)
Panoptic Segmentation and Labelling of Lumbar Spine Vertebrae using Modified Attention Unet
by: Pal, Rikathi, et al.
Published: (2024)
by: Pal, Rikathi, et al.
Published: (2024)
Exploring Explainability in Video Action Recognition
by: Saha, Avinab, et al.
Published: (2024)
by: Saha, Avinab, et al.
Published: (2024)
CLARIFY: A Specialist-Generalist Framework for Accurate and Lightweight Dermatological Visual Question Answering
by: Saha, Aranya, et al.
Published: (2025)
by: Saha, Aranya, et al.
Published: (2025)
Understanding the Transfer Limits of Vision Foundation Models
by: Huang, Shiqi, et al.
Published: (2026)
by: Huang, Shiqi, et al.
Published: (2026)
A Survey of Video Datasets for Grounded Event Understanding
by: Sanders, Kate, et al.
Published: (2024)
by: Sanders, Kate, et al.
Published: (2024)
Deep Unlearning: Fast and Efficient Gradient-free Approach to Class Forgetting
by: Kodge, Sangamesh, et al.
Published: (2023)
by: Kodge, Sangamesh, et al.
Published: (2023)
ToxVidLM: A Multimodal Framework for Toxicity Detection in Code-Mixed Videos
by: Maity, Krishanu, et al.
Published: (2024)
by: Maity, Krishanu, et al.
Published: (2024)
ToDo: Token Downsampling for Efficient Generation of High-Resolution Images
by: Smith, Ethan, et al.
Published: (2024)
by: Smith, Ethan, et al.
Published: (2024)
From Local Concepts to Universals: Evaluating the Multicultural Understanding of Vision-Language Models
by: Bhatia, Mehar, et al.
Published: (2024)
by: Bhatia, Mehar, et al.
Published: (2024)
Similar Items
-
Region Mixup
by: Saha, Saptarshi, et al.
Published: (2024) -
Language Models are Crossword Solvers
by: Saha, Soumadeep, et al.
Published: (2024) -
KisMATH: Do LLMs Have Knowledge of Implicit Structures in Mathematical Reasoning?
by: Saha, Soumadeep, et al.
Published: (2025) -
Evaluation of State-of-the-Art Deep Learning Techniques for Plant Disease and Pest Detection
by: Banerjee, Saptarshi, et al.
Published: (2025) -
Can Deep Learning Trigger Alerts from Mobile-Captured Images?
by: Sarkar, Pritisha, et al.
Published: (2025)