DIAGRAMS: A Review Framework for Reasoning-Level Attribution in Diagram QA
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Iyengar, Anirudh Iyengar Kaniyar Narayana, Kumar, Tampu Ravi, Suri, Manan, Bommireddy, Raviteja, Manocha, Dinesh, Mathur, Puneet, Gupta, Vivek |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DRAGON: A Benchmark for Evidence-Grounded Visual Reasoning over Diagrams
von: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Veröffentlicht: (2026)
von: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Veröffentlicht: (2026)
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
von: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Veröffentlicht: (2025)
von: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Veröffentlicht: (2025)
Follow the Flow: Fine-grained Flowchart Attribution with Neurosymbolic Agents
von: Suri, Manan, et al.
Veröffentlicht: (2025)
von: Suri, Manan, et al.
Veröffentlicht: (2025)
VisDoM: Multi-Document QA with Visually Rich Elements Using Multimodal Retrieval-Augmented Generation
von: Suri, Manan, et al.
Veröffentlicht: (2024)
von: Suri, Manan, et al.
Veröffentlicht: (2024)
ChartLens: Fine-grained Visual Attribution in Charts
von: Suri, Manan, et al.
Veröffentlicht: (2025)
von: Suri, Manan, et al.
Veröffentlicht: (2025)
Structured Uncertainty guided Clarification for LLM Agents
von: Suri, Manan, et al.
Veröffentlicht: (2025)
von: Suri, Manan, et al.
Veröffentlicht: (2025)
QUIETT: Query-Independent Table Transformation for Robust Reasoning
von: Najpande, Gaurav, et al.
Veröffentlicht: (2026)
von: Najpande, Gaurav, et al.
Veröffentlicht: (2026)
DocEdit-v2: Document Structure Editing Via Multimodal LLM Grounding
von: Suri, Manan, et al.
Veröffentlicht: (2024)
von: Suri, Manan, et al.
Veröffentlicht: (2024)
Generation-Time vs. Post-hoc Citation: A Holistic Evaluation of LLM Attribution
von: Saxena, Yash, et al.
Veröffentlicht: (2025)
von: Saxena, Yash, et al.
Veröffentlicht: (2025)
Learning Illumination Control in Diffusion Models
von: Anand, Nishit, et al.
Veröffentlicht: (2026)
von: Anand, Nishit, et al.
Veröffentlicht: (2026)
An Agentic Approach to Automatic Creation of P&ID Diagrams from Natural Language Descriptions
von: Gowaikar, Shreeyash, et al.
Veröffentlicht: (2024)
von: Gowaikar, Shreeyash, et al.
Veröffentlicht: (2024)
Quantum-Enhanced Distributed Sensor Fusion: Lower Bounds on Aggregation from Projection Noise to Heisenberg-Limited Byzantine-Tolerant Networks
von: Iyer, Vasanth, et al.
Veröffentlicht: (2026)
von: Iyer, Vasanth, et al.
Veröffentlicht: (2026)
Proof of Thought : Neurosymbolic Program Synthesis allows Robust and Interpretable Reasoning
von: Ganguly, Debargha, et al.
Veröffentlicht: (2024)
von: Ganguly, Debargha, et al.
Veröffentlicht: (2024)
Future of AI Models: A Computational perspective on Model collapse
von: Satharasi, Trivikram, et al.
Veröffentlicht: (2025)
von: Satharasi, Trivikram, et al.
Veröffentlicht: (2025)
Non-Invasive Qualitative Vibration Analysis using Event Camera
von: Bane, Dwijay, et al.
Veröffentlicht: (2024)
von: Bane, Dwijay, et al.
Veröffentlicht: (2024)
Multi-LLM QA with Embodied Exploration
von: Patel, Bhrij, et al.
Veröffentlicht: (2024)
von: Patel, Bhrij, et al.
Veröffentlicht: (2024)
A large language model-type architecture for high-dimensional molecular potential energy surfaces
von: Zhu, Xiao, et al.
Veröffentlicht: (2024)
von: Zhu, Xiao, et al.
Veröffentlicht: (2024)
ChartCitor: Multi-Agent Framework for Fine-Grained Chart Visual Attribution
von: Goswami, Kanika, et al.
Veröffentlicht: (2025)
von: Goswami, Kanika, et al.
Veröffentlicht: (2025)
On the $O(1/T)$ Convergence of Alternating Gradient Descent-Ascent in Bilinear Games
von: Nan, Tianlong, et al.
Veröffentlicht: (2025)
von: Nan, Tianlong, et al.
Veröffentlicht: (2025)
TraceBack: Multi-Agent Decomposition for Fine-Grained Table Attribution
von: Anvekar, Tejas, et al.
Veröffentlicht: (2026)
von: Anvekar, Tejas, et al.
Veröffentlicht: (2026)
Contested Route Planning
von: Černý, Jakub, et al.
Veröffentlicht: (2025)
von: Černý, Jakub, et al.
Veröffentlicht: (2025)
Mitigating Memorization in LLMs using Activation Steering
von: Suri, Manan, et al.
Veröffentlicht: (2025)
von: Suri, Manan, et al.
Veröffentlicht: (2025)
MTMed3D: A Multi-Task Transformer-Based Model for 3D Medical Imaging
von: Li, Fan, et al.
Veröffentlicht: (2025)
von: Li, Fan, et al.
Veröffentlicht: (2025)
Low Power Neuromorphic EMG Gesture Classification
von: Bezugam, Sai Sukruth, et al.
Veröffentlicht: (2022)
von: Bezugam, Sai Sukruth, et al.
Veröffentlicht: (2022)
Quantum circuit and mapping algorithms for wavepacket dynamics: case study of anharmonic hydrogen bonds in protonated and hydroxide water clusters
von: Saha, Debadrita, et al.
Veröffentlicht: (2024)
von: Saha, Debadrita, et al.
Veröffentlicht: (2024)
Inst4DGS: Instance-Decomposed 4D Gaussian Splatting with Multi-Video Label Permutation Learning
von: Lee, Yonghan, et al.
Veröffentlicht: (2026)
von: Lee, Yonghan, et al.
Veröffentlicht: (2026)
PACE: Data-Driven Virtual Agent Interaction in Dense and Cluttered Environments
von: Mullen, James, et al.
Veröffentlicht: (2023)
von: Mullen, James, et al.
Veröffentlicht: (2023)
Disentangling Polysemantic Neurons with a Null-Calibrated Polysemanticity Index and Causal Patch Interventions
von: Gupta, Manan, et al.
Veröffentlicht: (2025)
von: Gupta, Manan, et al.
Veröffentlicht: (2025)
Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations
von: Gupta, Manan, et al.
Veröffentlicht: (2026)
von: Gupta, Manan, et al.
Veröffentlicht: (2026)
Latent Phase-Shift Rollback: Inference-Time Error Correction via Residual Stream Monitoring and KV-Cache Steering
von: Gupta, Manan, et al.
Veröffentlicht: (2026)
von: Gupta, Manan, et al.
Veröffentlicht: (2026)
Do Vision-Language Models Understand Compound Nouns?
von: Kumar, Sonal, et al.
Veröffentlicht: (2024)
von: Kumar, Sonal, et al.
Veröffentlicht: (2024)
MedSR-Vision: Deep Learning Framework for Multi-Domain Medical Image Super-Resolution
von: Gurappa, Subhash, et al.
Veröffentlicht: (2026)
von: Gurappa, Subhash, et al.
Veröffentlicht: (2026)
Layered Graph Security Games
von: Černý, Jakub, et al.
Veröffentlicht: (2024)
von: Černý, Jakub, et al.
Veröffentlicht: (2024)
Raw2Event: Converting Raw Frame Camera into Event Camera
von: Ning, Zijie, et al.
Veröffentlicht: (2025)
von: Ning, Zijie, et al.
Veröffentlicht: (2025)
Visual Description Grounding Reduces Hallucinations and Boosts Reasoning in LVLMs
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
DocuBits: VR Document Decomposition for Procedural Task Completion
von: Lee, Geonsun, et al.
Veröffentlicht: (2024)
von: Lee, Geonsun, et al.
Veröffentlicht: (2024)
SLAT-Phys: Fast Material Property Field Prediction from Structured 3D Latents
von: Das, Rocktim Jyoti, et al.
Veröffentlicht: (2026)
von: Das, Rocktim Jyoti, et al.
Veröffentlicht: (2026)
Listen2Scene: Interactive material-aware binaural sound propagation for reconstructed 3D scenes
von: Ratnarajah, Anton, et al.
Veröffentlicht: (2023)
von: Ratnarajah, Anton, et al.
Veröffentlicht: (2023)
PhaseNet++: Phase-Aware Frequency-Domain Anomaly Detection for Industrial Control Systems via Phase Coherence Graphs
von: Bommireddy, Raviteja, et al.
Veröffentlicht: (2026)
von: Bommireddy, Raviteja, et al.
Veröffentlicht: (2026)
SPIRIT: Short-term Prediction of solar IRradIance for zero-shot Transfer learning using Foundation Models
von: Mishra, Aditya, et al.
Veröffentlicht: (2025)
von: Mishra, Aditya, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DRAGON: A Benchmark for Evidence-Grounded Visual Reasoning over Diagrams
von: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Veröffentlicht: (2026) -
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
von: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Veröffentlicht: (2025) -
Follow the Flow: Fine-grained Flowchart Attribution with Neurosymbolic Agents
von: Suri, Manan, et al.
Veröffentlicht: (2025) -
VisDoM: Multi-Document QA with Visually Rich Elements Using Multimodal Retrieval-Augmented Generation
von: Suri, Manan, et al.
Veröffentlicht: (2024) -
ChartLens: Fine-grained Visual Attribution in Charts
von: Suri, Manan, et al.
Veröffentlicht: (2025)