Causal Understanding For Video Question Answering
Fuente:
arXiv
Saved in:
| Main Authors: | Guda, Bhanu Prakash Reddy, Kulkarni, Tanmay, Sampath, Adithya, Sathyendra, Swarnashree Mysore |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Chasing the Tail: Effective Rubric-based Reward Modeling for Large Language Model Post-Training
by: Zhang, Junkai, et al.
Published: (2025)
by: Zhang, Junkai, et al.
Published: (2025)
Dynamic Strategy Planning for Efficient Question Answering with Large Language Models
by: Parekh, Tanmay, et al.
Published: (2024)
by: Parekh, Tanmay, et al.
Published: (2024)
Causal Question Answering with Reinforcement Learning
by: Blübaum, Lukas, et al.
Published: (2023)
by: Blübaum, Lukas, et al.
Published: (2023)
Quantifying First-Order Markov Violations in Noisy Reinforcement Learning: A Causal Discovery Approach
by: Mysore, Naveen
Published: (2025)
by: Mysore, Naveen
Published: (2025)
CoDi: Conversational Distillation for Grounded Question Answering
by: Huber, Patrick, et al.
Published: (2024)
by: Huber, Patrick, et al.
Published: (2024)
Evaluating Large Language Models on Solved and Unsolved Problems in Graph Theory: Implications for Computing Education
by: Kulkarni, Adithya, et al.
Published: (2026)
by: Kulkarni, Adithya, et al.
Published: (2026)
Multi-Dimensional Evaluation of Text Summarization with In-Context Learning
by: Jain, Sameer, et al.
Published: (2023)
by: Jain, Sameer, et al.
Published: (2023)
Mixture of Demonstrations for Textual Graph Understanding and Question Answering
by: Wu, Yukun, et al.
Published: (2026)
by: Wu, Yukun, et al.
Published: (2026)
Compositional 4D Dynamic Scenes Understanding with Physics Priors for Video Question Answering
by: Wang, Xingrui, et al.
Published: (2024)
by: Wang, Xingrui, et al.
Published: (2024)
Efficient Multi-Model Orchestration for Self-Hosted Large Language Models
by: Vangala, Bhanu Prakash, et al.
Published: (2025)
by: Vangala, Bhanu Prakash, et al.
Published: (2025)
Long-Context Long-Form Question Answering for Legal Domain
by: Kulkarni, Anagha, et al.
Published: (2026)
by: Kulkarni, Anagha, et al.
Published: (2026)
Question-Answering Dense Video Events
by: Qin, Hangyu, et al.
Published: (2024)
by: Qin, Hangyu, et al.
Published: (2024)
Evaluating Correctness and Faithfulness of Instruction-Following Models for Question Answering
by: Adlakha, Vaibhav, et al.
Published: (2023)
by: Adlakha, Vaibhav, et al.
Published: (2023)
FIQ: Fundamental Question Generation with the Integration of Question Embeddings for Video Question Answering
by: Oh, Ju-Young, et al.
Published: (2025)
by: Oh, Ju-Young, et al.
Published: (2025)
Understanding Network Behaviors through Natural Language Question-Answering
by: Xing, Mingzhe, et al.
Published: (2025)
by: Xing, Mingzhe, et al.
Published: (2025)
QMOS: Enhancing LLMs for Telecommunication with Question Masked loss and Option Shuffling
by: Guda, Blessed, et al.
Published: (2024)
by: Guda, Blessed, et al.
Published: (2024)
SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images
by: Shen, Jialu, et al.
Published: (2026)
by: Shen, Jialu, et al.
Published: (2026)
Towards Better Generalization in Open-Domain Question Answering by Mitigating Context Memorization
by: Zhang, Zixuan, et al.
Published: (2024)
by: Zhang, Zixuan, et al.
Published: (2024)
Learning Question-Aware Keyframe Selection with Synthetic Supervision for Video Question Answering
by: Kwon, Minchan, et al.
Published: (2026)
by: Kwon, Minchan, et al.
Published: (2026)
MapQA: Open-domain Geospatial Question Answering on Map Data
by: Li, Zekun, et al.
Published: (2025)
by: Li, Zekun, et al.
Published: (2025)
Towards Fine-Grained Video Question Answering
by: Dai, Wei, et al.
Published: (2025)
by: Dai, Wei, et al.
Published: (2025)
ChainReaction: Causal Chain-Guided Reasoning for Modular and Explainable Causal-Why Video Question Answering
by: Parmar, Paritosh, et al.
Published: (2025)
by: Parmar, Paritosh, et al.
Published: (2025)
KARMA-MV: A Benchmark for Causal Question Answering on Music Videos
by: Ghosh, Archishman, et al.
Published: (2026)
by: Ghosh, Archishman, et al.
Published: (2026)
Knowledge-Guided Time-Varying Causal Inference for Arctic Sea Ice Dynamics
by: Sampath, Akila, et al.
Published: (2026)
by: Sampath, Akila, et al.
Published: (2026)
Understanding and Supporting Formal Email Exchange by Answering AI-Generated Questions
by: Miura, Yusuke, et al.
Published: (2025)
by: Miura, Yusuke, et al.
Published: (2025)
Ego-Grounding for Personalized Question-Answering in Egocentric Videos
by: Xiao, Junbin, et al.
Published: (2026)
by: Xiao, Junbin, et al.
Published: (2026)
Commonsense Video Question Answering through Video-Grounded Entailment Tree Reasoning
by: Liu, Huabin, et al.
Published: (2025)
by: Liu, Huabin, et al.
Published: (2025)
Uncertainty-Guided Self-Questioning and Answering for Video-Language Alignment
by: Chen, Jin, et al.
Published: (2024)
by: Chen, Jin, et al.
Published: (2024)
VQA$^2$: Visual Question Answering for Video Quality Assessment
by: Jia, Ziheng, et al.
Published: (2024)
by: Jia, Ziheng, et al.
Published: (2024)
Explicit Abstention Knobs for Predictable Reliability in Video Question Answering
by: Ortiz, Jorge
Published: (2025)
by: Ortiz, Jorge
Published: (2025)
Semantic Event Graphs for Long-Form Video Question Answering
by: Dixit, Aradhya, et al.
Published: (2026)
by: Dixit, Aradhya, et al.
Published: (2026)
CogStream: Context-guided Streaming Video Question Answering
by: Zhao, Zicheng, et al.
Published: (2025)
by: Zhao, Zicheng, et al.
Published: (2025)
Eliminating the Language Bias for Visual Question Answering with fine-grained Causal Intervention
by: Liu, Ying, et al.
Published: (2024)
by: Liu, Ying, et al.
Published: (2024)
Optimizing Traffic Signal Control using High-Dimensional State Representation and Efficient Deep Reinforcement Learning
by: Francis, Lawrence, et al.
Published: (2024)
by: Francis, Lawrence, et al.
Published: (2024)
The Essence of Contextual Understanding in Theory of Mind: A Study on Question Answering with Story Characters
by: Zhou, Chulun, et al.
Published: (2025)
by: Zhou, Chulun, et al.
Published: (2025)
MIXRAG : Mixture-of-Experts Retrieval-Augmented Generation for Textual Graph Understanding and Question Answering
by: Liu, Lihui, et al.
Published: (2025)
by: Liu, Lihui, et al.
Published: (2025)
Encoding and Controlling Global Semantics for Long-form Video Question Answering
by: Nguyen, Thong Thanh, et al.
Published: (2024)
by: Nguyen, Thong Thanh, et al.
Published: (2024)
Enhancing Long Video Question Answering with Scene-Localized Frame Grouping
by: Yang, Xuyi, et al.
Published: (2025)
by: Yang, Xuyi, et al.
Published: (2025)
Open-Ended Multi-Modal Relational Reasoning for Video Question Answering
by: Luo, Haozheng, et al.
Published: (2020)
by: Luo, Haozheng, et al.
Published: (2020)
VideoQA-SC: Adaptive Semantic Communication for Video Question Answering
by: Guo, Jiangyuan, et al.
Published: (2024)
by: Guo, Jiangyuan, et al.
Published: (2024)
Similar Items
-
Chasing the Tail: Effective Rubric-based Reward Modeling for Large Language Model Post-Training
by: Zhang, Junkai, et al.
Published: (2025) -
Dynamic Strategy Planning for Efficient Question Answering with Large Language Models
by: Parekh, Tanmay, et al.
Published: (2024) -
Causal Question Answering with Reinforcement Learning
by: Blübaum, Lukas, et al.
Published: (2023) -
Quantifying First-Order Markov Violations in Noisy Reinforcement Learning: A Causal Discovery Approach
by: Mysore, Naveen
Published: (2025) -
CoDi: Conversational Distillation for Grounded Question Answering
by: Huber, Patrick, et al.
Published: (2024)