Neural-Symbolic VideoQA: Learning Compositional Spatio-Temporal Reasoning for Real-world Video Question Answering
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liang, Lili, Sun, Guanglu, Qiu, Jin, Zhang, Lizhong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VideoQA-SC: Adaptive Semantic Communication for Video Question Answering
von: Guo, Jiangyuan, et al.
Veröffentlicht: (2024)
von: Guo, Jiangyuan, et al.
Veröffentlicht: (2024)
Leveraging Static Relationships for Intra-Type and Inter-Type Message Passing in Video Question Answering
von: Liang, Lili, et al.
Veröffentlicht: (2025)
von: Liang, Lili, et al.
Veröffentlicht: (2025)
TUMTraffic-VideoQA: A Benchmark for Unified Spatio-Temporal Video Understanding in Traffic Scenes
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2025)
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2025)
QTG-VQA: Question-Type-Guided Architectural for VideoQA Systems
von: He, Zhixian, et al.
Veröffentlicht: (2024)
von: He, Zhixian, et al.
Veröffentlicht: (2024)
ReasVQA: Advancing VideoQA with Imperfect Reasoning Process
von: Liang, Jianxin, et al.
Veröffentlicht: (2025)
von: Liang, Jianxin, et al.
Veröffentlicht: (2025)
ENTER: Event Based Interpretable Reasoning for VideoQA
von: Ayyubi, Hammad, et al.
Veröffentlicht: (2025)
von: Ayyubi, Hammad, et al.
Veröffentlicht: (2025)
Video Flow as Time Series: Discovering Temporal Consistency and Variability for VideoQA
von: Song, Zijie, et al.
Veröffentlicht: (2025)
von: Song, Zijie, et al.
Veröffentlicht: (2025)
VideoQA in the Era of LLMs: An Empirical Study
von: Xiao, Junbin, et al.
Veröffentlicht: (2024)
von: Xiao, Junbin, et al.
Veröffentlicht: (2024)
UDVideoQA: A Traffic Video Question Answering Dataset for Multi-Object Spatio-Temporal Reasoning in Urban Dynamics
von: Vishal, Joseph Raj, et al.
Veröffentlicht: (2026)
von: Vishal, Joseph Raj, et al.
Veröffentlicht: (2026)
StreamingCoT: A Dataset for Temporal Dynamics and Multimodal Chain-of-Thought Reasoning in Streaming VideoQA
von: Hu, Yuhang, et al.
Veröffentlicht: (2025)
von: Hu, Yuhang, et al.
Veröffentlicht: (2025)
Reading Between the Lanes: Text VideoQA on the Road
von: Tom, George, et al.
Veröffentlicht: (2023)
von: Tom, George, et al.
Veröffentlicht: (2023)
Grounded Multi-Hop VideoQA in Long-Form Egocentric Videos
von: Chen, Qirui, et al.
Veröffentlicht: (2024)
von: Chen, Qirui, et al.
Veröffentlicht: (2024)
Understanding Complexity in VideoQA via Visual Program Generation
von: Eyzaguirre, Cristobal, et al.
Veröffentlicht: (2025)
von: Eyzaguirre, Cristobal, et al.
Veröffentlicht: (2025)
Beyond Isolated Facts: Synthesizing Narrative and Grounded Supervision for VideoQA
von: Liang, Jianxin, et al.
Veröffentlicht: (2025)
von: Liang, Jianxin, et al.
Veröffentlicht: (2025)
Neuro Symbolic Knowledge Reasoning for Procedural Video Question Answering
von: Fernando, Basura, et al.
Veröffentlicht: (2025)
von: Fernando, Basura, et al.
Veröffentlicht: (2025)
CCTVBench: Contrastive Consistency Traffic VideoQA Benchmark for Multimodal LLMs
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2026)
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2026)
Unbiased Scene Graph Generation by Type-Aware Message Passing on Heterogeneous and Dual Graphs
von: Sun, Guanglu, et al.
Veröffentlicht: (2024)
von: Sun, Guanglu, et al.
Veröffentlicht: (2024)
CaST-Bench: Benchmarking Causal Chain-Grounded Spatio-Temporal Reasoning for Video Question Answering
von: Zhang, Mingfang, et al.
Veröffentlicht: (2026)
von: Zhang, Mingfang, et al.
Veröffentlicht: (2026)
ActionAtlas: A VideoQA Benchmark for Domain-specialized Action Recognition
von: Salehi, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Salehi, Mohammadreza, et al.
Veröffentlicht: (2024)
Dissecting Multimodality in VideoQA Transformer Models by Impairing Modality Fusion
von: Rawal, Ishaan Singh, et al.
Veröffentlicht: (2023)
von: Rawal, Ishaan Singh, et al.
Veröffentlicht: (2023)
Perceive, Query & Reason: Enhancing Video QA with Question-Guided Temporal Queries
von: Amoroso, Roberto, et al.
Veröffentlicht: (2024)
von: Amoroso, Roberto, et al.
Veröffentlicht: (2024)
Align and Aggregate: Compositional Reasoning with Video Alignment and Answer Aggregation for Video Question-Answering
von: Liao, Zhaohe, et al.
Veröffentlicht: (2024)
von: Liao, Zhaohe, et al.
Veröffentlicht: (2024)
Video-in-the-Loop: Span-Grounded Long Video QA with Interleaved Reasoning
von: Wang, Chendong, et al.
Veröffentlicht: (2025)
von: Wang, Chendong, et al.
Veröffentlicht: (2025)
RoadSocial: A Diverse VideoQA Dataset and Benchmark for Road Event Understanding from Social Video Narratives
von: Parikh, Chirag, et al.
Veröffentlicht: (2025)
von: Parikh, Chirag, et al.
Veröffentlicht: (2025)
UpstreamQA: A Modular Framework for Explicit Reasoning on Video Question Answering Tasks
von: Nguyen, Jason, et al.
Veröffentlicht: (2026)
von: Nguyen, Jason, et al.
Veröffentlicht: (2026)
DocVideoQA: Towards Comprehensive Understanding of Document-Centric Videos through Question Answering
von: Wang, Haochen, et al.
Veröffentlicht: (2025)
von: Wang, Haochen, et al.
Veröffentlicht: (2025)
NeuS-QA: Grounding Long-Form Video Understanding in Temporal Logic and Neuro-Symbolic Reasoning
von: Shah, Sahil, et al.
Veröffentlicht: (2025)
von: Shah, Sahil, et al.
Veröffentlicht: (2025)
LeAdQA: LLM-Driven Context-Aware Temporal Grounding for Video Question Answering
von: Dong, Xinxin, et al.
Veröffentlicht: (2025)
von: Dong, Xinxin, et al.
Veröffentlicht: (2025)
MovieRecapsQA: A Multimodal Open-Ended Video Question-Answering Benchmark
von: Shaar, Shaden, et al.
Veröffentlicht: (2026)
von: Shaar, Shaden, et al.
Veröffentlicht: (2026)
Overview of TREC 2024 Medical Video Question Answering (MedVidQA) Track
von: Gupta, Deepak, et al.
Veröffentlicht: (2024)
von: Gupta, Deepak, et al.
Veröffentlicht: (2024)
STAIR: Spatial-Temporal Reasoning with Auditable Intermediate Results for Video Question Answering
von: Wang, Yueqian, et al.
Veröffentlicht: (2024)
von: Wang, Yueqian, et al.
Veröffentlicht: (2024)
Perception Test 2024: Challenge Summary and a Novel Hour-Long VideoQA Benchmark
von: Heyward, Joseph, et al.
Veröffentlicht: (2024)
von: Heyward, Joseph, et al.
Veröffentlicht: (2024)
Sports-QA: A Large-Scale Video Question Answering Benchmark for Complex and Professional Sports
von: Li, Haopeng, et al.
Veröffentlicht: (2024)
von: Li, Haopeng, et al.
Veröffentlicht: (2024)
TemporalDoRA: Temporal PEFT for Robust Surgical Video Question Answering
von: Carlini, Luca, et al.
Veröffentlicht: (2026)
von: Carlini, Luca, et al.
Veröffentlicht: (2026)
V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning
von: Cheng, Zixu, et al.
Veröffentlicht: (2025)
von: Cheng, Zixu, et al.
Veröffentlicht: (2025)
STEP: Enhancing Video-LLMs' Compositional Reasoning by Spatio-Temporal Graph-guided Self-Training
von: Qiu, Haiyi, et al.
Veröffentlicht: (2024)
von: Qiu, Haiyi, et al.
Veröffentlicht: (2024)
VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering
von: Meng, Yiran, et al.
Veröffentlicht: (2025)
von: Meng, Yiran, et al.
Veröffentlicht: (2025)
REVEAL: Relation-based Video Representation Learning for Video-Question-Answering
von: Chaybouti, Sofian, et al.
Veröffentlicht: (2025)
von: Chaybouti, Sofian, et al.
Veröffentlicht: (2025)
ViLA: Efficient Video-Language Alignment for Video Question Answering
von: Wang, Xijun, et al.
Veröffentlicht: (2023)
von: Wang, Xijun, et al.
Veröffentlicht: (2023)
VIRST: Video-Instructed Reasoning Assistant for SpatioTemporal Segmentation
von: Hong, Jihwan, et al.
Veröffentlicht: (2026)
von: Hong, Jihwan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
VideoQA-SC: Adaptive Semantic Communication for Video Question Answering
von: Guo, Jiangyuan, et al.
Veröffentlicht: (2024) -
Leveraging Static Relationships for Intra-Type and Inter-Type Message Passing in Video Question Answering
von: Liang, Lili, et al.
Veröffentlicht: (2025) -
TUMTraffic-VideoQA: A Benchmark for Unified Spatio-Temporal Video Understanding in Traffic Scenes
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2025) -
QTG-VQA: Question-Type-Guided Architectural for VideoQA Systems
von: He, Zhixian, et al.
Veröffentlicht: (2024) -
ReasVQA: Advancing VideoQA with Imperfect Reasoning Process
von: Liang, Jianxin, et al.
Veröffentlicht: (2025)