InterAct-Video: Reasoning-Rich Video QA for Urban Traffic
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vishal, Joseph Raj, Basina, Divesh, Patil, Rutuja, Gowda, Manas Srinivas, Naik, Katha, Yang, Yezhou, Chakravarthi, Bharatesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Eyes on the Road: State-of-the-Art Video Question Answering Models Assessment for Traffic Monitoring Tasks
von: Vishal, Joseph Raj, et al.
Veröffentlicht: (2024)
von: Vishal, Joseph Raj, et al.
Veröffentlicht: (2024)
UDVideoQA: A Traffic Video Question Answering Dataset for Multi-Object Spatio-Temporal Reasoning in Urban Dynamics
von: Vishal, Joseph Raj, et al.
Veröffentlicht: (2026)
von: Vishal, Joseph Raj, et al.
Veröffentlicht: (2026)
KAT to KANs: A Review of Kolmogorov-Arnold Networks and the Neural Leap Forward
von: Basina, Divesh, et al.
Veröffentlicht: (2024)
von: Basina, Divesh, et al.
Veröffentlicht: (2024)
SKoPe3D: A Synthetic Dataset for Vehicle Keypoint Perception in 3D from Traffic Monitoring Cameras
von: Pahadia, Himanshu, et al.
Veröffentlicht: (2023)
von: Pahadia, Himanshu, et al.
Veröffentlicht: (2023)
eSkiTB: A Synthetic Event-based Dataset for Tracking Skiers
von: Vinod, Krishna, et al.
Veröffentlicht: (2026)
von: Vinod, Krishna, et al.
Veröffentlicht: (2026)
eNavi: Event-based Imitation Policies for Low-Light Indoor Mobile Robot Navigation
von: Ramesh, Prithvi Jai, et al.
Veröffentlicht: (2026)
von: Ramesh, Prithvi Jai, et al.
Veröffentlicht: (2026)
Scale-Aware Vision-Language Adaptation for Extreme Far-Distance Video Person Re-identification
von: Rajbhandari, Ashwat, et al.
Veröffentlicht: (2026)
von: Rajbhandari, Ashwat, et al.
Veröffentlicht: (2026)
eTraM: Event-based Traffic Monitoring Dataset
von: Verma, Aayush Atul, et al.
Veröffentlicht: (2024)
von: Verma, Aayush Atul, et al.
Veröffentlicht: (2024)
MC-BEVRO: Multi-Camera Bird Eye View Road Occupancy Detection for Traffic Monitoring
von: Vaghela, Arpitsinh, et al.
Veröffentlicht: (2025)
von: Vaghela, Arpitsinh, et al.
Veröffentlicht: (2025)
SynTraC: A Synthetic Dataset for Traffic Signal Control from Traffic Monitoring Cameras
von: Chen, Tiejin, et al.
Veröffentlicht: (2024)
von: Chen, Tiejin, et al.
Veröffentlicht: (2024)
Event Quality Score (EQS): Assessing the Realism of Simulated Event Camera Streams via Distances in Latent Space
von: Chanda, Kaustav, et al.
Veröffentlicht: (2025)
von: Chanda, Kaustav, et al.
Veröffentlicht: (2025)
Event-based Graph Representation with Spatial and Motion Vectors for Asynchronous Object Detection
von: Verma, Aayush Atul, et al.
Veröffentlicht: (2025)
von: Verma, Aayush Atul, et al.
Veröffentlicht: (2025)
SEPose: A Synthetic Event-based Human Pose Estimation Dataset for Pedestrian Monitoring
von: Chanda, Kaustav, et al.
Veröffentlicht: (2025)
von: Chanda, Kaustav, et al.
Veröffentlicht: (2025)
Recent Event Camera Innovations: A Survey
von: Chakravarthi, Bharatesh, et al.
Veröffentlicht: (2024)
von: Chakravarthi, Bharatesh, et al.
Veröffentlicht: (2024)
Roundabout Dilemma Zone Data Mining and Forecasting with Trajectory Prediction and Graph Neural Networks
von: Satish, Manthan Chelenahalli, et al.
Veröffentlicht: (2024)
von: Satish, Manthan Chelenahalli, et al.
Veröffentlicht: (2024)
SEVD: Synthetic Event-based Vision Dataset for Ego and Fixed Traffic Perception
von: Aliminati, Manideep Reddy, et al.
Veröffentlicht: (2024)
von: Aliminati, Manideep Reddy, et al.
Veröffentlicht: (2024)
How Real is CARLAs Dynamic Vision Sensor? A Study on the Sim-to-Real Gap in Traffic Object Detection
von: Tan, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Tan, Kaiyuan, et al.
Veröffentlicht: (2025)
InterAct: Advancing Large-Scale Versatile 3D Human-Object Interaction Generation
von: Xu, Sirui, et al.
Veröffentlicht: (2025)
von: Xu, Sirui, et al.
Veröffentlicht: (2025)
InterAct: Capture and Modelling of Realistic, Expressive and Interactive Activities between Two Persons in Daily Scenarios
von: Huang, Yinghao, et al.
Veröffentlicht: (2024)
von: Huang, Yinghao, et al.
Veröffentlicht: (2024)
InterAct: A Large-Scale Dataset of Dynamic, Expressive and Interactive Activities between Two People in Daily Scenarios
von: Ho, Leo, et al.
Veröffentlicht: (2025)
von: Ho, Leo, et al.
Veröffentlicht: (2025)
Quantum trajectories and Page-curve entanglement dynamics
von: Ganguly, Katha, et al.
Veröffentlicht: (2025)
von: Ganguly, Katha, et al.
Veröffentlicht: (2025)
Transport in open quantum systems in presence of lossy channels
von: Ganguly, Katha, et al.
Veröffentlicht: (2024)
von: Ganguly, Katha, et al.
Veröffentlicht: (2024)
A Comprehensive Review of Leap Motion Controller-based Hand Gesture Datasets
von: Chakravarthi, Bharatesh, et al.
Veröffentlicht: (2023)
von: Chakravarthi, Bharatesh, et al.
Veröffentlicht: (2023)
Video-in-the-Loop: Span-Grounded Long Video QA with Interleaved Reasoning
von: Wang, Chendong, et al.
Veröffentlicht: (2025)
von: Wang, Chendong, et al.
Veröffentlicht: (2025)
VideoMathQA: Benchmarking Mathematical Reasoning via Multimodal Understanding in Videos
von: Rasheed, Hanoona, et al.
Veröffentlicht: (2025)
von: Rasheed, Hanoona, et al.
Veröffentlicht: (2025)
SEBVS: Synthetic Event-based Visual Servoing for Robot Navigation and Manipulation
von: Vinod, Krishna, et al.
Veröffentlicht: (2025)
von: Vinod, Krishna, et al.
Veröffentlicht: (2025)
TUMTraffic-VideoQA: A Benchmark for Unified Spatio-Temporal Video Understanding in Traffic Scenes
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2025)
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2025)
Detection of Micromobility Vehicles in Urban Traffic Videos
von: Sabri, Khalil, et al.
Veröffentlicht: (2024)
von: Sabri, Khalil, et al.
Veröffentlicht: (2024)
CCTVBench: Contrastive Consistency Traffic VideoQA Benchmark for Multimodal LLMs
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2026)
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2026)
VRR-QA: Visual Relational Reasoning in Videos Beyond Explicit Cues
von: Swetha, Sirnam, et al.
Veröffentlicht: (2025)
von: Swetha, Sirnam, et al.
Veröffentlicht: (2025)
Anomalous transport in long-ranged open quantum systems
von: Dhawan, Abhinav, et al.
Veröffentlicht: (2024)
von: Dhawan, Abhinav, et al.
Veröffentlicht: (2024)
ENTER: Event Based Interpretable Reasoning for VideoQA
von: Ayyubi, Hammad, et al.
Veröffentlicht: (2025)
von: Ayyubi, Hammad, et al.
Veröffentlicht: (2025)
YTCommentQA: Video Question Answerability in Instructional Videos
von: Yang, Saelyne, et al.
Veröffentlicht: (2024)
von: Yang, Saelyne, et al.
Veröffentlicht: (2024)
ReasVQA: Advancing VideoQA with Imperfect Reasoning Process
von: Liang, Jianxin, et al.
Veröffentlicht: (2025)
von: Liang, Jianxin, et al.
Veröffentlicht: (2025)
WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning
von: Zhang, Yuanhan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuanhan, et al.
Veröffentlicht: (2024)
Real-Time Vehicle Detection and Urban Traffic Behavior Analysis Based on UAV Traffic Videos on Mobile Devices
von: Zhu, Yuan, et al.
Veröffentlicht: (2024)
von: Zhu, Yuan, et al.
Veröffentlicht: (2024)
Video-R4: Reinforcing Text-Rich Video Reasoning with Visual Rumination
von: Tang, Yolo Y., et al.
Veröffentlicht: (2025)
von: Tang, Yolo Y., et al.
Veröffentlicht: (2025)
Adaptive Dense Evidence Refinement for Video Relational Reasoning for VRR-QA Challenge
von: Sun, Yuyang, et al.
Veröffentlicht: (2026)
von: Sun, Yuyang, et al.
Veröffentlicht: (2026)
Stable Cinemetrics : Structured Taxonomy and Evaluation for Professional Video Generation
von: Chatterjee, Agneet, et al.
Veröffentlicht: (2025)
von: Chatterjee, Agneet, et al.
Veröffentlicht: (2025)
AdaFuse-Det: Adaptive Cross-Modal Fusion of Event Cameras for Robust Object Detection in Low-Light RGB Imagery
von: Imandi, Raju, et al.
Veröffentlicht: (2026)
von: Imandi, Raju, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Eyes on the Road: State-of-the-Art Video Question Answering Models Assessment for Traffic Monitoring Tasks
von: Vishal, Joseph Raj, et al.
Veröffentlicht: (2024) -
UDVideoQA: A Traffic Video Question Answering Dataset for Multi-Object Spatio-Temporal Reasoning in Urban Dynamics
von: Vishal, Joseph Raj, et al.
Veröffentlicht: (2026) -
KAT to KANs: A Review of Kolmogorov-Arnold Networks and the Neural Leap Forward
von: Basina, Divesh, et al.
Veröffentlicht: (2024) -
SKoPe3D: A Synthetic Dataset for Vehicle Keypoint Perception in 3D from Traffic Monitoring Cameras
von: Pahadia, Himanshu, et al.
Veröffentlicht: (2023) -
eSkiTB: A Synthetic Event-based Dataset for Tracking Skiers
von: Vinod, Krishna, et al.
Veröffentlicht: (2026)