TB-Bench: Training and Testing Multi-Modal AI for Understanding Spatio-Temporal Traffic Behaviors from Dashcam Images/Videos
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Charoenpitaks, Korawat, Nguyen, Van-Quang, Suganuma, Masanori, Arai, Kentaro, Totsuka, Seiji, Ino, Hiroshi, Okatani, Takayuki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exploring the Potential of Multi-Modal AI for Driving Hazard Prediction
von: Charoenpitaks, Korawat, et al.
Veröffentlicht: (2023)
von: Charoenpitaks, Korawat, et al.
Veröffentlicht: (2023)
CoReTab: Improving Multimodal Table Understanding with Code-driven Reasoning
von: Nguyen, Van-Quang, et al.
Veröffentlicht: (2026)
von: Nguyen, Van-Quang, et al.
Veröffentlicht: (2026)
An Improved Method for Personalizing Diffusion Models
von: Zeng, Yan, et al.
Veröffentlicht: (2024)
von: Zeng, Yan, et al.
Veröffentlicht: (2024)
RefVSR++: Exploiting Reference Inputs for Reference-based Video Super-resolution
von: Zou, Han, et al.
Veröffentlicht: (2023)
von: Zou, Han, et al.
Veröffentlicht: (2023)
Rethinking Unsupervised Domain Adaptation for Semantic Segmentation
von: Wang, Zhijie, et al.
Veröffentlicht: (2022)
von: Wang, Zhijie, et al.
Veröffentlicht: (2022)
Open-vocabulary vs. Closed-set: Best Practice for Few-shot Object Detection Considering Text Describability
von: Hosoya, Yusuke, et al.
Veröffentlicht: (2024)
von: Hosoya, Yusuke, et al.
Veröffentlicht: (2024)
Rethinking Annotation for Object Detection: Is Annotating Small-size Instances Worth Its Cost?
von: Hosoya, Yusuke, et al.
Veröffentlicht: (2024)
von: Hosoya, Yusuke, et al.
Veröffentlicht: (2024)
Rethinking Open-Set Object Detection: Issues, a New Formulation, and Taxonomy
von: Hosoya, Yusuke, et al.
Veröffentlicht: (2022)
von: Hosoya, Yusuke, et al.
Veröffentlicht: (2022)
Cascaded Multi-Scale Attention for Enhanced Multi-Scale Feature Extraction and Interaction with Low-Resolution Images
von: Lu, Xiangyong, et al.
Veröffentlicht: (2024)
von: Lu, Xiangyong, et al.
Veröffentlicht: (2024)
MS-DPPs: Multi-Source Determinantal Point Processes for Contextual Diversity Refinement of Composite Attributes in Text to Image Retrieval
von: Sogi, Naoya, et al.
Veröffentlicht: (2025)
von: Sogi, Naoya, et al.
Veröffentlicht: (2025)
RP-SLAM: Real-time Photorealistic SLAM with Efficient 3D Gaussian Splatting
von: Bai, Lizhi, et al.
Veröffentlicht: (2024)
von: Bai, Lizhi, et al.
Veröffentlicht: (2024)
360° Image Perception with MLLMs: A Comprehensive Benchmark and a Training-Free Method
von: Tran, Huyen T. T., et al.
Veröffentlicht: (2026)
von: Tran, Huyen T. T., et al.
Veröffentlicht: (2026)
Temporal Insight Enhancement: Mitigating Temporal Hallucination in Multimodal Large Language Models
von: Sun, Li, et al.
Veröffentlicht: (2024)
von: Sun, Li, et al.
Veröffentlicht: (2024)
LAMS-Edit: Latent and Attention Mixing with Schedulers for Improved Content Preservation in Diffusion-Based Image and Style Editing
von: Fu, Wingwa, et al.
Veröffentlicht: (2026)
von: Fu, Wingwa, et al.
Veröffentlicht: (2026)
ECOSSISTEMAS DE REFERÊNCIA PARA RESTAURAÇÃO DE MATAS CILIARES: EXISTEM PADRÕES DE BIODIVERSIDADE, ESTRUTURA FLORESTAL E ATRIBUTOS FUNCIONAIS?
von: Marcio Seiji Suganuma
Veröffentlicht: (2013)
von: Marcio Seiji Suganuma
Veröffentlicht: (2013)
Comparando metodologias para avaliar a cobertura do dossel e a luminosidade no sub-bosque de um reflorestamento e uma floresta madura
von: Márcio Seiji Suganuma
Veröffentlicht: (2008)
von: Márcio Seiji Suganuma
Veröffentlicht: (2008)
Enriquecimento artificial da diversidade de espécies em reflorestamentos: análise preliminar de dois métodos, transferência de serapilheira e semeadura direta
von: Marcio Seiji Suganuma
Veröffentlicht: (2008)
von: Marcio Seiji Suganuma
Veröffentlicht: (2008)
Clustered Cystic Changes in Long-Term Follow-Up Thin-Section Computed Tomographic Findings in Fibrotic Nonspecific Interstitial Pneumonia
von: Masanori Akira, et al.
Veröffentlicht: (2024)
von: Masanori Akira, et al.
Veröffentlicht: (2024)
Machine Intelligence that Understands Visual and Linguistic Information and Interacts with Humans and Environments
von: Nguyen, Van Quang
Veröffentlicht: (2026)
von: Nguyen, Van Quang
Veröffentlicht: (2026)
Action-Agnostic Point-Level Supervision for Temporal Action Detection
von: Yoshida, Shuhei M., et al.
Veröffentlicht: (2024)
von: Yoshida, Shuhei M., et al.
Veröffentlicht: (2024)
Interpretable Traffic Responsibility from Dashcam Video via Legal Multi Agent Reasoning
von: Yang, Jingchun, et al.
Veröffentlicht: (2026)
von: Yang, Jingchun, et al.
Veröffentlicht: (2026)
Spatio-Temporal Data Enhanced Vision-Language Model for Traffic Scene Understanding
von: Ma, Jingtian, et al.
Veröffentlicht: (2025)
von: Ma, Jingtian, et al.
Veröffentlicht: (2025)
A Simple Channel Compression Method for Brain Signal Decoding on Classification Task
von: Ji, Changqing, et al.
Veröffentlicht: (2024)
von: Ji, Changqing, et al.
Veröffentlicht: (2024)
xMTrans: Temporal Attentive Cross-Modality Fusion Transformer for Long-Term Traffic Prediction
von: Ung, Huy Quang, et al.
Veröffentlicht: (2024)
von: Ung, Huy Quang, et al.
Veröffentlicht: (2024)
DashCop: Automated E-ticket Generation for Two-Wheeler Traffic Violations Using Dashcam Videos
von: Rawat, Deepti, et al.
Veröffentlicht: (2025)
von: Rawat, Deepti, et al.
Veröffentlicht: (2025)
TUMTraffic-VideoQA: A Benchmark for Unified Spatio-Temporal Video Understanding in Traffic Scenes
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2025)
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2025)
Density-Matrix Renormalization Group Study of Kitaev--Heisenberg Model on a Triangular Lattice
von: Shinjo, Kazuya, et al.
Veröffentlicht: (2015)
von: Shinjo, Kazuya, et al.
Veröffentlicht: (2015)
Efficient Traffic Prediction Through Spatio-Temporal Distillation
von: Zhang, Qianru, et al.
Veröffentlicht: (2025)
von: Zhang, Qianru, et al.
Veröffentlicht: (2025)
BSFT-like action from cohomomorphism
von: Totsuka-Yoshinaka, Jojiro
Veröffentlicht: (2025)
von: Totsuka-Yoshinaka, Jojiro
Veröffentlicht: (2025)
Estimation of Kinematic Motion from Dashcam Footage
von: Zhang, Evelyn, et al.
Veröffentlicht: (2025)
von: Zhang, Evelyn, et al.
Veröffentlicht: (2025)
Nexar Dashcam Collision Prediction Dataset and Challenge
von: Moura, Daniel C., et al.
Veröffentlicht: (2025)
von: Moura, Daniel C., et al.
Veröffentlicht: (2025)
Object Detection for Vehicle Dashcams using Transformers
von: Mustafa, Osama, et al.
Veröffentlicht: (2024)
von: Mustafa, Osama, et al.
Veröffentlicht: (2024)
Spatio-Temporal Self-Supervised Learning for Traffic Flow Prediction
von: Ji, Jiahao, et al.
Veröffentlicht: (2022)
von: Ji, Jiahao, et al.
Veröffentlicht: (2022)
Spatio-Temporal Partial Sensing Forecast for Long-term Traffic
von: Liu, Zibo, et al.
Veröffentlicht: (2024)
von: Liu, Zibo, et al.
Veröffentlicht: (2024)
Adaptive and Dynamic Spatio‐Temporal Network for Traffic Flow Forecasting
von: Ying Xing, et al.
Veröffentlicht: (2025)
von: Ying Xing, et al.
Veröffentlicht: (2025)
From Dashcam Videos to Driving Simulations: Stress Testing Automated Vehicles against Rare Events
von: Miao, Yan, et al.
Veröffentlicht: (2024)
von: Miao, Yan, et al.
Veröffentlicht: (2024)
Flat-Pack Bench: Evaluating Spatio-Temporal Understanding in Large Vision-Language Models through Furniture Assembly
von: Chetan, Aditya, et al.
Veröffentlicht: (2026)
von: Chetan, Aditya, et al.
Veröffentlicht: (2026)
An approximate Kappa generator for particle simulations
von: Zenitani, Seiji, et al.
Veröffentlicht: (2026)
von: Zenitani, Seiji, et al.
Veröffentlicht: (2026)
An Elementary Proof of the Nonexistence of Tarski Monster Groups of Exponent 3
von: Arai, Hiroshi
Veröffentlicht: (2025)
von: Arai, Hiroshi
Veröffentlicht: (2025)
MissBench: Benchmarking Multimodal Affective Analysis under Imbalanced Missing Modalities
von: Pham, Tien Anh, et al.
Veröffentlicht: (2026)
von: Pham, Tien Anh, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Exploring the Potential of Multi-Modal AI for Driving Hazard Prediction
von: Charoenpitaks, Korawat, et al.
Veröffentlicht: (2023) -
CoReTab: Improving Multimodal Table Understanding with Code-driven Reasoning
von: Nguyen, Van-Quang, et al.
Veröffentlicht: (2026) -
An Improved Method for Personalizing Diffusion Models
von: Zeng, Yan, et al.
Veröffentlicht: (2024) -
RefVSR++: Exploiting Reference Inputs for Reference-based Video Super-resolution
von: Zou, Han, et al.
Veröffentlicht: (2023) -
Rethinking Unsupervised Domain Adaptation for Semantic Segmentation
von: Wang, Zhijie, et al.
Veröffentlicht: (2022)