VACT: A Video Automatic Causal Testing System and a Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Haotong, Zheng, Qingyuan, Gao, Yunjian, Yang, Yongkun, He, Yangbo, Lin, Zhouchen, Zhang, Muhan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diffusion-Based Cross-Modal Feature Extraction for Multi-Label Classification
by: Lan, Tian, et al.
Published: (2025)
by: Lan, Tian, et al.
Published: (2025)
AI-Assisted Decision-Making for Clinical Assessment of Auto-Segmented Contour Quality
by: Wang, Biling, et al.
Published: (2025)
by: Wang, Biling, et al.
Published: (2025)
TASTE: A Designer-Annotated Multi-Dimensional Preference Dataset for AI-Generated Graphic Design
by: Zhu, Haonan, et al.
Published: (2026)
by: Zhu, Haonan, et al.
Published: (2026)
IMC 2024 Methods & Solutions Review
by: Gupta, Shyam, et al.
Published: (2024)
by: Gupta, Shyam, et al.
Published: (2024)
Remote sensing data imputation using deep learning for multispectral imagery
by: Liua, Shuang, et al.
Published: (2026)
by: Liua, Shuang, et al.
Published: (2026)
Landcover classification and change detection using remote sensing and machine learning: a case study of Western Fiji
by: Gurjar, Yadvendra, et al.
Published: (2025)
by: Gurjar, Yadvendra, et al.
Published: (2025)
Scalable Dynamic Origin-Destination Demand Estimation Enhanced by High-Resolution Satellite Imagery Data
by: Liu, Jiachao, et al.
Published: (2025)
by: Liu, Jiachao, et al.
Published: (2025)
Aortic root landmark localization with optimal transport loss for heatmap regression
by: Ishizone, Tsuyoshi, et al.
Published: (2024)
by: Ishizone, Tsuyoshi, et al.
Published: (2024)
Promoting AI Equity in Science: Generalized Domain Prompt Learning for Accessible VLM Research
by: Cao, Qinglong, et al.
Published: (2024)
by: Cao, Qinglong, et al.
Published: (2024)
Provable Contrastive Continual Learning
by: Wen, Yichen, et al.
Published: (2024)
by: Wen, Yichen, et al.
Published: (2024)
VideoGUI: A Benchmark for GUI Automation from Instructional Videos
by: Lin, Kevin Qinghong, et al.
Published: (2024)
by: Lin, Kevin Qinghong, et al.
Published: (2024)
Morphological Detection and Classification of Microplastics and Nanoplastics Emerged from Consumer Products by Deep Learning
by: Rezvani, Hadi, et al.
Published: (2024)
by: Rezvani, Hadi, et al.
Published: (2024)
The Solution for Temporal Action Localisation Task of Perception Test Challenge 2024
by: Han, Yinan, et al.
Published: (2024)
by: Han, Yinan, et al.
Published: (2024)
DocPTBench: Benchmarking End-to-End Photographed Document Parsing and Translation
by: Du, Yongkun, et al.
Published: (2025)
by: Du, Yongkun, et al.
Published: (2025)
CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos
by: Li, Xuchen, et al.
Published: (2025)
by: Li, Xuchen, et al.
Published: (2025)
ESA: Energy-Based Shot Assembly Optimization for Automatic Video Editing
by: Chen, Yaosen, et al.
Published: (2025)
by: Chen, Yaosen, et al.
Published: (2025)
Complex Mathematical Expression Recognition: Benchmark, Large-Scale Dataset and Strong Baseline
by: Bai, Weikang, et al.
Published: (2025)
by: Bai, Weikang, et al.
Published: (2025)
KM-GPT: An Automated Pipeline for Reconstructing Individual Patient Data from Kaplan-Meier Plots
by: Zhao, Yao, et al.
Published: (2025)
by: Zhao, Yao, et al.
Published: (2025)
AlphaEarth Satellite Embeddings for Modelling Climate Sensitive Diseases Towards Global Health Resilience
by: Nazir, Usman, et al.
Published: (2026)
by: Nazir, Usman, et al.
Published: (2026)
TinyBayes: Closed-Form Bayesian Inference via Jacobi Prior for Real-Time Image Classification on Edge Devices
by: Sardar, Shouvik, et al.
Published: (2026)
by: Sardar, Shouvik, et al.
Published: (2026)
Style-Based Neural Architectures for Real-Time Weather Classification
by: Ouattara, Hamed, et al.
Published: (2026)
by: Ouattara, Hamed, et al.
Published: (2026)
BiDepth: A Bidirectional-Depth Neural Network for Spatio-Temporal Prediction
by: Ehsani, Sina, et al.
Published: (2025)
by: Ehsani, Sina, et al.
Published: (2025)
Visual Error Patterns in Multi-Modal AI: A Statistical Approach
by: Wang, Ching-Yi
Published: (2024)
by: Wang, Ching-Yi
Published: (2024)
Massimo: Public Queue Monitoring and Management using Mass-Spring Model
by: Kumar, Abhijeet, et al.
Published: (2024)
by: Kumar, Abhijeet, et al.
Published: (2024)
Wilcoxon Nonparametric CFAR Scheme for Ship Detection in SAR Image
by: Meng, Xiangwei
Published: (2024)
by: Meng, Xiangwei
Published: (2024)
CanvasMAR: Improving Masked Autoregressive Video Prediction With Canvas
by: Li, Zian, et al.
Published: (2025)
by: Li, Zian, et al.
Published: (2025)
ColonScopeX: Leveraging Explainable Expert Systems with Multimodal Data for Improved Early Diagnosis of Colorectal Cancer
by: Sikora, Natalia, et al.
Published: (2025)
by: Sikora, Natalia, et al.
Published: (2025)
CausalVE: Face Video Privacy Encryption via Causal Video Prediction
by: Huang, Yubo, et al.
Published: (2024)
by: Huang, Yubo, et al.
Published: (2024)
A Large-Scale Benchmark of Cross-Modal Learning for Histology and Gene Expression in Spatial Transcriptomics
by: Gindra, Rushin H., et al.
Published: (2025)
by: Gindra, Rushin H., et al.
Published: (2025)
HumanVBench: Probing Human-Centric Video Understanding in MLLMs with Automatically Synthesized Benchmarks
by: Zhou, Ting, et al.
Published: (2024)
by: Zhou, Ting, et al.
Published: (2024)
Post-Training Quantization for Video Matting
by: Zhu, Tianrui, et al.
Published: (2025)
by: Zhu, Tianrui, et al.
Published: (2025)
OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation
by: Yuan, Shenghai, et al.
Published: (2025)
by: Yuan, Shenghai, et al.
Published: (2025)
Understanding Learning with Sliced-Wasserstein Requires Rethinking Informative Slices
by: Tran, Huy, et al.
Published: (2024)
by: Tran, Huy, et al.
Published: (2024)
MMSI-Video-Bench: A Holistic Benchmark for Video-Based Spatial Intelligence
by: Lin, Jingli, et al.
Published: (2025)
by: Lin, Jingli, et al.
Published: (2025)
SoccerHigh: A Benchmark Dataset for Automatic Soccer Video Summarization
by: Díaz-Juan, Artur, et al.
Published: (2025)
by: Díaz-Juan, Artur, et al.
Published: (2025)
Any-to-Bokeh: Arbitrary-Subject Video Refocusing with Video Diffusion Model
by: Yang, Yang, et al.
Published: (2025)
by: Yang, Yang, et al.
Published: (2025)
H2VU-Benchmark: A Comprehensive Benchmark for Hierarchical Holistic Video Understanding
by: Wu, Qi, et al.
Published: (2025)
by: Wu, Qi, et al.
Published: (2025)
Latent Knowledge-Guided Video Diffusion for Scientific Phenomena Generation from a Single Initial Frame
by: Cao, Qinglong, et al.
Published: (2024)
by: Cao, Qinglong, et al.
Published: (2024)
Rethinking Metrics and Benchmarks of Video Anomaly Detection
by: Liu, Zihao, et al.
Published: (2025)
by: Liu, Zihao, et al.
Published: (2025)
Parrot Mind: Towards Explaining the Complex Task Reasoning of Pretrained Large Language Models with Template-Content Structure
by: Yang, Haotong, et al.
Published: (2023)
by: Yang, Haotong, et al.
Published: (2023)
Similar Items
-
Diffusion-Based Cross-Modal Feature Extraction for Multi-Label Classification
by: Lan, Tian, et al.
Published: (2025) -
AI-Assisted Decision-Making for Clinical Assessment of Auto-Segmented Contour Quality
by: Wang, Biling, et al.
Published: (2025) -
TASTE: A Designer-Annotated Multi-Dimensional Preference Dataset for AI-Generated Graphic Design
by: Zhu, Haonan, et al.
Published: (2026) -
IMC 2024 Methods & Solutions Review
by: Gupta, Shyam, et al.
Published: (2024) -
Remote sensing data imputation using deep learning for multispectral imagery
by: Liua, Shuang, et al.
Published: (2026)