Finding the Trigger: Causal Abductive Reasoning on Video Events
Fuente:
arXiv
Saved in:
| Main Authors: | Le, Thao Minh, Le, Vuong, Do, Kien, Gupta, Sunil, Venkatesh, Svetha, Tran, Truyen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Progressive Multi-granular Alignments for Grounded Reasoning in Large Vision-Language Models
by: Le, Quang-Hung, et al.
Published: (2024)
by: Le, Quang-Hung, et al.
Published: (2024)
Predicting the Reliability of an Image Classifier under Image Distortion
by: Nguyen, Dang, et al.
Published: (2024)
by: Nguyen, Dang, et al.
Published: (2024)
Towards Agentic AI for Multimodal-Guided Video Object Segmentation
by: Tran, Tuyen, et al.
Published: (2025)
by: Tran, Tuyen, et al.
Published: (2025)
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
by: Le, Minh Khoa, et al.
Published: (2026)
by: Le, Minh Khoa, et al.
Published: (2026)
Planner-Refiner: Dynamic Space-Time Refinement for Vision-Language Alignment in Videos
by: Tran, Tuyen, et al.
Published: (2025)
by: Tran, Tuyen, et al.
Published: (2025)
Federated Domain Generalization with Latent Space Inversion
by: Palakkadavath, Ragja, et al.
Published: (2025)
by: Palakkadavath, Ragja, et al.
Published: (2025)
Enhancing Length Extrapolation in Sequential Models with Pointer-Augmented Neural Memory
by: Le, Hung, et al.
Published: (2024)
by: Le, Hung, et al.
Published: (2024)
SADL: An Effective In-Context Learning Method for Compositional Visual QA
by: Dang, Long Hoang, et al.
Published: (2024)
by: Dang, Long Hoang, et al.
Published: (2024)
Diverse Image Priors for Black-box Data-free Knowledge Distillation
by: Vo, Tri-Nhan, et al.
Published: (2026)
by: Vo, Tri-Nhan, et al.
Published: (2026)
Unified Framework with Consistency across Modalities for Human Activity Recognition
by: Tran, Tuyen, et al.
Published: (2024)
by: Tran, Tuyen, et al.
Published: (2024)
Improving Diversity in Black-box Few-shot Knowledge Distillation
by: Vo, Tri-Nhan, et al.
Published: (2026)
by: Vo, Tri-Nhan, et al.
Published: (2026)
Universal Multi-Domain Translation via Diffusion Routers
by: Kieu, Duc, et al.
Published: (2025)
by: Kieu, Duc, et al.
Published: (2025)
Learning Structural Causal Models from Ordering: Identifiable Flow Models
by: Le, Minh Khoa, et al.
Published: (2024)
by: Le, Minh Khoa, et al.
Published: (2024)
Stable Hadamard Memory: Revitalizing Memory-Augmented Agents for Reinforcement Learning
by: Le, Hung, et al.
Published: (2024)
by: Le, Hung, et al.
Published: (2024)
Score-based Integrated Gradient for Root Cause Explanations of Outliers
by: Nguyen, Phuoc, et al.
Published: (2026)
by: Nguyen, Phuoc, et al.
Published: (2026)
Revisiting the Dataset Bias Problem from a Statistical Perspective
by: Do, Kien, et al.
Published: (2024)
by: Do, Kien, et al.
Published: (2024)
Composite Concept Extraction through Backdooring
by: Ghosh, Banibrata, et al.
Published: (2024)
by: Ghosh, Banibrata, et al.
Published: (2024)
Test-Time Instance-Specific Parameter Composition: A New Paradigm for Adaptive Generative Modeling
by: Tran, Minh-Tuan, et al.
Published: (2026)
by: Tran, Minh-Tuan, et al.
Published: (2026)
Text-Enhanced Data-free Approach for Federated Class-Incremental Learning
by: Tran, Minh-Tuan, et al.
Published: (2024)
by: Tran, Minh-Tuan, et al.
Published: (2024)
Ensemble Learning for Vietnamese Scene Text Spotting in Urban Environments
by: Nguyen, Hieu, et al.
Published: (2024)
by: Nguyen, Hieu, et al.
Published: (2024)
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
Towards a text-based quantitative and explainable histopathology image analysis
by: Nguyen, Anh Tien, et al.
Published: (2024)
by: Nguyen, Anh Tien, et al.
Published: (2024)
Fantastic Targets for Concept Erasure in Diffusion Models and Where To Find Them
by: Bui, Anh, et al.
Published: (2025)
by: Bui, Anh, et al.
Published: (2025)
Online 3D Multi-Camera Perception through Robust 2D Tracking and Depth-based Late Aggregation
by: Le, Vu-Minh, et al.
Published: (2025)
by: Le, Vu-Minh, et al.
Published: (2025)
Beyond Surprise: Improving Exploration Through Surprise Novelty
by: Le, Hung, et al.
Published: (2023)
by: Le, Hung, et al.
Published: (2023)
ReasonVQA: A Multi-hop Reasoning Benchmark with Structural Knowledge for Visual Question Answering
by: Tran, Duong T., et al.
Published: (2025)
by: Tran, Duong T., et al.
Published: (2025)
Robust Deepfake Detection: Mitigating Spatial Attention Drift via Calibrated Complementary Ensembles
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
EDGER: EDge-Guided with HEatmap Refinement for Generalizable Image Forgery Localization
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
by: Le-Phan, Minh-Khoa, et al.
Published: (2026)
Enhancing Domain Adaptation through Prompt Gradient Alignment
by: Phan, Hoang, et al.
Published: (2024)
by: Phan, Hoang, et al.
Published: (2024)
Efficiently Assemble Normalization Layers and Regularization for Federated Domain Generalization
by: Le, Khiem, et al.
Published: (2024)
by: Le, Khiem, et al.
Published: (2024)
Black Swan: Abductive and Defeasible Video Reasoning in Unpredictable Events
by: Chinchure, Aditya, et al.
Published: (2024)
by: Chinchure, Aditya, et al.
Published: (2024)
LaVy: Vietnamese Multimodal Large Language Model
by: Tran, Chi, et al.
Published: (2024)
by: Tran, Chi, et al.
Published: (2024)
A Transformer-based Multimodal Fusion Model for Efficient Crowd Counting Using Visual and Wireless Signals
by: Cui, Zhe, et al.
Published: (2025)
by: Cui, Zhe, et al.
Published: (2025)
Generating Realistic Tabular Data with Large Language Models
by: Nguyen, Dang, et al.
Published: (2024)
by: Nguyen, Dang, et al.
Published: (2024)
Object Detection in Thermal Images Using Deep Learning for Unmanned Aerial Vehicles
by: Tu, Minh Dang, et al.
Published: (2024)
by: Tu, Minh Dang, et al.
Published: (2024)
KOPPA: Improving Prompt-based Continual Learning with Key-Query Orthogonal Projection and Prototype-based One-Versus-All
by: Tran, Quyen, et al.
Published: (2023)
by: Tran, Quyen, et al.
Published: (2023)
Erasing Undesirable Concepts in Diffusion Models with Adversarial Preservation
by: Bui, Anh, et al.
Published: (2024)
by: Bui, Anh, et al.
Published: (2024)
Multi-Reference Preference Optimization for Large Language Models
by: Le, Hung, et al.
Published: (2024)
by: Le, Hung, et al.
Published: (2024)
Exploring the Practicality of Federated Learning: A Survey Towards the Communication Perspective
by: Le, Khiem, et al.
Published: (2024)
by: Le, Khiem, et al.
Published: (2024)
Event-Enriched Image Analysis Grand Challenge at ACM Multimedia 2025
by: Tran, Thien-Phuc, et al.
Published: (2025)
by: Tran, Thien-Phuc, et al.
Published: (2025)
Similar Items
-
Progressive Multi-granular Alignments for Grounded Reasoning in Large Vision-Language Models
by: Le, Quang-Hung, et al.
Published: (2024) -
Predicting the Reliability of an Image Classifier under Image Distortion
by: Nguyen, Dang, et al.
Published: (2024) -
Towards Agentic AI for Multimodal-Guided Video Object Segmentation
by: Tran, Tuyen, et al.
Published: (2025) -
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
by: Le, Minh Khoa, et al.
Published: (2026) -
Planner-Refiner: Dynamic Space-Time Refinement for Vision-Language Alignment in Videos
by: Tran, Tuyen, et al.
Published: (2025)